< Terug naar vorige pagina

Publicatie

Sarcasm detection using an ensemble approach

Boekbijdrage - Boekabstract Conferentiebijdrage

We present an ensemble approach for the detection of sarcasm in Reddit and Twitter responses in the context of The Second Workshop on Figurative Language Processing held in conjunction with ACL 2020(1). The ensemble is trained on the predicted sarcasm probabilities of four component models and on additional features, such as the sentiment of the comment, its length, and source (Reddit or Twitter) in order to learn which of the component models is the most reliable for which input. The component models consist of an LSTM with hashtag and emoji representations; a CNN-LSTM with casing, stop word, punctuation, and sentiment representations; an MLP based on Infersent embeddings; and an SVM trained on stylometric and emotion-based features. All component models use the two conversational turns preceding the response as context, except for the SVM, which only uses features extracted from the response. The ensemble itself consists of an adaboost classifier with the decision tree algorithm as base estimator and yields F1-scores of 67% and 74% on the Reddit and Twitter test data, respectively.
Boek: 2nd Workshop on Figurative Language Processing, JUL09, 2020, ELECTR NETWORK
Pagina's: 264 - 269
ISBN:978-1-952148-12-5
Jaar van publicatie:2020
Trefwoorden:P1 Proceeding
BOF-keylabel:ja
Authors from:Higher Education
Toegankelijkheid:Closed