首页    期刊浏览 2024年12月03日 星期二
登录注册

文章基本信息

  • 标题:Exploring Stylometric and Emotion-Based Features for Multilingual Cross-Domain Hate Speech Detection
  • 本地全文:下载
  • 作者:Ilia Markov ; Nikola Ljubešić ; Darja Fišer
  • 期刊名称:Conference on European Chapter of the Association for Computational Linguistics (EACL)
  • 出版年度:2021
  • 卷号:2021
  • 页码:149-159
  • 语种:English
  • 出版社:ACL Anthology
  • 摘要:In this paper, we describe experiments designed to evaluate the impact of stylometric and emotion-based features on hate speech detection: the task of classifying textual content into hate or non-hate speech classes. Our experiments are conducted for three languages – English, Slovene, and Dutch – both in in-domain and cross-domain setups, and aim to investigate hate speech using features that model two linguistic phenomena: the writing style of hateful social media content operationalized as function word usage on the one hand, and emotion expression in hateful messages on the other hand. The results of experiments with features that model different combinations of these phenomena support our hypothesis that stylometric and emotion-based features are robust indicators of hate speech. Their contribution remains persistent with respect to domain and language variation. We show that the combination of features that model the targeted phenomena outperforms words and character n-gram features under cross-domain conditions, and provides a significant boost to deep learning models, which currently obtain the best results, when combined with them in an ensemble.
国家哲学社会科学文献中心版权所有