A three-way classifier trained on Reddit hate speech/counterspeech pairs predicts hater reentry and reentry type more accurately than a two-stage predictor, with linguistic features of counterspeech signaling different reactions.
A Study of Cyber Hate on Twitter with Implications for Social Media Governance Strategies
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
This paper explores ways in which the harmful effects of cyber hate may be mitigated through mechanisms for enhancing the self governance of new digital spaces. We report findings from a mixed methods study of responses to cyber hate posts, which aimed to: (i) understand how people interact in this context by undertaking qualitative interaction analysis and developing a statistical model to explain the volume of responses to cyber hate posted to Twitter, and (ii) explore use of machine learning techniques to assist in identifying cyber hate counter-speech.
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Echoes of Discord: Forecasting Hater Reactions to Counterspeech
A three-way classifier trained on Reddit hate speech/counterspeech pairs predicts hater reentry and reentry type more accurately than a two-stage predictor, with linguistic features of counterspeech signaling different reactions.