STATE ToxiCN provides the first span-level Chinese hate speech dataset with 9,533 target-argument-hateful-group quadruples and a 830-term annotated hateful slang lexicon, and baseline results show fine-tuned models outperform LLM APIs.
In Proceedings of the 15th international workshop on semantic evaluation (SemEval-2021), pages 59–69
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
STATE ToxiCN: A Benchmark for Span-level Target-Aware Toxicity Extraction in Chinese Hate Speech Detection
STATE ToxiCN provides the first span-level Chinese hate speech dataset with 9,533 target-argument-hateful-group quadruples and a 830-term annotated hateful slang lexicon, and baseline results show fine-tuned models outperform LLM APIs.