STATE ToxiCN provides the first span-level Chinese hate speech dataset with 9,533 target-argument-hateful-group quadruples and a 830-term annotated hateful slang lexicon, and baseline results show fine-tuned models outperform LLM APIs.
In Proceedings of the 2024 Conference on Empirical Methods in Natural Lan- guage Processing, pages 6012–6025, Miami, Florida, USA
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
STATE ToxiCN: A Benchmark for Span-level Target-Aware Toxicity Extraction in Chinese Hate Speech Detection
STATE ToxiCN provides the first span-level Chinese hate speech dataset with 9,533 target-argument-hateful-group quadruples and a 830-term annotated hateful slang lexicon, and baseline results show fine-tuned models outperform LLM APIs.