Pith. sign in

REVIEW 2 cited by

Target Span Detection for Implicit Harmful Content

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.19836 v2 pith:XQZIONZC submitted 2024-03-28 cs.CL

Target Span Detection for Implicit Harmful Content

classification cs.CL
keywords speechdetectionhatetargetcontentharmfulidentifyingimplicit
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Identifying the targets of hate speech is a crucial step in grasping the nature of such speech and, ultimately, in improving the detection of offensive posts on online forums. Much harmful content on online platforms uses implicit language especially when targeting vulnerable and protected groups such as using stereotypical characteristics instead of explicit target names, making it harder to detect and mitigate the language. In this study, we focus on identifying implied targets of hate speech, essential for recognizing subtler hate speech and enhancing the detection of harmful content on digital platforms. We define a new task aimed at identifying the targets even when they are not explicitly stated. To address that task, we collect and annotate target spans in three prominent implicit hate speech datasets: SBIC, DynaHate, and IHC. We call the resulting merged collection Implicit-Target-Span. The collection is achieved using an innovative pooling method with matching scores based on human annotations and Large Language Models (LLMs). Our experiments indicate that Implicit-Target-Span provides a challenging test bed for target span detection methods.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. When Does Span-Guided Detoxification Help? Human Preferences and Evaluator Diagnostics in a Controlled Comparison

    cs.CL 2026-07 conditional novelty 5.0

    Human preferences favor span-guided and unguided detoxification under complementary failure risks, with a large stratum association that automatic and LLM evaluators do not recover.

  2. AEGIS: Awareness-Enhanced Guidance for Iterative Safeguard

    cs.CL 2026-07 conditional novelty 5.0

    Marking offensive spans changes — but does not consistently improve — the toxicity–meaning trade-off in multilingual detoxification; the effect depends on the generator backbone and the language.