REVIEW 1 cited by
SGHateCheck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of Singapore
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
To address the limitations of current hate speech detection models, we introduce \textsf{SGHateCheck}, a novel framework designed for the linguistic and cultural context of Singapore and Southeast Asia. It extends the functional testing approach of HateCheck and MHC, employing large language models for translation and paraphrasing into Singapore's main languages, and refining these with native annotators. \textsf{SGHateCheck} reveals critical flaws in state-of-the-art models, highlighting their inadequacy in sensitive content moderation. This work aims to foster the development of more effective hate speech detection tools for diverse linguistic environments, particularly for Singapore and Southeast Asia contexts.
Forward citations
Cited by 1 Pith paper
-
Stylistic Evolution and LLM Neutrality in Singlish Language
Singlish changed cumulatively over a decade, and LLM-generated Singlish remains tied to particular time periods: realistic outputs carry temporal bias, while neutral outputs lose authenticity.
Discussion (0). Continue with ORCID to comment.