Pith. sign in

REVIEW 1 cited by

The Constant in HATE: Analyzing Toxicity in Reddit across Topics and Languages

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2404.18726 v1 pith:6RGOS52N submitted 2024-04-29 cs.CL

classification cs.CL
keywords communitiestopicstoxicitylanguagesacrosslanguageredditsignificant
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Toxic language remains an ongoing challenge on social media platforms, presenting significant issues for users and communities. This paper provides a cross-topic and cross-lingual analysis of toxicity in Reddit conversations. We collect 1.5 million comment threads from 481 communities in six languages: English, German, Spanish, Turkish,Arabic, and Dutch, covering 80 topics such as Culture, Politics, and News. We thoroughly analyze how toxicity spikes within different communities in relation to specific topics. We observe consistent patterns of increased toxicity across languages for certain topics, while also noting significant variations within specific language communities.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Understanding and Analyzing Inappropriately Targeting Language in Online Discourse: A Comparative Annotation Study

    cs.CL 2025-05 conditional novelty 4.0 of 10

    A comparative annotation study finds ChatGPT over-identifies inappropriate targeting in Reddit conversations and uncovers four new target categories beyond the standard hate speech classes.

Pith tools