Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:03:29.370944Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 1 inbound Pith citation observation for arXiv:2501.16750.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:03:29.370944Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T01:02:51.934157Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T01:02:53.997511Z
100 of 100 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3ad7f745-cbff-421d-9ec5-fc59bfb5fbbb · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://en.wikipedia.org/wiki/ Coleman-Liau_index
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03b1fd37-1304-4c6b-ad09-a10ebfb26c4c · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://github.com/unitaryai/detoxify
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4f50355-d933-4a08-88d7-26d878d1b1b5 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://gdpr-info.eu/
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dcf60a8-89d3-48d1-939a-e30de2bde689 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://chatgpt.com/g/g-w0y3CvDM 9-freddy-griffin
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a63bdbfd-cad7-412f-9a4b-8f93bc1429cd · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://chatgpt.com/g/g-JlQ9WBdHB-hate /
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd1fe658-5211-4357-b5f7-f4cc5a9b7645 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://chatgpt.com/g/g-87uTmBE65-ru de-gpt
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bbde68a-1663-43f4-a4ce-e354f24e7abe · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://www.perspectiveapi.com
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ac119e8-d12a-4ca5-a0b5-a24c13c232be · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://osf.io/edua3/
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8edefd31-f421-45fc-b7bc-f66f21388a73 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns https://lmsys.org/blog/2023-03-30-vicuna/
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebaeb6ad-bb19-4144-8aed-535777dd7068 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns ADL Task Force Issues Report Detailing Widespread Anti-Semitic Harassment of Journalists on Twitter During 2016 Campaign
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e659f29b-ae11-4dec-bf41-e49ed0e947be · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Online Hate and Harassment: The American Experi- ence 2023
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 350e5bd4-1ec6-429b-b67f-34892d7950fc · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Aunties, Strangers, and the FBI: Online Privacy Concerns and Experiences of Muslim- American Women
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef305288-9568-414e-bd57-75a81094b2ab · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Google’s Jigsaw was trying to fight toxic speech with AI
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a0f5ff9-abec-4355-a5b8-1cdd73640df6 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Robust Hate Speech Detection in Social Media: A Cross-Dataset Empirical Evaluation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 085a2934-d5cf-43a2-977e-d98ce24af45b · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns METEOR: An Auto- matic Metric for MT Evaluation with Improved Correlation with Human Judgments
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20622d68-2f80-4d1b-bf9d-8f72e7490f70 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns The Pushshift Reddit Dataset
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2799bf0-0fce-4029-a73a-5d097b6c6e68 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Nuanced Metrics for Measuring Unin- tended Bias with Real Data for Text Classification
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7793d8ca-21ff-48cf-997d-30f68d8a965d · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns HateGAN: Adversarial Generative-Based Data Augmentation for Hate Speech Detec- tion
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8e5201c-059b-43ca-bd76-21bab6bfc547 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns John, Noah Constant, Mario Guajardo- Cespedes, Steve Yuan, Chris Tar, Brian Strope, and Ray Kurzweil
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d5aa832-e0aa-453f-b517-7e8ed0cccfec · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Hate is not Binary: Studying Abusive Behavior of #GamerGate on Twitter
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a93cd6a6-24b0-4a66-8eb4-3a6261c457f8 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Christiano, Jan Leike, Tom B
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2df97a1e-c6ba-4887-bd58-49efbe1fe76c · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Hate Campaign
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 99b942b8-788a-41f9-a1d3-8136f5181f24 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Free dolly: Introducing the world’s first truly open instruction-tuned llm, 2023
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6bc594ad-b278-462e-b48c-96687bb8aec7 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Toxicity in ChatGPT: Analyzing Persona-assigned Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28de4e90-701c-4e55-b3dc-17e14de7a540 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns BERT: Pre-training of Deep Bidirectional Trans- formers for Language Understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation eea97b84-2711-4c36-8827-81c04e587bf6 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 613a99d5-4c1b-4084-b042-e22588cb4e20 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Paraphrase a text
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 76ba4c38-f8d0-461e-b25b-b4f8cff0cc8f · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Guide: Large Language Models-Generated Fraud, Malware, and Vulnerabilities
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 04699b60-c2af-4c3b-b9f2-8c5f1663c1fe · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Black-Box Generation of Adversarial Text Sequences to Evade Deep Learning Classifiers
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c7948ace-3c0f-4252-a6a9-4621348310dc · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns The Pile: An 800GB Dataset of Diverse Text for Language Modeling
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7352fa66-8c9f-435c-a322-6dd973b7e909 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8a497de1-fa71-4e2f-a6e4-cbf5a18d9def · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Hancock, and Zakir Durumeric
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3fe4fc75-27d8-45af-ab7c-d413cf709c09 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c02e3e85-f651-41d7-aad1-4cec723a7953 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Hess, Kelley P
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cdf58a07-63fb-49cd-8294-29708f51959c · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Deceiving Google's Perspective API Built for Detecting Toxic Comments
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7bd1e3f-72df-40d0-b70b-8a873d84cdfb · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Adversarial Example Generation with Syntactically Controlled Paraphrase Networks
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 867775fb-fa87-4bb4-ab05-25d298584f0d · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns High Accuracy and High Fi- delity Extraction of Neural Networks
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 18b20293-f9a8-4252-9dcc-d9f66218911d · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Is BERT Really Robust? A Strong Baseline for Natural Lan- guage Attack on Text Classification and Entailment
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68f67a60-d589-40c7-854f-bb23bb8592e4 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Toxic Comment Classification Challenge, 2017
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 850c24bd-9013-43db-a4cb-0eb5d49bbb12 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Jigsaw Unintended Bias in Toxicity Classification,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3560f2f0-ae31-4139-b05c-a2b323d259dc · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1119b65d-bc8f-44bd-80f4-5b0d50529dbe · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Content Analysis: An Introduction to Its Methodology
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fd7a6e8a-f462-44dc-b6ce-68e12e98ba1b · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Parikh, Nico- las Papernot, and Mohit Iyyer
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation af4b1e44-f981-4f32-92e8-7b956d279b36 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns TweetBLM: A Hate Speech Dataset and Analysis of Black Lives Matter-related Microblogs on Twitter
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5ba05853-35ad-47aa-80bb-f9575da20453 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns TextBugger: Generating Adversarial Text Against Real-world Applications
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4d76440f-0804-46ca-aff2-72fd0895ebf8 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd03df65-d1a3-4b55-95ec-c71b069e6fb4 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns A Holistic Approach to Undesired Content Detection in the Real World
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0a6cefe5-9337-45eb-991d-8652c1d33b41 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns HateX- plain: A Benchmark Dataset for Explainable Hate Speech De- 14 tection
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 515c39ea-1658-4802-a2eb-f09dec649daa · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns AI Trained on 4Chan Becomes ‘Hate Speech Machine’
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fbfdfdfb-1768-4df7-a7e6-fe61556966a1 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Mazurek, Florian Schaub, and Elissa M
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 76e81ca1-999c-42c7-93f7-be9c5dde4df2 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns The Challenge of Detecting Hate Speech
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation df5f1a68-3d31-4125-893d-ab986c018863 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 54135c04-8df8-43c4-97a4-01d6cffc24ee · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns What is hate speech
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4c2f3670-5515-45eb-8bde-5f11cd7172ee · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Handling Disagreement in Hate Speech Modelling
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e5466900-68dc-497c-b282-cf1c31e194ff · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns I Know What You Trained Last Summer: A Survey on Stealing Ma- chine Learning Models and Defences
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dd69ecf4-ebe1-418a-9767-1f520db815f3 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 566f0571-34f3-4fe8-8d60-856a0dd5e825 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Introducing GPTs
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation be5b8a04-8b95-4a9e-b688-049aef60b248 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns GPT-4 Technical Report
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 055fc79f-3b4f-4522-bd68-cddc1c2c34b6 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Offensive and Hateful Text Multiclassification
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4512db10-1eff-4c35-b2cd-94e30131da9d · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Facebook’s race-blind practices around hate speech came at the expense of Black users, new docu- ments show
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cc9834e9-ce0b-4570-978c-f1483673bc5a · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc3954aa-717b-4681-abc0-a62a2463d96d · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Gener- ating Natural Language Adversarial Examples through Proba- bility Weighted Word Saliency
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5d432944-2249-4239-95ba-b70f0ef1aa70 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Schuller
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 10176289-a493-4eb0-9720-25beae53e1c8 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Margetts, and Janet B
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5c772455-2cf9-45d1-aa32-3115a565f95c · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Sachdeva, Renata Barreto, Geoff Bacon, Alexander Sahn, Claudia von Vacano, and Chris J
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 574f0623-0853-4351-997b-924c24e0f4f9 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns TUBERAIDER: Attributing Coordinated Hate Attacks on YouTube Videos to their Source Communities
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3dd02bca-aeeb-470d-9dc9-b95fc2e2d14b · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Generative AI as a Vector for Harassment and Harm
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b43cb975-e63a-4db4-af32-ce6610a0568f · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Man who harassed black student online must deliver ‘sincere’ apology, renounce white supremacy
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c945fcff-def5-4e11-b494-4b0f1d8c8bb8 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Do Anything Now: Characterizing and Evaluat- ing In-The-Wild Jailbreak Prompts on Large Language Mod- els
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8cd9120a-fe98-49f3-845c-8e79e8e9c624 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns On Xing Tian and the Perseverance of Anti-China Sentiment Online
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 50ec7bac-64e6-4f6e-8608-02e913eced0b · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Model Stealing Attacks Against Inductive Graph Neural Networks
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation efd0ea7a-3324-4732-bbf3-3ddaf119f503 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Analyzing the Targets of Hate in Online Social Media
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4f3f5cb9-3a3b-48a8-877b-806ceca6bd09 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Learning to summarize from human feedback
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ef07ce3-c05c-4fd6-ac2a-f360ec628db1 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0c967b31-044a-49d2-a191-b7678e5d6db1 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Large-Scale Hate Speech Detection with Cross-Domain Transfer
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e1d50bf6-b164-47f5-9bf6-7978ac6476dc · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns LLaMA: Open and Efficient Foundation Language Models
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4913f0b0-d5dc-4d5d-8ad7-f434389b6bda · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Reiter, and Thomas Ristenpart
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a76556e8-7611-4441-91cb-99eca8b791e6 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Visualizing Data using t-SNE
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62f85124-fd18-4fc4-b752-6cbff76e2090 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Learning from the Worst: Dynamically Generated Datasets to Improve Online Hate Detection
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d3ad8727-8c8a-46d9-8470-645e194b289e · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Moderating New Waves of Online Hate with Chain-of- Thought Reasoning in Large Language Models
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a7abee6c-1f43-47e3-9d9c-7ef49454d838 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Vu, Alice Hutchings, and Ross J
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3ffcd2df-e5bf-4d0c-ba60-94956a9dc078 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns There’s so much responsibility on users right now: Expert Advice for Staying Safer From Hate and Harassment
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ee32c18a-1f36-4b04-b4aa-bc5d59673f41 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Challenges in Detoxifying Language Models
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10dcf881-889b-49a2-bdee-0601e963f804 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Not All Asians are the Same: A Disaggregated Approach to Identify- ing Anti-Asian Racism in Social Media
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cc1edeb2-f072-4d71-9e79-b69876745678 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Image-Perfect Imperfections: Safety, Bias, and Authenticity in the Shadow of Text-To-Image Model Evolution
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation afb446ee-7248-4722-8ea3-e0267bce9bd5 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Fight Fire with Fire: Fine-tuning Hate Detectors using Large Samples of Generated Hate Speech
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7d101714-7083-4b0b-9768-8b2465c0bf0d · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Baichuan 2: Open Large-scale Language Models
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e69bce8-4c74-4ab1-b41a-2832fbc0011e · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns GPT-4chan: This is the worst AI ever.https: //tinyurl.com/2s4jh5p4, 2022
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 854195b2-ce6d-410d-84fb-d210bef4997a · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns OpenAttack: An Open-source Textual Adversarial At- tack Toolkit
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 19b25f2e-afc1-4a23-8ff5-27052e604aa4 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns SecurityNet: Assess- ing Machine Learning Vulnerabilities on Public Models
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c9b65d46-ebc3-4c3d-990a-fc43901af425 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns OPT: Open Pre-trained Transformer Language Models
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb337bfa-4dfa-43ab-a3e2-8bd60141f990 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Generating Natural Adversarial Examples
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2d3426a3-2e04-4116-adc9-8c80586222b1 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04791f2d-d5d0-434a-8917-1add3bcf9edc · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns 2, 3, 5, 17, 18
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9cc12a26-8c36-4a77-8b09-b58c130cb60a · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Racism is a Virus: Anti-Asian Hate and Counterspeech in Social Media during the COVID-19 Crisis
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f44ee80e-a821-4bb2-b532-13a3b18b682a · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e8b5c314-e0f9-4f8d-847e-cd06e98960e6 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 988831f8-7a0a-4902-a5af-3abb1bb49485 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns Unresolved cited work
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 81c29158-d397-4a44-9abc-f140e901b128 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns identity attack
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c3846f5f-3496-42a5-8729-02c84077f5c3 · outbound
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns toxicity,
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 16bb6919-7b9a-4b14-bf1b-9cb3e6c48e0e · inbound
Are Today's LLMs Ready to Explain Well-Being Concepts? HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.