Pith. sign in

Paper Citation Record · LEDGER

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content

As of 8 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2509.12672.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.12672 v2

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T16:40:31.250299Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 618cc52e-9b6b-4065-b6ed-9e5c3d55ef56 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.116422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.116422Z digest=sha256:6d4377fe1150d700549cda5af61dfea4856ce437c574e9531cdabae176f948ad

Observation 46cdf308-c29a-430e-aa6a-821fff51889b · outbound

This paper cites write newline.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.120087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.120087Z digest=sha256:ad7452c7c30c9ce768453a2c3e7bd32f38c04530268329b7a5c3792ec370dc84

Observation d3614bcb-5916-431c-86c4-dcebb732c7e2 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.123713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.123713Z digest=sha256:e9bff5967d617df9c1d1e01e90caddb60ff216860e77a9857d35beac8aaef52c

Observation 6bc19f82-69af-4133-aee9-b35823587b2e · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.126880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.126880Z digest=sha256:551b4bb36f786e546245d490ac29cdd2c412c8ec245fc3cdd89912e94512f82e

Observation b74c5b7a-dd7d-4c18-81a7-a2f8c38e83cc · outbound

This paper cites Mechanistic Interpretability for AI Safety -- A Review.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Mechanistic Interpretability for AI Safety -- A Review

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.129847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.129847Z digest=sha256:8ce222c862705dfbd00c3fb34f9d23ad61c6b9ad5f3b1a42cdc41aea0391e758

Observation 072c8864-8fc0-4f1c-94cf-4981912c04ea · outbound

This paper cites Towards Building a Robust Toxicity Predictor.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Towards Building a Robust Toxicity Predictor

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.133163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.133163Z digest=sha256:acb8091f3db757ace06e3ed04aceac90dd2118916282e2fb03907cd791fc0d8c

Observation 247f0cce-2568-4a35-a002-e7f5d21532c6 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.136971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.136971Z digest=sha256:b2e1ca5a0cb1361460658fad55551e33f59bb9555ce335890009442a52785add

Observation 38b51520-2d7a-4a6f-9c57-3c6830720a52 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.140263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.140263Z digest=sha256:099bffba67ab87911a0ceb24309fb8acaa28aa9165554e2020d8806e0c42ea91

Observation 1390aade-ade4-4013-b84e-7deb1b97cd2f · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.143456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.143456Z digest=sha256:81a3b16252fdde7f2983fc96d0dc6fe01a66ce8be220fb30c68db3d680553b70

Observation a1e91bfd-b394-45b8-af01-0d6b5052dbca · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.146244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.146244Z digest=sha256:1ed0817cfc8641a23dd2e08acb1fb7c734b3cf6ee5810a64d28da3287f4ae546

Observation 7e2e582d-84cf-4b6f-af44-3536a0ba7fa2 · outbound

This paper cites Towards Automated Circuit Discovery for Mechanistic Interpretability.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Towards Automated Circuit Discovery for Mechanistic Interpretability

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.149122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.149122Z digest=sha256:ecd09195eb81975c6520d6ac1e1a9e91944d4ad4402b96c65b8e7bbd7361b8e8

Observation 1133c7a5-f396-47a1-9fb7-8b5b823790ba · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.153001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.153001Z digest=sha256:472941ff11b15524fedecd1d4370d4c005902dd9e65fa0af85adcb532170cbab

Observation 8163cfac-a653-45b1-8e2d-aaa7b94838d6 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.156541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.156541Z digest=sha256:6cb606368f73f3bd35957824b3f2463f50406a101b6c425961dba21132b81e87

Observation a59922d2-4c9d-4813-8857-b2829dc5399b · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.159467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.159467Z digest=sha256:c10be98a64a5576c6092cddc961766b6503c7e45f38868817bea68f4f73ddb4e

Observation 5aaad299-99cd-4808-9824-4eb97b536c99 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.162193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.162193Z digest=sha256:521d7f97c9e18f082180fa528a736fd5bca6dfa3d90397a1942b4b7442f7a49d

Observation 140d4b5f-71ba-4710-b87a-debafd5c72a4 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.165692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.165692Z digest=sha256:2aea2cacbb8cfe270d1de944554567a08ee9345e03da94e4614d350a4e41ef64

Observation af396301-d35d-4ef6-a04d-098c5fb597c9 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.168451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.168451Z digest=sha256:e2499d11e3641953b8b970824f1b7aa9aa055f6a4e2f97daa658b7846988f534

Observation 94337268-ddd3-4ed5-bd4a-6ba7e201cbd3 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.171648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.171648Z digest=sha256:780d94d97e4bab00849625d9c04bea397dc61d5809e5b6925e3392e9da818ef3

Observation 26eea721-1a69-4a69-a7b5-ac85284641c2 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.174811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.174811Z digest=sha256:9b8fe2ddd1aac6d327276fc97b13c17b62f994c773541fda9b0cc526693e7af5

Observation 50d3bb49-f4db-49aa-92d6-7022dbbacd81 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.178582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.178582Z digest=sha256:f20eeab233523827024a5af9c51324832adee8fcfb1b7af1f70ffe60440cf5e0

Observation cc9cd95e-1d06-44e1-9fe2-c423bae08b88 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.181797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.181797Z digest=sha256:61242c8c67db9729cb2dc922fc9ef62b67176cf5b4813df28348102875d07beb

Observation cf8c2438-d63d-4082-803e-d1245e896c3f · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.184837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.184837Z digest=sha256:32a1bbfd3ed5da329a008f8a2ec821f700571bd31a8cd4bd3aaf0faa586b9a65

Observation a04ecd64-ce82-4bcb-b222-78843b17a1a3 · outbound

This paper cites A New Generation of Perspective API: Efficient Multilingual Character-level Transformers.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content A New Generation of Perspective API: Efficient Multilingual Character-level Transformers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.187575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.187575Z digest=sha256:a543619716ebb4559c8221556c7ab976d89be6a36b8d2d0a9b6de89610a99c5b

Observation 63aee8c0-e72c-4077-a558-d254dc22d77c · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.190708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.190708Z digest=sha256:82fb7d8d32660e0c673594bcc0d55df6be66efce046c8c9cde1f3342154c4a04

Observation 1a99733c-38c0-447e-aaf8-0988630fe711 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.193428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.193428Z digest=sha256:8e2f51069d5cc999dbcf0a21977b75452bc826d3283186bf5eae4ee1e10e45fe

Observation c8773139-94a8-4f6e-a960-35935bf51f31 · outbound

This paper cites D.; and Finn, C.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content D.; and Finn, C

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.196217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.196217Z digest=sha256:beb08cc780f5769b1526b3756cfe140571efd477166c19146738993200eaaf4d

Observation b20839b2-68a0-4c99-9fbe-5a70fe119bb5 · outbound

This paper cites R.; Li, G.; and Crespi, N.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content R.; Li, G.; and Crespi, N

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.198929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.198929Z digest=sha256:bb32e9ce212e580f1095a58a7b704efcbfe332cfcf65db844097fd4828ab4892

Observation 1ff28abf-a7b1-45b6-8beb-025ff87e9613 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.202127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.202127Z digest=sha256:1adf86122738a1992897f0cebce59a0908d0fdaf18e5598ad104e54d619a1268

Observation 8b40be0b-e29d-4760-b373-2f63355c226f · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.204983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.204983Z digest=sha256:521fc59d800a4230338f59215b7ea6c3aa87a343978b9c54adab716f5ab42e2f

Observation ad8a4e2f-5d17-4a7f-9ef6-42f3e805f4a1 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.207748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.207748Z digest=sha256:9c5ff21307696aac5797557dded0b1a6beb877c757b303cd7c9759d721655d69

Observation e1b7ddb6-f539-44e8-a6dd-12d2629bb605 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.210447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.210447Z digest=sha256:dd37999843db3f62156dbbbeab7a06b1d4f0598f8d3e3bc7cd5e5681072237dc

Observation 293bf6f3-d479-4531-a742-f2abe67f3e14 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.213577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.213577Z digest=sha256:56a8a3ecc0fed2722262c578c7b2e4c0758cc40f9018d72a09f45396a1b228e0

Observation f0ec7491-8aa0-4f1d-8f17-166cf60effcd · outbound

This paper cites Token-Modification Adversarial Attacks for Natural Language Processing: A Survey.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Token-Modification Adversarial Attacks for Natural Language Processing: A Survey

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.216119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.216119Z digest=sha256:f65b201700ee77b169570f9d0b00bde3c947d021b4df275c983345178aec8f63

Observation cce6712f-334d-4d41-9a63-1f8030cf7501 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.219115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.219115Z digest=sha256:16fa1b248b24080e1ff9271ef795c061772e09b55f2f3df98fc6c942a2b87a33

Observation 8aaaba9b-adf5-48c8-a93f-43ac7d596243 · outbound

This paper cites o ck, F.; and Wagner, C. 2021. “Call me sexist, but.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content o ck, F.; and Wagner, C. 2021. “Call me sexist, but

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.222044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.222044Z digest=sha256:19d4f37700bec58aee6fa4de4a0bf6701353f60a8b625e993c49a116e9b03d44

Observation a2ae48b7-6b5d-4803-9419-1809e0dcd2f0 · outbound

This paper cites HowkGPT: Investigating the Detection of ChatGPT-generated University Student Homework through Context-Aware Perplexity Analysis.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content HowkGPT: Investigating the Detection of ChatGPT-generated University Student Homework through Context-Aware Perplexity Analysis

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.225010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.225010Z digest=sha256:96d58f490723872b05d2a789ac09d5b1f0a0c5218b04c4e665ddd4fc14e9e95d

Observation 86ca1c4e-0812-4f9a-83dd-4a60ca6afa21 · outbound

This paper cites Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Enhancing Adversarial Text Attacks on BERT Models with Projected Gradient Descent

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.227672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.227672Z digest=sha256:eb31ef84be2d5734ce81f21f516481b7ff8e8281c47e7d4c8e57261129d643ef

Observation 4b27728d-2133-4603-9263-342399895f24 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.230569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.230569Z digest=sha256:87a80923aad38bd4b86f49ba7a21625086859f47c9928b7c3b5d20fae9b786b1

Observation e72fdbae-b950-4752-92da-431093f8e751 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.233207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.233207Z digest=sha256:1d0a803eba017f38f29a4012ee7d95d34a4097a9fd5970c8911f36a0a0f25b77

Observation 7b88280b-d6f3-4674-a4ee-071c97d11a47 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.235984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.235984Z digest=sha256:e15e3f8556948e62145fac98cf47292ba5d821669ac4f03e6cd7a4e645164581

Observation 25799df4-b1cf-4a0a-8220-d889315bb4c5 · outbound

This paper cites S.; and Wong, D.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content S.; and Wong, D

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.238745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.238745Z digest=sha256:7a8988cc19cf5dd6194b3ca45cae35cd769cf546f852064e0921ec1a92b2d5d8

Observation d9849187-ef1a-42d2-ac4e-7c51ab90eeae · outbound

This paper cites S.; and Wong, D.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content S.; and Wong, D

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.241698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.241698Z digest=sha256:890224ce16a028bbd48105fcc1fa74dc2ca96e6bf79387b6b8adeb46567778c7

Observation 1fd0ebd5-059c-4277-aa79-70989e890b6a · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.244420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.244420Z digest=sha256:7cba116c284175ba609b1831e58272beec344638e8a5f819d5ee3c577122010b

Observation f925dcdf-a08b-4a7c-8316-fda551ff1c15 · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.247794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.247794Z digest=sha256:e8bd53b9bfe19ab43f37726c9c306470a734adf2b83cc0df4cd4a5a5ecd2287c

Observation cc5dfd60-e7b4-4b2a-9700-f45348ff631c · outbound

This paper cites an unresolved cited work.

Towards Inclusive Toxic Content Moderation: Addressing Vulnerabilities to Adversarial Attacks in Toxicity Classifiers Tackling LLM-generated Content Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T16:40:31.250299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:40:31.250299Z digest=sha256:eafe23f6acabc62c4d014139da28e54bb085fb3f6a118dab4df6c9f136d24bdb

Pith citing papers

No inbound Pith citation observations are available.