Pith. sign in

Paper Citation Record · LEDGER

Low-Resource Languages Jailbreak GPT-4

As of 22 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 88 inbound Pith citation observations for arXiv:2310.02446.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.02446 v2

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-17T09:24:13.911401Z

measured 143 of 143 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 88 of 88 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T06:05:12.200545Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact35
  • verified fuzzy14
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch5

External citation measurements

20
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 5e72d20d-cb82-4ea1-89b3-dc50d883d2a1 · outbound

This paper cites Jigsaw multilingual toxic comment classification.

Low-Resource Languages Jailbreak GPT-4 Jigsaw multilingual toxic comment classification

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.160455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:9c3e8e8beb761d1ad2151129bda02d435d667eaa0b7f2df3e48f4eff0462d5b7

Observation e9a0fc17-468f-4c68-bd18-7f1c68c99cce · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Low-Resource Languages Jailbreak GPT-4 Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.051786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:8c954090f13bdd0002c90aea370985849b343c25c7f586e47b8dca6c35bb59fa

Observation 0c7bde94-fae6-4b83-a86c-7ea8ff575dd3 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Low-Resource Languages Jailbreak GPT-4 Constitutional AI: Harmlessness from AI Feedback

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.056610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:978dde2dadfe06da8c5eaca7a52baf29e8c2d55b588c601f88815f06b37d6cd8

Observation 6b329777-af2e-41bb-a5a0-df452fbdbe7b · outbound

This paper cites A Multitask, Multilingual, Multimodal Evaluation of ChatGPT on Reasoning, Hallucination, and Interactivity.

Low-Resource Languages Jailbreak GPT-4 A Multitask, Multilingual, Multimodal Evaluation of ChatGPT on Reasoning, Hallucination, and Interactivity

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-17T19:58:48.184559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:1330fdb91bee1e15cfec92401df3aab71f27416a6d0da5882856be66ace5b049

Observation b5fb6dfc-815c-4b8a-ac26-79100018e9ff · outbound

This paper cites Building Machine Translation Systems for the Next Thousand Languages.

Low-Resource Languages Jailbreak GPT-4 Building Machine Translation Systems for the Next Thousand Languages

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.067395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:adcf4aca335439251f6a1ba41f5f67efd4154dbe7730e52376de0a59e07ba231

Observation 219cdd08-98ae-4c38-8b7d-cf470028a56c · outbound

This paper cites SeamlessM4T: Massively Multilingual & Multimodal Machine Translation.

Low-Resource Languages Jailbreak GPT-4 SeamlessM4T: Massively Multilingual & Multimodal Machine Translation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.072280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:9e65beb77f387512c5f0b3e8a7e114e5ecfde468bbde6c54f3b0529dc8c94a95

Observation 13f29bdb-d3c3-4d2b-a314-aae22063b8a2 · outbound

This paper cites Systematic inequalities in language technology performance across the world’s languages.

Low-Resource Languages Jailbreak GPT-4 Systematic inequalities in language technology performance across the world’s languages

Reference 7

Resolution
verified exact
doi, observed 2026-05-17T09:24:13.988038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:335553988f93a6a1eca4e09458b8bd0f5228cdfe9fabd366afa785e5fd8531cf

Observation aaa59a1f-99e0-40dd-96fc-eb6c8f354cf4 · outbound

This paper cites Are aligned neural networks adversarially aligned?.

Low-Resource Languages Jailbreak GPT-4 Are aligned neural networks adversarially aligned?

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.078179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:a5f403fb1234a1d059c67a685369895e16b33fd47b3f8ca28d04f5690459e7d5

Observation b01ba6a5-539a-49b6-ac68-ceef593ea4db · outbound

This paper cites Aim.

Low-Resource Languages Jailbreak GPT-4 Aim

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.195909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:b38197df04f398a93efcb6781a97139841659afc4c7418331f96d99aeda7504b

Observation 8b8efac4-acfa-4301-ba0d-53d29aa78aee · outbound

This paper cites Translatorbot.

Low-Resource Languages Jailbreak GPT-4 Translatorbot

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.156570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:d7cd95e9ba0c9a3eb832a5d81a22ddc16c51c91fd22df8b658c1594d28c9201a

Observation f591f313-3f61-4f47-a7e7-f7061bb8b0d7 · outbound

This paper cites How is chatgpt’ s behavior changing over time?.

Low-Resource Languages Jailbreak GPT-4 How is chatgpt’ s behavior changing over time?

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.152934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:fa1789f521f40a4aca512176707789746b3f3198fd9d5912ea76fc48022d6346

Observation 7a7bdd74-4784-4e57-8152-a089145dcfa0 · outbound

This paper cites CONAN - CO unter NA rratives through Nichesourcing: a Multilingual Dataset of Responses to Fight Online Hate Speech.

Low-Resource Languages Jailbreak GPT-4 CONAN - CO unter NA rratives through Nichesourcing: a Multilingual Dataset of Responses to Fight Online Hate Speech

Reference 12

Resolution
verified exact
doi, observed 2026-05-17T09:24:13.977286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:9808ba2ad7682f6e33afc6b263f4ce9cc8ac9816ae11f87d82f052dbfdf1d827

Observation 18cc0caf-5821-472c-80b6-9923e21a9ac2 · outbound

This paper cites Language support.

Low-Resource Languages Jailbreak GPT-4 Language support

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.202322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:220d0e38143b827370fdccfb7d1806da98da421372c0b9956e9f09e5d1cb2cc8

Observation e62dc9e1-ae2c-4462-b764-cd3397c9b422 · outbound

This paper cites No Language Left Behind: Scaling Human-Centered Machine Translation.

Low-Resource Languages Jailbreak GPT-4 No Language Left Behind: Scaling Human-Centered Machine Translation

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.083503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:1ab8d0a3871d81deca85e5abea95705af615c73de9e173f468fbc5082af70bcb

Observation 04d3ddc6-f5a7-4f7b-bdf8-64de9578d4e5 · outbound

This paper cites Multilingual Jailbreak Challenges in Large Language Models.

Low-Resource Languages Jailbreak GPT-4 Multilingual Jailbreak Challenges in Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.088763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:3b5bda76efd8e5fbb7274e04495d7f7ff7884c377ff741ca844871a6915030ba

Observation d61f5fa9-4e48-4a36-bd26-410023f9f859 · outbound

This paper cites Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned.

Low-Resource Languages Jailbreak GPT-4 Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.093690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:fe95d327ac04e3bbde0895581e96eaa6d57f9f3fc340361ce4c0e514254ad927

Observation ca8d4468-837a-4cf2-9522-c86a72686373 · outbound

This paper cites The Capacity for Moral Self-Correction in Large Language Models.

Low-Resource Languages Jailbreak GPT-4 The Capacity for Moral Self-Correction in Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:13.993100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:7b1112eb7d4ac388f9a566d5a3779f5c909a8f17763d214a2e0303731260e8f6

Observation 2b9bd19b-c2ff-43e8-bece-f59a28d16c28 · outbound

This paper cites RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models.

Low-Resource Languages Jailbreak GPT-4 RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.001152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:eb271f64e73487d7ba52eb082c45a1bf7321634ef24dfaffcf367d0070979a0c

Observation e7b74f7a-3319-40b6-9828-ee47aa8e1a69 · outbound

This paper cites ChatGPT Perpetuates Gender Bias in Machine Translation and Ignores Non-Gendered Pronouns: Findings across Bengali and Five other Low-Resource Languages.

Low-Resource Languages Jailbreak GPT-4 ChatGPT Perpetuates Gender Bias in Machine Translation and Ignores Non-Gendered Pronouns: Findings across Bengali and Five other Low-Resource Languages

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.006742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:d8c683772f3d530430b3776b204490711258659be0971f0aba9073132517f86e

Observation 0e89f9f4-fe39-4c91-8a25-cee6e716fd93 · outbound

This paper cites How Good Are GPT Models at Machine Translation? A Comprehensive Evaluation.

Low-Resource Languages Jailbreak GPT-4 How Good Are GPT Models at Machine Translation? A Comprehensive Evaluation

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.011839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:30f3c4759eba23103db296d2f6a510f1103b3d1874c95e46426021c94bf1a44d

Observation 0c0b38fe-4593-4552-94dd-9939bdb963cb · outbound

This paper cites CBBQ: A Chinese Bias Benchmark Dataset Curated with Human-AI Collaboration for Large Language Models.

Low-Resource Languages Jailbreak GPT-4 CBBQ: A Chinese Bias Benchmark Dataset Curated with Human-AI Collaboration for Large Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.017094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:c0a4b2e25fa4ea5e087abfc501c39d584b4fa2250168744e5c61af5fd1fdbc76

Observation cc8c84f2-b9fd-4b84-a239-02149dd89211 · outbound

This paper cites Adversarial Examples for Evaluating Reading Comprehension Systems.

Low-Resource Languages Jailbreak GPT-4 Adversarial Examples for Evaluating Reading Comprehension Systems

Reference 22

Resolution
verified exact
doi, observed 2026-05-17T09:24:13.958411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:e2c19677bce953a958c050e306be0a3f8cb17ec31184aaf70e5056be1862e0aa

Observation 04bd32f9-f9f1-4d67-a811-b4c28128b5be · outbound

This paper cites Automatically auditing large language models via discrete optimization.

Low-Resource Languages Jailbreak GPT-4 Automatically auditing large language models via discrete optimization

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.188704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:82da885b364313393aa22dafd591a66cdb666285165826bae73af5bfc777dda8

Observation 131b6eb3-5c38-44e4-a3ea-c702abe7d8a6 · outbound

This paper cites In: Zong, C., Xia, F., Li, W., Navigli, R.

Low-Resource Languages Jailbreak GPT-4 In: Zong, C., Xia, F., Li, W., Navigli, R

Reference 24

Resolution
malformed identifier
doi_truncated, observed 2026-05-17T09:24:13.966278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:544567be20f689fe89d7a35323c76a9102894f5a6fd0eb4a799f1c67e4c628dc

Observation 598b2da4-21a7-49c8-a406-61cd6a4a7fe3 · outbound

This paper cites ChatGPT Beyond English: Towards a Comprehensive Evaluation of Large Language Models in Multilingual Learning.

Low-Resource Languages Jailbreak GPT-4 ChatGPT Beyond English: Towards a Comprehensive Evaluation of Large Language Models in Multilingual Learning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.021973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:97ec3ae08bc6f6f1aa5fea2ede3656106addd53997f964d145366682b8d5742d

Observation a9e0449a-d2cf-4d77-a059-af389bc5cba1 · outbound

This paper cites Open Sesame! Universal Black Box Jailbreaking of Large Language Models.

Low-Resource Languages Jailbreak GPT-4 Open Sesame! Universal Black Box Jailbreaking of Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.027167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:0a0a290bd3ae05cfa63c47b8b79194c5131e6196e6b17cd20a3e36cf57eb09e8

Observation a5b85d78-6265-42c0-9c07-e5efde7decc1 · outbound

This paper cites RAIN: Your Language Models Can Align Themselves without Finetuning.

Low-Resource Languages Jailbreak GPT-4 RAIN: Your Language Models Can Align Themselves without Finetuning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.031548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:2578797a3daff49c3f77ac26f520269b0e6959a0bdea37dd117ca0c10f2e7a10

Observation 483431eb-9134-4d95-a57c-18e5efb866bd · outbound

This paper cites Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study.

Low-Resource Languages Jailbreak GPT-4 Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.036520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:598d3de120f5cb70b200e9465e583280e373b088271d8ad5cafa2d10a65368a1

Observation 46f48306-cf34-4f9e-88dd-044c93055af9 · outbound

This paper cites Black box adversarial prompting for foundation models.

Low-Resource Languages Jailbreak GPT-4 Black box adversarial prompting for foundation models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.184738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:14b5b0575cecb67de251ca2788459b6db346ee233d036ec13c3bea331e90f8df

Observation 9d73e60d-b501-4dee-a32b-1e4db372e2c8 · outbound

This paper cites Mitigating harm in language models with conditional-likelihood filtration.

Low-Resource Languages Jailbreak GPT-4 Mitigating harm in language models with conditional-likelihood filtration

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.042036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:1defce56460e5f65fb95d5ae9148b20dc8ffd755404fcb6ad17292cd11c534b9

Observation b99b48e9-72ed-4ed6-91a2-21029ab93649 · outbound

This paper cites Duolingo.

Low-Resource Languages Jailbreak GPT-4 Duolingo

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.164588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:cde46698f80b798a5e235c980a5576d42a4fb32afd896319734b82b805d356ae

Observation 81e4e847-266b-464c-8427-c6cc4a671b6f · outbound

This paper cites GPT-4 Technical Report.

Low-Resource Languages Jailbreak GPT-4 GPT-4 Technical Report

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.098278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:5e1bd23f76f064ba147894706aecf4659f945a31dbfa45adb4c14ea935c5d869

Observation 8ad5a889-6ccd-40e0-a9b0-9d196c7f40ff · outbound

This paper cites Government of iceland.

Low-Resource Languages Jailbreak GPT-4 Government of iceland

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.172750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:f7d3dbaa5deaf656d2156752c331d6fd107dfb9879e26ca5a6b771f3c17b362a

Observation 12a17a2d-2c7d-4e80-9324-431856a8b931 · outbound

This paper cites Training language models to follow instructions with human feedback.

Low-Resource Languages Jailbreak GPT-4 Training language models to follow instructions with human feedback

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.176969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:e3682577b891570875a429b50aab73476ceac25786b600afbd1f24a5847b1289

Observation ac28dea0-70dc-434f-8893-d5c0546ece81 · outbound

This paper cites Practical black-box attacks against machine learning.

Low-Resource Languages Jailbreak GPT-4 Practical black-box attacks against machine learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.180901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:376a398b7319557550be7addc4229ebce2592d1bf2e16a4d4be103f12c085ebb

Observation 0787857b-a8df-42d5-bf61-374879735b70 · outbound

This paper cites Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , year =.

Low-Resource Languages Jailbreak GPT-4 Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , year =

Reference 38

Resolution
metadata mismatch
doi, observed 2026-05-17T09:24:13.984406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:e1786880defa7613e1be4c2df703ec181c8831727ba4f9dba801a26bee4f6639

Observation 01370b94-dcf8-41bc-9c85-7ee731412362 · outbound

This paper cites Square one bias in NLP: Towards a multi-dimensional exploration of the research manifold.

Low-Resource Languages Jailbreak GPT-4 Square one bias in NLP: Towards a multi-dimensional exploration of the research manifold

Reference 39

Resolution
verified exact
doi, observed 2026-05-17T09:24:13.962385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:8e93ecab9c0b7fd78604aee1340604d774016157ee0cdc21a8fddcdc7c8e1558

Observation 60843afb-dff2-4f62-8c60-270121e8c023 · outbound

This paper cites Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models.

Low-Resource Languages Jailbreak GPT-4 Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.103926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:f375f3418beb332af96edaf85c193adfbb20e2832a659089879cf58f6cbd5920

Observation 7917c3f2-b8f1-4ea4-b9b0-9297702edb8b · outbound

This paper cites "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models.

Low-Resource Languages Jailbreak GPT-4 "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.108875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:9dbde0fc41655c6d36689dd698f0351c70aa5b0bd84afa4d4fca7002adcdd044

Observation 66951f9c-76a9-4887-ae4c-e98a35690c26 · outbound

This paper cites Why so toxic? measuring and triggering toxic behavior in open-domain chatbots.

Low-Resource Languages Jailbreak GPT-4 Why so toxic? measuring and triggering toxic behavior in open-domain chatbots

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.199218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:99dac69658a017cc469ecbf7f7d49a543ad5aff229df27cc66cc18328a6f1510

Observation acd2ece7-94dc-45ec-8189-b05f8640fc37 · outbound

This paper cites Universal Adversarial Attacks with Natural Triggers for Text Classification.

Low-Resource Languages Jailbreak GPT-4 Universal Adversarial Attacks with Natural Triggers for Text Classification

Reference 43

Resolution
verified exact
doi, observed 2026-05-17T09:24:13.980555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:0516dcddd4f6e9a55412450d3082ceea1aeb7c4bd6ab4f096eeec73f79359582

Observation 32623e82-77d5-40db-b16d-df78622b7075 · outbound

This paper cites Smith, and Luke Zettlemoyer.

Low-Resource Languages Jailbreak GPT-4 Smith, and Luke Zettlemoyer

Reference 44

Resolution
verified exact
doi, observed 2026-05-17T09:24:13.973601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:fecdd638dc67b24b4c5efd31e9fbc3507cc20f2833b427e27e53182a5edc2a5f

Observation 627295d0-7e0d-414f-bf2a-5336c4321232 · outbound

This paper cites ChatGPT is not a good indigenous translator.

Low-Resource Languages Jailbreak GPT-4 ChatGPT is not a good indigenous translator

Reference 45

Resolution
verified exact
doi, observed 2026-05-17T09:24:13.969936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:85cdc83482988b78c8b9a84fca7dfa2052ef3d07128bbef8c31cc83b4c93271b

Observation 78c6ff0a-89ac-4663-8bcd-12509abdb6fc · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Low-Resource Languages Jailbreak GPT-4 Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.113933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:fcf41ce129c644f9d9c5919927d63a0ee8e08d10e1888edd28a44fd4d848e160

Observation 0ff6e151-ef5f-4007-ab47-ae345e7a2c26 · outbound

This paper cites Translated unleashes full gpt-4 potential for businesses operating in languages other than english.

Low-Resource Languages Jailbreak GPT-4 Translated unleashes full gpt-4 potential for businesses operating in languages other than english

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.168804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:314f3bf0ed25867de783e8bfe45be013f9145ae7f9428892c2206176847782d0

Observation 72278891-c65c-46c9-9701-4a09aece17b0 · outbound

This paper cites Universal adversarial triggers for attacking and analyzing NLP.

Low-Resource Languages Jailbreak GPT-4 Universal adversarial triggers for attacking and analyzing NLP

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T09:24:14.192562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:750afd49a4db0bdb2439111b3c792fe3c093bb16c35bbf0a633fa1d78c6f31d2

Observation 8c99a8d6-11a8-41ac-a361-9c026a2860d2 · outbound

This paper cites doi: 10.18653/v1/D19-1221.

Low-Resource Languages Jailbreak GPT-4 doi: 10.18653/v1/D19-1221

Reference 49

Resolution
verified exact
doi, observed 2026-05-17T09:24:13.952968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:f17d5ea7b0693f6fcdef1f668806502d65935e22beeb72b54f84f450cf791364

Observation 0399a3eb-8607-4f73-8dc2-d8bbe1c13c06 · outbound

This paper cites All Languages Matter: On the Multilingual Safety of Large Language Models.

Low-Resource Languages Jailbreak GPT-4 All Languages Matter: On the Multilingual Safety of Large Language Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.119258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:52ae3d6665a0a6143b265be911d59ee4d14b9f5e777a90e199bf6cb37f7de467

Observation 49da06db-1474-4a09-96ad-1dec12bda0b5 · outbound

This paper cites Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs.

Low-Resource Languages Jailbreak GPT-4 Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.124720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:c25754e498204ef712570155cc2b00bc85f41a82e5ff89ae1c6cc2c466859ec0

Observation 1ccc29c5-7f42-49ea-96e3-e2803bffaaea · outbound

This paper cites Jailbroken: How Does LLM Safety Training Fail?.

Low-Resource Languages Jailbreak GPT-4 Jailbroken: How Does LLM Safety Training Fail?

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.129455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:cf5a5f3c4f785356153e469ab1811c01e7aff956e62aaff68266a447d1278500

Observation 5a10364c-fc48-4275-aaf5-b3028be8ea66 · outbound

This paper cites Ethical and social risks of harm from Language Models.

Low-Resource Languages Jailbreak GPT-4 Ethical and social risks of harm from Language Models

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.134418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:5adc111636b5f4999e7f98c5bd70f0eac6b1a863a640862446954cb805bb7636

Observation f7fbf205-4ce9-41bc-af92-228ed6990f92 · outbound

This paper cites Prompting Multilingual Large Language Models to Generate Code-Mixed Texts: The Case of South East Asian Languages.

Low-Resource Languages Jailbreak GPT-4 Prompting Multilingual Large Language Models to Generate Code-Mixed Texts: The Case of South East Asian Languages

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.139522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:055896da2163fa99d5fcaa941b388a61a7f39ce7df47c980f134bc1683477b6d

Observation f61039f2-9233-45ab-8844-f48251c0a235 · outbound

This paper cites GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher.

Low-Resource Languages Jailbreak GPT-4 GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.144341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:67173c00d04dce6ae9ddcb12c4797cab7cf0d8c108c7a6fdf986a2ae7db5e2cc

Observation 560fcd23-4622-413e-ae6f-a792a3e6af88 · outbound

This paper cites Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity.

Low-Resource Languages Jailbreak GPT-4 Red teaming ChatGPT via Jailbreaking: Bias, Robustness, Reliability and Toxicity

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.149204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:5994258b1c54419e176ca7fb9d593002b4eaaf9add9a2993413d8a3497fe505b

Observation 8bc99e31-5022-4f5d-8b1b-659d0ef8fd2e · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Low-Resource Languages Jailbreak GPT-4 Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:24:14.046876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:b277925b47672079bc9f56b70f5b76ac8be2e2738e68b801464013a9db212534

Pith citing papers

Observation 5c98929d-04de-423b-b2d4-a0d746e7a400 · inbound

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks cites this paper.

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks Low-Resource Languages Jailbreak GPT-4

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T17:11:00.639293Z digest=sha256:8d8892ceb8dc2a20258929a66f6b7f4bfc684a1e071b0f92c3918c55216063c2

Observation e8d8041c-31e1-4aaf-b901-14697704f486 · inbound

A StrongREJECT for Empty Jailbreaks cites this paper.

A StrongREJECT for Empty Jailbreaks Low-Resource Languages Jailbreak GPT-4

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T21:28:02.745230Z digest=sha256:a16e26a544295d32a3499c7f250e4deacc657cca0b73381e88868630d262c49e

Observation a0946d21-1906-4376-9f24-de60bb06111f · inbound

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models cites this paper.

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models Low-Resource Languages Jailbreak GPT-4

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-15T06:08:05.386345Z digest=sha256:9c3245dc61bb29182c86b4583f14443be740575ee510b22ab5c5eed1a6cc8e0b

Observation cb10f105-f8c0-45e0-9d87-63a9a4e5c86b · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Low-Resource Languages Jailbreak GPT-4

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:0d7e477ff44f4fa6a2f93448b84d3b8b3b53b72f8c39e45289a7c1a5ccd4c4c9

Observation 6db8d5bf-bcd9-4c2a-92b4-6db3aa4c1d04 · inbound

AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents cites this paper.

AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents Low-Resource Languages Jailbreak GPT-4

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-14T01:35:50.992477Z digest=sha256:8f00f750de609b4807dac0279d31de9f10d41e7a45597f38bd37ae2a137548fc

Observation 325a16ab-919d-4ae6-bb93-aa886cbbac89 · inbound

RV4Chatbot: Are Chatbots Allowed to Dream of Electric Sheep? cites this paper.

RV4Chatbot: Are Chatbots Allowed to Dream of Electric Sheep? Low-Resource Languages Jailbreak GPT-4

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T15:18:19.493168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:18:19.493168Z digest=sha256:3377c54b952765afa7c9eedbba633af778f12aea197735be5dc8b10c70f6934d

Observation 05ad4eab-bcea-4e40-bfd1-acdf4810ba78 · inbound

Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages cites this paper.

Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages Low-Resource Languages Jailbreak GPT-4

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T04:53:50.564436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:53:50.564436Z digest=sha256:71ce4810e6b0933843bfc8bde15411ca3dc6e7e61f7a72b186d435faa7d5f9d3

Observation 60f55bd8-5b5c-41fb-b07d-f676f96e7d5e · inbound

Aya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier cites this paper.

Aya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier Low-Resource Languages Jailbreak GPT-4

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T21:38:38.168765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:38:38.168765Z digest=sha256:24f01b5257fb069d90ef8db2a98e791cd8ab32d1a125de930b8bd6a62bf1dfd4

Observation 4a5f5dfe-42d3-458a-9933-33c402fa66f5 · inbound

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs cites this paper.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Low-Resource Languages Jailbreak GPT-4

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.105339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.105339Z digest=sha256:6b44a6c9e3061f7a58d10a3436e8c60c7d974b50561c37bb0c76e60d6f47dff9

Observation 3178df8e-113e-4b29-83db-0de3b2787764 · inbound

Large Language Model Safety: A Holistic Survey cites this paper.

Large Language Model Safety: A Holistic Survey Low-Resource Languages Jailbreak GPT-4

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T05:19:34.544425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:19:34.544425Z digest=sha256:b2658614e8c39567758bcdcf0bcd43ce190801e65e0e241a428fc3d389f4f4bb

Observation cea33b91-912b-4498-b599-90468491ec88 · inbound

CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers cites this paper.

CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers Low-Resource Languages Jailbreak GPT-4

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T01:03:12.321630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T01:03:12.321630Z digest=sha256:612cecc70a558f4a7874d83276e611d7542943add091baa5a588d3a9ace5e10c

Observation a76bef2b-afc8-45d7-b375-199041cb38a0 · inbound

Safeguarding Large Language Models in Real-time with Tunable Safety-Performance Trade-offs cites this paper.

Safeguarding Large Language Models in Real-time with Tunable Safety-Performance Trade-offs Low-Resource Languages Jailbreak GPT-4

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T22:35:27.002467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:35:27.002467Z digest=sha256:9366d9ee2634634005991307b23556068ea1f86a96293d30d4ee9b9fbac9cfc8

Observation e088c2cf-5dc4-4985-8771-c216233db363 · inbound

Self-Instruct Few-Shot Jailbreaking: Decompose the Attack into Pattern and Behavior Learning cites this paper.

Self-Instruct Few-Shot Jailbreaking: Decompose the Attack into Pattern and Behavior Learning Low-Resource Languages Jailbreak GPT-4

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T20:35:52.877810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:35:52.877810Z digest=sha256:c5f4eed6870f4ee29022f98029c23dd0b3ceddbe902b7b3169098dbfe7ee103b

Observation b95ecf7d-e00e-4012-a4b9-cbdb8aefe43f · inbound

Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints cites this paper.

Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints Low-Resource Languages Jailbreak GPT-4

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T20:34:36.603073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:34:36.603073Z digest=sha256:7cf1466b1afaffc4e98d74a0841057a9127abcd64eeaf970cb18701ba3601f38

Observation bc50e58e-37ee-492d-bc8c-52347c806179 · inbound

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities cites this paper.

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities Low-Resource Languages Jailbreak GPT-4

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-09T14:47:15.392539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:47:15.392539Z digest=sha256:aec24b51da798bf1f830d722debdc4f773bdfeebad3d13cf5347f69bb52c09b5

Observation c6ca5c90-bcec-4041-8a42-ba1e7e58f2ab · inbound

KDA: A Knowledge-Distilled Attacker for Generating Diverse Prompts to Jailbreak LLMs cites this paper.

KDA: A Knowledge-Distilled Attacker for Generating Diverse Prompts to Jailbreak LLMs Low-Resource Languages Jailbreak GPT-4

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T04:22:00.605573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:22:00.605573Z digest=sha256:df2cb51a7e8ca6b301dd143b0ac8cf85eb4188a19cd1d97555f803a2aaa23bee

Observation 0ef389d6-57e7-4057-a72b-6d2efaffec07 · inbound

Salamandra Technical Report cites this paper.

Salamandra Technical Report Low-Resource Languages Jailbreak GPT-4

Reference 213

Resolution
unresolved
no resolver link, observed 2026-08-08T04:58:34.959959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:58:34.959959Z digest=sha256:f3ef24f3f602277c875a7c2f80f30197fb519171ca52878f6ce7d4891ccbc62a

Observation f1577663-4b9b-47da-b1c0-c60ae7821c06 · inbound

Mind the Gap! Choice Independence in Using Multilingual LLMs for Persuasive Co-Writing Tasks in Different Languages cites this paper.

Mind the Gap! Choice Independence in Using Multilingual LLMs for Persuasive Co-Writing Tasks in Different Languages Low-Resource Languages Jailbreak GPT-4

Reference 139

Resolution
unresolved
no resolver link, observed 2026-08-07T21:10:57.781784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:10:57.781784Z digest=sha256:16e482577ead75c67d65fddb8a6abe804323e72db92a9f0d22d3914270492fd0

Observation ebe7a971-d1ae-4437-b7b3-138e698230a1 · inbound

The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions cites this paper.

The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions Low-Resource Languages Jailbreak GPT-4

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T23:08:02.401728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:08:02.401728Z digest=sha256:15f9e34096ff5efa437af0a7183a44ee94dcddc83eb5d43a64fe875f6911de32

Observation 730a3e3a-2ccb-4d8d-a0d3-d2bd2b570796 · inbound

LLM-Safety Evaluations Lack Robustness cites this paper.

LLM-Safety Evaluations Lack Robustness Low-Resource Languages Jailbreak GPT-4

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-23T01:27:21.257186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-23T01:26:45.402983Z digest=sha256:5381a58b8ee168f375b4324b46670a37b17be1d9f677fe7c449bb5c32f23bd89

Observation 1478c16c-eb79-4363-b9f9-85aeea398296 · inbound

CipherBank: Exploring the Boundary of LLM Reasoning Capabilities through Cryptography Challenges cites this paper.

CipherBank: Exploring the Boundary of LLM Reasoning Capabilities through Cryptography Challenges Low-Resource Languages Jailbreak GPT-4

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T06:05:12.200545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:05:12.200545Z digest=sha256:3779ff563fe7227a9d62b4b15f76441799d0458eceb0a7246ea9f7b16cf2e6d3

Observation 6e3891d3-7c1f-479f-962c-e7801b07c097 · inbound

JailbreaksOverTime: Detecting Jailbreak Attacks Under Distribution Shift cites this paper.

JailbreaksOverTime: Detecting Jailbreak Attacks Under Distribution Shift Low-Resource Languages Jailbreak GPT-4

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T05:57:58.376446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:57:58.376446Z digest=sha256:77abac130c4e92d5dc37e36e420bbd1375dfa4658388ab3eefb5e20928628c22

Observation 9842d1c3-4b93-464b-aded-7334c802a1be · inbound

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models cites this paper.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Low-Resource Languages Jailbreak GPT-4

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.854124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.854124Z digest=sha256:dc9330ffb286bf3a22b05742c851794100b67b60826285cab9e74f1a9a47eea6

Observation 2d7480d4-067b-49c8-a9d8-bc8226d1286f · inbound

Unmasking the Canvas: A Dynamic Benchmark for Image Generation Jailbreaking and LLM Content Safety cites this paper.

Unmasking the Canvas: A Dynamic Benchmark for Image Generation Jailbreaking and LLM Content Safety Low-Resource Languages Jailbreak GPT-4

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:39:22.213472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:39:22.213472Z digest=sha256:fd312181cf85a94b64a549e04395140f6eb29000b655eb62e05a068a6d6308d3

Observation cc920407-b97d-4470-b445-96c8b683fbf3 · inbound

A Survey of Attacks on Large Language Models cites this paper.

A Survey of Attacks on Large Language Models Low-Resource Languages Jailbreak GPT-4

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T20:34:34.526897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:34:34.526897Z digest=sha256:b6ddf75c96b89fa0288efaaf15a55e766b22fc9b45d46eacc8ff181763e70014

Observation b9de7364-2f94-42ad-9a4c-27242dac9181 · inbound

Improving LLM Outputs Against Jailbreak Attacks with Expert Model Integration cites this paper.

Improving LLM Outputs Against Jailbreak Attacks with Expert Model Integration Low-Resource Languages Jailbreak GPT-4

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T20:36:57.255813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:36:57.255813Z digest=sha256:ffa376dc56dbcec693a4cafc571b7e7f876e835476d8f040014828790048c0c4

Observation 93bf1d20-5cec-4421-a34e-59cce3628d5e · inbound

Towards medical AI misalignment: a preliminary study cites this paper.

Towards medical AI misalignment: a preliminary study Low-Resource Languages Jailbreak GPT-4

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:16.928485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:16.928485Z digest=sha256:b35d21f9360ba0bba3c68745635818912c71ec9e95fb014c7b2c763683fa2498

Observation 87021e4c-06e4-467e-89ff-3457213ba352 · inbound

Audio Jailbreak Attacks: Exposing Vulnerabilities in SpeechGPT in a White-Box Framework cites this paper.

Audio Jailbreak Attacks: Exposing Vulnerabilities in SpeechGPT in a White-Box Framework Low-Resource Languages Jailbreak GPT-4

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:27:06.090288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:27:06.090288Z digest=sha256:9ac0ab02f5c62a8d10542719ed82aed60c3cb4b40647f4fa498af05817e3cc81

Observation cd2759c2-ef6a-4d2e-82df-7b5953be472a · inbound

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline cites this paper.

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline Low-Resource Languages Jailbreak GPT-4

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:58:20.347882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:58:20.347882Z digest=sha256:c8b94bf376692c2e71e211a3ba7c322c4021a9f1cc8fd48e7badf01c00b86d58

Observation 3b49342b-9290-4a62-9060-ceded0969286 · inbound

Concealment of Intent: A Game-Theoretic Analysis cites this paper.

Concealment of Intent: A Game-Theoretic Analysis Low-Resource Languages Jailbreak GPT-4

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.616522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:53:40.616522Z digest=sha256:59b53460e99788a2219388f8ebd6c0eadc316714b4016d6ffafaa991a4ba25ed

Observation 9aaf56e4-1414-4bb6-8438-5260f479e868 · inbound

Expert Survey: AI Reliability & Security Research Priorities cites this paper.

Expert Survey: AI Reliability & Security Research Priorities Low-Resource Languages Jailbreak GPT-4

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:01.404747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:01.404747Z digest=sha256:ad235d0ef2f7fed178b37fc21bf8bfa8ea884f2fe424abb19b73a098a1d01da3

Observation 9a1dcd97-7d57-4a48-907f-8b3489064b64 · inbound

Compensating for Data with Reasoning: Low-Resource Machine Translation with LLMs cites this paper.

Compensating for Data with Reasoning: Low-Resource Machine Translation with LLMs Low-Resource Languages Jailbreak GPT-4

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:15:35.430375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:15:35.430375Z digest=sha256:faca264e8a51b081d821702ed24c873b1c00d240d9e076d5a06c7d00591d3060

Observation e7b22fde-4c56-45b1-8e2e-19c8d5f09556 · inbound

Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods cites this paper.

Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Low-Resource Languages Jailbreak GPT-4

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:37:08.222926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:37:08.222926Z digest=sha256:04e0db74f235f56239826607e6b058f14b324b0a457f8582444aca8d61aaa44f

Observation feee3bb0-1f51-4cd0-a73d-c45e8e78c78d · inbound

Cross-Modal Obfuscation for Jailbreak Attacks on Large Vision-Language Models cites this paper.

Cross-Modal Obfuscation for Jailbreak Attacks on Large Vision-Language Models Low-Resource Languages Jailbreak GPT-4

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T23:42:42.798103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:42:42.798103Z digest=sha256:e8214ef0221bef48a734d927770150989455f1fecfcbb819aebd96b90bb2a06c

Observation cfe6234d-f633-4e9d-966c-6efa0979c478 · inbound

Bridging the Gap with Retrieval-Augmented Generation: Making Prosthetic Device User Manuals Available in Marginalised Languages cites this paper.

Bridging the Gap with Retrieval-Augmented Generation: Making Prosthetic Device User Manuals Available in Marginalised Languages Low-Resource Languages Jailbreak GPT-4

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:29:48.997312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:29:48.997312Z digest=sha256:916cfb25cdd3c4c270a3bbbdd00a96b80a6a57b8abdc7462e45cae1a419687d2

Observation b3527ed0-15e6-4f3d-802b-ac57da5f1cef · inbound

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts cites this paper.

`For Argument's Sake, Show Me How to Harm Myself!': Jailbreaking LLMs in Suicide and Self-Harm Contexts Low-Resource Languages Jailbreak GPT-4

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:19.158580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:05:19.158580Z digest=sha256:aca18c17b220dc40456ad8c345684e18ec443bdce047fa5997e6f356ffee9ae8

Observation 6df9012a-c4cd-4d40-a6a2-5ea86222a8c9 · inbound

Beyond Weaponization: NLP Security for Medium and Lower-Resourced Languages in Their Own Right cites this paper.

Beyond Weaponization: NLP Security for Medium and Lower-Resourced Languages in Their Own Right Low-Resource Languages Jailbreak GPT-4

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T20:14:26.634219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:14:26.634219Z digest=sha256:df3010e86520a802e30c0324ab4a9969073202abbfc298e1299af1ed1f691fe9

Observation e392ee3c-fb5c-481c-9615-e62c0bd6c656 · inbound

Attention Slipping: A Mechanistic Understanding of Jailbreak Attacks and Defenses in LLMs cites this paper.

Attention Slipping: A Mechanistic Understanding of Jailbreak Attacks and Defenses in LLMs Low-Resource Languages Jailbreak GPT-4

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:11.798190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:11.798190Z digest=sha256:8de69148f2324744a93a057dbb14068df06ee7d918649b6b32944fef2f2d9165

Observation 70640b84-e213-4601-b2c0-aa0b955aa74c · inbound

CAVGAN: Unifying Jailbreak and Defense of LLMs via Generative Adversarial Attacks on their Internal Representations cites this paper.

CAVGAN: Unifying Jailbreak and Defense of LLMs via Generative Adversarial Attacks on their Internal Representations Low-Resource Languages Jailbreak GPT-4

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:46.549467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:18:46.549467Z digest=sha256:6eac4a3c79bdf7ce440ae5cb248055b5e3382435f41789bef3aeed8b83bf5c67

Observation f7ee0b76-e1ea-426c-b837-ad7777a81516 · inbound

Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation cites this paper.

Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation Low-Resource Languages Jailbreak GPT-4

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T19:26:13.755908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:26:13.755908Z digest=sha256:89f0ecdc8c4da45c1b1a5e5f97188340c24b00fe123768f335e9d604a960b409

Observation ddf51fb8-ea69-4843-b656-286c1056e48d · inbound

SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems cites this paper.

SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems Low-Resource Languages Jailbreak GPT-4

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:46.143521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:28:46.143521Z digest=sha256:85efc28ebd85c84d8af1734577c919f11e2c5954ae08bf09048a33e864fcd5bf

Observation 015af25f-f1e0-4a56-bfe7-fea6ed1beb49 · inbound

PRM-Free Security Alignment of Large Models via Red Teaming and Adversarial Training cites this paper.

PRM-Free Security Alignment of Large Models via Red Teaming and Adversarial Training Low-Resource Languages Jailbreak GPT-4

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T17:35:48.582594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:35:48.582594Z digest=sha256:67cc0751ccbab639b0535e118c7938cdcd77505d94c45f7f0e0e73de22a29aeb

Observation fd092592-9b9a-45b2-b9c7-b628184385d1 · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Low-Resource Languages Jailbreak GPT-4

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:38.034933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:38.034933Z digest=sha256:9df4a7e42767490dc1fcb10f4edb2cb4a2a6246f827d82b276b6cd504aacd3dd

Observation ea933e5c-774d-4f3e-af84-07d901cd2ee6 · inbound

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? cites this paper.

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? Low-Resource Languages Jailbreak GPT-4

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-05T16:00:52.265295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:00:52.265295Z digest=sha256:56dafce282b3aa39e48ddea90398f0b5e25fb2cc7a662c66849fa1e908c02bca

Observation a0678564-32d8-4534-93d4-5716c60d3a32 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Low-Resource Languages Jailbreak GPT-4

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.803021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.803021Z digest=sha256:d0698b44dd289cb15b5d97367619d3f44e1dcedfaf80b00709c28fd5a61dfb23

Observation a969486c-572c-44f7-aaab-60fbad43883b · inbound

AntiDote: Bi-level Adversarial Training for Tamper-Resistant LLMs cites this paper.

AntiDote: Bi-level Adversarial Training for Tamper-Resistant LLMs Low-Resource Languages Jailbreak GPT-4

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T16:25:35.748405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:25:35.748405Z digest=sha256:fb482c95df738b7a994b240c8baba58943d849286f99feb6bae760a720b42469

Observation a3a9ec3b-aa08-4e34-92ef-7fe2bdb95827 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Low-Resource Languages Jailbreak GPT-4

Reference 219

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.150266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.150266Z digest=sha256:d3ae0883eb7ed14369252ba8c44430306b2d93b9c2ed05d831e8163d4d385f50

Observation 557b2729-488e-4a06-986c-0e27793bd88f · inbound

ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs cites this paper.

ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs Low-Resource Languages Jailbreak GPT-4

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-18T01:55:37.975623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T01:54:22.995178Z digest=sha256:5d7d66f681748f6363a2c84084d5d083ea6d3183df0b8b294c234098a8425289

Observation 0a168ba9-4614-456a-b49b-c1513d2e07a0 · inbound

SelfGrader: LLM Jailbreak Detection via Anchored Token-Level Logits cites this paper.

SelfGrader: LLM Jailbreak Detection via Anchored Token-Level Logits Low-Resource Languages Jailbreak GPT-4

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T21:43:46.728512Z digest=sha256:4ac765ff16ce70c07c302fc068d10123e84bfee01b7aafd78854b859cc79d474

Observation e9833494-62af-4690-86b2-ec0dcd530029 · inbound

TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs cites this paper.

TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs Low-Resource Languages Jailbreak GPT-4

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T16:02:52.006859Z digest=sha256:28f5816fd0a5536395d1daf1a4348fff3d21f9bc2f704e576a87f46a84fecf7d

Observation e075710c-92a5-4d79-8e88-d7f33173d648 · inbound

LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety cites this paper.

LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety Low-Resource Languages Jailbreak GPT-4

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T15:02:57.349931Z digest=sha256:4d0cedff3ecfd53f61a943a521b7d4479472d67718faee8a60eecb4196091edc

Observation 6c23ba0a-308f-48a1-9558-2eea7772ee62 · inbound

COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling cites this paper.

COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling Low-Resource Languages Jailbreak GPT-4

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-10T01:14:16.831333Z digest=sha256:aa3acc9766505103a657915f06c3eefa95818b87ce986553f37aa0fe90d45b40

Observation 7ee8a589-60cd-426c-bbe5-c1aace63aa1b · inbound

Automation-Exploit: A Multi-Agent LLM Framework for Adaptive Offensive Security with Digital Twin-Based Risk-Mitigated Exploitation cites this paper.

Automation-Exploit: A Multi-Agent LLM Framework for Adaptive Offensive Security with Digital Twin-Based Risk-Mitigated Exploitation Low-Resource Languages Jailbreak GPT-4

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T11:45:21.608035Z digest=sha256:9e7f3eb766f799bd954b94e860000eeb1d06a2c0aa7209a835bfb5f41e07cb48

Observation 1a14223e-619f-4373-9e0e-8f1d3e813f72 · inbound

Attention Is Where You Attack cites this paper.

Attention Is Where You Attack Low-Resource Languages Jailbreak GPT-4

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-09T19:54:41.445447Z digest=sha256:1e0eb6e6afa27a0d95f7817289fc5bb2911d1e4d7bf930bec04415dcce81ffd9

Observation 1e51af16-11b0-450e-93af-2e7b0e919c8b · inbound

A Theoretical Game of Attacks via Compositional Skills cites this paper.

A Theoretical Game of Attacks via Compositional Skills Low-Resource Languages Jailbreak GPT-4

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-09T19:07:13.287966Z digest=sha256:7bb2e6d0400f5288b21abdc0fad3e16cbf6c71d3e927eaaf77df5a1f328b0406

Observation db57c284-d282-44e8-8613-4e3b15a9fb7e · inbound

Multilingual Safety Alignment via Self-Distillation cites this paper.

Multilingual Safety Alignment via Self-Distillation Low-Resource Languages Jailbreak GPT-4

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-08T19:35:56.059362Z digest=sha256:a433abab70523e45128f52d098cb5be68f3409885f66a2771f4565349cb11213

Observation 3b739871-9309-4df5-80c6-7e6a5f004c0d · inbound

Multilingual Safety Alignment via Self-Distillation cites this paper.

Multilingual Safety Alignment via Self-Distillation Low-Resource Languages Jailbreak GPT-4

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-11T01:08:32.264867Z digest=sha256:78e67f5af84ab51876c40d4a09b5c5e4297c63e51543ad92d502961992d2cc0f

Observation 1d09e206-3ed3-49f3-b659-baae2f5652be · inbound

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours cites this paper.

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours Low-Resource Languages Jailbreak GPT-4

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T16:06:18.057868Z digest=sha256:0144e083991ae05db49f6aa9ae308974ff7a3737b7f4b9db6fc680a6004838d9

Observation 88731c6b-a13a-4b79-bd28-2d68ae420da2 · inbound

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue cites this paper.

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue Low-Resource Languages Jailbreak GPT-4

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T11:17:19.079380Z digest=sha256:03f34102215c19be7d41845e2b569c058763d60109aa2cd82826e866305a1a8d

Observation b6d226c8-474d-485d-90d1-cdd3e69bc946 · inbound

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue cites this paper.

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue Low-Resource Languages Jailbreak GPT-4

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T07:53:06.928500Z digest=sha256:ac3c6b4073e039068e1e395715c6f9c1bfc56e1d31ec7de09ac60f4febf21c34

Observation a30bdb1c-f63e-4508-bb95-f4f20759d76c · inbound

Few-Shot Truly Benign DPO Attack for Jailbreaking LLMs cites this paper.

Few-Shot Truly Benign DPO Attack for Jailbreaking LLMs Low-Resource Languages Jailbreak GPT-4

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-13T07:06:46.387088Z digest=sha256:48fc69367c41a84a23971f4e928ee6c14731777271f2807e72252cbdfe32b6e3

Observation e5133c3d-5357-48e3-98d0-2801510da929 · inbound

Certified Robustness under Heterogeneous Perturbations via Hybrid Randomized Smoothing cites this paper.

Certified Robustness under Heterogeneous Perturbations via Hybrid Randomized Smoothing Low-Resource Languages Jailbreak GPT-4

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:14.203470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T20:23:24.004872Z digest=sha256:99230611ed6fc55e8e9d439a25c6d1457e5b36e874e6b386fc48f0af89214925

Observation 64eb1c64-7824-44ed-895b-90e5aa7615e0 · inbound

From AI-Generated Content to Agentic Action: Security and Safety Threats in Generative AI cites this paper.

From AI-Generated Content to Agentic Action: Security and Safety Threats in Generative AI Low-Resource Languages Jailbreak GPT-4

Reference 149

Resolution
verified exact
local_arxiv, observed 2026-05-20T18:08:50.503214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T18:08:24.901025Z digest=sha256:09cac3ce8d09ccd46a4b5c60e843658ebfafc98ee86bf6f987bdcd1fe609edef

Observation 81ed05c5-27fb-4e8e-8bcf-36d042269f9b · inbound

Multilingual jailbreaking of LLMs using low-resource languages cites this paper.

Multilingual jailbreaking of LLMs using low-resource languages Low-Resource Languages Jailbreak GPT-4

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T10:33:12.153016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:b2c7e1e3b461ed13e85948406115fd6ee0940c1b0ce47a9e44c0eec44581d0d6

Observation 5f91e606-6411-494d-b76b-97e309ba57d0 · inbound

LASH: Adaptive Semantic Hybridization for Black-Box Jailbreaking of Large Language Models cites this paper.

LASH: Adaptive Semantic Hybridization for Black-Box Jailbreaking of Large Language Models Low-Resource Languages Jailbreak GPT-4

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-21T04:49:35.465825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T04:48:19.926845Z digest=sha256:6bc9c0f0949a9bd8b2f1bdd9040631be1ddd884dc516832852baa84302d075a0

Observation 3c8e8c2d-22da-4878-818e-c73a971c8684 · inbound

SomaliBench Eval: Measuring English-to-Somali Refusal Gaps in Open-Weight Language Models cites this paper.

SomaliBench Eval: Measuring English-to-Somali Refusal Gaps in Open-Weight Language Models Low-Resource Languages Jailbreak GPT-4

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:34:01.325996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T22:33:37.250679Z digest=sha256:04b7e8262629b82d3dcc083e725ac0975dfb5ad88520d6e4bfe97e33686b1b19

Observation d4844bf5-f01b-4adb-acfa-fe29f720b1a3 · inbound

Evolving Skill-Structured Attack Memory Enhances LLM Jailbreaking cites this paper.

Evolving Skill-Structured Attack Memory Enhances LLM Jailbreaking Low-Resource Languages Jailbreak GPT-4

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T07:13:16.290123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:10:50.007951Z digest=sha256:77441e08466dca13619394e9507c4dc440aab4368e71ab65e211e104ce7f0611

Observation f0e5f82d-8fb4-4c90-945d-3324333f93d9 · inbound

Low-Resource Safety Failures Are Action Failures, Not Representation Failures cites this paper.

Low-Resource Safety Failures Are Action Failures, Not Representation Failures Low-Resource Languages Jailbreak GPT-4

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-28T17:02:24.050953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T16:59:51.969896Z digest=sha256:665516bd96590a8e092bf7d61786dcccbacc85985b52f1b4a3361654edf8dd55

Observation 851b7c7a-e4ef-49d9-8c90-c4f5eb351590 · inbound

Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations cites this paper.

Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations Low-Resource Languages Jailbreak GPT-4

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T17:02:24.386957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T16:56:53.975854Z digest=sha256:9f5ba3d09e663d730951b14800d7cc262a5ba24423ce79bf6d38cf710231e485

Observation 9d2e5706-ef58-49f8-8c6a-a8215af88a7d · inbound

Inference-Time Vulnerability Beyond Shallow Safety: Alignment Along Generation Trajectories cites this paper.

Inference-Time Vulnerability Beyond Shallow Safety: Alignment Along Generation Trajectories Low-Resource Languages Jailbreak GPT-4

Reference 3

Resolution
malformed identifier
local_arxiv, observed 2026-07-02T08:26:47.879196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T06:05:57.528433Z digest=sha256:bdd8649f6dfec7442e8683c5c687e865b3fad1052d1c77273bf0231d8ee1da39

Observation 876a4ca2-4fe7-4d9b-9814-3d24ef178c63 · inbound

CHASE: Adversarial Red-Blue Teaming for Improving LLM Safety using Reinforcement Learning cites this paper.

CHASE: Adversarial Red-Blue Teaming for Improving LLM Safety using Reinforcement Learning Low-Resource Languages Jailbreak GPT-4

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:06:55.641922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T02:34:26.334078Z digest=sha256:35b18e714b0ab1d3a2d5054dbca9866d446eea0bbcb8875cb13cc1292b512598

Observation 0ca34582-8e8d-4677-8e07-024afe7f940c · inbound

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks cites this paper.

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks Low-Resource Languages Jailbreak GPT-4

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T13:26:59.338990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T01:16:07.252429Z digest=sha256:f3b813eae95f6d3dfd29d79925f7772cb8a018795c29a7de02ad9a302a4be74c

Observation 3dde908f-93b3-4a35-86e7-dcaa763179da · inbound

Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics cites this paper.

Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics Low-Resource Languages Jailbreak GPT-4

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-07-02T17:37:14.905991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T21:55:48.561400Z digest=sha256:56234f29a5b036139bdc1aef6e5824e337938ada334c1a3bf1a6a461d43e5d26

Observation 30fbb572-a222-46f2-bcf6-90395448acc2 · inbound

Safety is Contextual, LLM-Judges Are Not: Navigating the Rigid Priors of Evaluators cites this paper.

Safety is Contextual, LLM-Judges Are Not: Navigating the Rigid Priors of Evaluators Low-Resource Languages Jailbreak GPT-4

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-06-27T21:41:18.502143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T21:39:26.268337Z digest=sha256:6b0ca5743ac2384d4cc7bba02582445f9d1858d7911ddb461a49e4f15e0ddf90

Observation 869b28d2-4ea0-40a9-b7e0-ca0b14518301 · inbound

Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis cites this paper.

Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis Low-Resource Languages Jailbreak GPT-4

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-03T01:07:30.331495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T16:48:54.802860Z digest=sha256:68bd66731a13fc31b216a250b32cbfec383a4c348732505d636752d489fbee8d

Observation 3c35fb39-c1c1-4144-8e43-6bb0965f3bf2 · inbound

Grammar-Constrained Decoding Can Jailbreak LLMs into Generating Malicious Code cites this paper.

Grammar-Constrained Decoding Can Jailbreak LLMs into Generating Malicious Code Low-Resource Languages Jailbreak GPT-4

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:58:06.191553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T09:14:50.233629Z digest=sha256:1f869d8fe97fb53d10ed187295c8ef0aa7a3901bb93109405c2f6eee72688934

Observation 87300ae7-8f77-4c6c-9b13-c244b35237bf · inbound

PHANTOM: A Large-Scale Dataset of Multimodal Adversarial Attacks for Vision-Language Models cites this paper.

PHANTOM: A Large-Scale Dataset of Multimodal Adversarial Attacks for Vision-Language Models Low-Resource Languages Jailbreak GPT-4

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-07-04T17:09:59.495564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-25T23:52:02.327522Z digest=sha256:504b8dd233d3d22dfa6765381d00a6702822e14f7df1bb158352b48bd5866b95

Observation ae8e0d9e-e5ae-4280-8be1-391f641c46f5 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Low-Resource Languages Jailbreak GPT-4

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T19:50:11.071589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-25T20:58:53.119386Z digest=sha256:62283beff71615ab35c509d1ca92bf2f55c0103548e62700121005989eb47bda

Observation 4b18ef6e-a960-4f42-ad66-af71683a75e7 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Low-Resource Languages Jailbreak GPT-4

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T10:16:38.577010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:16:38.577010Z digest=sha256:7bc0e978872c69b56a9f3c83f65f7a3f0442bc511e1bade36ae276273b03adb9

Observation b8b7a5d9-06d6-497d-ac9b-cd082d2e9672 · inbound

Mitigating Taint-Style Vulnerabilities in MCP Servers via Security-Aware Tool Descriptions cites this paper.

Mitigating Taint-Style Vulnerabilities in MCP Servers via Security-Aware Tool Descriptions Low-Resource Languages Jailbreak GPT-4

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-07-09T10:26:11.124368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-09T10:22:23.782469Z digest=sha256:4df82e397307bc5064fca1b8ead1609ffda0005dd0d36747cfc1dfced747e021

Observation 22beb086-42d9-4e21-83f2-1bb05cbe5467 · inbound

Safety That Does Not Transfer: Cross-Lingual Clinical Correctness Drift in Deployable Medical Language Models cites this paper.

Safety That Does Not Transfer: Cross-Lingual Clinical Correctness Drift in Deployable Medical Language Models Low-Resource Languages Jailbreak GPT-4

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T18:33:08.006467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:33:08.006467Z digest=sha256:432270be23cdf34d06596e53ff3a37c2e9691041c7e36198df6b0088a494acef

Observation 4b4f6118-97a6-4bab-b21e-601db75eb50f · inbound

QuantiBias: Benchmarking Quantization-Induced Bias in LLMs cites this paper.

QuantiBias: Benchmarking Quantization-Induced Bias in LLMs Low-Resource Languages Jailbreak GPT-4

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T08:38:56.164741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:38:56.164741Z digest=sha256:522af66bd0b8f84d107c086107a1e5bf968bb817506bf49276b32b189b004ceb

Observation 8e123e44-bafe-4111-ac4e-b8a13c14c091 · inbound

Inspect India Evals: An Open Benchmarking Framework for Evaluating Large Language Models in the Indian Linguistic and Cultural Context cites this paper.

Inspect India Evals: An Open Benchmarking Framework for Evaluating Large Language Models in the Indian Linguistic and Cultural Context Low-Resource Languages Jailbreak GPT-4

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T02:40:45.091898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:40:45.091898Z digest=sha256:bfd215e469e0820205d6b0eefe52a6f208520f03958b3a36ce923fe240722280

Observation da228762-a317-4082-a49b-946363d9df69 · inbound

Same violence, different answer: how AI responds to coercive control against women across languages cites this paper.

Same violence, different answer: how AI responds to coercive control against women across languages Low-Resource Languages Jailbreak GPT-4

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T00:14:01.267746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:14:01.267746Z digest=sha256:bee81f9127b559392d3da60a043e75082073469edd7eaaebe247436c92824228

Observation 71d8efd2-22c5-40c3-b32e-9750cd307da2 · inbound

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity cites this paper.

Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity Low-Resource Languages Jailbreak GPT-4

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T00:48:49.383630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T00:48:49.383630Z digest=sha256:7a2b507b5dbc8dd8854a1c440844e14d98ce247b66c414d77e8e4ad30171ae00

Observation d5e71f74-3316-4a22-91b4-87fd5cdccce7 · inbound

Cross-Lingual Bias in Large Language Models: A Comparative Analysis of English and Swahili cites this paper.

Cross-Lingual Bias in Large Language Models: A Comparative Analysis of English and Swahili Low-Resource Languages Jailbreak GPT-4

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T17:11:46.281928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:11:46.281928Z digest=sha256:ea0f92a36b9e20398774c606804235d5d4c106bf826ce31f224e82e367a7c9cc

Observation 99a6199e-6ae3-40ec-b128-1a5692750599 · inbound

Who Bridges Safety? Identifying and Targeting Cross-Lingual Shared Safety Pathways cites this paper.

Who Bridges Safety? Identifying and Targeting Cross-Lingual Shared Safety Pathways Low-Resource Languages Jailbreak GPT-4

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T23:44:39.628652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T23:44:39.628652Z digest=sha256:e7423c63b5dd1f144c0244e1c9679c00e98240fb1c023117300947d0943717e7

Observation 749fd45f-f6c6-445f-80e0-106a461a5a3b · inbound

From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch cites this paper.

From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch Low-Resource Languages Jailbreak GPT-4

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T04:15:50.149734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:15:50.149734Z digest=sha256:d1b181b8bf374521be2203951091cad7ab763db3fdd0ccef312b73356aba81a5