Pith. sign in

Paper Citation Record · LEDGER

Multilingual and Multi-Accent Jailbreaking of Audio LLMs

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2504.01094.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.01094 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:21.669746Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1d275f16-d8b2-42f6-b61b-fca787d0a70d · inbound

Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs cites this paper.

Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:21.669746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:21.669746Z digest=sha256:479ee6ce2ccc01bd8e1b90bb05fd63bfdc4a87b82189f2f1b203a78697618627

Observation aa10ee02-ed1f-4879-b6a2-3e92b6ea1eea · inbound

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey cites this paper.

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:34:53.256580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T13:32:57.771753Z digest=sha256:0eb2cfea0914249d1d23b782d2d9c10ac934146e9843db043593b7de5dd5f361

Observation f562dc3d-8d52-44ce-8fa5-e62091a1c5ca · inbound

Evaluating Robustness of Large Audio Language Models to Audio Injection: An Empirical Study cites this paper.

Evaluating Robustness of Large Audio Language Models to Audio Injection: An Empirical Study Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:57.602243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:57.602243Z digest=sha256:ac99657d1873c81593dd496d6c71fd5a6fe39dd5ac1f93a23f4e1a0a64a8c4e3

Observation dc74c2b6-91c8-4bd3-8740-8cff811e3c33 · inbound

A Red Teaming Roadmap Towards System-Level Safety cites this paper.

A Red Teaming Roadmap Towards System-Level Safety Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:23.270108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:23.270108Z digest=sha256:62f1c4563546646d2046fe4d994c6358faab3fd1cfa4361a5998d4e29780a902

Observation abc7a546-3fd6-4739-9b6b-23f4c0705f25 · inbound

Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World cites this paper.

Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:12.137325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:46:12.137325Z digest=sha256:5e8e62f482615d7a1ed0383ab554446a9d2dc1a0dfde5bd16650cab7e66139ad

Observation 8791365a-6029-4713-9822-8f0587a7ca06 · inbound

A Systematic Study of Cross-Modal Typographic Attacks on Audio-Visual Reasoning cites this paper.

A Systematic Study of Cross-Modal Typographic Attacks on Audio-Visual Reasoning Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:38:02.915828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T17:34:10.089555Z digest=sha256:902b8c03f272963440ea2a2b55508ce630108f70714414050bdf9f0772cfb435

Observation e0bd95a3-f524-4d0e-ac72-356170a5c189 · inbound

VoxSafeBench: Not Just What Is Said, but Who, How, and Where cites this paper.

VoxSafeBench: Not Just What Is Said, but Who, How, and Where Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:24:22.104211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T10:19:28.041282Z digest=sha256:299e9fcdc337abef9e1823585bc835e3e068bfd06a06a6504ac4a5035c474c73

Observation b2eb458b-528f-4785-a8f2-f95ac6d69e7f · inbound

Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs cites this paper.

Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:02:24.881577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T08:01:25.938248Z digest=sha256:026b5446b6852b2386e821cfe91ce63d826ad463e92051698f24a6810fba0fa9

Observation 066019e8-969a-4a55-a9b3-8d8457c3e525 · inbound

Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization cites this paper.

Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:45:44.045138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T18:08:16.051117Z digest=sha256:4ef473dc42ed6008a0da1e61b30518e9039146b1146c55fef148044c79cbeb37

Observation 65e4bafc-91e0-49ce-8ad4-11f1dad79184 · inbound

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook cites this paper.

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 164

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:39:49.038313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T07:38:23.099479Z digest=sha256:810fc28998f7b9eeca36f8ee8bf9fb0e38df1f45150ae5e40dc20154da6d86ba

Observation 836a884d-d24b-4801-ad12-fbf407a6b0b9 · inbound

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents cites this paper.

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 177

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:15:05.654669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T22:11:44.891731Z digest=sha256:d31af8078b2bb869825b67f3873566379ab0439e1e957d7de76865d3b944f712

Observation d6bf5e62-1c9f-4606-86dd-ef8e68fa1cfe · inbound

RedVox: Safety and Fairness Gaps in Speech Models Across Languages cites this paper.

RedVox: Safety and Fairness Gaps in Speech Models Across Languages Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 144

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:59:52.812797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T04:37:00.399470Z digest=sha256:10761f4989876079c895ca59e88baa10508a2d157d0b40273ef4036e0d24a3ec