Pith. sign in

Paper Citation Record · LEDGER

Multilingual and Multi-Accent Jailbreaking of Audio LLMs

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2504.01094.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.01094 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:21.669746Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1d275f16-d8b2-42f6-b61b-fca787d0a70d · inbound

Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs cites this paper.

Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:21.669746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:21.669746Z digest=sha256:a08bcea7d3ae4f6e5d2c4c4285732c2375fcab0dfcaab0ad745fcf7bf15e44fe

Observation aa10ee02-ed1f-4879-b6a2-3e92b6ea1eea · inbound

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey cites this paper.

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:34:53.256580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T13:32:57.771753Z digest=sha256:bb8d540659ce04bad30023b196db5b4fbe32eb7924951b3aa6c572978a9c40d7

Observation f562dc3d-8d52-44ce-8fa5-e62091a1c5ca · inbound

Evaluating Robustness of Large Audio Language Models to Audio Injection: An Empirical Study cites this paper.

Evaluating Robustness of Large Audio Language Models to Audio Injection: An Empirical Study Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:57.602243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:57.602243Z digest=sha256:7b22fa1af584748d9832f1f35d6f4fe0d15efbfa7d2029a4b338e8532b6f4ff5

Observation dc74c2b6-91c8-4bd3-8740-8cff811e3c33 · inbound

A Red Teaming Roadmap Towards System-Level Safety cites this paper.

A Red Teaming Roadmap Towards System-Level Safety Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:23.270108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:23.270108Z digest=sha256:5af9cabd432638b5cddb2356769b3276c3433a67002a4bd97478583af238f45c

Observation abc7a546-3fd6-4739-9b6b-23f4c0705f25 · inbound

Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World cites this paper.

Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:12.137325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:46:12.137325Z digest=sha256:824f21d6a58bf7acdc259305d92532b0dcae7d503053a4484025f96317910e66

Observation 8791365a-6029-4713-9822-8f0587a7ca06 · inbound

A Systematic Study of Cross-Modal Typographic Attacks on Audio-Visual Reasoning cites this paper.

A Systematic Study of Cross-Modal Typographic Attacks on Audio-Visual Reasoning Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:38:02.915828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T17:34:10.089555Z digest=sha256:76e38a4fb503b2a5c57936baf4b85c002f381642d0d37e07aa11cb1dc45a479c

Observation e0bd95a3-f524-4d0e-ac72-356170a5c189 · inbound

VoxSafeBench: Not Just What Is Said, but Who, How, and Where cites this paper.

VoxSafeBench: Not Just What Is Said, but Who, How, and Where Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:24:22.104211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T10:19:28.041282Z digest=sha256:93016b6e5784c361ff94b55910fed2afe73355c47059470e7197d2d8160fc632

Observation b2eb458b-528f-4785-a8f2-f95ac6d69e7f · inbound

Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs cites this paper.

Benign Fine-Tuning Breaks Safety Alignment in Audio LLMs Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:02:24.881577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T08:01:25.938248Z digest=sha256:83492b03016d76992ab641e2c6a62523c921d0b78ec5e591b3e00e4abab99028

Observation 066019e8-969a-4a55-a9b3-8d8457c3e525 · inbound

Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization cites this paper.

Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient Optimization Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:45:44.045138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-08T18:08:16.051117Z digest=sha256:3003c0c730595fbf0e9536aa86a703c91826b15b989245d1ac2b356dfc2fdf76

Observation 65e4bafc-91e0-49ce-8ad4-11f1dad79184 · inbound

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook cites this paper.

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 164

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:39:49.038313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T07:38:23.099479Z digest=sha256:d7e013fc4be4f4ae5dba8f62c554702ac924f4700c3e039d582d7a78c9f887d5

Observation 836a884d-d24b-4801-ad12-fbf407a6b0b9 · inbound

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents cites this paper.

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 177

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:15:05.654669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-30T22:11:44.891731Z digest=sha256:653d7e59f0b260d2031c50f7a88312a824b3847b64b4142e54497f430dc2b9f0

Observation d6bf5e62-1c9f-4606-86dd-ef8e68fa1cfe · inbound

RedVox: Safety and Fairness Gaps in Speech Models Across Languages cites this paper.

RedVox: Safety and Fairness Gaps in Speech Models Across Languages Multilingual and Multi-Accent Jailbreaking of Audio LLMs

Reference 144

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:59:52.812797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-26T04:37:00.399470Z digest=sha256:5ae474c0ca1f8db5b8728f0f12a9e58827eea24f6c1716a6849c0a06716d50c0