Pith. sign in

Paper Citation Record · LEDGER

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models

As of 13 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 2 inbound Pith citation observations for arXiv:2411.14842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14842 v2

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:53:50.291447Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:08:53.614645Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T13:01:24.303798Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact1
  • verified fuzzy1
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ba5d5b06-610e-45d7-b0b0-45b07683fe57 · outbound

This paper cites online" 'onlinestring :=.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.181090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.181090Z digest=sha256:11349b6af2e4c653855a492732f12c3357514fab082678afba72ef2bf2a1860d

Observation 10b626ae-cc60-4653-99b9-08db4fc775d5 · outbound

This paper cites write newline.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.184464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.184464Z digest=sha256:2cfbbded90422df2cf50d5e720f3d8f62cc6bf0093fff248b8226e4b7826dc98

Observation 03958182-a0ae-4e07-9242-fe66ca3dda32 · outbound

This paper cites https://mindgard.ai/resources/audio-based-jailbreak-attacks-on-multi-modal-llms?hs_amp=true Audio-based jailbreak attacks on multi-modal llms.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models https://mindgard.ai/resources/audio-based-jailbreak-attacks-on-multi-modal-llms?hs_amp=true Audio-based jailbreak attacks on multi-modal llms

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:53:50.562562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.187536Z digest=sha256:735536fae5b4e9216b58c1d4b178e566c9df4427aee86c267bb7dca46c62a321

Observation 648746cc-f3a7-46ef-b7e1-a66bf1f98a50 · outbound

This paper cites GPT-4 Technical Report.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models GPT-4 Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.190254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.190254Z digest=sha256:b23f3a1e55edc52d4e3214e8fb8669761323f7fc8479ef1a7e2bf8f8b53821b3

Observation cc86c316-5d8f-44fe-a60a-5e5e3d4e7047 · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Common Voice: A Massively-Multilingual Speech Corpus

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.193004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.193004Z digest=sha256:f8e84ded355e285ad36b17c192f10c178560a81050e7befe00279922442ac8c3

Observation 37f13612-8d29-464a-83c4-2f0d060d96c6 · outbound

This paper cites AudioLM: a Language Modeling Approach to Audio Generation.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models AudioLM: a Language Modeling Approach to Audio Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.195710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.195710Z digest=sha256:8927a463556cb31d34a933372e610f62c6b8cd3a007821df683ea56afe88d8ac

Observation 96eff583-1f1d-48ee-b94c-d9ee0a3e4208 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.555001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.198493Z digest=sha256:6ffa4349333d6c6f7ba28540f0986c1d8a9cbdf021340d241b4268a855620892

Observation fcf0152e-b1f2-4192-95ba-797a2a41b387 · outbound

This paper cites X-LLM: Bootstrapping Advanced Large Language Models by Treating Multi-Modalities as Foreign Languages.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models X-LLM: Bootstrapping Advanced Large Language Models by Treating Multi-Modalities as Foreign Languages

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.200951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.200951Z digest=sha256:cdc88ed208e91feca8925521077755938f74506a6fc591e70ffae353064130f9

Observation d0cc565e-8671-4409-8d88-c0eef7461e5c · outbound

This paper cites Qwen2-Audio Technical Report.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Qwen2-Audio Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.203848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.203848Z digest=sha256:aa0cdb19e31c8a2d06a8d147c2d9e9df51edfd0f3fb8a3e26a212d4ab3d95355

Observation b9dd0ca6-42d6-4fdf-8f04-b1419870419b · outbound

This paper cites LLaMA-Omni: Seamless Speech Interaction with Large Language Models.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models LLaMA-Omni: Seamless Speech Interaction with Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.206384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.206384Z digest=sha256:dbe562f407ffc5d5f2cb1bdc3db08f39ccf62b986e40d108ad386ae4f7835ea8

Observation 627023ff-a988-4368-a316-a63aa45e0422 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.209199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.209199Z digest=sha256:9fe80ccd4128cf1ad4dd6179c4351124bf0430dd2ff409112348d7b93788f21a

Observation 486790f9-0b5a-40f8-afbc-20ba81c3b34b · outbound

This paper cites Crafting Adversarial Examples For Speech Paralinguistics Applications.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Crafting Adversarial Examples For Speech Paralinguistics Applications

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.211503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.211503Z digest=sha256:21739c6da3458d6357243e30db7d84fd175c46c8737a1c732b4c49c34393a071

Observation a7406b9c-ca49-476d-83c2-231ae9e05159 · outbound

This paper cites Explaining and Harnessing Adversarial Examples.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Explaining and Harnessing Adversarial Examples

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.214015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.214015Z digest=sha256:c919a46b0ad5b3a12af7e5a75eb0ea3e791638f820d906d05752f9c184379a2e

Observation 477936b0-7483-4144-a6dc-0903cd6b9dd3 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.216630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.216630Z digest=sha256:d41e7eaeb11d813fafc277c942e22d8bb6846b2a9686cf9d04f5f796b484ff05

Observation bc1c7d80-3e03-47e2-815f-e82673e8cd43 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.539989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.218844Z digest=sha256:9eb5d79334e1d91294406cda88cc0a71fb9d74a1e065a38b2a0201c32017d328

Observation 42422993-fcd3-4c08-b84e-3dd72e361457 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.531981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.221167Z digest=sha256:e556ea27b91b5f163d7aeaef5a273edb10701aea1521930f709035353368ebb5

Observation 222b03e1-8048-4644-bdb2-827943b144b8 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.223514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.223514Z digest=sha256:2823e012fc03b21315950a4fc3a07b3e72fc707df60bf616d3c4b5606923b66c

Observation 60bab187-7eec-47aa-bbee-82ed448c3745 · outbound

This paper cites Practical Attacks on Voice Spoofing Countermeasures.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Practical Attacks on Voice Spoofing Countermeasures

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-12T14:53:50.378095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.225865Z digest=sha256:07e3048bcbd4628ba371a2c5461ffb8cc0b932694cfb9ac64dca7d040c5160eb

Observation 11b5b3a4-84e1-4dd1-b689-04f7dd265e23 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.519528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.228627Z digest=sha256:ad6f508bbac0ecb9343e66f8c110c0510da68d991001ffeaae5dd5a5746c522d

Observation 7795793c-73df-4ed6-80c1-8a583cd36271 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.512298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.230799Z digest=sha256:826d449cfc29257c324d3f5bbb260db22ed649caa1435e98755c1bcb5eed4e15

Observation e0bb5fff-4783-4c6e-a3a4-eda8094b08b7 · outbound

This paper cites o pf, Yannic Kilcher, Dimitri von R \.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models o pf, Yannic Kilcher, Dimitri von R \

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.233179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.233179Z digest=sha256:936fc818884fcf81dfced9e79d86eb0fb82570d1eba3bec6b87f0db0cecc34b1

Observation 43218a79-3d53-46e9-8701-1f2bce8624e5 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.235548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.235548Z digest=sha256:1e14ed6af396b3d893da042b9d997e37b2fd36f492046bff7cb8a8df0166b09e

Observation 38bf092c-26d5-4826-bb45-b7eb49dcc8c1 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.237798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.237798Z digest=sha256:a29153ef85b58810197ab9ece6170da579312e321ca0ad152e651ce6c46414a2

Observation f25f35c9-0600-405d-854c-58c93ff06b6e · outbound

This paper cites TVQA: Localized, Compositional Video Question Answering.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models TVQA: Localized, Compositional Video Question Answering

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.240142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.240142Z digest=sha256:f183180793151762e17fe8542561464cdb066afa2be45dfa4c0d3ca2e6af5fcd

Observation 3d40e55a-6c83-4881-bc8c-648778f03a46 · outbound

This paper cites BERT-ATTACK: Adversarial Attack Against BERT Using BERT.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models BERT-ATTACK: Adversarial Attack Against BERT Using BERT

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.242789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.242789Z digest=sha256:4a8a865bf7738d6c2cc2f54a472885805c768bee32f7b446d877c9ee95ed70b1

Observation bd491c59-430d-4cb9-acc9-46c016f2f114 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.245376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.245376Z digest=sha256:7f05eb42397d04a74c538419360c59b1c4a0edffd6870df9e56fde8799c49bd6

Observation dfa25e81-5557-42c7-b925-8def239575c5 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.490297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.247705Z digest=sha256:c5e073a6c380e2d5ddca3b16343cc9ceb7c1e0dc4bf9d25f2f8a02dfb148e4cc

Observation f2dfba72-681a-42a7-9238-80a840cd5757 · outbound

This paper cites Universal Adversarial Perturbations for Speech Recognition Systems.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Universal Adversarial Perturbations for Speech Recognition Systems

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.250123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.250123Z digest=sha256:b4d416a17cb4fbbbe4c4b9abdd57f1d60986b24160a4316f51fae3e72f749dbc

Observation b11c9d48-3d80-42e9-93a1-58213be636e4 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.482953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.252728Z digest=sha256:5a2c07c9873a928eebcfbfe8d99a618144212598c98865756a36df9e9255ec12

Observation f9366436-e76f-4f1f-8760-ae3355f798c9 · outbound

This paper cites MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.255144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.255144Z digest=sha256:52e72fcfd472d6c78901140271cb914c814959fca4811d02181bb181dccfc69d

Observation d719191c-745e-40de-9674-334d7756c878 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.257739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.257739Z digest=sha256:12799a5476c36cbc5a5fc380602506be977fe1d486099fdf81719e6a6635c5bf

Observation e29aff72-4f58-4383-a4eb-300dc7d3922f · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.260054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.260054Z digest=sha256:abcd6a8416d7ada2b0fe3bd5358bb0aa34f9e38fda8fa2b1a7667d3cfc380b2b

Observation 28da6201-a48e-4fca-b5d3-aa3929698ff2 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.262421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.262421Z digest=sha256:94a9b02955a9738e813d4a1f5425cb4cb85cc8fa6a10d2ddc7d92e5a4bf3b898

Observation 6563df85-7341-406a-860c-5133be0ec1cd · outbound

This paper cites Survey of Vulnerabilities in Large Language Models Revealed by Adversarial Attacks.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Survey of Vulnerabilities in Large Language Models Revealed by Adversarial Attacks

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.264719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.264719Z digest=sha256:3dc42e358626b83e630b85636aff292d1124a76f6e09f8334ef4045ae9db8278

Observation 0d80c757-b79b-4f3a-80e6-be0e7c0472e9 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.467619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.267548Z digest=sha256:2663a215bf6ff3280d469caeb03fddb6ecb2fc8ef50068834edf290c38aef966

Observation eb30e7a8-8cb1-4b0d-aebd-af82e3ce047a · outbound

This paper cites Intriguing properties of neural networks.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Intriguing properties of neural networks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.269950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.269950Z digest=sha256:e4503397f398a056c14bd65acba4302c45e7387117fe9d9cc58ee29de12da26e

Observation 8e68ac8a-b0bc-4e98-8101-4533f3045b4e · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.272249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.272249Z digest=sha256:c0c17ea6e657e587e220989a305e7584170605c7397b4b3bac1848b9b2803db1

Observation 34cbb73a-4218-4fa6-bdb3-3f9828aa8201 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.460451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.274792Z digest=sha256:2b4b31acda55f0c110869a45e54694f07f51fee94d6a53c64a53a19bd8d49a89

Observation 9973a1e1-1a7e-4873-b5cd-f824c26b7f68 · outbound

This paper cites EDA: Easy Data Augmentation Techniques for Boosting Performance on Text Classification Tasks.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models EDA: Easy Data Augmentation Techniques for Boosting Performance on Text Classification Tasks

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.277008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.277008Z digest=sha256:2547c62d539b76d3f4cd984b8c2c262689280f525c33a5c1436aa3f7ba7509fa

Observation dcf30d46-b378-44b1-b844-e9e79cd55611 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.279212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.279212Z digest=sha256:e51d7f8b4d6a766dfebb31742264b16d14f7a0731a161a685919acd3f74b92fa

Observation 2845429d-25e5-4aa7-a5da-d30a9c8155a6 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.449323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.281571Z digest=sha256:f8b7a879daa75ee2052b583c4e49828a65a87ba79ebbb8fa0880d10712399217

Observation 24ad5ad3-8b5a-4d78-804e-b78ca199ed99 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.441791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.284062Z digest=sha256:ae82ab8224683ec66e984764dee2553eacd34c93ceba5f38237dfffa1f54f801

Observation a6fc6891-7878-449c-ad9a-34aa68a21cc8 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.286582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.286582Z digest=sha256:aaacdceeb3f729af563300d71fcc53cae256144828601f79e95a7e7b74b2183c

Observation 4dc5f1dd-5a8f-4ef7-ad56-ec30331b9091 · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:53:50.434539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T14:53:50.289070Z digest=sha256:8291331617fdc77500402ac575154cbedd016ded1ccfaccdd63ce74abf060c6e

Observation 39a28f5e-ad5d-4226-9991-0e9a31096deb · outbound

This paper cites an unresolved cited work.

Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:50.291447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:53:50.291447Z digest=sha256:95f89b2b091d906246b19a43671035b10d53beecae8029430a9e3c423f15acdb

Pith citing papers

Observation 5fc71e28-71a2-4d99-8525-194bb4fceb9a · inbound

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models cites this paper.

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:53.614645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:08:53.614645Z digest=sha256:f3280232f5d598662cf972d80d84e6d86f97183b257747704578e4d79170997f

Observation 7cf92bd0-5ea8-4629-95ae-e39e65d899e1 · inbound

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs cites this paper.

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs Who Can Withstand Chat-Audio Attacks? An Evaluation Benchmark for Large Audio-Language Models

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:01:24.305685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-18T12:57:04.450462Z digest=sha256:e328317aa0b0f7a042b4d8fd86dd8f0b4ff5f14889e22610fd6a9c8207398d30