Pith. sign in

Paper Citation Record · LEDGER

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition

As of 4 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 1 inbound Pith citation observation for arXiv:2605.02782.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.02782 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T18:06:47.859998Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T16:57:11.349967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact10
  • verified fuzzy34
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 48dd6110-208d-40dd-822c-182dc72dcd16 · outbound

This paper cites Attention is All you Need , url =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Attention is All you Need , url =

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.854601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:40b189ff6c7a6b30d570550fbd7ae100a00a4cc0968387c40ca9214800a6b5f6

Observation abfbf2bf-15ea-4e16-bd20-bb0be9f09530 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T18:57:28.886773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:cfb5d2a17e010369e6ad471c65b6350188f784331e594a1134d7f005fbef7a2d

Observation 25f871f0-514e-4a1d-b2f9-725f1bf5a208 · outbound

This paper cites 2024 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , eprint=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.843242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:835bd2f1c13b087fc0591583480e5ea4394d9bab76cab63863eaf372cbd632fd

Observation 36730354-62dc-4052-bb4d-7221d63a623e · outbound

This paper cites 2022 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2022 , eprint=

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.846544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:d2e4c6d22e04d623b23fb28cb15bd46038c5c18627dcd80364657a2dff21b754

Observation 382eb0ea-8551-47ea-aca8-ec0f3a8e9594 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.835612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:372e9c95eefb93bf0a76107bae4f733e97e4cc409a21e83b81fe67b8d53a1db9

Observation ef4927ab-c9a5-46fd-925a-068d7e2ed7d7 · outbound

This paper cites 2024 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , eprint=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.839378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:265e89ef5bae70d9b23f41a035741389fb92da48b75a7fb1d5fead9a9d530f45

Observation a8570081-f8b5-49f8-a4c0-77433598c369 · outbound

This paper cites Huang and Kenneth Watkin and Simone Frame , year =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Huang and Kenneth Watkin and Simone Frame , year =

Reference 7

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.429906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:ef6938445971600dc228e1ce002647c2764b2961bbd31b25e8c4a3ff5c8905e4

Observation 04cb6bed-f215-41a4-911b-333432346655 · outbound

This paper cites The TORGO database of acoustic and articulatory speech from speakers with dysarthria , volume =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition The TORGO database of acoustic and articulatory speech from speakers with dysarthria , volume =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.850173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:1a9cc16bdaefcb70f91e25a7dd8236a68f0dc8dffe5cb393ecb11004b780519e

Observation 9eee9828-d0d9-42c1-b5a3-2afd7d957b40 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.859162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:8b302d2cd0c9b56afd46272d69921b5e860a23e93083dece18685e36f561b92b

Observation 00b9c6d1-0fe8-46fc-8ab1-067e2e28de5c · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Robust Speech Recognition via Large-Scale Weak Supervision

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-08T18:08:53.438480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:bbff250658a6a4236d2695a68a13d6c521762c259a90ed88a26bac5f71288f0d

Observation 1e592c9d-a193-465e-ac09-909bf9770593 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Advances in Neural Information Processing Systems , volume=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.875307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:90fd0ff526bc6752df999ee7ff6264925351ac74b64c29d8c1331d0aa7ffbd26

Observation 5f66ad36-d3be-400e-82e4-f5463f295cbe · outbound

This paper cites Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:45:44.665734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:a38899f39db0c0f8e8febf25c4d36192d88e1d79169bf1494a5b06ff38368a70

Observation cfffa3e2-59ab-4fe8-8c71-e4ff17edcb33 · outbound

This paper cites Qwen2-Audio Technical Report.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen2-Audio Technical Report

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T02:14:45.520303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:2a98f5538fc58f5d8f61856e511dcd8a12c35b55dbc1c592676f193c0c967826

Observation cc47f93b-d4b6-46d3-98cc-71af1ef3e089 · outbound

This paper cites Qwen3-Omni Technical Report.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen3-Omni Technical Report

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:20:38.220353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:2c7aafe7db2d9c53333a6677c0e6081bf8601afba1ca6b06357fa771e48f803e

Observation b09c780d-ef32-4bd3-835d-1fb4239e4ed6 · outbound

This paper cites Qwen3-ASR Technical Report.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen3-ASR Technical Report

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T13:59:09.287362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:e21bdc5261b9fd38f3b7c2583e2c26827635be2949c3d6b544679b74e3eda93f

Observation 1f382364-325a-42ce-b87c-474eeecb3c28 · outbound

This paper cites an unresolved cited work.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-26T05:51:46.867064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:6b70647fd64d7079f61d60a1d3927257b1909691520f1a175bd89ab532233677

Observation afc141da-f4c4-4e2a-b833-c8411c26ee84 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.870905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:82c10c78c74a0e416eeafbd51fceb0cfc386094bf65ff6157de2dbaf9f8bf21c

Observation 759082b8-cc24-4121-bb31-9339ad7052b4 · outbound

This paper cites 2026 , howpublished=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2026 , howpublished=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.808050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:e40329b5d7eceeadffbc63711eab51229e0003db9f460ca112f83c0007d8142b

Observation 815abc3d-fa9b-4526-a1cb-69c235784603 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.820714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:e39ac396375b89d8aae43a36340c868271d3d8cc448dfcc224a25a3b83ddf5ad

Observation 9e2681ed-8247-4c3e-9b71-7ab2d1ea266f · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:07:27.685083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:ed2ef055928722ad197be0f98115c9815a177cd170380c07e5114e5cb4ff9602

Observation 0606d955-6152-4192-92aa-d46fb9b50092 · outbound

This paper cites 2026 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2026 , eprint=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.827925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:47626e8307a7af9e960d38e939c07581af5b597b197256d5761a67cc210427a7

Observation a69bc087-7e50-403c-b22f-74aaddf284c1 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.831903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:38ccf9e944e49d994fb07ac280a65ac53b47875302b0d5b1c1effd86b2ea032c

Observation 17fb27cc-669d-4ee4-8bf8-e2a22b41866e · outbound

This paper cites 2023 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2023 , eprint=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.800656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:7011c750004222c6b37833214b2829731a67aa13de5a9914283eefd6fcc9e97e

Observation 8494eeaa-22d6-4b7f-a522-ee7c528f3cf9 · outbound

This paper cites an unresolved cited work.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-05-26T05:51:46.789546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:63bf07770f7401e4341a36dc470261505a83b225818485eb755004136dbf6c5c

Observation a3f07603-28fe-4459-8fc2-99edfeb344cc · outbound

This paper cites 2020 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2020 , eprint=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.781921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:d7e68679d6d9363d437e5cccd6b9f28094b396ac7926c51746201e2a11ffa603

Observation 3138e4ef-d2fa-4e0f-affc-18745287b4e0 · outbound

This paper cites 2019 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2019 , eprint=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.778031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:0efb80a3fa8269b14cc8bd71f1691220b0b79449c64ccc55f2abbfc6cf6fd769

Observation f819bf0d-ead0-479e-af11-d0ace543000a · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.769333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:5f495362a8420493623f1bf2bab1a45abd03321746cba74e6e70a6fb25745864

Observation 82b25f50-6f04-4bd9-a567-13d95ccd4b42 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.773634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:8e33d89845ef79fde663a3412a8342cde8d3e2ea7ec36ed93e938dfbfaeb4405

Observation 47434211-2451-46ad-b537-05a2aad74924 · outbound

This paper cites 2020 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2020 , eprint=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.785680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:502f851c3b28282b337133ed6352a9b5097a51a7b0abeb7aafd461b5bc66c122

Observation dae98243-70fb-40e5-a150-56d58d85fa51 · outbound

This paper cites State-transition interpolation and map adaptation for HMM-based dysarthric speech recognition.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition State-transition interpolation and map adaptation for HMM-based dysarthric speech recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.793600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:23b81a0463c71688da03c61101bc5ac2ea668c078a174013c4cca2d5cf1afac6

Observation d8a3375b-1393-40c3-995f-9075ae3be6bd · outbound

This paper cites Estimation of Phoneme-Specific HMM Topologies for the Automatic Recognition of Dysarthric Speech , volume =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Estimation of Phoneme-Specific HMM Topologies for the Automatic Recognition of Dysarthric Speech , volume =

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.797546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:bc94d84db58c9213486fba246aabf979fbcc12870d935a7590b8a550a60ce571

Observation f9fb7c86-9838-4fd6-9a3e-fac60b4e4ccd · outbound

This paper cites Proceedings of Interspeech 2016 , pages =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Proceedings of Interspeech 2016 , pages =

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.765101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:264483f3c077fbf21caf758f014248e396d18410aa0d40f3ac62acccf9736126

Observation 461654aa-6da6-4048-92e6-80c647603d53 · outbound

This paper cites Proceedings of Interspeech 2023 , pages =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Proceedings of Interspeech 2023 , pages =

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.804302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:f5ad0299dc9102aa59d122c0cc40f2d056b51ae6ff6fed4f122c5a84d366629e

Observation eaf12820-967b-4bd4-b1bc-4aa1b86833ce · outbound

This paper cites Two-stage data augmentation for improved ASR performance for dysarthric speech , journal =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Two-stage data augmentation for improved ASR performance for dysarthric speech , journal =

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-08T18:08:53.426444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:e5831930a12bf745907df80122f1a8806550ca62cb05e26ba62d944501e5f772

Observation 43211658-e092-48a1-a61e-a4f84dfb50f8 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.824345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:9a0157049f6d3a542201c430cf21ebb968d39d0fb3e08435fb393b5f4de059c9

Observation cd90e2c1-92c2-4c4b-b8ec-92aea2305c8f · outbound

This paper cites 2024 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , eprint=

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.879111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:d912473daff362f3709ace54205433fd42005c4917f67e435b3c6c5b302ebb5c

Observation bdbdef6e-27b8-4b95-9824-665614cd2726 · outbound

This paper cites 2024 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , eprint=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.816987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:e3741fe1c2750a958f505d41013e48b337e89d9cc50893aca5f9c7cec38104d3

Observation bb9911bf-b8dd-4e10-a4bc-14ad09125cdd · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.812244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:502252ce04fe236c44fd7078d921727b1faa1c5342fe90a2f42df07452e6ed57

Observation 89fc1373-6409-4701-b7ee-bfb6bc4d0f5d · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.748914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:cc7c6ebda7721dbd2257eafb44037eb7283516ccd6d99b2c964c0a514caba97c

Observation 3521aa13-a46a-4964-af84-f56b5b91586a · outbound

This paper cites Urbina and Peter D.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Urbina and Peter D

Reference 40

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.433304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:f1a820df32548e133dcef3326b4b264a9b3dab2f13ccfcd425c1e81f8d49e424

Observation a0f455ef-ccc3-45eb-9e62-687c2a082f53 · outbound

This paper cites 2025 , howpublished=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , howpublished=

Reference 41

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.416248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:f7d07ee42bd8c3827623b6f5516788b983a388928cdb5ba095f351e6e68b0264

Observation 9e1483b4-ded0-4f9d-9019-0125eab270ec · outbound

This paper cites 2024 , howpublished=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , howpublished=

Reference 42

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.412609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:a8225062653f16e720f91a4cb71540f20b0289751607629bfb8474af519beb12

Observation b81bdc27-f711-4033-b93e-b126e8981e4f · outbound

This paper cites 2021 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2021 , eprint=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.753016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:577cee6c3fc4b3ac7c2538fcb4881e1ce22292986b2bc18f8bce758ec08d423b

Observation 0ec50091-831a-430d-b2c4-7fef8c9c57a1 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.756634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:90a1261ccbf7989940ed59039a79187fa79ffb9942780bbbfd6eee8ec2a9ea18

Observation aa6951c3-96c7-4afd-b77e-7d82356f5c13 · outbound

This paper cites an unresolved cited work.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Unresolved cited work

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-08T18:08:53.421100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:b004669de9cc0fd0eed7fa806c8a625f72bbb9fdad73e04dc72aa10aa2c4ccf5

Observation 1a996b29-27d5-42a3-8a09-cd3a1776eaad · outbound

This paper cites title =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition title =

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.760948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:9d6bbdacafe6160a2b2157b6b49a94f09e084d4390c434be1613cd3a2e014eb3

Observation 6b7df4a9-0a82-41b1-b131-09227a6b6037 · outbound

This paper cites and Nelson, P.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition and Nelson, P

Reference 47

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.408648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:85f46a04216d7a6018918ba4d7e2c0689baefe751a3fa5ea4a0c587ba5fdd674

Observation effc897b-06b1-4270-bd68-ecdd91200da2 · outbound

This paper cites Population: The 2012 National Health Interview Survey (NHIS) , author =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Population: The 2012 National Health Interview Survey (NHIS) , author =

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.745153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:b8a5be4320df92bfd1646b452dc43e5ba172f027e0084834dc7863272d664e9e

Observation b951dcaf-82a2-4d5c-8d66-6fcbb76bfa96 · outbound

This paper cites Variational Low-Rank Adaptation for Personalized Impaired Speech Recognition , year=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Variational Low-Rank Adaptation for Personalized Impaired Speech Recognition , year=

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.741151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:a7db818b83ddb37575c5745a1f2a60662bf359cbc5a3760b63122572e5e593a8

Observation b6eb8d56-71e8-43f0-8c4d-c461aa698559 · outbound

This paper cites 2012 , publisher=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2012 , publisher=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.863170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:7d10a4879f1da17d288f1e70711587fada5963feb01419cfd578ad60eb43737b

Pith citing papers

Observation 475bb423-9c3c-4c9b-aff9-f39783c5f9f6 · inbound

ESCUCHA: A Spanish Speech Benchmark for Heterogeneous Acoustic Conditions cites this paper.

ESCUCHA: A Spanish Speech Benchmark for Heterogeneous Acoustic Conditions When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T16:57:11.349967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:57:11.349967Z digest=sha256:73314421eaea6f9989b46fbac5afb9027de1c5006da4bb50c8b199fab99bf575