Pith. sign in

Paper Citation Record · LEDGER

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains

As of 1 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 1 inbound Pith citation observation for arXiv:2606.10838.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.10838 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T11:35:16.646773Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T11:35:16.646773Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T07:57:44.596905Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved36
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation acd5e373-7857-4386-bbf5-07fba0ae9dfb · outbound

This paper cites Therefore, research is shifting to- wards adaptation of pre-trained large language models (LLM) for speech recognition and understanding.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Therefore, research is shifting to- wards adaptation of pre-trained large language models (LLM) for speech recognition and understanding

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:28d4f8d4dd50aff5231f744f2d3c193ec43d0c33d34a1dcdb31619af2be6eb77

Observation 9a1743c7-d57e-4f03-85a2-1688951c6e5f · outbound

This paper cites Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T07:57:44.598243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:4758d528b900f42c9af8eba74eed7b201b424fca200c01ad9c643a62188e0516

Observation 7b8adb40-daff-40b0-aa79-2917692a0321 · outbound

This paper cites audio-only.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains audio-only

Reference 3

Resolution
malformed identifier
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:c2bfcf6ec5fa08012eebf9fa12488f6cfa620afdb62c56a73648e1bffb06dfdc

Observation e381f047-59fc-4796-962b-7e8ea919df76 · outbound

This paper cites use the context.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains use the context

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:f058499eab00ca77110ac439a6c60da726bb60f417db1b4881139d4ce712bc9a

Observation ea0d76a0-27a8-4635-a2bf-8e2a1e38a3f7 · outbound

This paper cites Hence, we should not expect massive WER improvements, especially when segments are rather short.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Hence, we should not expect massive WER improvements, especially when segments are rather short

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:32cd5fddaa369497f75b80e3343289215e1e97b61322dd706474c03af6bb5054

Observation 8a1e48c3-bcee-4131-a022-dd9ca0bbaba1 · outbound

This paper cites We also introduced a pipeline and dataset that pair metadata with contextual transcript errors and correction rationales.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains We also introduced a pipeline and dataset that pair metadata with contextual transcript errors and correction rationales

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:15a50bc6976014dccf3a20951656e7ec9e5a88b6264f3e705688ba679e9ed46f

Observation f784c8bd-53c0-4df0-a516-049f0f4a5b85 · outbound

This paper cites an unresolved cited work.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:4f39ee04ac4fe069c18952710d8f8bb2ed2b1b12abad79e3ca060e5f87f3fb00

Observation 020b58d4-090a-4f0c-9044-9173d43025b2 · outbound

This paper cites No part of the manuscript’s content or ideas was produced by generative AI.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains No part of the manuscript’s content or ideas was produced by generative AI

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:87f5e6d5ff413771b5667660e18deddc7be5721291bce9df2534d883847383b9

Observation d165279a-20bc-4274-9b35-5125fdcb84b0 · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Robust speech recognition via large-scale weak su- pervision,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:3c159eababc89d20c3f7695742a766909958ac0d9eea9b81e20aa8685bba93de

Observation eab2204e-8c00-4b1b-a18f-9702595a57b3 · outbound

This paper cites OWSM v4: Improving open whisper-style speech models via data scaling and cleaning,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains OWSM v4: Improving open whisper-style speech models via data scaling and cleaning,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:9874b6c8a84f80077f4e136568731a5a2033ff657490955fcd75a5f4653b59fd

Observation b4342c52-138b-495d-b4d1-540952a43f17 · outbound

This paper cites Less is more: Accu- rate speech recognition & translation without web-scale data,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Less is more: Accu- rate speech recognition & translation without web-scale data,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:4750106ed7c26b13c6966b50d7e8a400a28483b3739bdd663fbcb67e699b6ca1

Observation d455614f-ef5a-47f3-85ea-c30083f859f0 · outbound

This paper cites Contextualized end-to-end automatic speech recognition with intermediate bias- ing loss,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Contextualized end-to-end automatic speech recognition with intermediate bias- ing loss,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:727ae628cf0fc5b7697524eec862fc1c658366bd54526b2f692bf397c1415574

Observation 223b4596-07f9-4cc7-abdb-33ed92d9db3f · outbound

This paper cites BR- ASR: Efficient and scalable bias retrieval framework for contex- tual biasing ASR in speech LLM,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains BR- ASR: Efficient and scalable bias retrieval framework for contex- tual biasing ASR in speech LLM,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:34ec80891125328bae41ac1568df24bbd8f606573b160329c52acf8fa24b2d2c

Observation 33b18c3a-27f2-443e-a7ee-7ed0964e4051 · outbound

This paper cites Contextual biasing speech recognition in speech-enhanced large language model,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Contextual biasing speech recognition in speech-enhanced large language model,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:069ea3dac0bce76d7109e0cb930eea6ac9bfac25573db59b30a4c80f73063be9

Observation 1e4361a6-2b13-4d4a-ae45-1ba3897b0215 · outbound

This paper cites Improving domain-specific ASR with LLM-generated contextual descriptions,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Improving domain-specific ASR with LLM-generated contextual descriptions,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:ac99029b3658e78689e4051f458e8dd76b82ae1ac597dc63439fed879926a517

Observation 89c05a2f-3aad-4286-b1f4-ef00acaee108 · outbound

This paper cites MaLa- ASR: Multimedia-assisted LLM-based ASR,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains MaLa- ASR: Multimedia-assisted LLM-based ASR,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:f39d0e9ff7bf54a71109a1b2c4cf841ed0609bc4f3853c41f5e9b8e8a934759b

Observation 802a4f8d-684f-40f1-bd6f-74c85405b212 · outbound

This paper cites Contextual biasing of named-entities with large lan- guage models,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Contextual biasing of named-entities with large lan- guage models,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:7726b4a061cfaecf26e12feb90b753c3666e1c8c573d7476bd9bfbe89a8b414e

Observation a85f323b-93d5-4371-b69e-de5908bdb9e2 · outbound

This paper cites Listen again and choose the right answer: A new paradigm for automatic speech recognition with large language models,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Listen again and choose the right answer: A new paradigm for automatic speech recognition with large language models,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:347012f2fedf15ea3a697adf18503b6531d00d28afa5cb1711e096c225cfb206

Observation 3db1e8d5-4731-406e-999a-4f36eda89892 · outbound

This paper cites Towards interfacing large language models with asr systems using confidence measures and prompting,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Towards interfacing large language models with asr systems using confidence measures and prompting,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:f4cc71d8a315be4e12957e74a72d36d5be49a6dae57a4b4f911d558214017ee4

Observation 87c3bb3f-1cff-4a71-8c80-5be085830b21 · outbound

This paper cites Predicting compact phrasal rewrites with large language models for ASR post edit- ing,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Predicting compact phrasal rewrites with large language models for ASR post edit- ing,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:f6f5a81dca107f416048ce2895f7cfb29085251b49d79c23cdee300030082b45

Observation cefe8e71-cb86-42f0-8935-72eda160ebac · outbound

This paper cites Chain-of-thought prompting elic- its reasoning in large language models,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Chain-of-thought prompting elic- its reasoning in large language models,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:ea9d45e4329f6a213031f10e3112549269004c114fda539950a7b04346dac031

Observation 85227b8a-768c-43ce-aea9-730a1382386f · outbound

This paper cites Distilling step-by-step! outperforming larger language models with less training data and smaller model sizes,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Distilling step-by-step! outperforming larger language models with less training data and smaller model sizes,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:2f49fab342681e6a25ddc0b9f06db275675de701d0ffb61c90f2d7754a32b55d

Observation a323edff-fbb5-404f-a8a7-d13e1c1228c0 · outbound

This paper cites Speech recognition rescoring with large speech-text foundation models,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Speech recognition rescoring with large speech-text foundation models,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:d2966eb0d4a4782b37d54f829e0cdf794236caf6e71723cbe55d076a6a847fa6

Observation 9259eaa1-fe73-416a-898d-290458286e7b · outbound

This paper cites Phonetically-augmented discriminative rescoring for voice search error correction,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Phonetically-augmented discriminative rescoring for voice search error correction,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:f4e90cd29d972ec3b849953672a06f40b133e160757ccb34ee2803f2248be36c

Observation 68dde5ba-dfa9-4e03-8391-48a514d37a25 · outbound

This paper cites SALSA: Speedy ASR-LLM synchronous aggregation,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains SALSA: Speedy ASR-LLM synchronous aggregation,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:d8711800c9a3c5dd6a9b98c5cc6cd074a260ec6592a499db384f2f45e7c392f4

Observation 545d17d0-bfff-41fb-acd5-a98076934f6c · outbound

This paper cites Skip-Salsa: Skip synchronous fusion of ASR LLM de- coders,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Skip-Salsa: Skip synchronous fusion of ASR LLM de- coders,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:39e2a60a2487923994d657740eeeee543c6c2fab1fb07b85e607ed2a021f51ca

Observation 35eda3da-a87d-4dcb-8e26-e4993f24c6de · outbound

This paper cites SAKURA: On the multi-hop reasoning of large audio-language models based on speech and audio information,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains SAKURA: On the multi-hop reasoning of large audio-language models based on speech and audio information,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:1c68e7b02d1c8cf9a85ddff7de70caea2bb9719187330a3e2d63bd8c751b02c8

Observation 686169e5-8eb6-4272-a024-a89dfb3179e7 · outbound

This paper cites MMAU: A mas- sive multi-task audio understanding and reasoning benchmark,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains MMAU: A mas- sive multi-task audio understanding and reasoning benchmark,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:14e76c1d87bff54e4d73cdb6d6e6363d29f92af34e28785ae790f08612c65099

Observation 2017c47c-82d9-43f0-a6f4-89138b205358 · outbound

This paper cites Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Audio-CoT: Exploring Chain-of-Thought Reasoning in Large Audio Language Model

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:57:44.604366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:07d04c7d4b9daf295ebf5b889bf0dc571a036b6b14a3bb13e1927e67b3b3573b

Observation 3e6019b8-a250-4d77-9ec6-2806549b1c65 · outbound

This paper cites Audio- reasoner: Improving reasoning capability in large audio language models,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Audio- reasoner: Improving reasoning capability in large audio language models,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:6ce28181ea221e9ff759fec460db0fea719294fcd047d8e1b6fac161458b231e

Observation 1e5b62ca-2395-481a-b670-fb09cca10408 · outbound

This paper cites Can large audio-language models truly hear? tackling hallucinations with multi-task assessment and stepwise audio reasoning,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Can large audio-language models truly hear? tackling hallucinations with multi-task assessment and stepwise audio reasoning,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:3f22ab8f012e19413c379ec0525d3df000f0e61bbccce0974ea69a8ea3677a62

Observation d3024e32-e923-4cb2-88a3-3c95211dd8d1 · outbound

This paper cites DeSTA: Enhancing speech language mod- els through descriptive speech-text alignment,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains DeSTA: Enhancing speech language mod- els through descriptive speech-text alignment,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:09a61bd3e838578157eeb6ed33107b69eaa594c5859e07f269e6f55eb9885c80

Observation c695309d-20f2-4b9d-a53e-d3b0b92d4c23 · outbound

This paper cites Developing instruction- following speech language model without speech instruction- tuning data,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Developing instruction- following speech language model without speech instruction- tuning data,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:55b21883181e9d4d4c0893df21649ba6b02ba100f82fdabe349f0862948162fa

Observation e8599cf0-eef5-487d-8c9c-9839136ce988 · outbound

This paper cites Desta2. 5-audio: Toward general-purpose large audio language model with self-generated cross-modal alignment.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Desta2. 5-audio: Toward general-purpose large audio language model with self-generated cross-modal alignment

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-03T07:57:44.601381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:bc31096e9869607e825d2072bf648e5c0acaf0f6beeb6e57757db4f61a731088

Observation 18ee4180-e07f-45d2-994d-a7e493e06cc4 · outbound

This paper cites GAMA: A large audio-language model with advanced audio understanding and complex reasoning abilities,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains GAMA: A large audio-language model with advanced audio understanding and complex reasoning abilities,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:40c968332b3d4ebed99a8bf86ae1b87be15487b296cb53c46379313cd960e92f

Observation f3062b14-86a2-46f7-8d3e-ce76488ef62b · outbound

This paper cites GigaSpeech: An evolving, multi-domain ASR cor- pus with 10,000 hours of transcribed audio,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains GigaSpeech: An evolving, multi-domain ASR cor- pus with 10,000 hours of transcribed audio,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:72f23280ef18531e7fb6cd87ec76b156e24756601ff0d4a5f4a048203aabddbd

Observation e7320777-0e57-4ee5-a56c-f88cd74721c1 · outbound

This paper cites SlideSpeech: A large scale slide-enriched audio-visual corpus,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains SlideSpeech: A large scale slide-enriched audio-visual corpus,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:a222d09e8e656a0b554fdcf1b7ffb84658374285fada65aa63fc41d94a2fecfc

Observation d8e43eff-2ebf-45a5-aa2b-efd4b4c9af5e · outbound

This paper cites SlideA VSR: A dataset of paper explanation videos for audio-visual speech recognition,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains SlideA VSR: A dataset of paper explanation videos for audio-visual speech recognition,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:12f363b3a3c363a1d5f2a3b18fc0b6b631a876d32b223b41af784d9491b4f99b

Observation a955a042-d024-4580-90f4-9bbe6c244734 · outbound

This paper cites M 3A V: A multimodal, multigenre, and multi- purpose audio-visual academic lecture dataset,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains M 3A V: A multimodal, multigenre, and multi- purpose audio-visual academic lecture dataset,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:012842013bdb24f6ecb80dc1d2007e8731659151a020979cddbfaa0154b5608d

Observation 82170b27-1925-4d80-bca0-df0dbf45eda2 · outbound

This paper cites BERT: Pre- training of deep bidirectional transformers for language under- standing,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains BERT: Pre- training of deep bidirectional transformers for language under- standing,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:fc3187ebe0982241a4915d5e5eb9b1cb87c29fb483ba84e2f3d786917da8bd3e

Observation 57b9fb9d-ddd8-4c1a-8e00-c663128684de · outbound

This paper cites QLoRA: Efficient finetuning of quantized llms,.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains QLoRA: Efficient finetuning of quantized llms,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-27T11:35:16.646773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:b0a39cb90e5979c9a4e29c9a6bdb1ccbb217c29bdbfb9d450b1aeda2c0480fae

Pith citing papers

Observation 9a1743c7-d57e-4f03-85a2-1688951c6e5f · inbound

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains cites this paper.

Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains Towards Deep Contextual Reasoning from Broad Descriptions for ASR with Speech-LLM via Metadata-Driven Reasoning Chains

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T07:57:44.598243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-27T11:35:16.646773Z digest=sha256:4758d528b900f42c9af8eba74eed7b201b424fca200c01ad9c643a62188e0516