Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2405.08295.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:30:32.222856Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 474d7e28-d8cc-4621-ae66-e608e21415dd · inbound
Qwen2-Audio Technical Report SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 62792d3a-d3f9-48b5-badd-2459f82d3802 · inbound
Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6f095213-9732-49f6-9f03-231262498176 · inbound
Qwen2.5-Omni Technical Report SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c9bfbf4a-251a-4423-a3d4-c7be0a6c46ad · inbound
On The Landscape of Spoken Language Models: A Comprehensive Survey SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e9316b4-b1fd-426f-8937-3caf555a31f6 · inbound
Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 443ccc8e-7d58-43a8-ba77-17d99a880be4 · inbound
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0a06db11-99ae-49f8-8285-e4a32689457d · inbound
Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdb3b362-ebbd-455e-8dd6-55c2a7f4ad4f · inbound
Unlocking Speech Instruction Data Potential with Query Rewriting SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b91d2310-470c-4a35-a5e0-5ddf7a956343 · inbound
Your Spending Needs Attention: Modeling Financial Habits with Transformers SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89172613-8aed-4848-822a-cc168751ec30 · inbound
EmoSLLM: Parameter-Efficient Adaptation of LLMs for Speech Emotion Recognition SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f54078c4-3e0f-4f78-99d6-2b2fa9376b3c · inbound
TokenVerse++: Towards Flexible Multitask Learning with Dynamic Task Activation SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44ed0d62-41c9-427e-9093-19ea7800fef0 · inbound
Enhancing Speech Large Language Models through Reinforced Behavior Alignment SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24cce4b3-79af-445d-9dcd-fb45577f3dec · inbound
Direct Simultaneous Translation Activation for Large Audio-Language Models SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9a25b118-325e-4ad3-a8db-c625e6a19e72 · inbound
A Simple Method to Enhance Pre-trained Language Models with Speech Tokens for Classification SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7950c080-4db3-4b22-9a33-64b1ead58181 · inbound
EchoingPixels: Aliasing-Resistant Joint Token Reduction for Audio-Visual LLMs SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f701f70d-084b-4191-93fb-9320c4c6b1fa · inbound
Reducing Prompt Sensitivity in LLM-based Speech Recognition Through Learnable Projection SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9e4419e8-9bf8-4396-862e-f2a4021a6b9c · inbound
AUHead: Realistic Emotional Talking Head Generation via Action Units Control SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb44bd26-3154-4e60-aea1-c6402245406d · inbound
RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8875ffd5-8b02-434c-9207-648105f67899 · inbound
VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bcc2be79-4c71-42fa-8a17-0086fe8db9dc · inbound
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 113
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60960240-19b8-47ba-b252-6a48a7d72167 · inbound
PlanRAG-Audio: Planning and Retrieval Augmented Generation for Long-form Audio Understanding SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed66c343-5d3e-4173-8a20-25e1c6f7bd57 · inbound
Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 141
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6754d537-dcf4-4b2d-b544-b7a8c614195d · inbound
Enhancing BEST-RQ Pseudo-Label Quality through Online Refinement for Automatic Speech Recognition SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.