Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2505.13032.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T19:02:41.013510Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation c0ea1ff9-b90e-4dea-b21d-9771092d2392 · inbound
MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 152eed71-91e6-42da-9dec-b4226d8ff683 · inbound
When Silence Matters: The Impact of Irrelevant Audio on Text Reasoning in Large Audio-Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation aa23bad6-c2cb-4114-9261-b27d6cec8f14 · inbound
Robustness assessment of large audio language models in multiple-choice evaluation MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f721397d-e7b9-4569-98c5-5247d43af320 · inbound
AQA-TTRL: Self-Adaptation in Audio Question Answering with Test-Time Reinforcement Learning MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79175ecb-6a78-446b-aaad-b157ba509d01 · inbound
VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c171446b-c51c-4493-9a76-50842f872cfe · inbound
Revisiting Audio-language Pretraining for Learning General-purpose Audio Representation MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0641492d-ef98-48b5-a9d1-1df261d82689 · inbound
ORCA: Open-ended Response Correctness Assessment for Audio Question Answering MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 474c89ad-ff40-4a63-afa2-9db29c20e7f5 · inbound
Omni2Sound: Towards Unified Video-Text-to-Audio Generation MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6ac55984-e2be-4519-9b7a-c4ef584c94d0 · inbound
AQUA-Bench: Beyond Finding Answers to Knowing When There Are None in Audio Question Answering MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8471acb8-d7c3-46b0-a26b-e8b30fd7cbc5 · inbound
OmniFysics: Towards Physical Intelligence Evolution via Omni-Modal Signal Processing and Network Optimization MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5c781189-f558-491a-a1d5-8b5e42eafa9b · inbound
OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 78be80f0-2027-42b8-a2c5-09f9b1782fdd · inbound
Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bd1e90bb-740a-4a33-983b-3a7427757246 · inbound
Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14ffd3cf-7f30-4117-b8fc-8452e52c17f8 · inbound
Qwen3.5-Omni Technical Report MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d36301f1-b210-4179-b7d9-79d79e33735f · inbound
Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4a92f963-fa4a-4ffd-bede-1c716414a507 · inbound
Task-Aware Answer Preservation under Audio Compression for Large Audio Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cc39da04-14be-4fa5-b2a3-352957fc7b32 · inbound
Towards Fine-Grained Multi-Dimensional Speech Understanding: Data Pipeline, Benchmark, and Model MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 91a4b310-96d0-4f1a-88dd-ba250d6a20d8 · inbound
ViMU: Benchmarking Video Metaphorical Understanding MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dff1ec23-b6eb-4dae-99cf-683e73bdfa0e · inbound
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 184
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ed348d9c-01ba-4551-b5f0-89b367c0933b · inbound
A Survey of Audio Reasoning in Multimodal Foundation Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 121
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3946f81a-ca8f-4bb7-9500-2be86240a790 · inbound
EchoDistill:Alignment Noisy-to-Clean Self-Distillation for Robust Audio LLMs MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 106a13cf-e029-45a9-8dbc-0b8c15f6618b · inbound
Escape the Language Prior: Mitigating Late-Stage Modality Collapse in Audio Reasoning via Modality-Aware Policy Optimization MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f915d858-d9f7-49f3-9296-81babc305094 · inbound
Audio-Mind: An Auditable Agentic Framework for Audio Understanding MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 69d7a8f7-47e5-4d33-b6c2-5d1e157bf6b8 · inbound
A Unified and Reproducible Experimentation Framework for Speech Understanding MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 09b4c3d6-a41a-4c77-866a-a0fc1bb40cae · inbound
PolySpeech-100: A Large-Scale Benchmark for Speech Understanding Across 100+ Languages and Dialects MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c1785c1d-8c56-4343-a04d-f990d3fcfe49 · inbound
MOSS-Audio Technical Report MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 332d5daa-7dc9-4bb5-82d2-4df03117a0d1 · inbound
VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 999708c9-b3b1-45c6-9cc8-a69c06fd7859 · inbound
TinyGiantALM: A Compact Audio-Language Model for Intent-Aware Reasoning under Resource Constraints MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ecd17084-6b4e-461d-8879-8a95dce18a54 · inbound
RAIL: Rethinking Auditory Intelligence in Large Audio-Language Models with a CHC-Grounded Benchmark MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d864592e-ca3e-4336-b532-a45832ade299 · inbound
A Closer Look at Failure Modes in Temporal Understanding of Large Audio-Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1d713cd4-f323-4f0e-af45-1d2104b5379d · inbound
Continuous Audio Thinking for Large Audio Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5623885a-4d30-45b6-94ec-e4999e929b02 · inbound
From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4c716412-4af6-47e6-85a1-ce211034605b · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7a251b6c-e59c-4191-87b0-26578b22e099 · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7186ecf6-6145-494f-9dab-65ba9f74b0a5 · inbound
Empowering Long-form Omni-modal Understanding with Robust Audio Perception MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 845f6050-4f64-412b-ab73-8ac9afaa3b9a · inbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee5013f1-6efc-402a-b10f-21f3f7de974c · inbound
Summary of DCASE 2026 Task 5: Audio-Dependent Question Answering MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79395fd8-cb8f-4eb4-a852-a748d76f965b · inbound
X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 212
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e2e166e-f451-4bf5-9455-58dc409ab545 · inbound
Weak-to-Strong On-Policy Distillation MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 143
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e2a3dbc-9498-453c-a23b-34a3e2ccf11a · inbound
Hear, Invoke, and Understand: A Skill-Calling Multimodal Agent for Large Audio Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.