Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2503.11197.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:13.461419Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T20:07:21.344787Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9f43c0ce-d66c-4ef9-a17f-0fc807f4c018 · inbound
From System 1 to System 2: A Survey of Reasoning Large Language Models Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 292
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7b9ad463-7fbf-40ee-8a5f-8390c59237f0 · inbound
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02f72257-f736-4ef4-9ac9-4788cfbaf892 · inbound
Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 08b7bb0e-fa30-455a-9a89-729280f973ba · inbound
Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 171
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a3b6bae-50c7-4904-b015-a960435a0c0e · inbound
Group Relative Policy Optimization for Speech Recognition Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b56048b-e90a-4405-b888-06f9d163bd80 · inbound
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 262
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 513d9acc-d37d-431a-9187-d31604016403 · inbound
AQA-TTRL: Self-Adaptation in Audio Question Answering with Test-Time Reinforcement Learning Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e30c90c-79f5-444e-a87c-a4d8b13e6018 · inbound
Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f292c1d-1083-4035-b4a8-109329d1dec2 · inbound
LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 946b3ddd-5e8a-452d-8d87-9c6b7ded12a0 · inbound
VERITAS: A Multi-Agent Co-Scientist for Verifiable Image-Derived Hypothesis Testing Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8293b213-4cda-4579-b637-2a6d80e6fd61 · inbound
Why Your Tokenizer Fails in Information Fusion: A Timing-Aware Pre-Quantization Fusion for Video-Enhanced Audio Tokenization Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aceb0ce8-818d-4344-8794-de122493689c · inbound
TinyMU: A Compact Audio-Language Model for Music Understanding Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e4d01878-c041-45ed-a53e-f627b8b95a83 · inbound
Audio-DeepThinker: Progressive Reasoning-Aware Reinforcement Learning for High-Quality Chain-of-Thought Emergence in Audio Language Models Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 044e02f1-1000-4e55-9cbb-922572d725cf · inbound
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cf1683a5-dbf9-4766-9b64-1d903fbb5bdf · inbound
Beyond Content: A Comprehensive Speech Toxicity Dataset and Detection Framework Incorporating Paralinguistic Cues Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7e39656b-afa1-4cab-afec-3ed85f069a99 · inbound
A Survey of Audio Reasoning in Multimodal Foundation Models Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ff16a0b4-21dd-4882-99d2-2b437f8372ca · inbound
Learning When to Think While Listening in Large Audio-Language Models Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4399527a-8e7c-4c25-82d8-6525afc60902 · inbound
LaSR: Context-Aware Speech Recognition via Latent Reasoning Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 908f10dc-06eb-4515-9197-97752fbe235a · inbound
FSA-GRPO: Teaching Auditory LLMs to Use Few-shot Demonstrations Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7cfff068-1d25-4f50-8ef7-0f8da95b5d03 · inbound
VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1b580945-e9cd-485e-94c7-0b3bb827eaf8 · inbound
Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5924d362-6311-4e55-906c-f7e5c3a8e008 · inbound
Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18352f4d-a462-4945-a5a9-133ff75fc6eb · inbound
X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f8e5bc0-3f3e-42a5-aa65-05b8214630aa · inbound
Weak-to-Strong On-Policy Distillation Reinforcement Learning Outperforms Supervised Fine-Tuning: A Case Study on Audio Question Answering
Reference 129
Source-reported events for the cited work
Unavailable: canonical work link unavailable.