Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 36 inbound Pith citation observations for arXiv:2503.03983.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:05:27.317435Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation ad9a24e6-3bfd-487b-8adb-81255a67fd56 · inbound
BLAB: Brutally Long Audio Bench Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87703216-962f-4ed3-be43-2d13eac1b752 · inbound
MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faaf1430-0060-4ed1-99b9-2e4d8b354f35 · inbound
Reducing Object Hallucination in Large Audio-Language Models via Audio-Aware Decoding Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f78fe07d-4bdf-4cf2-8357-5353d91130da · inbound
CMI-Bench: A Comprehensive Benchmark for Evaluating Music Instruction Following Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6eb32c1-0faf-4999-86af-f60a5b54c463 · inbound
AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9daf461b-eab3-46d9-8e7c-7a68f1a3d7a1 · inbound
SLAP: Siamese Language-Audio Pretraining Without Negative Samples for Music Understanding Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2482effc-92ad-4225-a722-cc2ace70b7e1 · inbound
MixAssist: An Audio-Language Dataset for Co-Creative AI Assistance in Music Mixing Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc18c3e4-6860-40e5-82f7-e582062a4df1 · inbound
Step-Audio 2 Technical Report Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 790b06dd-25e5-44b1-8942-25f3c8f43102 · inbound
BoSS: Beyond-Semantic Speech Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce6d2a7f-0377-48f8-9a32-3893dc6f4dcb · inbound
LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ca8e5ab-f5dc-433b-90a8-29a0b028f2ad · inbound
WoW-Bench: Evaluating Fine-Grained Acoustic Perception in Audio-Language Models via Marine Mammal Vocalizations Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da3cfc8-93d9-40e5-9552-ac757a222577 · inbound
FastSLM: Hierarchical Temporal Abstraction for Efficient Long-Form Speech Adaptation Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e029be1-4592-4b6f-a584-2dbfc5d5dded · inbound
OmniFysics: Towards Physical Intelligence Evolution via Omni-Modal Signal Processing and Network Optimization Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2652559e-6df6-4a18-abe9-e3b7c8d37e5c · inbound
SpotSound: Enhancing Large Audio-Language Models with Fine-Grained Temporal Grounding Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 59f72fc5-5d63-4c36-8b5c-8dbe497526bb · inbound
Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bf4b9018-0016-4edd-9b2e-362ec6250864 · inbound
HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 79924e11-1dd8-426b-ba6e-5a07000b7e4f · inbound
TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8016a153-88ae-481b-8b36-36e61735ea75 · inbound
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 129
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e8ae7566-cd02-4b84-b8da-3f3f3ebf7645 · inbound
Escape the Language Prior: Mitigating Late-Stage Modality Collapse in Audio Reasoning via Modality-Aware Policy Optimization Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d9c56024-a811-499c-bbbc-f8ed687dbe10 · inbound
Audio Interaction Model Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d46a9922-240c-492b-a61c-4259fec8c112 · inbound
TinyGiantALM: A Compact Audio-Language Model for Intent-Aware Reasoning under Resource Constraints Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3956a836-13a2-4a95-a7af-7bd398394a8c · inbound
Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 178befd7-2047-465e-a124-f93262f31937 · inbound
RAIL: Rethinking Auditory Intelligence in Large Audio-Language Models with a CHC-Grounded Benchmark Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a152d5c4-cb10-4a61-a202-8841013b14ac · inbound
Continuous Audio Thinking for Large Audio Language Models Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 24985778-cf4a-4593-9557-8599f6dd6b6d · inbound
The Hidden Evolution of Disguised Visual Context inside the VLM Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b147afc1-d365-467a-aac4-435a39917103 · inbound
RFM-Editing 2: Text-Guided Audio Editing with Rectified Flow Matching and Coarse-to-Fine Diffusion Transformers Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 42d05f29-ac5f-4c73-a39c-3b055e3b27a1 · inbound
RFM-Editing 2: Text-Guided Audio Editing with Rectified Flow Matching and Coarse-to-Fine Diffusion Transformers Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 85c4457d-8f24-4970-8c7b-1551134ddaa4 · inbound
RFM-Editing 2: Text-Guided Audio Editing with Rectified Flow Matching and Coarse-to-Fine Diffusion Transformers Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6db14df-ebe8-4c4c-8a2c-7ec07b92283f · inbound
ALM2Vec: Learning Audio Embeddings for Universal Audio Retrieval with Large Audio-Language Models Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a17a7f5c-b979-4e1d-bbc6-f5190ad4fc7f · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 87ae1241-5a59-47df-934c-0d21baebe401 · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea274fe0-9202-4966-b8b5-ad83df298189 · inbound
Empowering Long-form Omni-modal Understanding with Robust Audio Perception Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ada37a26-5187-4f53-b72a-338eaae07299 · inbound
X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6893a8c-bff4-45e4-9398-5ea127d52293 · inbound
Weak-to-Strong On-Policy Distillation Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed269b75-4fc8-4ee2-8213-14e5013f8b44 · inbound
Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6654286-5e25-4ae2-b20f-3e67995c7063 · inbound
VoxZip: Semantic-Anchored Temporal KV Cache Compression for Long-Context Audio Inference Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.