Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T10:35:43.312600Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2605.27976.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T10:35:43.312600Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 851e009d-8c6c-4938-89f4-c8790ec6d149 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Phi-4-mini technical report: Compact yet powerful multimodal language models via mixture-of-loras
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42f812bf-b1e7-4d67-a60d-22cb7fe0b643 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding BLAB: brutally long audio bench
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0620e2f9-239f-487e-be99-74052a52bdf7 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Pyannote.audio: Neural building blocks for speaker diarization
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d78f6d0-4b47-4c62-92db-9c0ad324e38f · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Qwen-audio: Advancing universal audio understanding via unified large-scale audio-language models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7e7acb4-b88c-4c6c-82cf-5ee75273ffbf · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Qwen2-audio technical report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e04493a2-55a0-4500-be36-9c3eb09b70a8 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Sakshi, Oriol Nieto, Ramani Duraiswami, and Dinesh Manocha
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac1407f-ed96-4e3d-bc0b-7ec2e6790ae5 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Sakshi, Jaehyeon Kim, Wei Ping, Rafael Valle, Dinesh Manocha, and Bryan Catanzaro
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e84f7e1-f91d-4b23-ac57-affe6d032fe9 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Audio flamingo 3: Advancing audio intelligence with fully open large audio language models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c809f7b-43f4-402d-84ed-f14f232cec2a · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Audiomarathon: A comprehensive benchmark for long-context audio understanding and efficiency in audio llms
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 569bf9ae-8437-4e5a-bb32-c8e03654e8e6 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Baichuan-omni-1.5 technical report
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f7499b1a-c5a1-4c88-b7c5-01ca3a55ca6e · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Audio flamingo: A novel audio language model with few-shot learning and dialogue abilities
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e077b7d2-3f8b-4bd9-87ac-dc83eec48aa8 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Chronosaudio: A comprehensive long-audio benchmark for evaluating audio-large language models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17384832-b48b-4050-bafa-e35836cdbddb · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding MMAR: A challenging benchmark for deep reasoning in speech, audio, music, and their mix
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f10b74e5-fc46-4133-a5c3-5b3b604c34bd · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding GPT-4o System Card
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9b0012c3-023d-49cf-997a-18b56063d347 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e640e11-715b-4601-9465-099bc6d385c5 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Sakshi, Utkarsh Tyagi, Sonal Kumar, Ashish Seth, Ramaneswaran Selvakumar, Oriol Nieto, Ramani Duraiswami, Sreyan Ghosh, and Dinesh Manocha
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe581de5-a0be-43d4-b6c3-899fc0ca8364 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding SALMONN: towards generic hearing abilities for large language models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86c1bdb6-b4b4-400b-84e1-efcc5e2af339 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Gemini 2.5: Pushing the frontier with advanced reasoning, multimodality, long context, and next generation agentic capabilities
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a7bf7e8-7c82-45dc-9587-e1334619446b · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Moss-audio technical report
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e47dbdf-7f19-412b-b7f8-24016b634eda · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Qwen3.5-omni technical report
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66141ead-8d4f-4b70-a333-6aa7552b0b66 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Qwen3-omni technical report
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 952cb64f-03ac-40c5-8e9d-e348decc72ce · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8f964fd-55a8-4320-8dce-2ae5ecb38e3f · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Mimo-audio: Audio language models are few-shot learners
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94ee65de-4ed3-421b-be04-48b765a604db · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Qwen2.5-omni technical report
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e6fdf44-b206-4cf4-b8ef-80a14c3ea324 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Longspeech: A scalable benchmark for transcription, translation and understanding in long speech
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 484ce664-ced3-42cb-8c16-1b2d548db4c0 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Air-bench: Benchmarking large audio-language models via generative comprehension
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23dffe9d-da5f-4b31-8e18-932863f078be · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b4ea9a28-8e27-4944-a7c5-62dcac7abff5 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Salmonn-omni: A standalone speech LLM without codec injection for full-duplex conversation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bc62ca5-3a29-40f9-b947-1dddcf747746 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding @esa (Ref
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5537e823-2b50-45d2-8b31-25bb367fe59c · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21ae1eb0-6eac-43e8-b9a2-f671c77de499 · outbound
VoiceGiraffe: A Benchmark for Extreme Long-Context Audio-Language Understanding Victoria Beckham
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
No inbound Pith citation observations are available.