Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T09:59:40.094302Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 2 inbound Pith citation observations for arXiv:2607.27109.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T09:59:40.094302Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T09:59:37.074456Z
A source-named dated measurement, never combined with another source.
Source: cited_works
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4df5a8f2-4c4e-4e81-9671-28d979ddf88d · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning With the development of AudioLLMs [1, 2, 3, 4], captions are moving from brief descriptions to more open-ended and fine- grained audio understanding
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33dc2a45-5619-480f-ac5d-a244681dbbe5 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ebe64a4-17a9-4994-8ffb-c20251a3a30d · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Describe this audio in detail
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457195ae-7de4-477f-90e5-e6c37cf3068f · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Fine-grained dimensions are first averaged within each capability category
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da110d05-a343-443d-accd-34f524c2101f · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning MMAC comprises 6 capability categories and 15 fine-grained dimensions and evaluates the coverage and cor- rectness of target information in free-form captions
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4538e26d-3de4-4ec6-a8ed-3361f2c1bed4 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Qwen2.5-Omni Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9cb9266-893e-485d-81c3-4b35982de869 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Midashenglm: Efficient audio understanding with general audio captions,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a66660a1-e030-4b13-8f54-1c65590f6c64 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Qwen3.5-Omni Technical Report
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01bdb891-2d6c-4875-bf15-5b1f57c14925 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Step-Audio 2 Technical Report
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ca77e76-f96c-4760-9395-668feb1d181b · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Wavcaps: A chatgpt-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc00fe66-ec57-4814-93cc-77dd1b99c26f · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Audiocaps: Generating captions for audios in the wild,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0a623c7-9f0c-452c-b461-77f1938847e5 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Clotho: An audio captioning dataset,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d54386a7-ab99-4a16-9be1-0c793ecfb129 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Mmau: A massive multi-task audio understanding and reasoning bench- mark,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fd35f9f-6b1c-412c-88b8-c60e5a01d0c1 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Mmar: A chal- lenging benchmark for deep reasoning in speech, audio, music, and their mix,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32a8f721-39f8-4df6-8a9d-2f322c4f6eb2 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Omni- captioner: Data pipeline, models, and benchmark for omni de- tailed perception,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15e8ee66-d3cb-4d2d-85ab-0b6bb9d42045 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50680b97-8baf-4b2d-a749-9cf712a1c33b · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Secap: Speech emotion captioning with large language model,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97cf3800-31bf-44e1-9a05-1fa0d8b901e3 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning MMSU: A massive multi-task spoken language understanding and reason- ing benchmark,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c693153-dad8-4601-af31-ad398c47d5ec · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f63db418-9cfa-4023-ada0-145a761ef7d2 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning speechocean762: An open-source non-native english speech corpus for pronuncia- tion assessment,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 415f8c7a-8eaa-4b13-a064-94b4dd13eb7a · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Childman- darin: A comprehensive mandarin speech dataset for young children aged 3-5,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87861c04-79c7-4f37-8865-a0c9a206f123 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Seniortalk: A chinese conversation dataset with rich annotations for super- aged seniors,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 520b7c69-5b58-4530-9bc9-a26e1c276a85 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Gi- gaspeech: An evolving, multi-domain asr corpus with 10,000 hours of transcribed audio,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7436b06f-52c5-4b9e-8c40-47e9d1c11b32 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Recent ad- vances in speech language models: A survey,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 360e4d38-1518-40b6-8b6a-fa193003d94a · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Audio set: An on- tology and human-labeled dataset for audio events,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d605888a-3331-46f1-aeb1-23d53d3ffba0 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Esc: Dataset for environmental sound classi- fication,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 176197ea-f851-4f4e-87f1-1cb4f261c4e7 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Fsd50k: an open dataset of human-labeled sound events,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5547dc01-76ca-49ba-a192-4d64127443e2 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Kespeech: An open source speech dataset of mandarin and its eight sub- dialects,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ec54b8f-1607-4a82-8cbe-ac8cc6ab211c · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Qwen3-Omni Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c05541b1-a981-4e64-b355-7a994614cc3a · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Coig-cqia: Qual- ity is all you need for chinese instruction fine-tuning,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcbc62df-6379-4001-b08f-c0c8a0356ae0 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Mustard: Mastering uniform synthesis of theorem and proof data,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a062c1b-e461-4751-9cf6-359854244a11 · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning Indextts2: A break- through in emotionally expressive and duration-controlled auto-regressive zero-shot text-to-speech,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a34e2e1-d804-4004-a826-4dbfd3230fec · outbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning The problem ofm rankings,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a26f578e-b308-4672-ac91-6103ded435d9 · inbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33dc2a45-5619-480f-ac5d-a244681dbbe5 · inbound
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.