Pith. sign in

Paper Citation Record · LEDGER

MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2508.13992.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.13992 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T10:13:42.305221Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3c050704-71de-49be-ad09-3bd994b9a4db · inbound

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents cites this paper.

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T10:13:42.305221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:13:42.305221Z digest=sha256:bc719bc8a57f0fbcdf537de6135a1acb36518198a42a5b24ccba0c842c083d8b

Observation 630738de-3364-4576-8c49-4a3858100d7a · inbound

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering cites this paper.

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T19:38:23.429787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T19:38:23.429787Z digest=sha256:ce5d8ed6e265e2a4caa9c10a71b066a07ee7abe874e28d1d8e8c36c38455c829

Observation 62228609-ff36-4d97-b1fe-632e0f933498 · inbound

Generating Synthetic Doctor-Patient Conversations for Long-form Audio Summarization cites this paper.

Generating Synthetic Doctor-Patient Conversations for Long-form Audio Summarization MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:10:50.371597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:41:23.645440Z digest=sha256:29abd3dc01449df937741fbb8c53763e99d03f47c3018660dd68da0a402c084e

Observation 2c85a7f3-9755-430e-a390-5a9abadc296b · inbound

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering cites this paper.

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:11:01.470349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:45:51.528645Z digest=sha256:5b98a786ef6fb26fbee749c3188fd86214e9b743ae2a2beb981bdcca8832638e

Observation f7cc0b70-022f-4ca1-a213-d6822ddf94b9 · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:28.269718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T14:10:03.707886Z digest=sha256:fa5717b0843afe93537edb1040fed289011e4b9db31406027255e0b5a657f563

Observation 5d451620-3915-40e8-a784-61bf883bc12c · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T21:18:46.566338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:18:46.566338Z digest=sha256:31f9e9846febbdf176ddfc2e1a998a319d483da0878b86aec828e7d98078e48d

Observation 8eaed678-2693-4dd2-91fc-152624e51389 · inbound

VoxSafeBench: Not Just What Is Said, but Who, How, and Where cites this paper.

VoxSafeBench: Not Just What Is Said, but Who, How, and Where MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:24:22.023192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T10:19:28.041282Z digest=sha256:3b41ece8150dcae80de24199ba76f790ade6c5f25b85b3c50c4cac63d58ac47f

Observation a6acffad-05ba-4093-9044-286feb5b7a2b · inbound

HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models cites this paper.

HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:14.474889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T05:20:30.823304Z digest=sha256:2704fede2c32b3a034b6f71f779dbe627ba8ed0d00080c548ea5145b9e5670f8

Observation 2c35ff54-432d-418c-b216-dc49fbaf54a3 · inbound

All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation cites this paper.

All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:16:36.228584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T17:39:38.234052Z digest=sha256:5b5aab434479fded3495d65d23aa2c243b12b1f1f001cb2f2098e1ca9e6ccdc6

Observation 51ff07d7-eeee-47dc-ac61-93f1c282e891 · inbound

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models cites this paper.

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:41:26.485335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T14:29:18.348031Z digest=sha256:e572a426a2118f32f1ebbbe4638a1b1f1aa5495b591a8d211e994846951c013c

Observation f82d1dc7-c589-4291-9b7c-7d3c7c483e3a · inbound

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio cites this paper.

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T20:17:04.813458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T18:15:54.497023Z digest=sha256:5d7692804a414cd4028f72bf428fd749bf553741ffa8c7e87787f0037cf33a08

Observation d3485401-3041-438b-91fc-3b663a6dcae9 · inbound

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio cites this paper.

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T07:55:30.217657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T07:50:29.976958Z digest=sha256:73804e8e0b5e409d77413cb9846713f5d1f7c75f211bd0269d83580e63e910f6

Observation 7bdcc89a-5dfc-409e-8b91-172ad3794571 · inbound

Benchmarking LLMs on the Massive Sound Embedding Benchmark (MSEB) cites this paper.

Benchmarking LLMs on the Massive Sound Embedding Benchmark (MSEB) MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:46:07.386028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T17:15:02.257046Z digest=sha256:c8767e66138592020aa722ab0e82bd39fe5655f569be924c2526d9fe596a4633

Observation df12a7d6-0bc4-40fa-9432-44a744bc39f7 · inbound

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook cites this paper.

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 195

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:39:48.882480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T07:38:23.099479Z digest=sha256:8b7deda9dcf0ca31ee0aaf699ee0b2c34873e8029159f68899454f93985061aa

Observation 0f4e6d22-4ce6-4acd-93b9-ee1fdeeaae7a · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:09:24.456660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:3498b7da9c01c4b5f8fb4cb9d1a2984c401484ab64177a49a77f29a6941f4d9a

Observation 01b27112-de83-423d-b388-3ec646801966 · inbound

Raon-Speech Technical Report cites this paper.

Raon-Speech Technical Report MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-13T08:22:44.347941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T08:22:44.347941Z digest=sha256:ed76f92fb418eb75e5ad014e9c975db22399ea6d79198accb1be33e7c4aea92c

Observation 09ccbcf8-05eb-4c71-b591-bc88a95f6dd3 · inbound

Escape the Language Prior: Mitigating Late-Stage Modality Collapse in Audio Reasoning via Modality-Aware Policy Optimization cites this paper.

Escape the Language Prior: Mitigating Late-Stage Modality Collapse in Audio Reasoning via Modality-Aware Policy Optimization MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:53:47.588897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T17:45:10.339950Z digest=sha256:1277f27c63e9ae75d8b903f1b36540b6531eaeda40148831029c589b49f71241

Observation 107c5050-9c44-47d3-9984-f432ec791da5 · inbound

Audio-Mind: An Auditable Agentic Framework for Audio Understanding cites this paper.

Audio-Mind: An Auditable Agentic Framework for Audio Understanding MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-29T10:13:17.567135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-29T10:03:54.653164Z digest=sha256:8015f6e3460a1df38cc3e565b421d1c674ce5295d3d512a48827123480f23c61

Observation d7bcb2bb-69a6-48ba-85d5-6edbed107792 · inbound

MOSS-Audio Technical Report cites this paper.

MOSS-Audio Technical Report MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T00:56:25.050583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T13:05:29.813707Z digest=sha256:3283f04f7c59dd6e8ba2bfb664b479e1569e8982452f9f9e0f301c2a429dac97

Observation b2a49e64-eb0b-47c9-a1cf-8e50004d6314 · inbound

VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track cites this paper.

VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:07:21.319265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T21:02:19.300441Z digest=sha256:b2fa3d989b5fb8597938a60a593dae2e1219ce0c795e34ecad85814ff2390c3b

Observation c9bcd5b7-c3dc-42bb-bed5-7161c73f42ea · inbound

A Closer Look at Failure Modes in Temporal Understanding of Large Audio-Language Models cites this paper.

A Closer Look at Failure Modes in Temporal Understanding of Large Audio-Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:39:01.662599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T23:25:46.349380Z digest=sha256:7d7c74049172b677cbaaf825c866dcb612e856b209ae3f89724fe8838a5e23e7

Observation bdac4518-886a-45ec-b611-4374ad04782f · inbound

Comparative Reasoning: Making an Audio Language Model Better at Comparing Emotions cites this paper.

Comparative Reasoning: Making an Audio Language Model Better at Comparing Emotions MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:00:00.285874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-25T23:18:55.383851Z digest=sha256:fd2d6b7990d3281b9aa8459a4530728afce69e7db456c3da9ffc4d2a3c58f5be

Observation bddd3957-e5ba-4efb-827e-f0128e321565 · inbound

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models cites this paper.

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T20:10:07.952661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-25T20:36:24.901454Z digest=sha256:1b28e6ea281f1fd7148adf4d65563b0f1f202bcbd2abffd3df968bbe53d1876a

Observation 244e041f-d3f7-43e2-9324-10d11def72b9 · inbound

Hearing Like Humans? Sound Symbolism and Perceptual Alignment in Speech Language Models cites this paper.

Hearing Like Humans? Sound Symbolism and Perceptual Alignment in Speech Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-14T13:49:26.320341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T13:49:26.320341Z digest=sha256:e69f4838e28bd44fa3a730d3a1182ae839905c7681a5009d29280db1ca7875ed

Observation 6725413c-ccf2-44a3-92cf-440a8c763c73 · inbound

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models cites this paper.

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T05:20:02.638568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:20:02.638568Z digest=sha256:7afacd8629c5990f952c3c620c35fae6a4cee43923e07c68dd8f51eb1a935ece

Observation db2dece3-a474-4384-933b-90fc9997a4e1 · inbound

Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning cites this paper.

Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T10:40:25.893951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:40:25.893951Z digest=sha256:39032704aac1be25df4195bb66d5cc47a029e9b5010f9f5457ae7c99c797e3a2