Pith. sign in

Paper Citation Record · LEDGER

MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2508.13992.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.13992 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T10:13:42.305221Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3c050704-71de-49be-ad09-3bd994b9a4db · inbound

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents cites this paper.

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T10:13:42.305221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:13:42.305221Z digest=sha256:f6eac92dc5122c58514f5eeeec22d102ec863aeefcc12097a6592a0e0ee624df

Observation 630738de-3364-4576-8c49-4a3858100d7a · inbound

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering cites this paper.

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T19:38:23.429787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T19:38:23.429787Z digest=sha256:9f40f1a25d4c8616f9a52abd50b365613c64c7efc9b5dd4e633df2f1a9e0da85

Observation 62228609-ff36-4d97-b1fe-632e0f933498 · inbound

Generating Synthetic Doctor-Patient Conversations for Long-form Audio Summarization cites this paper.

Generating Synthetic Doctor-Patient Conversations for Long-form Audio Summarization MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:10:50.371597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T18:41:23.645440Z digest=sha256:ecf120d560117d49d36ed1ba281cc88a435d29a915330084dc68a142b58a798f

Observation 2c85a7f3-9755-430e-a390-5a9abadc296b · inbound

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering cites this paper.

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:11:01.470349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:45:51.528645Z digest=sha256:c241181c7404bcebd3bea7c7ed62303b97389f9f8d6a600e6781ea94dd6900ea

Observation f7cc0b70-022f-4ca1-a213-d6822ddf94b9 · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:28.269718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T14:10:03.707886Z digest=sha256:cb54ea72da368026396fb0c9749aa5e39dd920a9ac2dac8bdd77391bf38cd978

Observation 5d451620-3915-40e8-a784-61bf883bc12c · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T21:18:46.566338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:18:46.566338Z digest=sha256:22746b14c5d4d553a62089ec07ecbfb6b37d8ed54fdb92ad722f4ebd87724c16

Observation 8eaed678-2693-4dd2-91fc-152624e51389 · inbound

VoxSafeBench: Not Just What Is Said, but Who, How, and Where cites this paper.

VoxSafeBench: Not Just What Is Said, but Who, How, and Where MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:24:22.023192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T10:19:28.041282Z digest=sha256:ae38f9c54203b15dfb2c272f763495b8ad87f34333d7934e5574d17ff05e2553

Observation a6acffad-05ba-4093-9044-286feb5b7a2b · inbound

HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models cites this paper.

HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:14.474889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T05:20:30.823304Z digest=sha256:65d4a2f08ae14eb587c6938765b5b3f8aca3bd79f1ae217049478bd76abc53db

Observation 2c35ff54-432d-418c-b216-dc49fbaf54a3 · inbound

All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation cites this paper.

All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:16:36.228584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T17:39:38.234052Z digest=sha256:59a3517d016212aa6d45662bf6bdbc41f36804347193eb3a922134780e3883a4

Observation 51ff07d7-eeee-47dc-ac61-93f1c282e891 · inbound

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models cites this paper.

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:41:26.485335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T14:29:18.348031Z digest=sha256:fa7448fc4f2d8722a0582d5cedd144d0b45762e4f30958768338b7cb228f4bb8

Observation f82d1dc7-c589-4291-9b7c-7d3c7c483e3a · inbound

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio cites this paper.

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T20:17:04.813458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T18:15:54.497023Z digest=sha256:a20e1d06f329efbec41f74da6369db44d05d2665b1665bfca398d87a83098432

Observation d3485401-3041-438b-91fc-3b663a6dcae9 · inbound

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio cites this paper.

MedMosaic: A Challenging Large Scale Benchmark of Diverse Medical Audio MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T07:55:30.217657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-01T07:50:29.976958Z digest=sha256:108d39908ee036f4d8afe3c1f6e22a7d947a8ed519e0a3b8107696dff15c8850

Observation 7bdcc89a-5dfc-409e-8b91-172ad3794571 · inbound

Benchmarking LLMs on the Massive Sound Embedding Benchmark (MSEB) cites this paper.

Benchmarking LLMs on the Massive Sound Embedding Benchmark (MSEB) MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:46:07.386028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:15:02.257046Z digest=sha256:ade367a107377f420e01a183081f0de9e2f8fab7fac854724bdb7184eb52c3fc

Observation df12a7d6-0bc4-40fa-9432-44a744bc39f7 · inbound

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook cites this paper.

A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 195

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:39:48.882480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T07:38:23.099479Z digest=sha256:f2487a03192fc276c80e4609b282133e4b4c14bc817799a540bbf848413d3e54

Observation 0f4e6d22-4ce6-4acd-93b9-ee1fdeeaae7a · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:09:24.456660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:3ddd45b18e07226ac131084af324ad0e0f6013fee0fda448f279a458d947ec49

Observation 01b27112-de83-423d-b388-3ec646801966 · inbound

Raon-Speech Technical Report cites this paper.

Raon-Speech Technical Report MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-13T08:22:44.347941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T08:22:44.347941Z digest=sha256:f772a3e71b7852dac87e505da7a334f9991592bf72f4555504777c55facabe35

Observation 09ccbcf8-05eb-4c71-b591-bc88a95f6dd3 · inbound

Escape the Language Prior: Mitigating Late-Stage Modality Collapse in Audio Reasoning via Modality-Aware Policy Optimization cites this paper.

Escape the Language Prior: Mitigating Late-Stage Modality Collapse in Audio Reasoning via Modality-Aware Policy Optimization MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:53:47.588897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T17:45:10.339950Z digest=sha256:3ea04c39138ea9197c5891aba9ab0edb8451ff93677af8625485ae5b57fe0372

Observation 107c5050-9c44-47d3-9984-f432ec791da5 · inbound

Audio-Mind: An Auditable Agentic Framework for Audio Understanding cites this paper.

Audio-Mind: An Auditable Agentic Framework for Audio Understanding MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-29T10:13:17.567135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-29T10:03:54.653164Z digest=sha256:93805cf06a233f50fb40cd92986bc45672403a5d7fe931e40aed0c120f7accc0

Observation d7bcb2bb-69a6-48ba-85d5-6edbed107792 · inbound

MOSS-Audio Technical Report cites this paper.

MOSS-Audio Technical Report MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T00:56:25.050583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T13:05:29.813707Z digest=sha256:fd80924dec8529b075e7bcc8018ca95b4dcb9f58aa738419070e2bbe3517e7f5

Observation b2a49e64-eb0b-47c9-a1cf-8e50004d6314 · inbound

VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track cites this paper.

VISA: A Visual Information Strengthened Audio-Reasoning System for the Interspeech 2026 ARC Agent Track MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:07:21.319265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T21:02:19.300441Z digest=sha256:bfad1221cfc87e8a603bb22bd5449faae8848831ce778ab81c76b9c278f009c2

Observation c9bcd5b7-c3dc-42bb-bed5-7161c73f42ea · inbound

A Closer Look at Failure Modes in Temporal Understanding of Large Audio-Language Models cites this paper.

A Closer Look at Failure Modes in Temporal Understanding of Large Audio-Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:39:01.662599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T23:25:46.349380Z digest=sha256:99c8d4de12470aeed9ba56b533677fae0e0e801375f203bfb1926e523bcaeede

Observation bdac4518-886a-45ec-b611-4374ad04782f · inbound

Comparative Reasoning: Making an Audio Language Model Better at Comparing Emotions cites this paper.

Comparative Reasoning: Making an Audio Language Model Better at Comparing Emotions MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:00:00.285874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-25T23:18:55.383851Z digest=sha256:7e39087aeba97c8ac487ce304233d30756f10ff72f0fbf1917fa83beaaea046d

Observation bddd3957-e5ba-4efb-827e-f0128e321565 · inbound

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models cites this paper.

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T20:10:07.952661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-25T20:36:24.901454Z digest=sha256:44c638afa46bc970c82734c7f3db23abe4b21d7b5687aba539289f95f830ffc8

Observation 244e041f-d3f7-43e2-9324-10d11def72b9 · inbound

Hearing Like Humans? Sound Symbolism and Perceptual Alignment in Speech Language Models cites this paper.

Hearing Like Humans? Sound Symbolism and Perceptual Alignment in Speech Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-14T13:49:26.320341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T13:49:26.320341Z digest=sha256:3b8e73253cd17573ffc7f9c47269afffffd894e5fc6d79bb9d4e492edfe7c7ab

Observation 6725413c-ccf2-44a3-92cf-440a8c763c73 · inbound

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models cites this paper.

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T05:20:02.638568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:20:02.638568Z digest=sha256:403bb8a7144222b266146b7c41bc144cffe0a5d8dbf49cc0cc5cc85faeba511a

Observation db2dece3-a474-4384-933b-90fc9997a4e1 · inbound

Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning cites this paper.

Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T10:40:25.893951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:40:25.893951Z digest=sha256:e712735c091f09feaa7dd523b07ca549a25caeec6e8034a3a44fdede358b80d0