Pith. sign in

Paper Citation Record · LEDGER

MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2603.14145.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.14145 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T04:44:31.241489Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T01:19:20.302601Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 724358ec-6603-49ad-870a-5b38603ecda7 · inbound

Omni-DeepSearch: A Benchmark for Audio-Driven Omni-Modal Deep Search cites this paper.

Omni-DeepSearch: A Benchmark for Audio-Driven Omni-Modal Deep Search MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-06-23T03:12:46.666575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:21:52.972107Z digest=sha256:4da6b83aafeffe390368d19b3c2b60e48c160d67666580c3cd89ba4d1f078122

Observation dd41a5e6-ea4f-46ab-aac5-42ed7b9f8941 · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-23T03:12:46.666575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T13:36:44.071188Z digest=sha256:df4f6e19df03d81848fde64063c8701d3deae37ea87acaf45c0f1daa3bba07d6

Observation b8f0d194-da6e-44d1-a777-9fda581f2eb7 · inbound

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction cites this paper.

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-07-04T01:19:20.305457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-04T01:11:42.073993Z digest=sha256:38c4f6be2576b5eb584764e9b53477f77e72ae08366f622cb5935a06ececc49f

Observation a90fcf1b-9876-49bd-8bc0-fba53e39d7d2 · inbound

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers cites this paper.

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:02:34.210757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T18:59:51.362554Z digest=sha256:2418bc5bfb42fd997c30fe2177f2ccd12bf1627ba89098f3f342432061ad98f7

Observation ab674eb8-4067-4913-b5fe-7310d949bc8e · inbound

OmniHalluc-L: Counterfactual Benchmarking and Modality-Perturbation Reliability Calibration for Long-Form Omni Hallucination cites this paper.

OmniHalluc-L: Counterfactual Benchmarking and Modality-Perturbation Reliability Calibration for Long-Form Omni Hallucination MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-02T06:06:41.682559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T07:31:51.780636Z digest=sha256:2209a7232ca694261dbc0ff9c2fbe9b8ae7f05b984ce44057ddf1b85e57266ba

Observation bcaaf31d-f7a1-40a9-8182-09787e190c61 · inbound

OmniFocus: Query-Guided Modality-Balanced Token Compression for Omni-Modal Large Language Models cites this paper.

OmniFocus: Query-Guided Modality-Balanced Token Compression for Omni-Modal Large Language Models MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T05:13:10.430838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:13:10.430838Z digest=sha256:ace0589752cff1d97ec65f1c8f56a1cc28da3d030646daebeb0eba1b2ac59533

Observation c97dc29e-8fc5-4ad4-86e4-5d6e0b14bfbf · inbound

Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos cites this paper.

Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T21:22:42.469251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:22:42.469251Z digest=sha256:4a94b283ba58d0a90b54157b932a5b0cd22aaf45ad5c9e89ecae82d158a79c69

Observation 39b824c6-e3b6-42c1-96db-5e0500105f8f · inbound

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model cites this paper.

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 156

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:14.261099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:14.261099Z digest=sha256:d1b2541543d935b00d21fba83251d169e7a06fbabc19ce5bfafdfddc1d45c915

Observation 16bdf053-cf44-4b4f-9257-a8077c613024 · inbound

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent cites this paper.

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T04:44:31.241489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:44:31.241489Z digest=sha256:e234ed2534914b1155a0ac0bcf6b1c8c31cd0ab992ec59ac3a59c537cc47fef6