Pith. sign in

Paper Citation Record · LEDGER

CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2401.11944.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.11944 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:19:05.167967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T04:32:32.672779Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 623d5a10-821a-485a-a79f-434c82c4544c · inbound

DeepSeek-VL: Towards Real-World Vision-Language Understanding cites this paper.

DeepSeek-VL: Towards Real-World Vision-Language Understanding CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:58:54.760355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T17:58:54.177359Z digest=sha256:4f8ea0b05e5b9b5bbfde7373f776e068f7fc043402ed9298521b02585dc797b3

Observation e18477da-9ef8-4c40-889f-ce3c5099ed29 · inbound

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning cites this paper.

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 131

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:33:26.885629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T20:33:26.613927Z digest=sha256:bca5841ffdc7ad3abc0039b2901ef24348b0efe63d81a40f97e9774433d81d52

Observation 68fc532c-d9aa-4e30-9e14-86a8f87eb304 · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 247

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:32.676271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:73da42fde00314187f6be8885baaae8925faa7715ec2316de3a7aa30166c1d79

Observation 8dd42996-6022-4616-9e13-33f11b1ae93e · inbound

FlagEvalMM: A Flexible Framework for Comprehensive Multimodal Model Evaluation cites this paper.

FlagEvalMM: A Flexible Framework for Comprehensive Multimodal Model Evaluation CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:44.959144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:44.959144Z digest=sha256:d6edd5f3f4b1de2359fd6d6c552496d030a498ebe8a5a7d9080f43007fc908cf

Observation 5aff777f-413d-4142-9836-e8741bc8053b · inbound

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos cites this paper.

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T04:22:56.098198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:22:56.098198Z digest=sha256:7ee03ab582f32e880cc5e631b8e4074045259e98c61b3ebca8859eb73ddecf8c

Observation b1b73b5f-c11d-462a-87d9-feb48b3446d9 · inbound

Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? cites this paper.

Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:19:05.167967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:19:05.167967Z digest=sha256:7e25caa2afc39a9bcbba635d927719b6d8dd493818e27ee747006a8e6ec43800

Observation ed9c1492-3022-4bba-a86d-b40b58a226c2 · inbound

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique cites this paper.

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:04:55.471250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:04:55.471250Z digest=sha256:83f7398d9a69597ff29c515fcdb3472c969e6b6a9d3297396c3dc714021658b8

Observation 4666b248-d71e-4a20-91fa-9d307d542e3f · inbound

From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping cites this paper.

From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:25:59.778943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:11:50.265856Z digest=sha256:7e8aabb4fb048d5eea4a5ba2244766191eea99eb18cc84053fe6e5515ad266bb

Observation 6b981cda-3a3a-4f1d-aa23-b32c0d90ebe5 · inbound

OProver: A Unified Framework for Agentic Formal Theorem Proving cites this paper.

OProver: A Unified Framework for Agentic Formal Theorem Proving CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:48:23.548793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-20T14:43:46.517807Z digest=sha256:fdd89a6d31e4e06c6d19664044f98e5c851f16a1b0dee427f2a889cabbf75564