Pith. sign in

Paper Citation Record · LEDGER

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models

As of 22 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2508.06009.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06009 v1

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:03:10.178730Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

70 of 70 outbound references displayed

  • verified exact2
  • verified fuzzy3
  • unresolved65
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a3f53cd3-5ee1-46be-8659-9225c646233f · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.136528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.479730Z digest=sha256:8bdf38d75bd9be6ad9ffc39d2ce2cf38471647b84f56b70c77870a5b622b8e1b

Observation b38bca70-330c-4537-bf98-e86dfc74ca3c · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.125949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.548855Z digest=sha256:a4f3127ef4184caab06c5cccddeb73adfad7bfb162682f798627b0e3b021cacb

Observation 7e50cb86-3a88-40d2-bf01-b7e465fa5f7b · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.116421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.630453Z digest=sha256:29ea05696ece758f41270f4efaca0b454f9d753afb1217ef9839aa9d70de6fdc

Observation 661b3ec9-63de-46cb-bba7-6d56bd1afe86 · outbound

This paper cites D.; and Ammanabrolu, P.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models D.; and Ammanabrolu, P

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:03:11.107598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.698730Z digest=sha256:5097b3ef09b4c34b42c912df03a3c236f91a9c1ced54f29d6892180ea0f81582

Observation 019de136-a5fb-49e7-b721-ea8225ea87f3 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.098134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.782415Z digest=sha256:032f8c127ad89064deb505a3abec48f689023691409ecf15540449f0ab3f6270

Observation 67e0bbaa-98a7-4d03-be3f-2c61303d8811 · outbound

This paper cites S.; Bahaj, S.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models S.; Bahaj, S

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:03:11.090252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:05.902707Z digest=sha256:ad75c73fd3f4c46cf6442358642f1c8b01b9ed5d4ddf05ee7dcc6a71af2ed5dc

Observation 67ae9123-8417-4077-898e-4bc46c7a5b94 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.024471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.024471Z digest=sha256:ccea74d20410034bf7f42d57e9b9cc74f23411ca15d38d3ccf131756ae8535a0

Observation a0f2249e-6e3a-4c6a-9816-39f714cb71d7 · outbound

This paper cites Qwen2.5-VL Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Qwen2.5-VL Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.129793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.129793Z digest=sha256:93b9d2040bc70b4bed0911f19d71216d2d4375d1d415dfd38cad732939702873

Observation e144e3fe-8be4-40c8-a66c-c85a2f0d0488 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.080301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.189698Z digest=sha256:792688d409bbcd574279a782b322107a4d8ba729363b215d889616aba2c3a00c

Observation 7256ba6d-cb3d-432f-b1ec-11ae92db001e · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.069911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.274113Z digest=sha256:ca76ee1d4b6ce702f4d65f347fde31c910347fad68f5ae6f605055b484b26397

Observation a97923c0-6d6c-47cb-a13c-450e89675f53 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.054461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.405881Z digest=sha256:6ad788db18aeb592bd3d74dc44a2933ed7571db60bb302cd4a32eba6b5d2e5bc

Observation 2f187234-bd6e-4fc0-917d-39b20e354fcd · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.044836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.513232Z digest=sha256:c7733258bb9211cac59b672bd4a877d6ed55701daf10e2d38596c7fd7a684a81

Observation 510143d2-1596-4e11-a7df-93bb2fa1d1c5 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.033015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:06.534871Z digest=sha256:adb00a1ba271be6ce03dd7042299a288f4e760ac871ba95f922e31b259646a76

Observation 52ab568b-7125-4c5f-8753-0f7549bbe07d · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.609642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.609642Z digest=sha256:2d81803868ca79057c921b9979b191ad0ea6c3297936521297e0b32ef95684f6

Observation 7463ee3b-7a8b-4b4c-8b17-39c278af18b2 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.826107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.826107Z digest=sha256:9fe599bc93de453d06589144ac1167c78e440b4dc91d795e24ebb40d03b2cf14

Observation f661969b-fbb4-4be4-87a8-f55ab6236a67 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Training Verifiers to Solve Math Word Problems

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:06.901216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:06.901216Z digest=sha256:632b3a2f179aac6c4373ed1a1b7f93822ab0e010fb4dcd9385219d2e0e798bd3

Observation 4bf6a102-a30a-416f-b249-14055def1dc4 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.061721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.061721Z digest=sha256:9b57d4897117ab1501423fb42e7dcf7fb17d3eecc6d4ea11bca8dc1681eddb5e

Observation bb1a9588-635b-4ae0-9df9-a79c863ab9a3 · outbound

This paper cites OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.169525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.169525Z digest=sha256:9732b1e2b6472e3f901327898e574e6d9709a362ea40c4a324ed6502dd20e6fc

Observation e26bc2f0-7d3a-407e-8a53-524173d99e44 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.244354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.244354Z digest=sha256:fe5e1a4f1ea0ea8e37e791527df9f2e1842f9cbd663a94919a529d057a8b8249

Observation cb3702ed-2f72-4b02-b978-4fae2d648b4f · outbound

This paper cites OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.345505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.345505Z digest=sha256:8986ce892d521264a609b68deaf1d824f7ba844b36f474e6fce44f664797bebb

Observation 0c7f0ac6-c693-4f66-b267-b8e515ff00ee · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.437202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.437202Z digest=sha256:60fb148a114496ef0be130e38c3d13b68bb6c625f360d2c9adf51e9d93f2a286

Observation e0211619-0a7c-4cb0-a0f0-f8ff398862f6 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.545378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.545378Z digest=sha256:bb3660e83d65f349f644e2f7e67cf7d2bce5dec214549c46bed1378072f3210b

Observation b1f05515-02b9-4a4c-9d86-6a0c6faa08c2 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Measuring Mathematical Problem Solving With the MATH Dataset

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.632051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.632051Z digest=sha256:67ba49dfa4dea35c0f23d3911e0e09d8e3e530d3cd050b1d350e30143d4ca320

Observation 899150c5-4673-40c3-b6a9-e19a38c1dc7a · outbound

This paper cites Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.734635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.734635Z digest=sha256:699063c8d6c01243a38916a9e66eccbd6815c3bfeaa12aaa5e6a9ec41fcf8456

Observation 429db48e-c95c-4a12-bdd3-388ffca76dd4 · outbound

This paper cites OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models OCR-Reasoning Benchmark: Unveiling the True Capabilities of MLLMs in Complex Text-Rich Image Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:07.854784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:07.854784Z digest=sha256:2abca100b496a7f6d753eb435f6d32a2c246a2a565c55d16852091ba7017cb69

Observation 0fdc2059-8eac-43f0-8bff-108a0c209620 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.023470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:07.949013Z digest=sha256:a45aac46e1d2eb8d9179974d24da8323ec09708752e63458f7899474156f5974

Observation f4afe626-f042-4eb8-948c-6478f66f7fef · outbound

This paper cites VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.012180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.012180Z digest=sha256:f5e10241682e2a90e08885a063118f767a9809a44cc269357719806ba736ed9b

Observation 5f44f766-ee04-4484-9033-28be5d49da4d · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.014370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:08.096139Z digest=sha256:3487f0f8646650c197620962693627cc8e93bee2f20ecd5c250c1559b49ec710

Observation e446c60f-acc8-4f68-9a7d-af9aa250ad29 · outbound

This paper cites SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.173192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.173192Z digest=sha256:9e194cbc60bf7902d1ad1bc4d7d0de24b5bf030a7f6706a7e31d9d0be23657aa

Observation 7458803d-b9af-4f7d-9139-0725b7e7c1df · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 30

Resolution
verified exact
raw_fallback, observed 2026-08-05T23:03:10.609040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:08.264837Z digest=sha256:894091cf500141a2bb93ce5dc10a0fce8aab12bda448a437948bd01be9d6163c

Observation 3e9f344b-1c98-4941-af2c-1bf31b044908 · outbound

This paper cites CMMaTH: A Chinese Multi-modal Math Skill Evaluation Benchmark for Foundation Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models CMMaTH: A Chinese Multi-modal Math Skill Evaluation Benchmark for Foundation Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.364386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.364386Z digest=sha256:76738eaa30f2eadd8277b6b9501f4c5623ce564e6ea2f82abeb2acd158580bce

Observation 65025616-6cc4-4899-8aaf-823dfe965e17 · outbound

This paper cites DeepSeek-V3 Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models DeepSeek-V3 Technical Report

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.435992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.435992Z digest=sha256:03f9f933ae6aeef1be7ad9ba2dfa618b76295ed23050401ee66d67ef79b75b67

Observation 8066f43e-ca67-4b67-8ca5-d290acc7a358 · outbound

This paper cites Focus Anywhere for Fine-grained Multi-page Document Understanding.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Focus Anywhere for Fine-grained Multi-page Document Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.515997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.515997Z digest=sha256:7ee301cf5e8aaec0e0743e9db29b6498dad1c2f8574bf5f6d3e76ac2d33a4c36

Observation 380ac013-5ac1-43fe-b9ce-d221e9b4ed3c · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:11.004621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:08.606167Z digest=sha256:c60677dbbeb3fba42d8e27f7e929d9ab36fb791e58212749cd7972340b05d350

Observation 5e9ad217-3b3c-4ffe-bc04-8c7437c425c5 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.992808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:08.696090Z digest=sha256:4917cb604fa8d66fa2048a8e544224c193f564a3755cf9db56224411d457ad19

Observation accb2ee4-d4c1-4e0a-8015-82579f104c99 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.779737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.779737Z digest=sha256:20a4c20b732c04149c72a939471ebba7e2c6430cca61485a00f4fdb3af50bbfc

Observation 45c6c99f-4dc9-45fe-87ad-1a2aef8d2741 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:08.881627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:08.881627Z digest=sha256:c9e8bbbf362ea1413e280d7de2b562d25b71def158bb150b7b2839786c283ea8

Observation e2016163-bd89-441e-a1da-aa3a6e5fb986 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:09.004368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:09.004368Z digest=sha256:4afca6ad630ab5812d2948a67a269192a8a9cd41473ef2c8c696f62156f0c298

Observation 0f940b45-b68a-4cf9-acee-4d2492a24848 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:09.168731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:09.168731Z digest=sha256:9819809ff65cabacff458374b7fce5304e800c730f5c51f6c6050c914e79138b

Observation c1d456e7-75c6-4dac-8b03-f18346f8f951 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.968289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.300212Z digest=sha256:6073cd2f6814cd35753266115469a65afd8bf18eeaf80a5f9ed6aae27ba25fbb

Observation 907d46af-5286-4823-b493-aba88da5abf9 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.960560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.418539Z digest=sha256:5856c18a0dcba371533aefd3694ed6260a6c142eaee387591fc24d154a673b63

Observation 999c36b2-a78d-47d1-b25d-8364c7a3cbe8 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.952770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.585949Z digest=sha256:25e762a3f579ea7d81f3c166eebd23f9345fcaf75ce138f5f3b6fd54b85e713c

Observation cbb68054-14f4-434d-bc7c-2dd937822bb7 · outbound

This paper cites Skywork-R1V3 Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Skywork-R1V3 Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:09.720243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:09.720243Z digest=sha256:c1787000c3e2d90fe7256696f9f861171fc8afd8f4477a9d2fde3557664d9cc8

Observation f226de3c-ad66-4d4e-854a-0c91a587c763 · outbound

This paper cites W.; Tay, Y.; Ruder, S.; Zhou, D.; et al.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models W.; Tay, Y.; Ruder, S.; Zhou, D.; et al

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:03:10.944633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.892367Z digest=sha256:e9cc3f3e3756fa0cb78aecc8465caaff786aec64e6b9989b18b10f8005a311d9

Observation 13955807-f937-4d3e-adea-6aa6c22de0b5 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.937022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:09.990297Z digest=sha256:57998d3fb764650ead72c2730216e22fd0ef4350bbbb8811c127ecb17af721f8

Observation 7d52edc9-2494-4423-99e0-35a83f14b4f1 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.928498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.101605Z digest=sha256:153ea5b47730466eb0c90558a56967cecb8e90866b545b4cfc9bdb92946b5315

Observation 71aee071-de03-4200-b863-2ceaaa125bea · outbound

This paper cites MiMo-VL Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models MiMo-VL Technical Report

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.109377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.109377Z digest=sha256:bd373e0953e7753b628ab6b3cd9a9133301423069725ddfe956e3f430a1d6cdb

Observation 2ea5c41e-9f20-4505-9e8e-01755079951f · outbound

This paper cites Gemma 3 Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Gemma 3 Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.112701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.112701Z digest=sha256:744b04c58ad05b5ea13525c050359b071ec89eaa68bc2a971ffe84347f94daa9

Observation 5e485052-805b-4905-8a77-1f30ab78ac14 · outbound

This paper cites Kimi-VL Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Kimi-VL Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.115570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.115570Z digest=sha256:0fdd4bb6bb05c42e6c30942c73d7b1425c3f394a7d454ed3034a53964037aab3

Observation 1e0d2528-1958-4e3f-b77f-1efc4f76eaf6 · outbound

This paper cites Kwai Keye-VL Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Kwai Keye-VL Technical Report

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.118939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.118939Z digest=sha256:ecca08869cc1b7f4ecf84974024647b4faa80421bd6a11938358af38cbf7db42

Observation 618dd905-d1d5-49ec-99d7-26c90ea45130 · outbound

This paper cites VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.122029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.122029Z digest=sha256:b15f56735b0307a8a61ea2bd930bfe36574cf140698e415bfdd31f1eea8a2cbb

Observation 260dc3eb-3791-42c0-8cfe-437c4713d715 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.920534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.125020Z digest=sha256:b46a92d2802dbc0e6e6d4bc74bb4aea8c301ae7a967ba39c47c3089c2d521637

Observation 3871fa7b-7502-439a-94b1-7fcaea51847c · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.911319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.128111Z digest=sha256:564c371327976d9c6ff8168193aee2a5b8879d668d926e8f0e7c38200fe10ff3

Observation 2a694aaf-8b1a-4551-b2ec-17d8ecf284fb · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.902563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.131052Z digest=sha256:be62c961edf233901ec4091628ddedf6452f61ab20abf904cb6a64dab5ad7131

Observation f055fe9b-4553-40eb-bc57-d58d410b375a · outbound

This paper cites SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.133767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.133767Z digest=sha256:07d175980b6d0df48e90856aabc5bf4c9b6679b025c88d8472fb4a38eff2ee09

Observation 41cbcac5-d37f-48cf-91ae-e0388f1b2d00 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.137008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.137008Z digest=sha256:1ab8445471ca8539fe052b670dc8c0e039de123c2800f80440ede092125bcee5

Observation 1bd96295-23af-4c3e-a5db-9dc5ee03bd66 · outbound

This paper cites Benchmarking Multimodal Mathematical Reasoning with Explicit Visual Dependency.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Benchmarking Multimodal Mathematical Reasoning with Explicit Visual Dependency

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-05T23:03:10.382590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.139666Z digest=sha256:02e3358b22e7b108b40bb8d5a3ad2c4edf992da0f57572978ba4731371b137e6

Observation 4fb2e29a-2ccb-4639-9c21-9808a76826c5 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.142900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.142900Z digest=sha256:f10f0e0136510858462b67d74a8e14f99bdfe16d74976d869fcb0c78b69fe68c

Observation 9ce07fde-0d99-462f-9dcc-29e2e4c88fe9 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.893404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.146186Z digest=sha256:0c4aeb78c3f9ff5b821a59e045430c7090c1abc6871bf510d61ff243c5478aa4

Observation a7d2f13b-9698-4f37-8ec8-bb2072bd816f · outbound

This paper cites LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.149147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.149147Z digest=sha256:3ed8c783b9023e6b3b14df823a1e58c836b52d614dc9f66143dc6151a7da5792

Observation d42b273b-e04a-4d0d-b3fa-ccdea184d038 · outbound

This paper cites SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.152722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.152722Z digest=sha256:0afb9d81f73fb75f8c5046385142f5680f10a78d1649ea973733ccac1dbbea8a

Observation ac6f4c5d-f2f3-4688-ac28-1377088bed02 · outbound

This paper cites Qwen3 Technical Report.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Qwen3 Technical Report

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.155851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.155851Z digest=sha256:752b89d30118765d6c220ccf224b9a1cb63cd4d1188d420439984e05f8764b65

Observation 81f9e82d-4567-4976-9de0-b0ba9a4a3bb7 · outbound

This paper cites WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.158919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.158919Z digest=sha256:8d0e98e07545f1eefd22fde23ad718f7d9c9ed76d7265fa9bb6c6ad67449b868

Observation 0b2729ac-0c6d-42a0-a2a7-401f20bcb0e0 · outbound

This paper cites CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.162023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.162023Z digest=sha256:c91aa2ac3cb593a50aa92b4882f01666f9933461420d8a9cb4aef1419cf284e0

Observation ab2b6813-c8b3-4b9b-b62b-dc614589eade · outbound

This paper cites Benchmarking Reasoning Robustness in Large Language Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Benchmarking Reasoning Robustness in Large Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.164969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.164969Z digest=sha256:54ba0a5d7fc439fc4c3d602fca0d89e291935a90e616082fa6ada2d8887c259e

Observation 91144bb5-3f36-47df-adb1-01aa8ccf59ce · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.885047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.167996Z digest=sha256:66343d223029adab4bb52ea4f1f9f087dbac5099d0ed966c77b8bf5e92eef388

Observation 3547a3a2-b23f-4526-acf2-6c9d0fb60d39 · outbound

This paper cites an unresolved cited work.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-05T23:03:10.875416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-05T23:03:10.170967Z digest=sha256:f56173e8c1f858dfcb5c3e8a4636fe1f977c1493f0c956c7e771e6c8a6b87baf

Observation 1cf70cc8-b296-4921-858d-26877ae6012a · outbound

This paper cites Multimodal Table Understanding.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models Multimodal Table Understanding

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.173220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.173220Z digest=sha256:a04125f10fbeb7cd695f39d25fadea55b71a2ee3090374d274ea31718424d1e3

Observation c7cda48d-4955-4a87-b877-5afcf3d4c5f3 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.175885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.175885Z digest=sha256:352454108848d69330d3da8fddbf70bb443b162de5aebd9adbd7304625ed8ed8

Observation 803706d1-06d3-4b48-8f87-64fdea209998 · outbound

This paper cites DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models.

MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:10.178730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:03:10.178730Z digest=sha256:0f8ebcd65500142ecb551c6fbdae62eabf9fb89ea37c47904d35ae3622f43e66

Pith citing papers

No inbound Pith citation observations are available.