Pith. sign in

Paper Citation Record · LEDGER

MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2404.09486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.09486 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:55:59.096567Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T04:32:33.203813Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5664a034-6115-451b-ab8f-f9d17a57bcc0 · inbound

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing cites this paper.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.361657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.361657Z digest=sha256:086d07887fe337e128fc2458ed933c54ae047a663444323e3272c8c71fa14d25

Observation 1d070049-fa1a-48e2-a934-e1df4690b0d6 · inbound

ScratchEval: Are GPT-4o Smarter than My Child? Evaluating Large Multimodal Models with Visual Programming Challenges cites this paper.

ScratchEval: Are GPT-4o Smarter than My Child? Evaluating Large Multimodal Models with Visual Programming Challenges MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T10:46:57.710361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:46:57.710361Z digest=sha256:a2ac6b3e0810048c1bd685f853813954764c4b73da0068837ac7e627a8cacfc1

Observation 1d77b88b-84ad-4574-97f9-5c4fd4d58148 · inbound

An Exploratory Study of ML Sketches and Visual Code Assistants cites this paper.

An Exploratory Study of ML Sketches and Visual Code Assistants MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T13:16:19.365083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:16:19.365083Z digest=sha256:22f51797c3d262b644d5cf47ca7a5e81ca4d7de7c30fa0d25d70f049d00312b5

Observation c46d5a91-088d-4f2a-b0a3-079f476777ea · inbound

Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark cites this paper.

Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:18:59.607732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:18:59.607732Z digest=sha256:1778df4e426d5e4b19428c4e8db599770f63109ebeec2f41ed06ed024849ae36

Observation 8d10f1bd-dfb7-48b5-a7ad-057fc629b2f2 · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:33.207510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:432a8404a1f97e07b3b57222dfbe7fc7b693683e73d1b8948368e4461bb82f99

Observation 824db424-6a76-4b01-a105-2f0a24cb35e2 · inbound

AutoP2C: An LLM-Based Agent Framework for Code Repository Generation from Multimodal Content in Academic Papers cites this paper.

AutoP2C: An LLM-Based Agent Framework for Code Repository Generation from Multimodal Content in Academic Papers MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T05:55:59.096567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:55:59.096567Z digest=sha256:999485622d1102f7f2aeb800e2ae9499edf87e9175108bd3f97101e65b9df496

Observation 24f9be77-7c62-422b-953f-dc208a8fff6e · inbound

Knowledge Augmented Complex Problem Solving with Large Language Models: A Survey cites this paper.

Knowledge Augmented Complex Problem Solving with Large Language Models: A Survey MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T23:56:02.083139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:56:02.083139Z digest=sha256:e055fff5116a9d248926632dad7a9d18450b9a5575570e5b288cffd6e214c19a

Observation 573dda8e-ab3b-45d6-b919-5468638ee61b · inbound

VS-Bench: Evaluating VLMs for Strategic Abilities in Multi-Agent Environments cites this paper.

VS-Bench: Evaluating VLMs for Strategic Abilities in Multi-Agent Environments MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:57:16.244535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T11:57:08.314088Z digest=sha256:ffec2c468b900e53b4aeff1baef1051201b5143a76234ab73dec2adff5e6d925

Observation 61f359e5-b834-465e-a679-61d5bec19075 · inbound

SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design cites this paper.

SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:26:15.088404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:26:15.088404Z digest=sha256:717b5101f06e74ad10a64a7d2fca895055cbe0cf4b19e863d72b978446bb637f

Observation f3450736-3817-48e6-9ca6-50b2f2fcdf67 · inbound

Multilingual Multimodal Software Developer for Code Generation cites this paper.

Multilingual Multimodal Software Developer for Code Generation MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:18:53.801064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:18:53.801064Z digest=sha256:99115a65c3512e76778a3bd3e7838bb16d3d786d23fa8a983d427480c514028d

Observation 64eefe0e-4b8c-49fe-b5a0-279ab978312f · inbound

FairReason: Balancing Reasoning and Social Bias in MLLMs cites this paper.

FairReason: Balancing Reasoning and Social Bias in MLLMs MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T11:09:04.748471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:09:04.748471Z digest=sha256:c69d1d64880a4de045664e9f346b5b47901a1cd7405a413efd208976784ff129

Observation 119c5331-5187-448b-b7d6-c6acb141cb7f · inbound

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models cites this paper.

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:37.655941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:37.655941Z digest=sha256:7ec61ef84047d7df6af1314a5f36fedd0227120c3021d32b35a7b4d3d0e22b88

Observation 72a4aa70-42da-4596-8e02-576f46a02e9d · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:09.695133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:09.695133Z digest=sha256:f0f0bc08f24ce4e7be3d243e628d1e09ae656757e648ecb1fd20cfd634cfb1ab

Observation e70609b5-6e4b-4f82-8c92-4f85e7a6378f · inbound

SVRepair: Structured Visual Reasoning for Automated Program Repair cites this paper.

SVRepair: Structured Visual Reasoning for Automated Program Repair MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T06:11:24.737429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T06:11:24.737429Z digest=sha256:c1b0b512936845d60c32ad2f9b51e92ef58c4562f81b716ad3e525c721f90cdb

Observation 6054ffdd-f7e2-4d57-9dc4-a73ca2ba512b · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:34.456470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T20:12:46.385646Z digest=sha256:b7338318cefa4a827a2ad26bcde98b898905d1088033dbcb102e82146c8b4a8e

Observation 4a46117c-ad85-4711-bc31-ea1d734d265d · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:31:25.174665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T10:30:06.829915Z digest=sha256:58b9b00f79dae9467b38a15fb0073bcf1e8229451be3285be40632321f416ea9

Observation d477d8c0-a048-43cf-9752-c1307870ddc2 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T22:00:06.512766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:00:06.512766Z digest=sha256:e600c6f17aa7470716ff03a2b4a27a234f72ea016bc437e2571524d1568f89ea