Pith. sign in

Paper Citation Record · LEDGER

HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2410.12381.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.12381 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:55:05.985112Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:56:55.430688Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 781a1daf-41e5-4c0d-a0d0-9b7e712c6709 · inbound

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models cites this paper.

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T20:55:05.985112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:55:05.985112Z digest=sha256:b216c786d2531f3b856b251755fbeb17355343e3c19c5f6f6bb3d1b99e26917c

Observation 2e2a2458-608b-4513-a8aa-b9be07b56ce9 · inbound

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization cites this paper.

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-16T00:19:20.545535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T00:19:20.462455Z digest=sha256:48b1005f558b4e99b3b1231436ddaf4c7ba34cb3701dfd8e9d36c1c20920b369

Observation edb1385a-aa8f-4b6a-a67f-e6175f0dace0 · inbound

MMR-V: What's Left Unsaid? A Benchmark for Multimodal Deep Reasoning in Videos cites this paper.

MMR-V: What's Left Unsaid? A Benchmark for Multimodal Deep Reasoning in Videos HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:35.576700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:52:35.576700Z digest=sha256:a120457295c878b9e6d6dad51f4e5bed2f617a032fd9a7575ebca914adf7f22a

Observation 662d4919-af9f-4c31-a24b-d5b8e1316c15 · inbound

SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design cites this paper.

SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:26:15.141311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:26:15.141311Z digest=sha256:85bc52d3a643bafcda1367cb6d709af5cc312f4e82c9589f51ca92225566f46d

Observation a29e0267-9a48-4c26-9cce-465bded6ffc4 · inbound

Multilingual Multimodal Software Developer for Code Generation cites this paper.

Multilingual Multimodal Software Developer for Code Generation HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:58.380932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:19:58.380932Z digest=sha256:50b0997e4e243cfa6051ab8a1e06d44afab4578fd72fee89760932241cf79dca

Observation 6e2acc19-30d8-4d7d-a8d5-7394338b1587 · inbound

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models cites this paper.

VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding Models HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:39.584064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:39.584064Z digest=sha256:3e56c8cf741a9ee43c28041bb7a24d1e6f164464bccb528191a4ccb63f8e090d

Observation e14d6276-bb43-4220-bc27-33d47dda05a8 · inbound

Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks cites this paper.

Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T11:56:10.967394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:56:10.967394Z digest=sha256:71e7620e5bb472e25cc09350ab8012065b17b04711acca5b33ccacbef801a4cd

Observation d6088590-dd16-4f8c-ae1a-f242e0762ab0 · inbound

VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs cites this paper.

VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-03T20:23:09.675588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:23:09.675588Z digest=sha256:bd8dba4e7d3e7657fa82fdb2e737964f35b488026a3ab217f9be989cd9b293ab

Observation 967a24f9-866d-458e-a97c-5cb9a03ecbd1 · inbound

ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation cites this paper.

ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T09:45:31.514666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:45:31.514666Z digest=sha256:743af333caacea5bbc56cfa59f4be94f8b0d0cf007454faa924ae0f7c72e1e4a

Observation a154fba6-4f9c-4295-a26d-e4c9b0a721cd · inbound

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles cites this paper.

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:51:15.396365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T07:51:13.362986Z digest=sha256:39b1ecf0e053fbf9c4182aa03d5633bc3e969f0ea0dd4f6abb62e104c660ae15

Observation 85e648c6-9d82-47f6-90e6-c274ed0a3d76 · inbound

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction cites this paper.

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:56:55.432310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T02:46:50.373450Z digest=sha256:d9e68202ad2ffbb91d12dfa4ec6abbff2845c59b8bc7ffc6fdb805000d5169c5