Pith. sign in

Paper Citation Record · LEDGER

VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2504.10342.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.10342 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:00:59.247883Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:09:43.184299Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3e1fca7e-9e72-4ef5-bbf0-8dbbbf594ba0 · inbound

Human-Aligned Bench: Fine-Grained Assessment of Reasoning Ability in MLLMs vs. Humans cites this paper.

Human-Aligned Bench: Fine-Grained Assessment of Reasoning Ability in MLLMs vs. Humans VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:00:59.247883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:00:59.247883Z digest=sha256:9201c1e1eabb35a9c8a74ca7a2205398bd9bd3fb3eb2c94995a01a3afc6244dc

Observation e64bb1aa-a903-4e8b-87cd-0c766d05cfd5 · inbound

Evaluating the Logical Reasoning Abilities of Large Reasoning Models cites this paper.

Evaluating the Logical Reasoning Abilities of Large Reasoning Models VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-15T20:52:26.710386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:52:26.710386Z digest=sha256:456c98343ad412a96edb159f128ac2f203216a6553c7150673cfd3b349059817

Observation dc85e25c-5131-46e0-a594-a8fb0bacfe77 · inbound

Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models cites this paper.

Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:50:32.927998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:50:32.927998Z digest=sha256:26f7627ca59a7315032f985ce93775d5c523e8ee5f69042a9e9449af89ceb5f8

Observation e8aab8db-8a7b-4c70-acd6-5180a716456b · inbound

MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs cites this paper.

MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:59.800576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:35:59.800576Z digest=sha256:a4b7413f3dd0cac1c61f775b2d73782002374526768a7a93855f26be95c69218

Observation 0299d2e3-07dc-4d17-91fe-4aac498c87bb · inbound

VisualSphinx: Large-Scale Synthetic Vision Logic Puzzles for RL cites this paper.

VisualSphinx: Large-Scale Synthetic Vision Logic Puzzles for RL VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:18.894158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:18.894158Z digest=sha256:4aa2c89ffe6a504a04cde8a89e343786f766d198e60bdf4044e06692b4c3bd2b

Observation 1bd79913-9728-404c-a379-a26deb8e6c40 · inbound

PyVision: Agentic Vision with Dynamic Tooling cites this paper.

PyVision: Agentic Vision with Dynamic Tooling VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:31:31.285831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:31:31.285831Z digest=sha256:493763c8abe47a5af0604351c24e4e4fdf36a3899ab8c91c811ee28fdf5f7e70

Observation 157d773d-61dd-4579-a0a3-75a727390367 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:34.725580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T20:12:46.385646Z digest=sha256:61f3e2ec7d0b54046e51ed97ee6e25a91a42fb97cd8389e457fe4ba06d57ca72

Observation 14012aca-e9b3-4b09-ba82-9391edd17e32 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:31:25.285893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T10:30:06.829915Z digest=sha256:bc2d7b6d15d2f24dbd08d3cc088f434d60dcd8003407b9dec036a9ec14ae395b

Observation 36dc0cea-f4ce-4c15-8b78-2988d4cbaf90 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T22:00:09.028700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:00:09.028700Z digest=sha256:2097d55a07a9ebb0b4967c4152b2706b3c7262bca94b1a0b9b143a1157a478af

Observation 28dea71a-aa68-40be-9c65-92537819bd3b · inbound

Limits of Spatial Imagery Reasoning in Frontier LLM Models cites this paper.

Limits of Spatial Imagery Reasoning in Frontier LLM Models VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-13T19:19:52.557913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T19:19:52.557913Z digest=sha256:61f318bc7e81bc1e2e2d13bcdd57085b75df199d184eff5d71e1906b4f784968

Observation 8c9eab83-c05a-49e5-9dea-746e01f64a12 · inbound

SALLIE: Generation-Free Hidden-State Detection of Jailbreaks and Prompt Injections Across Text and Vision cites this paper.

SALLIE: Generation-Free Hidden-State Detection of Jailbreaks and Prompt Injections Across Text and Vision VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:30:51.295449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T19:04:46.426969Z digest=sha256:dcc6542f3d8540ba13b4294cb36056e2ed5b9e330fb2d55bbeda9b811a6d827b

Observation 32710ea2-6bc0-4cea-a032-ea0f33c017a4 · inbound

TraversalBench: Challenging Paths to Follow for Vision Language Models cites this paper.

TraversalBench: Challenging Paths to Follow for Vision Language Models VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:36:02.685510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T15:25:38.215013Z digest=sha256:e1b7dd222c9c50ebd2a24b5892097709336d05633992364bac72bff2f2905105

Observation 7a59ac05-a125-454e-ae97-d8497ee5de87 · inbound

Reinforcing Multimodal Reasoning Against Visual Degradation cites this paper.

Reinforcing Multimodal Reasoning Against Visual Degradation VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:06:25.204696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T04:37:21.451146Z digest=sha256:f89627897d58a088646096518b4ebd1eb45d13bcf7a4821c60ef484cc3035e07

Observation 17dca504-fc9a-489b-bffe-4dcc82decf55 · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:17:29.199028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T07:14:48.918959Z digest=sha256:1e615c0f0fac0f22bfdac23f3ede3aa4d1dd8d24553c2dbb7a4f94de12ca93f9

Observation 048ff051-86d8-4ffa-9f85-f94b1affb024 · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:48:00.572342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-14T21:47:50.595481Z digest=sha256:6312817c8eab0338ba8a35ce76a04fe0d48c61d22d6edd234435a5b7c7d6ff4b

Observation 4cd7d506-de3d-4848-95b4-e600246450b2 · inbound

Semantic-Enriched Latent Visual Reasoning cites this paper.

Semantic-Enriched Latent Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:43:05.875829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T06:40:55.537488Z digest=sha256:f7ddce1121aa0b9b68d874f1385970de8a5daa1705321e448b2c29a091604eb7

Observation db4ee2f6-c4f7-4e38-97bb-966aba36fff3 · inbound

Semantic-Enriched Latent Visual Reasoning cites this paper.

Semantic-Enriched Latent Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:55:00.779176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T18:48:04.370230Z digest=sha256:7e7babf4fd61d0a5c169d470bb51ef27294ea7916712442e64889e637a7bd210

Observation 6b9996d1-738b-4331-a89d-f3355faf8ac1 · inbound

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning cites this paper.

StemBind: When MLLMs Get Lost Between Rules and Instances in Abstract Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:12:50.317499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T23:15:56.598968Z digest=sha256:a11357cb0e57b3d9d66ae352d92b3d0a3186cc5eafaa3bce9d3ff4f9901118aa

Observation 0d4afd7a-2761-4209-b41e-ba36b85a72e4 · inbound

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients cites this paper.

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 141

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:56.116501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T01:08:52.981296Z digest=sha256:53f825350d2a2f4045aa1c423a674303c1edaa6ee5851a586c46bba89864301b

Observation 91627fbf-52d5-4376-b522-4392b0c57f4a · inbound

Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do cites this paper.

Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:09:43.185822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T10:27:43.396559Z digest=sha256:058cba53817469534c6a358a9fe3f9fec0a1dbdc67713307147d0b58fa7c2ef5

Observation 843e4628-03f0-49c5-80f1-819965140d95 · inbound

PACE: A Proxy for Agentic Capability Evaluation cites this paper.

PACE: A Proxy for Agentic Capability Evaluation VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T13:38:18.468017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-07-03T13:34:39.350893Z digest=sha256:3ef49558d93737b4e40f851daf4aacaeb53cb6030d05e26eee13d123e1a1faa4

Observation ef855c3d-5503-47e4-8adb-f9d5f82bc823 · inbound

PACE: A Proxy for Agentic Capability Evaluation cites this paper.

PACE: A Proxy for Agentic Capability Evaluation VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-12T08:29:58.561496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T08:29:58.561496Z digest=sha256:f1ceaf48b5637de489c3d7f97db540599716f7f1acb00cc343430f8ea19cead2

Observation ca6f6a0e-2630-43d9-9ccf-325d95d6ef81 · inbound

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning cites this paper.

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T11:47:26.293957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:47:26.293957Z digest=sha256:ec9bde3ca7cd6f689d4a88ac6f5b286b43d9876742e66e95ca229e5b7fa19359

Observation 6b89ba06-8336-4c64-8b10-f79a53b39f0d · inbound

Beacon: Knowing When and How to Perform Agentic Visual Reasoning cites this paper.

Beacon: Knowing When and How to Perform Agentic Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T02:45:28.647722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:45:28.647722Z digest=sha256:51d6510ff13efc0e66fca7257c673d8519e2f6dd5e845234bc635575318b4dad

Observation 911b0bbe-095d-4888-802d-e7e9880745fa · inbound

LUT: Latent Utility Training for Visual Reasoning cites this paper.

LUT: Latent Utility Training for Visual Reasoning VisualPuzzles: Decoupling Multimodal Reasoning Evaluation from Domain Knowledge

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T00:28:16.463177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:28:16.463177Z digest=sha256:1811e8b089d1a0681289b2fbe42ac56b31567bc783828f79c0804e9e7bde0337