Pith. sign in

Paper Citation Record · LEDGER

PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 33 inbound Pith citation observations for arXiv:2501.16411.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.16411 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 33 of 33 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:14:08.671259Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T18:06:25.898385Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d5e2afbd-c8a4-42e7-98b2-3df09b155d46 · inbound

Abstract 3D Perception for Spatial Intelligence in Vision-Language Models cites this paper.

Abstract 3D Perception for Spatial Intelligence in Vision-Language Models PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:45:24.431565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T22:43:16.761970Z digest=sha256:e95275df04a78b646903f26f13ec0d1b0343cbff7372dbd50a81f0f2f1f7aa85

Observation 78f07fa2-0a71-46e7-ab1d-4b349425d388 · inbound

SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios cites this paper.

SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T21:14:08.671259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:14:08.671259Z digest=sha256:ca87a70fc58480b2b2b25b97a5f486e31aff10bf8ba1248f092332a4d04a6cf1

Observation 466b7f2c-449d-402f-b6c3-6a4b6c17a26f · inbound

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models cites this paper.

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:59:08.633140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T05:55:11.495430Z digest=sha256:0b7bf20faa1adeb5927e9b60b70a374905a80760a276c10289c6b9943e7a8deb

Observation 88b3674e-346b-4206-a479-2edee83aa4e4 · inbound

CoSPlan: Corrective Sequential Planning via Scene Graph Incremental Updates cites this paper.

CoSPlan: Corrective Sequential Planning via Scene Graph Incremental Updates PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T17:17:05.467145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:17:05.467145Z digest=sha256:9328945c725fa1314178a2b71a955c0ddcab419ca45de9a8f8c63dca0411b3eb

Observation b379fa1c-68a0-47ce-829c-38684ac4ae24 · inbound

V-REX: Benchmarking Exploratory Visual Reasoning via Chain-of-Questions cites this paper.

V-REX: Benchmarking Exploratory Visual Reasoning via Chain-of-Questions PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T16:47:24.462509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:47:24.462509Z digest=sha256:3220d6c0e02d9db03176c5f6bf56d3bb5ec368d25cc4fee1eb8d24b2e4b1a498

Observation 4ad7f2f2-4cb9-40be-a877-71b08816f498 · inbound

Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control cites this paper.

Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:00:24.008404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T16:57:19.490074Z digest=sha256:6c5dec9797221b79363d00a3dc79d68d05f0051eaa0ba1faab431c4c05b58464

Observation c8afffaa-b52f-4a36-8ee1-4d4446130432 · inbound

SCP: Spatial Causal Prediction in Video cites this paper.

SCP: Spatial Causal Prediction in Video PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:50:11.227229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T16:47:44.523606Z digest=sha256:7b760dab5b5024e383a7176de869ec8326315fb25a1499e25678205d71f85ee2

Observation 253d7d01-5785-4c20-8e07-4864a5d3dbbe · inbound

Multimodal Language Models Cannot Spot Spatial Inconsistencies cites this paper.

Multimodal Language Models Cannot Spot Spatial Inconsistencies PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:13:24.768838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T23:12:27.333405Z digest=sha256:de5be6c42e6ec89d2510553ef2a26db98aca6a286cd766f7736cde85459d8471

Observation 4b761ef7-61fa-437e-8043-744e943c169f · inbound

Dual-Cluster Memory Agent: Resolving Multi-Paradigm Ambiguity in Optimization Problem Solving cites this paper.

Dual-Cluster Memory Agent: Resolving Multi-Paradigm Ambiguity in Optimization Problem Solving PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:46:04.347333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T00:37:55.147350Z digest=sha256:3621ca7dec36d8b1d278af227427135d67a9c67ba2666903b561d961f5d08867

Observation c6304c87-1c74-40c7-9399-168e0d4138f5 · inbound

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving cites this paper.

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 96

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:21:03.894039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-09T22:04:19.654714Z digest=sha256:078f386bf19e06ae64761c9d968e5cb67b0965bc14016803c9f719507eec1cbf

Observation 935b2493-3073-40e1-bb43-26b7cdfca23a · inbound

Grounding Video Reasoning in Physical Signals cites this paper.

Grounding Video Reasoning in Physical Signals PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:34:07.272291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T22:29:21.240010Z digest=sha256:fcc4e9f6f43ca5c91529c886f91e4f981c2d265708c07ab352504ee0d256e898

Observation f65b4cce-0ad5-4843-b705-438b4bd60786 · inbound

From Priors to Perception: Grounding Video-LLMs in Physical Reality cites this paper.

From Priors to Perception: Grounding Video-LLMs in Physical Reality PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:21:08.313979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T17:41:23.233366Z digest=sha256:a75e724a86ecca1ebb391f85baf63ffab5e74b0bc998996db460e00e2f93f7f5

Observation 4ef3b87a-c965-4d11-b366-5270fa6a8c92 · inbound

Tracing the Arrow of Time: Diagnosing Temporal Information Flow in Video-LLMs cites this paper.

Tracing the Arrow of Time: Diagnosing Temporal Information Flow in Video-LLMs PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:05:53.207696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T03:04:27.841522Z digest=sha256:eeb085eb243add4d1b3cdb42d676f13f630280a3bf0c769db0cd534f05b778e6

Observation 2d14ffac-878a-4c03-aa2e-43ef42ff1033 · inbound

Quantitative Video World Model Evaluation for Geometric-Consistency cites this paper.

Quantitative Video World Model Evaluation for Geometric-Consistency PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:14:53.082576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T03:11:03.060052Z digest=sha256:f8c131b6c0bb05b25652dc25ad08e6e611cbe64b9c447611d285c3c5f699d6da

Observation 6ab7aa76-0d97-4a16-b022-74df7f7b3ab3 · inbound

PhysBrain 1.0 Technical Report cites this paper.

PhysBrain 1.0 Technical Report PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:37:39.921164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T16:34:44.204055Z digest=sha256:026a5a15520cf33443f68f741d8651703f9c4662aa8c50d077d2ddc503c101b8

Observation a84fa5a1-28c9-4685-9564-e6224bcd8179 · inbound

Evidence of a Cognitive Shift in AI Education: How Students Are Rethinking Human Intelligence? cites this paper.

Evidence of a Cognitive Shift in AI Education: How Students Are Rethinking Human Intelligence? PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T01:33:56.534774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-21T01:31:17.669186Z digest=sha256:d2d00442cecb427d6c05df75e3d7ac499778158d1f402d8ccea281063d00007d

Observation 150f97be-9eaa-4875-9b79-17369e18171e · inbound

GeoWorld-VLM: Geometry from World Models for Vision-Language Models cites this paper.

GeoWorld-VLM: Geometry from World Models for Vision-Language Models PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T17:58:49.557468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T17:57:00.909897Z digest=sha256:ccae97a1355d96a608033d52c6af60d3cd80455ab63af411e074d9f6def8a063

Observation 637f56f8-b75c-46b0-8cd7-9eef946ff3c5 · inbound

GeoWorld-VLM: Geometry from World Models for Vision-Language Models cites this paper.

GeoWorld-VLM: Geometry from World Models for Vision-Language Models PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:05:00.630085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T19:02:05.125937Z digest=sha256:cdf115483bd7a31cd169ec8fb066cb6be0dc2628e72a0bb7cae3b669e488833d

Observation 6fdda63d-6d25-48c9-a74e-4b9c0c25729a · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:53:13.228601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T10:52:22.778489Z digest=sha256:89610ff712700f9ac6818f51991dd02ebb20a7c99b42ef188ac1ab8aaab74342

Observation 20dda940-1c8b-48fd-9f8f-0d7591ca878a · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T14:55:48.810709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T18:25:17.831116Z digest=sha256:895324ef427f40503e2810adf721833285d9cf5eeb1904f1e6a67f0afe687462

Observation e87495d2-7025-455f-9942-efeb2ac24e45 · inbound

$\Delta$ynamics: Language-Based Representation for Inferring Rigid-Body Dynamics From Videos cites this paper.

$\Delta$ynamics: Language-Based Representation for Inferring Rigid-Body Dynamics From Videos PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:14:40.989000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T06:14:26.824489Z digest=sha256:6f48c661734d2d71256302cf586727ccfd851bd3c1ef8fc597dede80e3f44fe0

Observation 5ead1bbb-04be-4290-af55-1656208acec8 · inbound

World Models in Words: Auditing Physical State-Transition Commitments in Vision-Language Models cites this paper.

World Models in Words: Auditing Physical State-Transition Commitments in Vision-Language Models PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:23:15.377444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T08:18:45.804447Z digest=sha256:8c67b5943204a5b530bb354b25e73b8fa130399b160141bd4f99ee3959e69acd

Observation 6b6d6710-75cd-4995-8649-3336d863e737 · inbound

Benchmarking Single-Factor Physical Video-to-Audio Generation cites this paper.

Benchmarking Single-Factor Physical Video-to-Audio Generation PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.382683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T07:41:56.917119Z digest=sha256:f3b0a40adbdbf253f7e4637082fcf3d18d19432d715544f92676723209c21fd3

Observation dbf71278-e1e7-40ad-8abb-55308a286db9 · inbound

Physically Viable World Models: A Case for Query-Conditioned Embodied AI cites this paper.

Physically Viable World Models: A Case for Query-Conditioned Embodied AI PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T09:13:16.514166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T06:55:57.801162Z digest=sha256:8d735325eb4a42eab328145e5f6e72152f439b27ff9809b063817f998b82336e

Observation 99ed44cd-9b49-4b80-b54e-42c1ffa78105 · inbound

Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs cites this paper.

Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:57:06.745261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T23:15:38.962013Z digest=sha256:7ab6b4e66751dc21582fc80dacd4640ed0a72f804538cc606a3764545bf2ef0a

Observation 03672721-5de8-42b2-84a0-91793d035cc0 · inbound

Physics in 2-Steps: Locking Motion Priors Before Visual Refinement Erases Them cites this paper.

Physics in 2-Steps: Locking Motion Priors Before Visual Refinement Erases Them PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:36:57.143795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T01:58:00.336605Z digest=sha256:215de50508ebb9aab5860db1efefca81ecd707bead92734936320278edf94bbd

Observation d8f0afaf-4af6-4b32-b691-5135409340bd · inbound

ChronoPhyBench: Do MLLMs Truly Understand the World or Merely Exploit Language Priors? cites this paper.

ChronoPhyBench: Do MLLMs Truly Understand the World or Merely Exploit Language Priors? PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:27:22.649811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T20:23:18.667677Z digest=sha256:e9290caa5be454b2eb09a806e0e77b1fe79993346a63c4986d3399661fced8f0

Observation 75b7dd1d-4257-4bb6-9d38-070252d0be41 · inbound

RigPI: Dynamic Parameter Identification of Rigid Body via VLM-Seeded Differentiable Simulation cites this paper.

RigPI: Dynamic Parameter Identification of Rigid Body via VLM-Seeded Differentiable Simulation PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:39:59.976044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-25T23:35:13.949488Z digest=sha256:ee8a2566a3722b92e35a5bb20ad10a11a2edd29971bfc5b1a374f736ee545cad

Observation 8fac81bd-607e-4c0c-a713-b8d29f4845d8 · inbound

RigPI: Dynamic Parameter Identification of Rigid Body via VLM-Seeded Differentiable Simulation cites this paper.

RigPI: Dynamic Parameter Identification of Rigid Body via VLM-Seeded Differentiable Simulation PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:19:50.512556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T05:22:09.771975Z digest=sha256:3ec6006c7d81d4b679227fc9e4291c198b57ac72ca471986addcf1680af74766

Observation 3de5026f-5932-46d8-afa2-1f54c803ea22 · inbound

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping cites this paper.

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T14:17:02.469095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-02T14:16:39.649823Z digest=sha256:3e7d555f87b1f0b5f35336962d03fed205b074f49f034f1dbc874215908f3fa7

Observation e04f13be-3fad-462a-862b-e35eaebd4cb3 · inbound

Does AI Understand Imaging? A Systematic Benchmark of Agentic AI for Computational Imaging Tasks cites this paper.

Does AI Understand Imaging? A Systematic Benchmark of Agentic AI for Computational Imaging Tasks PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-09T18:06:25.899637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-09T17:59:03.869202Z digest=sha256:b460edf7407a51893c8b4acffc06e8502d805aac38a0179f3043cbfc98e3c71a

Observation 8299f42c-fbf6-4c4a-bca0-c50e31843b7c · inbound

RetroHolmes: When Semantic Plausibility Fails Retrospective Physical Process Reasoning cites this paper.

RetroHolmes: When Semantic Plausibility Fails Retrospective Physical Process Reasoning PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T07:27:08.281893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:27:08.281893Z digest=sha256:4cfce12fa8802a29b6b2a61df87e4e9733a5a69c62a44898f7fd5b32f8aedadb

Observation 1c6b2d5e-c3fd-4ab7-922d-cf3aabae0e3b · inbound

Apple-$\pi$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence cites this paper.

Apple-$\pi$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T21:07:24.575800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:07:24.575800Z digest=sha256:2981850ee70416886e9dcdb1d926f046f4e60eef86119d73ae79a6aab82afef2