Pith. sign in

Paper Citation Record · LEDGER

VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 53 inbound Pith citation observations for arXiv:2503.06800.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.06800 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 53 of 53 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:36:25.525906Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:19:51.020711Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 101b1ba2-fbfd-4999-b886-9c911a6af147 · inbound

VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models cites this paper.

VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:05.096694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:05.096694Z digest=sha256:f347c317e044aa14b23ba1ef399b021bd2a7af79f436a4de299661b0caf16da3

Observation 5952afe4-b42d-4238-8458-97351e4b5921 · inbound

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models cites this paper.

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:50.339759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:50.339759Z digest=sha256:dd22281ec308878cbe60dba092b177b288b5daa71b99319d12ec7d2d501c713b

Observation 5408ffb6-a523-4df4-b901-5475e57ec79a · inbound

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models cites this paper.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:00.441134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:00.441134Z digest=sha256:e393c49a1eda49b8bfe467cea1ee42af42eca413d37cddd78228a2b4335abebe

Observation 8814fc66-f840-41c2-83fd-f5297fd85d97 · inbound

Human Preference-Aligned Concept Customization Benchmark via Decomposed Evaluation cites this paper.

Human Preference-Aligned Concept Customization Benchmark via Decomposed Evaluation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:59:25.198610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:59:25.198610Z digest=sha256:d94f4d7cc20ceabdee6f3fd923fb52fb95aab2b720b027cbfee73d05f8097408

Observation 3512bfe1-4b8d-4c9d-99a2-9424d273ac60 · inbound

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility cites this paper.

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.343732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T12:55:42.679016Z digest=sha256:12b651d13d638cb7dadf03e62a1324e4945a24192d965a81e4854e97101fb706

Observation feea1a15-9db2-4ab4-8253-1d0e571742b9 · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:01:13.918325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:c3b4bf80172350016190d28722958b8bfbad21da528f36401b095b2ce02fdeaa

Observation 500a90b3-6903-4d27-8f9b-9b60434e040c · inbound

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models cites this paper.

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:59:08.743022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T05:55:11.495430Z digest=sha256:6b6ca7d5418f9765883a4746ff5897b61fa2244ecce75bec9dc83981f840fbc4

Observation 2eaf1f6a-91c2-4d27-b10a-c018831d97ce · inbound

Generative Action Tell-Tales: Assessing Human Motion in Synthesized Videos cites this paper.

Generative Action Tell-Tales: Assessing Human Motion in Synthesized Videos VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T19:11:47.098371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:11:47.098371Z digest=sha256:1f2fc05bfaad0c1da4931e2ec76261d215418568741feef83d75cd32891693ac

Observation 7661578d-9147-4d74-b3f5-cc93263bf7bd · inbound

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models cites this paper.

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:00:27.237634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T17:57:57.263574Z digest=sha256:c2f81503c66de7262cd1680bfa4fa913dd955a51c87bd456578455cafe135f7f

Observation 8a6a0aeb-f9d3-42e4-80ff-be1b46426c19 · inbound

ProPhy: Progressive Physical Alignment for Dynamic World Simulation cites this paper.

ProPhy: Progressive Physical Alignment for Dynamic World Simulation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:08:47.998118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T01:05:08.087136Z digest=sha256:fec69f00dcb6d9d3c371bfaf569cf0b3ebd1ebae5963cb5a81bf08611508ed29

Observation 98814dd8-c6bb-49dc-8692-50fc179782af · inbound

Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control cites this paper.

Agentic Physical AI toward a Domain-Specific Foundation Model for Nuclear Reactor Control VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:00:23.976691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T16:57:19.490074Z digest=sha256:39db092cd052a8c4d48ea49f227cfc9f5c42df536311aa49c3f20f92be300399

Observation 4ed3be5e-8e4f-4170-a9ab-4d69a8c8e3c6 · inbound

Self-Refining Video Sampling cites this paper.

Self-Refining Video Sampling VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T14:40:14.478534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T14:37:57.167882Z digest=sha256:6851f5fdfef5e4d3f0658b5c2dd415c9ed3dbc083d96db507bc236e51e3a1e5c

Observation c657b1b5-45b9-42b6-b311-c1d46f519cb2 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 170

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.471254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:620df8541c81e1f52fde81aca6b8143bf0b9271478f91c3bddfa167ee34853a4

Observation ddbbf14e-72a5-420a-abb7-34edde3c7d7f · inbound

MoRight: Motion Control Done Right cites this paper.

MoRight: Motion Control Done Right VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:26:01.050690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:38:03.776766Z digest=sha256:7faef875b23328fd070547bceaab9233372fb4965388d8e6f44d1b0a9c317ba2

Observation 6f66a35d-6b39-4dba-991d-3817f187d81b · inbound

PhysInOne: Visual Physics Learning and Reasoning in One Suite cites this paper.

PhysInOne: Visual Physics Learning and Reasoning in One Suite VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:25:59.537119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:39:48.066744Z digest=sha256:9ba8c7072909a78e3811da5321b70de6728a7737477774565ae3f16fdb2f76f4

Observation 9864ebcd-f30e-4cc2-9c58-0e545b7f4650 · inbound

How Far Are Video Models from True Multimodal Reasoning? cites this paper.

How Far Are Video Models from True Multimodal Reasoning? VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:51:03.643358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T02:44:52.920816Z digest=sha256:e18d90e4c2dfa21cde6facf868aa311e653cf1f23396297cc0e56fa55dea7a55

Observation 199032ec-d394-47bc-b93f-e3e2d2bee3c2 · inbound

BRITE: A Benchmark for Reliable and Interpretable T2V Evaluation on Implausible Scenarios cites this paper.

BRITE: A Benchmark for Reliable and Interpretable T2V Evaluation on Implausible Scenarios VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:41:33.649224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-09T21:12:33.209353Z digest=sha256:6094181b4c5721794df270e89801feb8c323b8c764a525c390c12fcf8e063ac9

Observation 236bef27-0804-4a57-a7eb-11e53c67c88e · inbound

Do Joint Audio-Video Generation Models Understand Physics? cites this paper.

Do Joint Audio-Video Generation Models Understand Physics? VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:50:56.767419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:12:04.230076Z digest=sha256:5f82a438463b03c04bd5e6e17945800889ec5d10f54ae5ccb1be694b85f471dd

Observation 68ca9a9f-28a4-45f6-97c7-000e535ba148 · inbound

Do Joint Audio-Video Generation Models Understand Physics? cites this paper.

Do Joint Audio-Video Generation Models Understand Physics? VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:45:08.203083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T23:39:22.070629Z digest=sha256:ac4ba6ab1aeeec8f1f29fe75b1e76346e0ab269a7cc6e15356238b12cf442fa1

Observation c4348ac0-7666-440b-b235-55241b7a45f7 · inbound

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models cites this paper.

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:40:54.609388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:21:52.861714Z digest=sha256:c420ad16d9bf3ae1b3931fafc7b1debeb678b091827222244b61393ff6c03d65

Observation ae39e17c-04af-48de-bf8a-868a40164b10 · inbound

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models cites this paper.

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:15:08.007606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T23:13:31.195562Z digest=sha256:11820b7fb63aa9ca5beb7326121b1f9260afdecb4c1f321e3de3feb1e3ed86af

Observation 42134aa3-572d-4f86-9217-752a2cba9445 · inbound

PhyGround: Benchmarking Physical Reasoning in Generative World Models cites this paper.

PhyGround: Benchmarking Physical Reasoning in Generative World Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:21:27.838360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T05:17:30.010064Z digest=sha256:1a2a9e09a105b3e13b3b9d4a156e8bfe75f2f3c6782982eae31891f9cf678f35

Observation c6006889-d1ff-4cf3-818a-0b4733a448b9 · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:22.162702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T06:00:31.582714Z digest=sha256:eb991941240deeeec0f47571a17f0e9daf7d3a2b94c096eda13e200be03b6b3b

Observation 2dac5e39-8a77-4e2b-a560-ced9b61bfc1b · inbound

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating cites this paper.

CaC: Advancing Video Reward Models via Hierarchical Spatiotemporal Concentrating VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:55:45.294459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T22:38:43.102769Z digest=sha256:eb97ff63d5e1802a7239c87363f200daa7d7eb024a490e39973fb3d97e52b416

Observation b5c2cfd1-1775-4c66-bdd3-46d10217d19f · inbound

DriveCtrl: Conditioned Sim-to-Real Driving Video Generation cites this paper.

DriveCtrl: Conditioned Sim-to-Real Driving Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:15:04.778745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T21:06:06.548538Z digest=sha256:f5562747eb7fb2526ce470b8a5c7a71e7e9aa5f56e2c370088c779dda844c626

Observation 881ec02c-7a65-4267-a803-556ef6b74eb5 · inbound

NEWTON: Agentic Planning for Physically Grounded Video Generation cites this paper.

NEWTON: Agentic Planning for Physically Grounded Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:13.762531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T10:56:22.343333Z digest=sha256:9426c7d6d357c12ee869aaef15996747cfe711a6724b7c377c7ef8379e4aa77f

Observation a98e831c-1b5f-4ca3-91d7-63573d7e2c39 · inbound

PhyWorld: Physics-Faithful World Model for Video Generation cites this paper.

PhyWorld: Physics-Faithful World Model for Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.595472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:e07e2be3f2839351546903cf1d0fb6a7a41641bb887b58e8e0f70ef0208eff97

Observation 5cd09e5e-1276-445c-953d-e2da461e1bdb · inbound

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models cites this paper.

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:40:23.190574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T04:39:22.400458Z digest=sha256:58afa5b15ad3dd77b40361ff292cdc3249250af6989ee174a051a55055ee831b

Observation c747a6d3-4c52-4ddd-b01f-5d5acc06334a · inbound

LaMo: Self-Supervised Latent Motion Priors for Physical Realism in Video Generation cites this paper.

LaMo: Self-Supervised Latent Motion Priors for Physical Realism in Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:45:20.481958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T04:42:32.717968Z digest=sha256:c7a114ae4b2222b36066be0f940c6a0b1a8e27d2f9b0c147b63798d81c134590

Observation 3916976f-d93f-406c-9aca-fa3e7b81f9c7 · inbound

Tempered Self-Similarity Alignment for Physically Plausible Video Generation cites this paper.

Tempered Self-Similarity Alignment for Physically Plausible Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:44:38.371299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T11:39:06.597513Z digest=sha256:9aa552f1ce16245344d1ecba793dc10eb6bcdbf74d773885e4a860d389b1c3cd

Observation 31fa3a2e-2dd7-4854-ba29-969d8e717dd3 · inbound

WBench: A Comprehensive Multi-turn Benchmark for Interactive Video World Model Evaluation cites this paper.

WBench: A Comprehensive Multi-turn Benchmark for Interactive Video World Model Evaluation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:14:02.269296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T22:57:08.381846Z digest=sha256:293a8d6bdadff7fa1700335bd73adc503941c6694c6e09e26f2ca3193dfe2275

Observation 2dc8fd89-4e71-47a0-8c81-0f5c24d823d7 · inbound

What-If World: A Causal Benchmark for General World Models in Embodied Scenarios cites this paper.

What-If World: A Causal Benchmark for General World Models in Embodied Scenarios VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.403328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T18:23:22.987086Z digest=sha256:7bfc43cbdb02e90f4abc25c761689904bceee32993af73d62a18aae0e16de70a

Observation 928a2668-cecb-4338-97a9-034a1af2aa1c · inbound

Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation cites this paper.

Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:03:26.633426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:55:24.689338Z digest=sha256:9faa84054ef1e7edf3ac36a2c72422d2e2b8aaca242590110091ce23002b8cdd

Observation b092f7a9-95a8-4f84-9bd4-cd2b18a83278 · inbound

YoCausal: How Far is Video Generation from World Model? A Causality Perspective cites this paper.

YoCausal: How Far is Video Generation from World Model? A Causality Perspective VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:33:15.610723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T08:27:03.674229Z digest=sha256:dd6a3bb1dcd40405a84733a33cb292c1ee32e74831182ceb95ec1528e9f27c20

Observation 2a5c3853-3881-45ff-821b-facdf2267b1e · inbound

OptiWorld: Optimal Control for Video World Generation under Physical Constraints cites this paper.

OptiWorld: Optimal Control for Video World Generation under Physical Constraints VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:32:35.562695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T19:02:51.848742Z digest=sha256:32247abf3804d4e01aa136e0d2984a7e92fbd6cdd157cd564f83f9b565e7781e

Observation c199f755-4fc8-4cc0-8f1d-1a8b4870040a · inbound

MPMWorlds: Material-Point-Method Simulations for Inferring and Extrapolating Physical Dynamics cites this paper.

MPMWorlds: Material-Point-Method Simulations for Inferring and Extrapolating Physical Dynamics VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:16:24.684102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T12:19:27.221596Z digest=sha256:ad46fddf598063608e29f9f721363c772ae435443963c10f93c50dd49e3ca4a5

Observation d8879029-c4bb-428d-9030-f6d732fd5ae2 · inbound

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment cites this paper.

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:16:44.758956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T07:02:37.291472Z digest=sha256:3dbb777832f8a021c23176cdc512b5279f3d9ab14567896dd363f884a3f8901c

Observation db7ea51f-8389-43de-bceb-5e4a862c72a0 · inbound

Current World Models Lack a Persistent State Core cites this paper.

Current World Models Lack a Persistent State Core VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:31.017133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T17:33:41.461245Z digest=sha256:0ccaac74fc176f78c77c0fde4030bbc6d54f5b240bf0669e0f816116ae3e2022

Observation 7e2c8aa6-6fec-46e8-a275-c5d9d023e5a7 · inbound

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation cites this paper.

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:45.415761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T09:04:23.965554Z digest=sha256:2909bd709100c58d4e95fa23d31f7382d4c6dd218a06382f414cbc441bd43a98

Observation 03148a37-6e8a-4ca4-a273-bb1a937b2c8b · inbound

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation cites this paper.

PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:19:51.022182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:16:53.011837Z digest=sha256:1f212dcd6264b51fccb2dbe8ee7b9dca0808dfdb7d66d52d971b850ffbc09279

Observation 5b8ddd1a-9cf9-47f0-88fd-f3abfedb3e5a · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:25:57.779239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T02:03:45.564122Z digest=sha256:f0db85ceaae2d4e0711aa3b25ffe2dd6324296c3a434ef1053571a8bb7189561

Observation 776a5435-d44c-4177-9a53-4db6cf6b2e37 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:35:40.499609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T06:25:58.872140Z digest=sha256:7ce5db633d7deef3a67d3b84fe23b25b9656d3dd6105c7fbc64e6a08c35f6f73

Observation 9cd71609-d385-4009-b531-d27eaf95c46a · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:22.806256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T20:52:28.444524Z digest=sha256:43676a891f0af778321b4c5a68aa81ca5b2c448441d101c8279937687d6ec78f

Observation fbfc599f-82d6-4ac7-9073-49c0a272493f · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T22:49:00.899176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T22:44:16.272541Z digest=sha256:a79b53961876d1bae098264043984349fbee5e99c8fe84181217ed9de7bd80ae

Observation 06aee8cc-d380-4e35-81c5-407ff4d20c72 · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T17:14:19.770867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:14:19.770867Z digest=sha256:783b4c186b0956b729aacd120192790dd91c246cf873b2bf2cbcd2a46cfd5a2d

Observation 911c1311-d029-402b-81d2-40a9d4c148af · inbound

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments cites this paper.

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T10:00:09.187494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:00:09.187494Z digest=sha256:d631d793d462621e0c5fd3a3b7aba60cf5ced819e84d6f536656c5489b4db3da

Observation 42136782-eba8-4cb4-8ca0-4f81411e752f · inbound

Apple-$\pi$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence cites this paper.

Apple-$\pi$: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T21:07:22.906584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:07:22.906584Z digest=sha256:065b2d2d67af1a8aeeb1017f15b7ff1b6675251277fc3f56605b359e1b5f0e7c

Observation 087d6028-ce69-4bf7-bf1d-e1f662b0f4e2 · inbound

When Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video Generation cites this paper.

When Physical Preferences Meet Semantic Constraints: Physical and Semantic Direct Preference Optimization for Text-to-Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T19:33:03.502183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:33:03.502183Z digest=sha256:64b5ba2e0c448591edef7a533f363f40801f5de6873d49223e6c56eb74a359ab

Observation 5920ead6-c2f8-48fb-9f7d-3462801e86f6 · inbound

Thinking in Video: Can Video Generators Really Reason About the Real World? cites this paper.

Thinking in Video: Can Video Generators Really Reason About the Real World? VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T17:47:06.272600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:47:06.272600Z digest=sha256:804c6b4dffaa4a43bb963cf69bc74d52ce96936961645ca971d6859105a388b4

Observation a3f15e8b-0fbc-4760-a54c-db2a3b675253 · inbound

Learning Explicit Physical Parameter Control and Benchmarking for Video Generation cites this paper.

Learning Explicit Physical Parameter Control and Benchmarking for Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T14:02:10.808123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:02:10.808123Z digest=sha256:6f7bb2c7573b3aa70fd2d4173fecc3c22d4ccdf86614795d390102f72bd82004

Observation a6a7e5c4-8726-45ce-a12b-a2e89635b934 · inbound

Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision cites this paper.

Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T02:49:57.720734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:49:57.720734Z digest=sha256:ceb78b21f40e06ad9a3edbfa3ccc84d9d57a58968b0373e48cc38489811ab360

Observation 73a0adac-4e87-49e6-b669-293c9e9c6aea · inbound

VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System cites this paper.

VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T08:28:04.629823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:28:04.629823Z digest=sha256:ad452bd5a5185107c7acd7022ba94942dccb56bb1809e7ca648381522b9a8c4f

Observation 824e9a33-f5e3-4323-b33b-9064aae12a26 · inbound

GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models cites this paper.

GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T20:36:25.525906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:36:25.525906Z digest=sha256:704e48f568f1ea11352111d706825f44d2f7dd0b50fad0fcde6328f8899b1d8a