Pith. sign in

Paper Citation Record · LEDGER

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models

As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2507.15824.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15824 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:27:03.245735Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved28
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dd05d835-036a-4d73-a28f-484ef0c0855e · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:00.219696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:00.219696Z digest=sha256:1d475aa7e3a1e7ab7886b5e1df82049788e450db89014fc5e387f9acb62ce9cd

Observation 6d342f1d-ad5c-427e-b5bd-651c97f9a05e · outbound

This paper cites Impossible Videos.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Impossible Videos

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:00.284729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:00.284729Z digest=sha256:2901f3682f2dfeb93d399a8a7b82a39514ad3337b03cfa37d5b440c7126f838c

Observation ab5d0084-29aa-49f1-8b60-e22e7e23eb9f · outbound

This paper cites VideoPhy: Evaluating Physical Commonsense for Video Generation.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models VideoPhy: Evaluating Physical Commonsense for Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:00.383675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:00.383675Z digest=sha256:c7b5cd2b198de68f1f58661ed7ff3f8c7757894abc99cac56f9db427d6003be2

Observation 5408ffb6-a523-4df4-b901-5475e57ec79a · outbound

This paper cites VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:00.441134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:00.441134Z digest=sha256:1d97867b3c6e0fb60e9dc3226a0346354d30499deaf117bd84575f25e7b7d97a

Observation 0cd747e2-bf7b-4097-86cf-cb7ba2c3b498 · outbound

This paper cites PIQA: Reasoning about Physical Commonsense in Natural Language.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models PIQA: Reasoning about Physical Commonsense in Natural Language

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:00.511632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:00.511632Z digest=sha256:ef7775944532c9b1c89bde13e8fa064af1044a6bf4c9636268c541365da867f3

Observation 68e9ff5b-fc7e-4e3a-ab78-bdd91f96d9c4 · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:04.190158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:00.600323Z digest=sha256:955f4221706d43cea5dc85c63026babcce188ab351b9c24b74e16a3813c0a478

Observation 06aaccd1-4251-4057-a12d-a536cd6bce6d · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:04.168628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:00.668523Z digest=sha256:4cf609628eeb69087e46cade971302d767e1a4cd63a9626ea4b31f3d84147b90

Observation 443a8733-255a-4e0c-a0c6-9a7208b88e23 · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:04.152771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:00.772417Z digest=sha256:67148a7b089fbace57b71fe6af8c3ce53d2a3e00df942368887ba33ffbb856b1

Observation 203b6b81-9be3-4821-981a-47ddeeb5ec0b · outbound

This paper cites Gemini 2.5 Pro.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Gemini 2.5 Pro

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:04.137144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:00.825849Z digest=sha256:b55d965090ae2e7d8bf4e216887dd9b787f49a022578ed74e2740213bcb2dca4

Observation b1f2d662-8e71-42d6-872c-e817f9b13f2b · outbound

This paper cites Gemini 2.5 Flash.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Gemini 2.5 Flash

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:04.116230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:00.897645Z digest=sha256:8abb8f3a27bfa333c9513648f59b4ce6ccc143970419479c92b1a8e4baf4c84b

Observation 49e730d7-c9ba-457e-b563-20e622c9018f · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:04.097902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:00.996067Z digest=sha256:c1cb3e64f0a180dbe4e4676445f97c4eeb9ba51941ecbe8d5e9ba85800cb0d79

Observation e9be4478-8ace-4043-8239-d84d8ea79e36 · outbound

This paper cites LTX-Video: Realtime Video Latent Diffusion.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models LTX-Video: Realtime Video Latent Diffusion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:01.156060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:01.156060Z digest=sha256:42db1c36cb1ef420b12ecb761cc2b57df6ea7f57f54b036fa509b570b25cef41

Observation a72f13bd-15f5-4954-abc2-fc2f681ef209 · outbound

This paper cites Pre-Trained Video Generative Models as World Simulators.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Pre-Trained Video Generative Models as World Simulators

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:01.234367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:01.234367Z digest=sha256:2db48b145661035ad83d4e6898af8de1788fb4b325df819a2dbf7b0085e38038

Observation af6b60e2-c63c-45b3-8c6b-cdf8d396e8b1 · outbound

This paper cites Hessel, A.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Hessel, A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:04.064421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:01.325311Z digest=sha256:1128e21f1e1e3f84e8e5c3d56fde642e586b11a4d2a03fc6172998f2cf2564b6

Observation 6ec4b984-e348-489b-be63-be1a310b892a · outbound

This paper cites GAIA-1: A Generative World Model for Autonomous Driving.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models GAIA-1: A Generative World Model for Autonomous Driving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:01.391505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:01.391505Z digest=sha256:0fa4b974403571373c1057cc3044860bb649b3ffaf8830a1639dd576f0dc5d16

Observation 4efc76c7-0511-40a7-895d-2155be20621c · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:01.441313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:01.441313Z digest=sha256:4e2aec6a3fd7bee5802b8ec304e1ece249302ecadbafcb6a51ff0b7c98a3e24a

Observation 746c52b3-4100-4d68-a5ad-ffe926aa39fd · outbound

This paper cites Huang, H.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Huang, H

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:04.049667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:01.495108Z digest=sha256:a8f7e00c3c58d862f08a7274da5a134cafeacdfca4b0a5eb1f7c7635b13b604f

Observation d3e8b807-4e35-4a51-9bb8-da81b28bb603 · outbound

This paper cites Huang, Y.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Huang, Y

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:04.032429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:01.576906Z digest=sha256:ecdb3cd5f252d5e4dcc300827f8c84b67e4ce30fc39b9c54cd0a4199f642ef6e

Observation fa7dc48a-ed60-4b45-a638-f94f94f50c2b · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:01.639928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:01.639928Z digest=sha256:3a6bba5bb9355a06333d64c149c981470c751a75482359c9705470b2d2c899b6

Observation 06750cd7-951f-4862-acd7-51c0aaeca1bc · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:04.009403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:01.712330Z digest=sha256:979716fecd70317445dc7efd997ec17b5e41c36c17e81410fe580020cc9a98c2

Observation 8697019d-a8fc-476f-adc8-4201986812b1 · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:03.986899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:01.801312Z digest=sha256:98520f6dbc4048850af4aff331026c0897eafbe1a14b00d18f9d8b3aae709e63

Observation 2dac2889-bf4f-47b0-b914-8a8391058d78 · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:03.967150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:01.865776Z digest=sha256:f53851554fb7d541eb5344f291acdea1362941abeaa849774212d8de654b7dec

Observation 78df88ff-b6a1-4c01-b8c2-34c05b4f63a7 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:01.918575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:01.918575Z digest=sha256:8b6f029cfd7a79dbaff13fbdaaf39e2969efc38b99fb0491b91ff292a0f29935

Observation 27178bec-ac2b-4eb9-81d7-9bd31f9552fe · outbound

This paper cites Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:01.977216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:01.977216Z digest=sha256:199473e1b90eed2f359093484910400e5e3bcd7e53059235b66cf1e40cb8d6b5

Observation c9861670-97bb-4563-b7f1-469724aed6ba · outbound

This paper cites Do generative video models understand physical principles?.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Do generative video models understand physical principles?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:02.052076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:02.052076Z digest=sha256:12087d3208f758feec9fa34307705841ccb64bbc9f1dc5c1b0e043fc378011b0

Observation dd65b5ba-7e32-48b4-8fda-d7d3f0d98802 · outbound

This paper cites Sora: Generating videos from text, 2025.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Sora: Generating videos from text, 2025

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:03.949098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:02.115473Z digest=sha256:135451314524ed27a22298fe681576ba4329e34c9c8445bc11cfb0712b29f2c0

Observation 83d1dd11-8178-4bb1-b895-5aff8986a35f · outbound

This paper cites Gen-3 Alpha.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Gen-3 Alpha

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:03.931477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:02.234677Z digest=sha256:f0c51e67eb4d03dc1c86189f86897218e2b032eca2d37752f4310d00386ea154

Observation 2ca34fe6-b039-42c7-8c55-7e16e86dfc42 · outbound

This paper cites Salimans, I.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Salimans, I

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:03.911290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:02.385510Z digest=sha256:77ddb84ac7f6e5a554bb330370d81c934db0673c9020bf21d8a64107bf748c61

Observation bc2a468d-5cd6-4fd3-bffa-f615d734e8e7 · outbound

This paper cites Magi-1: Autoregressive video generation at scale, 2025.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Magi-1: Autoregressive video generation at scale, 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:27:03.893828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:02.553838Z digest=sha256:c526119be9a7d84997dcdf2a746d9b997ca197a9d0dfea04d9536a6a07b54276

Observation b4ea5221-cc9e-4341-8817-fc93eac75d0a · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:02.698103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:02.698103Z digest=sha256:40d07e47621f291eb2aeb95f8647ca1a94f392045477d544a581b5adaea676f8

Observation 6f08ae1b-8495-4218-9ab9-857cabf24641 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:02.864867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:02.864867Z digest=sha256:b53c8e55e422412c1df0f4a4e325cc39ee0c80426de564b6411fd0b8eb00970b

Observation fe0eb035-5071-4400-b47a-e6a17f6bbff8 · outbound

This paper cites WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:02.986657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:02.986657Z digest=sha256:ad6f5242583caf0abc127e77604f4f737128336777c68dbf5d6a641915b1af68

Observation b6464209-3ee2-44bb-89d8-af07f222587f · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:03.876651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:03.109495Z digest=sha256:280bcd2cbd935dac64d5d3228d23a8c4d7b9907bb62a903ad84c69d06979d1c1

Observation 8b847a99-dda7-4fbd-a632-715a95042b47 · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:03.857513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:03.230414Z digest=sha256:29ccb24d3b0a2871b5f5c114bb1143038798b5331c4a7d0b81477f93e57873f9

Observation 144296a1-b1b6-42ab-a870-9fc2f61ec084 · outbound

This paper cites MAGVIT: Masked Generative Video Transformer.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models MAGVIT: Masked Generative Video Transformer

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:03.235158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:03.235158Z digest=sha256:dec8a050648f0c7da1545f5d77b147c0a1cbecaeece4e728b276afbe416d585f

Observation 7b25baba-56db-4676-8722-d9ea2ea759f3 · outbound

This paper cites VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:03.240621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:03.240621Z digest=sha256:bbb7802b8651ef2a80374d56362a5087e825f0910df5e72761cad0adf9db95e7

Observation b82d2d0f-ccd5-4d98-a987-6b562a6e5018 · outbound

This paper cites are” rather than what agents can “do.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models are” rather than what agents can “do

Reference 37

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T15:27:03.834877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:03.245735Z digest=sha256:8e8816fe19a83b985434b3fa4c1bf8340fe0f5bfc1c699b39a8420dcb16499b0

Observation 9e4b52cf-bc2b-4562-9aac-05d737f56107 · outbound

This paper cites an unresolved cited work.

Can Your Model Separate Yolks with a Water Bottle? Benchmarking Physical Commonsense Understanding in Video Generation Models Unresolved cited work

Reference 2025

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:27:04.082066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:27:01.065176Z digest=sha256:445d21a62c818991371dc2b0f69d9b75059f8e9544516c93ccddbe7d8976e3b9

Pith citing papers

No inbound Pith citation observations are available.