Pith. sign in

Paper Citation Record · LEDGER

PhyWorld: Physics-Faithful World Model for Video Generation

As of 12 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 1 inbound Pith citation observation for arXiv:2605.19242.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.19242 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-20T07:28:20.248452Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-08T07:10:33.826140Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T07:14:45.189730Z

Reference resolution

64 of 64 outbound references displayed

  • verified exact37
  • verified fuzzy25
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a243100c-b0b7-43c1-b2e2-334f733df2c4 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

PhyWorld: Physics-Faithful World Model for Video Generation Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.628793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:6ebb9225c5dd738e26b9faa9b4b8c2e3659a86adc31083304552db711ef65fda

Observation dbac46c9-33bb-4c4c-8dfe-5866633ec6b0 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

PhyWorld: Physics-Faithful World Model for Video Generation CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.615118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:7255c4ec9016b90bb13b5f8ad7fb34c93b8f32c97bc4b8ef9d4d2c176cc50776

Observation 69863998-34f2-4cb3-8c98-2f504caefe43 · outbound

This paper cites Video models are zero-shot learners and reasoners.

PhyWorld: Physics-Faithful World Model for Video Generation Video models are zero-shot learners and reasoners

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.612602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:10b745fec5e0ec8523ccf1ac3c4fe0ac47577e517cbb022114d513e425e0dbc3

Observation 4f61161b-69bc-4e99-b58a-21de0deb8575 · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

PhyWorld: Physics-Faithful World Model for Video Generation Cosmos World Foundation Model Platform for Physical AI

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.573153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:34ef5883b7a26351f3baec8735c5767464eda36647db05f7b13ae49f57c8fa7f

Observation ef4eaf4e-8cce-4cba-aa86-4fef9eae7410 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

PhyWorld: Physics-Faithful World Model for Video Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.578777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:d7ccef4127fde865d823647fc0d5f85bfa1092da365433f2c8232ecb973677c7

Observation 47e5b0af-03d1-4c8e-a2a4-10e4aacd9b11 · outbound

This paper cites VBench: Comprehensive benchmark suite for video generative models.

PhyWorld: Physics-Faithful World Model for Video Generation VBench: Comprehensive benchmark suite for video generative models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.999229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:e79b108b5beadd21cf169f6b680802481615ff692341f953252a026bc61a4a8f

Observation b0eb3ff4-f7a9-4687-b518-b561f545bd77 · outbound

This paper cites Understanding world or predicting future? a comprehensive survey of world models.ACM Computing Surveys, 58(3):1–38.

PhyWorld: Physics-Faithful World Model for Video Generation Understanding world or predicting future? a comprehensive survey of world models.ACM Computing Surveys, 58(3):1–38

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.997551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:76034cacf64238699dd2994aea6bd9338415af4ba2f18f3e4fb62c1cce0e1d3d

Observation b168b687-71c0-4596-935d-8fa870bf346f · outbound

This paper cites A Comprehensive Survey on World Models for Embodied AI.

PhyWorld: Physics-Faithful World Model for Video Generation A Comprehensive Survey on World Models for Embodied AI

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T01:14:24.976532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:1b30bde03a9f601ab686a47dd817801ac3f657e99eb50371a03034beb4a6cb26

Observation 72ea1d9f-91aa-4c88-bc1c-f78fecaa06e2 · outbound

This paper cites Simulating the visual world with artificial intelligence: A roadmap.

PhyWorld: Physics-Faithful World Model for Video Generation Simulating the visual world with artificial intelligence: A roadmap

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.589867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:71ff0816bf8cce34df0e56bdf4cd4dcfcedfee4a4c34cfd32e933717a6082353

Observation dace664c-aa55-47b6-b0a2-dd75fa405e78 · outbound

This paper cites A Survey: Learning Embodied Intelligence from Physical Simulators and World Models.

PhyWorld: Physics-Faithful World Model for Video Generation A Survey: Learning Embodied Intelligence from Physical Simulators and World Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.610198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:df2ae11a6d2a0c6395097dcc7f3065a17459d9a5b28f4df4b35481c340cb2d78

Observation fb8c8f22-cbd2-476d-9e90-5a9de5290a75 · outbound

This paper cites Open-source multimodal moxin models with moxin-vlm and moxin-vla.

PhyWorld: Physics-Faithful World Model for Video Generation Open-source multimodal moxin models with moxin-vlm and moxin-vla

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.553889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:b58e8952452fb7609a95db03df16f74add044cc968f1d280aee8485a0050a047

Observation bde30324-447d-45fa-b538-054dd951d0a9 · outbound

This paper cites 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement.

PhyWorld: Physics-Faithful World Model for Video Generation 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.545335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:62ec08bc5b5ca6f2cecfd8cbd65d21269102675e0fd56bfe1390bc0c16974338

Observation 10f0360c-90d3-41a6-8c6c-d45776729e8b · outbound

This paper cites Exploring the Evolution of Physics Cognition in Video Generation: A Survey.

PhyWorld: Physics-Faithful World Model for Video Generation Exploring the Evolution of Physics Cognition in Video Generation: A Survey

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.537094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:283a5869f2c3fefab4dde24203eb0c83e995ade2d96c09709c4abcd993b8ee8a

Observation caff9054-6f0d-4667-a594-d415a215614b · outbound

This paper cites Generative Physical AI in Vision: A Survey.

PhyWorld: Physics-Faithful World Model for Video Generation Generative Physical AI in Vision: A Survey

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.539755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:aa1a62502f15f4693d8c48bebb34fbd34c2b8fc5ae39b2eb158d77ebdd9a11d2

Observation bd2c83cc-563e-45d0-aa7c-f1d7083a7278 · outbound

This paper cites From specialist to generalist: A comprehensive survey on world models.Authorea Preprints.

PhyWorld: Physics-Faithful World Model for Video Generation From specialist to generalist: A comprehensive survey on world models.Authorea Preprints

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:24.000912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:15e375489e6b7c6b35442bb1f89238686f2916d48ae9e0ff985fda87facae1ca

Observation f909970c-5778-4f22-acf2-a95a4dfbdb3f · outbound

This paper cites Learning to model the world: A survey of world models in artificial intelligence.Authorea Preprints.

PhyWorld: Physics-Faithful World Model for Video Generation Learning to model the world: A survey of world models in artificial intelligence.Authorea Preprints

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.995798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:9b5760c76afcc01a95d89d530fd03ba34c232d6e6cf9d1f0a80dff3b6d725436

Observation f4b160fd-9140-4d17-838c-7e05c5f750ca · outbound

This paper cites Squat: Quant small language models on the edge.

PhyWorld: Physics-Faithful World Model for Video Generation Squat: Quant small language models on the edge

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:24.004617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:e3f0b870c922cfcd207016963ca0da5aa976648619404d387cae21e5b7e5b0db

Observation 040b0f30-9bcc-47a4-8d07-a98bcc4d105a · outbound

This paper cites Pruning foundation models for high accuracy without retraining.

PhyWorld: Physics-Faithful World Model for Video Generation Pruning foundation models for high accuracy without retraining

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.990665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:c656231582934f6add1cefb410fca090f6445f9258ca414b533c74ab6b6c6c0f

Observation dcf7a5e2-958c-4585-8bf4-f5da809eca32 · outbound

This paper cites Search for efficient large language models.

PhyWorld: Physics-Faithful World Model for Video Generation Search for efficient large language models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.994066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:507835781da79d90d7f7c8a867f6fe99d5cf8ce0aefbcba77628d473f5b9257c

Observation 1e838db4-0b61-47bb-b7e2-79f236ccbfe5 · outbound

This paper cites Quartdepth: Post-training quantization for real-time depth estimation on the edge.

PhyWorld: Physics-Faithful World Model for Video Generation Quartdepth: Post-training quantization for real-time depth estimation on the edge

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:24.002714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:e52c943d78f425877ffd98e9a435a9191a446a837ed1d71d0c5f003f2648b12e

Observation 4b87c62e-455c-45df-bf3b-1db012641db0 · outbound

This paper cites Hierarchical World Models as Visual Whole-Body Humanoid Controllers.

PhyWorld: Physics-Faithful World Model for Video Generation Hierarchical World Models as Visual Whole-Body Humanoid Controllers

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.548310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:9717504beba25dc6af5c17051a69dcab774fb50168d661776ee010c93528d7a6

Observation e1e3d42c-8b2a-46a1-9028-35a7909cd14e · outbound

This paper cites Learning latent action world models in the wild.

PhyWorld: Physics-Faithful World Model for Video Generation Learning latent action world models in the wild

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.533053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:bb6b8ffbc392abc31e8f11ec206d44f68736fdff173f08772a01eeb13a725327

Observation 4acf341c-0fd4-47f1-a9f5-413b98451037 · outbound

This paper cites arXiv preprint arXiv:2601.10553 , year=.

PhyWorld: Physics-Faithful World Model for Video Generation arXiv preprint arXiv:2601.10553 , year=

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.620861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:7b7169a785a774b035aaa4dba12f81cd767ea807a45acb8688dc2fcc44d198b7

Observation 018a42f3-e417-4a93-a6f4-d1d56cdc4cb7 · outbound

This paper cites Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning.

PhyWorld: Physics-Faithful World Model for Video Generation Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.600904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:c0fffe4143723035dd67c3bb8b20a7bf67b86755a9f42e1b9e39a9026cf2c1b0

Observation b91a31d6-730a-49cb-88f7-831c0919cd02 · outbound

This paper cites Cambrian-S: Towards Spatial Supersensing in Video.

PhyWorld: Physics-Faithful World Model for Video Generation Cambrian-S: Towards Spatial Supersensing in Video

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.521126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:2536ababbbf938b83a1956bce1f63cd9f2300f6b1d199547d80df2b53ae1ede0

Observation 5b37fdd0-2aa8-41d2-aac4-09e63f2f7e24 · outbound

This paper cites Vagen: Reinforcingworldmodelreasoningformulti-turnvlm agents.arXivpreprint.

PhyWorld: Physics-Faithful World Model for Video Generation Vagen: Reinforcingworldmodelreasoningformulti-turnvlm agents.arXivpreprint

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.524250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:57b84513d1c520ad3b2947e363ea8e57b48859f89b47a984054b97aca8b1a2be

Observation a6f9cb59-6c50-44d8-86a7-e073954cb883 · outbound

This paper cites arXiv preprint arXiv:2601.03782 (2026).

PhyWorld: Physics-Faithful World Model for Video Generation arXiv preprint arXiv:2601.03782 (2026)

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.530191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:2df29989ee67a9c5d739bc35eac093b1b5100b64eba7406bc65ced5a764aabca

Observation 903b7c8e-3ab8-4d44-9234-6385fd89c480 · outbound

This paper cites Sparse learning for state space models on mobile.

PhyWorld: Physics-Faithful World Model for Video Generation Sparse learning for state space models on mobile

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.992468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:c2d9d8f16985fb20912dc6bf345c244ee4d92e7d3c05ff1bd4bc8ed84c4f4258

Observation 6e57d839-7afd-4229-abb2-bd774f092a6a · outbound

This paper cites Exploring token pruning in vision state space models.

PhyWorld: Physics-Faithful World Model for Video Generation Exploring token pruning in vision state space models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.985417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:b67af8be90dea7df7f84220643460e6fe461ca99ee481960b6529d420d8b9b13

Observation a74ace31-f8d9-4ed5-bb9b-435b044f7d52 · outbound

This paper cites Rethinking token reduction for state space models.

PhyWorld: Physics-Faithful World Model for Video Generation Rethinking token reduction for state space models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.987055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:3fe0e813973a38c58483f68b082f3441f3207c9675b6f26b6ef43f96a3242296

Observation d49bb6d2-5ba1-4e59-b39d-69c208eecbc3 · outbound

This paper cites Cocopie: enabling real-time ai on off-the-shelf mobile devices via compression-compilation co-design.

PhyWorld: Physics-Faithful World Model for Video Generation Cocopie: enabling real-time ai on off-the-shelf mobile devices via compression-compilation co-design

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.988868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:16f804c9b036df1346555b8b3b2da07cbbd6618a6d0513dcbc9239ef558573a0

Observation 44484835-d87a-40a3-b97c-de2df62e74e1 · outbound

This paper cites Effective moe-based llm compression by exploiting heterogeneous inter-group experts routing frequency and information density.

PhyWorld: Physics-Faithful World Model for Video Generation Effective moe-based llm compression by exploiting heterogeneous inter-group experts routing frequency and information density

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.604061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:98911d568f8015322ca8dafd88f985519d7a4c76868b5f0397032d9ebc221245

Observation df3c2012-f541-4a6c-887a-c845b770dbd9 · outbound

This paper cites Causal World Modeling for Robot Control.

PhyWorld: Physics-Faithful World Model for Video Generation Causal World Modeling for Robot Control

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.607462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:453e5ee0607657794709c8b88831dc6b9a8c81b9ba669ecb7079f95fff609241

Observation dcb20b8a-4475-4a44-bbec-5350b4b7ddf1 · outbound

This paper cites Advancing Open-source World Models.

PhyWorld: Physics-Faithful World Model for Video Generation Advancing Open-source World Models

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.618043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:51dd618117de261a23028ca92ed9cddc53e3394bb552bae23cd3a9de3cb01903

Observation d1936401-242f-4e89-94b2-bf03d872615a · outbound

This paper cites VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting.

PhyWorld: Physics-Faithful World Model for Video Generation VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-09T02:20:27.970763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:2daee745bbe0a81cf730bd00f9ccc600a3d6a9e0e6d665ad9f115bdc9fa68b12

Observation 90d011c2-cb5b-4d9c-aa3e-cce469f53a64 · outbound

This paper cites AdaWorld: Learning Adaptable World Models with Latent Actions.

PhyWorld: Physics-Faithful World Model for Video Generation AdaWorld: Learning Adaptable World Models with Latent Actions

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.598106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:2f7a8963bd00dbec2da648e90c4ab945173479eda07562022403b35748d28b9c

Observation fe07e540-b813-4295-b783-984b027b1a96 · outbound

This paper cites Fastcar: Cache attentive replay for fast auto-regressive video generation on the edge.

PhyWorld: Physics-Faithful World Model for Video Generation Fastcar: Cache attentive replay for fast auto-regressive video generation on the edge

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.981356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:02e2b6e974df007de284bd34c395e836a63dfe018ad70f232c88c23c036c3ad2

Observation a5e7048f-1c29-493f-8d75-2eece2f1ddfc · outbound

This paper cites Numerical pruning for efficient autoregressive models.Proceedings of the AAAI Conference on Artificial Intelligence, 39(19):20418–20426, Apr.

PhyWorld: Physics-Faithful World Model for Video Generation Numerical pruning for efficient autoregressive models.Proceedings of the AAAI Conference on Artificial Intelligence, 39(19):20418–20426, Apr

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:24.008451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:024d91c8b815d252a53db25758161671182d44d2736ea8e9d5676336bcd5a8e6

Observation aaf5a6e6-c5a0-44c4-b525-ba72b4946597 · outbound

This paper cites Lazydit: Lazy learning for the acceleration of diffusion transformers.Proceedings of the AAAI Conference on Artificial Intelligence, 39(19):20409–20417, Apr.

PhyWorld: Physics-Faithful World Model for Video Generation Lazydit: Lazy learning for the acceleration of diffusion transformers.Proceedings of the AAAI Conference on Artificial Intelligence, 39(19):20409–20417, Apr

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.979461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:1c6733cbc3f25257662b99578c1467a47688b5d1b8c19f86511d77a9122ee7d4

Observation 0b2ddcb4-849e-4149-aaf6-62ce12478855 · outbound

This paper cites Epona: Autoregressive Diffusion World Model for Autonomous Driving.

PhyWorld: Physics-Faithful World Model for Video Generation Epona: Autoregressive Diffusion World Model for Autonomous Driving

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.556616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:0e9aefa9730289920fa84d5d99d2b1cb55ab0bb695161d68c75417da6e06c73d

Observation a89f8daa-b429-4626-8d22-727e5ddee000 · outbound

This paper cites Hieramp: Coarse-to-fine autoregressive amplification for generative dataset distillation.

PhyWorld: Physics-Faithful World Model for Video Generation Hieramp: Coarse-to-fine autoregressive amplification for generative dataset distillation

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.626220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:65ab24eb03a1398a2b2865c1952d732e9e84e0d760efa115d4c4f75263c5c3da

Observation 1e7163e2-7600-4f81-a03c-ac8720f9bc56 · outbound

This paper cites Taming diffusion for dataset distillation with high representativeness.

PhyWorld: Physics-Faithful World Model for Video Generation Taming diffusion for dataset distillation with high representativeness

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:24.010056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:dab124b2f090b9e432e458dded9e349bb5b44fce4d30223dba9c4f896ed38ea9

Observation 278ab391-940f-4a33-9bd1-50592993c81e · outbound

This paper cites Fast and memory-efficient video diffusion using streamlined inference.

PhyWorld: Physics-Faithful World Model for Video Generation Fast and memory-efficient video diffusion using streamlined inference

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:24.006406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:d16f42bd390cdd95d699c5de739ed5be627d94220f48dd8e76bb2f98e65a4049

Observation 08cedce0-791c-491f-a7bd-6765f6157c5c · outbound

This paper cites DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions.

PhyWorld: Physics-Faithful World Model for Video Generation DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.527071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:69835ea3c5482b9573256249c97a3dd1fe459dda65a416e998a5185ae315b183

Observation 4210e69b-de55-4697-bab6-793cc986516a · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

PhyWorld: Physics-Faithful World Model for Video Generation Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.542421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:435c94cb147f3b46c4b325949c5f9a327800c074eaddbf7bbab6a68db08e5d98

Observation 4fff17bc-8e7a-47d8-a392-21834bc480ee · outbound

This paper cites Self-Forcing++: Towards Minute-Scale High-Quality Video Generation.

PhyWorld: Physics-Faithful World Model for Video Generation Self-Forcing++: Towards Minute-Scale High-Quality Video Generation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.551107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:ac947e74f74a3ef9b2fecd009575679a0dc80f96a34f908b2b93d3d3de53ae42

Observation f13a3dc9-a0cc-4598-af60-67ada1d25852 · outbound

This paper cites Longcat-video technical report.

PhyWorld: Physics-Faithful World Model for Video Generation Longcat-video technical report

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.977922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:af9d5dd226f3d7dbd07a314466e6effde35f60d986f29064e06bb6378b1658ee

Observation c54966d6-c06a-4c14-9e8f-8c99236634af · outbound

This paper cites LongLive: Real-time Interactive Long Video Generation.

PhyWorld: Physics-Faithful World Model for Video Generation LongLive: Real-time Interactive Long Video Generation

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.587053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:2a17e08592421b70451a0f9ff16bf58e8a54b4dce5f4f569c679e3a03cfe7eac

Observation 2b44d8fc-f696-42c0-8da9-ba4b92d84f1f · outbound

This paper cites Longcat-next: Lexicalizing modalities as discrete tokens.

PhyWorld: Physics-Faithful World Model for Video Generation Longcat-next: Lexicalizing modalities as discrete tokens

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.581660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:c92d80e7eae3d0f4df9c96fd0aa15ae2264e2b095e66f13ab84fdbfd353112ab

Observation ebd4de47-58b6-4b64-895e-80c36d70732b · outbound

This paper cites Do generative video mod- els understand physical principles? InProceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, pages 948–958.

PhyWorld: Physics-Faithful World Model for Video Generation Do generative video mod- els understand physical principles? InProceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, pages 948–958

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.975853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:a03effbc28b24be5a73e91982482c40af51d1c0647cbea403a821332b4a5e931

Observation 1c47e46a-24a2-4735-bc32-eb160440f816 · outbound

This paper cites Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation.

PhyWorld: Physics-Faithful World Model for Video Generation Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.584301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:012a1ec115b6f415d00903e946bb62ee37ee4efb055fab6649b779367500a21a

Observation a98e831c-1b5f-4ca3-91d7-63573d7e2c39 · outbound

This paper cites VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation.

PhyWorld: Physics-Faithful World Model for Video Generation VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.595472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:0a64bf69f711e89395974fbb27791084f181343af1ddfb0faa465bd8a7480662

Observation 0a2e3f03-c26b-4e87-80a9-c43cf7e7d9fb · outbound

This paper cites WorldModelBench: Judging Video Generation Models As World Models.

PhyWorld: Physics-Faithful World Model for Video Generation WorldModelBench: Judging Video Generation Models As World Models

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T07:33:07.576004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:33c823b2c942e5c95d9fd08a968f7bdb3b841ecfeb7113fcedf2b260124e6d39

Observation 8ca144e1-5ed4-4925-b76a-61507c2f3ad5 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.Advances in neural information processing systems, 36:53728–53741.

PhyWorld: Physics-Faithful World Model for Video Generation Direct preference optimization: Your language model is secretly a reward model.Advances in neural information processing systems, 36:53728–53741

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.977679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:2bd3d0528bc9aad6865801a6d7015e6d076151f281adaf95992906e750fc67da

Observation 9a4c9848-b501-4c95-ac9b-99a6d48e60a3 · outbound

This paper cites Learning transferable visual models from natural language supervision.

PhyWorld: Physics-Faithful World Model for Video Generation Learning transferable visual models from natural language supervision

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.984115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:daa5ff22d2823f28647b8729d6c8abb932db88ee06710a6ba05bc493038c3f6e

Observation 1590316b-66d8-4cf6-a53f-35724b89802b · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

PhyWorld: Physics-Faithful World Model for Video Generation OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.570212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:7dc96fe9e96f59f9f3abee7da9aa47516f381e2ee0c7331b376887a404a2b663

Observation 418e5d35-e002-4b34-b98a-f936a8df6214 · outbound

This paper cites Revisiting weak-to-strong consistency in semi-supervised semantic segmentation.

PhyWorld: Physics-Faithful World Model for Video Generation Revisiting weak-to-strong consistency in semi-supervised semantic segmentation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.974436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:9011df7155ef8b4734b9275ead8857ecfe3ba5706e27939af01f78d240375774

Observation 0b22deb0-df4a-4801-b95f-74480cac1358 · outbound

This paper cites Flow Matching for Generative Modeling.

PhyWorld: Physics-Faithful World Model for Video Generation Flow Matching for Generative Modeling

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.567644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:d16f6db1c1e31dcc15a3df030e0dd915a30b543ca851601086b90d6869134d75

Observation 364595ac-1e4c-4150-8781-e49882f28ef8 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

PhyWorld: Physics-Faithful World Model for Video Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.972242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:e2cd8beebf999dbba60b9dd58b51c27ffa50694aea2dc180080459afa5ceb12d

Observation d8f82144-93c8-4830-abcc-f777dab8c6dc · outbound

This paper cites Qwen3.5: Towards native multimodal agents, February 2026.

PhyWorld: Physics-Faithful World Model for Video Generation Qwen3.5: Towards native multimodal agents, February 2026

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:23.968557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:3e395a920354f8534d30aa14cd559474b80247b561307e6be36da3b947c32f21

Observation 302fea01-1e51-455b-8828-5b43743b2392 · outbound

This paper cites Diffsynth-studio.https://github.com/datawhalechina/diffsynth-studio.

PhyWorld: Physics-Faithful World Model for Video Generation Diffsynth-studio.https://github.com/datawhalechina/diffsynth-studio

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T07:33:24.011679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:97c346fc3ca3cff2b1f3a1a7891dc4d22343d5148f6471134a1832f6cecda34c

Observation 2749e119-40d4-4cd0-ace9-ca4057a14112 · outbound

This paper cites LTX-2: Efficient Joint Audio-Visual Foundation Model.

PhyWorld: Physics-Faithful World Model for Video Generation LTX-2: Efficient Joint Audio-Visual Foundation Model

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:33:07.592407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:358a18f4c0c93aca8c67f1d202924f03af8fe9cdf875d48c7864a0b5b5df1978

Observation 6f0e6895-1a37-4f47-92a9-4a038a87770a · outbound

This paper cites Omniweaving: Towards unified video generation with free-form composition and reasoning.https://arxiv.org/abs/2603.24458.

PhyWorld: Physics-Faithful World Model for Video Generation Omniweaving: Towards unified video generation with free-form composition and reasoning.https://arxiv.org/abs/2603.24458

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.559530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:72fa91a6a05f35194bfef422c7285e81b4bc0c3d8422005c44d7477248544096

Observation d1e93c55-63b7-496a-bac3-7e435c202e64 · outbound

This paper cites World Simulation with Video Foundation Models for Physical AI.

PhyWorld: Physics-Faithful World Model for Video Generation World Simulation with Video Foundation Models for Physical AI

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T07:33:07.562468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:9f49f2deffed07fbcd74bf4f95f9fef0d556dc87fa37e28e0e3c8e5fee5bf994

Pith citing papers

Observation 3701c19d-d1f4-4fc6-8a03-473ff01b4756 · inbound

A Definition and Roadmap for World Models cites this paper.

A Definition and Roadmap for World Models PhyWorld: Physics-Faithful World Model for Video Generation

Reference 74

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T07:14:45.191861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-07-08T07:10:33.826140Z digest=sha256:9414eebf7024d275a4c773b3ec546fb3f8b79ebf351851578a4a547c6f59ee80