Pith. sign in

Paper Citation Record · LEDGER

VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2503.23368.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.23368 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:34:16.094906Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:20:05.840081Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2a742e0c-8837-4f4f-8bb2-41edbf20e9c1 · inbound

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality cites this paper.

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:31.233075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:51:31.233075Z digest=sha256:a6eaccfec242206c897c51184cd94ed9c958ac2ec952c726e1b86a80027b51b7

Observation 84627d0a-877e-42ef-af92-e3f490eb5a14 · inbound

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility cites this paper.

Enhancing Physical Plausibility in Video Generation by Reasoning the Implausibility VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.394265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T12:55:42.679016Z digest=sha256:ff84de1e7eb8bc4f0b51f3baa8ad38d69e7fcbcee34d633b619b6e65d0755349

Observation ff32deee-3011-4c9c-815f-bc563eef382b · inbound

Self-Refining Video Sampling cites this paper.

Self-Refining Video Sampling VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:40:14.482113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T14:37:57.167882Z digest=sha256:4a3ae67c508aa315bb3c298b2a0c9744d09e7fa4a58f68fe31a40e563bad8b3c

Observation b431a42f-d677-4464-a42b-c033990e3954 · inbound

Vision Language Models Cannot Reason About Physical Transformation cites this paper.

Vision Language Models Cannot Reason About Physical Transformation VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-15T13:27:51.848177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:27:51.848177Z digest=sha256:998f0529115afc29b962219667d8fa872aa6127572414bc27ff1988527bd2b6f

Observation 79bd86d6-9257-4fc5-bd67-4474dcee7d62 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 300

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.719526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:0bfbab79f36533ab7fdc3d567dc745e3045022e001d088b3abbad93ace2bb724

Observation a083338f-2e36-480f-9393-1af2799bc958 · inbound

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving cites this paper.

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 137

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:26.419010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T04:13:37.421188Z digest=sha256:22d80f3c6856ded2cfdc5e1930028fe79f948e448df7129bf84095a8958c8cdc

Observation 494e971f-48f1-4b80-bfb2-dec0392206a3 · inbound

OptiWorld: Optimal Control for Video World Generation under Physical Constraints cites this paper.

OptiWorld: Optimal Control for Video World Generation under Physical Constraints VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:32:35.539164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T19:02:51.848742Z digest=sha256:230d8450366d30fdc59294edae96b11644ae56aaa088cb51becfcbd1fe475173

Observation 8cbd19d9-eb58-466c-ac65-eeab1a9da6da · inbound

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation cites this paper.

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:20:05.841727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-25T21:33:38.643889Z digest=sha256:ae69035f16016bd8b3b2b7190b27ccd1596ca6fb050a65c58a7052f0d4903a48

Observation 32225ca1-f466-4642-acc8-ea40ea9fd754 · inbound

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards cites this paper.

muSync-GS: Physics-Synchronized Driving Video Synthesis for Weather and Geometric Road Hazards VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T00:34:16.094906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:34:16.094906Z digest=sha256:07ebc01f2f0c6804d593878521e006689877d91068dd736beeb06676c27c4269