Pith. sign in

Paper Citation Record · LEDGER

VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2408.02629.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.02629 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T16:56:33.961865Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T07:56:04.692157Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d3ea8e14-b555-4cb7-8a88-adabe49cab4e · inbound

Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation cites this paper.

Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:40:00.032419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T14:39:59.870039Z digest=sha256:ef4f35fe3dd27039cd29c4d591c8a3219b222c59dadc4fb3fcf8a3b0440caee3

Observation 390de76e-87ca-478f-acca-2019bac7a998 · inbound

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models cites this paper.

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:00:27.203120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T17:57:57.263574Z digest=sha256:39321604015003c7d6df152e4e040756fd145a32db8a331197f8edd0d9bdf7f5

Observation e9a2f9df-bd54-4244-9eed-4e6bf4ae1808 · inbound

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models cites this paper.

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:28:05.580957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T16:25:03.743594Z digest=sha256:129204e68dc76819a3de1761f59a9f4791045b2c3bf185eeda88f633ddb749b0

Observation c9616a39-1e50-41c6-94d1-c11ac9fe2134 · inbound

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models cites this paper.

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-21T16:04:14.709987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T16:01:52.150950Z digest=sha256:1217321dddf7ebe6db231c4a72f7056cd066cdb44957c682868cde9d9a7e05cf

Observation c77cf558-2b39-49cc-b1f4-fae8875e98cb · inbound

VDCook:DIY video data cook your MLLMs cites this paper.

VDCook:DIY video data cook your MLLMs VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:56:19.221954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T16:52:36.882218Z digest=sha256:6d060ea840e8668776cfa2de40c63e7f929e065672409620bc15d1f4e234164a

Observation 08fdc825-9988-48cb-86a3-653e8224fdc7 · inbound

Can Video Diffusion Models Predict Past Frames? Bidirectional Cycle Consistency for Reversible Interpolation cites this paper.

Can Video Diffusion Models Predict Past Frames? Bidirectional Cycle Consistency for Reversible Interpolation VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T16:56:33.961865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:56:33.961865Z digest=sha256:8707876454710561b4191ea77d205d303e649d099188e2a6b6159b5ca0ddee59

Observation c20ae787-89ea-405c-92c4-52f261aa6189 · inbound

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing cites this paper.

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:45:50.458209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T18:54:47.451103Z digest=sha256:57b741f7336d02f30df096a9a4b9c9a8fc8cb35d42dda9fcb3bf323f76de0ea3

Observation e1562e97-2246-4edc-921f-17f561d6d65e · inbound

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer cites this paper.

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:46:49.069232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T04:15:20.045060Z digest=sha256:17274d0999d8f613bdc5c8601a3ae50bf401978377a71a4c67d4013bbf0680c6

Observation cdfcdd4e-e450-4e2a-a17b-86b549ad0622 · inbound

Variance Reduction for Expectations with Diffusion Teachers cites this paper.

Variance Reduction for Expectations with Diffusion Teachers VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:53:57.919961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T04:51:57.800006Z digest=sha256:d9201a9428c625523845839afeb56f87fa80172e3fe7273af7eb666d5cf21704

Observation 3e10c653-0cca-4927-a583-1d6cceba155a · inbound

Variance Reduction for Expectations with Diffusion Teachers cites this paper.

Variance Reduction for Expectations with Diffusion Teachers VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:24.685953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T05:41:10.076206Z digest=sha256:2e564f065c5d2750be454d22ffad700f034e7efb48f1ca8afe6682bdb1ae822d

Observation 1fa789f5-e1b1-4950-8b96-b313095709a1 · inbound

Toward Native Multimodal Modeling: A Roadmap cites this paper.

Toward Native Multimodal Modeling: A Roadmap VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:04:01.932824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T22:58:38.610609Z digest=sha256:69ca53abd0c4a3518f570e6ba723c0dfafc38ba0b17afe195d37a98ec3991ba5

Observation f9d4710d-4d57-43be-8766-893118d790e6 · inbound

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation cites this paper.

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.776236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T11:04:49.366127Z digest=sha256:62b3ba42f73fc29cfbb68481949f2b4b9684b0bb92709f0690d805cbfe819030

Observation 0f8ec3eb-70b6-4884-b4e6-72129c28253c · inbound

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment cites this paper.

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:16:44.728428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T07:02:37.291472Z digest=sha256:b034d27d705e7adc8107fb174885ae39dc73c7e3eeec4bd76e519a2a861826ef

Observation a6023803-aa52-45f6-997c-ca561d221cb9 · inbound

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation cites this paper.

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 62

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T00:07:28.154541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T17:30:25.371658Z digest=sha256:8c42f4535942cb04259a8dc469141a26e50214023b7d98f7d11f854bd0e39ece

Observation dcd6ee96-b556-420b-b064-51f048eb7cce · inbound

Kairos: A Regret-Aware Native World-Action Model Stack for Physical AI cites this paper.

Kairos: A Regret-Aware Native World-Action Model Stack for Physical AI VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:28:44.842107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T04:02:53.110012Z digest=sha256:633a57c101dd74172af02d36eca09fc308e55b3860ad69962ea30b4ef7aa8343

Observation 5f133fd3-a200-47b8-b965-12e576e6ac11 · inbound

Infinite Worlds with Versatile Interactions cites this paper.

Infinite Worlds with Versatile Interactions VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-07-09T07:56:04.693407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-09T07:51:36.802801Z digest=sha256:e97e83ec4676560307827d16d21490a4d76516e40ccbaa29badcb4e0028ec46f

Observation 500fa9be-3d54-4470-8bef-a7b993569353 · inbound

Group-of-Latents: Perceptual Video Compression at Extreme Bitrates via Masked Latent Generative Modeling cites this paper.

Group-of-Latents: Perceptual Video Compression at Extreme Bitrates via Masked Latent Generative Modeling VidGen-1M: A Large-Scale Dataset for Text-to-video Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T14:30:46.262409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:30:46.262409Z digest=sha256:ec6e7e8680d4e951b80b1f654d03961044800a85ac9b92972099b622d058468c