Pith. sign in

Paper Citation Record · LEDGER

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation

As of 21 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.18789.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.18789 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T14:24:46.123867Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 693973c7-37fe-447c-abe1-2a7cd9eabb6e · outbound

This paper cites VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.067822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.067822Z digest=sha256:ff31b53c81f0cf7223d11f233417a8bbea63dbc9967bfe0866d5a7a34afe13ff

Observation 0d35a70b-b969-4c19-8217-0940589ede0c · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.074247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.074247Z digest=sha256:676ddc51e96c2252293ae4f6fd8b41306fb4c8fc99f8d6a9d54cf65dd459a6b0

Observation 8a13afa6-1a26-45c5-acc6-b8c4bcb271d9 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Imagen Video: High Definition Video Generation with Diffusion Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.080717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.080717Z digest=sha256:f4cdc78d8ba7b5055390ede062cdb8069cc72bffeafd06d1c385a0c47b2e4c21

Observation 413adb3b-7dc5-4ed3-b8d5-8b34deb0ba8c · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Latte: Latent Diffusion Transformer for Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.092789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.092789Z digest=sha256:145cb53e1d4315e605340aa96cf5f82f7327370c370525df7ec294a082760e13

Observation c9b6c558-2227-4405-a9c4-9c87dd2f600e · outbound

This paper cites Scaling Data-Constrained Language Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Scaling Data-Constrained Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.095609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.095609Z digest=sha256:60ee3c69aae6d0df10d0ad9dbd893b3348794bac8058810faf582197dce10976

Observation 58b82d0c-4db1-4c0e-a33d-613e99a7c9d2 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Movie Gen: A Cast of Media Foundation Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.098413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.098413Z digest=sha256:b813fbe2bd78be5f4daf5dad34101ec7834accb07457e8222a06a5fec87b7dc2

Observation ae690f68-872e-40e2-b250-fb8819585b03 · outbound

This paper cites The layout bet, June 2026.https://reve.com.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation The layout bet, June 2026.https://reve.com

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.101322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.101322Z digest=sha256:892dbb1327ec3a111839a27dcc713c8aa56f09151a2222852772b9b27ae005b1

Observation dc336a12-9133-4c2f-ac8d-cf08f5064aef · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.104418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.104418Z digest=sha256:eb1c93157c3966b6d0d1ceeed9033b325df7b1147ffed9cc12e15381635e784b

Observation 2d70f25d-88dd-472f-a9bd-ff865892816d · outbound

This paper cites Kimi-VL Technical Report.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Kimi-VL Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.107195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.107195Z digest=sha256:91ff73b99f40ded7c00383a7b99c51fd085d7f60e092c14ebceadd64f49f133a

Observation 2c439224-b816-4c81-a181-933ffe2909de · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.109969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.109969Z digest=sha256:222948eb5b80d0c217c0546a088242c6fd8b1b1f028892f8952a7ab69119c890

Observation 74041b2e-5c19-4a61-af00-d706787a76e5 · outbound

This paper cites Qwen3 Technical Report.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Qwen3 Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.112606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.112606Z digest=sha256:6490c1463fe4372c0bc836023aa4be9fd556cca4267b2c52da8b4ef93cc851ea

Observation 769ef145-26bf-4f10-98fb-44334ddf91d5 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.115337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.115337Z digest=sha256:e7472f2008824ef5d1f8df8c7a752cbdf227505c4766b5046ecf78e78f0a9794

Observation 22de8f3c-08d4-4808-a226-b81939f715aa · outbound

This paper cites Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.118299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.118299Z digest=sha256:9b8f9b95d09c87f599bbc82f2dc989c6669f4f49e829d92c5828963322c49c0c

Observation b3bf8dc0-9717-4763-8671-a51cf65f0121 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Open-Sora: Democratizing Efficient Video Production for All

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.121154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.121154Z digest=sha256:801d6ae5253995ab13024c650e9a5667d80c6c678881191432f379dfa97cb327

Observation 3d381a34-1eec-44a7-ba60-4cb63def9d77 · outbound

This paper cites 4 lists all attributes used in the Moving Alphabet dataset and their possible values.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation 4 lists all attributes used in the Moving Alphabet dataset and their possible values

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.123867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.123867Z digest=sha256:dc5e7978372829877a84a4806bba8d75664ce53d21c1d2c89013efa96b0708b5

Observation b2c9eb7a-83df-48e9-b52f-0adfc932bb45 · outbound

This paper cites Scaling Laws for Neural Language Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Scaling Laws for Neural Language Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.086843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.086843Z digest=sha256:3e997b1c9e8f00509c50e73617836194e8170a9ac907f04b88b518ecd05a5599

Observation b6b3cb75-426b-4279-ab51-ca0b52e2215d · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.089947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.089947Z digest=sha256:3bd07797edd067b1654bb43461d77b28a89cb7cf92862569076b5e15a47c9eca

Observation ccea6b25-c376-4495-8069-b4f379bc0dea · outbound

This paper cites Classifier-Free Diffusion Guidance.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Classifier-Free Diffusion Guidance

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.077417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.077417Z digest=sha256:4902764bda7635cb1319615c0388ca464e74adb9971f353955ff7ce9b74ba225

Observation 73280caa-25ab-44dd-b35d-df6f459a6509 · outbound

This paper cites T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.083996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.083996Z digest=sha256:9bc22de29f489d0d7df2cf39179f76d37f935ea613cfd5bf9ff757dca2c371e0

Observation fecae1f6-d53c-4734-abb8-b07da78f8706 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.064539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.064539Z digest=sha256:50d63d4385f2164872087c8ff638c3768f6c11143c3d72728cd17fc3a34023fa

Observation 1607a6e6-7344-473d-bc6e-1917f964994a · outbound

This paper cites Lumiere: A Space-Time Diffusion Model for Video Generation.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation Lumiere: A Space-Time Diffusion Model for Video Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.060545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.060545Z digest=sha256:049bc524846a0eb1ff0ba518b8f5e831cf2b181fb5c4cfebb826b43c62428cd2

Observation f06cfb1f-d129-48dc-bc79-00a4a6f15663 · outbound

This paper cites TinyStories: How Small Can Language Models Be and Still Speak Coherent English?.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation TinyStories: How Small Can Language Models Be and Still Speak Coherent English?

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.071187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.071187Z digest=sha256:2c32e3e91fee0260b9170690db21cc323d59237412d91eb4c44bcaa878349e27

Pith citing papers

No inbound Pith citation observations are available.