Pith. sign in

Paper Citation Record · LEDGER

Towards A Better Metric for Text-to-Video Generation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2401.07781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.07781 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:52.645716Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T08:17:45.921041Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2bc85961-b5da-4802-9ecb-18e9262208b7 · inbound

SafeMVDrive: Multi-view Safety-Critical Driving Video Synthesis in the Real World Domain cites this paper.

SafeMVDrive: Multi-view Safety-Critical Driving Video Synthesis in the Real World Domain Towards A Better Metric for Text-to-Video Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:52.645716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:52.645716Z digest=sha256:0516129eb93915f79331edec7660ead5f9a1203d516c54e7aab916a45c1a36c5

Observation c48c9115-7c42-4700-92cb-3b8a15840695 · inbound

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation cites this paper.

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation Towards A Better Metric for Text-to-Video Generation

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:35.615058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:35.615058Z digest=sha256:e50952cd435bb5bf9a6fc4320447f7d77c8595d81da4884d777d535b1ba5450d

Observation d94fbeeb-4f90-4105-8084-6e11203b2c68 · inbound

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation cites this paper.

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation Towards A Better Metric for Text-to-Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:15.889647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:15.889647Z digest=sha256:73ae716d4993ce71234d2e3d7694b5e818c3cd29784528ce73db1a88269bc0d1

Observation e648002b-7a69-4abb-8ff1-b27fb60f6bd6 · inbound

Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation cites this paper.

Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation Towards A Better Metric for Text-to-Video Generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-05T17:19:20.415679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:19:20.415679Z digest=sha256:18563d427356f0dad585da3b392ae90721bc414a141df83f3f2f915e80d7de1d

Observation 00d1c516-17f7-4edc-ab3c-1c61dd76310f · inbound

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models cites this paper.

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models Towards A Better Metric for Text-to-Video Generation

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:10:52.960076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:39:26.420463Z digest=sha256:197f56da80208d268e1fadd26a387440f0538c9007814245ed50ce32cdab221b

Observation d82f33b4-5749-48de-9e74-4788ad43eb00 · inbound

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation cites this paper.

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation Towards A Better Metric for Text-to-Video Generation

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-03T08:17:45.922557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T10:46:56.871174Z digest=sha256:5af0b77a2d42dc588bb7a9e0765135cf9fc2198c112ffc2d81ba57b45938c71b

Observation 8739ceaa-b1c5-42ac-b2c1-4e45991434c3 · inbound

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models cites this paper.

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models Towards A Better Metric for Text-to-Video Generation

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:54:34.937252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T09:51:45.272494Z digest=sha256:24480d4928176ffa0d7885c9a26a5fa7d137e6149f41a11d16bcb16487e141b8

Observation 0beb3a96-74c9-4d75-83ef-09340d4870a0 · inbound

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models cites this paper.

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models Towards A Better Metric for Text-to-Video Generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-12T11:18:58.558989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:18:58.558989Z digest=sha256:7be1044e2d2c8efbb7247831f6523aa8c096f3abd143b0f15765c24bbdb9eb67

Observation 71475724-7945-4ddc-a12c-9e501990faf3 · inbound

ParticleGen: A Multi-Agent System for Particle Effects Generation cites this paper.

ParticleGen: A Multi-Agent System for Particle Effects Generation Towards A Better Metric for Text-to-Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T02:48:05.322308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:48:05.322308Z digest=sha256:544f444cef72a187c28c1b120e2e5d09f9edad7c36c7fe78a2c2cdecb83f3181