Pith. sign in

Paper Citation Record · LEDGER

T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2407.14505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.14505 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:30:52.375066Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:39:50.643820Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4314d572-971a-4298-8ccd-03925a8e9be6 · inbound

Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation cites this paper.

Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:40:00.025640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T14:39:59.870039Z digest=sha256:3ba3aa265ce539470ca6c3937e2eb0ee40253b8d7d79c924e1fce66b06cc52b8

Observation 991a8122-6b02-489c-83ee-e497003b3735 · inbound

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement cites this paper.

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:25:29.416102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T08:25:01.468957Z digest=sha256:3b07bfaa4c119f78bf460d5b8162f85387887affc7558f2f645e586677d32468

Observation c3e7287d-1a9d-48af-9ee4-ebd163dfbe9e · inbound

VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness cites this paper.

VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-14T18:42:03.138921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T18:42:02.940250Z digest=sha256:2563dcbf51b96ef9199c2bcab1103e93799381f97534b5b68abf6e6ecc2421e9

Observation dbaf1bcc-bd34-48b7-88c3-6562d37530b3 · inbound

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback cites this paper.

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:56:58.291372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T18:56:39.735334Z digest=sha256:d2c1710389dc3ba6a48a48e2b325fb40be93cca48683ce16bc96b0d8d71396ee

Observation e9c22929-79e6-4272-95b3-dc5859d79c97 · inbound

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models cites this paper.

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:52.375066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:52.375066Z digest=sha256:3c2f7104f81009cd72ec560cb0a5c00bf31a971721a6f72ac5c837699b5e33b4

Observation 2c7118b9-ed99-4b24-a5e8-b6ac002c0669 · inbound

Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation cites this paper.

Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:28:42.060128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T21:28:41.904725Z digest=sha256:6baa0e44d587c512ef00e253dc8859c95b2bf1d07a64434ac2a4f63c44d15f1b

Observation 48546cb2-cd8b-442f-891b-28f49d5ca583 · inbound

FaceAnonyMixer: Cancelable Faces via Identity Consistent Latent Space Mixing cites this paper.

FaceAnonyMixer: Cancelable Faces via Identity Consistent Latent Space Mixing T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 2007

Resolution
unresolved
no resolver link, observed 2026-08-05T23:17:31.309031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:17:31.309031Z digest=sha256:cbe1cfc3be9c52dd0c9a4b993c775e14c69ab84e845d685101968893d6acb7b0

Observation d487a349-1bf0-491c-aebc-9b6f35056bb6 · inbound

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling cites this paper.

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:15:54.355622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T05:13:42.934115Z digest=sha256:0d0627559a57b10da073a6ae5d8fd2c94dc635b98ddcbd4a574816ea80cfbfad

Observation da9d1af9-c0c1-4893-98ca-f018a65e10ee · inbound

VISTA: A Controllable Platform for Generating and Auditing Egocentric Assistance Scenarios cites this paper.

VISTA: A Controllable Platform for Generating and Auditing Egocentric Assistance Scenarios T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T14:24:57.691309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:24:57.691309Z digest=sha256:731f9cd71e8336553012bffb8cc31c44d4ae5e09c9675216ea4d65044ddc8f53

Observation cac5b9e1-84df-4e74-adba-e8b1e2f91ecc · inbound

PhyMotion: Structured 3D Motion Reward for Physics-Grounded Human Video Generation cites this paper.

PhyMotion: Structured 3D Motion Reward for Physics-Grounded Human Video Generation T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:49:41.450400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T02:49:21.291716Z digest=sha256:885033ee2b576a2cde052094fc27f4293207c22b68284dd41f995f393c8dd54e

Observation 1893dce7-7ebd-43a9-ba47-f36ad7896b2b · inbound

Quantitative Video World Model Evaluation for Geometric-Consistency cites this paper.

Quantitative Video World Model Evaluation for Geometric-Consistency T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:14:53.044179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T03:11:03.060052Z digest=sha256:5eccf742448781027f49efa5c7dc9caf1a02a030b7b1c71f525fc6d4c65e6c49

Observation 0248baeb-cf3c-4be0-96fe-2dc3abe2c403 · inbound

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models cites this paper.

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.902525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:46:28.469913Z digest=sha256:a1e689ebf0dc82f2df81ea006b3b9db3bf984c4fd11af6b8cab373c4448415d2

Observation d57fe3b7-a99a-4b5a-86bb-b7540b419bf8 · inbound

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation cites this paper.

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:12:25.192683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T17:05:57.685728Z digest=sha256:fb7aa16bfc9acb340e1dadaa0058b375b2d9cd1857cba91bd24bf8ba55e726f2

Observation 1ae5a96b-ff39-49e1-9527-3163a8c5ce69 · inbound

ResilPhase: Plug-and-Play Phase Mapping and Noise-Resilient Macro-Trajectory Extrapolation for Diffusion Acceleration cites this paper.

ResilPhase: Plug-and-Play Phase Mapping and Noise-Resilient Macro-Trajectory Extrapolation for Diffusion Acceleration T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:39:50.645701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-26T05:03:22.182138Z digest=sha256:eb25a038b5f86019018809f8fb5ab997e2c6d357ee9746199ac6cf6c75b39a30

Observation 5c6e1ef0-5f68-456e-91c7-3e8ea85a550a · inbound

OTCache: Optimal Transport for Geometry-Aware Caching in Diffusion Models cites this paper.

OTCache: Optimal Transport for Geometry-Aware Caching in Diffusion Models T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:45:39.517240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T06:21:22.905093Z digest=sha256:ebb19980dd7e7ec3bdd6c52637cdd9a5d7fc6efc2c559f2f3516bad8eaa860b5

Observation bde72b70-1241-49ce-9aa6-f821b323897d · inbound

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence cites this paper.

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 172

Resolution
unresolved
no resolver link, observed 2026-07-14T03:51:24.547781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T03:51:24.547781Z digest=sha256:b3071169063da7634bac7c629482759d0242fa67f09da7c1789e5c9db56d37fa

Observation afc33395-3294-46e4-92b3-c5ad9d5b2ddc · inbound

KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation cites this paper.

KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T02:56:21.282757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:56:21.282757Z digest=sha256:9e73c3a83f372946b8c7ff98d6ebf6c007ecb7dafbedd225947d10d4ced5b495

Observation 2a76ccff-3e84-469b-ac40-77cb35f8d457 · inbound

Multi-Dimensional Quality Assessment for AI-Generated Human-Centric Videos: Dataset and Model cites this paper.

Multi-Dimensional Quality Assessment for AI-Generated Human-Centric Videos: Dataset and Model T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T20:07:42.500888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:07:42.500888Z digest=sha256:35b4995fa367d3bf024753bf0c3f57fc8d862977717d7f5af94b8aea0023f837