Pith. sign in

Paper Citation Record · LEDGER

TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2406.08656.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.08656 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:46.458920Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:49:31.003751Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 24e2b302-5f59-4efa-9066-1467aaa50ebf · inbound

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback cites this paper.

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:56:58.222912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T18:56:39.735334Z digest=sha256:60853415fc6dfcfae47e7ff9566f67054fbe954befb3957b7a03f4538589417a

Observation f9d3a925-7cf2-41e0-8c54-ce28d147afb5 · inbound

ARGUS: Hallucination and Omission Evaluation in Video-LLMs cites this paper.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.629628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.629628Z digest=sha256:ff062e3ca98a07936817d252803ac1ada70c03e5414e7b063e0146088be5dd2b

Observation ab23c3cf-1d1e-496d-9759-c525e4513816 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:46.458920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:46.458920Z digest=sha256:4c0634fbc2fae1149befa0839b192924ab5f17d5df41f03f6d16ea55266a0d28

Observation f2308e59-9a08-42ce-a73a-64f2db7c1330 · inbound

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation cites this paper.

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:26.566236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:26.566236Z digest=sha256:002b0cea18db6da13a52a4e22c3c4501b3308dc92f468b5856a2f837376aca30

Observation 070e9f70-c0f7-4fc6-813a-ad34bd17ea0e · inbound

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding cites this paper.

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T13:01:58.045344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:01:58.045344Z digest=sha256:92a076025c4faffc16bc99612ec94c2c46f00ae0a14abdb7425df44907479482

Observation f7706508-171a-4b27-b1a5-33634d400074 · inbound

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding cites this paper.

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T06:36:33.811054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:36:33.811054Z digest=sha256:9a8578bd60ef24403e9d0c3457b9ef82a25ee39f16930850c146fcf16c1f7474

Observation eed2fe64-e570-4751-8fc1-4a7c37f37f1a · inbound

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation cites this paper.

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:03:37.455762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T03:02:26.084185Z digest=sha256:6c083d8d55c371dea4dc30660a6d875cdca22a0a7d2df6771ca1a23ea2bc0bc2

Observation 319912a3-7652-4639-ac05-94c56e90beeb · inbound

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation cites this paper.

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:19:50.198674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T06:15:32.881140Z digest=sha256:733f491af4a329f188091032b58af09c41dc9ad81c74080115ee02189a9f823a

Observation d9c6c93f-4d63-40ca-a3a5-deb230cecd16 · inbound

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models cites this paper.

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:40:23.130772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T04:39:22.400458Z digest=sha256:d7022caf64287a934d128848881b0696c0bb8edd16d4193fa436af15c6ec4c63

Observation 42c44632-6e7c-4385-a01e-cb5d6b8075c6 · inbound

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV cites this paper.

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:00.793852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T22:52:38.330851Z digest=sha256:6932cf3159659b7013fce4dd532e3bfdb97087318cfc7797a08c5dc8250b5d7c

Observation b2015940-614c-4655-a5ad-580580669802 · inbound

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends cites this paper.

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 134

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:06:14.094042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T17:29:18.513507Z digest=sha256:2318fab3c2d97ed42ce02f47f089c289ed8cb893f938c71ceca2aa6794ed5b50

Observation 01349f55-1a4f-44c8-86dd-85674ac5176b · inbound

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency cites this paper.

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:57:38.101059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T13:34:25.079037Z digest=sha256:e51e066b9191ed06c4ecd459160f881540d4650c1645bf496f4b7ebb3702334a

Observation 1287a9a5-e700-4982-99b2-cbe2f21ee39a · inbound

Current World Models Lack a Persistent State Core cites this paper.

Current World Models Lack a Persistent State Core TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:31.006734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T17:33:41.461245Z digest=sha256:b4fc8b908669baa268573c186ffaf6800112359208919800f672326325ff836c

Observation e92bca83-aa2c-4ec9-8b0f-80b4951b5ec4 · inbound

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence cites this paper.

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 115

Resolution
unresolved
no resolver link, observed 2026-07-14T03:51:24.547781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T03:51:24.547781Z digest=sha256:71e028d118f61d747f1e958cd7afc5d2da76ae1293231af0962478f5dc1e8f6c

Observation 763678e9-f1d8-4bc6-81bc-4e197037d136 · inbound

VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation cites this paper.

VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T04:57:34.021666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:57:34.021666Z digest=sha256:a461a1cc89f51fc684401386991b80abff0e21349202dbfcfdb11d036d5a3f09