Pith. sign in

Paper Citation Record · LEDGER

TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2406.08656.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.08656 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:46.458920Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:49:31.003751Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 24e2b302-5f59-4efa-9066-1467aaa50ebf · inbound

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback cites this paper.

We'll Fix it in Post: Improving Text-to-Video Generation with Neuro-Symbolic Feedback TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:56:58.222912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T18:56:39.735334Z digest=sha256:f96f3dfb4e39da62dfe72796707c20645c7802065e431ec3e3395abfb51d0a28

Observation f9d3a925-7cf2-41e0-8c54-ce28d147afb5 · inbound

ARGUS: Hallucination and Omission Evaluation in Video-LLMs cites this paper.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.629628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.629628Z digest=sha256:202d5a72526255ab6ad46fd5b4312efee5bd5dc1f17e4cacd0b5160ba68c05b1

Observation ab23c3cf-1d1e-496d-9759-c525e4513816 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:46.458920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:46.458920Z digest=sha256:a7be72b1fedf192be80ef246ccc7e3cf7bc1bfc6795e9d8a6ccb0a2338ba0661

Observation f2308e59-9a08-42ce-a73a-64f2db7c1330 · inbound

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation cites this paper.

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:26.566236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:26.566236Z digest=sha256:30802dc1040ce8339278658bb9313614b549ccdb11e3384d4c1c61426d0a755f

Observation 070e9f70-c0f7-4fc6-813a-ad34bd17ea0e · inbound

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding cites this paper.

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T13:01:58.045344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:01:58.045344Z digest=sha256:7fff5a587b6430ecb825d4905622c1361cf40e89a91bd5c5f2e43dabd267c8ba

Observation f7706508-171a-4b27-b1a5-33634d400074 · inbound

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding cites this paper.

NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T06:36:33.811054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:36:33.811054Z digest=sha256:ee75ef2c86fb021c075858254c8e6f85f12d46356dda2a32c4fee78cdede6ea6

Observation eed2fe64-e570-4751-8fc1-4a7c37f37f1a · inbound

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation cites this paper.

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:03:37.455762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T03:02:26.084185Z digest=sha256:634601b15c57266564429631a43da6fa1a8fdf55ac80e1a657bbe5b6a6d8a266

Observation 319912a3-7652-4639-ac05-94c56e90beeb · inbound

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation cites this paper.

RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:19:50.198674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T06:15:32.881140Z digest=sha256:51482ad1e5b9f6438348108877700fffdfb96e74992fc68cc958953b8f30cba8

Observation d9c6c93f-4d63-40ca-a3a5-deb230cecd16 · inbound

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models cites this paper.

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:40:23.130772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T04:39:22.400458Z digest=sha256:9d3248397dee36acfe595818d400b6101ac72a413e52560425fc909cd19c5a1f

Observation 42c44632-6e7c-4385-a01e-cb5d6b8075c6 · inbound

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV cites this paper.

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:00.793852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T22:52:38.330851Z digest=sha256:af42f9ef52a05d0d3c087ea8536c77a72dceb0d6972a72cce443fe91b2c298b0

Observation b2015940-614c-4655-a5ad-580580669802 · inbound

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends cites this paper.

Towards Interactive Video World Modeling: Frontiers, Challenges, Benchmarks, and Future Trends TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 134

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:06:14.094042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T17:29:18.513507Z digest=sha256:d7f4a25d16a44cce2c2a8db49ecbce640ceb950390773e17b9965ae1c498e78f

Observation 01349f55-1a4f-44c8-86dd-85674ac5176b · inbound

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency cites this paper.

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:57:38.101059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T13:34:25.079037Z digest=sha256:cc391e23d439154ea2e5a11bb5bfd5fb461ce4c2546fd19a673c95e422bb8cfe

Observation 1287a9a5-e700-4982-99b2-cbe2f21ee39a · inbound

Current World Models Lack a Persistent State Core cites this paper.

Current World Models Lack a Persistent State Core TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:31.006734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:33:41.461245Z digest=sha256:58e06030f3ef8b3bbae7b6e95a1b84b042ff8b81b94fbd3bfc5ed6be502aa4f4

Observation e92bca83-aa2c-4ec9-8b0f-80b4951b5ec4 · inbound

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence cites this paper.

From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 115

Resolution
unresolved
no resolver link, observed 2026-07-14T03:51:24.547781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T03:51:24.547781Z digest=sha256:5474274e4637d6cee4a1d711edd62dafbe9eec92c37e951a8d96d23d861ec1fc

Observation 763678e9-f1d8-4bc6-81bc-4e197037d136 · inbound

VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation cites this paper.

VGIF-Score: Interpretable and Diagnostic Evaluation of Spatio-Temporal Instruction Following in Video Generation TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T04:57:34.021666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:57:34.021666Z digest=sha256:07aa7ea885fbaf22ec5eb442cd9de4a506147ee0b69a4b90566f024217420d27