Pith. sign in

Paper Citation Record · LEDGER

Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2305.10874.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.10874 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:36:48.967450Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T11:34:37.702024Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 532c46f3-6ed1-43f3-8103-8712c7bd60d1 · inbound

VideoPhy: Evaluating Physical Commonsense for Video Generation cites this paper.

VideoPhy: Evaluating Physical Commonsense for Video Generation Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 104

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:34:37.704069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T11:34:37.599691Z digest=sha256:5f0fa84c92c2e1ae8d5a3e3edd33467fade23a82fa4ae17b46cb586dd12cc6e3

Observation 0c99a30f-fd9e-415b-ab35-3235bac5faa3 · inbound

UNICA: A Unified Neural Framework for Controllable 3D Avatars cites this paper.

UNICA: A Unified Neural Framework for Controllable 3D Avatars Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 67

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T19:48:11.255188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T19:47:13.052204Z digest=sha256:f4d5d54fb3827de804e61cf8a14f0aa3a212d5e47b06cfc2c7d3c83b209e0922

Observation 4a9695ad-c41e-40c6-b1a4-01a26f35a35b · inbound

Detecting AI-Generated Videos with Spiking Neural Networks cites this paper.

Detecting AI-Generated Videos with Spiking Neural Networks Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:31:12.328250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T16:15:42.113951Z digest=sha256:c98806a040a1b4a4e6061214f7fba3d0984c5be7be4d9c2e9e9979f8c8778269

Observation 824d2951-fae7-4812-972e-8320785a452d · inbound

Detecting AI-Generated Videos with Spiking Neural Networks cites this paper.

Detecting AI-Generated Videos with Spiking Neural Networks Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-03T02:25:01.702234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:25:01.702234Z digest=sha256:f2905a09c17081b10024207e6843849cfbe7083aded6ef30a2305e341df39c84

Observation 1f389f1e-0f8a-4008-899c-9103a5bf502d · inbound

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation cites this paper.

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 207

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:48.054091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:23:48.054091Z digest=sha256:ec08163397f040cb798af0786e80b9d10203d2c8f24a284d797e6cd347355166

Observation e7c544ed-a4d6-4341-b448-2587a11b1bdc · inbound

Retrieval-Driven Training-Free AI-Generated Video Attribution cites this paper.

Retrieval-Driven Training-Free AI-Generated Video Attribution Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T16:36:48.967450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:36:48.967450Z digest=sha256:cde492bedbf519d8682359f4df50138f8d16fed327e94cbd2131f46d52c95969

Observation 8c187cd9-4c89-44ba-8d5c-d7695cb93257 · inbound

RAID: Towards Robust AI-Generated Image Detection with Bit-Reversed Images cites this paper.

RAID: Towards Robust AI-Generated Image Detection with Bit-Reversed Images Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T16:20:04.584508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:20:04.584508Z digest=sha256:a84e7d701a58aeadaa02e3591df672336972811d9b5abfbca918fabc3c729dce