Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2502.03897.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:07:28.480478Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T05:56:40.900641Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a65d3b03-373a-4102-baa0-0f2224156e3b · inbound
JWB-DH-V1: Benchmark for Joint Whole-Body Talking Avatar and Speech Generation Version 1 UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e3600b3-5b03-43bb-aa81-f95b258b9571 · inbound
UniVerse-1: Unified Audio-Video Generation via Stitching of Experts UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe916510-e314-4523-bf4a-49e286c8d953 · inbound
Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86a4669c-fe76-4651-b85e-aec31796905f · inbound
PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 750fb1d0-85b6-4f75-8733-d431af5e8148 · inbound
PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 195a3271-7eee-405d-a737-f4f90e7b23f6 · inbound
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36151287-89fb-4b24-b8e5-a1965a16360e · inbound
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5472edbe-229b-4705-bada-76c3b06531ac · inbound
SyncDPO: Enhancing Temporal Synchronization in Video-Audio Joint Generation via Preference Learning UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 432bf714-ac4c-41fc-8b02-6979188d94b2 · inbound
MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38385add-b7cf-4a4e-ad03-983f4a143131 · inbound
MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e276bec-38a2-4b47-87aa-98b4275b5842 · inbound
Inference-Time Scaling for Joint Audio-Video Generation UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.