Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2406.19388.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T19:59:27.912849Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T00:04:22.552662Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ca402966-e9f1-484e-8561-a58021840176 · inbound
AudioSetCaps: An Enriched Audio-Caption Dataset using Automated Generation Pipeline with Large Audio and Language Models Taming Data and Transformers for Audio Generation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62c79e91-2061-402c-a6bf-6f4d89e674ee · inbound
HunyuanVideo: A Systematic Framework For Large Video Generative Models Taming Data and Transformers for Audio Generation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1077e4aa-094b-498e-829b-60aa55850084 · inbound
AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation Taming Data and Transformers for Audio Generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a5f7d87-ce14-433a-bbdb-35d30f0e8508 · inbound
MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis Taming Data and Transformers for Audio Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 252aef73-a3a0-4225-925a-957d08ea0c46 · inbound
ETTA: Elucidating the Design Space of Text-to-Audio Models Taming Data and Transformers for Audio Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c6ef354-ca3f-4c70-aeca-c4fac11cd4bd · inbound
LAVCap: LLM-based Audio-Visual Captioning using Optimal Transport Taming Data and Transformers for Audio Generation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d902a54c-7d83-4850-9ab8-e32caae0b953 · inbound
Audio-Language Models for Audio-Centric Tasks: A Systematic Survey Taming Data and Transformers for Audio Generation
Reference 114
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab9921c6-7f27-4813-8fdb-9715031c8ae5 · inbound
CLAP-ART: Automated Audio Captioning with Semantic-rich Audio Representation Tokenizer Taming Data and Transformers for Audio Generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b22a89d-8fc6-43e5-8b3d-9b9d56dbf6b8 · inbound
Ovi: Twin Backbone Cross-Modal Fusion for Audio-Video Generation Taming Data and Transformers for Audio Generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation eb5cee6c-a1db-4e75-8407-12de4b8c8d7c · inbound
Omni2Sound: Towards Unified Video-Text-to-Audio Generation Taming Data and Transformers for Audio Generation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c4e9a791-cfa6-48e6-9005-ff73212cf8d5 · inbound
UNISON: A Unified Sound Generation and Editing Framework via Deep LLM Fusion Taming Data and Transformers for Audio Generation
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation aa35f474-cc64-41c6-bfd8-57ea10a6eed6 · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence Taming Data and Transformers for Audio Generation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0f5bfcfd-e825-45fc-ad6d-8e3cd4f569ec · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence Taming Data and Transformers for Audio Generation
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0098b65d-aebd-41f6-95de-d804a312b0de · inbound
VoxAudio: Vocalized Audio Synthesis via Multi-Reward Autoregressive Flow Matching Taming Data and Transformers for Audio Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.