Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 61 inbound Pith citation observations for arXiv:2301.12503.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:58:34.577052Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T11:59:50.469153Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 4244a785-d3d1-4482-8f44-0db941ed88c7 · inbound
DGSNA: Dynamic Generative Scene-based Noise Addition method AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 833355d6-6bc7-45e7-b768-abac05b45178 · inbound
How Far Are We from Generating Missing Modalities with Foundation Models? AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 401ae98b-4024-41b3-ad2e-c5b3acd1d977 · inbound
Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3384e3c5-011a-4867-b15e-0451f2ba050a · inbound
Hear-Your-Click: Interactive Object-Specific Video-to-Audio Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8afa710d-257f-449e-a116-60b3e028de26 · inbound
ADMC: Attention-based Diffusion Model for Missing Modalities Feature Completion AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1a90ca5-8730-466f-96ed-44324b7d35cd · inbound
Diffusion Models for Time Series Forecasting: A Survey AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9139e409-ca8c-43d2-9ce7-0a57256d1a80 · inbound
CHORDS: Diffusion Sampling Accelerator with Multi-core Hierarchical ODE Solvers AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 035f79e7-9d79-4d62-83f6-e690c2962950 · inbound
SonicGauss: Position-Aware Physical Sound Synthesis for 3D Gaussian Representations AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 511d80a0-aa90-46af-93dc-954e0950af30 · inbound
Flow Matching Policy Gradients AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d212bfcf-8c24-4765-b564-a318947e4b00 · inbound
Aether Weaver: Multimodal Affective Narrative Co-Generation with Dynamic Scene Graphs AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eae0bf53-89bc-4d9b-adb5-4745a5dce62d · inbound
Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d75771a-4415-4798-b175-2d463d71af6d · inbound
AudioGen-Omni: A Unified Multimodal Diffusion Transformer for Video-Synchronized Audio, Speech, and Song Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64148a22-c14a-43ff-a71a-acb113314683 · inbound
Via Score to Performance: Efficient Human-Controllable Long Song Generation with Bar-Level Symbolic Notation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0fee5c5-370f-447c-94d1-65406a061189 · inbound
Inference-time Scaling for Diffusion-based Audio Super-resolution AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 986d2272-7240-4034-913a-4c095f3bd0ca · inbound
ASAudio: A Survey of Advanced Spatial Audio Research AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b40285ca-fcc7-4c76-86fb-d4cc7f16590e · inbound
A Sharp KL-Convergence Analysis for Diffusion Models under Minimal Assumptions AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0ef9092-7b6a-45a5-a957-900a8070585c · inbound
Audio-Guided Visual Editing with Complex Multi-Modal Prompts AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01204637-e03e-46bd-86aa-ef69bf55b79c · inbound
WaveLLDM: Design and Development of a Lightweight Latent Diffusion Model for Speech Enhancement and Restoration AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29112a42-28a7-48ca-a089-aee4158f9105 · inbound
AudioMoG: Guiding Audio Generation with Mixture-of-Guidance AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cf197551-189e-4413-bfbb-75800a2a5d56 · inbound
Taming Text-to-Sounding Video Generation via Advanced Modality Condition and Interaction AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba08cfd-6e72-4426-9c96-74149d070b56 · inbound
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cab15e9-428c-48d5-8f1c-1590be20ebf8 · inbound
Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7f33acef-db98-43a1-926c-d821ecfb8475 · inbound
Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 34d78cc4-a91d-4143-a1b0-bab32f6be7fa · inbound
Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad9ddeab-612d-4c4d-806d-6591e5de0714 · inbound
JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ad26e427-935b-4ab3-bf78-e31e4327eb8f · inbound
Dual-End Consistency Model AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a5fe170f-ff47-4c70-9b9d-e78aba20a207 · inbound
Dual-End Consistency Model AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8e4dd18-5751-40fe-bcf0-56855f426025 · inbound
Diffusion Models Memorize in Training -- and Generalize in Inference AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6f6fbe12-1056-481f-a394-17d714e80274 · inbound
Making Separation-First Multi-Stream Audio Watermarking Feasible via Joint Training AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d8d74b6-99e3-412b-a589-d61ee8b439db · inbound
Woosh: A Sound Effects Foundation Model AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 01351671-3543-4f01-a0f1-0ac538ec27e2 · inbound
FoleyDesigner: Immersive Stereo Foley Generation with Precise Spatio-Temporal Alignment for Film Clips AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7a567944-8d5b-41e9-8911-b1b29291eb45 · inbound
AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 29ea7c1a-3e59-4755-a1a7-e76c6cf6c59b · inbound
Latent Fourier Transform AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 01797019-b68d-4ae6-ac3f-1015173d9bd7 · inbound
ATRIE: Adaptive Tuning for Robust Inference and Emotion in Persona-Driven Speech Synthesis AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 39d88a97-f0b0-48b8-9ac7-7861846c0866 · inbound
Fast Text-to-Audio Generation with One-Step Sampling via Energy-Scoring and Auxiliary Contextual Representation Distillation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2d7908dd-aed0-4ad3-87ba-adc684434783 · inbound
Stage-adaptive audio diffusion modeling AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9213e690-cbd2-494a-b509-1434c2bb22a9 · inbound
Remix the Timbre: Diffusion-Based Style Transfer Across Polyphonic Stems AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a8c7fa61-6d59-4926-9f7a-96b99a4fd859 · inbound
DiffATS: Diffusion in Aligned Tensor Space AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a5609b5b-8f1c-4de5-8d38-02ceb2028256 · inbound
HapticLDM: A Diffusion Model for Text-to-Vibrotactile Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e51d76d-72e3-48e3-9143-9a44fa42d66a · inbound
PoDAR: Power-Disentangled Audio Representation for Generative Modeling AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 875fd0db-094b-4cac-a577-9f4e223b7d81 · inbound
WavFlow: Audio Generation in Waveform Space AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9e9be8ae-4fc9-4935-9355-e05862a46376 · inbound
EigeNet: Geometry-Informed Multi-Modal Learning for Few-shot Novel View RIR Prediction AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 917ed80e-e9b2-4838-8169-7486d16d7403 · inbound
Mental Damage: Caption Poisoning Attacks on Retrieval-Augmented Text-to-Music Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d8b7e1ac-56f6-48d5-9137-35997613f8c2 · inbound
dMoE: dLLMs with Learnable Block Experts AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b957d6e4-935a-412e-a3f5-2f0b3e30d1e4 · inbound
Entropy as a Structural Prior: How a Log-Barrier on DiT Belief Space Drives Musical Diversity and Development AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2987213b-e846-46ff-9bad-cb73d2f87e50 · inbound
Flow Matching with In-Context Priors for Out-of-Distribution Brain Dynamics AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b2be330d-b942-43c6-81c4-adebd15d638f · inbound
Net-Ev$^2$: A Generative Simulator for Network Event Evolution AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6553f382-ad31-49dd-8943-a4627ad387a8 · inbound
STAR-VAE: Structured Topology-Aware Regularization for Audio Reconstruction and Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 201afcc3-58ad-49f4-b226-f6f680c2fc2d · inbound
ARIA: Adaptive Region-Based Importance Allocation for Conditional Diffusion Distillation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1f58b6b3-7635-4d6a-b9b4-127dc290554b · inbound
MAVIN: Multi-Shot Audio-Visual Generation with Customized Narrative Control AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d870af0f-cf64-45bc-add2-1507338a977b · inbound
MAVIN: Multi-Shot Audio-Visual Generation with Customized Narrative Control AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09db36e7-d0b1-4508-9387-bb85347cb181 · inbound
ALM2Vec: Learning Audio Embeddings for Universal Audio Retrieval with Large Audio-Language Models AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eda6be9b-5c73-4779-a2c7-3b4ba7304da7 · inbound
An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6b618a80-5245-450a-b915-ba7b6712dcc8 · inbound
SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46dabe77-610a-4901-8c3c-bfea8c1bbebd · inbound
Dance to Music Generation leveraging Pre-training with Unpaired data and Contrastive Alignment AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0eb786b-3987-46b9-8017-92c2c3547882 · inbound
FlashDiff: Efficient Regional Execution and Scheduling for Diffusion Model Serving AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 961127b7-5d90-47d2-a72e-207eddba16ef · inbound
Analytic Distribution of Classifier-Free Guidance for Schedule Design AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09755735-db25-461a-a381-ef03fee5233a · inbound
RPPNet: Perceptually-Grouped Rhythm-Pitch Primitives for Long-Term Structure Melody Generation via Boundary-Aware Modeling AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bbb7541-3030-47ac-9cc4-64cb3ef81e9f · inbound
Amortized Moment Matching for Visual Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e6509e6-dac5-48c8-9391-0c7876ebc211 · inbound
Exploring Efficient Waveform Diffusion Models for Foley Sound Generation AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c16e57b-094a-4a42-a1dd-9b50a8f604f0 · inbound
AI-Based Sound Effect Generation: A Narrative Review of Generative Models Across Input Modalities AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.