Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-01T03:45:40.636290Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2606.31259.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-01T03:45:40.636290Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
51 of 51 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5ea8592d-e260-4e9d-846d-c89534dc9c97 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Make-an-audio: Text-to-audio generation with prompt- enhanced diffusion models,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e832f22f-5a78-44ff-b0b8-063574399ea2 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Audiogen: Textually guided audio generation,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4a620aa2-6883-4dea-aee5-5764793845de · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation AudioLDM: Text-to-audio generation with latent diffusion models,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5b884f80-2ba9-48b0-abf5-378989279e80 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Audioldm 2: Learning holistic audio generation with self-supervised pretraining
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 414dd74d-f4dd-452a-8876-22be12576b5d · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Any-to-any generation via composable diffusion,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 49ad1c5a-e797-4ad2-92b8-5d4e199f14f4 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Auffusion: Leveraging the power of diffusion and large language models for text-to-audio generation,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 44652a9a-ebf0-4e02-a6cd-f58d8276cf61 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 28a7277b-f8bf-4cfc-8ef4-3f017db3587e · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Denoising diffusion probabilistic models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4d8ee700-e503-4295-9a56-83ddbf834e01 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Denoising diffusion implicit models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ea950e4a-e0df-44d3-b54e-605c408b205c · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e81d83d3-b117-46d2-8dc2-ed5211fa980c · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Dpm-solver++: Fast solver for guided sampling of diffusion probabilistic models,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1c8c73f1-8741-4dfb-8095-ac27e46cd834 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Elucidating the design space of diffusion-based generative models,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b85ed6f7-fc2e-45d7-b3a3-a855b02a826e · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Analytic-dpm: an analytic estimate of the optimal reverse variance in diffusion probabilistic models,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8a6807da-9f65-4cc3-b049-c459aae6b750 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Consistency models,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 100b490a-db43-47b8-a455-8d17d3293fc5 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Consistencytta: Accelerating diffusion-based text-to-audio generation with consistency distillation,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7e86ca38-588f-42c1-b013-dbf301d508dd · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Audiolcm: Efficient and high-quality text-to-audio generation with minimal inference steps,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 29755e60-187a-4df3-8d9b-4a82e9429a5b · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Instructpix2pix: Learning to follow image editing instructions
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0ce01166-648a-45b2-9f0d-a3d5aa2231c5 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Ts2f: Text-assisted speech-to-face generation,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 147d604e-9d83-49ac-b59a-af2cdb335604 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation AudioCaps: Generating captions for audios in the wild,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4c4083aa-ab37-4cd7-9df1-2c2dc7bbe03d · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation WavCaps: A ChatGPT-assisted weakly-labelled audio captioning dataset for audio-language multimodal research,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d91e1fe3-7281-4a62-bf7f-fdb63386e5b0 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Audiosetcaps: Enriched audio captioning dataset generation using large audio language models,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation adc3b1f5-3410-4d24-9607-0b7b30dd42f7 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Journeydb: A benchmark for generative image understanding,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 94db4798-f07f-480b-98ea-bb6f2d572426 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Laion- 5b: An open large-scale dataset for training next generation image-text models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 026bd4cc-cd01-408d-a812-8b1271a30ada · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Prolific- dreamer: High-fidelity and diverse text-to-3d generation with variational score distillation,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a234b00f-cd1a-41e5-87d0-0ac3ff12296f · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Swiftbrush: One-step text-to-image diffu- sion model with variational score distillation,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation aaa167fc-e9b2-4255-bd6e-f0f614ed01c2 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Swiftbrush v2: Make your one-step diffusion model better than its teacher,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0c131e16-0f6b-41d6-b6c7-5afd49369a9a · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Clotho: An audio captioning dataset,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bc36acda-8e77-44d1-a3f6-0dcf8547c4e7 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Audiolm: A lan- guage modeling approach to audio generation,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1d1c1db6-e338-40be-8b43-b8c7fdd30f06 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Diffsound: Discrete diffusion model for text-to-sound generation,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3ab76f2b-28f4-4b3c-9d41-7c62a4abd398 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Improved techniques for training score- based generative models,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ce9a2467-275f-41c7-8a8c-ceb53551769b · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Text-to-audio gen- eration using instruction guided latent diffusion model,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation da459659-1138-4f96-9d95-2aca7d617869 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Diffwave: A versatile diffusion model for audio synthesis,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 789d25a4-d230-4375-8818-5a5c360ce065 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Wave- grad: Estimating gradients for waveform generation,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 87601a5d-2d35-48db-ab41-d9713bb0306b · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation High- resolution image synthesis with latent diffusion models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation be6dbe49-c11d-4e59-b529-b57578e5adfd · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Pseudo numerical methods for diffusion models on manifolds
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a0c3e79c-78bb-4cf8-9702-3bcc8732884e · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Progressive distillation for fast sampling of diffusion models,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 223603cc-ab08-4380-889e-67d9ff11200d · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Dreamfusion: Text-to-3d using 2d diffusion,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0d4044a5-481a-4642-a0e2-76a719ae5961 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Score jacobian chaining: Lifting pretrained 2d diffusion models for 3d generation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bcfc807b-7570-4dd2-ae27-26f39dc238c4 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Magic3d: High-resolution text-to- 3d content creation,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 77f13f2e-3f87-4f52-82fe-41b860d7d387 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Latent-nerf for shape-guided generation of 3d shapes and textures,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation dd4afe92-b547-4ade-b482-e9efa7288a82 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Fantasia3d: Disentangling geometry and appearance for high-quality text-to-3d content creation,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 99fd5030-4c2e-4868-8fcb-2642dc5ad66b · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation LoRA: Low-rank adaptation of large language models,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a4f16c29-6ad6-4037-9475-9f61d0ec565a · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Nonlinear total variation based noise removal algorithms
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 96f26093-8399-48bd-8371-45709339df8a · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation A duality based approach for realtime tv-l 1 optical flow
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0203c610-71f9-4a7f-a7ed-88409025fc71 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Classifier-free diffusion guidance
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8fffec2a-fe23-46f6-8929-0262afc4f7b1 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Decoupled weight decay regularization,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 01c32da1-7a78-4783-a7e8-5acf733b18d1 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Panns: Large-scale pretrained audio neural networks for audio pattern recognition
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f7563967-269c-47d5-8f3e-81c7b10add25 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Cnn architectures for large-scale audio classification,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e1f9f09c-2b45-46e7-8b2e-877fe5b466fa · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c9f9d3cf-b0fb-4d07-9553-ba8f8c833a22 · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Supercharged one-step text-to-image diffusion models with negative prompts,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a9eb15af-54cc-46b1-bf8e-393e7099ffdf · outbound
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation Prompt-to-prompt image editing with cross-attention control,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
No inbound Pith citation observations are available.