Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 70 inbound Pith citation observations for arXiv:2209.15352.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:27:56.712678Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
54
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation efeb8960-e2a4-4923-a911-adfcd1d04b37 · inbound
Shap-E: Generating Conditional 3D Implicit Functions AudioGen: Textually Guided Audio Generation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ead98b7d-9584-4ad0-a34c-abdfe64ff9ed · inbound
AudioPaLM: A Large Language Model That Can Speak and Listen AudioGen: Textually Guided Audio Generation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9a1c68fa-3bca-4550-bce5-b4e39b77681c · inbound
AudioPaLM: A Large Language Model That Can Speak and Listen AudioGen: Textually Guided Audio Generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 32b56a7a-2b0d-42f3-9330-44fc7b8f3be1 · inbound
Movie Gen: A Cast of Media Foundation Models AudioGen: Textually Guided Audio Generation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 369c61ec-e708-4499-9a1c-23c07b26287f · inbound
DGSNA: Dynamic Generative Scene-based Noise Addition method AudioGen: Textually Guided Audio Generation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 04286985-c320-4d5c-b573-024e5a7f7941 · inbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows AudioGen: Textually Guided Audio Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdbf71f0-986e-453b-b4e8-fc928c46b627 · inbound
YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls AudioGen: Textually Guided Audio Generation
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a677f8b-f682-4499-a7d2-9425e63bd858 · inbound
SILA: Signal-to-Language Augmentation for Enhanced Control in Text-to-Audio Generation AudioGen: Textually Guided Audio Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b0328f-621c-491d-a3ed-870cadf9102b · inbound
VinTAGe: Joint Video and Text Conditioning for Holistic Audio Generation AudioGen: Textually Guided Audio Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52c9a419-3ef7-4f45-a1da-78829d2f1d5c · inbound
Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey AudioGen: Textually Guided Audio Generation
Reference 207
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40aa2056-1b35-4ee2-bf0d-1365d5d79202 · inbound
ETTA: Elucidating the Design Space of Text-to-Audio Models AudioGen: Textually Guided Audio Generation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 648dee04-b661-4c3c-8687-9cad3c27db25 · inbound
Diffusion Generative Modeling for Spatially Resolved Gene Expression Inference from Histology Images AudioGen: Textually Guided Audio Generation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0423f0dc-9b3c-48f9-9046-375a6fed13ca · inbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment AudioGen: Textually Guided Audio Generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f174bf9-4794-4b90-9a4b-16154374c1d1 · inbound
Latent Swap Joint Diffusion for 2D Long-Form Latent Generation AudioGen: Textually Guided Audio Generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea6fdf99-c86f-4979-9a02-e9f0147347e9 · inbound
Watermarking across Modalities for Content Tracing and Generative AI AudioGen: Textually Guided Audio Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a95b9ff-de8a-497d-8d58-def212839c4e · inbound
PAST: Phonetic-Acoustic Speech Tokenizer AudioGen: Textually Guided Audio Generation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abc09efc-c13d-4be5-8260-8f0b94acdb03 · inbound
SpecMaskFoley: Steering Pretrained Spectral Masked Generative Transformer Toward Synchronized Video-to-audio Synthesis via ControlNet AudioGen: Textually Guided Audio Generation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebbb6bb1-5a6c-4a0b-8b3c-05f49645f4d4 · inbound
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation AudioGen: Textually Guided Audio Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd6f89eb-99a1-4a4f-b2c3-2541d03efd3c · inbound
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec AudioGen: Textually Guided Audio Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0beca6d6-5409-4aff-8881-f8ef0005ec92 · inbound
RPRA-ADD: Forgery Trace Enhancement-Driven Audio Deepfake Detection AudioGen: Textually Guided Audio Generation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e14ef2e-2a4b-44a6-abc1-7ad47759d318 · inbound
Uncertainty-o: One Model-agnostic Framework for Unveiling Uncertainty in Large Multimodal Models AudioGen: Textually Guided Audio Generation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd1ab871-4070-40d1-bc54-9ce02b7be156 · inbound
UmbraTTS: Adapting Text-to-Speech to Environmental Contexts with Flow Matching AudioGen: Textually Guided Audio Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62f42dd3-c0c1-4d0c-9e21-5dd102e3b288 · inbound
Scaling Self-Supervised Representation Learning for Symbolic Piano Performance AudioGen: Textually Guided Audio Generation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22a2f17d-6983-401a-9a7e-2026ed796308 · inbound
AudioBERTScore: Objective Evaluation of Environmental Sound Synthesis Based on Similarity of Audio embedding Sequences AudioGen: Textually Guided Audio Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd74fb05-b515-4cb9-b2c6-78d2cfbc1ee7 · inbound
Improving GANs by leveraging the quantum noise from real hardware AudioGen: Textually Guided Audio Generation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb3f769f-f12c-4815-b75c-438ab21490ca · inbound
MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation AudioGen: Textually Guided Audio Generation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fec46f35-ef4d-44e9-9579-51fb1c6d56bd · inbound
SonicGauss: Position-Aware Physical Sound Synthesis for 3D Gaussian Representations AudioGen: Textually Guided Audio Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30c34803-a6cb-4e46-873f-06920348492f · inbound
AudioGen-Omni: A Unified Multimodal Diffusion Transformer for Video-Synchronized Audio, Speech, and Song Generation AudioGen: Textually Guided Audio Generation
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d21322c-37df-411b-af3b-985273ffd8b9 · inbound
ASAudio: A Survey of Advanced Spatial Audio Research AudioGen: Textually Guided Audio Generation
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4353d9b0-39b4-408c-8ec1-a990bf116fc7 · inbound
Text2Weight: Bridging Natural Language and Neural Network Weight Spaces AudioGen: Textually Guided Audio Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9a47c8d-e2c4-48d8-97b0-bf3fb57334d4 · inbound
Conflicting Scores, Confusing Signals: An Empirical Study of Vulnerability Scoring Systems AudioGen: Textually Guided Audio Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 699117b6-2f64-4149-9fd1-165eaaf6c739 · inbound
CompLex: Music Theory Lexicon Constructed by Autonomous Agents for Automatic Music Generation AudioGen: Textually Guided Audio Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1260ad0-06c3-4bd2-87ee-c4e0c0a3b362 · inbound
Effectively obtaining acoustic, visual and textual data from videos AudioGen: Textually Guided Audio Generation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4ad2945-d372-4a1a-bec9-ff1e605254a2 · inbound
Segment Transformer: AI-Generated Music Detection via Music Structural Analysis AudioGen: Textually Guided Audio Generation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2ed3cc5-127e-42e8-a0e2-220132449bde · inbound
Testing chatbots on the creation of encoders for audio conditioned image generation AudioGen: Textually Guided Audio Generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdc799a6-0e8a-482a-ad42-19bc116d3940 · inbound
AudioMoG: Guiding Audio Generation with Mixture-of-Guidance AudioGen: Textually Guided Audio Generation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8fba2990-c8a4-471d-be22-abe733918f0e · inbound
Quantifying the Climate Risk of Generative AI: Region-Aware Carbon Accounting with G-TRACE and the AI Sustainability Pyramid AudioGen: Textually Guided Audio Generation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ab5f087a-00b7-403a-a110-52256ddb52ad · inbound
Quantifying the Climate Risk of Generative AI: Region-Aware Carbon Accounting with G-TRACE and the AI Sustainability Pyramid AudioGen: Textually Guided Audio Generation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 434f0367-9f12-4872-b9ea-5c08903e7b22 · inbound
Two-Dimensional Quantization for Geometry-Aware Audio Coding AudioGen: Textually Guided Audio Generation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e3eb7a0d-662c-4de7-b6b1-ad3bb442c4fb · inbound
Language Models as Semantic Teachers: Post-Training Alignment for Medical Audio Understanding AudioGen: Textually Guided Audio Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5bc56e09-265e-4bf7-a799-5dc6ab6f22e4 · inbound
Omni2Sound: Towards Unified Video-Text-to-Audio Generation AudioGen: Textually Guided Audio Generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b6bc8e9f-6678-40cd-be8f-813d203bcdbe · inbound
SemanticAudio: Audio Generation and Editing in Semantic Space AudioGen: Textually Guided Audio Generation
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 061b5dad-dfc2-42d9-bc6e-291fe3375a92 · inbound
FoleyDirector: Fine-Grained Temporal Steering for Video-to-Audio Generation via Structured Scripts AudioGen: Textually Guided Audio Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 654bf81e-37a0-47ea-8326-c82201367047 · inbound
Woosh: A Sound Effects Foundation Model AudioGen: Textually Guided Audio Generation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 97460d33-4d3c-4005-8e54-fe74b6302896 · inbound
FoleyDesigner: Immersive Stereo Foley Generation with Precise Spatio-Temporal Alignment for Film Clips AudioGen: Textually Guided Audio Generation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9a480914-74fe-46ee-a9de-7aff9550d9dc · inbound
Language-Guided Multimodal Texture Authoring via Generative Models AudioGen: Textually Guided Audio Generation
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 15c4c702-8f2f-457e-aed8-38a4f0da5e68 · inbound
Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing AudioGen: Textually Guided Audio Generation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation fa340242-55c2-4653-9b71-0126a681a4cb · inbound
Stage-adaptive audio diffusion modeling AudioGen: Textually Guided Audio Generation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6a672a1a-68e4-4080-ac5a-2e8f03a5d69c · inbound
PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization AudioGen: Textually Guided Audio Generation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c41a3266-46f2-4bac-b4ee-d0f84219272b · inbound
PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization AudioGen: Textually Guided Audio Generation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6587fdd1-bbcd-4cfe-8f7d-8d60d6f629d7 · inbound
HapticLDM: A Diffusion Model for Text-to-Vibrotactile Generation AudioGen: Textually Guided Audio Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 311d23eb-a810-499c-a0a7-241c32dba0f2 · inbound
WavFlow: Audio Generation in Waveform Space AudioGen: Textually Guided Audio Generation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0e1f002f-0aad-40d4-8a01-d0b593e31204 · inbound
From Prompts to Context: An Ontology-Driven Framework for Human-Generative AI Collaboration AudioGen: Textually Guided Audio Generation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f989c01c-5fc3-4c94-af0e-9f6dd47e0fc0 · inbound
Exploring LLMs for South Asian Music Understanding and Generation AudioGen: Textually Guided Audio Generation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f9cb6270-4fc0-4fba-8be1-7451caaa2f12 · inbound
HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis AudioGen: Textually Guided Audio Generation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 215416f7-9adc-446b-bb67-eef53a16984e · inbound
AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation AudioGen: Textually Guided Audio Generation
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e27c08c9-823f-49c2-8761-251e7f0b3cad · inbound
STAR-VAE: Structured Topology-Aware Regularization for Audio Reconstruction and Generation AudioGen: Textually Guided Audio Generation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f812876c-b041-46ca-b479-de859aacba17 · inbound
AudioCALM: Continuous Autoregressive Language Modeling for Universal Audio Generation AudioGen: Textually Guided Audio Generation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e81382e8-ef1a-4f3f-928f-ca254135b169 · inbound
DTM-Codec: Dynamic Token Masking for VFR Speech Coding with Efficient Boundary Selection AudioGen: Textually Guided Audio Generation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 577fed90-655a-45b5-8f7e-3dd254524b18 · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence AudioGen: Textually Guided Audio Generation
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a63fd72a-20b8-4faa-9bc0-8e0282303d1e · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence AudioGen: Textually Guided Audio Generation
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c63dd82-4e28-4882-8ed9-21601106bb7b · inbound
Efficient Text-to-Audio Generation via Pruning AudioGen: Textually Guided Audio Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83e62017-cf58-492c-8725-d771dc9ac622 · inbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models AudioGen: Textually Guided Audio Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b2fa850-bb71-4998-8e76-cabf201c68f4 · inbound
Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers AudioGen: Textually Guided Audio Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0610341-293a-487a-a718-a52ed9be0656 · inbound
AI-Based Sound Effect Generation: A Narrative Review of Generative Models Across Input Modalities AudioGen: Textually Guided Audio Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb6fb21a-5080-4518-a440-6a289ae14ec0 · inbound
AudioScape-TTA: A Structured Soundscape Benchmark for Fine-Grained Text-to-Audio Evaluation AudioGen: Textually Guided Audio Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 437ad67b-7280-40ed-9e1b-59ef718edfb1 · inbound
Dramarrator: Object-Based Audio Editing for Audio Drama Production from Books AudioGen: Textually Guided Audio Generation
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 561c049e-96ad-480d-9eca-3549fced96a6 · inbound
MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection AudioGen: Textually Guided Audio Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e7819bb-46a5-4095-844b-9aa85c87df2c · inbound
HarmoniDPO: Video-guided Audio Generation via Preference-Optimized Diffusion AudioGen: Textually Guided Audio Generation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 207557e9-24ec-461c-bbb6-0a732429bc3f · inbound
VoxAudio: Vocalized Audio Synthesis via Multi-Reward Autoregressive Flow Matching AudioGen: Textually Guided Audio Generation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.