Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:05:29.628800Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2508.02391.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:05:29.628800Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T07:49:20.194990Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T05:56:40.912157Z
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 48db7d8e-d0e6-4105-95d2-2a4543c366b7 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution MusicLM: Generating Music From Text
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c06cd0e7-530e-43a7-9125-70cbc86b3060 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3673279d-2d10-4603-b4dc-35e525b8aeb8 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Musicldm: Enhancing novelty in text-to-music generation using beat-synchronous mixup strategies
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5c9013d8-5612-40b4-bc5f-5524d676e8f3 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Wavlm: Large-scale self-supervised pre- training for full stack speech processing.IEEE Journal of Selected Topics in Signal Processing, 16(6):1505–1518, 2022
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95d9772d-68ba-4190-8073-53d2ae2d242e · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Qwen2-Audio Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 518c59ab-ee3e-47fc-94d1-251728d6e871 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Directly Fine-Tuning Diffusion Models on Differentiable Rewards
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2748de54-886e-422f-9f23-377135b90ad9 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Clap learning audio concepts from natural language supervision
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af261e11-cd76-4626-a8df-d031500f5e34 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution FunASR: A Fundamental End-to-End Speech Recognition Toolkit
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 152c3e68-9ab9-4f09-9092-41f59eebed06 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffaedfe7-3c15-4547-98b4-6cd388f1a795 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution NU-Wave 2: A General Neural Audio Upsampling Model for Various Sampling Rates
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c62d301-5dc5-4ce7-b698-6ed0dfdbf324 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e50f301c-ad54-4c9a-bd2a-4f9ae10809a1 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Classifier-Free Diffusion Guidance
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 023b804b-cf22-47f1-9408-d224b0a9f3f5 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis.Advances in neural information processing systems, 33:17022–17033, 2020
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 014c09e7-7854-42f0-aa36-8d28a67ce2ca · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution NU-Wave: A Diffusion Probabilistic Model for Neural Audio Upsampling
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 827cfac1-42dc-47e1-b3ff-6d89c3876da1 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Noise-free optimization in early training steps for image super- resolution
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab16529c-e383-45ff-aaf6-4716bb46018c · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Audiosr: Versatile audio super-resolution at scale
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0fee5c5-370f-447c-94d1-65406a061189 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2bc81cf-3b27-4e90-a3f2-6b936bfc3ae5 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Neural Vocoder is All You Need for Speech Super-resolution
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e068ebb-6c8c-4317-93d5-a1dc6e83723c · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution VoiceFixer: Toward General Speech Restoration with Neural Vocoder
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d29ae95f-9b46-41b3-9a01-5effb4bc06e2 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.Advances in Neural Information Processing Systems, 35:5775–5787, 2022
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f11eca33-6ed1-4d3b-91e2-a8800b66484e · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Scaling inference time compute for diffusion models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d014c32a-f416-493b-9aee-f7fc60768fa7 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Uncertainty-driven loss for single image super-resolution.Advances in Neural Information Processing Systems, 34:16398–16409, 2021
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 44414bcc-ba4b-4cb4-8992-01c2a6a7a066 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dd6f9d9-f804-468b-800a-17064386a1f9 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Esc: Dataset for environmental sound classification
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7474cf9f-2393-45cf-b483-29ddd7c48efa · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Robust speech recognition via large-scale weak supervision
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68004d0d-dddb-45cc-b8d5-53f7c625913f · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution FastSpeech 2: Fast and High-Quality End-to-End Text to Speech
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd7b94e0-ed22-4fe2-941e-7ac4b062b715 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Fastspeech: Fast, robust and controllable text to speech.Advances in neural information processing systems, 32, 2019
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5bfd7618-59e3-4446-b30a-2a323305096f · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution High- resolution image synthesis with latent diffusion models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8139e27-ff8e-4e4d-86e3-ba24bdd26ffb · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da5f8c3c-c005-4d6c-a313-048a8742480e · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bad2a83-4fc0-408d-a093-cf9fda9f040e · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Denoising Diffusion Implicit Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 421a2ab1-1d8c-48ed-b174-a78fcbeacf4c · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Score-Based Generative Modeling through Stochastic Differential Equations
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e073ebd2-2993-4a1a-ac33-63b1d0d1c62b · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution AudioX: A Unified Framework for Anything-to-Audio Generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5df9d974-cdcf-4a0d-a221-71806c1dde49 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Meta Audiobox Aesthetics: Unified Automatic Quality Assessment for Speech, Music, and Sound
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c40a1b6a-db08-4293-b150-fb6fbc2bd88b · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Towards robust speech super-resolution.IEEE/ACM transactions on audio, speech, and language processing, 29:2058–2066, 2021
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fee83341-ce77-442d-9a4f-0cef3b860c8f · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Esrgan: Enhanced super-resolution generative adversarial networks
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f7a061b5-c1ea-4829-a27e-8209352f00f1 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Difix3d+: Improving 3d reconstructions with single-step diffusion models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae902334-f46f-4de1-8968-28ec79d9aa6c · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47215184-baae-4547-bdcc-39ff7bc441d1 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Flashspeech: Efficient zero-shot speech synthesis
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f925a107-9820-4bfc-ba08-83c3990f4c71 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Comospeech: One-step speech and singing voice synthesis via consistency model
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 751e1231-605c-4b79-85a1-fd9ae8a5d8d5 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98348ee1-b86b-421a-afa6-e30ad42aff16 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution GAN Vocoder: Multi-Resolution Discriminator Is All You Need
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9876195-b761-4a36-9249-f6796d051134 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Conditioning and sampling in variational diffusion models for speech super-resolution
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f1ca83c3-1f35-4bbc-a6ef-48de064b3670 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Uncertainty-guided perturbation for image super-resolution diffusion model
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c1d3dd04-2823-48ce-95c4-3e9ca5587cd6 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Flashvideo: Flowing fidelity to detail for efficient high-resolution video generation.arXiv preprint arXiv:2502.05179, 2025
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6725dd0-fa92-42f8-93ce-be202f53b31d · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution Inference-time scaling of diffusion models through classical search.arXiv preprint arXiv:2505.23614, 2025
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d5bd33c-6925-46c0-856b-49e492e51a65 · outbound
Inference-time Scaling for Diffusion-based Audio Super-resolution In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59270332-aa16-4238-8b3f-21ab1d15d682 · inbound
Inference-Time Scaling for Joint Audio-Video Generation Inference-time Scaling for Diffusion-based Audio Super-resolution
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.