Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:54:55.275856Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2608.09571.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:54:55.275856Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation afd02cc8-07cd-496d-a50f-5e442ff4c8c2 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4eed1274-ef77-48e1-b345-c3ae0df48eb6 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation V oicebox: Text-guided multilingual universal speech generation at 8 scale.Advances in neural information processing systems, 36:14005–14034, 2023
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fb1a21ac-4182-42ba-84e7-1f03920daa17 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation F5- TTS: A fairytaler that fakes fluent and faithful speech with flow matching
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c3e5f785-c4a8-424a-8b6e-b1cbf54b13c0 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Tangoflux: Super fast and faithful text to audio generation with flow matching and clap-ranked preference optimization
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4fa110ad-eb7a-4c90-9dca-66fc47c1d481 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Plumbley
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c268a083-01d3-4d96-ae51-f753edf7da14 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation cdbad638-fd23-48e1-bf05-11a5343a5750 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Audiobox: Unified Audio Generation with Natural Language Prompts
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a466b657-6ca1-489c-921d-c11eb11d7be7 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Dasheng AudioGen: A Unified Model for Generating Coherent Audio Scenes from Text
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f59e3dc8-afac-4863-b9ef-1593a0f9b875 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation UniMoE-Audio: Unified speech and music generation with dynamic- capacity mixture-of-experts
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8e6278e1-319a-45e0-9363-943639f4f052 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Mandic, Wenwu Wang, and Mark D
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a826b188-4b5a-4d18-bb33-6e4b0ca9c963 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation UniSonate: A unified model for speech, music, and sound effect generation with text instructions
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 79092f9d-7ac3-4d2d-b7e0-8a2dbf9f2cda · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Make-An-Audio 2: Temporal-Enhanced Text-to-Audio Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d82903fd-9735-4cd8-81ec-05f31fbda54c · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Text-to-audio generation using in- struction guided latent diffusion model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 54b05279-be64-43e9-8226-51d3f4c08fac · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Freeaudio: Training-free timing plan- ning for controllable long-form text-to-audio generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e04a2ccd-3fd9-4258-a87f-8fe0dd487153 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity.Journal of Machine Learn- ing Research, 23(120):1–39, 2022
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62964b0e-7417-4d5b-9b98-07b190c2032c · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Scaling Diffusion Transformers to 16 Billion Parameters
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ea80760-c232-497a-bed6-eb0602fcff01 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Switch diffu- sion transformer: Synergizing denoising tasks with sparse mixture-of-experts
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f5509802-f7a0-46ce-8b60-2855f3ce56b7 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation EC-DIT: Scaling diffusion transformers with adaptive expert-choice routing
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 91b25063-e225-4f9a-92d6-762f8998988d · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c0ebbd6-d290-41f6-9327-a19b0d50143d · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Scalable diffusion mod- els with transformers
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ec6c0a83-9fa1-445a-aa46-3864d9e1d3cb · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Learning transferable visual models from natural language supervision
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8a2bcaef-be1c-44b5-9a36-b84fec2768e0 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e10443b-a2af-4cb4-9da5-27e2fbf33a07 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Step-Audio: Unified Understanding and Generation in Intelligent Speech Interaction
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0669497a-4ea5-47c6-8599-cf94c0d5df74 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Simple and controllable music generation.Advances in neural information processing systems, 36:47704–47720, 2023
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ef7088ac-e0e6-4331-a5c5-bff7ee09a496 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation SegTune: Structured and fine-grained con- trol for song generation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a41bb708-5063-4773-a12d-b90f59ac082e · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Qwen2.5-Omni Technical Report
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fda2e9ef-d7ff-4da9-a101-f60dad53afd7 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Parker, CJ Carr, Zack Zukowski, Josiah Taylor, and Jordi Pons
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 386c4899-330e-4995-a944-fd706fcd1e3d · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Audiox: A unified framework for anything-to-audio generation
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ddc464a5-37d5-489e-bdb3-37f041586904 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Higgs Audio V2: Redefining expressiveness in audio generation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4ae51aae-4b50-4398-bbf1-ffeb860a1e57 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation UniFlow- Audio: Unified flow matching for audio generation from omni-modalities.CoRR, abs/2509.24391, 2025
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 955e52c4-3a3b-455c-8728-ad26f9cbfee5 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Supervised learning of universal sentence representations from natural language inference data
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abcac6ea-9023-46a7-83e5-f42822daecd6 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Classifier-Free Diffusion Guidance
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6004f259-5b16-45c1-a7ea-f7031326e77c · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Eliminating oversaturation and artifacts of high guidance scales in diffusion models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ac4144c2-224e-423e-9d08-114090bf528a · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 371dd502-aa54-4b93-8057-8f167f221832 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Librispeech-pc: Benchmark for evaluation of punc- tuation and capitalization capabilities of end-to-end asr models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b2798beb-52d1-42af-ac10-55cdf9072415 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation AudioCaps: Generating captions for audios in the wild
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7f6a365e-c732-4d60-b0d2-c35994108945 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation MusicLM: Generating Music From Text
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b68f9f84-a28b-4284-9d43-227b77b4bca7 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation The Song Describer Dataset: a Corpus of Audio Captions for Music-and-Language Evaluation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9242d031-4e69-4b04-9d86-ddf022d84c86 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Qwen3-VL Technical Report
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cb6a07a-4fb6-4c10-b20e-18cc239d67a0 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Robust speech recognition via large-scale weak supervision
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c7d98b08-56b2-46b5-bba9-9d43b7952fa7 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation rain” may be an event in sfx, “an urban street in a downpour
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f7f18d1f-6b9b-4db2-b40a-3db806700a16 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation rule-based templates
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1fa24644-4753-455f-b26b-eb41a89cd3be · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Do not include markdown, explana- tions, or extra text
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0010179d-e9ce-4f5e-bd02-2954138956d6 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Do not add new keys
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 364a779a-e06d-4041-af44-06bacacb3304 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation - Do not paraphrase, expand, or rewrite it
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 28fde77a-e01a-4113-815d-914948893213 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation - Otherwise leave it empty
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 750581f5-7a9b-4b7d-be7b-b099686e335e · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Other- wise leave it empty
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 465acee9-987c-47c6-bfc0-4dab1cd5d5c7 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ace219d6-1fef-41df-969f-24e92c3f7d29 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation - Otherwise leave it empty
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 353f1487-1fbf-4872-a937-cbd532615bd3 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 83f9cdc4-50cd-4be2-bf37-b6357e747476 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6330a75d-ffd2-4610-a6f9-473999d3b4e5 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 889cc983-b815-4dd9-a3e1-56cfdcf16dad · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4c9b4f35-0a32-4976-af3c-7a1910d000f2 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation 1", "2",
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4038b4a8-cdc6-468c-835d-7d4d25138f2f · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0d1a757a-e876-4dc8-b70c-6f66be8fce40 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation summary":
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 57e5ad33-fe50-4da2-b06f-b3c4ed53a099 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Audiobox: Unified Audio Generation with Natural Language Prompts
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2ee54d8-1f0b-469a-98b8-2ae550ef66e5 · outbound
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation Qwen2.5-Omni Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.