Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:04:51.686582Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2508.11326.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:04:51.686582Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-29T10:28:18.202974Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9738eb00-04f2-48bf-92c0-b27cd17b3e05 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts URL https://elevenlabs.io/
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ce9e6f66-ac65-4b5c-9169-07ff32a861b1 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts URL https://www.minimaxi.com/
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 89f4c621-ff5f-4dcc-822d-b331d18b5808 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Prompttts: Controllable text-to-speech with text descriptions
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8c8023d0-3447-4634-b9fe-92b3eadf8ecc · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Textrolspeech: A text style control speech corpus with codec language text-to-speech models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e089af01-3335-4723-9b80-2e7ee8217d75 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Prompttts 2: Describing and generating voices with text prompt
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1bdee4dd-a422-45a5-97ce-36ccaa271d3e · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Libritts-p: A corpus with speaking style and speaker identity prompts for text- to-speech and style captioning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fc63fbb-c694-47ce-96fa-7e5e5aa45525 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Speechcraft: A fine-grained expressive speech dataset with natural language description
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604d43fb-5ff3-4590-b857-a3b1e3df899f · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Natural language guidance of high-fidelity text-to-speech with synthetic annotations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad880798-f23f-401c-86a7-7da9281cad6e · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Scaling rich style-prompted text-to-speech datasets
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da4489d1-bfb3-4e5a-88bc-db1fc786b8b6 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts GPT-4 Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b68494d3-07d7-4dba-a3df-bf733c33b7f2 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0be7c83-54b5-4e2c-b38c-2317bd581025 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 971ec484-d792-4683-9284-c4dd997917f3 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8030ec9-5c4e-4c8b-b9de-b51f1d09328a · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts EmoVoice: LLM-based Emotional Text-To-Speech Model with Freestyle Text Prompting
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf596226-53b9-4921-ab0f-3282f7c3cfe8 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Qwen3 Technical Report
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf9a4a57-a755-4743-aae5-8ba47bbd97e6 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Investigating the Catastrophic Forgetting in Multimodal Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8fb27d0-048f-43b8-8811-f6ae30b3d965 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Mono-internvl: Pushing the boundaries of monolithic multimodal large language models with endogenous visual pre-training
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b2ff2d1-6018-4309-95ab-bf44defa8f83 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Investigating the Catastrophic Forgetting in Multimodal Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43dfd901-5a48-435a-97c5-6aa2d561ba95 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Vlmo: Unified vision- language pre-training with mixture-of-modality-experts
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6313292f-a74b-4422-952c-043a5e25126a · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts EVEv2: Improved Baselines for Encoder-Free Vision-Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0af524c-c2bd-43bc-b1e0-e4d2c969f240 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Simbert: Integrating retrieval and generation into bert
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 813005c5-8fe4-4834-b72a-2f5f23fc6ff8 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 695d6b7c-a62b-4a93-bf27-bbd25450c258 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Fastspeech: Fast, robust and controllable text to speech
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation faa15446-d0e2-479a-84e9-03b0b46f337b · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac9cca37-f696-416f-8db7-12761b0f7f4a · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts BERT: pre-training of deep bidirectional transformers for language understanding
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2c2df928-bc5f-48f1-b625-13fb9153f588 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts MiniMax-Speech: Intrinsic Zero-Shot Text-to-Speech with a Learnable Speaker Encoder
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88f43fff-0582-44ed-bff3-0f42b3fb910a · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Scaling vision-language models with sparse mixture of experts
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dae82b05-3ee8-421c-9566-a108eea12dd3 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 653e739d-550c-4d36-abff-8a836d6e1634 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bc8c221-afe1-4c67-97d8-90be3bb98c38 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts The Llama 3 Herd of Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d17d0c2a-3b7c-4648-827f-3ffda6d2e73f · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Orpheus tts: Towards human-sounding tts
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe17e825-4a1a-4234-a1b4-7f458d45e298 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Hubert: Self-supervised speech representation learning by masked prediction of hidden units
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cf0c18c0-5f8a-4972-b220-3ebce6c1707b · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Qwen2.5-1M Technical Report
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66e3afe6-3357-4c14-8fde-7939efa6c934 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Neural discrete representation learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc0ecd74-90e6-4083-b7bc-ac851aba7832 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Soundstream: An end-to-end neural audio codec
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ed51057-2fdc-4987-a663-7e456b2448e6 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts High-fidelity audio compression with improved RVQGAN
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9accb710-3a08-439c-b5ad-e7e7fbbba6c0 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Spark-TTS: An Efficient LLM-Based Text-to-Speech Model with Single-Stream Decoupled Speech Tokens
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdc1c6f7-2d9a-4f7d-9bd7-3bf548064f4d · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Self-supervised learning with random-projection quantizer for speech recognition
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ecd9b8f4-018b-449c-80fd-7bc44ad28dfb · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Parker, CJ Carr, Zack Zukowski, Josiah Taylor, and Jordi Pons
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dc488d9-eef8-4a29-9e0e-fef17261ee57 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Albergo, Nicholas M
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6b3cc7b-26e8-46e9-9a56-07636ec06f1b · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Emilia: A large-scale, extensive, multilingual, and diverse dataset for speech generation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be96fe05-0c7d-4492-9fbe-fc61f0d13e93 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Elucidating the de- sign space of diffusion-based generative models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8854b402-7093-4e14-9a29-82a68a516603 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts URL https://doi.org/10.1109/TASLP.2021
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d35de97-2743-4564-959d-2c91eed98a3a · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a804652-6dfb-4a4a-9f30-ad5deddc4abd · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts URL https://doi.org/10.1109/ ICASSP49357.2023.10096285
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5145ede-99f9-4cef-bc1e-8e700f33f581 · outbound
MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts doi: 10.18653/V1/N19-1423
Reference 4186
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c51c5e4-1523-44d3-a286-74cf8ac7a6d0 · inbound
Unified Synthesis of Compositional Speech and Sound from Free-Form Text Prompts MoE-TTS: Enhancing Out-of-Domain Text Understanding for Description-based TTS via Mixture-of-Experts
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.