Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:46:23.404901Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2412.01100.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:46:23.404901Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:46:23.292385Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-12T04:46:23.543943Z
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1190a25e-e7b2-4c45-889b-8d3fc6a7410d · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 On the one hand, TTS systems need to accurately replicate the target speakers’ voice, including their timbre, pitch, and prosody, using voice cloning techniques
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a5953255-227d-4271-9788-83a2cc98c164 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 592e28c6-cd5f-4b2a-98ac-9cdff6417d90 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Firstly, we will overview the text represen- tations and the discrete speech tokens, and then introduce the LLaMA-based codec language model with a delay pattern
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6f26b3b5-fbf7-41de-9a15-3bd8a680e28b · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Model Configurations We use the open-source models and parameters of MT5-base, DAC and HuBERT, with both DAC and HuBERT configured for a sampling rate of 16 kHz
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 01c61244-b2b2-41d0-90ba-e3291cb72429 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 嗯” (En- glish translation: “um
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b7fa6797-065c-45fa-8d5d-8d8bf1dff17b · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 We propose a LLaMA-based codec language model with a delay pattern for spontaneous style voice cloning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 542d5b6a-0118-423b-9794-2f06bfeb2511 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 139d5379-45c2-4cf5-bd1c-874b58db1a35 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Deep voice 3: Scaling text- to-speech with convolutional sequence learning,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b710d84c-195f-4657-98c3-a383d04cd232 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Neural voice cloning with a few samples,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ea60877d-f449-4e08-bac4-d82e9ed42ed0 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Adaspeech: Adaptive text to speech for custom voice,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2be492a3-b375-428d-9fd9-c5743916fb77 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d446060d-6a77-46c0-81dd-6940d5873c3c · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d293bd2d-c7d7-4380-9a81-fddc37b9b8b6 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speechx: Neural codec language model as a versatile speech transformer,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 52a4f702-c67b-4136-bb99-3b2fed255e00 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 ELLA-V: Stable Neural Codec Language Modeling with Alignment-guided Sequence Reordering
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9536cd50-92d2-420f-b1a8-48d92887b299 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 High Fidelity Neural Audio Compression
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56d4e3c2-ab15-4282-8e69-e1e06b5951f9 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Soundstream: An end-to-end neural audio codec,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2336ffaf-b83f-4884-8c33-73fefbbd7960 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Conversational end-to-end tts for voice agents,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9ace13ad-7a94-4da0-b69b-b61de73813ed · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 End-to-end text-to-speech based on latent representa- tion of speaking styles using spontaneous dialogue,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e941ab01-b81a-4b56-82c3-6611513a1fc5 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Spontts: modeling and transferring spontaneous style for tts,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ab184b98-a3ca-4c99-96f8-ebee51f4f082 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Controllable Context-aware Conversational Speech Synthesis
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation eef021a3-549c-473f-8d2f-0f8ff685f326 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Spontaneous Style Text-to-Speech Synthesis with Controllable Spontaneous Behaviors Based on Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8198aee0-864b-4a99-aae9-39ad1aece243 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 WenetSpeech4TTS: A 12,800-hour Mandarin TTS Corpus for Large Speech Generation Model Benchmark
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 713e2b56-1c94-4773-bb5a-85dd25ea7cc5 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Wenetspeech: A 10000+ hours multi- domain mandarin corpus for speech recognition,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d7bf9a78-9f93-4ed1-baf8-69e0a9c4cdcc · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 mt5: A massively multilingual pre-trained text-to-text transformer,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a8c3f4cf-12b0-42e2-b03f-74b31a86af06 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 HuBERT: Self-Supervised Speech Rep- resentation Learning by Masked Prediction of Hidden Units,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f66521be-d789-4b3c-8d13-68af8feaa5a6 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 High-fidelity audio compression with improved rvqgan,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0f5ec975-e862-44f8-8c43-08b91f42b660 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 LLaMA: Open and Efficient Foundation Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9cbc3d8-0b36-4985-b679-2ddd1768a059 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Flashattention: Fast and memory-efficient exact attention with io-awareness,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 30144310-756d-41df-ba16-52faf1b7a0e8 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Speak, read and prompt: High-fidelity text-to-speech with min- imal supervision,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b902a4c1-e7f8-48ac-8890-11e9b59feea2 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Simple and controllable music gen- eration,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d6439b49-a38e-4d9c-82ea-8e7aacd6f301 · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Natural language guidance of high-fidelity text-to-speech with synthetic annotations
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39d733d5-23be-4d16-a63b-ccb8ab33b86a · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 Stay on topic with Classifier-Free Guidance
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 724b1e42-e0be-4518-9a48-49440cf836ad · outbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 V oxinstruct: Expressive human instruction-to-speech generation with unified multilingual codec language modelling,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a5953255-227d-4271-9788-83a2cc98c164 · inbound
The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024 The Codec Language Model-based Zero-Shot Spontaneous Style TTS System for CoVoC Challenge 2024
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.