Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:02:03.077732Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 4 inbound Pith citation observations for arXiv:2507.21138.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:02:03.077732Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T06:36:31.523597Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T23:19:02.942476Z
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1116d9ff-1bb6-45fe-8c3b-2d690506a98d · outbound
TTS-1 Technical Report The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6165c63-f5b3-490c-908c-7f7aa6a4e881 · outbound
TTS-1 Technical Report Yodas: Youtube-oriented dataset for audio and speech
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9a8e03bc-ef3b-494a-8385-5600549cb4dc · outbound
TTS-1 Technical Report Emilia: An extensive, multilingual, and diverse speech dataset for large-scale speech generation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a5392520-ba95-430c-a8b1-d442e5564146 · outbound
TTS-1 Technical Report Styletts 2: Towards human-level text-to-speech through style diffusion and adversarial training with large speech language models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d993e778-bc6a-4974-a47b-8ab4e75e5f8b · outbound
TTS-1 Technical Report Fastspeech: Fast, robust and controllable text to speech
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 448ecfc7-dea8-42b5-b8e6-c5ee136542be · outbound
TTS-1 Technical Report Better speech synthesis through scaling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9573b28-862d-43f8-928b-4878a58d2629 · outbound
TTS-1 Technical Report Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e204a9e3-3816-48fe-84cf-4e4f73ff3a0c · outbound
TTS-1 Technical Report V oicebox: Text-guided multilingual universal speech generation at scale
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3fb837f9-85f5-43b5-add2-e03a99597d36 · outbound
TTS-1 Technical Report Language models are unsupervised multitask learners
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c2d9832-c784-4272-b9a4-311689e7b0d2 · outbound
TTS-1 Technical Report Training Compute-Optimal Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d96c6f96-d5ae-4aa6-b75b-63251abf5781 · outbound
TTS-1 Technical Report Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc5e8aef-029b-4609-8889-5fef279eb141 · outbound
TTS-1 Technical Report MiniMax-Speech: Intrinsic Zero-Shot Text-to-Speech with a Learnable Speaker Encoder
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aed9d272-a437-42ff-bec8-c77046dd3603 · outbound
TTS-1 Technical Report Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3c777f1-cd98-4493-aae4-b99d5ea637dd · outbound
TTS-1 Technical Report CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2cd3a78-d85f-44a1-9c4b-f1c80281eac1 · outbound
TTS-1 Technical Report CosyVoice 2: Scalable Streaming Speech Synthesis with Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd31abf3-cdc3-4957-9bde-e4ca9251251f · outbound
TTS-1 Technical Report Redpajama: an open dataset for training large language models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 669d28f8-b554-4228-a0ee-fdb73850def1 · outbound
TTS-1 Technical Report Open instruction generalist (oig) dataset
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1f6b538f-24df-4cef-b876-74eb4f3f71dc · outbound
TTS-1 Technical Report DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a853971-6149-46e6-9863-4098efe1267e · outbound
TTS-1 Technical Report Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61e20d6f-5552-4c0c-a9a8-b4a2d805e76e · outbound
TTS-1 Technical Report Wavlm: Large-scale self-supervised pre-training for full stack speech processing
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation befd186d-6c49-493d-8b30-f788cea54faa · outbound
TTS-1 Technical Report Dnsmos p
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d900436-35b3-4f17-b713-eba6b6f78fa3 · outbound
TTS-1 Technical Report Lora: Low-rank adaptation of large language models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdf5b87d-9598-4cab-88a2-5c88c38d010a · outbound
TTS-1 Technical Report Deep residual learning for image recognition
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 089d6aab-8a5f-4074-bdbd-071b1dd9a10c · outbound
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e452096-9f3d-4855-8a96-101daea28de3 · outbound
TTS-1 Technical Report Matrix multiplication background user’s guide
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63165417-09a3-4464-b43e-d01040b7e49f · outbound
TTS-1 Technical Report Initializing new word embeddings for pretrained language models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dee61c16-a693-413d-baec-a8d71bc82094 · outbound
TTS-1 Technical Report BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7372153-22da-4005-9a06-a28caf869736 · outbound
TTS-1 Technical Report Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 331a3e81-ee3b-4418-9bf2-aac95671b668 · outbound
TTS-1 Technical Report High Fidelity Neural Audio Compression
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfac856f-3877-4bf0-9bdb-d925ad897b44 · outbound
TTS-1 Technical Report Adding Instructions during Pretraining: Effective Way of Controlling Toxicity in Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2d0e952-b052-4dc2-8a25-3c9367e1cfe1 · outbound
TTS-1 Technical Report A neural probabilistic language model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9133858a-f8b9-420b-8e82-19916efa00be · outbound
TTS-1 Technical Report FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29c6c949-fa70-4f7c-8811-4deab52cb8f1 · outbound
TTS-1 Technical Report Adam: A Method for Stochastic Optimization
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc53bb10-0b74-41a7-93a0-bac436025d36 · outbound
TTS-1 Technical Report PyTorch Distributed: Experiences on Accelerating Data Parallel Training
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d962fbbe-3828-4961-9347-cf0e52fb0db0 · outbound
TTS-1 Technical Report PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e27dba3b-4a19-4c67-a1b4-2ed803652cb5 · outbound
TTS-1 Technical Report Parler-tts
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a5c05da-c5d0-4796-bf6d-9303960c483b · outbound
TTS-1 Technical Report Zero: Memory optimizations toward training trillion parameter models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52ded5a3-2265-49cd-b401-61758dbf0392 · outbound
TTS-1 Technical Report CosyVoice 3: Towards In-the-wild Speech Generation via Scaling-up and Post-training
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3640e72a-bbf6-47a0-89f5-b238dca3c766 · outbound
TTS-1 Technical Report Enhancing Zero-shot Text-to-Speech Synthesis with Human Feedback
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a88a8cd-6997-471a-8f21-bc960cf587e3 · outbound
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63c76d66-c01b-4edd-9a8a-2caebbdb197f · outbound
TTS-1 Technical Report DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ca58588-1f1f-4943-8194-60f8dbc61c15 · outbound
TTS-1 Technical Report Direct preference optimization: Your language model is secretly a reward model
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f720f069-b85f-48e8-afe8-ae422da1397e · outbound
TTS-1 Technical Report Understanding R1-Zero-Like Training: A Critical Perspective
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9b904ce-8b94-46ad-9c3d-99590258e76e · outbound
TTS-1 Technical Report Robust speech recognition via large-scale weak supervision
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d2518ba-0c5b-4908-ace2-7d94f7171699 · outbound
TTS-1 Technical Report HuggingFace's Transformers: State-of-the-art Natural Language Processing
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd80dfc2-0d2c-400a-818f-ebd963c3eeab · outbound
TTS-1 Technical Report PyTorch: An Imperative Style, High-Performance Deep Learning Library
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a606eda4-cfd8-44ee-85bf-7b99585a2b9b · outbound
TTS-1 Technical Report Pytorch lightning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c88b760d-9f09-4e4e-a31c-fb9f6ae7cc7b · outbound
TTS-1 Technical Report Efficient memory management for large language model serving with pagedattention
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64fed19a-4d58-49a6-867f-8977988efecd · inbound
Evaluating and Rewarding LALMs for Expressive Role-Play TTS via Mean Continuation Log-Probability TTS-1 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b76085d3-b265-4c38-beb9-aa07f37b5812 · inbound
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cca52b6c-b63b-414c-9517-cb5c8bc46c2a · inbound
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc3f5c45-0029-4f23-a0c6-f49bfa10d585 · inbound
Reliable Neural-Codec Text-to-Speech by ASR Self-Verification and Distillation: Near-Zero Catastrophic Failures Across Models and Codecs TTS-1 Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.