Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2406.05551.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:20.716820Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T03:27:35.627140Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a37e1d39-7169-43b7-b78e-807172d6bac5 · inbound
F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 119
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 409fb9d5-cb23-48db-a668-1044ea7a94e1 · inbound
Accelerating Flow-Matching-Based Text-to-Speech via Empirically Pruned Step Sampling Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda5b317-bf10-4a5b-9491-bd3be58499d9 · inbound
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9a8e6f38-969d-4cdd-bd82-df9d885b2315 · inbound
DMOSpeech 2: Reinforcement Learning for Duration Prediction in Metric-Optimized Speech Synthesis Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fbfafa0-f539-453a-95f2-15b3161ffe4c · inbound
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad2d839b-f8e5-48f7-82d1-2b7dae1305fa · inbound
INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 49e4c55a-ae76-4246-988d-bf6b3d735819 · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5cc8b4ba-b737-4398-9219-3a8d3acb7bc2 · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 138
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 742112f1-e411-4b25-99dd-9cc91516fdfb · inbound
V.O.I.C.E (Voice, Ownership, Identity, Control, Expression): Risk Taxonomy of Synthetic Voice Generation From Empirical Data Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 054760b5-1092-498f-a46d-067d866ca8d2 · inbound
Taming Audio VAEs via Target-KL Regularization Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation efba0c9f-6502-4976-9253-af1178854f41 · inbound
dots.tts Technical Report Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6a57ac6d-c5ad-41e0-949a-4c268b81d625 · inbound
HoliDubber: Holistic Video Dubbing for Complex Acoustic Scenes via Text-Guided Audio Synthesis Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 888525bb-4f74-4615-9b35-cae730747879 · inbound
ReGen: Hierarchical Multi-Prompt Representation Generation for Efficient Waveform Diffusion Models Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9da3a33a-10a0-4a4d-93d7-2e50bb0c1f16 · inbound
Stable Autoregressive Speech Generation with Low-Frame-Rate High-Dimensional Continuous Tokens Autoregressive Diffusion Transformer for Text-to-Speech Synthesis
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.