Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T18:16:12.952995Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2412.08117.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T18:16:12.952995Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8f711594-054f-4465-b33b-08feedcfa46a · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Generative adversarial networks,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b3d9a24-927e-49c2-b31d-8966c22dab90 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Auto-Encoding Variational Bayes
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df2212d3-512f-44a9-b4eb-3266bc94de1c · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation High-resolution image synthesis with latent diffu- sion models,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 720960e0-eec0-4241-93e2-80741af4382f · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7793e79e-f2e5-444f-86c5-1434dd298a43 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Imagen Video: High Definition Video Generation with Diffusion Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a410033-4c77-4bd8-9410-ed42a2f2ea06 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Tacotron: Towards End-to-End Speech Synthesis
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fb9159b-4922-41f7-9d6b-bf0387be0815 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Fastspeech: Fast, robust and controllable text to speech,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ba92aff8-39ac-46cf-b0ce-2aa378d59444 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Stylespeech: Parameter-efficient fine tuning for pre-trained controllable text-to- speech,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1aa292d8-1074-403e-871c-d66fb6638db9 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation A Survey on Audio Diffusion Models: Text To Speech Synthesis and Enhancement in Generative AI
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dca5006-86dc-4fda-ad66-2cd78ebf63e9 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Diffvoice: Text-to-speech with latent diffusion,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 46ba7a1e-1aaf-4d62-9f4f-a52768b8babd · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation FastSpeech 2: Fast and High-Quality End-to-End Text to Speech
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5167604-be12-4ec3-8600-53b7c9ad8409 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation RAVE: A variational autoencoder for fast and high-quality neural audio synthesis
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4452742f-99ad-484f-93d4-bd6b6a6209bc · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Near-perfect-reconstruction pseudo-qmf banks,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9a2ae4ed-71a9-434c-be2c-6c8cfe5022d9 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation DDSP: Differentiable Digital Signal Processing
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ac8cbd3-03e8-4a65-a23e-ce9710f0a6ec · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Attention is all you need,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ee89e17-3e0b-44b8-9b61-b8b59403121a · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation DiffWave: A Versatile Diffusion Model for Audio Synthesis
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91d0d1b1-92b8-4264-a173-22593ea28d62 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Denoising diffusion prob- abilistic models,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef073fe5-4583-4c29-bfb8-01d8622775aa · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Chinese mandarin female corpus,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 00b0e12f-de65-47b2-a4c6-5fc694175c6d · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Mel-cepstral distance measure for objective speech quality assessment,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e9f7d144-ad7a-433b-baf2-703df0e7e363 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 86cd6324-93e4-4bac-adf1-32dd78709c14 · outbound
LatentSpeech: Latent Diffusion for Text-To-Speech Generation Robust speech recognition via large- scale weak supervision,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.