Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:25:27.963778Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 2 inbound Pith citation observations for arXiv:2507.00227.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:25:27.963778Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:25:25.532629Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T21:25:28.221289Z
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 62ae4ed7-aadc-4108-929e-47e32adb34e9 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 49afb146-51f8-45db-b45b-9e62d7e0db1f · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Overall Pipeline The pipeline follows the architecture of ToucanTTS [22,23] due to its modularity and open-source implementation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9f358991-7879-422c-88a6-e17c4e762c14 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Datasets Training Data: For this work, we constrain ourselves to read speech, leaving experiments on conversational speech and other more challenging scenarios for future work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 36a8c919-1b5c-4d38-85a9-b55fc6f53735 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation efa6ef73-0ed5-4a80-8573-fa8b4f8c2b25 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Better speech synthesis through scaling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e3d7a7a-6229-458b-aee6-151703c289b9 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60e604e7-d9d6-4595-b079-2b58e419f91e · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Naturalspeech: End-to-end text-to- speech synthesis with human-level quality,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7bf20346-5421-4bb7-b7dc-288481467bc3 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis The models were trained for 100k steps with a batch size of 32, which allowed all models to converge
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 82a9bb4d-1a27-4e73-8901-e1fab26881cb · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis DelightfulTTS 2: End-to-End Speech Synthesis with Adversarial Vector-Quantized Auto-Encoders
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 584a3be3-77cc-4d51-9898-a16971698376 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis NaturalSpeech 2: Latent Diffusion Models are Natural and Zero-Shot Speech and Singing Synthesizers,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 60025069-ceb7-4c55-a07a-21a884d9afff · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis DelightfulTTS: The Microsoft Speech Synthesis System for Blizzard Challenge 2021,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 557055bb-dfe4-4e1c-afb2-f67bedb5357e · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis FastSpeech: fast, robust and controllable text to speech,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c5f6263b-378c-4b2b-8d67-c6a74ca4b328 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis A vector quantized approach for text to speech synthesis on real-world spontaneous speech,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 12e1a565-b37b-4476-95f3-ed08d75bc41c · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Prosody Is Not Identity: A Speaker Anonymization Approach Using Prosody Cloning,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3606ef3f-76b6-416d-a74c-7b655f6b3dc1 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis PoeticTTS - Control- lable Poetry Reading for Literary Studies,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8f0e7379-065e-4a69-be68-17a0be58255e · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49542191-d96c-4e3d-b266-6d82137c50fe · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Towards Controllable Speech Synthesis in the Era of Large Language Models: A Systematic Survey
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e802ac51-b664-4305-b18a-33aac17ab649 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Rectified flow: A marginal preserving approach to opti- mal transport,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 242c652e-5f93-4d53-9255-93f37d509e64 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis FastSpeech 2: Fast and High-Quality End-to-End Text to Speech,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6e20df70-d04a-48c3-85fd-ea2f07fb9a7c · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis FastPitch: Parallel text-to-speech with pitch pre- diction,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f9a864fb-5f9b-4004-9df1-bafdeb42aba1 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Variational inference with normal- izing flows,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2090e922-1f97-4c9f-bc2e-b63a297f009c · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Flow Matching for Generative Modeling,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 779bf251-b542-40dd-843b-4d7e562ca3ab · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 59365757-9d0b-446c-8ee0-cca33bebc2d8 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Language-Agnostic Meta-Learning for Low-Resource Text-to-Speech with Articulatory Features,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8f081286-6e61-4b29-bef8-b04928499a20 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cb83357-74db-4dc2-b2bf-27db27d741ad · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Varianceflow: High-quality and controllable text-to-speech using variance information via nor- malizing flow,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3e3a62d0-6f2a-47c5-a44b-15eb5ad24685 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 048829ea-3db1-4189-b01e-022b9e6e6f6e · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis The IMS Toucan system for the Blizzard Challenge 2023,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 45cbab1c-4d07-4b45-a091-589cf8fafecc · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Meta Learning Text-to-Speech Synthesis in over 7000 Languages,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7bd51685-5e4c-4780-95e7-84e69a297bd8 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis This dataset is com- prised exclusively of read speech in English and features 2,456 speakers
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation db497eca-0ddc-452a-b11c-9af99834d4de · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Con- former: Convolution-augmented Transformer for Speech Recog- nition,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4b29a346-0c94-4e90-bac6-4cfa8a31275d · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Exact Prosody Cloning in Zero- Shot Multispeaker Text-to-Speech,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 415de59b-ad34-4d83-8309-563006018323 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis ECAPA- TDNN: Emphasized Channel Attention, Propagation and Ag- gregation in TDNN Based Speaker Verification,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 42d95302-b371-44ce-85d5-f76c3b45f116 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis SpeechBrain: A General-Purpose Speech Toolkit
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 353c162b-3dc2-4405-815a-4947c14847fb · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Matcha-TTS: A fast TTS architecture with conditional flow matching,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 00f2ba9f-5f93-4624-b434-7c35ec4fd020 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Lib- riTTS: A Corpus Derived from LibriSpeech for Text-to-Speech,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bf9bc176-fd1d-4dbc-9c79-ee70e4dd65c0 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis The Ryerson Audio-Visual Database of Emotional Speech and Song (RA VDESS): A dy- namic, multimodal set of facial and vocal expressions in North American English,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 543afebb-e387-4652-b980-3adcf089e46d · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis ADEPT: A Dataset for Evaluating Prosody Transfer,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 27d9d0f8-92aa-4c9b-8c88-bac2c33b65c8 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Scalable diffusion models with transform- ers,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99558b67-758e-4e19-b44c-682672e7fe52 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Divergence measures based on the shannon entropy,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation dbd4f5c4-3d6a-4fc5-9a2e-3c672fec2720 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis On information and sufficiency,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b718218a-af35-42a3-9d89-517e3504650c · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Use of ranks in one-criterion variance analysis,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 85f92009-20f1-448b-8b24-28193523c985 · outbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Multiple comparisons using rank sums,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 62ae4ed7-aadc-4108-929e-47e32adb34e9 · inbound
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f5e865ca-1348-4678-84c3-c0977397f6ec · inbound
Bridging the Stability-Expressivity Gap: Synthetic Data Scaling and Preference Alignment for Low-Resource Spoken Language Models Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.