Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:41:01.512970Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2412.20048.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:41:01.512970Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
75 of 75 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c23cb365-eee1-4a1b-bb8a-07ca5bd05705 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation The amazing benefits of being bilingual,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5d15a981-3d6f-4e69-a3a7-70b09b9846fd · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Disentangled representation learning for multilingual speaker recogni- tion,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8216c3ff-4d1a-4620-a9b8-fd20ab69d79a · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Crosslingual and multilingual speech recognition based on the speech manifold,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2b16526b-174f-48db-b2f1-7ea9f1eaa3fc · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Distilling a pretrained language model to a multilingual asr model,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2b79a606-c8fd-4e75-98e3-43b4e84b1814 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Joint ASR and language identification using RNN-T: An efficient approach to dynamic language switching,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation eed4ff27-cd21-4963-ba27-574002257665 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Joint unsupervised and supervised learning for context-aware language iden- tification,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 349dcebc-a7fa-433e-97cd-28dc3958c96b · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Fastpitch: Parallel text-to-speech with pitch prediction,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bdb4b98c-6964-4637-9bcc-d30e1d0d7339 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Speaker adaptive text-to-speech with timbre-normalized vector-quantized feature,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a604f3c6-c4ef-4ebd-896c-415750cedae8 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation EfficientTTS 2: Variational end-to-end text-to-speech synthesis and voice conversion,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 05260672-3867-48f6-89a2-bf5c14eae1b5 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Bytes are all you need: End-to-end multilingual speech recognition and synthesis with bytes,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a11dc593-9a18-48e6-9cf7-05aff151bfdd · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Improve cross-lingual text-to- speech synthesis on monolingual corpora with pitch contour informa- tion,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8302a362-5913-4d46-8a29-daf86d124842 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Language-agnostic meta-learning for low-resource text-to-speech with articulatory features,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation af47961d-8ebb-45e2-a4ea-6e4228fde64b · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Learning to speak fluently in a foreign language: Multilingual speech synthesis and cross-language voice cloning,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8d41d9eb-99b6-4812-a974-2a26a56b0778 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Disentan- gled speaker and language representations using mutual information minimization and domain adaptation for cross-lingual TTS,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c6f464f7-6ff1-4529-8ab7-2d9d93c1e1a1 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation GenerTTS: Pronunciation disentanglement for timbre and style generalization in cross-lingual text-to-speech,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cd8acda7-6487-45a4-82f8-9eba3e9e4432 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation DSE-TTS: Dual speaker embedding for cross-lingual text-to-speech,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 93e48c2c-23e7-41ff-8d96-4703c38e8ec7 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation ZMM-TTS: Zero-shot Multilingual and Multispeaker Speech Synthesis Conditioned on Self-supervised Discrete Speech Representations
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 24b0fff1-493f-4e57-bbd5-7cbf40f7e594 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Unit selection in a concatenative speech synthesis system using a large speech database,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 91598463-fd05-41fc-9cb8-7aa59d3722e9 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Statistical parametric speech synthesis,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6d1d7ac2-cff8-4500-a8f4-be36d07032a7 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Naturalspeech: End-to-end text-to-speech synthesis with human-level quality,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 30070c5f-f1c9-4403-9a58-cf8fa98d6ec9 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Harmonic-net: Fundamental frequency and speech rate controllable fast neural vocoder,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 17f12e01-95b9-44b2-8e78-61efbc5e51c5 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Fregrad: Lightweight and fast frequency-aware diffusion vocoder,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5b22e679-ee23-42cb-bb09-a32f3753b78f · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation TriniTTS: Pitch-controllable end-to-end TTS without external aligner
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 86d5654b-fd8c-4a89-b268-c97125f338ff · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Hierspeech: Bridging the gap between text and speech by hierarchical variational inference using self-supervised representations for speech synthesis,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 614d5c98-5ce2-4392-8be2-08ff8d25809c · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation WaveNet: A Generative Model for Raw Audio
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bceb40e-42d5-4061-a722-705378184bec · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Deep voice: Real-time neural text-to-speech,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c9318edb-cb90-4bc6-b523-4e1e1278648e · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Natural TTS synthesis by conditioning wavenet on mel spectrogram predictions,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 57adf011-d731-4c2d-8505-c9b089076f6b · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Fastspeech 2: Fast and high-quality end-to-end text to speech,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 199ac34f-19af-4db2-a874-85977630b03b · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Matcha- TTS: A fast TTS architecture with conditional flow matching,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bb6e89df-9fc1-48d8-ac11-f7b31e9a642c · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Multispeech: Multi-speaker text to speech with transformer,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a10192b2-9bf5-40e9-b767-bef9dbf8dbc8 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Lightspeech: Lightweight and fast text to speech with neural architecture search,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 92c323f5-604a-4040-ab6b-ff61414a0bcd · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Phonological features for 0-shot multilingual speech synthesis,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 85d6b405-fd35-4a37-adff-08282372a955 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Text-inductive graphone-based language adaptation for low-resource speech synthesis,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7e7c046c-2cf5-42ac-b3f2-7cfea3e0e140 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation BERT: Pre-training of deep bidirectional transformers for language understanding,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7af7583e-39dd-4fd3-acf9-644ea4cd4064 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Domain-adversarial training of neural networks,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 785c3b1b-eafd-4333-8647-7bc47c845abd · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Learning disentangled representations via mutual information estimation,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 36afed2e-9437-4b04-a4dc-48e4d2cfb84e · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation SANE-TTS: Stable and natural end-to-end multilingual text-to-speech,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 97c9a134-6f77-4bdc-b621-9e75747d8f33 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Crossspeech: Speaker-independent acoustic representation for cross- lingual speech synthesis,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a47dc9af-30fa-4784-a4a9-756411a563de · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Invariant risk minimization,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dc79d5a0-1db8-437b-8f03-c3ec0ddb7f7d · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Domain generalization with mixstyle,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b1a2639b-98e0-4e28-95e7-d75dd4702376 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation One TTS alignment to rule them all,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d0c5ad9d-09b4-4dde-967d-f6f1541e0898 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Conformer: Convolution-augmented transformer for speech recognition,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3ba2697f-2546-41d8-ae46-6659e9c2772f · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Feature-critic networks for heterogeneous domain generalization,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a66522b6-11f1-4ab5-b738-483249044bb8 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Generspeech: Towards style transfer for generalizable out-of-domain text-to-speech,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c4ece250-45d7-4872-beb6-ebd9ae1e1680 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation PV AE-TTS: Adaptive text-to-speech via progressive style adaptation,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fcfee05f-68af-4a65-998e-6b01c8c3a593 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Style normalization and restitution for domain generalization and adaptation,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 602d7743-9ae7-4ad6-b30e-d0c9a3de8ea6 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Prosospeech: Enhancing prosody with quantized vector pre-training in text-to-speech,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 01b8547f-fdc7-467c-a438-a8e4011503cd · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Diffprosody: Diffusion-based latent prosody generation for expressive speech synthesis with prosody conditional adversarial training,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 26cfbfa1-825f-4daa-a197-d4e51f33f2c6 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation pYIN: A fundamental frequency estimator using probabilistic threshold distributions,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f75cf353-7153-4d34-88dc-12db8c1bde0a · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Neural analysis and synthesis: Reconstructing speech from self-supervised representations,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 760857da-4f2d-4ad3-b8d5-19ec24273660 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Exploring wav2vec 2.0 on speaker verification and language identification,
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 71799840-ac50-45c1-893c-407a7d1fd3a2 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Let there be sound: Reconstructing high quality speech from silent videos,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7cf0cd41-4470-4132-a999-417d4e9ebbae · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Scaling speech technology to 1,000+ languages,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d3ef05fd-f846-4cc0-a9b2-c146ec400b44 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Wav2vec 2.0: A framework for self-supervised learning of speech representations,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0f4223bf-e28b-4eee-8e8c-47859fd98892 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Connection- ist temporal classification: labelling unsegmented sequence data with recurrent neural networks,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation acaa876b-4f89-4927-9574-4d563ef18ac5 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation The LJ speech dataset,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a57298f3-6db4-4c26-bd39-344a6ea1c185 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation CSTR VCTK corpus: English multi-speaker corpus for CSTR voice cloning toolkit,
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5c4ab24-b74c-4038-a6a3-f30616549ad2 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation The BIAOBEI dataset,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0e24fc1d-ff54-4eac-9a95-82c80d341ab7 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation AISHELL-3: A multi- speaker Mandarin TTS corpus,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation db8ce6fb-d726-4e9a-8b6f-d6c4797a14b3 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation CSS10: A collection of single speaker speech datasets for 10 languages,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b73612f4-d09b-4e05-8e46-eb109fdb70ce · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation JSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a2e9241-8696-4630-ae8c-87cb57ecb2ae · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Multi-speaker TTS data,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3972753a-281c-48cd-8d64-58c05baed105 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Phonemizer: Text to phones transcription for multiple languages in python,
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 83f93c23-0ea3-49da-bc02-620346aed726 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation NANSY++: Unified voice synthesis with neural analysis and synthesis,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation eb81c13e-73ad-42df-bdce-8e2274db605a · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Language modeling with gated convolutional networks,
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dd1b4f6f-e28d-4026-b08d-8bcceceaadab · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Fre-GAN: Adversarial frequency-consistent audio synthesis,
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 966db735-0d84-4e0b-bec3-82e608b80322 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation UTMOS: Utokyo-sarulab system for voicemos challenge 2022,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3704482b-c728-455c-ad3e-694cd1fe8531 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation The blizzard challenge 2005: Evaluating corpus-based speech synthesis on common databases,
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e17948f1-e53e-4637-a248-cd30562989d9 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation USAT: A universal speaker-adaptive text-to-speech approach,
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d5973154-1774-4f7a-9d78-36add20b08ae · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Dual-branch modeling based on state-space model for speech enhancement,
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5b5b390d-242c-4881-949c-f96c4c20a761 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation V oicegrad: Non-parallel any-to-many voice conversion with annealed langevin dy- namics,
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7f978c48-dc69-4df5-913d-dc6478aadf7d · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Robust speech recognition via large-scale weak super- vision,
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 192c509f-3006-45f7-88c2-f350c5d28d0d · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Visualizing data using t-SNE,
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 96e5ca60-1ce3-47fe-b1c8-a37cf7fd8237 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99f0a98a-20ff-4fe6-ac2b-0712939f2c70 · outbound
CrossSpeech++: Cross-lingual Speech Synthesis with Decoupled Language and Speaker Generation XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.