Pith. sign in

Paper Citation Record · LEDGER

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

As of 8 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 3 inbound Pith citation observations for arXiv:1907.04448.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1907.04448 v2

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-25T00:09:35.661121Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:29.459675Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-25T00:10:07.071622Z

Reference resolution

36 of 36 outbound references displayed

  • verified exact5
  • verified fuzzy27
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ae3d3a5-6cbd-4435-b177-51c8f00bf4b9 · outbound

This paper cites prosody, by conditioning synthesis on la- tent representations [8–12] in addition to text.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning prosody, by conditioning synthesis on la- tent representations [8–12] in addition to text

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.660334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:709810d7adff6d3abcd108f385257db3fd0095a4b1eb54dc68f595aaeae1d64c

Observation 6fd71456-8b9d-4435-b97c-96a539929a35 · outbound

This paper cites Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T00:10:07.075022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:fd3e144026d4c0e159886696fc90d5ac005301e94646a4e24dd3bc053782afe8

Observation 6cfd8f0d-51a0-4201-b5a4-346c0b1007e9 · outbound

This paper cites an unresolved cited work.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:10:07.642502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:74e39d4e9db64c5d7357b43c68988ec5c921a83255e609e3afe38595fa9aa724

Observation fa48e1d4-4bc6-477f-925c-40da5a1e5ccf · outbound

This paper cites heavyaccented.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning heavyaccented

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.589630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:8b9e9b16ca3f0b29fe6838094495961eb2355902a997446529fffa972260febc

Observation 0cb291dc-c2f5-4781-906a-d237027c3a92 · outbound

This paper cites an unresolved cited work.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:10:07.648339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:930d7a57e1edb8b8706f32a76450395d01dbc27e1d9317aeea4342d4795da620

Observation 13dd437f-953e-4e43-9586-426de24c3f0f · outbound

This paper cites an unresolved cited work.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:10:07.680552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:2a90e72ff2117191e874182b3c91aae22d547edffc9cdee192c05d7c8568e789

Observation dad61d25-0bb5-4b23-9366-265181562df0 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning WaveNet: A Generative Model for Raw Audio

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.041425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:079cb3d59b3c5e31e150469003cbddcfccbea1f8dee985ae8b1559294c9786d5

Observation da5c7525-49ce-40ba-b433-1f8767d18ee9 · outbound

This paper cites Tacotron: A fully end-to-end text-to-speech synthesis model.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Tacotron: A fully end-to-end text-to-speech synthesis model

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.576486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:1205a5badf39f273c1908d762e991e6689361e558d9cf3ced45f5658755aa784

Observation 4a71613e-5913-4fa7-8311-5a843e6d3c49 · outbound

This paper cites DeepVoice2: Multi-speakerneuraltext- to-speech.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning DeepVoice2: Multi-speakerneuraltext- to-speech

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.572278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:1ee3c6208eabc78f12dbffffe837ce57d70de91337422cfdcf59c6866c99050c

Observation b6187871-b060-4c34-adce-e4132faea9e0 · outbound

This paper cites Neuralvoice cloning with a few samples.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Neuralvoice cloning with a few samples

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.667956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:710f7bf9517036c7ce8ad5cbb56df82598e80c94fd9d2b32e9795bc9dbfec279

Observation f73a218d-7ba3-46ec-ba2c-d62a60168311 · outbound

This paper cites Transfer learn- ing from speaker verification to multispeaker text-to-speech syn- thesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Transfer learn- ing from speaker verification to multispeaker text-to-speech syn- thesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.602137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:cd226b7368b8253fff75f30fdad367e743eec6472a8375c269ba33135c2ed032

Observation 463ce1de-833a-4c65-bdd9-02d7ffd90e7f · outbound

This paper cites Fitting new speakers based on a short untranscribed sample.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Fitting new speakers based on a short untranscribed sample

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.672328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:dfcbf44264445474b91dc261583c6e84d5c327fc1190b3ac7e84784612627740

Observation 50f1348c-36f9-4db1-9bd1-b67c37892f87 · outbound

This paper cites Sample Efficient Adaptive Text-to-Speech.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Sample Efficient Adaptive Text-to-Speech

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.062820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:e3296e3102d7f8a4d35bf2427ade44d54be56e4b35c2b533034d216117c29d2f

Observation 1546fa51-115f-4a65-97e2-6a89c39779eb · outbound

This paper cites Style tokens: Unsupervised style modeling, control and transfer in end-to-end speechsynthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Style tokens: Unsupervised style modeling, control and transfer in end-to-end speechsynthesis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.605944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:d3a5fa0399e27ac8cb745e5f5492e4aad07419a9fc74985881117e5b4dd01d87

Observation 4e89d297-4da6-4fea-a261-56f9edce538e · outbound

This paper cites Towards end- to-endprosodytransferforexpressivespeechsynthesiswithTaco- tron.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Towards end- to-endprosodytransferforexpressivespeechsynthesiswithTaco- tron

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.618551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:abb964dd0f3654d526c7ab7790f996f47883f64f7e0930061aef02f3c81c6db5

Observation 82dfebc3-3a6a-44a0-a101-eff418e3ecf1 · outbound

This paper cites Expressive speech synthesisviamodelingexpressionswithvariationalautoencoder.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Expressive speech synthesisviamodelingexpressionswithvariationalautoencoder

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.626764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:2e429bffd248870ce01510104a7c32b0a3c6ca208eafc0d4c95db3cd24d37048

Observation 9ba06a2b-ac69-44c8-8be9-d88c3e4cc5f6 · outbound

This paper cites Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.055571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:02ef16f72733e8203a4bf5198b21bf83f156652930ff2368d77e48b8d71717e9

Observation 743c9309-9d1c-40f6-af1a-c75a2ed2f734 · outbound

This paper cites Hierarchical generative modeling for controllable speech synthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Hierarchical generative modeling for controllable speech synthesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.594414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:960f8a62d5bf9b79b5a61b6cc99e8f2c884d5ef272788467aee1fdf7e45b42e7

Observation 60b12332-d185-49e3-8474-e24b70eafd40 · outbound

This paper cites Statistical parametric speech syn- thesis based on speaker and language factorization.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Statistical parametric speech syn- thesis based on speaker and language factorization

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.598354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:77c2529a5a4b54db07d4666ff0c63f71a9d8102eb6692578d4a26f83da7fa90b

Observation bcdd3e0e-eea6-43b6-b75f-ad844e25f929 · outbound

This paper cites Multi-language multi-speaker acoustic model- ingforLSTM-RNNbasedstatisticalparametricspeechsynthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Multi-language multi-speaker acoustic model- ingforLSTM-RNNbasedstatisticalparametricspeechsynthesis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.609809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:bbf313f3e6f86a2cc16580fdc4365b5053cb28e4dcb179e58748ddd7bde2016e

Observation fa1e41ff-b3e3-4599-ba23-1f2a46f9ef41 · outbound

This paper cites A light-weight method of building an LSTM-RNN-based bilingual TTS system.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning A light-weight method of building an LSTM-RNN-based bilingual TTS system

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.614306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:eaf2ca9d2e64376e90b88a5cef6963abbb3dd7941b764ecc35a049de8c6302b2

Observation 0835da2a-2a5b-47f4-a6b5-5bf4cae7d335 · outbound

This paper cites Learning pronunciation from a foreign language in speech synthesis networks.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Learning pronunciation from a foreign language in speech synthesis networks

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-25T00:10:07.048463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:cc4acfda4600b70bc7e602d8b3b918af08865d5271b44268e387d474ce117000

Observation 9cb4e778-c9bb-4a9f-8c20-c173c7e27d25 · outbound

This paper cites Unsupervisedpolyglottexttospeech.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Unsupervisedpolyglottexttospeech

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.676164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:f207779941a173e9502440e69cf08e5c7fe4857a6a4dd0892a9e5302b95d0f01

Observation 343089da-d838-405c-abdc-153c89ec3e02 · outbound

This paper cites WORLD: a vocoder- based high-quality speech synthesis system for real-time applica- tions.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning WORLD: a vocoder- based high-quality speech synthesis system for real-time applica- tions

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.664180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:275b66401de9f9d28e9aaa1e3a62a2cad71d93d146aa7a297147fdccdb0ce011

Observation 21e1c771-0432-47df-b516-f1a73f85d01b · outbound

This paper cites Bytesareallyou need: End-to-end multilingual speech recognition and synthesis with bytes.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Bytesareallyou need: End-to-end multilingual speech recognition and synthesis with bytes

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.638343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:cbb40f20e6a3033290280a684c24c6754c38f8e904e13ff6abd6b9980f0c7835

Observation 0c01e1a0-242a-4369-9fc8-35d5b94496fe · outbound

This paper cites Natural TTS synthesis by conditioning WaveNet on mel spectrogram predic- tions.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Natural TTS synthesis by conditioning WaveNet on mel spectrogram predic- tions

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.652174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:9e776f419db82534a158e6479eba5477b52a4cda8a138890028a440b59cfb9b4

Observation 1ef0ec5c-fd0e-462e-b864-08100d5ee0b7 · outbound

This paper cites Auto-encodingvariationalBayes.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Auto-encodingvariationalBayes

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.656015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:8465966ab3492117cba04857b0f1f72fb861dd40285a817b243975872cbba665

Observation ca15ffc2-a01b-4c46-9772-50af92048019 · outbound

This paper cites Efficient neural audio synthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Efficient neural audio synthesis

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.695580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:64b7cbeb41ea6f79a01f78621d7a4f6d9adbc9cc3189a60179d0d206e42d5f33

Observation 586a05dc-7512-40bb-af36-0f5f68adfad3 · outbound

This paper cites Char2wav: End-to-endspeechsyn- thesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Char2wav: End-to-endspeechsyn- thesis

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.622605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:18cb4b5a31032725c5f3ce2d532a770f9d0e6a1d9ce047ad6022b98f44ea047f

Observation 79bd0c19-510b-4d92-b149-7d37f90d1262 · outbound

This paper cites Deep Voice 3: Scaling text-to-speech with convolutional sequence learning.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Deep Voice 3: Scaling text-to-speech with convolutional sequence learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.684223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:f4d79e9717080054452570d573f836d86cc30f9d6183a5e19605c35df01793e0

Observation 6f3d1a08-bbbf-4049-b058-c4a0b3db7883 · outbound

This paper cites Representation Mixing for TTS Synthesis.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Representation Mixing for TTS Synthesis

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:10:07.068668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:fba476b69964f3f764adf101abd9cd920241b8e5602dc0bc1d4ca7cde90dde60

Observation bc52e1f1-98cb-413c-8928-22b3002d98e4 · outbound

This paper cites Data-oriented methods for grapheme-to-phoneme conversion.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Data-oriented methods for grapheme-to-phoneme conversion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.631868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:e22f949ff48091b06d3e84d7ecd2d12f9e585b51e0b8a359d611b81d46093714

Observation 99409a6a-0220-44ba-8d48-ba0d1edc939b · outbound

This paper cites Domain- adversarial training of neural networks.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Domain- adversarial training of neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.687837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:51a194e60131c321ddb07119f92e3739e5ce54f69ddcb9cc77bd69abb71915e8

Observation d10ac97c-0db0-4a90-bd2b-2e4c23a7c895 · outbound

This paper cites Disentangling correlated speaker and noise for speech synthesis via data augmentation and adversarial factor- ization.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Disentangling correlated speaker and noise for speech synthesis via data augmentation and adversarial factor- ization

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.691566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:c51cdbe9971699a3aadf3f969a21926c54caad52357dba7300f03b11ea58825c

Observation bd3acc62-3653-41ab-9dda-91fbb8c19d58 · outbound

This paper cites Cross-lingual speaker discrimination usingnaturalandsyntheticspeech.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Cross-lingual speaker discrimination usingnaturalandsyntheticspeech

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.584968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:c00cef0be31265cff4ae7e4c5f773c92cdf49c8ee957ea2413f6157b6af21f2a

Observation 900e2af5-09e5-4d88-b18d-184bac19d2a1 · outbound

This paper cites Generalized end- to-end loss for speaker verification.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Generalized end- to-end loss for speaker verification

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:10:07.580729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:166b797d3d10701dee7b92d1c24d3504d9731443fbd91fa80c49bcf867f29afb

Pith citing papers

Observation 6fd71456-8b9d-4435-b97c-96a539929a35 · inbound

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning cites this paper.

Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T00:10:07.075022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T00:09:35.661121Z digest=sha256:fd3e144026d4c0e159886696fc90d5ac005301e94646a4e24dd3bc053782afe8

Observation d54c8481-fba0-4c4e-bfaa-35dfdcb78861 · inbound

Optimizing Multilingual Text-To-Speech with Accents & Emotions cites this paper.

Optimizing Multilingual Text-To-Speech with Accents & Emotions Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:29.459675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:29.459675Z digest=sha256:37cfc8745eafd327dab5878d855ad6d677a2d205d171950d21d4df56811fd331

Observation b5f15223-498d-4312-b7d4-6d7df3408c0e · inbound

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation cites this paper.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.408435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.408435Z digest=sha256:849866967507669b7c750d8641a763cbdfcaa1336a672163d418de357fbfd666