Pith. sign in

REVIEW 2 cited by

End-to-end Text-to-speech for Low-resource Languages by Cross-Lingual Transfer Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1904.06508 v2 pith:6HMTTLLS submitted 2019-04-13 cs.CL cs.LGcs.SDeess.AS

classification cs.CLcs.LGcs.SDeess.AS
keywords datalanguageslanguagemappingpairedsourcetargetend-to-end
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

End-to-end text-to-speech (TTS) has shown great success on large quantities of paired text plus speech data. However, laborious data collection remains difficult for at least 95% of the languages over the world, which hinders the development of TTS in different languages. In this paper, we aim to build TTS systems for such low-resource (target) languages where only very limited paired data are available. We show such TTS can be effectively constructed by transferring knowledge from a high-resource (source) language. Since the model trained on source language cannot be directly applied to target language due to input space mismatch, we propose a method to learn a mapping between source and target linguistic symbols. Benefiting from this learned mapping, pronunciation information can be preserved throughout the transferring procedure. Preliminary experiments show that we only need around 15 minutes of paired data to obtain a relatively good TTS system. Furthermore, analytic studies demonstrated that the automatically discovered mapping correlate well with the phonetic expertise.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Characterization of Speech Similarity Between Australian Aboriginal and High-Resource Languages: A Case Study on Dharawal

    eess.AS 2025-09 conditional novelty 5.0 of 10

    Dharawal speech embeddings from a 107-language model rank Latin, Maori, Korean, Thai, and Welsh as the most similar high-resource languages, though confusion-based and geometry-based rankings differ.

  2. LIMBA: An Open-Source Framework for the Preservation and Valorization of Low-Resource Languages using Generative Models

    cs.CL 2024-11 conditional novelty 3.0 of 10

    LIMBA is a proposed pipeline that combines collection, grammatical tagging, translation, speech, and generative modules to build language models for low-resource languages, with preliminary Sardinian experiments.

Pith tools