Pith. sign in

REVIEW 1 cited by

Learning to adapt: a meta-learning approach for speaker adaptation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1808.10239 v1 pith:PZN6CHOI submitted 2018-08-30 cs.CL

classification cs.CL
keywords adaptationadaptingspeakerweightsacousticlhucmeta-learnermeta-learning
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The performance of automatic speech recognition systems can be improved by adapting an acoustic model to compensate for the mismatch between training and testing conditions, for example by adapting to unseen speakers. The success of speaker adaptation methods relies on selecting weights that are suitable for adaptation and using good adaptation schedules to update these weights in order not to overfit to the adaptation data. In this paper we investigate a principled way of adapting all the weights of the acoustic model using a meta-learning. We show that the meta-learner can learn to perform supervised and unsupervised speaker adaptation and that it outperforms a strong baseline adapting LHUC parameters when adapting a DNN AM with 1.5M parameters. We also report initial experiments on adapting TDNN AMs, where the meta-learner achieves comparable performance with LHUC.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. In-Context Learning Boosts Speech Recognition via Human-like Adaptation to Speakers and Language Varieties

    cs.CL 2025-05 conditional novelty 5.0 of 10

    Providing 12 in-context audio-text examples reduces Phi-4-Multimodal's average word error rate by 19.7% relative across English varieties, with the largest gains for low-resource accents.

Pith tools