Pith. sign in

REVIEW 1 cited by

MERLIon CCS Challenge: A English-Mandarin code-switching child-directed speech corpus for language identification and diarization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.18881 v1 pith:TYFBZBAL submitted 2023-05-30 eess.AS

classification eess.AS
keywords corpuslanguagespeechchallengemerlionchild-directedcode-switchingdiarization
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

To enhance the reliability and robustness of language identification (LID) and language diarization (LD) systems for heterogeneous populations and scenarios, there is a need for speech processing models to be trained on datasets that feature diverse language registers and speech patterns. We present the MERLIon CCS challenge, featuring a first-of-its-kind Zoom video call dataset of parent-child shared book reading, of over 30 hours with over 300 recordings, annotated by multilingual transcribers using a high-fidelity linguistic transcription protocol. The audio corpus features spontaneous and in-the-wild English-Mandarin code-switching, child-directed speech in non-standard accents with diverse language-mixing patterns recorded in a variety of home environments. This report describes the corpus, as well as LID and LD results for our baseline and several systems submitted to the MERLIon CCS challenge using the corpus.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Speaker Role and Language Diarization for Analyzing Multilingual Interviews for Language Proficiency of Older Adults

    eess.AS 2026-08 conditional novelty 6.0 of 10

    In multilingual interviews with older adults in India, automatically extracted respondent speech ratio and intended-language usage predict human proficiency ratings, and the pattern survives fully automatic diarization.

Pith tools