REVIEW 3 cited by
SPRING-INX: A Multilingual Indian Language Speech Corpus by SPRING Lab, IIT Madras
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
India is home to a multitude of languages of which 22 languages are recognised by the Indian Constitution as official. Building speech based applications for the Indian population is a difficult problem owing to limited data and the number of languages and accents to accommodate. To encourage the language technology community to build speech based applications in Indian languages, we are open sourcing SPRING-INX data which has about 2000 hours of legally sourced and manually transcribed speech data for ASR system building in Assamese, Bengali, Gujarati, Hindi, Kannada, Malayalam, Marathi, Odia, Punjabi and Tamil. This endeavor is by SPRING Lab , Indian Institute of Technology Madras and is a part of National Language Translation Mission (NLTM), funded by the Indian Ministry of Electronics and Information Technology (MeitY), Government of India. We describe the data collection and data cleaning process along with the data statistics in this paper.
Forward citations
Cited by 3 Pith papers
-
NIRANTAR: Continual Learning with New Languages and Domains on Real-world Speech Data
A real-world continual learning benchmark for multilingual ASR built from 3,250 hours of Indian language speech shows that no current CL method performs consistently across language- and domain-incremental scenarios.
-
Recognizing Every Voice: Towards Inclusive ASR for Rural Bhojpuri Women
Using 25-30 seconds of audio per speaker from 100 rural Bhojpuri women, synthetic speech augmentation cuts ASR word error on the new SRUTI benchmark by 4.7 points.
-
Technical report: Impact of Duration Prediction on Speaker-specific TTS for Indian Languages
In a five-language zero-shot TTS study, no single duration prediction strategy dominates: speaker-prompted durations help some languages, infilling durations help others, and results vary by metric.
Discussion (0). Sign in to comment.