REVIEW 5 cited by
Coswara -- A Database of Breathing, Cough, and Voice Sounds for COVID-19 Diagnosis
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The COVID-19 pandemic presents global challenges transcending boundaries of country, race, religion, and economy. The current gold standard method for COVID-19 detection is the reverse transcription polymerase chain reaction (RT-PCR) testing. However, this method is expensive, time-consuming, and violates social distancing. Also, as the pandemic is expected to stay for a while, there is a need for an alternate diagnosis tool which overcomes these limitations, and is deployable at a large scale. The prominent symptoms of COVID-19 include cough and breathing difficulties. We foresee that respiratory sounds, when analyzed using machine learning techniques, can provide useful insights, enabling the design of a diagnostic tool. Towards this, the paper presents an early effort in creating (and analyzing) a database, called Coswara, of respiratory sounds, namely, cough, breath, and voice. The sound samples are collected via worldwide crowdsourcing using a website application. The curated dataset is released as open access. As the pandemic is evolving, the data collection and analysis is a work in progress. We believe that insights from analysis of Coswara can be effective in enabling sound based technology solutions for point-of-care diagnosis of respiratory infection, and in the near future this can help to diagnose COVID-19.
Forward citations
Cited by 5 Pith papers
-
HEARTS: Benchmarking LLM Reasoning on Health Time Series
A 110-task benchmark across 20 health signal modalities shows current LLMs underperform specialized models and depend on simple heuristics rather than robust time-series reasoning.
-
HPP-Voice: A Large-Scale Evaluation of Speech Embeddings for Multi-Phenotypic Classification
A 30-second counting task, embedded with speaker-identification models, predicts male sleep apnea (AUC 0.64) and shows gender- and condition-specific model rankings across a new 7,188-recording clinical speech benchmark.
-
CoughViT: A Self-Supervised Vision Transformer for Cough Audio Representation Learning
A self-supervised masked-spectrogram pretraining method for cough audio produces representations that match or exceed AudioSet-pretrained AST on three cough classification tasks.
-
GeHirNet: A Gender-Aware Hierarchical Model for Voice Pathology Classification
A gender-aware two-stage classifier using ResNet-50 on Mel spectrograms achieves 97.63% accuracy for six voice pathologies across four public datasets.
-
Cough Classification using Few-Shot Learning
A prototypical-network few-shot classifier on cough spectrograms reaches about 72% three-class accuracy and is deemed equivalent to binary classifiers within a generous 15-point margin.
Discussion (0). Continue with ORCID to comment.