Pith. sign in

A Generalist Audio Foundation Model for Comprehensive Body Sound Auscultation

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Accurate and efficient auscultation-based diagnostics are vital for early disease detection, especially in resource-limited settings where specialized clinical expertise is scarce. Traditional auscultation, which heavily depends on clinician experience, suffers from significant inter-observer variability, while existing AI models often falter due to the limitations of non-representative training data. In this study, we introduce AuscultaBase, a novel AI-driven diagnostic framework that harnesses self-supervised and contrastive learning techniques alongside large-scale, multi-source data integration to advance body sound analysis. By generating robust feature representations, AuscultaBase markedly enhances performance in abnormality detection, disease classification, and activity recognition tasks. Comprehensive evaluations on our newly established benchmark, AuscultaBench, demonstrate that AuscultaBase consistently outperforms state-of-the-art methods across key performance metrics, underscoring its potential as a scalable and cost-effective tool for clinical screening and early disease intervention. The code and model checkpoint has been released in https://github.com/applewpj/AuscultaBase.

citation-role summary

background 1

citation-polarity summary

fields

eess.AS 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

Towards Pre-training an Effective Respiratory Audio Foundation Model

eess.AS · 2025-05-21 · conditional · novelty 5.0

General audio pre-training (AudioSet) outperforms respiratory-specific pre-training for respiratory sound tasks, and further pre-training on combined AudioSet plus respiratory data yields the best results on the OPERA benchmark.

citing papers explorer

Showing 1 of 1 citing paper.

  • Towards Pre-training an Effective Respiratory Audio Foundation Model eess.AS · 2025-05-21 · conditional · none · ref 32 · internal anchor

    General audio pre-training (AudioSet) outperforms respiratory-specific pre-training for respiratory sound tasks, and further pre-training on combined AudioSet plus respiratory data yields the best results on the OPERA benchmark.