Pith. sign in

REVIEW 1 cited by

Exploring Pre-trained General-purpose Audio Representations for Heart Murmur Detection

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2404.17107 v1 pith:DYINHSJE submitted 2024-04-26 eess.AS cs.SD

classification eess.AScs.SD
keywords heartaudiogeneral-purposelearningpre-trainedavailablerepresentationssound
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

To reduce the need for skilled clinicians in heart sound interpretation, recent studies on automating cardiac auscultation have explored deep learning approaches. However, despite the demands for large data for deep learning, the size of the heart sound datasets is limited, and no pre-trained model is available. On the contrary, many pre-trained models for general audio tasks are available as general-purpose audio representations. This study explores the potential of general-purpose audio representations pre-trained on large-scale datasets for transfer learning in heart murmur detection. Experiments on the CirCor DigiScope heart sound dataset show that the recent self-supervised learning Masked Modeling Duo (M2D) outperforms previous methods with the results of a weighted accuracy of 0.832 and an unweighted average recall of 0.713. Experiments further confirm improved performance by ensembling M2D with other models. These results demonstrate the effectiveness of general-purpose audio representation in processing heart sounds and open the way for further applications. Our code is available online which runs on a 24 GB consumer GPU at https://github.com/nttcslab/m2d/tree/master/app/circor

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Exploring Finetuned Audio-LLM on Heart Murmur Features

    eess.AS 2025-01 reject novelty 5.0 of 10

    A finetuned Qwen2-Audio audio LLM achieves high accuracy on several heart murmur features, but the paper's claim of state-of-the-art performance is contradicted by its own tables for grading and murmur classification.

Pith tools