Pith. sign in

REVIEW

Optimizing Audio Augmentations for Contrastive Learning of Health-Related Acoustic Signals

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.05843 v1 pith:GKZZO2XW submitted 2023-09-11 cs.LG cs.SDeess.AS

classification cs.LGcs.SDeess.AS
keywords audiohealthlearningacousticaugmentationsnfnetslowfastacoustics
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Health-related acoustic signals, such as cough and breathing sounds, are relevant for medical diagnosis and continuous health monitoring. Most existing machine learning approaches for health acoustics are trained and evaluated on specific tasks, limiting their generalizability across various healthcare applications. In this paper, we leverage a self-supervised learning framework, SimCLR with a Slowfast NFNet backbone, for contrastive learning of health acoustics. A crucial aspect of optimizing Slowfast NFNet for this application lies in identifying effective audio augmentations. We conduct an in-depth analysis of various audio augmentation strategies and demonstrate that an appropriate augmentation strategy enhances the performance of the Slowfast NFNet audio encoder across a diverse set of health acoustic tasks. Our findings reveal that when augmentations are combined, they can produce synergistic effects that exceed the benefits seen when each is applied individually.

Discussion (0). Sign in to comment.

Pith tools