Pith. sign in

REVIEW 1 cited by

Generative Deep Learning and Signal Processing for Data Augmentation of Cardiac Auscultation Signals: Improving Model Robustness Using Synthetic Audio

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.10125 v2 pith:EOWXLELV submitted 2024-10-14 cs.SD eess.ASeess.SP

classification cs.SDeess.ASeess.SP
keywords modelperformanceaudioaugmentedauscultationcardiacclassificationdata
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Accurately interpreting cardiac auscultation signals plays a crucial role in diagnosing and managing cardiovascular diseases. However, the paucity of labelled data inhibits classification models' training. Researchers have turned to generative deep learning techniques combined with signal processing to augment the existing data and improve cardiac auscultation classification models to overcome this challenge. However, the primary focus of prior studies has been on model performance as opposed to model robustness. Robustness, in this case, is defined as both the in-distribution and out-of-distribution performance by measures such as Matthew's correlation coefficient. This work shows that more robust abnormal heart sound classifiers can be trained using an augmented dataset. The augmentations consist of traditional audio approaches and the creation of synthetic audio conditionally generated using the WaveGrad and DiffWave diffusion models. It is found that both the in-distribution and out-of-distribution performance can be improved over various datasets when training a convolutional neural network-based classification model with this augmented dataset. With the performance increase encompassing not only accuracy but also balanced accuracy and Matthew's correlation coefficient, an augmented dataset significantly contributes to resolving issues of imbalanced datasets. This, in turn, helps provide a more general and robust classifier.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Scaling to Multimodal and Multichannel Heart Sound Classification with Synthetic and Augmented Biosignals

    cs.SD 2025-09 conditional novelty 5.0 of 10

    Fine-tuning Wav2Vec2 on diffusion-generated and augmented biosignals improves abnormal heart sound classification across single-channel, PCG+ECG, and multichannel PCG data, with reported state-of-the-art benchmark numbers.

Pith tools