Pith. sign in

REVIEW 2 cited by

Adversarial Fine-tuning using Generated Respiratory Sound to Address Class Imbalance

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.06480 v1 pith:Y7XQXM3Q submitted 2023-11-11 cs.SD cs.LGeess.AS

classification cs.SDcs.LGeess.AS
keywords respiratorysoundadversarialfine-tuningdatamethodaddressapproach
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep generative models have emerged as a promising approach in the medical image domain to address data scarcity. However, their use for sequential data like respiratory sounds is less explored. In this work, we propose a straightforward approach to augment imbalanced respiratory sound data using an audio diffusion model as a conditional neural vocoder. We also demonstrate a simple yet effective adversarial fine-tuning method to align features between the synthetic and real respiratory sound samples to improve respiratory sound classification performance. Our experimental results on the ICBHI dataset demonstrate that the proposed adversarial fine-tuning is effective, while only using the conventional augmentation method shows performance degradation. Moreover, our method outperforms the baseline by 2.24% on the ICBHI Score and improves the accuracy of the minority classes up to 26.58%. For the supplementary material, we provide the code at https://github.com/kaen2891/adversarial_fine-tuning_using_generated_respiratory_sound.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles

    cs.SD 2025-05 conditional novelty 5.0 of 10

    Soft-label distillation from same-architecture teacher ensembles improves respiratory sound classification and sets a new ICBHI score of 64.39, though gains are partly due to test-set-based selection of settings.

  2. Adaptive Differential Denoising for Respiratory Sounds Classification

    eess.AS 2025-06 conditional novelty 3.0 of 10

    An Adaptive Differential Denoising network achieves a 65.53% average score on ICBHI 2017 respiratory sound classification, surpassing the previous best by 1.99%.

Pith tools