Using Whisper to auto-accept recordings that transcribe perfectly, and crowdsourcing only the rest, cut validation costs by about 43% in a German speech dataset with quality close to a fully crowd-validated French dataset.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Speech Foundation Models and Crowdsourcing for Efficient, High-Quality Data Collection
Using Whisper to auto-accept recordings that transcribe perfectly, and crowdsourcing only the rest, cut validation costs by about 43% in a German speech dataset with quality close to a fully crowd-validated French dataset.