Pith. sign in

REVIEW

Assessing Evaluation Metrics for Speech-to-Speech Translation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2110.13877 v1 pith:K3E74SCD submitted 2021-10-26 cs.CL cs.SDeess.AS

classification cs.CLcs.SDeess.AS
keywords translationlanguagesspeech-to-speechevaluationstandardizedevaluatemetricspreviously
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Speech-to-speech translation combines machine translation with speech synthesis, introducing evaluation challenges not present in either task alone. How to automatically evaluate speech-to-speech translation is an open question which has not previously been explored. Translating to speech rather than to text is often motivated by unwritten languages or languages without standardized orthographies. However, we show that the previously used automatic metric for this task is best equipped for standardized high-resource languages only. In this work, we first evaluate current metrics for speech-to-speech translation, and second assess how translation to dialectal variants rather than to standardized languages impacts various evaluation methods.

Discussion (0). Continue with ORCID to comment.

Pith tools