Pith. sign in

REVIEW 1 cited by

Machine Translation Decoding beyond Beam Search

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2104.05336 v1 pith:YARMFBBF submitted 2021-04-12 cs.CL cs.LG

classification cs.CLcs.LG
keywords searchbeamdecodingmachinemctsmethodmetrictranslation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Beam search is the go-to method for decoding auto-regressive machine translation models. While it yields consistent improvements in terms of BLEU, it is only concerned with finding outputs with high model likelihood, and is thus agnostic to whatever end metric or score practitioners care about. Our aim is to establish whether beam search can be replaced by a more powerful metric-driven search technique. To this end, we explore numerous decoding algorithms, including some which rely on a value function parameterised by a neural network, and report results on a variety of metrics. Notably, we introduce a Monte-Carlo Tree Search (MCTS) based method and showcase its competitiveness. We provide a blueprint for how to use MCTS fruitfully in language applications, which opens promising future directions. We find that which algorithm is best heavily depends on the characteristics of the goal metric; we believe that our extensive experiments and analysis will inform further research in this area.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ScaffoldGPT: A Scaffold-based GPT Model for Drug Optimization

    q-bio.BM 2025-02 reject novelty 5.0 of 10

    A scaffold-prompted GPT with two-phase pretraining, reinforcement fine-tuning, and reward-guided decoding reports improved drug-optimization scores on COVID and cancer benchmarks, though evaluation and training share ...

Pith tools