Pith. sign in

REVIEW 1 cited by

The University of Edinburgh's Neural MT Systems for WMT17

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1708.00726 v1 pith:5K3U25MP submitted 2017-08-02 cs.CL

classification cs.CL
keywords systemstranslationarchitecturesbiomedicalczechdeepedinburghenglish
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This paper describes the University of Edinburgh's submissions to the WMT17 shared news translation and biomedical translation tasks. We participated in 12 translation directions for news, translating between English and Czech, German, Latvian, Russian, Turkish and Chinese. For the biomedical task we submitted systems for English to Czech, German, Polish and Romanian. Our systems are neural machine translation systems trained with Nematus, an attentional encoder-decoder. We follow our setup from last year and build BPE-based models with parallel and back-translated monolingual training data. Novelties this year include the use of deep architectures, layer normalization, and more compact models due to weight tying and improvements in BPE segmentations. We perform extensive ablative experiments, reporting on the effectivenes of layer normalization, deep architectures, and different ensembling techniques.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SEE: Sememe Entanglement Encoding for Transformer-bases Models Compression

    cs.LG 2024-12 conditional novelty 4.0 of 10

    A sememe- and morpheme-based tensor product embedding layer compresses transformer embedding parameters by up to 80x while keeping BLEU close to the uncompressed model.

Pith tools