Pith. sign in

REVIEW 6 cited by

New Trends for Modern Machine Translation with Large Reasoning Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2503.10351 v2 pith:F5Q67YPA submitted 2025-03-13 cs.CL

classification cs.CL
keywords translationlrmsreasoningcontextmodelsbeyondcontextualcultural
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recent advances in Large Reasoning Models (LRMs), particularly those leveraging Chain-of-Thought reasoning (CoT), have opened brand new possibility for Machine Translation (MT). This position paper argues that LRMs substantially transformed traditional neural MT as well as LLMs-based MT paradigms by reframing translation as a dynamic reasoning task that requires contextual, cultural, and linguistic understanding and reasoning. We identify three foundational shifts: 1) contextual coherence, where LRMs resolve ambiguities and preserve discourse structure through explicit reasoning over cross-sentence and complex context or even lack of context; 2) cultural intentionality, enabling models to adapt outputs by inferring speaker intent, audience expectations, and socio-linguistic norms; 3) self-reflection, LRMs can perform self-reflection during the inference time to correct the potential errors in translation especially extremely noisy cases, showing better robustness compared to simply mapping X->Y translation. We explore various scenarios in translation including stylized translation, document-level translation and multimodal translation by showcasing empirical examples that demonstrate the superiority of LRMs in translation. We also identify several interesting phenomenons for LRMs for MT including auto-pivot translation as well as the critical challenges such as over-localisation in translation and inference efficiency. In conclusion, we think that LRMs redefine translation systems not merely as text converters but as multilingual cognitive agents capable of reasoning about meaning beyond the text. This paradigm shift reminds us to think of problems in translation beyond traditional translation scenarios in a much broader context with LRMs - what we can achieve on top of it.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 6 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Translation with Thought: Difficulty-Adaptive Reasoning via Reinforcement Learning for Multi-Domain Machine Translation

    cs.CL 2026-07 conditional novelty 6.0 of 10

    A two-stage SFT+RL recipe (TwT) that allocates reasoning depth by input difficulty matches large reasoning models on auto-metric MT quality with 32-60% fewer tokens, but the auto-metric edge is partly the training objective.

  2. LatentMT: Machine Translation with Latent Reasoning

    cs.CL 2026-07 conditional novelty 6.0 of 10

    A 2.6B looped language model with per-pair LoRA adapters matches or beats 8B-14B MT systems on 32 language pairs, with recurrent-step gains saturating after the first few steps.

  3. TransEvalnia: Reasoning-based Evaluation and Ranking of Translations

    cs.CL 2025-07 conditional novelty 6.0 of 10

    TransEvalnia, a reasoning-based LLM prompt pipeline for translation evaluation, matches or outperforms MT-Ranker on most WMT pairs and produces human-approved explanations.

  4. TAT-R1: Terminology-Aware Translation with Reinforcement Learning and Word Alignment

    cs.CL 2025-05 conditional novelty 6.0 of 10

    Word-alignment rewards for RL-trained translation raise terminology accuracy on RTT from 54.42 to 56.42 TA without hurting general translation quality.

  5. How Well Do Large Reasoning Models Translate? A Comprehensive Evaluation for Multi-Domain Machine Translation

    cs.CL 2025-05 conditional novelty 5.0 of 10

    Large reasoning models such as OpenAI-o1, DeepSeek-R1, and Gemini-2.0-Flash-Thinking score higher than traditional LLMs on semantic quality metrics in complex and document-level translation, but lag on BLEU and in ter...

  6. TULUN: Transparent and Adaptable Low-resource Machine Translation

    cs.CL 2025-05 conditional novelty 5.0 of 10

    An open-source platform that adds glossary and translation-memory-guided LLM post-editing on top of neural machine translation reports large gains on Tetun and Bislama and modest gains on six FLORES languages.

Pith tools