Pith. sign in

REVIEW 2 cited by

An Efficient Approach for Studying Cross-Lingual Transfer in Multilingual Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.20088 v1 pith:WHDR3AFH submitted 2024-03-29 cs.CL

classification cs.CL
keywords languagelanguagestransfertargetapproachmultilingualbeneficialconsistently
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The capacity and effectiveness of pre-trained multilingual models (MLMs) for zero-shot cross-lingual transfer is well established. However, phenomena of positive or negative transfer, and the effect of language choice still need to be fully understood, especially in the complex setting of massively multilingual LMs. We propose an \textit{efficient} method to study transfer language influence in zero-shot performance on another target language. Unlike previous work, our approach disentangles downstream tasks from language, using dedicated adapter units. Our findings suggest that some languages do not largely affect others, while some languages, especially ones unseen during pre-training, can be extremely beneficial or detrimental for different target languages. We find that no transfer language is beneficial for all target languages. We do, curiously, observe languages previously unseen by MLMs consistently benefit from transfer from almost any language. We additionally use our modular approach to quantify negative interference efficiently and categorize languages accordingly. Furthermore, we provide a list of promising transfer-target language configurations that consistently lead to target language performance improvements. Code and data are publicly available: https://github.com/ffaisal93/neg_inf

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Training Bilingual LMs with Data Constraints in the Targeted Language

    cs.CL 2024-11 conditional novelty 6.0 of 10

    Higher-quality auxiliary English pretraining data improves target-language performance for languages close to English (about 2% on translated QA tasks), but not for distant languages, when target-language data is limi...

  2. Generative AI and linguistic diversity in academic writing and publishing: Perspectives from World Englishes

    cs.CL 2026-07 conditional novelty 4.0 of 10

    GenAI in academic writing reinforces dominant English hierarchies while remaining a possible site of resistance if designed, governed, and used for linguistic diversity.

Pith tools