Pith. sign in

REVIEW 2 cited by

Do Multilingual Language Models Capture Differing Moral Norms?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2203.09904 v1 pith:NFXUN4GH submitted 2022-03-18 cs.CL

classification cs.CL
keywords languagesmodelsmoralmultilingualnormspotentiallycapturedata
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Massively multilingual sentence representations are trained on large corpora of uncurated data, with a very imbalanced proportion of languages included in the training. This may cause the models to grasp cultural values including moral judgments from the high-resource languages and impose them on the low-resource languages. The lack of data in certain languages can also lead to developing random and thus potentially harmful beliefs. Both these issues can negatively influence zero-shot cross-lingual model transfer and potentially lead to harmful outcomes. Therefore, we aim to (1) detect and quantify these issues by comparing different models in different languages, (2) develop methods for improving undesirable properties of the models. Our initial experiments using the multilingual model XLM-R show that indeed multilingual LMs capture moral norms, even with potentially higher human-agreement than monolingual ones. However, it is not yet clear to what extent these moral norms differ between languages.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Lost in the Tower of Babel: The Adverse Effects of Incidental Multilingualism in LLMs

    cs.CL 2026-05 unverdicted novelty 6.0 of 10

    Incidental multilingualism from uneven web training makes LLMs unequal, brittle, and opaque across languages.

  2. Large Language Models as Mirrors of Societal Moral Standards

    cs.AI 2024-12 conditional novelty 3.0 of 10

    Current open language models, especially English-centric GPT-2 and OPT, show weak or negative correlations with cross-national moral survey values, and the best model, BLOOMZ-560M, reaches only small positive correlations.

Pith tools