Pith. sign in

Paper Citation Record · LEDGER

MPO: Multilingual Safety Alignment via Reward Gap Optimization

As of 15 August 2026, this Paper Citation Record lists 95 of 95 outbound references and 1 inbound Pith citation observation for arXiv:2505.16869.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16869 v1

Coverage vector

measured 95 of 95 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:57:35.041116Z

measured 96 of 96 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:40:28.745740Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:40:30.509912Z

Reference resolution

95 of 95 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved94
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 497f8a2f-4be6-44e1-808b-3a73f6558898 · outbound

This paper cites online" 'onlinestring :=.

MPO: Multilingual Safety Alignment via Reward Gap Optimization online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.456676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.456676Z digest=sha256:5c874adff92ce8573b0ca3d11e3dbb8f1d4e29006dc1aeb594affc5a18949cbe

Observation 7d8ff080-d982-4309-874d-06424dada9cc · outbound

This paper cites write newline.

MPO: Multilingual Safety Alignment via Reward Gap Optimization write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.506655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.506655Z digest=sha256:b41b1bda83a2e9eb6f6b604824aa9045f35517942e5f7d75991f94b22ef925d3

Observation 8ba7ea19-07cc-4635-b2b1-c6ff5343566f · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.560690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.560690Z digest=sha256:580cd952cc2e7e5accf05cb038c149b706a8ed30fd70901f40c1a49e839083b1

Observation 101214ea-f3e7-4c4c-8879-a7ccb09cf00a · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.638462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.638462Z digest=sha256:87c1e8e1441c639e4b63a9755522ba758ab911001ebc2eae11f1451097606739

Observation feed01be-d34d-4d8e-953f-9564d0c08325 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.710928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.710928Z digest=sha256:f83be0230ac1bc98c0f28ba4725a0b361bd6d901f4a7d83fdbea873e8e6b3999

Observation 1665e839-c1eb-424d-bf25-d19453b49c01 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.789514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.789514Z digest=sha256:68cc7b974851955bc281f0d58edd2dd08bf09770da82a3c87c7ef188d4630dca

Observation 6dbbf828-86a8-466a-a296-16f26c5d8f29 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.858897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.858897Z digest=sha256:85b61fae27cbf2bbde533de786fb6e098a5ad5ab86ad90fac1cce9bf4490b356

Observation 0fec29b8-de1f-4aa1-b0f0-702c132c6e64 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.928984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.928984Z digest=sha256:7c9da0f1dd07a462e5e2f073503745774c7f5f9f71fb701f1f4b21ed43d054a5

Observation 12f72d56-2192-4907-8718-1538c2549aa7 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.018578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.018578Z digest=sha256:f49a5a01b5e5faf76cb1973fd45614c05ad5809e5a7a4024c480d968c4ef279d

Observation a4d63cd6-ccee-4efe-850f-e0f976fb48aa · outbound

This paper cites High-Dimension Human Value Representation in Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization High-Dimension Human Value Representation in Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.083248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.083248Z digest=sha256:54ad6111fa64e7098099667c2cf13186bb13c7c20e9f301841c49b78c5957c9a

Observation c0f07532-770e-437a-b29d-857be1e23e6a · outbound

This paper cites Towards Scalable Automated Alignment of LLMs: A Survey.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Towards Scalable Automated Alignment of LLMs: A Survey

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.146437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.146437Z digest=sha256:4ee2667635596448887aa16676d02da330bfa832d85033b205ef885b68f7e044

Observation cf229c52-f0ba-4940-9581-9a17bc746fcd · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.231163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.231163Z digest=sha256:1c1c7b186b52ec183d63351025988f2eb5684c6abd9561eab64af2f43f3a23bf

Observation 83067b22-3ba4-4077-ac0e-50b004ede170 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.237978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.237978Z digest=sha256:4612fa044850eca919ba9b989fcd770f73daa2d36a13975b6b136ccab90a7256

Observation e5f97a93-f886-4890-919f-1b733718782f · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.242409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.242409Z digest=sha256:e8bed41d97bc434b867d6a3a5207b014799ba69c10f1f93036bcc037bcaace31

Observation e64fe45c-a05b-42b3-9423-1ce6b4f9afc4 · outbound

This paper cites No Language Left Behind: Scaling Human-Centered Machine Translation.

MPO: Multilingual Safety Alignment via Reward Gap Optimization No Language Left Behind: Scaling Human-Centered Machine Translation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.247094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.247094Z digest=sha256:afe942f169e9879576ff1acd88ec7e24278c4c963946a852d8574aaa9406ccaa

Observation bbf6d2b8-c686-4ffa-98e0-b047983cf746 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.351130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.351130Z digest=sha256:9a12cc371e1eaf4114770c3377c389b23ff843a99abb31e8c73a915aad407414

Observation 082310df-b97d-440d-a0f3-52e56812b61f · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

MPO: Multilingual Safety Alignment via Reward Gap Optimization RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.425024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.425024Z digest=sha256:4127fbd5816a843769cbf665977b78e3801dd3f7f8da4d67b91c64cb07a05a6f

Observation ba87895d-fe0a-472c-8d90-a6c3d2c913ec · outbound

This paper cites The Llama 3 Herd of Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization The Llama 3 Herd of Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.545583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.545583Z digest=sha256:528c6cb2af079e8c1c41199741459c5e9921de2ec1c652c1b4f05df7c2b6b257

Observation 64a5bca5-f1c3-44f1-b169-e5b946ac97e5 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

MPO: Multilingual Safety Alignment via Reward Gap Optimization KTO: Model Alignment as Prospect Theoretic Optimization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.700478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.700478Z digest=sha256:a7623cdba89a63f9b006866e3b85fc509ff4fa3efa61fb96b28a284890233382

Observation 194804bc-7aad-4459-96e8-2fabf0380569 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.829780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.829780Z digest=sha256:3e8a66dd8417a0da4e856f78dc98d855a0f76136bec8138ac7cc570055f05a2d

Observation e9ed2007-fb39-4570-989e-af31e701df1a · outbound

This paper cites LLMs Lost in Translation: M-ALERT uncovers Cross-Linguistic Safety Inconsistencies.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LLMs Lost in Translation: M-ALERT uncovers Cross-Linguistic Safety Inconsistencies

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.999616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.999616Z digest=sha256:f5be1fda2d22e88b01c9029f0025173281198b6dfa3a739cf29d242fc1d98b68

Observation 1519a9e1-e67b-41f0-9275-d2a1a4517c22 · outbound

This paper cites Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.179526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.179526Z digest=sha256:cd34642a98291863ba4717caccbadb4bd4120ecfae62a1c483c1bafa9675ee88

Observation 99d9834b-ceb7-477c-80bf-12b9c591f53f · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MPO: Multilingual Safety Alignment via Reward Gap Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.326794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.326794Z digest=sha256:212803f8512381ad950671c0dd2c251274855c73a97d96474e3656ed8f65590e

Observation 37779663-14c7-4844-8c03-42accb2ae545 · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Direct Language Model Alignment from Online AI Feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.468933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.468933Z digest=sha256:81a051369be768c0f0db78cf9eb3d31b683b55ce75d4cff3e52b6b1a7cf536d3

Observation 27c48240-013e-4924-a6c1-a6c84b511234 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.632423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.632423Z digest=sha256:b88755d56083a0a0212741e264874bd55dcc6f6debf1d89ee84493fe623b4f46

Observation cc08fea9-a032-4b1b-ad7d-9390616c367e · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.777808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.777808Z digest=sha256:b0c590b967be03ea07b1004c63fb29ee116529c8378586888bb1034f9a3f9caf

Observation 85152293-05a5-459a-a684-bd08637f235d · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.925931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.925931Z digest=sha256:d34517d22f038533f9394d7f4497ed5622c3a783d4887d00dd393ec44215366d

Observation f838adc4-e756-4e9d-b604-4ffe7b8e99cb · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:27.095791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:27.095791Z digest=sha256:c831e735a5a6a4cff1287c64714f6e4385ffbe4966bbffa57068ffb38eed3159

Observation f7d3b16e-c508-4aae-a349-25bdfd37f0d1 · outbound

This paper cites Cross-lingual Transfer of Reward Models in Multilingual Alignment.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Cross-lingual Transfer of Reward Models in Multilingual Alignment

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:57:35.859818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:27.179834Z digest=sha256:acd36b9ec234fdd288645df50bf97f9eb0e8b23a94ad9a1dd139485329be5483

Observation 6018cc7a-1e68-4d15-b2b2-eebc2bfd3394 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:42.643686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:27.271141Z digest=sha256:6f2ebf486fb2834520dc32c7a42f39a749b38e9e7af21508e4b73d24a0cbd554

Observation 5473e744-cacd-48aa-9108-265a17e89494 · outbound

This paper cites Large Language Models Are Cross-Lingual Knowledge-Free Reasoners.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Large Language Models Are Cross-Lingual Knowledge-Free Reasoners

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:27.400236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:27.400236Z digest=sha256:4cc897b6edacc6da021ddb5de2c10c2671805adee5270e1375c49989bc729e19

Observation 9fc2c1ab-27eb-43a6-b243-7544099e38f0 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:42.282021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:27.535161Z digest=sha256:c71f866aabc145ac1cb700db3598c32b7a7e8254e3427a7407df19b1dfb0c8bb

Observation 59fbe1ef-b02a-452d-87d1-4cc44a803892 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:41.996094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:27.723880Z digest=sha256:4deb40e3f0154f2329cfbaf20de6a69180b2a8fe660c02065b1f2db2aeb7005d

Observation 795a9b5c-9ae4-401d-9fa3-edf03b080874 · outbound

This paper cites PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference.

MPO: Multilingual Safety Alignment via Reward Gap Optimization PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:27.885966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:27.885966Z digest=sha256:6a7c98ef3f2468c6a2c9ca9bc38ffc8a6304aca5eb054336c75454b7fcccaea8

Observation 77a6aafa-f178-4bd3-81f8-03eed46059fd · outbound

This paper cites Mistral 7B.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Mistral 7B

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:27.997377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:27.997377Z digest=sha256:1bd98bbf56fe9f978d496db7a73ce00bd8f89089554c83e8ce74a3d76bab26df

Observation 056051c0-a66a-4d0d-9f71-50200e298291 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.119956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.119956Z digest=sha256:f489be02387f53db79a7145f603a7445f7e9616fce358560fa1552bfada03723

Observation 6c708738-64a1-4aa6-a72f-3d5b4a647a6c · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:41.722080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:28.280548Z digest=sha256:551e234a053d7b4361ba71d48bf8db2a852f960d20f1c8a647ca55a4b573cda1

Observation 60600cb9-d31f-4b29-a573-cad06695cfcd · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.395465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.395465Z digest=sha256:5681db9a0d972a25f4d65005763ce7c264e87424d0ad62613b10ec8393da3c7f

Observation a4fc7b8f-d925-4471-ba70-edfcde93d668 · outbound

This paper cites A Cross-Language Investigation into Jailbreak Attacks in Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization A Cross-Language Investigation into Jailbreak Attacks in Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.564672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.564672Z digest=sha256:6c7b18c337ee15852d4d367dd357231e0c4bb6dd401097be66a8f7486a9ec39a

Observation f6997dc8-bd11-4a18-84a8-8ffee63b9d80 · outbound

This paper cites XTRUST: On the Multilingual Trustworthiness of Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization XTRUST: On the Multilingual Trustworthiness of Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.698812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.698812Z digest=sha256:fa8cb6721f85f5e0eed5ad3e736c90c36c1ac354717fc832fb6abac69ee0c8d1

Observation 38a24934-5e02-4b88-859c-2fb5741f9951 · outbound

This paper cites Is Translation All You Need? A Study on Solving Multilingual Tasks with Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Is Translation All You Need? A Study on Solving Multilingual Tasks with Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.808252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.808252Z digest=sha256:32c8e6a11f5061cbe0b4986b9b1915f49d58a7e182671ab764e4d56ad40b0bcf

Observation 6852c46f-e6fb-4d63-972d-aac640c962ee · outbound

This paper cites LiPO: Listwise Preference Optimization through Learning-to-Rank.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.962243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.962243Z digest=sha256:df634eb102b4bf3f39c5a0593e245af8a6f7491fe07c553cb13aa153a52a3936

Observation 6756cf52-6e94-4b84-9ade-83efaeb08cc7 · outbound

This paper cites SimPO: Simple Preference Optimization with a Reference-Free Reward.

MPO: Multilingual Safety Alignment via Reward Gap Optimization SimPO: Simple Preference Optimization with a Reference-Free Reward

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:29.123942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:29.123942Z digest=sha256:6e4250bc7aefe1aa29761cb75e9f8dc5f47b3d8d60e2ea1b98b2f95fcc77697b

Observation b342a0da-fb0a-4358-b2ad-f005862cef9a · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:41.424966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:29.232511Z digest=sha256:6e7fd859cca7e87e8b7710b146dbe9de8aeda5787aeeb48ebed5182feecdc51c

Observation b099331d-2fcc-4156-baf3-2a36470ed34d · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:41.127182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:29.400722Z digest=sha256:5164cb50e18d2ea6487c557b16e8117f93db850e0ee57b8dc1abeab306e332a3

Observation a2decc0f-9639-4fac-ade4-88e2134b45ba · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:29.548473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:29.548473Z digest=sha256:ca2b70da2214c82fb9eaa5a01da6eb01d6a12d63bba2d1aecb134e10b0718356

Observation fbf0c15e-e879-4f34-b07e-e854e4c04e5f · outbound

This paper cites Disentangling Length from Quality in Direct Preference Optimization.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Disentangling Length from Quality in Direct Preference Optimization

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:29.682556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:29.682556Z digest=sha256:3938adb05f9141ac37f498502cde958c06eebf148350b7fac6483064c0090d99

Observation ef3b0565-4d01-47e9-94ed-130ff597e838 · outbound

This paper cites Towards Understanding the Fragility of Multilingual LLMs against Fine-Tuning Attacks.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Towards Understanding the Fragility of Multilingual LLMs against Fine-Tuning Attacks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:29.830816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:29.830816Z digest=sha256:b162da479a1a28de5c3c6f691b609b1b8584b6c0f5f98d2c8cfdc1d774cf4488

Observation 384696c4-80c4-4884-bf73-dfba6942dba1 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:40.768641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:29.980995Z digest=sha256:a83088a95ac9f17e345f70349118f364dc21aeaa5be3bf78e9f91ee47b207080

Observation 4f0628ee-8b9d-4e35-8257-3542d0c58060 · outbound

This paper cites Multilingual Large Language Model: A Survey of Resources, Taxonomy and Frontiers.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Multilingual Large Language Model: A Survey of Resources, Taxonomy and Frontiers

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:30.121689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:30.121689Z digest=sha256:b16f702e3f8b4489fc888a8938eaef37ae4ee808e548b757dd22a0ae8f832787

Observation e3675735-891f-41a6-ae51-9e41d4ac6c9e · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:30.298858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:30.298858Z digest=sha256:ed30e26c2272e42f77e8ebb943cda85bd8bd5baf9fdfadff60d22e4df224e6b8

Observation e6d4bbd5-aa18-4d50-9cd0-67a9bf82f9e4 · outbound

This paper cites Empowering Multi-step Reasoning across Languages via Tree-of-Thoughts.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Empowering Multi-step Reasoning across Languages via Tree-of-Thoughts

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:30.449159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:30.449159Z digest=sha256:872d7ab8c659386d07a511a7742b9ea62087dbc532525630e0f79bd7d84546cc

Observation 88739229-8381-4a4e-85a7-b44baa333377 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:40.466984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:30.555277Z digest=sha256:7ce507aa2949c8845fcf073ba10f931d1fa8eded1ed1949a69e54b459587d010

Observation e58de1f1-48ac-4cc7-aad1-602ef4fb59a4 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:40.070819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:30.639912Z digest=sha256:0195702f3f936d7bc073c4546b981c4efdb3ffecba347f61954fe606ef7e285b

Observation c8df08d7-fa7c-40af-90c7-3558c8025d87 · outbound

This paper cites Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:30.769929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:30.769929Z digest=sha256:14b758a81994eb0913c8f4bef64cb810182b7456aa859c9f14ae20072c576200

Observation 8fc3bb9f-9dd6-404b-a077-7627ad067481 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:39.626610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:30.899444Z digest=sha256:9937cf97693ca9a5c09522f14280f0134eb59ade461f65cc51fac9ed679afe88

Observation 4de1f0e2-1706-42eb-8b7b-7bc3a58cc799 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:39.329803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:31.003328Z digest=sha256:7ca2530078ba2cae1b8761f85207ab1d81514e78c4f2f664a5c87cc648ead947

Observation ca5ac8a9-eedd-41b3-9250-b13f49a36cb8 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.111817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.111817Z digest=sha256:37602c2a863ae0ddcb7a427c78c9e06981758486496d94cbac9e77160252691c

Observation 6b1f4f8d-a4ec-484e-8542-d04f3331f5f3 · outbound

This paper cites A Long Way to Go: Investigating Length Correlations in RLHF.

MPO: Multilingual Safety Alignment via Reward Gap Optimization A Long Way to Go: Investigating Length Correlations in RLHF

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.186800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.186800Z digest=sha256:c4fb07247489e1f064641895e7436e1090043d252087c49aa370dfab7ba42fa3

Observation 874b778a-127d-433c-a680-7b50424dacba · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:39.093873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:31.263035Z digest=sha256:67a94015617d8694c05a1e822794d6383479b1d90b8ae646d6a004e74907c72b

Observation a8d89031-663a-45ff-bf7d-17a612ed4ac4 · outbound

This paper cites Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.334455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.334455Z digest=sha256:0852b497b7916438efa27d60e4075e93e4c79a8d520c0b619a9e55f8e7a11b53

Observation 53a3215f-2213-4b49-b398-321db234a88a · outbound

This paper cites A Roadmap to Pluralistic Alignment.

MPO: Multilingual Safety Alignment via Reward Gap Optimization A Roadmap to Pluralistic Alignment

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.395223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.395223Z digest=sha256:abb835d3d2de651fe1392e5829c4ff644d475256fd523e8462eea4405997f589

Observation 9087105a-3423-49a5-bdc5-2a49c9149be9 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Gemma 2: Improving Open Language Models at a Practical Size

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.492367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.492367Z digest=sha256:13ce240307ed45565c67f6a8918078d7f16dea682b7b020a794fd93542a46ee9

Observation e4fcae0e-8798-4418-a31d-f7ec5f352c37 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LLaMA: Open and Efficient Foundation Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.599783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.599783Z digest=sha256:7cd90fd4ba00a220250666ec6ade2a75ec2a9deba223118394d827bec4a626b4

Observation c3bb1a22-1236-41ba-8813-f422d8b6ad35 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.698375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.698375Z digest=sha256:48f58c43c9d9098f0fe8c0b59140e40ca7f5a8fe56abc32427722686dc2133b0

Observation 193c4fe7-28da-4ff6-8839-9e5ff3ab131a · outbound

This paper cites Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.771057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.771057Z digest=sha256:c9b49d31d4e793f43a9849f8c763b4d8012458f9f46741e9ea7836aac19ef332

Observation 6f0334cc-3ba2-4641-91e9-6318de668604 · outbound

This paper cites The Hidden Space of Safety: Understanding Preference-Tuned LLMs in Multilingual context.

MPO: Multilingual Safety Alignment via Reward Gap Optimization The Hidden Space of Safety: Understanding Preference-Tuned LLMs in Multilingual context

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.833322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.833322Z digest=sha256:5e94913f69e3050eaf8c915773903a46cb0e22d5012d8262003d281f0dfdfc53

Observation 987cc69d-84de-4dab-9a7e-544ac5e73b2c · outbound

This paper cites Secrets of RLHF in Large Language Models Part II: Reward Modeling.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Secrets of RLHF in Large Language Models Part II: Reward Modeling

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.942943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.942943Z digest=sha256:9deb1b0f5d9a51953a62bcec23b81f0c669cf46be0059f817ff0b48d6ca8a9fd

Observation f0a359a9-4859-48a0-a56e-c182e408bc33 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.142229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.142229Z digest=sha256:10d0e81ac434b175dc8a4b84ec256aa74922f99ecfbcb1170c8c91840a859c24

Observation ebc27f3d-3bfd-4ef4-8a68-cc192a5a3bd2 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:38.802200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:32.231104Z digest=sha256:01e7e28b7c067281d5e566e3a4ba7254faaaa79b3424ff2870fa5f60d1e74c26

Observation aa0da394-8c4e-45e9-b174-c1d39e70a3b9 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.346784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.346784Z digest=sha256:10ca4dc1b6e674eeffc7aa2cf4fa8f799389c8c23abb8984e0fb11e99fdb336a

Observation e2f53a85-0537-4637-9bee-cbda48e618e1 · outbound

This paper cites AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization.

MPO: Multilingual Safety Alignment via Reward Gap Optimization AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.415979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.415979Z digest=sha256:0e3cd301ff539ed2fa21168ee4a1662935c50b1f12c61285491bcb14e016c4f8

Observation a3c0d95a-bbe3-4757-a3eb-a969bc83ffcc · outbound

This paper cites Self-Play Preference Optimization for Language Model Alignment.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Self-Play Preference Optimization for Language Model Alignment

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.545303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.545303Z digest=sha256:ff7f58325043b85ac889075f216e035416d8277dcfc5c4f8b9e06005f9701f08

Observation 04a79900-4a94-4c93-a31e-935cc8f187b0 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:38.556358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:32.641698Z digest=sha256:e419993a0f329a711a4862bd5501ce2056dd4edf979f6f9c1e3b69b8d67b416e

Observation da0f5f37-1f7c-4240-8dd9-790546c4e401 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:38.263191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:32.766000Z digest=sha256:f96551097eb1af5fe7664a1ad5e6fa16f7d6f82ad2c0860514c71b264523a832

Observation bb475938-3de4-4752-894b-27492d3a2540 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:38.022571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:32.862330Z digest=sha256:2e759d713151ab83b85b855a2f744f07b0e7d2b12d84ba799675d614d6969197

Observation 007606b7-0f54-4940-9364-ad318506c3f5 · outbound

This paper cites Qwen2.5 Technical Report.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Qwen2.5 Technical Report

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.944119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.944119Z digest=sha256:94e3606505a0d09718b8a365642a859668a99af25b419a9a3c96d2d3d39f1a01

Observation fd3fe851-ca50-4459-86c0-eb3284e68f77 · outbound

This paper cites Language Imbalance Driven Rewarding for Multilingual Self-improving.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Language Imbalance Driven Rewarding for Multilingual Self-improving

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.067658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.067658Z digest=sha256:c4d3a0f6f4fe3165682db2618a50e17f527a0e3cfd8647c5b77a6eb857941415

Observation 5c602530-3321-4d28-9ae1-49550853a256 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:37.745551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.135246Z digest=sha256:6aacc75cf0da36ff4eb5794dfe0ca4596e62be807b44a2fc85b23d3a62a9ae26

Observation 459f7de1-50a4-4bb5-9cef-4d70061f5343 · outbound

This paper cites LIMO: Less is More for Reasoning.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LIMO: Less is More for Reasoning

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.230366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.230366Z digest=sha256:346b7f574a6d2bfc58320998902e9e7fdedde074f6b3b299f5acb778d5c19d46

Observation a0e9820d-b1af-4b14-b87b-4f900dcabd39 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:37.446554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.310420Z digest=sha256:e1ce926e2f215543ed1abd481c509493176396a57165af1d8a6b54080650cc5c

Observation bf04a478-53f7-4b9a-a795-3d434d2f7093 · outbound

This paper cites Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual Understanding.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual Understanding

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.437094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.437094Z digest=sha256:14554af2d27985d1e8ec48b8635691022384250e179ea865693e6bc4a00abf2b

Observation ded7f472-70ec-4d01-873b-b688b89c4823 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:37.243436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.534688Z digest=sha256:04548503d96474b57566cd774a62022a5961e355769095c5be8659d4945689ba

Observation 762e112c-6610-4c75-b25f-36d6adf96db2 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.625408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.625408Z digest=sha256:f86550872ed0131c51573cf673be1f8061e212343206c07126bd11ae605e68ca

Observation 242a42a7-3c4e-4fa1-8577-1062d1e4b100 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:37.073604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.723261Z digest=sha256:2d51be21ce79fa11ac869b67098e59adab921510fa871b04bc1b0b2f91a5d117

Observation aa8c5c34-ff28-4625-a657-1dcddd708dd8 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.909689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.844818Z digest=sha256:7ed0a799d81f92a0097103e2e99d9257197e9ccd9261208bd390a0f5530f7c59

Observation 43b95a2c-098d-411b-bbbd-227a6a8f75c3 · outbound

This paper cites Lens: Rethinking Multilingual Enhancement for Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Lens: Rethinking Multilingual Enhancement for Large Language Models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.978484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.978484Z digest=sha256:f854d0b7f22ad2f943ba1b1c6dbfc768578063eda0026b3fdb0434b9d451df99

Observation 0996711a-353a-46f7-8ebf-30c5f8f3ac44 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 88

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.751044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:34.104597Z digest=sha256:26d32504afeed83a39aa2db3bf3f62b34317b0658bc72d5555806e54db0af02c

Observation 18767a1e-5f4c-48c5-a02f-551c7436fa5a · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.593780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:34.237330Z digest=sha256:da992026a013c968befb687c94f37f61f15ad40645b14f2351c60369d1ee228b

Observation dda1918c-bb5b-4f74-bd1a-c2ca0568c75f · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:34.362206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:34.362206Z digest=sha256:26afc0c31f0200a4246f9352f1ebace8b876a9a834920a82ac8db3bfd855ef3e

Observation e49b728a-c5be-4450-b589-1482d296fa00 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.434984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:34.468892Z digest=sha256:259ecbef21e4bb4b1b0455b1a6b3f8e03223b5568ae51340a72c841df16a2d89

Observation ce17b9a4-ba83-4304-98ee-e311c9930fe7 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.264145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:57:34.623028Z digest=sha256:9048d746c1ec5aa0ba7ec7faa4f9dbb880055f332b6ffa71aa0c237f89f7bbd5

Observation 71142e03-05df-428c-8bc1-c5e1e97b0b1f · outbound

This paper cites DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization.

MPO: Multilingual Safety Alignment via Reward Gap Optimization DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:34.760038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:34.760038Z digest=sha256:c77e68a203f7808aaf94897b2203bb477837cd88ecfdc3af515fb9002878ac10

Observation d101f4e2-d9d3-45d7-85ec-90e4347e6dea · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Fine-Tuning Language Models from Human Preferences

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:34.857649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:34.857649Z digest=sha256:abf77cb02473aaa568bebfbcbcd884d8c10790f3dae7d38054ae838ca43b949a

Observation 956e5001-0899-4a2b-8aed-3b38f92ab001 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Representation Engineering: A Top-Down Approach to AI Transparency

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:35.041116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:35.041116Z digest=sha256:db7135c1a7f99593d87467e32f4fa2c78ff5bcf4126acb6c67023a8aa4938237

Pith citing papers

Observation f39c6eb5-946f-4bb9-a14d-4d9151c1402f · inbound

The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It cites this paper.

The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It MPO: Multilingual Safety Alignment via Reward Gap Optimization

Reference 146

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:40:30.600320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:40:28.745740Z digest=sha256:9a146b5dabfdf884b891069da3e735c9c1f0c99bfceac0e692e176c1d6f6ef3d