Pith. sign in

Paper Citation Record · LEDGER

MPO: Multilingual Safety Alignment via Reward Gap Optimization

As of 8 August 2026, this Paper Citation Record lists 95 of 95 outbound references and 1 inbound Pith citation observation for arXiv:2505.16869.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16869 v1

Coverage vector

measured 95 of 95 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:57:35.041116Z

measured 96 of 96 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:40:28.745740Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:40:30.509912Z

Reference resolution

95 of 95 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved94
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 497f8a2f-4be6-44e1-808b-3a73f6558898 · outbound

This paper cites online" 'onlinestring :=.

MPO: Multilingual Safety Alignment via Reward Gap Optimization online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.456676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.456676Z digest=sha256:f136dd0bdce9b7fe3995fd9ab5283397f478821e5fe454b2212b3193fc48b6b4

Observation 7d8ff080-d982-4309-874d-06424dada9cc · outbound

This paper cites write newline.

MPO: Multilingual Safety Alignment via Reward Gap Optimization write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.506655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.506655Z digest=sha256:96d277485b7fd9a473da0c6580a8baf12175f6f88edeb6cf767f6169cd5e937b

Observation 8ba7ea19-07cc-4635-b2b1-c6ff5343566f · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.560690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.560690Z digest=sha256:fc441485802e2ab85034e583c48accf92d04423e083d2b544f60977a94ec70f4

Observation 101214ea-f3e7-4c4c-8879-a7ccb09cf00a · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.638462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.638462Z digest=sha256:8189a9fccf51afa5f28db2d844fa6d9e1d301176ad18cf047d6e48fc90a0d3ad

Observation feed01be-d34d-4d8e-953f-9564d0c08325 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.710928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.710928Z digest=sha256:b109ccdcb6cd91306d96fa24ddeaaa130a8b1abb7a9538337639575fd8b433bf

Observation 1665e839-c1eb-424d-bf25-d19453b49c01 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.789514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.789514Z digest=sha256:0af60359699901c92abd42579fd701893277efe4f8c75afb82904515998c9ab1

Observation 6dbbf828-86a8-466a-a296-16f26c5d8f29 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.858897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.858897Z digest=sha256:53ceeee9d5004a68ae9506397029bc6500151e98f5c01ab324b267591eb83284

Observation 0fec29b8-de1f-4aa1-b0f0-702c132c6e64 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:24.928984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:24.928984Z digest=sha256:c0f2ad8344c45f7545d0bbe9f886b98b1ba4bc7b89504931d30768f28c2f8cb0

Observation 12f72d56-2192-4907-8718-1538c2549aa7 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.018578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.018578Z digest=sha256:b8986c13cfc6da7f60da8d768c8eac0f3a7e426d4b8452ef4dfc44e6ac444281

Observation a4d63cd6-ccee-4efe-850f-e0f976fb48aa · outbound

This paper cites High-Dimension Human Value Representation in Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization High-Dimension Human Value Representation in Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.083248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.083248Z digest=sha256:6d1dcc2e1eb1a34e14ac1e3b2e6365158a521a1a2bd10fb2f6298bf7c5e8ee62

Observation c0f07532-770e-437a-b29d-857be1e23e6a · outbound

This paper cites Towards Scalable Automated Alignment of LLMs: A Survey.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Towards Scalable Automated Alignment of LLMs: A Survey

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.146437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.146437Z digest=sha256:6c3384ab2325d629b2dd983076ab20c2b0e5aa9c8b4d3bfee9f85bc794a96029

Observation cf229c52-f0ba-4940-9581-9a17bc746fcd · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.231163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.231163Z digest=sha256:b3809eefad7c7f7498d63301db8b6b55e0421bc053592c2bc2090abf99344713

Observation 83067b22-3ba4-4077-ac0e-50b004ede170 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.237978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.237978Z digest=sha256:2033cea192589be521e88bd01ab837ecbeef5afad556d7e068e2838d19beb509

Observation e5f97a93-f886-4890-919f-1b733718782f · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.242409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.242409Z digest=sha256:181dc669a18a3c40c4b5a840a8728a279d75686485808363b663d4ea70ba10fb

Observation e64fe45c-a05b-42b3-9423-1ce6b4f9afc4 · outbound

This paper cites No Language Left Behind: Scaling Human-Centered Machine Translation.

MPO: Multilingual Safety Alignment via Reward Gap Optimization No Language Left Behind: Scaling Human-Centered Machine Translation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.247094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.247094Z digest=sha256:29501f34c0ebcd723e30792da47366ae9a9819487dcfeef2765fe7c0a125027b

Observation bbf6d2b8-c686-4ffa-98e0-b047983cf746 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.351130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.351130Z digest=sha256:6fb18205b44d3705e021fd110affa6fca073ba38be686c7225d75bf433f16619

Observation 082310df-b97d-440d-a0f3-52e56812b61f · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

MPO: Multilingual Safety Alignment via Reward Gap Optimization RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.425024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.425024Z digest=sha256:ab27b74f59ca521074f8373ca8b16087c4b5b22e366e5a7a3fe21b320bb0a369

Observation ba87895d-fe0a-472c-8d90-a6c3d2c913ec · outbound

This paper cites The Llama 3 Herd of Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization The Llama 3 Herd of Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.545583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.545583Z digest=sha256:a6fa4a9d16e88d621f621871bb2b0d69e096285e87bff714d9718aef4c050554

Observation 64a5bca5-f1c3-44f1-b169-e5b946ac97e5 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

MPO: Multilingual Safety Alignment via Reward Gap Optimization KTO: Model Alignment as Prospect Theoretic Optimization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.700478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.700478Z digest=sha256:ffd22253c57964459acd7fa48fe5657cd6dc15f62a15495c5b6bbadb6e27322a

Observation 194804bc-7aad-4459-96e8-2fabf0380569 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.829780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.829780Z digest=sha256:cd97af89d68685008201b1e93115b2b36aa39d88d5ccfa54f00dd4cc4514a624

Observation e9ed2007-fb39-4570-989e-af31e701df1a · outbound

This paper cites LLMs Lost in Translation: M-ALERT uncovers Cross-Linguistic Safety Inconsistencies.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LLMs Lost in Translation: M-ALERT uncovers Cross-Linguistic Safety Inconsistencies

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:25.999616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:25.999616Z digest=sha256:196f7211c3ba6894853f64ff6065f8c93ba6c63272d929f5840702bcdd1c790f

Observation 1519a9e1-e67b-41f0-9275-d2a1a4517c22 · outbound

This paper cites Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.179526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.179526Z digest=sha256:f3b86cd82d060608179d44f652c3504fbf1a2941490efdea1f9d7d70023e1eea

Observation 99d9834b-ceb7-477c-80bf-12b9c591f53f · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MPO: Multilingual Safety Alignment via Reward Gap Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.326794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.326794Z digest=sha256:a6e3dbb55345324cd97867c4a34c25a1adee1ad382032c5899fe8b9bdbb8fd66

Observation 37779663-14c7-4844-8c03-42accb2ae545 · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Direct Language Model Alignment from Online AI Feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.468933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.468933Z digest=sha256:2b73163cbe2e00fe6bde48dea281ee66d6c34e400f236b0d0649669e501e1b7f

Observation 27c48240-013e-4924-a6c1-a6c84b511234 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.632423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.632423Z digest=sha256:b26714e7f90f29b12be98f874ca3a8c71788d56087c8c82d02ab09439110ce77

Observation cc08fea9-a032-4b1b-ad7d-9390616c367e · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.777808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.777808Z digest=sha256:71f241d1548320e48ef24227064712a98bc6377c8959094ae2ad2ede6ab9f2a9

Observation 85152293-05a5-459a-a684-bd08637f235d · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:26.925931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:26.925931Z digest=sha256:8ea927e4a03fd39079fdfe2d98a9eed97ab147f2899501589bb4bdfa7e2f36c1

Observation f838adc4-e756-4e9d-b604-4ffe7b8e99cb · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:27.095791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:27.095791Z digest=sha256:3f0b08a8ff46d1a778575d6df031930358b61ae46745a1adf9ac533f516a0856

Observation f7d3b16e-c508-4aae-a349-25bdfd37f0d1 · outbound

This paper cites Cross-lingual Transfer of Reward Models in Multilingual Alignment.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Cross-lingual Transfer of Reward Models in Multilingual Alignment

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:57:35.859818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:27.179834Z digest=sha256:5458d717772dce89c66b352512a14a35eefcbf825aadff946ccbc364518bc626

Observation 6018cc7a-1e68-4d15-b2b2-eebc2bfd3394 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:42.643686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:27.271141Z digest=sha256:d0e6ba10397a183a4c5042d9d28f0ea14567499ffc73f5aa641f5955c03169ba

Observation 5473e744-cacd-48aa-9108-265a17e89494 · outbound

This paper cites Large Language Models Are Cross-Lingual Knowledge-Free Reasoners.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Large Language Models Are Cross-Lingual Knowledge-Free Reasoners

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:27.400236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:27.400236Z digest=sha256:2651e3d39112a30b660a611f59f4dc4e9572489b00e4d6f18badce1b90144167

Observation 9fc2c1ab-27eb-43a6-b243-7544099e38f0 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:42.282021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:27.535161Z digest=sha256:695de1c65b8ae2229f1b81fe97372793a6c8f817d1cfc8f3cc1aca4f7025415a

Observation 59fbe1ef-b02a-452d-87d1-4cc44a803892 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:41.996094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:27.723880Z digest=sha256:1b9c465e8bcbc792ffe6e00ebb668a9912e9bb1f824c6715ff15f74f480ba9de

Observation 795a9b5c-9ae4-401d-9fa3-edf03b080874 · outbound

This paper cites PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference.

MPO: Multilingual Safety Alignment via Reward Gap Optimization PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:27.885966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:27.885966Z digest=sha256:6590d70609123ed615b27a8361766f513008370d3144021cbfb54433be06252e

Observation 77a6aafa-f178-4bd3-81f8-03eed46059fd · outbound

This paper cites Mistral 7B.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Mistral 7B

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:27.997377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:27.997377Z digest=sha256:448ffac4eddd8fc4944b2607cde4af38b37cd88eb0a2df87632184033c628063

Observation 056051c0-a66a-4d0d-9f71-50200e298291 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.119956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.119956Z digest=sha256:ed9506714aef336c00b597cbca9d1cd02ac919c23f4c3f9483bdb4251c65766c

Observation 6c708738-64a1-4aa6-a72f-3d5b4a647a6c · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:41.722080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:28.280548Z digest=sha256:34b7a6a643a930832006ede04ae96e48bc9297c5bbcadcc7420567e05da7d446

Observation 60600cb9-d31f-4b29-a573-cad06695cfcd · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.395465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.395465Z digest=sha256:ebfacac5ffc9ce507f604e3fc5c89ca86ec7e3ebe5e876c56b6b254b85ff80c4

Observation a4fc7b8f-d925-4471-ba70-edfcde93d668 · outbound

This paper cites A Cross-Language Investigation into Jailbreak Attacks in Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization A Cross-Language Investigation into Jailbreak Attacks in Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.564672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.564672Z digest=sha256:e35acf01948362f143faa12c591584c1c639a952004f8cbbacf99ff2db1dc367

Observation f6997dc8-bd11-4a18-84a8-8ffee63b9d80 · outbound

This paper cites XTRUST: On the Multilingual Trustworthiness of Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization XTRUST: On the Multilingual Trustworthiness of Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.698812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.698812Z digest=sha256:91e151de174cb47179244c6f27df9dc61c7ef7fbaa56475c182f2ce5257d7f58

Observation 38a24934-5e02-4b88-859c-2fb5741f9951 · outbound

This paper cites Is Translation All You Need? A Study on Solving Multilingual Tasks with Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Is Translation All You Need? A Study on Solving Multilingual Tasks with Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.808252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.808252Z digest=sha256:d4ff12b11cab7bd0af136f81f6280b4f72cd73921e4f094e11ac39f4d0c06ce6

Observation 6852c46f-e6fb-4d63-972d-aac640c962ee · outbound

This paper cites LiPO: Listwise Preference Optimization through Learning-to-Rank.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.962243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.962243Z digest=sha256:d0f7421fbc0f2c306c5e90707611f2c779294a7f3d87e35ef813c493e80b0ae4

Observation 6756cf52-6e94-4b84-9ade-83efaeb08cc7 · outbound

This paper cites SimPO: Simple Preference Optimization with a Reference-Free Reward.

MPO: Multilingual Safety Alignment via Reward Gap Optimization SimPO: Simple Preference Optimization with a Reference-Free Reward

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:29.123942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:29.123942Z digest=sha256:c2a50c5721539028e725055fb719499c31beb54f7b1881b219125f8589d117c2

Observation b342a0da-fb0a-4358-b2ad-f005862cef9a · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:41.424966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:29.232511Z digest=sha256:cde0d90958d0ebc02e4eabe208080de64c51514c9d778adf5d3f45c8b5c17c51

Observation b099331d-2fcc-4156-baf3-2a36470ed34d · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:41.127182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:29.400722Z digest=sha256:dde733c47cd7f11f1dd058af163e7b5b6cc2d3d4d46134c37f662ed93dc0a872

Observation a2decc0f-9639-4fac-ade4-88e2134b45ba · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:29.548473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:29.548473Z digest=sha256:00fe1a471c547302918fe6779994a8ec5d10c3ae8ed48de56683d6d248bfd06f

Observation fbf0c15e-e879-4f34-b07e-e854e4c04e5f · outbound

This paper cites Disentangling Length from Quality in Direct Preference Optimization.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Disentangling Length from Quality in Direct Preference Optimization

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:29.682556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:29.682556Z digest=sha256:00e2bf19f084c293fa3b72203ef8ec15e307208252f78eebaeb71425db667272

Observation ef3b0565-4d01-47e9-94ed-130ff597e838 · outbound

This paper cites Towards Understanding the Fragility of Multilingual LLMs against Fine-Tuning Attacks.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Towards Understanding the Fragility of Multilingual LLMs against Fine-Tuning Attacks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:29.830816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:29.830816Z digest=sha256:8363d0f935229de2c697dbc4b5d6f4fae8086393dbcbda2ef824590759868fb8

Observation 384696c4-80c4-4884-bf73-dfba6942dba1 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:40.768641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:29.980995Z digest=sha256:037500e2f6a47bd1b1615c4736b9d0984e81565eca089a3738c1eadc53cc5f11

Observation 4f0628ee-8b9d-4e35-8257-3542d0c58060 · outbound

This paper cites Multilingual Large Language Model: A Survey of Resources, Taxonomy and Frontiers.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Multilingual Large Language Model: A Survey of Resources, Taxonomy and Frontiers

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:30.121689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:30.121689Z digest=sha256:881acc262ea4a8ea3e5d61d05d6ae186966dcd136a7560c1391581fa9909bdbf

Observation e3675735-891f-41a6-ae51-9e41d4ac6c9e · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:30.298858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:30.298858Z digest=sha256:56688a1405ab2b1d1a0b5b442c4f56544089eaac7b8c98db246a5517c593a69d

Observation e6d4bbd5-aa18-4d50-9cd0-67a9bf82f9e4 · outbound

This paper cites Empowering Multi-step Reasoning across Languages via Tree-of-Thoughts.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Empowering Multi-step Reasoning across Languages via Tree-of-Thoughts

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:30.449159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:30.449159Z digest=sha256:29f44ed3aca03c8600f909743fc40196e2533c4df90218d90057bb3be7a3afc1

Observation 88739229-8381-4a4e-85a7-b44baa333377 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:40.466984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:30.555277Z digest=sha256:a2eff25750e9502cee576a028379eebf43aa65b09b7d25b50a4af8a1251e2eb9

Observation e58de1f1-48ac-4cc7-aad1-602ef4fb59a4 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:40.070819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:30.639912Z digest=sha256:a7fd590eb308aa9a531eb0721a1ac3b7f267d808c9fab4c59c7d26608f110654

Observation c8df08d7-fa7c-40af-90c7-3558c8025d87 · outbound

This paper cites Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:30.769929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:30.769929Z digest=sha256:d076ad63cc9d749a20ba594d0bd7bcd23b6816d3bc83a200e577d67b5de395c9

Observation 8fc3bb9f-9dd6-404b-a077-7627ad067481 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:39.626610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:30.899444Z digest=sha256:1b06ad7c8333afdc66d747528641f60ce1c350d35a5d39b58acb762ee150da6e

Observation 4de1f0e2-1706-42eb-8b7b-7bc3a58cc799 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:39.329803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:31.003328Z digest=sha256:771e1704158927023daebbd5fd5a93045a7843159a19b8a2995e9908763fdfb6

Observation ca5ac8a9-eedd-41b3-9250-b13f49a36cb8 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.111817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.111817Z digest=sha256:aa43c5bc92eb15b726d737477a55aa57264614f5438a7486925e484562d31dec

Observation 6b1f4f8d-a4ec-484e-8542-d04f3331f5f3 · outbound

This paper cites A Long Way to Go: Investigating Length Correlations in RLHF.

MPO: Multilingual Safety Alignment via Reward Gap Optimization A Long Way to Go: Investigating Length Correlations in RLHF

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.186800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.186800Z digest=sha256:9f83ff37eb050fd6d76066b889a2511cf75a9947788e287e664007c3f7cb904f

Observation 874b778a-127d-433c-a680-7b50424dacba · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:39.093873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:31.263035Z digest=sha256:f029af27b8917c68ce605c6b0f71c0825ef4a4143e451f7de2f7ba3caf2bcb88

Observation a8d89031-663a-45ff-bf7d-17a612ed4ac4 · outbound

This paper cites Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.334455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.334455Z digest=sha256:97da0d06f124a8981540e307f4e56b63b5718c574cdb791d34bcd1a373e3eab9

Observation 53a3215f-2213-4b49-b398-321db234a88a · outbound

This paper cites A Roadmap to Pluralistic Alignment.

MPO: Multilingual Safety Alignment via Reward Gap Optimization A Roadmap to Pluralistic Alignment

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.395223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.395223Z digest=sha256:d9b82b44d373f64ee43571c69dcb23a7817fc88d446a6318ff37ce99e6c7df30

Observation 9087105a-3423-49a5-bdc5-2a49c9149be9 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Gemma 2: Improving Open Language Models at a Practical Size

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.492367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.492367Z digest=sha256:4f59d049f7f26d25a61069c4cd276843a0eef3fba095d44295db98065dc1d10e

Observation e4fcae0e-8798-4418-a31d-f7ec5f352c37 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LLaMA: Open and Efficient Foundation Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.599783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.599783Z digest=sha256:0075878e4e665f0c11fa5c95866094d3a11ea204968f805c9fc85f93784c5608

Observation c3bb1a22-1236-41ba-8813-f422d8b6ad35 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.698375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.698375Z digest=sha256:81057cc902a53b57de82cbd28ca39b3885ae9a26d902cceabb1da0e8ca782e9d

Observation 193c4fe7-28da-4ff6-8839-9e5ff3ab131a · outbound

This paper cites Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.771057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.771057Z digest=sha256:6b680520d80b1a627609fff0b40ace3eaf645612e6729dd99f0ca57cbfc021f6

Observation 6f0334cc-3ba2-4641-91e9-6318de668604 · outbound

This paper cites The Hidden Space of Safety: Understanding Preference-Tuned LLMs in Multilingual context.

MPO: Multilingual Safety Alignment via Reward Gap Optimization The Hidden Space of Safety: Understanding Preference-Tuned LLMs in Multilingual context

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.833322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.833322Z digest=sha256:5d5f56b60f2a03797e4bcc2fa0a834c44a379d83638e11c55b7d1a0c1bbf85e5

Observation 987cc69d-84de-4dab-9a7e-544ac5e73b2c · outbound

This paper cites Secrets of RLHF in Large Language Models Part II: Reward Modeling.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Secrets of RLHF in Large Language Models Part II: Reward Modeling

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:31.942943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:31.942943Z digest=sha256:f0784441f6bc83d539cdd08653cdc7cfb43b99218a0a40a2cd4ce8669fc8229d

Observation f0a359a9-4859-48a0-a56e-c182e408bc33 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.142229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.142229Z digest=sha256:217c636051a9efedf03742a6faade207f87ff5e2e468a4a4b01d2895edae77a4

Observation ebc27f3d-3bfd-4ef4-8a68-cc192a5a3bd2 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:38.802200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:32.231104Z digest=sha256:2aefe82a8414fdffcc65c28f5ebee2c18e4dcde68ec1b5adaef65f40ee5b3a01

Observation aa0da394-8c4e-45e9-b174-c1d39e70a3b9 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.346784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.346784Z digest=sha256:6073f963f1d2d1ad5eccf3fba2e053961512a1a6789d22fc3130328cce8f86d4

Observation e2f53a85-0537-4637-9bee-cbda48e618e1 · outbound

This paper cites AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization.

MPO: Multilingual Safety Alignment via Reward Gap Optimization AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.415979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.415979Z digest=sha256:0e570405bd6f9c4508dc381a5e87cd38dae61bcd424f760fbd4d0e5245a6a1f1

Observation a3c0d95a-bbe3-4757-a3eb-a969bc83ffcc · outbound

This paper cites Self-Play Preference Optimization for Language Model Alignment.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Self-Play Preference Optimization for Language Model Alignment

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.545303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.545303Z digest=sha256:1d4e419051d5a4ffa1aa2b3560d3d760f5926c794190c2774f569d8916da94a9

Observation 04a79900-4a94-4c93-a31e-935cc8f187b0 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:38.556358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:32.641698Z digest=sha256:53d3dbfc3ee3d6c7abcff8ae6e7fc1d2f4eb91c3718de8c69bc771f9a292d10e

Observation da0f5f37-1f7c-4240-8dd9-790546c4e401 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:38.263191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:32.766000Z digest=sha256:81025305438bf892453f2b1d06d21bb7b8ad29d357ee4b0514f6c6f1dec96fdd

Observation bb475938-3de4-4752-894b-27492d3a2540 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:38.022571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:32.862330Z digest=sha256:d09dd15628dd449af12fa5465f901711d0780064b8c5665da36fda0b70ff05fd

Observation 007606b7-0f54-4940-9364-ad318506c3f5 · outbound

This paper cites Qwen2.5 Technical Report.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Qwen2.5 Technical Report

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:32.944119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:32.944119Z digest=sha256:76f8881e6e7c4fe9a1350617d73f5975d22fcc1eaae8efb35d63a08ca8e24e43

Observation fd3fe851-ca50-4459-86c0-eb3284e68f77 · outbound

This paper cites Language Imbalance Driven Rewarding for Multilingual Self-improving.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Language Imbalance Driven Rewarding for Multilingual Self-improving

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.067658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.067658Z digest=sha256:04204f6c525ca7b6d59407120ec71750614f9e705d55f76cae4b9dcd51fb28b3

Observation 5c602530-3321-4d28-9ae1-49550853a256 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:37.745551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.135246Z digest=sha256:55424f68552c8c1456ddf0d1bd477bb5dc7e45abb3935d757b65aaa05d20ec9b

Observation 459f7de1-50a4-4bb5-9cef-4d70061f5343 · outbound

This paper cites LIMO: Less is More for Reasoning.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LIMO: Less is More for Reasoning

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.230366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.230366Z digest=sha256:cbd95103ab4e2de93d5699880228b1313fe23592ad79085a62e13204c30b61c7

Observation a0e9820d-b1af-4b14-b87b-4f900dcabd39 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:37.446554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.310420Z digest=sha256:8814acb39f8afbb499d559d2613acba83188ca5dfaba812168cf143bc62bbf59

Observation bf04a478-53f7-4b9a-a795-3d434d2f7093 · outbound

This paper cites Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual Understanding.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual Understanding

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.437094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.437094Z digest=sha256:61a7cc47859c7334beace047f4ac998e85578c872b11f0e9305391e2f87b5dd2

Observation ded7f472-70ec-4d01-873b-b688b89c4823 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:37.243436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.534688Z digest=sha256:4c925fe0e452f4ef0b8333c2e225d5532a980c2fa75c481b7d496cce51984f9e

Observation 762e112c-6610-4c75-b25f-36d6adf96db2 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.625408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.625408Z digest=sha256:f5373787c6bb3ca972ca2dc14d4932ef10e9ed5f787b58a9339cd23a0a649160

Observation 242a42a7-3c4e-4fa1-8577-1062d1e4b100 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:37.073604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.723261Z digest=sha256:3925b66b693f0da221636348b37b90e7e7391b76b396b02c1b605503f299dfa5

Observation aa8c5c34-ff28-4625-a657-1dcddd708dd8 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.909689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:33.844818Z digest=sha256:c4fb629ec20df9508c7e6ca1501a2ed11ba3f4eafb4e4700150a3912e745cf9a

Observation 43b95a2c-098d-411b-bbbd-227a6a8f75c3 · outbound

This paper cites Lens: Rethinking Multilingual Enhancement for Large Language Models.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Lens: Rethinking Multilingual Enhancement for Large Language Models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:33.978484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:33.978484Z digest=sha256:b7fa5abcd8c282581e60859b6ea5b2342b91f5fddd2330fba56045ae786eddbf

Observation 0996711a-353a-46f7-8ebf-30c5f8f3ac44 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 88

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.751044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:34.104597Z digest=sha256:02f19d784a9b22e72f35b2a020348d8c07bdc4cd2c32052c1c980731a337955e

Observation 18767a1e-5f4c-48c5-a02f-551c7436fa5a · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.593780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:34.237330Z digest=sha256:315e3b3c674593542a743a506ab17b702aae40e232cd0554a856a01eade21c47

Observation dda1918c-bb5b-4f74-bd1a-c2ca0568c75f · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:34.362206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:34.362206Z digest=sha256:c52f5379cce935f9d9073ea3699fcb4defdf6cc6df7cd706afe18f8436cf6f42

Observation e49b728a-c5be-4450-b589-1482d296fa00 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.434984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:34.468892Z digest=sha256:27c128df36f09bd70b0260ee5348a64deb9010377b8145370331fab8719f839d

Observation ce17b9a4-ba83-4304-98ee-e311c9930fe7 · outbound

This paper cites an unresolved cited work.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:57:36.264145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:57:34.623028Z digest=sha256:c84019ebe8945e6f6fa304812b5d64f1cf0d334d086fc18957f871c2dc515349

Observation 71142e03-05df-428c-8bc1-c5e1e97b0b1f · outbound

This paper cites DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization.

MPO: Multilingual Safety Alignment via Reward Gap Optimization DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:34.760038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:34.760038Z digest=sha256:9c3c173131eab99630ed1320afa2e946d86232fee5b6ff822706b729cf1ef2ba

Observation d101f4e2-d9d3-45d7-85ec-90e4347e6dea · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Fine-Tuning Language Models from Human Preferences

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:34.857649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:34.857649Z digest=sha256:8ced46e0dabc0d93e299b669199ea8f243642185df92cc96a6f0b827e2f1163a

Observation 956e5001-0899-4a2b-8aed-3b38f92ab001 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

MPO: Multilingual Safety Alignment via Reward Gap Optimization Representation Engineering: A Top-Down Approach to AI Transparency

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:35.041116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:35.041116Z digest=sha256:773daaff321e5af925a9b467c096534647b60c098510ce3eea92aca5855ed44e

Pith citing papers

Observation f39c6eb5-946f-4bb9-a14d-4d9151c1402f · inbound

The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It cites this paper.

The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It MPO: Multilingual Safety Alignment via Reward Gap Optimization

Reference 146

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:40:30.600320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:40:28.745740Z digest=sha256:a7bb652eaf4e37cf9bc9bd174346f7f0851f1e87658d2e864f08c7cc86702c0b