Pith. sign in

Paper Citation Record · LEDGER

Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2402.10958.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.10958 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T16:18:40.567531Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4ad7b4d5-f69b-4d43-a61b-cdfbbbb65288 · inbound

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment cites this paper.

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:38:39.836064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-24T00:38:36.992597Z digest=sha256:395f79bdea59c4db38148c5ba0d744c3d1ca9880b3f0305938298846a52e5c76

Observation 649dace3-01ac-4ac4-8cca-b3a0e6853cb3 · inbound

On Almost Surely Safe Alignment of Large Language Models at Inference-Time cites this paper.

On Almost Surely Safe Alignment of Large Language Models at Inference-Time Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T16:18:40.567531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:18:40.567531Z digest=sha256:e50a62f5d14e4637cb7e1ea67dab263be5daa859e27eab5d6de29ccee497e555

Observation b6ce621d-1846-4d13-b133-bd353dcd43ae · inbound

Reviving The Classics: Active Reward Modeling in Large Language Model Alignment cites this paper.

Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T11:47:17.652900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:47:17.652900Z digest=sha256:0a4452b5e9f60f48421653f0334d3bacf06a8161191f7d1733a11be30af97a81

Observation 2a9e79ab-9ede-473a-910b-1ae2daf2d4d9 · inbound

ChartSketcher: Reasoning with Multimodal Feedback and Reflection for Chart Understanding cites this paper.

ChartSketcher: Reasoning with Multimodal Feedback and Reflection for Chart Understanding Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:56.873323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:56.873323Z digest=sha256:3900071262b1ada2df78b773069d3f8c8974a1182691e16bf36f417dd969cf4d

Observation ac69e7f0-10ed-4d9b-8ad0-9b06eba55910 · inbound

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities cites this paper.

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-06T16:34:25.272362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:34:25.272362Z digest=sha256:e6b0393dc6cb8db46cf414546c8cb4d2df082853224d66b19e21edf8c089f03e

Observation 3202da34-0f25-437e-b9e8-d4835183269e · inbound

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects cites this paper.

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:30.106598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-18T08:11:31.181704Z digest=sha256:a4b0dda7787d3184138a84f8af79adfa4e0c64e72634f2e7544ed4ad433e099b

Observation 77500ca3-58fd-49d1-8ecc-9497e8e1e968 · inbound

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection cites this paper.

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:36:35.294034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T20:33:09.627731Z digest=sha256:d2c959dee6f2ec97ce2796b8c809f3838b5f28521e9ec9ae685c8f6676e1fb9f

Observation 89595ea6-5a5b-4048-83d3-84593ecc7695 · inbound

Generating Place-Based Compromises Between Two Points of View cites this paper.

Generating Place-Based Compromises Between Two Points of View Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.278619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-08T03:36:31.695964Z digest=sha256:c878c3430066d1a89a03bf164b9cccfff84295907b2ad83796e3806cc64956ef

Observation 73bcca5e-ca59-4b78-a8ec-22f726ff617d · inbound

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph cites this paper.

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:30:58.410493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T02:26:33.933264Z digest=sha256:1cba543312f55bcc50c95fb3ae36565ce0fb54db7ec03c08e738e95be71c6000

Observation 4d9973fc-054e-46c5-b440-63a5e531ee52 · inbound

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation cites this paper.

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:34:26.735555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T08:30:37.016856Z digest=sha256:0b68fd450bbda04e53360e3734ba4cf8fe81fd3092ee507034cf2e434b8b10b9