Pith. sign in

Paper Citation Record · LEDGER

Boosting Direct Preference Optimization with Penalization

As of 9 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2606.12505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.12505 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T10:36:26.784333Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7f08a3ab-acc4-4139-a3f0-a6de9ecacf1c · outbound

This paper cites Advances in Neural Information Processing Systems , year =.

Boosting Direct Preference Optimization with Penalization Advances in Neural Information Processing Systems , year =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:d5a3fab6317b0cb85f1e8db7996ff134e90c07ce61f1b626d60932a900d1a55f

Observation a9f42c09-776d-4800-9898-5a795eeed333 · outbound

This paper cites Advances in Neural Information Processing Systems , year =.

Boosting Direct Preference Optimization with Penalization Advances in Neural Information Processing Systems , year =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:a98ad4be761d693c3376c302ec7a735408372fade0978a780de5c1394e7188f9

Observation 5130cb78-ddb5-498a-bd8c-c3da3ccb94cc · outbound

This paper cites The Method of Paired Comparisons , author =.

Boosting Direct Preference Optimization with Penalization The Method of Paired Comparisons , author =

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:bff8497be6c470bc1e1a0f481dc53dcc297cb694b10f5f8cbcfde2e802a3317a

Observation ab993950-bcd0-419a-abd5-46c9b0ea0700 · outbound

This paper cites Advances in Neural Information Processing Systems , year =.

Boosting Direct Preference Optimization with Penalization Advances in Neural Information Processing Systems , year =

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:4f0abab4a667874c3da53f984c03e01800113659a9ce2973aeecb5a66ceb59ed

Observation cbf36011-e588-4292-ba28-6f978a891b3c · outbound

This paper cites an unresolved cited work.

Boosting Direct Preference Optimization with Penalization Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:e59986e5572bb877099c060576816e768e460d2acb2e531e320b9a6aba13d036

Observation e45fb5f7-635f-49fb-b37f-e5cfc184354a · outbound

This paper cites AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization.

Boosting Direct Preference Optimization with Penalization AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T08:57:48.047259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:07de988777194aa211b410474ae016fa7f9f892abdd979ddd88cb2d10740979f

Observation 8568ac25-9b9f-456f-a1c4-3ed628a0daa3 · outbound

This paper cites 2024 , eprint =.

Boosting Direct Preference Optimization with Penalization 2024 , eprint =

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:eacda18792d7a535d225377adb638ce6186c59a7af5823cfe8cfac3e8a0bf0b8

Observation 5b4e59c9-ad7a-40a2-ae4a-26a77d7d28b7 · outbound

This paper cites Simplicity prevails: Rethinking negative preference optimization for llm unlearning.

Boosting Direct Preference Optimization with Penalization Simplicity prevails: Rethinking negative preference optimization for llm unlearning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T08:57:48.044433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:772b52f63bd83ddf235d7b0136c41aebf4097886bb3cc46ff41ef2f7f62c867e

Observation 3a1863a5-77cd-4b9f-8b88-380e69480e7d · outbound

This paper cites Conference on Language Modeling , year =.

Boosting Direct Preference Optimization with Penalization Conference on Language Modeling , year =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:28fb4537b3e30c23e576375cc5c1bc567e497c9e56b6e3a491cd184c6c928792

Observation ea16de6e-408c-4a4b-9262-820b28c7d855 · outbound

This paper cites an unresolved cited work.

Boosting Direct Preference Optimization with Penalization Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:6cfeb8155c1c542fb45bf98498ddb8c1bb02dd465cc8bfa448f036de8e5c4652

Observation c5fd4ed7-f508-44b9-bae5-3a38819dded8 · outbound

This paper cites 2024 , eprint =.

Boosting Direct Preference Optimization with Penalization 2024 , eprint =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:a5d3420421f27b8129168c92abf7655e9bcfd6e5c3ae5bf01945f71104627efc

Observation 7ac105ee-4bbd-4558-a586-1809a1db502d · outbound

This paper cites 2024 , eprint =.

Boosting Direct Preference Optimization with Penalization 2024 , eprint =

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:81c374d3c544fbc871357af617a3a96a040f88564d92b4eaf96ad327472be540

Pith citing papers

No inbound Pith citation observations are available.