Pith. sign in

Paper Citation Record · LEDGER

Boosting Direct Preference Optimization with Penalization

As of 18 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2606.12505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.12505 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T10:36:26.784333Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7f08a3ab-acc4-4139-a3f0-a6de9ecacf1c · outbound

This paper cites Advances in Neural Information Processing Systems , year =.

Boosting Direct Preference Optimization with Penalization Advances in Neural Information Processing Systems , year =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:25ecceb1d8a07e42f591edc63f5adb6994a01c6f61c5f84870c7d54d11c60577

Observation a9f42c09-776d-4800-9898-5a795eeed333 · outbound

This paper cites Advances in Neural Information Processing Systems , year =.

Boosting Direct Preference Optimization with Penalization Advances in Neural Information Processing Systems , year =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:2f2f463c2095cdba298bbc395d6c9824d8d446dc854739d50a87bc5651b76628

Observation 5130cb78-ddb5-498a-bd8c-c3da3ccb94cc · outbound

This paper cites The Method of Paired Comparisons , author =.

Boosting Direct Preference Optimization with Penalization The Method of Paired Comparisons , author =

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:093833a4da7d8527b94a3802120f3328fe44547d3d13701972a12eaef44b4afb

Observation ab993950-bcd0-419a-abd5-46c9b0ea0700 · outbound

This paper cites Advances in Neural Information Processing Systems , year =.

Boosting Direct Preference Optimization with Penalization Advances in Neural Information Processing Systems , year =

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:2e0ebd443f8fc9a6072d23d98e637d8a63db57d3421efd125852e9988c973aee

Observation cbf36011-e588-4292-ba28-6f978a891b3c · outbound

This paper cites an unresolved cited work.

Boosting Direct Preference Optimization with Penalization Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:01b8c7263ef6beedb699b687d4a45c11935a5fcc402db0c51dc10cf3f4bce163

Observation e45fb5f7-635f-49fb-b37f-e5cfc184354a · outbound

This paper cites AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization.

Boosting Direct Preference Optimization with Penalization AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T08:57:48.047259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:09874e02577e951af03c719fde425a451b878e1697ed59e4a822183fc905338e

Observation 8568ac25-9b9f-456f-a1c4-3ed628a0daa3 · outbound

This paper cites 2024 , eprint =.

Boosting Direct Preference Optimization with Penalization 2024 , eprint =

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:608b931db4f82e2797c1b1ae2cd0e530a6a7df90f4d71e6e6993d88350b7e197

Observation 5b4e59c9-ad7a-40a2-ae4a-26a77d7d28b7 · outbound

This paper cites Simplicity prevails: Rethinking negative preference optimization for llm unlearning.

Boosting Direct Preference Optimization with Penalization Simplicity prevails: Rethinking negative preference optimization for llm unlearning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T08:57:48.044433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:1d51d42e7a5f72ac02e97e8354f46647eec14f98de8c08fe4261d2ee92a0df84

Observation 3a1863a5-77cd-4b9f-8b88-380e69480e7d · outbound

This paper cites Conference on Language Modeling , year =.

Boosting Direct Preference Optimization with Penalization Conference on Language Modeling , year =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:3f22e6b05e518e9950f9a9b99f28a05e359f8855f521d389ed6b75c3518aff48

Observation ea16de6e-408c-4a4b-9262-820b28c7d855 · outbound

This paper cites an unresolved cited work.

Boosting Direct Preference Optimization with Penalization Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:6250e0c5c52e5c245c21b5e21da3eeeb75853cc3cdb5d63d8b384caf41656453

Observation c5fd4ed7-f508-44b9-bae5-3a38819dded8 · outbound

This paper cites 2024 , eprint =.

Boosting Direct Preference Optimization with Penalization 2024 , eprint =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:71bd7078aad47a10f40660135ef42193537e828ce0d34ec2536d3eed2686ef83

Observation 7ac105ee-4bbd-4558-a586-1809a1db502d · outbound

This paper cites 2024 , eprint =.

Boosting Direct Preference Optimization with Penalization 2024 , eprint =

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-27T10:36:26.784333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-27T10:36:26.784333Z digest=sha256:f47f46fad70467431ddb5c845a5958b2ea276fa859937b950028837d42971669

Pith citing papers

No inbound Pith citation observations are available.