Pith. sign in

Paper Citation Record · LEDGER

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable)

As of 8 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 0 inbound Pith citation observations for arXiv:2507.07855.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07855 v4

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:52:38.826438Z

measured 71 of 71 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

71 of 71 outbound references displayed

  • verified exact2
  • verified fuzzy18
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 965346be-dd5c-4541-a403-540a244512fa · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:40.129033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:33.515815Z digest=sha256:3c60089ccd7bc3700d81670240566113f282470c7ba44d6fb64fa93cfbedb434

Observation 81fccb27-e677-4950-bdec-22a72fff3fde · outbound

This paper cites Alfano, S.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Alfano, S

Reference 2

Resolution
verified exact
raw_fallback, observed 2026-08-06T18:52:39.375745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:33.577577Z digest=sha256:e584c9aba31401c6bf0f13b0fce5bd7f9c8a4a09123899a3db7bf1fb54567246

Observation 80ac216e-2012-4d87-b9ef-03a0217c83c2 · outbound

This paper cites Amari and H.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Amari and H

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:40.116234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:33.693550Z digest=sha256:cb7be1a446ca52418ad012b90bab0c8e60b7fe7fca984196239827c1d3ebf088

Observation 9a063b77-0699-426f-9c4c-ff02860b399c · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:40.103357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:33.787534Z digest=sha256:80fdf138f09751be85e7a664c9f8137f2104a8be1a8f26f9351f58b1c44b962c

Observation d8a5099f-8521-45ab-8fd5-6f6d36949e01 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:40.091265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:33.863589Z digest=sha256:283241c188a5ea71c41ad79fdf122244f1d664943358bbebf18f288679242f3a

Observation 5acc9787-cfeb-4a03-a9ef-73a41d7caafc · outbound

This paper cites Bao and N.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Bao and N

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:40.078926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:33.934402Z digest=sha256:a08986a85888a1fc574f3b661a130e9631025341fd151839b7b630feafa05dca

Observation 75ca5948-f075-4972-aad3-1adc2626e9d7 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:40.065896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.039406Z digest=sha256:9765855a2669d992543a23dbabeea795870007329bc56a2efe848acd27d92c10

Observation 02724bf2-6657-422a-991a-34a97ae0e295 · outbound

This paper cites Blondel, A.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Blondel, A

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:40.053423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.197979Z digest=sha256:cebea9f050410cbc21998463152afc84093254a687edac1ed6e3680634986f1d

Observation 9f798169-da66-4976-988e-1f5427eabfa9 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:40.040610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.263682Z digest=sha256:759c78bbde45e660df5d06400fce6e893194cca42c80696dd8248ff420de82e5

Observation 02b8a3f9-e2c3-4d3b-9b5f-1ec9f1231fd8 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:40.027737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.366115Z digest=sha256:881519846cfe6261d86811aac04e8e62bd2e6586addcc49128c5df5dd1e8d14f

Observation 98eb9345-e399-4045-bbdb-d7f2cdf0b977 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:40.014204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.469748Z digest=sha256:f564e3c2bb740cdcf0f27aa861e15c65fc5e36adce98778c3a837b84e202206f

Observation ec68244f-d308-4ea8-8d49-23025093a4f6 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.999814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.542796Z digest=sha256:6a3fc2a1e4dcf675b9d2db9d6ff8e6817ded4aff252d0cb98d2383cb646b3337

Observation c8a8fa34-7bb3-4ca3-9565-c24ec8214238 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.985297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.639454Z digest=sha256:4d28c18f2755de98473873335eabcb60d73deff5307a88711a7b1e4a580b02ec

Observation d5c96902-2571-47f1-997d-aba9d37f84e3 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.971992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.793576Z digest=sha256:a3a2cbe2cb953e2d2420ca87b69d89377dace0ce735fdd8ece6190eceb79c316

Observation 3aba22c1-3715-4e1a-8acc-44cf78ad5ff3 · outbound

This paper cites Doignon and J.-C.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Doignon and J.-C

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.958070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.870271Z digest=sha256:d3678c122d96fb310f375d6460ac5f0e776cafca1aca7e5702ae1505d6c4f61f

Observation e6e2ce3d-05c8-46d8-99fd-a7221203e38f · outbound

This paper cites Ethayarajh, W.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Ethayarajh, W

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.945039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:34.946745Z digest=sha256:9ef7f8f747c28dc2151c54da0e922f2da1093eccffca9d3ff68e07a925b948ea

Observation b4cd4805-d636-4f9b-999f-2bf556a9bfbe · outbound

This paper cites Gneiting and A.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Gneiting and A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.931857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:35.057746Z digest=sha256:bfb39c15f0bbcde1a9d356bb0c86219b308baf8dc502b1dcb5ebf2e218439a36

Observation 32cf51ff-128e-4429-903c-9e18fb93f4dc · outbound

This paper cites The Llama 3 Herd of Models.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) The Llama 3 Herd of Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:35.156728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:35.156728Z digest=sha256:d49609ff8b7eb51ac6ac688872592e8ffcc4aad9c949409272183b248704424c

Observation 0b4a9192-b9bf-4986-bc94-d7c986731020 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:35.231062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:35.231062Z digest=sha256:635bfacd157d570c11cc7ee63164b038dd1e2d30202e3b94fe1dc614baa9028f

Observation 8a966e47-8557-4927-a89b-76ffe185a115 · outbound

This paper cites AlphaPO: Reward Shape Matters for LLM Alignment.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) AlphaPO: Reward Shape Matters for LLM Alignment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:35.347941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:35.347941Z digest=sha256:8f9b2d71a0b6026064bbec60961346218f96665f209b4ef5fc8a62d761093e42

Observation 0b463851-738e-4987-a460-caae629ac756 · outbound

This paper cites Hastie, R.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Hastie, R

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.909162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:35.450714Z digest=sha256:8fa53198f33579bebfa0026ae63692f102ec841a68cc46e1d9cad4714c430989

Observation 5229e8a3-3a7e-4622-99b7-61e3fb4701fb · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.896574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:35.600101Z digest=sha256:ea97b949b3649b6d89d1b7f5f4ef77e7be8ae35df11c3c7d05d2b389e5028dbb

Observation 8a5b238e-5927-42cc-bb41-8ba86e955dda · outbound

This paper cites Huang, W.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Huang, W

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.883467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:35.670121Z digest=sha256:c34067f77e81db01c7da8a54cc5ed36d419d0a4248a464f04802387de4fd3b03

Observation f4ce977c-d379-4f66-ae69-9b6c4f9dd52b · outbound

This paper cites Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:35.793726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:35.793726Z digest=sha256:6c6b85f310fd13f6fbc63f978d67f1246d7bd486ef1d336a8919c1a12f5a1b94

Observation 54ec5f14-ff72-42b4-be43-e7ced6156b14 · outbound

This paper cites Kakade, A.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Kakade, A

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.870296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:35.884856Z digest=sha256:1530e4ac38d519776ee1884a1a5ab398408becc877edf43933da184f17a3e884

Observation 207c34d9-a568-4712-a475-d5c6144f41cf · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.857109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:35.950676Z digest=sha256:e59760db1477a447b28b82dc54d9e9924fed50653ee7dc0fe9b31ba9a653f3ae

Observation 1a545f81-5c72-4731-a87e-71a3aea4a00b · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.843801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.013640Z digest=sha256:5361075c43d4cf8fbcba5fbb74ae40011c1b35e3cefa5e0598870fc3ef30ab17

Observation d2c9642c-e399-49e0-952e-8cf8db96b42a · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.830449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.081485Z digest=sha256:e2d0c76bd7be239ab32bd87903c97f0742abc14d75b24cc782604851d234deb9

Observation d31023e5-7897-444b-a6ca-3cbc8878500c · outbound

This paper cites Direct Preference Knowledge Distillation for Large Language Models.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Direct Preference Knowledge Distillation for Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:36.169261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:36.169261Z digest=sha256:3231944512192d2300cd9af55da53bfeb040decbcf49f216fd864fe03c0e3b77

Observation 90904036-3fb3-48b8-8851-53fe7c53e6a8 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.817390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.262670Z digest=sha256:8255eae273bb539a216db7e1258acd88bc0b1fb76d4532b72b7876d6778f9d85

Observation 8a726a5b-64ff-4d3e-b2b1-3d1b9561d547 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.804388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.334781Z digest=sha256:7bce1909ff0aa74d15da680c5e2ef99a9baf1f8a743f8d2c143b139cd3506058

Observation 84697ba7-17f3-401a-adc4-40948e8236d8 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.791682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.408635Z digest=sha256:0d60abe6d548333ad79660a583c9b4c652fe4b8b23316de1c8262b65889fc0c0

Observation 8ba6d603-4793-466b-9b3d-f8b415ccc5aa · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.778771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.471253Z digest=sha256:d924856f8fc59b6db942d7b9851c25ed74c07b061364fc9b4168e1e357272192

Observation 076ddc84-144c-4949-81a0-2359698b150f · outbound

This paper cites McCarthy.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) McCarthy

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.765690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.561030Z digest=sha256:96df959ac31c2153ec2c71392aa403c34675ad2b8b1d9381442e8cc8e7e434cc

Observation 30c1683d-dab5-4f38-9829-ae7326836ebf · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.752867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.626119Z digest=sha256:7a4893c3d8e97e98d6f2754e80c6890ea6e6cd1fca6820ba48d8be76abc0bb0d

Observation 2d8a6005-dce8-457c-94ce-b0e00586c136 · outbound

This paper cites Mitchell.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Mitchell

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.740328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.721352Z digest=sha256:2f2b3e9e42812048a8706c9a6fd914b430ef32957fff8fdf38a9e7ce464d1b7b

Observation 640d0667-d9b3-4cf8-a373-16a7155fd115 · outbound

This paper cites Nock and A.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Nock and A

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.726540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.788286Z digest=sha256:0b316c13a90f6f8626d593bbc69c9d0b434d1fdd5bd3e2f59f6e833316839dcd

Observation 47acef93-28e0-467b-8706-75b236f882bc · outbound

This paper cites Nock and F.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Nock and F

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.713336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.854967Z digest=sha256:22f64ff07b173ae0f7c23f86b5754742d6a623f4ae9c76fd7ecc97e09ede8a46

Observation 19057d04-3f49-4fe8-a5dd-af41e8ea68d9 · outbound

This paper cites Nock and F.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Nock and F

Reference 39

Resolution
verified exact
doi, observed 2026-08-06T18:52:38.889685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:36.915265Z digest=sha256:5d35b70674bc118267bb19370e8b0c8f6452e190c0eb6ce7832d855f4cf55cb0

Observation c59ab0c2-4ffc-48e9-ba54-f641f9820612 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.699922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.001364Z digest=sha256:b8b83ad1bdf6e7ce22d3bbc0c311e40063cbef76e38161dfa4e7c8428689374d

Observation d48ec375-fdcf-4214-8efa-b92bb651f903 · outbound

This paper cites Nemotron-4 340B Technical Report.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Nemotron-4 340B Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:37.079112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:37.079112Z digest=sha256:28136127ba40c2d4a945187a2d94a1fc3c5a7aa507169076cefc87be5ca1e0d3

Observation 6d9a245f-f9fd-455f-8b32-49a6a3e425fb · outbound

This paper cites Ouyang, J.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Ouyang, J

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.687241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.152941Z digest=sha256:198a05c3cc1efdc0038f89787f217b26d4f841118aa6173df2db2c4270d92839

Observation b96596a0-c6c6-45de-8988-40fcf3df7371 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:37.224508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:37.224508Z digest=sha256:b60b34a73c2a97ec0037c73059ecf529620882426f8b244da0f98579f7758c14

Observation f56d7776-02e4-497b-9014-63d4212ff29a · outbound

This paper cites Rafailov, A.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Rafailov, A

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.673035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.280470Z digest=sha256:bdbdedc8c88a1672d81aa9a30b67f17c4cf8d8a0d8cad0e8676556a550643aea

Observation c9a5e42b-2936-4a46-93f6-1f6c17925fc3 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.659448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.389498Z digest=sha256:d2dc1b15b9b0a9ac0b1515915f3048598f898e219916a801b811c84a1c4538db

Observation 4d5ac30a-a30f-43e0-915a-d8c703a5fb74 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:37.471398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:37.471398Z digest=sha256:e1ac5beab5034747a1cb39b32ffe687db12ecb8b9d251ed8099ce4074a244921

Observation cbb500e1-6af3-4c61-9eea-1934bd815d86 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:37.535852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:37.535852Z digest=sha256:0053b2f3ce02af52b6eefcc3ffe5f80917c6f91b530db23598533e5667adc749

Observation 0d563e8f-888e-405b-992a-05b6de337c58 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.646073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.599101Z digest=sha256:f6c817dc5f5e32ab9f20871dea6fd37532bf6ce7cbb72d3c95561d6684e1a55b

Observation ac251eed-27f3-4bb2-8ba0-6b38c10425e9 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.633197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.665105Z digest=sha256:7facac617522933c0df9c5d8a3517b584f718f5d036f3d1aa6b4c8aa1d059453

Observation 3ce4ba05-615a-4030-be25-2691ab56a28b · outbound

This paper cites Slocum, A.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Slocum, A

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.619757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.723880Z digest=sha256:83b525146ab17fecad9589f20dc69c871b7ef665dd6750e303d63cbf629fab14

Observation 8812f057-e68f-4d3e-b688-476d9e86ca73 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.607057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.830259Z digest=sha256:e34718e7d485f9a9c26e470875a5d59d38729d9539c8bafd455a2f749cf01d78

Observation bd040ad2-ddcd-4ea5-8881-49e96b1d3e82 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:37.893399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:37.893399Z digest=sha256:d1f31f002090a652499cf1ccf5a32b28e700637797ce454ff9bbac3e7e760733

Observation dac90df9-7ad8-44c4-9c78-95413555e016 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.594541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:37.954555Z digest=sha256:7a34840765763be3ef372f80b51932a7e5ce3bd92144afbe1e6da8b555d4a0c2

Observation 7f341fc4-d89b-4a78-bcab-887fe94274b5 · outbound

This paper cites Sypherd, R.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Sypherd, R

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.580769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.019369Z digest=sha256:0bef0e066cf5e501cde7df9e54d45d7a6b435bc194e7ef575c925b9149ae3a9c

Observation 9daebe3b-71e2-4a3b-bbe4-99ffe76e8547 · outbound

This paper cites Tunstall, E.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Tunstall, E

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:52:39.566845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.147360Z digest=sha256:c0b149362d0ae9aa2bc0cbbf731d5d7e99c588e6a6f7c8584669f45b104d0c41

Observation 88ca329d-0891-469e-9bfd-9ba8e26cdf17 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.553032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.268692Z digest=sha256:48cb0d339e4ac318066fa32fef3010a5b9c83c28ca108bc693862e57ab57d6e7

Observation 50fd11cb-35ba-45bb-bacb-9fb4b9a50010 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.538286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.437209Z digest=sha256:97bd3df04d01b21b2a7d219f2d5713453c0baa2c718f4e7abb33174923a88bc2

Observation bfc3f260-911f-4edc-ac4c-2c45d148e19b · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.524264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.563817Z digest=sha256:5e3244f5b1789a6bd43108b1b086f4dbc499a0919e1c148651c4da63ca7cf9dc

Observation adc61e49-10df-49f2-96e6-7f0df74fd2f1 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.510989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.682104Z digest=sha256:0e9de6fd7cd3979900b6f40c6920d907b1ac73aa282a628aec1b8c85412e9958

Observation e22ddd89-11b9-4ec6-8db6-9870ea80c6b6 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.496283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.778837Z digest=sha256:663845b72aa377f4607dd6ea1ad5648d9fde5ceffde10057d1ddc305269f3e63

Observation 08a406d6-ea26-4385-ad56-1768411060d0 · outbound

This paper cites Qwen2.5-Omni Technical Report.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Qwen2.5-Omni Technical Report

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:38.783043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:38.783043Z digest=sha256:9ac0e7d5cd1eedfece7350265df06414cac49e629dbcededefeb858ca12f54d6

Observation 87803603-1cad-408c-b14e-b59054fc6f47 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.480895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.787672Z digest=sha256:2e6a7ca54adadfc4eb105ef16fd42dc952eaf58880881a9b42eb81e1d7ac782e

Observation 0881ed7b-d103-4605-bd91-21ebd84de778 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.465569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.791473Z digest=sha256:4dccfd78dfe05a67547507c38dc53e5a0a6851f5fffdb9d57fa6afde4d7d6f63

Observation ec96c192-f2de-446e-a537-58949b0618cb · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.451379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.795690Z digest=sha256:e2c4c6df31559d08706968ff3e82935840fd78e24aa806a053a775430f1e5067

Observation 87ebc61f-baca-466e-855a-47b090309761 · outbound

This paper cites Beyond Bradley-Terry Models: A General Preference Model for Language Model Alignment.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Beyond Bradley-Terry Models: A General Preference Model for Language Model Alignment

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:38.799899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:38.799899Z digest=sha256:ecd0c59ee6984ed4ce795d046a5a5789277395e54f2a8842ec43382883cc1b6b

Observation 57dd9b03-2a44-46d3-b077-a2ffcc84e3d3 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.436613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.804815Z digest=sha256:d027042283c64844fbffac60402f1d4059e68c8ce9871c76aa0d0e5093c6d24f

Observation 6b833df4-92a0-4977-88b7-29c850e08e7f · outbound

This paper cites SLiC-HF: Sequence Likelihood Calibration with Human Feedback.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) SLiC-HF: Sequence Likelihood Calibration with Human Feedback

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:38.808890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:38.808890Z digest=sha256:87cae93730f4c063596fd6e509635f2e17e4cbf5fd20c44130ac7c6c98a9bfd5

Observation 0e55d8d0-b384-45ba-917f-38e4578e5718 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:52:39.421914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T18:52:38.813440Z digest=sha256:419be17ab71d470c27cce42323f97c93a42ea6e58d0731cc7076b3847fd78948

Observation c8c75fc2-f3d9-4982-935b-29a743119ac3 · outbound

This paper cites @esa (Ref.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) @esa (Ref

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:38.817370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:38.817370Z digest=sha256:b871cc150735a77b7f911f9a467a98a84f604c81e02128e4f0b7e3f3dee4e822

Observation acb13461-2315-4188-8636-15f03c2cd226 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:38.822165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:38.822165Z digest=sha256:cf745325e184f86e01d75468d7b464502802b2921adf6a5da14f9e498410ac4e

Observation df42e405-82d6-4b7c-92ba-e6aeddbf3765 · outbound

This paper cites an unresolved cited work.

DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable) Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T18:52:38.826438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:52:38.826438Z digest=sha256:01314ff44310c80d819ea34f15eb9fe5b2958db1cf1bebb1c91c1e86efaa942d

Pith citing papers

No inbound Pith citation observations are available.