Pith. sign in

Paper Citation Record · LEDGER

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models

As of 9 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:2506.07165.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07165 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:48:04.530071Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T14:02:11.084514Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T14:02:39.792894Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact1
  • verified fuzzy14
  • unresolved29
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 509ef113-699f-40ce-8791-51ccddbc9a56 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.380172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.380172Z digest=sha256:bdd94c79eb02fb5cb5e29597b2829f60f83c1744be22c5a3ec7484370b1ccdad

Observation 6a775cc4-d1d3-4e49-8a9b-cb25c3ccb4a9 · outbound

This paper cites - (2) Acknowledges both but slight deviations.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models - (2) Acknowledges both but slight deviations

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.384800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.384800Z digest=sha256:8a537e8bb1ca71367b6261dd563780327902364244da85951628e566afc454f3

Observation 1b1a3283-9453-4c75-b10b-7acce1c8d99c · outbound

This paper cites Unified Preference Optimization: Language Model Alignment Beyond the Preference Frontier.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unified Preference Optimization: Language Model Alignment Beyond the Preference Frontier

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:48:04.758497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.333706Z digest=sha256:24cac35675b825d6f2200097f8131ff3450682b79ff7a00a6a38da35d1977d5c

Observation 71380829-229d-44c8-9d61-8c582f171bdb · outbound

This paper cites Based the instruction following rule and given my answer to an instruction, your role is to provide specific and constructive score for me.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Based the instruction following rule and given my answer to an instruction, your role is to provide specific and constructive score for me

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.206741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.394469Z digest=sha256:a57bb2a48851ea243e4d00ad52535754fadaa14f9d3bea8595f5ced29b00a541

Observation 86162f79-067f-4f88-b0c0-d2f4e8348557 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.909769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.489631Z digest=sha256:a3122e34ed99904638c2fe368674a4287985a28cc4321c29d232cc9147ef7476

Observation 4aab7911-1989-4d87-9051-75447098113c · outbound

This paper cites Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.350188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.350188Z digest=sha256:6c461da79f8a769dba50f8286c9d60b7e34e0c013ab17605d36efb91cad31de9

Observation fbff26fe-c0c9-46e4-bedc-349cecf533f1 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.804950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.520670Z digest=sha256:09f1c9f631d2b35026327c20628998a4b098bef825a7f16bbc455c0f2eb0aa87

Observation 603471d7-87ff-4375-9720-010016c55d82 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.258845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.360084Z digest=sha256:fd36d5dcf5e5c7ec7c97c9e5961ebaf3fd611eb39a3968ce7af15b27c7914d0b

Observation 5fbb7736-bb7b-42f0-86de-4dcedd084167 · outbound

This paper cites Al ter na tively, you can use a nav iga tion app like Google Maps or Waze to get the most ac cu rate and up -to -date di rec tions.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Al ter na tively, you can use a nav iga tion app like Google Maps or Waze to get the most ac cu rate and up -to -date di rec tions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:04.774067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.530071Z digest=sha256:4ff0089cf4e6de7ae30296f8db33404571228f0c08f8091c43da4fe0a29defd1

Observation a6c9871d-7ba5-4d27-ba1d-24fa4ab96c53 · outbound

This paper cites RRHF: Rank Responses to Align Language Models with Human Feedback without tears.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models RRHF: Rank Responses to Align Language Models with Human Feedback without tears

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.370319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.370319Z digest=sha256:a8945a6674dc6e7917cd86410684c7173699a038a0ebe91143d640cdf3960969

Observation 1f6ca1b9-2929-4fd4-a204-d7da3e8cf2d7 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.375331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.375331Z digest=sha256:dc55f63cf929b5e40e5d2a683d2ddb692f9d9852ece8c1cd3388436ad56dfed9

Observation d9e92826-9fb6-46ae-8ebc-a6341bdc5fa9 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.389549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.389549Z digest=sha256:c540e77dfe620046860f58c553a1f1dfe1a23ccb304f289f00a2aacf88ab438b

Observation ed371d59-8070-42de-9b40-2ce71ec1f177 · outbound

This paper cites The response completely missed the essence of what the user wanted.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models The response completely missed the essence of what the user wanted

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.192139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.399006Z digest=sha256:fedb150b9b5e9b96a2f9c77efcbc51fa6f8f562865a546733598957c821d8fcb

Observation 0cd52a9f-1b84-4ad1-a6bd-9cbb8fd30d50 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.177921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.403467Z digest=sha256:6e9957195ec877d96ff6346b2e5e5b581bd458b36c70365b7e683f882b9d4db2

Observation cd53e2c5-9267-4541-9df4-ee050e8602fb · outbound

This paper cites The response did not fully satisfy what the user was looking for.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models The response did not fully satisfy what the user was looking for

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.163942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.408360Z digest=sha256:6cd521a24899b4d576ad408071a9ee65e0a5fe18c47b76ad7156c3153afce2e6

Observation aa9b1903-cb42-4c0a-814c-254e76fae5c4 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.150165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.413250Z digest=sha256:b895f6d716e73bcde81415a741fcbc712336ede9bb0d4220b618a2e0130d88a6

Observation cbc3e889-c355-4e81-9ff2-9165a1ad4ad2 · outbound

This paper cites Based the helpfulness rule and given my answer to an instruction, your role is to provide specific and constructive score for me.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Based the helpfulness rule and given my answer to an instruction, your role is to provide specific and constructive score for me

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.136403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.417719Z digest=sha256:9c74ff005939f5119346db66bcc2eaedc9233ee5c051786b9597f32626d23b28

Observation 32325681-7ec1-4dd6-9bdf-f0ea1f3c9990 · outbound

This paper cites All information provided is wrong, false or hallucinated.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models All information provided is wrong, false or hallucinated

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.122505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.422395Z digest=sha256:c92b316e86610202627222768f63364ba0fa62c597faa1462faabd809f0449bb

Observation 52821ed3-340a-4e97-a005-215c70773b3c · outbound

This paper cites The response may contain multiple instances of hallucinations, false information, misleading information, or irrelevant information.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models The response may contain multiple instances of hallucinations, false information, misleading information, or irrelevant information

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.108613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.426726Z digest=sha256:311862c63f01b3aa84613e8cb3be2fe9b004ca8c3a5edd7a6caff7dcc7f548a9

Observation 58aa1c99-ebb1-497d-9d3e-ebc5d1b74f66 · outbound

This paper cites The response may miss some details, contain misleading information, or minor hallucinations, but is more or less aligned with what the prompt asks for.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models The response may miss some details, contain misleading information, or minor hallucinations, but is more or less aligned with what the prompt asks for

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.094666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.431058Z digest=sha256:28e6ab8e1b1b42c616cca8e9aecfd99a0069a57e5db2541b44704adfe8d93228

Observation 7fd95f00-d9a4-45b7-9049-a5a45085108f · outbound

This paper cites It contains no misleading information or hallucinations.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models It contains no misleading information or hallucinations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.079708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.435493Z digest=sha256:d3d98880714cb49666780ce86f96a0a207f52cbf1eb3bfb38f6a1a9f18a3cbcc

Observation 790f3e81-e2ec-499a-aab5-3cfcc4769e69 · outbound

This paper cites preference.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models preference

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.065370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.439958Z digest=sha256:1d6d2828f0de8f465d0d596bed14633812759972bd6fd9463bb20db61d913fc2

Observation 0011564e-56b0-4aab-abea-a22f639a26d9 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.049744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.445478Z digest=sha256:eb89bd242addf806ec3c399bb5b6f82d664be0befe194b0c4905820a0db684f0

Observation 3f29bfbe-07b8-4c9f-a11d-1b34ce1daa2a · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.035054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.449913Z digest=sha256:51ee28819f44ae00700aa9edc1b8bdb6e77ba6730436bbd5526a1425c4571fb8

Observation dce37129-54b0-45fb-8324-9574dd80a10a · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.020579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.454405Z digest=sha256:93fc6c81bed2a8c870a23248b9f8964dbf08f506f7604cf71e01a027e30f1095

Observation 4af259a0-678c-44b6-8a77-ec21b6b0bcc1 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.006108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.458745Z digest=sha256:9a539ea32daa987f76f45fa1f8ebbbb1953ba373297a35d4811d0f571ecb437f

Observation 48b3e66d-2b68-455e-8ba5-0529e032da40 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.992129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.463266Z digest=sha256:9682d646005e9f09dc99b4f99a3276b71866a6c3c3fbaab708a06fce4f0b6318

Observation 40e32249-f472-4723-8fff-fd3817ceeb14 · outbound

This paper cites “markdown“‘markdown. This is an example of a code block in Markdown. You can see that it is formatted to look like it’s not part of the regular text flow.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models “markdown“‘markdown. This is an example of a code block in Markdown. You can see that it is formatted to look like it’s not part of the regular text flow

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:04.978143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.467611Z digest=sha256:50ce5db9d214dad7457b8fbc2b62a8351ced60aa6bbb0d7b7587cc4763df7486

Observation 0d84453a-c1c0-4a46-8b93-550b18215763 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.963674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.472349Z digest=sha256:78c3244308ecc1695d5561b795d5e3fcc6be8d281a9c696cb52f9a7e46817e26

Observation 76909b5b-1482-4ff9-8e44-9f0fcfd9bfd7 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.950425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.476696Z digest=sha256:e9a2af7501e8403ffaa58cf3f9b8faf9d18e9a73b40162e2a0ffb7f723bcbe9c

Observation ebadc2cd-9ee1-40cb-b856-13ae76e0a176 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.937069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.480870Z digest=sha256:8de8843d28d323e3d9296a35445ac6ed986aadaba369d7368930881a9fe703ca

Observation b862183f-ab10-4734-86ba-1069bdc073a7 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.923359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.485127Z digest=sha256:94c572afdcfd12f3fa05661afaf7b1fc4ac981f14b34a9d562e9dba01ad2afee

Observation 41fc8d2a-bdb7-4ad1-a839-ac4003e13cda · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.895292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.494306Z digest=sha256:87f80cd605afb2967a401b585bfb4eb9f5f2d980a3031e76984661be6a9edf35

Observation 46f191fa-7d8e-42ec-a79a-c6b79afd8bfe · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.881178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.498733Z digest=sha256:086499a00eb67653d562d4acbb721d4eb16358cdb65fd9f7becd7a7fb7b49db4

Observation 8905fa09-1251-4ee1-ad62-a8170e0d10ff · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 39

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T05:48:04.866254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.502975Z digest=sha256:fb1c2a094a11327a180119b498fefcdcd0e208f83329a032af66b2a71060bb27

Observation 0666dcdc-6eb4-4a53-9d2a-e916603e2a7f · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.850567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.507704Z digest=sha256:7f8961070a189da06eb0396e5581e42deca0051704ab0f60c8aba6e1e9c7b78b

Observation 4b3002d7-2969-41ed-bdf5-5b0cce6667fe · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.834072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.511997Z digest=sha256:19f83eaa0ca67291bbc9070cd2689e3ea9a94b1412f1d2a5a57f19898ee1218d

Observation 76cc2232-739b-4d7b-8197-e6c39dab316a · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.819470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.516378Z digest=sha256:9cab6867ca1b4c11b13d1f8d77fafaba82da0c4f323c29e9e69a4eb42361e15a

Observation d97ec2f6-8e34-480a-9c64-b69b5dd026f1 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.788887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.525653Z digest=sha256:6832612dd80da9b3abbce98879eae36a036e46371fb607bc17b7564381d629a3

Observation a13a5ca8-7485-485b-9904-bfad94f57cc4 · outbound

This paper cites Rethinking Bradley-Terry Models in Preference-Based Reward Modeling: Foundations, Theory, and Alternatives.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Rethinking Bradley-Terry Models in Preference-Based Reward Modeling: Foundations, Theory, and Alternatives

Reference 1027

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.364900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.364900Z digest=sha256:2872f84efb78436b5899429183263c606bc2e605e7e6afbaa454d1fe7d4ce665

Observation 10c276b0-1f1d-4f37-ba7b-f54921c2139b · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models KTO: Model Alignment as Prospect Theoretic Optimization

Reference 1983

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.339130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.339130Z digest=sha256:9b261c9ebbe6093917c57d88a6d41e95bfa4a7ec91a25247267385dc84b66712

Observation 46dcb65d-320c-4102-bf03-cf7d6de321d5 · outbound

This paper cites Ryan Park, Rafael Rafailov, Stefano Ermon, and Chelsea Finn.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Ryan Park, Rafael Rafailov, Stefano Ermon, and Chelsea Finn

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.273911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.355394Z digest=sha256:01e8e80ba1ecc538d4ae1e5b58ce0d788f0d447d8ee2f56b36738b49bbf2be20

Observation b655c1f6-dca9-4a28-bdb0-b8ca25589367 · outbound

This paper cites Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.345187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.345187Z digest=sha256:fbc16f225f5a548699786a3879c5cf9dcc5dfb7262317c9c1918e518461bd29d

Observation 46af202c-d500-4828-b593-0f2cd9e422e8 · outbound

This paper cites Mohammad Gheshlaghi Azar, Zhaohan Daniel Guo, Bi- lal Piot, Rémi Munos, Mark Rowland, Michal Valko, and Daniele Calandriello.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Mohammad Gheshlaghi Azar, Zhaohan Daniel Guo, Bi- lal Piot, Rémi Munos, Mark Rowland, Michal Valko, and Daniele Calandriello

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.303657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.324206Z digest=sha256:dda1e0754b0ba57848a4f313d50410dbbaf23eac33521e733e719736f4121e43

Observation 91b171db-f386-42d6-88ab-6aa8242fb34e · outbound

This paper cites Anirudhan Badrinath, Prabhat Agarwal, and Jiajing Xu.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Anirudhan Badrinath, Prabhat Agarwal, and Jiajing Xu

Reference 4455

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.288823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:48:04.329332Z digest=sha256:33ea2edbaa62d3273131f816deece3d5271816c610b3c540037d84055eb8009c

Pith citing papers

Observation 7f8dbf7d-2cd1-4c5d-8052-3748243927f9 · inbound

Failure Modes of Maximum Entropy RLHF cites this paper.

Failure Modes of Maximum Entropy RLHF AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:02:39.796329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T14:02:11.084514Z digest=sha256:59ed4ea2abd7fe7a927fbee1130a2e0bb976a0a92ca098458eea46193ff755f0