Pith. sign in

Paper Citation Record · LEDGER

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment

As of 14 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2411.10534.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.10534 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T19:39:52.676052Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 69793b3b-3e62-4f68-a6ac-631df9c59531 · outbound

This paper cites Beyond Preferences in AI Alignment.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Beyond Preferences in AI Alignment

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.595975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.595975Z digest=sha256:efc1726dbb61733ff6d6615fade4b6f78ee0b316fb06076c0d70fcd6f57845b9

Observation e683f59b-4f15-4a41-b53b-5d14e70e1558 · outbound

This paper cites Deliberative Technology for Alignment.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Deliberative Technology for Alignment

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.600574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.600574Z digest=sha256:3a7f4f24007c4814cb52c65f417bb0b065232d51574ee8cf57bb5246477aeeaf

Observation b2c0cebe-2234-4b17-9b8f-eb0c71e25f36 · outbound

This paper cites Manning, and Chelsea Finn.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Manning, and Chelsea Finn

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.604691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.604691Z digest=sha256:525a2a74011d4a601da959e637ad768902687b152474267b1c5d88f5ed6ad4d0

Observation dbca8042-4ca2-457e-8409-dbe1128d72f1 · outbound

This paper cites Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.612999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.612999Z digest=sha256:515e164377328883a08f398c1f51e367c63a942a1cbfa60e4831576919286cc9

Observation e5c260d2-ff12-4dda-896f-2cba2f51396e · outbound

This paper cites Direct preference-based policy optimization without reward modeling.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Direct preference-based policy optimization without reward modeling

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.966308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T19:39:52.617004Z digest=sha256:ad9cae153971d3ee68162cd2b0062034c673c8fc8ec7b8b5337a18e7555c9b74

Observation 4f4fdb30-2753-4bae-b614-7210da193576 · outbound

This paper cites Beyond Reverse KL: Generalizing Direct Preference Optimization with Diverse Divergence Constraints.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Beyond Reverse KL: Generalizing Direct Preference Optimization with Diverse Divergence Constraints

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.620622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.620622Z digest=sha256:c2a4adbd798c8590b1ee6ad98239305c9a9231ce23a11a05c021bd6a0213a1f4

Observation b507da81-d45b-4eaa-b2c8-9f6a244236d1 · outbound

This paper cites an unresolved cited work.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:39:52.955176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T19:39:52.624164Z digest=sha256:0d074a4462610f384be4381ae38c7f879221a927554f2703116fe7ec86b5315c

Observation 03eb4060-2812-4518-9ed5-1bf46321e9ea · outbound

This paper cites Learning to summarize with human feedback.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Learning to summarize with human feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.627591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.627591Z digest=sha256:78401be21430c9860e09c138788e7d9854c2ca2573d51cb6f68937973d484960

Observation 27a118fc-0d2c-40a1-a6d2-45b105be6539 · outbound

This paper cites Deep reinforcement learning from human preferences.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Deep reinforcement learning from human preferences

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.630421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.630421Z digest=sha256:d2114be71276d88b8970471997807a95f02da76f01178974b5180f3e35255075

Observation 109095c4-f7d5-435a-a3e9-b3c8894cb8db · outbound

This paper cites Improving alignment of dialogue agents via targeted human judgements.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Improving alignment of dialogue agents via targeted human judgements

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.633285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.633285Z digest=sha256:c1887cc0516198758da7b136b894d071fda0f50e34439f18857628a0978f4e60

Observation 41a79c77-0549-4c67-b638-67f13c1c4dd3 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.636422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.636422Z digest=sha256:5dcb91fe0b38fb7a91e93f184add249bb7004d43413e71f4382a13e89d652d14

Observation 58982c2f-1cd4-4ff3-867a-2eee19425b61 · outbound

This paper cites Jury learning: Integrating dissenting voices into machine learning models.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Jury learning: Integrating dissenting voices into machine learning models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.930031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T19:39:52.640437Z digest=sha256:59ea1df9b1e5d1c68d69d1c03e31cb0b6002028b7ee5721da7b330ee90a7019e

Observation 6931099f-7b9a-4544-acea-90ba83baaa35 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Constitutional AI: Harmlessness from AI Feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.643877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.643877Z digest=sha256:89544a090244e5a12323d3d07a8638e8213dce26dbe33f150ff3e099d1680c13

Observation 9196ab73-006f-42b9-a3b6-0872b8104ab6 · outbound

This paper cites Liao, Esin Durmus, Alex Tamkin, and Deep Ganguli.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Liao, Esin Durmus, Alex Tamkin, and Deep Ganguli

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.647474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.647474Z digest=sha256:f8c67d3b870216710df16ca73a2e2772346121ecf9322547411d34eb75ec6944

Observation 590043db-3736-4e0c-a7c8-ac12182dae2d · outbound

This paper cites Rule- based rewards for language model safety, 2024.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Rule- based rewards for language model safety, 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.918756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T19:39:52.650573Z digest=sha256:eb65b8111a89124ab50348066cdddf829908c785f5347c7955073c64e42f5be8

Observation f4469d5c-a084-4ae8-a1cc-d61c59f13394 · outbound

This paper cites Specific versus General Principles for Constitutional AI.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Specific versus General Principles for Constitutional AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.653860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.653860Z digest=sha256:13f06ef2939788d4d694c13eefb59fd7f80680b2e22560f701a244360cda9dde

Observation 8f3849b7-91c4-4410-88bb-3776415f1b7e · outbound

This paper cites Rule-based reinforcement learning for efficient robot navigation with space reduction.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Rule-based reinforcement learning for efficient robot navigation with space reduction

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.907710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T19:39:52.657688Z digest=sha256:2097ad0a15a67fbffd358053ac0835c03b2b1722441427bd86f1514a3835a78e

Observation a1956bef-af8c-4402-b8f9-560006877d8f · outbound

This paper cites Democratic Policy Development using Collective Dialogues and AI.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Democratic Policy Development using Collective Dialogues and AI

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.661399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.661399Z digest=sha256:9fa27abc437f09e716fa67e6ca991abcc669b97c818131dfd22c1541f89cedd8

Observation ca1a902e-34d3-46ed-9c2e-90d8e603b608 · outbound

This paper cites Inverse Constitutional AI: Compressing Preferences into Principles.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Inverse Constitutional AI: Compressing Preferences into Principles

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.665058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.665058Z digest=sha256:1ed8ac7238e9771d1562a5032185ff30cfa787b3487f4c37aecd4608e35ac69d

Observation e859dcb8-59b1-4005-bc0c-84ab48944b0b · outbound

This paper cites Qiu, Michael Varga, and Aviv Ovadya.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Qiu, Michael Varga, and Aviv Ovadya

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.895654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T19:39:52.668925Z digest=sha256:2ca26413fc6d7fa70f8f7ea1d980b42fbf549bee860ed83f8a0ad0a4bfedee84

Observation d49ab429-8548-4a11-b8f1-c3c01680bdba · outbound

This paper cites Mémoire sur les élections au scrutin.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Mémoire sur les élections au scrutin

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.884725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T19:39:52.672404Z digest=sha256:8c32ec7064f1c41af52875db35659d6b14357937aaf52663daaf4c902ee710ee

Observation 68ead250-29e0-41b8-b584-055f85b030b8 · outbound

This paper cites super-majority.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment super-majority

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.676052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.676052Z digest=sha256:44000d2690c50922f01f230be5577697645b2aec69d6d4063216b149636e9a6d

Observation 748ec504-d1ab-41de-90ee-2339fd5f3b81 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.608618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.608618Z digest=sha256:2cde7617a473882493f82b5b556797a10dcef4702f724f8e2b6d31314e4a8d87

Pith citing papers

No inbound Pith citation observations are available.