Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2312.08358.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T18:15:15.978088Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 1f529d69-7010-4884-8f41-2a12335c311a · inbound
Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 171
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23783648-074d-49db-887b-bd1fdb950640 · inbound
Beyond the Binary: Capturing Diverse Preferences With Reward Regularization Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0531dae-9d5b-44ea-8ec1-4e8c62143004 · inbound
Test-Time Alignment via Hypothesis Reweighting Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 60d3818d-8a1a-445f-bb9a-ecb28acf6327 · inbound
Clone-Robust AI Alignment Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d07cd496-0aa3-4874-aa36-cd583d2aca4a · inbound
Examining Alignment of Large Language Models through Representative Heuristics: The Case of Political Stereotypes Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4dada43-e749-4918-a3d3-c68069742f65 · inbound
Jackpot! Alignment as a Maximal Lottery Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 809a079d-4dea-4418-bbbf-009af9095b07 · inbound
CTR-Driven Advertising Image Generation with Multimodal Large Language Models Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d19d2f2-1c7b-4921-a9fd-08f0ea002ba9 · inbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 699e2fc7-0744-4fe0-9bf4-189127a6bc72 · inbound
Theoretical Tensions in RLHF: Reconciling Empirical Success with Inconsistencies in Social Choice Theory Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 129db947-a33a-4513-93e2-67f981934587 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e20a2f2a-12a6-4e2e-8637-5519f87425ea · inbound
Active Query Selection for Crowd-Based Reinforcement Learning Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1abf0f85-8f73-4ef6-801e-8190f4448a73 · inbound
RLHF May Not Reflect Genuine Preferences Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4fc5fe13-7443-4442-8716-83ee117809d9 · inbound
Efficient Personalization of Generative User Interfaces Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 06cf540a-7373-4691-b261-15b869a81079 · inbound
Relative Principals, Pluralistic Alignment, and the Structural Value Alignment Problem Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a4e556fa-f511-4cbe-b4cd-41ca01a10e4c · inbound
Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 186
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7c1484f3-54f5-4e27-b71b-48038f27c286 · inbound
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 215
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.