Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T11:47:17.656530Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 2 inbound Pith citation observations for arXiv:2502.04354.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T11:47:17.656530Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:46.666638Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T16:34:25.763556Z
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 174bb4b2-35d2-4388-948b-939e6546f868 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 574f42ca-56cb-4807-8531-c3e721a74159 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0ff8392-f026-409e-b95a-8d437f5e521a · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Constitutional AI: Harmlessness from AI Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 421a4be3-99bf-4c32-88da-4b2d20da9c7f · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation def6c9fb-8d3f-4735-8a6f-57c21c1d186c · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment and Broderick, T
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 73be46cb-f446-40c5-bedd-122907673054 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment and Broderick, T
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c3a7be71-321d-4e9c-a0f1-b9e1e4c87fcc · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment and Verdinelli, I
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fdfe3e72-a9f0-48fd-a50f-37861c61d6bf · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment F., Leike, J., Brown, T., Martic, M., Legg, S., and Amodei, D
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d478326-31d9-4d69-ae24-51535ff72f5e · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment I., Santos, S
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e538dc2d-97f4-4c23-8dd9-ab195e63b329 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Reward Model Ensembles Help Mitigate Overoptimization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0647d1c7-36b5-4203-bac7-7901b6c74ce9 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment and Lindley, D
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f369319e-ccf3-45c4-84da-8814aa6852dd · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fabc0822-4643-490c-bc48-f3d97a5d7106 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment RLHF Workflow: From Reward Modeling to Online RLHF
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7222187e-abbb-4a67-bbde-25f6e9312ecc · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment The Llama 3 Herd of Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9202e05-ea5c-4e0c-9332-e853598b3078 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation eb7c576b-ffe9-4c0b-842d-b2cfb32a03f7 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment BoNBoN Alignment for Large Language Models and the Sweetness of Best-of-n Sampling
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a118c3d7-72b8-488b-a82e-5df92d1e56ba · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Bayesian Active Learning for Classification and Preference Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9867a51b-9551-426f-8040-bc6394a2b309 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a7a444ad-3e2a-4116-b84a-7a03e2cea1a5 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 44e5298d-977f-4d99-9569-d0ba02dd4071 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Y., Chandu, K., Dziri, N., Kumar, S., Zick, T., Choi, Y., Smith, N
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 05800bff-4fb0-469b-abc4-9d7c6776d3bc · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f60a0c9-e24e-48f1-9307-f3c5565b4fa2 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Statistical Rejection Sampling Improves Preference Optimization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19a4e2fc-69d7-4c01-b1a9-f1fe15219fd6 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Uncertainty-aware Reward Model: Teaching Reward Models to Know What is Unknown
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74abd0b1-95f9-4f5a-8379-137c39ff5c3c · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4216feaf-9aff-4e0b-b0be-172324f8c281 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4f263f93-3b0b-4669-90f3-6a7a378ae7f7 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Active Preference Learning for Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03b7d3ca-8995-4aaa-b5ea-ee8e23a7e105 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ad69f174-c741-4499-9a5d-ebdfd45bbb82 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8b5965a5-7467-42f9-bcd3-6b122384601f · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 33633452-c62a-4b11-be23-3328cb1c3c85 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Active Learning for Convolutional Neural Networks: A Core-Set Approach
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 630da295-44db-4efc-95eb-386572fc8886 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63a778b0-0f9f-4b9d-96c1-f592b4fc4b51 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3e6f97b-f54d-4bf2-b662-4252f89b9899 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e4c1d78b-fcf4-4cdd-a783-98b08dbba329 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Rethinking Bradley-Terry Models in Preference-Based Reward Modeling: Foundations, Theory, and Alternatives
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b591223-f09c-4cc6-aca1-575fccf4d264 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Gemma: Open Models Based on Gemini Research and Technology
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50417d67-ee5a-4c7d-a818-9f41f644a5dc · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b8e3b49-8a77-4747-8ce7-b0eb02a0a7c5 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ca588570-0b37-4682-85fb-bc775174bb3e · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment and Chris Glaze, B
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c0bdc399-e989-4d4e-92bf-b379aed4f951 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 791982b8-e588-4887-bc77-7e90d866e91f · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Secrets of RLHF in Large Language Models Part II: Reward Modeling
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f269a302-d70b-4d9f-af17-0cb8f115d568 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd4ef48f-0cae-4d75-8fc2-22714ee90684 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment MetaMetrics: Calibrating Metrics For Generation Tasks Using Human Preferences
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68877fec-7230-405e-b79d-0453e484ba40 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Iterative Preference Learning from Human Feedback: Bridging Theory and Practice for RLHF under KL-Constraint
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ed37742-e476-45a4-a5c3-fec5ba2729d7 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6ce621d-1846-4d13-b133-bd353dcd43ae · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9dd4566-6569-4195-97c2-5c64f7de45c9 · outbound
Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Overcoming Reward Overoptimization via Adversarial Policy Optimization with Lightweight Uncertainty Estimation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13b03f5f-35e5-4c53-8118-8a89fa713ed7 · inbound
OpenReview Should be Protected and Leveraged as a Community Asset for Research in the Era of Large Language Models Reviving The Classics: Active Reward Modeling in Large Language Model Alignment
Reference 135
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f256eb6-64ed-4697-934f-42d9e1153fb3 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Reviving The Classics: Active Reward Modeling in Large Language Model Alignment
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.