Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T17:15:27.898978Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2509.11026.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T17:15:27.898978Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T12:47:37.888316Z
A source-named dated measurement, never combined with another source.
Source: cited_works
18 of 18 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 178274f0-4833-4e24-b46d-ebaa34776a55 · outbound
Rethinking Human Preference Evaluation of LLM Rationales REV: Information-Theoretic Evaluation of Free-Text Rationales
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05322b96-4ea8-4209-8bb7-3085c46d7962 · outbound
Rethinking Human Preference Evaluation of LLM Rationales Deep reinforcement learning from human preferences
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a6dd9a8-749d-476d-b2fd-8ef1e599dc4a · outbound
Rethinking Human Preference Evaluation of LLM Rationales ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f6839e-b693-4227-8801-8a1b6fb6d790 · outbound
Rethinking Human Preference Evaluation of LLM Rationales Are Machine Rationales (Not) Useful to Humans? Measuring and Improving Human Utility of Free-Text Rationales
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2996221c-3aa0-4582-8b8d-822dc2b55ca9 · outbound
Rethinking Human Preference Evaluation of LLM Rationales 2 OLMo 2 Furious
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d60bde18-2647-4359-b013-92f3d56a5127 · outbound
Rethinking Human Preference Evaluation of LLM Rationales Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8e998bf-7eae-4b59-bb1d-43c97d681ccf · outbound
Rethinking Human Preference Evaluation of LLM Rationales ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea864c49-05f0-41b0-bb51-7c3494cd308c · outbound
Rethinking Human Preference Evaluation of LLM Rationales Tailoring Self-Rationalizers with Multi-Reward Distillation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd8281ac-fab5-489a-81d8-c777d494e750 · outbound
Rethinking Human Preference Evaluation of LLM Rationales PINTO: Faithful Language Reasoning Using Prompt-Generated Rationales
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a58348b-4a58-4cc3-ad89-25bfcc4479f0 · outbound
Rethinking Human Preference Evaluation of LLM Rationales Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3858e017-d4a9-4b83-8191-6be19cb770b7 · outbound
Rethinking Human Preference Evaluation of LLM Rationales cc/paper files/paper/2016/file/10a5ab2db37feedfdeaab192ead4ac0e-Paper.pdf
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e27b531-77ab-4288-b346-651276ab82ed · outbound
Rethinking Human Preference Evaluation of LLM Rationales Scott M Lundberg and Su-In Lee
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5de3ddad-9dde-4a3e-ab0f-98e9b50cd13a · outbound
Rethinking Human Preference Evaluation of LLM Rationales Explain Yourself! Leveraging Language Models for Commonsense Reasoning
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3243e181-b116-4136-88f6-dbe0913ea7a6 · outbound
Rethinking Human Preference Evaluation of LLM Rationales DecipherPref: Analyzing Influential Factors in Human Preference Judgments via GPT-4
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f740dcd0-d594-4249-84bb-70da27720075 · outbound
Rethinking Human Preference Evaluation of LLM Rationales Reframing Human-AI Collaboration for Generating Free-Text Explanations
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c877e8ca-d434-4b45-a792-6b23e27f6ff8 · outbound
Rethinking Human Preference Evaluation of LLM Rationales Faithfulness Tests for Natural Language Explanations
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9c9a3ff-c413-4d5f-845b-cc6fadf33474 · outbound
Rethinking Human Preference Evaluation of LLM Rationales Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b80c55b-79d6-4cf8-ac00-58ff29db57d8 · outbound
Rethinking Human Preference Evaluation of LLM Rationales Unresolved cited work
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50e6b86a-e429-4e72-a484-99f4ea2e78e5 · inbound
CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents Rethinking Human Preference Evaluation of LLM Rationales
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.