Pith. sign in

Paper Citation Record · LEDGER

Rethinking Human Preference Evaluation of LLM Rationales

As of 19 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2509.11026.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.11026 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:15:27.898978Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:47:37.888316Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 178274f0-4833-4e24-b46d-ebaa34776a55 · outbound

This paper cites REV: Information-Theoretic Evaluation of Free-Text Rationales.

Rethinking Human Preference Evaluation of LLM Rationales REV: Information-Theoretic Evaluation of Free-Text Rationales

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.861042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.861042Z digest=sha256:6bf1ac6e3b9366d2b5339c7a567f270af792508d96035dc6e874b0310b17c937

Observation 05322b96-4ea8-4209-8bb7-3085c46d7962 · outbound

This paper cites Deep reinforcement learning from human preferences.

Rethinking Human Preference Evaluation of LLM Rationales Deep reinforcement learning from human preferences

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.866035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.866035Z digest=sha256:a5ec9cff1fbd41ce06a29de6133f8cec2320c8a5388fd5c8077e9ebf8eb55ab2

Observation 4a6dd9a8-749d-476d-b2fd-8ef1e599dc4a · outbound

This paper cites ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning.

Rethinking Human Preference Evaluation of LLM Rationales ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.868447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.868447Z digest=sha256:7812f2cb6613ef7519ca4df3f06f2a2d95174c1512e7527cf3d4e2a972c7938e

Observation 71f6839e-b693-4227-8801-8a1b6fb6d790 · outbound

This paper cites Are Machine Rationales (Not) Useful to Humans? Measuring and Improving Human Utility of Free-Text Rationales.

Rethinking Human Preference Evaluation of LLM Rationales Are Machine Rationales (Not) Useful to Humans? Measuring and Improving Human Utility of Free-Text Rationales

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.875563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.875563Z digest=sha256:802615c679d9b082bf6d4dc34d5ed534808c7fffb8f7fdbcfece74a73153b412

Observation 2996221c-3aa0-4582-8b8d-822dc2b55ca9 · outbound

This paper cites 2 OLMo 2 Furious.

Rethinking Human Preference Evaluation of LLM Rationales 2 OLMo 2 Furious

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.882256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.882256Z digest=sha256:e37983c0be25c558be90114bd0ac38b7ea1845f80ad3db65967e313951608d7e

Observation d60bde18-2647-4359-b013-92f3d56a5127 · outbound

This paper cites Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al.

Rethinking Human Preference Evaluation of LLM Rationales Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.884639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.884639Z digest=sha256:98878c9faa627953d340bf469391e5918dd614a21595bf569d69660bbc5730bb

Observation b8e998bf-7eae-4b59-bb1d-43c97d681ccf · outbound

This paper cites ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness.

Rethinking Human Preference Evaluation of LLM Rationales ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.887155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.887155Z digest=sha256:2984ebbbc4edb4bd1bcea413d75e701b510b992ceb682bc7c95ed387d51a9b1e

Observation ea864c49-05f0-41b0-bb51-7c3494cd308c · outbound

This paper cites Tailoring Self-Rationalizers with Multi-Reward Distillation.

Rethinking Human Preference Evaluation of LLM Rationales Tailoring Self-Rationalizers with Multi-Reward Distillation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.891750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.891750Z digest=sha256:61d085af1f5070a3fbfe2304812741bd6fd32ddcef9cb52cceeeff5ee9dde28b

Observation fd8281ac-fab5-489a-81d8-c777d494e750 · outbound

This paper cites PINTO: Faithful Language Reasoning Using Prompt-Generated Rationales.

Rethinking Human Preference Evaluation of LLM Rationales PINTO: Faithful Language Reasoning Using Prompt-Generated Rationales

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.894185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.894185Z digest=sha256:d64d55e0947f4672669797ac4b874ce83b01c8fbbeb1f98bed091b4f83dab9b3

Observation 9a58348b-4a58-4cc3-ad89-25bfcc4479f0 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Rethinking Human Preference Evaluation of LLM Rationales Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.898978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.898978Z digest=sha256:6679b0732042e1a636a2add43ffd4c61fc1be0c97383fb581c01704e793fdeb1

Observation 3858e017-d4a9-4b83-8191-6be19cb770b7 · outbound

This paper cites cc/paper files/paper/2016/file/10a5ab2db37feedfdeaab192ead4ac0e-Paper.pdf.

Rethinking Human Preference Evaluation of LLM Rationales cc/paper files/paper/2016/file/10a5ab2db37feedfdeaab192ead4ac0e-Paper.pdf

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.880107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.880107Z digest=sha256:b8cfde43893ab7f833e6cdbbdb5673d8851b6ee186fe1d9db0614a0edf232fea

Observation 1e27b531-77ab-4288-b346-651276ab82ed · outbound

This paper cites Scott M Lundberg and Su-In Lee.

Rethinking Human Preference Evaluation of LLM Rationales Scott M Lundberg and Su-In Lee

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.877843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.877843Z digest=sha256:143e5efcc2ba5bb65f2763aebd82b03d6311e5447668cacb9c0a1a6ba891f2e5

Observation 5de3ddad-9dde-4a3e-ab0f-98e9b50cd13a · outbound

This paper cites Explain Yourself! Leveraging Language Models for Commonsense Reasoning.

Rethinking Human Preference Evaluation of LLM Rationales Explain Yourself! Leveraging Language Models for Commonsense Reasoning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.889372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.889372Z digest=sha256:4dafac53223bc01c0579bd944e843bd5198cd3ae99a93eae001b953d72270e94

Observation 3243e181-b116-4136-88f6-dbe0913ea7a6 · outbound

This paper cites DecipherPref: Analyzing Influential Factors in Human Preference Judgments via GPT-4.

Rethinking Human Preference Evaluation of LLM Rationales DecipherPref: Analyzing Influential Factors in Human Preference Judgments via GPT-4

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.873148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.873148Z digest=sha256:0e8664863ca047e84dea6852f19ad9c9898655be33560f0badb01f36e34060c7

Observation f740dcd0-d594-4249-84bb-70da27720075 · outbound

This paper cites Reframing Human-AI Collaboration for Generating Free-Text Explanations.

Rethinking Human Preference Evaluation of LLM Rationales Reframing Human-AI Collaboration for Generating Free-Text Explanations

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.896357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.896357Z digest=sha256:4a2997c452f385d6ea001985b2d66623c1f255584507774779e15dfabb63a612

Observation c877e8ca-d434-4b45-a792-6b23e27f6ff8 · outbound

This paper cites Faithfulness Tests for Natural Language Explanations.

Rethinking Human Preference Evaluation of LLM Rationales Faithfulness Tests for Natural Language Explanations

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.857734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.857734Z digest=sha256:c1bd4d9d87a225ce14804babb73fdba9555b7666842b7fd896ff75596c3f4af3

Observation b9c9a3ff-c413-4d5f-845b-cc6fadf33474 · outbound

This paper cites Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference.

Rethinking Human Preference Evaluation of LLM Rationales Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.863520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.863520Z digest=sha256:59e174d0c71d74f422fda3d168bff3c4323e6b997472c99049d3770dcb792bac

Observation 1b80c55b-79d6-4cf8-ac00-58ff29db57d8 · outbound

This paper cites an unresolved cited work.

Rethinking Human Preference Evaluation of LLM Rationales Unresolved cited work

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.870816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.870816Z digest=sha256:238bcc8be188b2f0718b71c481713e344eb8fb2fba7c9bd7b60c19e26313f281

Pith citing papers

Observation 50e6b86a-e429-4e72-a484-99f4ea2e78e5 · inbound

CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents cites this paper.

CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents Rethinking Human Preference Evaluation of LLM Rationales

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T12:47:37.888316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:47:37.888316Z digest=sha256:1ae4f9f6a47d522840c1fde65a3fccd6e5ac733d2b42c33c54fb511106fd9f97