Pith. sign in

Paper Citation Record · LEDGER

Guiding LLM Decision-Making with Fairness Reward Models

As of 10 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2507.11344.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.11344 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:16:36.596455Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T03:04:43.641141Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cc0a3585-45aa-400c-838d-eacbb62a00c1 · outbound

This paper cites Measuring Gender and Racial Biases in Large Language Models.

Guiding LLM Decision-Making with Fairness Reward Models Measuring Gender and Racial Biases in Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:32.884393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:32.884393Z digest=sha256:8bdb38f65739a552335f78f6db0bbf8f9533a4331f405cf0805454997cc28cb1

Observation 4590bfba-9322-4d44-873f-0f74b1f7c85a · outbound

This paper cites Machine bias.

Guiding LLM Decision-Making with Fairness Reward Models Machine bias

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:40.945369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:32.959005Z digest=sha256:e7f131b2ad63e55e0181a42642a3611dafc508bc1c8201b3846f93aed63ae521

Observation 6913c2f1-197a-417e-ad0d-12708f06defb · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Guiding LLM Decision-Making with Fairness Reward Models Constitutional AI: Harmlessness from AI Feedback

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:33.065618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:33.065618Z digest=sha256:63e7ef39d09da20eec39559e6e1e3a0649eefb9ce585a67225bdee72df6fbfae

Observation bef5e4b2-8799-4ec8-9c16-6493857ec85b · outbound

This paper cites Fairness and Machine Learning.

Guiding LLM Decision-Making with Fairness Reward Models Fairness and Machine Learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:40.754525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:33.157482Z digest=sha256:187d70b9bca05de7488d3b6e5d9aa59661d419c2836185038c92d230f6adc8a5

Observation 3eb81b7a-0873-4112-a139-5ff3577d6dc2 · outbound

This paper cites Scaling test-time compute with open models, 2024.

Guiding LLM Decision-Making with Fairness Reward Models Scaling test-time compute with open models, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:33.257599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:33.257599Z digest=sha256:cba577c95173e70a673160ef1445fac8de80565cc1fb0618fb8b1a6730bee9ff

Observation c4fe649e-1b8b-4dc1-89ca-8e56190bfcda · outbound

This paper cites On the dangers of stochastic parrots: Can language models be too big? In Proceedings of the 2021 ACM conference on fairness, accountability, and transparency, 2021.

Guiding LLM Decision-Making with Fairness Reward Models On the dangers of stochastic parrots: Can language models be too big? In Proceedings of the 2021 ACM conference on fairness, accountability, and transparency, 2021

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:40.598198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:33.326485Z digest=sha256:30b404e254b4f646ff17ee6f0a1215e8ba1fe31e37bdc187e7a7988990f72be5

Observation dc893af8-465b-4303-8687-e6e0658b8567 · outbound

This paper cites Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment.

Guiding LLM Decision-Making with Fairness Reward Models Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:33.404516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:33.404516Z digest=sha256:d7369c1b0c20652abf4e82a8582238af93361327e3671be6dcb7df217f60b57a

Observation 92d27a1c-2a61-4ff2-90cc-dfa72e7dab6e · outbound

This paper cites Nuanced metrics for measuring unintended bias with real data for text classification.

Guiding LLM Decision-Making with Fairness Reward Models Nuanced metrics for measuring unintended bias with real data for text classification

Reference 8

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T17:16:37.508219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:33.483416Z digest=sha256:91cc186b5794739d30b87b59be0c8bcdd4ce90a815d9a7a92af99985445621cf

Observation 4262d806-a84a-47fd-b2a0-ee476739ab06 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Guiding LLM Decision-Making with Fairness Reward Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:33.556042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:33.556042Z digest=sha256:4fc52fd863f9c1fb1a8c53b06666c6d74d58520eba4eef0155a5e25d44978cc2

Observation 239b4c31-6913-46f3-9ac9-11e8803335b2 · outbound

This paper cites Alphamath almost zero: Process supervision without process.

Guiding LLM Decision-Making with Fairness Reward Models Alphamath almost zero: Process supervision without process

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:40.391505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:33.639423Z digest=sha256:ccde8df1ebc4eb11255d73d35cc2281f937c394ce9cff47055c3e56613758179

Observation dbf97b6b-53b1-4d69-a55f-9a8f909d4e79 · outbound

This paper cites Fair prediction with disparate impact: A study of bias in recidivism prediction instruments.

Guiding LLM Decision-Making with Fairness Reward Models Fair prediction with disparate impact: A study of bias in recidivism prediction instruments

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:40.258768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:33.697395Z digest=sha256:8aac90da6942e08190ff1a79468d72c35aa368c153095dd3577e6c5eed236214

Observation 343f7e75-8f04-4cd2-94e2-ecf33192fb98 · outbound

This paper cites Bias in bios: A case study of semantic representation bias in a high-stakes setting.

Guiding LLM Decision-Making with Fairness Reward Models Bias in bios: A case study of semantic representation bias in a high-stakes setting

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:33.731709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:33.731709Z digest=sha256:b0562f2bea8dbd57fa45b5d6d2f419fc55b261e96316288047c62c0128c7d3ae

Observation 00a43689-dcc9-4192-bb08-549b22fc27af · outbound

This paper cites Evaluation of A frican A merican language bias in natural language generation.

Guiding LLM Decision-Making with Fairness Reward Models Evaluation of A frican A merican language bias in natural language generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:33.808567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:33.808567Z digest=sha256:b8ab2f18be7930e920a19f107f3d3d0f185202b5d0922b5cb0dfce63cf192904

Observation f03e6d9e-b996-492b-aec9-b8026dcf5c23 · outbound

This paper cites The accuracy, fairness, and limits of predicting recidivism.

Guiding LLM Decision-Making with Fairness Reward Models The accuracy, fairness, and limits of predicting recidivism

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:40.061471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:33.911279Z digest=sha256:bdae7d8faf8cf2fd72ede83cac21c83bdf09e3d530db6389ceb73e8e44459a7a

Observation b10125fd-8937-4a2b-992b-c60ee033dfe7 · outbound

This paper cites Gaebler, Sharad Goel, Aziz Huq, and Prasanna Tambe.

Guiding LLM Decision-Making with Fairness Reward Models Gaebler, Sharad Goel, Aziz Huq, and Prasanna Tambe

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:33.978846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:33.978846Z digest=sha256:4448523c40faadaa6518487371a3dc8c25f1d2d06ad28c882b757c47fa94b048

Observation 0451b8b3-e928-4242-98f0-b59fb8bd9c21 · outbound

This paper cites Gallegos, Ryan A.

Guiding LLM Decision-Making with Fairness Reward Models Gallegos, Ryan A

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:34.020800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:34.020800Z digest=sha256:a4bcf84507ded837a6f86fd071826dbfc1d9dc80c396047682b5747515a61a36

Observation a90cb941-9963-4151-ac92-b617a720e87a · outbound

This paper cites Debiasing pre-trained language models via efficient fine-tuning.

Guiding LLM Decision-Making with Fairness Reward Models Debiasing pre-trained language models via efficient fine-tuning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:34.109032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:34.109032Z digest=sha256:c7aaef3280ad73244c637ea47ad7d776c5aef1df67668304ee9bb37dbc8d8dcb

Observation 7b79ff65-f079-4643-808a-ac6998336767 · outbound

This paper cites Bias in Large Language Models: Origin, Evaluation, and Mitigation.

Guiding LLM Decision-Making with Fairness Reward Models Bias in Large Language Models: Origin, Evaluation, and Mitigation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:34.192370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:34.192370Z digest=sha256:f69bb56a21897c2fb080f477fb74152dd76841e3e4aa079eb4852a2e31008217

Observation bd06d18d-8403-4b1c-ab91-237c5b3ab031 · outbound

This paper cites Equality of opportunity in supervised learning.

Guiding LLM Decision-Making with Fairness Reward Models Equality of opportunity in supervised learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:39.817344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:34.276605Z digest=sha256:336180a79bb0708cea4f823f77523f35a26dc0e0c73a19cc8787c8529327597d

Observation 38241025-f83b-4642-b5c0-b8ff424c9674 · outbound

This paper cites V- ST ar: Training verifiers for self-taught reasoners.

Guiding LLM Decision-Making with Fairness Reward Models V- ST ar: Training verifiers for self-taught reasoners

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:39.593074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:34.383923Z digest=sha256:d1fac64aac229b3a42c062f2bdfc2f25736d8c20059fa5c3cdb3dfb583879140

Observation 947a5ebd-3ab4-4f84-bdae-3aba1c02254e · outbound

This paper cites Prompting Techniques for Reducing Social Bias in LLMs through System 1 and System 2 Cognitive Processes.

Guiding LLM Decision-Making with Fairness Reward Models Prompting Techniques for Reducing Social Bias in LLMs through System 1 and System 2 Cognitive Processes

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:34.481000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:34.481000Z digest=sha256:62d4c6fc7ba1b6b296f53197af0c19a76c87dbbed59692969ceb94cd8bebb1f2

Observation 2332d80b-b2c0-4c99-a2ea-c78009ee6401 · outbound

This paper cites Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting.

Guiding LLM Decision-Making with Fairness Reward Models Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:34.600067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:34.600067Z digest=sha256:3d2ca272d8de9f324522fa21651ee9d21a35cc0bd5f1b72e20447ca23e978571

Observation a4162730-f66a-4fd6-a73f-a2ff60492451 · outbound

This paper cites Gender bias and stereotypes in large language models.

Guiding LLM Decision-Making with Fairness Reward Models Gender bias and stereotypes in large language models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:39.391977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:34.673173Z digest=sha256:ef2dec9b83e25a3dd7829cdded2ed3aa67aba01cd80f30814e0a134382c2a22b

Observation 53e20214-70cc-4ab7-9c6c-fc3770a5e16d · outbound

This paper cites When do pre-training biases propagate to downstream tasks? a case study in text summarization.

Guiding LLM Decision-Making with Fairness Reward Models When do pre-training biases propagate to downstream tasks? a case study in text summarization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:34.751145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:34.751145Z digest=sha256:0c9f4d52de8c63ccec62fa6257cb947a71fe798068ac74cb1d884b9a2cfc7bd5

Observation be259f34-6ba6-4adb-8604-6aafe16ff023 · outbound

This paper cites Let's verify step by step.

Guiding LLM Decision-Making with Fairness Reward Models Let's verify step by step

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:34.831656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:34.831656Z digest=sha256:c86e87ab56884f726003a99a03b8ae1d3be54a48771ab9a64e0a42b903b6c3c5

Observation 190bf357-95f7-4a39-aaa5-303b39125cf7 · outbound

This paper cites Debiasing large language models with structured knowledge.

Guiding LLM Decision-Making with Fairness Reward Models Debiasing large language models with structured knowledge

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:34.904157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:34.904157Z digest=sha256:2189bde1d66a3043a3e932d202beb7b2ff05f1b0a1c93ef9ca4575a57471de34

Observation 2598aec9-9663-4997-98ea-eabef42f61dc · outbound

This paper cites Fairness-guided few-shot prompting for large language models.

Guiding LLM Decision-Making with Fairness Reward Models Fairness-guided few-shot prompting for large language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:39.143969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:34.978458Z digest=sha256:b4b3c6a2e13264fd4f8da334092a8ffeee0d2f469f3663a9f01219c5de709f54

Observation 9462b191-baa9-4f9c-94b9-b23826ce9b3b · outbound

This paper cites Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models.

Guiding LLM Decision-Making with Fairness Reward Models Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:16:37.148249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:35.106211Z digest=sha256:9b402ff51141c02d2b11b4e2ae284b4bd42efe8f751d99d8a2bac75fbf6b22a4

Observation 859b9ccb-1d88-421a-a44e-d564ed349bdf · outbound

This paper cites GPT-4 Technical Report.

Guiding LLM Decision-Making with Fairness Reward Models GPT-4 Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.181824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.181824Z digest=sha256:513c676af056806af22fdb82a47725922effde248ab4426ac6426c2dd107f564

Observation 783d5837-a48e-499d-a173-cb7bbc3cbe03 · outbound

This paper cites Bias in word embeddings.

Guiding LLM Decision-Making with Fairness Reward Models Bias in word embeddings

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.257114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.257114Z digest=sha256:e1863c8a6375c3b4eb452b5b63a1063c359a4b1f9e4737a6370b3f2477829930

Observation 9e44a4c7-fa98-46e4-9acf-a30085eee4c7 · outbound

This paper cites BBQ : A hand-built bias benchmark for question answering.

Guiding LLM Decision-Making with Fairness Reward Models BBQ : A hand-built bias benchmark for question answering

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.331592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.331592Z digest=sha256:bd70d67d949b5b1541c35c5b36b6913553b461462c3ae5f6f7f48fc474a32e7a

Observation 9cc90f10-3013-4f17-84e3-a0c5d69a5fa5 · outbound

This paper cites Divine LL a MA s: Bias, stereotypes, stigmatization, and emotion representation of religion in large language models.

Guiding LLM Decision-Making with Fairness Reward Models Divine LL a MA s: Bias, stereotypes, stigmatization, and emotion representation of religion in large language models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.390548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.390548Z digest=sha256:c42dd4f43ceca4723dd938ee4f94eb33cf217ab1ceeb285d84525a4b0c3c1797

Observation 4ffbd460-dc6e-49e7-ad19-a74ea5f316da · outbound

This paper cites Proximal Policy Optimization Algorithms.

Guiding LLM Decision-Making with Fairness Reward Models Proximal Policy Optimization Algorithms

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.420119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.420119Z digest=sha256:2019623f309c4195d5b44e58494813bff88e8faa3ed66d641e5e727c455e4aa6

Observation e61e2c6e-e099-4736-b1a5-13274a76cb76 · outbound

This paper cites On second thought, let ' s not think step by step! bias and toxicity in zero-shot reasoning.

Guiding LLM Decision-Making with Fairness Reward Models On second thought, let ' s not think step by step! bias and toxicity in zero-shot reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.468808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.468808Z digest=sha256:3a1c7b2118ae831ce87caa349fc5f5c81ac8733e33f3c181a4f42a3cc2436313

Observation 245ea5f8-203a-4e87-bff7-c98f74010643 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Guiding LLM Decision-Making with Fairness Reward Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.536089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.536089Z digest=sha256:f0fcce5a4a672e563c60544d2cb1520fc5ef0328ee60c707479dc477c43565a2

Observation 030d9034-1f56-404f-8756-deceb6222eda · outbound

This paper cites Scaling LLM test-time compute optimally can be more effective than scaling parameters for reasoning.

Guiding LLM Decision-Making with Fairness Reward Models Scaling LLM test-time compute optimally can be more effective than scaling parameters for reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.586343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.586343Z digest=sha256:a4f404c8f832807620babe8f241830cb76615e990464ed6eb194a2508b08b0bd

Observation 95250af8-c97c-46be-a444-562177b8bedc · outbound

This paper cites Unveiling Gender Bias in Terms of Profession Across LLMs: Analyzing and Addressing Sociological Implications.

Guiding LLM Decision-Making with Fairness Reward Models Unveiling Gender Bias in Terms of Profession Across LLMs: Analyzing and Addressing Sociological Implications

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.620778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.620778Z digest=sha256:144788e923677a0cbcff8bc91cd2b6648bfcb7306ea9ab04b94aa636f0f45993

Observation 8608611d-f52e-4665-9817-335e646aebb2 · outbound

This paper cites Toward self-improvement of llms via imagination, searching, and criticizing.

Guiding LLM Decision-Making with Fairness Reward Models Toward self-improvement of llms via imagination, searching, and criticizing

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:38.931655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:35.716278Z digest=sha256:67ddea8f3f5acf25ad4c447ed4355bd096bf9736238ed30360cd75113eb38379

Observation d7875b3e-8f1c-4a5f-8efc-69c0403a9a3d · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Guiding LLM Decision-Making with Fairness Reward Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.783003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.783003Z digest=sha256:19e9c34cf7b2e64777984b66ea780ec936ab51bcad7dee821cf8f26a271cd4de

Observation 6c9db766-0caa-4e05-9fec-c551aea0eb4b · outbound

This paper cites an unresolved cited work.

Guiding LLM Decision-Making with Fairness Reward Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:16:38.742876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:35.833352Z digest=sha256:b31dfc25d86e72a77b9a81ae26296039a6fb88d3008356ada967a5096250dee9

Observation 3dc15b25-52bb-478c-806a-b86504c164a0 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Guiding LLM Decision-Making with Fairness Reward Models Solving math word problems with process- and outcome-based feedback

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.915640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.915640Z digest=sha256:3b0e92e3473d6c291c16490857541b29a1c8adb0bda5eaec6f9b2989d8d4ded5

Observation b12e44ce-f502-4ebc-bb0c-56257c9e3fd6 · outbound

This paper cites ``kelly is a warm person, joseph is a role model'': Gender biases in LLM -generated reference letters.

Guiding LLM Decision-Making with Fairness Reward Models ``kelly is a warm person, joseph is a role model'': Gender biases in LLM -generated reference letters

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:35.989318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:35.989318Z digest=sha256:2f1ce7aeedc72d8f4deb5ac26ca1e26c0d72fafef8004413cf64e2af7b7a4609

Observation 29756c0c-8d77-4321-9d5a-05cec139544b · outbound

This paper cites Alphazero-like tree-search can guide large language model decoding and training.

Guiding LLM Decision-Making with Fairness Reward Models Alphazero-like tree-search can guide large language model decoding and training

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:38.444517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:36.045452Z digest=sha256:41a62d11edb4daa3b7d376c4a9f4ff80e090c1f1b3efc7dfa6a8b051766dc4fd

Observation e181325b-299b-43e9-bb30-c8b23d172bf4 · outbound

This paper cites Math-shepherd: Verify and reinforce LLM s step-by-step without human annotations.

Guiding LLM Decision-Making with Fairness Reward Models Math-shepherd: Verify and reinforce LLM s step-by-step without human annotations

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:36.088433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:36.088433Z digest=sha256:7b33fa5cd67655f3b1ff15c8e3257b690281d529d7c0b6449e1409e1420f97ad

Observation 7c31743d-7da5-4686-854c-00a05a8c5382 · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

Guiding LLM Decision-Making with Fairness Reward Models Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:36.146401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:36.146401Z digest=sha256:bd17b943e7cd25a8017a5fb40d697ef5f0c538f8a02498f1138bcba1c3dd7e96

Observation 772ded21-26f9-4b76-b258-59b652fba045 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Guiding LLM Decision-Making with Fairness Reward Models Chain-of-thought prompting elicits reasoning in large language models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:36.231787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:36.231787Z digest=sha256:dec21227e297f2d773b29d8dfa72c976b1e27a0c1350b8a2757c3f42866924c7

Observation 88bccc54-0090-4147-bc83-1e6642df3a25 · outbound

This paper cites Blueprint for an ai bill of rights: Making automated systems work for the american people, 2022.

Guiding LLM Decision-Making with Fairness Reward Models Blueprint for an ai bill of rights: Making automated systems work for the american people, 2022

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:38.256514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:36.304543Z digest=sha256:c82db444382f9010321a8e6670d35533e20afb6272ceba4fb48ad2131843d99f

Observation c54b5ecf-f6f7-4f39-9b13-2616c96509ba · outbound

This paper cites Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval, page 1578–1590.

Guiding LLM Decision-Making with Fairness Reward Models Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval, page 1578–1590

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:38.048677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:36.347103Z digest=sha256:c8cfe583ce84feef8ed34a663b249e3856a6f7584937d8263f6b4f0500c52cad

Observation 1d5547e0-35a6-46fd-99a2-8b8afe9eff28 · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Guiding LLM Decision-Making with Fairness Reward Models Tree of thoughts: Deliberate problem solving with large language models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:37.868649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:36.376580Z digest=sha256:0f725a511078dcc8a81c2fbcc0e92821b357c9a4d65c7b8dfc75ff3f615dad1d

Observation cb3dabd6-abbe-4d12-8f41-a7466d145002 · outbound

This paper cites Scaling Relationship on Learning Mathematical Reasoning with Large Language Models.

Guiding LLM Decision-Making with Fairness Reward Models Scaling Relationship on Learning Mathematical Reasoning with Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T17:16:36.465705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:16:36.465705Z digest=sha256:cece11430c73a6c57fdeee0d75f416db52cdbbc1e6d3a0574b6858fcfe1aac3c

Observation 2b9ef331-a297-4a24-8175-a6f9d936f503 · outbound

This paper cites Star: Bootstrapping reasoning with reasoning.

Guiding LLM Decision-Making with Fairness Reward Models Star: Bootstrapping reasoning with reasoning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:37.780741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:36.543036Z digest=sha256:7679d2848afa24b45c868ba868214d1aaa0e358e5cc3d530127200a56606d983

Observation a49fec15-69e9-48ca-ac85-03ab78d9f5f8 · outbound

This paper cites Towards effective discrimination testing for generative ai.

Guiding LLM Decision-Making with Fairness Reward Models Towards effective discrimination testing for generative ai

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:16:37.675829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T17:16:36.596455Z digest=sha256:ee47ebdcd8215f311e115173f491693af06cb7c4c21b588a395995b547ffd74f

Pith citing papers

Observation 9610f636-a3ec-4be5-8de9-6de0de29ed12 · inbound

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation cites this paper.

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation Guiding LLM Decision-Making with Fairness Reward Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T03:04:43.641141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:04:43.641141Z digest=sha256:8a0839783ffec98c986650b9b18cb319568b48317454fbb0837009de44383828