Pith. sign in

Paper Citation Record · LEDGER

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time

As of 10 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2505.23729.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23729 v2

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:44:21.494870Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-07T16:37:58.860183Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T23:36:36.042184Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 891401d7-bf5f-47c7-9988-f9ac8ed458a0 · outbound

This paper cites GPT-4 Technical Report.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.523685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.523685Z digest=sha256:06d5e29e8c4683742e94a7e4f045499f6ce97c64a1d9cd58bf0bef7e7f1df1ae

Observation ee71dafe-1d62-4a28-8fbe-9dbfed2a7323 · outbound

This paper cites A General Language Assistant as a Laboratory for Alignment.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time A General Language Assistant as a Laboratory for Alignment

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.570449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.570449Z digest=sha256:1b916d0f819650f2c41b69711c7871ac57e9ac0f991b6d257461adbef6b44ce7

Observation 337d917b-3659-4729-a39a-759dfda77eb1 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.670874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.670874Z digest=sha256:4862d840d1c49bfeb035766d23c5e47c00fc4317f00fb187aec452c433192e8b

Observation 1c67c8a4-6df2-4479-8e87-fd1722545232 · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.715883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.715883Z digest=sha256:22c6f67e6bcc6c6ada1d50a33e97e3bc73032f8aff3af02729d2428c64d0a8ef

Observation 3402ccd8-14fa-4630-85fb-68f2df3180cb · outbound

This paper cites Transfer Q Star: Principled Decoding for LLM Alignment.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Transfer Q Star: Principled Decoding for LLM Alignment

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.777137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.777137Z digest=sha256:cf3f3989fdafc99f05d1e4e434fd82d3ed93dad0c6ad3e9a5840e2c702d4f447

Observation c8a48e54-160c-4bad-be69-9dc30f79c723 · outbound

This paper cites Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.813333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.813333Z digest=sha256:d6d08263ce502723d16bffd190a6279f67c46885d035c81e1b74851ac173dd8b

Observation 2b558640-201d-451f-865b-bbbced157d82 · outbound

This paper cites Safe rlhf: Safe reinforcement learning from human feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Safe rlhf: Safe reinforcement learning from human feedback

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:44:23.483038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:19.848413Z digest=sha256:5215244890f3deb6ba857e43983b25651cf7ee593fbfd9593b2f9f817a4c2d0e

Observation f8a887c8-851c-46d3-9390-2239f5b8c0b1 · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.890936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.890936Z digest=sha256:bc18e88ca78badc5502e99ae640449b4ee88b978711b7adf26cfdccf2c789cd4

Observation 54d9b4db-9fb7-4734-9ae9-fff583816ab7 · outbound

This paper cites LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.930035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.930035Z digest=sha256:5e41f6e9add9d5e32be81dfc70d75b3299dfef3ee53668061374af53fb613908

Observation ffc54cc7-49e8-4bb7-aa0c-757e345d2a40 · outbound

This paper cites Improving alignment of dialogue agents via targeted human judgements.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Improving alignment of dialogue agents via targeted human judgements

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.989378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.989378Z digest=sha256:5a1529f8a5778457899be64ea8a36aeb54ff5248e5ac323d8a564b4674dc9f5a

Observation ceadafea-0f80-4831-9f45-f568c50a5e4a · outbound

This paper cites Y., Sengupta, S., Bonadiman, D., Lai, Y.-a., Gupta, A., Pappas, N., Mansour, S., Kirchhoff, K., and Roth, D.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Y., Sengupta, S., Bonadiman, D., Lai, Y.-a., Gupta, A., Pappas, N., Mansour, S., Kirchhoff, K., and Roth, D

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.011975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.011975Z digest=sha256:ec3d5d0beaf2bb04a3b26344cee6cb567d647a153ccc83f4de92002ed0afe6f7

Observation 61a98890-281d-4596-b4b6-033b3d05fab0 · outbound

This paper cites One-Shot Safety Alignment for Large Language Models via Optimal Dualization.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time One-Shot Safety Alignment for Large Language Models via Optimal Dualization

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:44:21.936517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:20.062006Z digest=sha256:f05179a2cc681f7c739cce9234ff911292df8e2ad034e9c242635393b16ba3e4

Observation 4abcb6a1-a83f-4656-9288-eb089b4004cb · outbound

This paper cites Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.132271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.132271Z digest=sha256:29054d585ac7b374673080eaa527e0cdc3d70e9284833e30b9fecf86d4ae070c

Observation 75620268-4462-4d24-b277-de0d52d9690e · outbound

This paper cites PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.174221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.174221Z digest=sha256:71fd2f94700d87c323622cb50e82a8091178991342ddc644187c5b5d54ba31f9

Observation c3258aed-6092-456c-8e05-fe4a1ca085de · outbound

This paper cites ARGS: Alignment as Reward-Guided Search.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time ARGS: Alignment as Reward-Guided Search

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.211826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.211826Z digest=sha256:8c1fe487eb011dea8ea901f6a93a9603114352e697f1c38b112d0861494a3988

Observation e7c2b313-9bb8-48d7-872c-9e44e93999a4 · outbound

This paper cites Chain of Hindsight Aligns Language Models with Feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Chain of Hindsight Aligns Language Models with Feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.272639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.272639Z digest=sha256:5e62735de59aa9112dea900a9a59dc344eb0dd0fd452cd568472b16d8ac6b791

Observation c4a2c54a-ce43-422a-8f00-1e73656fbb5c · outbound

This paper cites L., Daly, R.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time L., Daly, R

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.378678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.378678Z digest=sha256:a06230f2ef1f8754b29cd57243f4d91f4aa68b709d4ae527d390df07f0983710

Observation 22dfaf66-14ba-4494-a9a8-68b624cc6a72 · outbound

This paper cites Controlled Decoding from Language Models.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Controlled Decoding from Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.437302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.437302Z digest=sha256:d8dd9359638a5b2f1fa5c6f545a1cc5ed2051070d6ee1bea56616e37b3197513

Observation 15dd78a4-4dfc-4755-8df5-edf90a34b73e · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time WebGPT: Browser-assisted question-answering with human feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.495851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.495851Z digest=sha256:7657bae510856e49925027f8034d3f8f5b05900b1d7dfac822c3d4d122c47da8

Observation 65e6c99c-b9f3-4de9-8e2f-26867fc6a3e0 · outbound

This paper cites and Ozdaglar, A.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time and Ozdaglar, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:44:23.367957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:20.531442Z digest=sha256:edd080c7d92080860399515d014ba409490424112c172ff757fc34be5793eb49

Observation 203f129a-4eb8-476a-89f7-a4d8295f03b5 · outbound

This paper cites Training language models to follow instructions with human feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Training language models to follow instructions with human feedback

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.625568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.625568Z digest=sha256:e5432f2bc5fa9c67c111601a365f700f694b2f4ab4b5fbb413c76a10e8301b3f

Observation 40b10372-d403-445e-adc2-30945b7b7ae3 · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.716463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.716463Z digest=sha256:e3546da27aa9ce4abcf29fb0fe5d34d2891bf96e35254a54ced5f0241a4f2367

Observation 5ebeae0c-190a-4e3c-80ce-7337da08b8dd · outbound

This paper cites D., Ermon, S., and Finn, C.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time D., Ermon, S., and Finn, C

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:44:23.230777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:20.846734Z digest=sha256:45f9dca4de29ebcd0cb24c066f1a2121766ba2660494a0d1393489440b7c714e

Observation 13965ca7-aa9b-49e0-8301-5245044ce286 · outbound

This paper cites D., Ermon, S., and Finn, C.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time D., Ermon, S., and Finn, C

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.920937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.920937Z digest=sha256:cd4346f94cb649d97443d6dcd1e67094a2730bd61cfc6a7a8e9e34b3ed46c073

Observation 0db0aa03-e18c-47eb-abb0-17911a8b4930 · outbound

This paper cites A., and Du, S.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time A., and Du, S

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:44:23.044947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:20.960306Z digest=sha256:0b7c69550e4f9bd6c14d188eb9ed2689a63aad708b5005709c47a0338b20f570

Observation 741c5dab-a3fa-48ca-bc1b-d8a5fa8bd3ea · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:44:22.879347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:21.004124Z digest=sha256:a97b50d104010d9f36db89cc27abd3b4f3dd2fbef20fcc138d13b8fb94e29c50

Observation dd9d3a93-36f1-439e-9f1b-53c7170e4048 · outbound

This paper cites S., Tang, X., and Bogunovic, I.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time S., Tang, X., and Bogunovic, I

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:21.054411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:21.054411Z digest=sha256:d499e5b43e9d94173d160aa355f183638469a22ba4f7b6e50eca7aac3cbbe72d

Observation 24c55602-d5f2-498e-abe4-7162cafa0752 · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:44:22.669996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:21.100780Z digest=sha256:cde289087e8e5046ecbf42e78334adf28c121ab3cef144e61c254044cf698496

Observation f7f74f73-2063-42e7-9102-3bc219834832 · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:44:22.525760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:21.296915Z digest=sha256:44247a35f79f2e97508fd47deb31a1f181ab964d30a0ba8fbbf450ec320eae70

Observation 51c96f80-76e3-460b-8126-d49253d2555d · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:44:22.365506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T12:44:21.338251Z digest=sha256:ff20d18a5752199c1dcaa1727d25f8f34e65ab0178ad1cdd034477d25ac5837d

Observation 7a22bb6e-889f-4edf-9875-f3ce74226827 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:21.380872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:21.380872Z digest=sha256:2f8d86707fcd4713b0671e1313301ce9ccb8743d54cfc18383a462b510268430

Observation c601bb67-6395-46bb-bcef-466d80a3ef90 · outbound

This paper cites M., and Wolf, T.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time M., and Wolf, T

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:21.446086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:21.446086Z digest=sha256:27ff83f64124b75e67cfbfcf3893fb6fa14a68f764f8c33b68ae482c12390fbf

Observation 95c77aef-c487-46b2-87fe-2c1dbbb47ae5 · outbound

This paper cites write newline.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time write newline

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:21.494870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:21.494870Z digest=sha256:91eccd8d9089ef4d385888a6b25e5a48c89e0aed8c16bcd0aa1f3eb1affe62f3

Pith citing papers

Observation 8f601549-2e01-452d-b487-9eec480b1316 · inbound

Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest cites this paper.

Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:36.044970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-07T16:37:58.860183Z digest=sha256:9e614278a581fd873bdef24fb91531f2efb78137d833792c2a23792569a88a92