Pith. sign in

Paper Citation Record · LEDGER

Prompt Optimization with Human Feedback

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2405.17346.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.17346 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:00:22.168152Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T11:39:47.126608Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a0f5423e-9fe3-47f0-8768-6d997c86d5ad · inbound

Meta-Prompt Optimization for LLM-Based Sequential Decision Making cites this paper.

Meta-Prompt Optimization for LLM-Based Sequential Decision Making Prompt Optimization with Human Feedback

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-09T18:00:22.168152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:00:22.168152Z digest=sha256:e078f869631dbaa75f5bf27313759faa8d31bc49f270a20f1e91dd82b119885f

Observation 5575a231-c644-4859-81b7-f862f73dcddb · inbound

Federated Linear Dueling Bandits cites this paper.

Federated Linear Dueling Bandits Prompt Optimization with Human Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T16:44:30.641818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:44:30.641818Z digest=sha256:0cde3b3cfeee0fc8f260a38c979788cd4ac3f48d94c3fcd322544b3f6f3964a1

Observation 13f79b30-3afd-46c1-8894-1d1359c4d537 · inbound

Large Language Model-Enhanced Multi-Armed Bandits cites this paper.

Large Language Model-Enhanced Multi-Armed Bandits Prompt Optimization with Human Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T16:35:28.419518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:35:28.419518Z digest=sha256:5205649f45fdab03eb95e3d1cbe73657442f76bb47a115389e0d99893c473268

Observation 9ecf1f23-fa15-468a-a815-5e36c844cb73 · inbound

Online Clustering of Dueling Bandits cites this paper.

Online Clustering of Dueling Bandits Prompt Optimization with Human Feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T13:34:48.775041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T13:34:48.775041Z digest=sha256:a524877f241d21a3598dbf869cbf2f18c7e54f7c01d6d9703e461680f37c0d17

Observation c10bf89f-d525-48b5-8f5b-0b0875d8261d · inbound

Aligning LLMs by Predicting Preferences from User Writing Samples cites this paper.

Aligning LLMs by Predicting Preferences from User Writing Samples Prompt Optimization with Human Feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:21.687814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:30:21.687814Z digest=sha256:59f166fe7bda7d025cd86bcb4b0ddd959a66f904634a0a8f8d4cce6c3f9efee8

Observation d7feab75-c05f-44d6-ab4b-0c84ef5797b9 · inbound

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems cites this paper.

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems Prompt Optimization with Human Feedback

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:21:42.516204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T23:21:42.029285Z digest=sha256:26437dfbd90304f231de971216afb4a9582fc7e56969cad3881a882d7a2ed579

Observation 817e95e2-066b-46db-9f69-96c35017a432 · inbound

Deploying AI for Signal Processing education: Selected challenges and intriguing opportunities cites this paper.

Deploying AI for Signal Processing education: Selected challenges and intriguing opportunities Prompt Optimization with Human Feedback

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-04T19:59:17.505827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:59:17.505827Z digest=sha256:65dab883bfcd692e90f7b158fe581c64741fd23177b4dd23c509993d01575af8

Observation 38e76b08-d09e-4250-9d1b-ac2dad4b7d2b · inbound

T-POP: Test-Time Personalization with Online Preference Feedback cites this paper.

T-POP: Test-Time Personalization with Online Preference Feedback Prompt Optimization with Human Feedback

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-04T13:52:11.749286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:52:11.749286Z digest=sha256:c4c7cd941d20e2d4f7811b229e3bb5c0e037c24c49c306e0f81ef5c38218e2cf

Observation 97cf0a86-6e3f-4c9c-b6d1-f8a1d1443454 · inbound

Memory in the Age of AI Agents cites this paper.

Memory in the Age of AI Agents Prompt Optimization with Human Feedback

Reference 111

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:18:20.579829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T18:18:19.911342Z digest=sha256:b78701c79c8ba7972af99e5368ee93f7204c3151c4d1f38cc88d61f5dcb46178

Observation aad03aa5-4fb9-49d5-be28-facea8046975 · inbound

MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks cites this paper.

MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks Prompt Optimization with Human Feedback

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T19:22:25.382765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T19:22:25.382765Z digest=sha256:a57c77782e6b4c0a0970e33b553c07501adbd92aae8ac876a68a5cd1eabb46cd

Observation ddea6f1e-5af0-4a9c-a382-43158a26c7a9 · inbound

OOPrompt: Reifying Intents into Structured Artifacts for Modular and Iterative Prompting cites this paper.

OOPrompt: Reifying Intents into Structured Artifacts for Modular and Iterative Prompting Prompt Optimization with Human Feedback

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:01:05.272416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T02:29:16.582341Z digest=sha256:8f50487a2851a5e5a91f00080727381ddbbac2e2dbeb41dea348243fbcefdea9

Observation 64f0cfa4-71ea-431f-b402-deea27d8aa0d · inbound

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis cites this paper.

LiteOdyssey: A Lightweight Reasoning AI Agent for Interpretable Rare-Disease Diagnosis Prompt Optimization with Human Feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T13:53:50.979339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:53:50.979339Z digest=sha256:f68852ad8dbf509abe981717f9fa8f03371a5ad73441091abb10992635f0d5ff

Observation 1c40d86e-b88a-4742-8ad7-d83ceafc67c9 · inbound

Towards Spec Learning: Inference-Time Alignment from Preference Pairs cites this paper.

Towards Spec Learning: Inference-Time Alignment from Preference Pairs Prompt Optimization with Human Feedback

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-07-04T11:39:47.128804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T07:49:36.816100Z digest=sha256:877b00388f636ac490deb8bd43c0fca36e77f858dc3d8ec304b816f7481c537d

Observation eec38900-c83a-4a05-b344-9b7c2e8dcda4 · inbound

Towards Spec Learning: Inference-Time Alignment from Preference Pairs cites this paper.

Towards Spec Learning: Inference-Time Alignment from Preference Pairs Prompt Optimization with Human Feedback

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:04:39.417519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T10:17:33.176525Z digest=sha256:7656e05a7bbeb50a473936f63c127d2c652047c8503c8b158a4745b038c807cb