Pith. sign in

Paper Citation Record · LEDGER

Solving the Inverse Alignment Problem for Efficient RLHF

As of 20 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2412.10529.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.10529 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:55:14.514424Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 07c87c11-fa96-48a2-9e95-02d4caa3a219 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Solving the Inverse Alignment Problem for Efficient RLHF Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.399260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.399260Z digest=sha256:bd3bc6be6067a64b18f180eab0a5531104bd1f1db081d1e67891e843fbab9b60

Observation c499bb23-9f36-4e99-8080-eacd0bf158e8 · outbound

This paper cites an unresolved cited work.

Solving the Inverse Alignment Problem for Efficient RLHF Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.407227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.407227Z digest=sha256:4726f6be75e0d2c89e6c1de83baa9af90cc293e860ce36911d12af7c813e2545

Observation 718eef58-0d97-4394-8095-5bd116d189ae · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Solving the Inverse Alignment Problem for Efficient RLHF BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.413400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.413400Z digest=sha256:c9ed1f7ac82a95b72a02d8ad058f69ff6aeff6db441c00753a5484822c10f73b

Observation 135f7846-d08d-45df-a25d-07eb56de92bc · outbound

This paper cites On the Sensitivity of Reward Inference to Misspecified Human Models.

Solving the Inverse Alignment Problem for Efficient RLHF On the Sensitivity of Reward Inference to Misspecified Human Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.418381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.418381Z digest=sha256:340cf7eada76d5a6494f3565600bdffaab2067b62bf5a82d9eafd81a08e22922

Observation 418aed5e-c154-45f6-9a0f-ef2cc201f03c · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Solving the Inverse Alignment Problem for Efficient RLHF Adam: A Method for Stochastic Optimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.424376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.424376Z digest=sha256:0300cc2d16eb7589601312c45094ea002bc7ba714f427bbb33f1b68ae48c7dfc

Observation f398cc95-f2a7-411e-910d-76a946347aab · outbound

This paper cites Models of human preference for learning reward functions.

Solving the Inverse Alignment Problem for Efficient RLHF Models of human preference for learning reward functions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.429741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.429741Z digest=sha256:e07a134e7600b904601d7eee812f65e4a7712eed0d74501a14bf4c49e6db6fa8

Observation 369f9ae1-2567-4743-af88-144ccb68a72f · outbound

This paper cites an unresolved cited work.

Solving the Inverse Alignment Problem for Efficient RLHF Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:55:15.200644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-11T15:55:14.436590Z digest=sha256:c43c441ddae969decfb03904e1aad65e79ff101aca3b9a9ce464dbf119b21cc8

Observation 715fa89e-d95b-4611-b27f-8216179a7733 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Solving the Inverse Alignment Problem for Efficient RLHF RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.441779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.441779Z digest=sha256:5709f62dccbf7d2aba88b738914e842fbc68b43fd756d068d3a16509f615294e

Observation cc78cb53-29be-42d4-83af-d1313c893a74 · outbound

This paper cites Training language models to follow instructions with human feedback.

Solving the Inverse Alignment Problem for Efficient RLHF Training language models to follow instructions with human feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.446632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.446632Z digest=sha256:40b083c186dc63f02e659e2fe7ea6b33c791b157f6d2c2fd4930778d78c268d8

Observation a8f90746-c415-4a97-97d1-077ba8ce998a · outbound

This paper cites an unresolved cited work.

Solving the Inverse Alignment Problem for Efficient RLHF Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.452073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.452073Z digest=sha256:b5a57f1405150f3a034ba55e88d03be5637a4dcc9a3473a9bd6446f6d98a9ad7

Observation 8840e782-3196-432a-9ed7-dc356ce3859e · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

Solving the Inverse Alignment Problem for Efficient RLHF Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.456371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.456371Z digest=sha256:f4007f25d875e572c0be2b96c28f147d42ba3bc103e99375e036240af381a88d

Observation 3afdc876-f7ec-42e6-af5d-90c4ee1588a9 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Solving the Inverse Alignment Problem for Efficient RLHF Proximal Policy Optimization Algorithms

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.461402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.461402Z digest=sha256:1e590bc2b9749271e5f8e7db15cd9f9175b20d529807beebc8084f49aa62a7d1

Observation ef2a9466-bebd-4926-8336-bf23dd87f1fc · outbound

This paper cites SALMON: Self-Alignment with Instructable Reward Models.

Solving the Inverse Alignment Problem for Efficient RLHF SALMON: Self-Alignment with Instructable Reward Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.466628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.466628Z digest=sha256:4d5ef6ce723da43a3d52e0076ef0be2b5e2584bf72f3e0d88bbfc21e47b097fd

Observation 699d589d-bd1a-49db-8cbd-9c0051b1a8e5 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Solving the Inverse Alignment Problem for Efficient RLHF Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.474557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.474557Z digest=sha256:8b5dbe14c70b5e620831212e8be370c3c67665bbb954b6f76bfec96739aedfb8

Observation df5f3f17-5997-4c7f-a644-b7df2a879671 · outbound

This paper cites an unresolved cited work.

Solving the Inverse Alignment Problem for Efficient RLHF Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:55:15.165304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-11T15:55:14.484126Z digest=sha256:1c77382ae3b35809978d925418a8d86a1f09e7c671f4c6e3875779fd3349ea03

Observation 4f31d886-6ea6-4106-bbab-67ccbb15feb1 · outbound

This paper cites an unresolved cited work.

Solving the Inverse Alignment Problem for Efficient RLHF Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:55:15.143893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-11T15:55:14.489830Z digest=sha256:1408146f0c2a2238b81b52ccbc936602623e8748a2d5c3d12fa645dd7bbff4db

Observation 96f0f89b-53c5-4a16-919a-2d875d3b71d2 · outbound

This paper cites Bayesian Reward Models for LLM Alignment.

Solving the Inverse Alignment Problem for Efficient RLHF Bayesian Reward Models for LLM Alignment

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.495738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.495738Z digest=sha256:d05442f2b139b4c37c79796b8730e218566ee41d3a781681b15fd695f22d2c12

Observation 4ae95ce3-8b36-46ca-9134-70117210163d · outbound

This paper cites online" 'onlinestring :=.

Solving the Inverse Alignment Problem for Efficient RLHF online" 'onlinestring :=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.504316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.504316Z digest=sha256:c35171770006013ad5cf1c656b665e52f58dda37f969be53698927237a1d8864

Observation cec60575-ca6d-4e77-8662-1f307e5bf472 · outbound

This paper cites write newline.

Solving the Inverse Alignment Problem for Efficient RLHF write newline

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T15:55:14.514424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:55:14.514424Z digest=sha256:9a69fbdc25976b9da28e97111c26726b152d1ee64c9e0b6d2da50f60b2d619e2

Pith citing papers

No inbound Pith citation observations are available.