Pith. sign in

Paper Citation Record · LEDGER

Mitigating the Alignment Tax of RLHF

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2309.06256.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2309.06256 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:11:28.144815Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

9
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7f479c68-b81b-4744-8438-a6a7a5cf0883 · inbound

Rethinking Mixture-of-Agents: Is Mixing Different Large Language Models Beneficial? cites this paper.

Rethinking Mixture-of-Agents: Is Mixing Different Large Language Models Beneficial? Mitigating the Alignment Tax of RLHF

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T18:11:28.144815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:11:28.144815Z digest=sha256:c3733a705cb761151974c1e3b415e82ba484afd43a81945179420ed7d6e24491

Observation d1c608ac-143c-4e23-8001-10d4687a0371 · inbound

Compromising Honesty and Harmlessness in Language Models via Deception Attacks cites this paper.

Compromising Honesty and Harmlessness in Language Models via Deception Attacks Mitigating the Alignment Tax of RLHF

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T05:42:43.505543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:42:43.505543Z digest=sha256:592e98dfce599a2f26ce2749683fcf4c50e8cdf94a0123bda995e0495a4a9f90

Observation 02b76a65-7a54-43de-9bbc-5c7ad1d15192 · inbound

ExeSQL: Self-Taught Text-to-SQL Models with Execution-Driven Bootstrapping for SQL Dialects cites this paper.

ExeSQL: Self-Taught Text-to-SQL Models with Execution-Driven Bootstrapping for SQL Dialects Mitigating the Alignment Tax of RLHF

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:56:03.062616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:56:03.062616Z digest=sha256:a71c918010519cf56e67cabb5e5a54f5d2cc09e199f4ccd09e91049280d5441f

Observation 68814b27-f0a7-4b71-bc79-75bc0a926d01 · inbound

Understanding Overadaptation in Supervised Fine-Tuning: The Role of Ensemble Methods cites this paper.

Understanding Overadaptation in Supervised Fine-Tuning: The Role of Ensemble Methods Mitigating the Alignment Tax of RLHF

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:09.918152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:40:09.918152Z digest=sha256:f4313c0b951e1943f8c0ee91331c1e48f51be813def01d34e18a38c463d16557

Observation 1b1f0d2a-e239-4670-9539-70369bca1236 · inbound

Bradley-Terry and Multi-Objective Reward Modeling Are Complementary cites this paper.

Bradley-Terry and Multi-Objective Reward Modeling Are Complementary Mitigating the Alignment Tax of RLHF

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:31.112415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:51:31.112415Z digest=sha256:a78f4b260d6399a358b3995e5226bd687cc9ca13f7f85ee7d8c56af8742c551d

Observation cf35a2c0-128f-4efe-90fa-6dd0c16c0c5a · inbound

Cycle Context Verification for In-Context Medical Image Segmentation cites this paper.

Cycle Context Verification for In-Context Medical Image Segmentation Mitigating the Alignment Tax of RLHF

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:24:50.745763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:24:50.745763Z digest=sha256:c6d225888e2929e7d2504d5f1ce59dcbcfc6dd8ae47fbc4572c60d9b75061654

Observation 46881b43-642c-4772-b5c1-270725d1f111 · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models Mitigating the Alignment Tax of RLHF

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:14.879078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:14.879078Z digest=sha256:31013d97fdc06ab7afe0b74e0f5621879e8512959e0ae9010d5c9a451aeb58a1

Observation 652a0f17-da08-49ed-a08c-15566b8bdf03 · inbound

Beyond Correctness: Harmonizing Process and Outcome Rewards through RL Training cites this paper.

Beyond Correctness: Harmonizing Process and Outcome Rewards through RL Training Mitigating the Alignment Tax of RLHF

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T22:40:43.282501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T22:38:57.833414Z digest=sha256:4dde2d015851887a84ae8dffd481448497462f65fdcf5ad4aae2aeae808cf9ac

Observation 1f34472f-2b41-4fc4-af15-a05165cda945 · inbound

Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware Pruning cites this paper.

Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware Pruning Mitigating the Alignment Tax of RLHF

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T21:01:50.125128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:01:50.125128Z digest=sha256:8a18f0df7eee7e5954c841990d99a318fce687660783cf3e915bbd8a2cc58e02

Observation 9c6b2640-fb55-435b-b9a5-497f3548e30e · inbound

CapTrack: Multifaceted Evaluation of Forgetting in LLM Post-Training cites this paper.

CapTrack: Multifaceted Evaluation of Forgetting in LLM Post-Training Mitigating the Alignment Tax of RLHF

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:45:26.492450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T06:40:51.046965Z digest=sha256:9470f64cb973ebe9e196aa6d8e0740f3c87093e921457f5241c70fe5d4fb5d84

Observation c9986d8d-5bad-41a8-8adf-dfdb044c48a7 · inbound

Generative AI Technologies, Techniques & Tensions: A Primer cites this paper.

Generative AI Technologies, Techniques & Tensions: A Primer Mitigating the Alignment Tax of RLHF

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:41:01.385099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T05:39:12.226558Z digest=sha256:b23c111acbf72965938de725611a475032da7373d6fe1fb67f37e509580545fc

Observation 787e0af1-5a2b-480c-a4b2-6d089dd5ee3c · inbound

OLLM: Options-based Large Language Models cites this paper.

OLLM: Options-based Large Language Models Mitigating the Alignment Tax of RLHF

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:16:08.545128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T02:04:12.174834Z digest=sha256:fe67981ba863e0ec7d5331647ddf9c41a24db4f362139d3dc3654d19dd7bf1c8

Observation 24c97580-deb9-4128-80cc-55e119410d26 · inbound

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion cites this paper.

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion Mitigating the Alignment Tax of RLHF

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:57:06.339401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:03:10.263663Z digest=sha256:af37b2383cd3c139259846bbdb44ac41df9a1a5e9d2cd435e8ad1db23a8cee20

Observation f868b637-308a-4d11-9a44-e1df3252cccb · inbound

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion cites this paper.

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion Mitigating the Alignment Tax of RLHF

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:12:58.908959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T21:12:06.989077Z digest=sha256:5f2cb31c1e96b6c99bcfba0395e5e0f6d096b391fb34f3473d6eb94f5bb84e5a

Observation 8718851e-af55-470f-9c06-cccf72d2c979 · inbound

Learning, Fast and Slow: Towards LLMs That Adapt Continually cites this paper.

Learning, Fast and Slow: Towards LLMs That Adapt Continually Mitigating the Alignment Tax of RLHF

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:07:18.462997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T05:00:31.452781Z digest=sha256:95d260584f932d574b598540a31df9000c2eb4ad375ef47498e13c4e72b4045b

Observation 4ac69b6d-a040-4ae5-bfad-47abb874b41b · inbound

Learning, Fast and Slow: Towards LLMs That Adapt Continually cites this paper.

Learning, Fast and Slow: Towards LLMs That Adapt Continually Mitigating the Alignment Tax of RLHF

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:19:45.701852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T05:19:05.368681Z digest=sha256:820dda5af696632dca7036cf2d9442805e26e8d27d36a9cd57a97eeabd4ba723

Observation c8d86648-e791-47ff-a418-5c284e9d5afa · inbound

Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training cites this paper.

Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training Mitigating the Alignment Tax of RLHF

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:25:33.131993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T08:21:41.008505Z digest=sha256:591c450623f9d9ca11fac907aad6dfdd9603aeb9ae4203199a2d39bcb4d06b1e

Observation 8a5c7bdf-874e-4fb0-929b-75d6f025ea84 · inbound

Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training cites this paper.

Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training Mitigating the Alignment Tax of RLHF

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T19:20:54.974570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T19:20:54.974570Z digest=sha256:454b6b7635ac28db8b0b7714ac1f25c0f2c2a924c1fd18767bfa79e6dd7f2afc

Observation 140ebab7-adb2-4c5f-857d-4c936101c097 · inbound

ARMOR: Adaptive Retriever Optimization for Low-Resource Telecom Question Answering cites this paper.

ARMOR: Adaptive Retriever Optimization for Low-Resource Telecom Question Answering Mitigating the Alignment Tax of RLHF

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-30T04:54:16.249019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T04:53:59.203935Z digest=sha256:eb5969d6c6899b39a8972af92f56c0b598d526bc286c788a19e770bb3c98327d

Observation 71b40abc-9420-4341-99ca-719257dc5299 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Mitigating the Alignment Tax of RLHF

Reference 249

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:95462973407fb9e507cb8609e326c7effcaee71170deb868bd3d1c3d7cd8b028

Observation 737093c4-4e6b-43e1-8156-8c8d0465eb86 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Mitigating the Alignment Tax of RLHF

Reference 250

Resolution
unresolved
no resolver link, observed 2026-08-02T08:41:01.426952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:41:01.426952Z digest=sha256:ecbeec62fe5acbec90267196f7af7f8055f7222b9ccd759823120f39f1328288

Observation 66d3b4ac-d4b1-49c5-8d8c-4c78377c6b24 · inbound

SOS-LoRA: Static Orthogonal-Subspace Low-Rank Adaptation with Fixed Multi-Scale Scaling cites this paper.

SOS-LoRA: Static Orthogonal-Subspace Low-Rank Adaptation with Fixed Multi-Scale Scaling Mitigating the Alignment Tax of RLHF

Reference 157

Resolution
unresolved
no resolver link, observed 2026-08-02T09:51:03.639953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T09:51:03.639953Z digest=sha256:2dd1572b073462b602eabad947ca715c9c734b848f26d6ad56b85a714ef772ab

Observation 2ce3cc3a-5cad-4056-8208-3bf5fef146ab · inbound

A Taxonomy of Cognitive Capability Gaps in Generative and Agentic AI cites this paper.

A Taxonomy of Cognitive Capability Gaps in Generative and Agentic AI Mitigating the Alignment Tax of RLHF

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-04T04:53:01.794216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:53:01.794216Z digest=sha256:fdb33514d57853fb46f24858cb7c4370682c4df29efbdf31cc8821594948c349