Pith. sign in

Paper Citation Record · LEDGER

ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.11889.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.11889 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T18:53:15.648678Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:29:44.263692Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3e8d1165-3c6e-4798-9d66-753133cc7da7 · inbound

Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM cites this paper.

Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T18:53:15.648678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:53:15.648678Z digest=sha256:c25323ce8798844ef2f0d23d2af987ace6e9a8460ad2d5621d9cd94ab0a7e03b

Observation 123dee81-6968-497b-8966-f3b1c026d33a · inbound

On Almost Surely Safe Alignment of Large Language Models at Inference-Time cites this paper.

On Almost Surely Safe Alignment of Large Language Models at Inference-Time ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T16:18:40.699323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:18:40.699323Z digest=sha256:e5eceabc84a8c278a6727fe8cbe318fa32b7656d797b6f5671a23c315d874838

Observation 1253af54-4495-4576-919c-9a0b5f21ff5b · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 214

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.324707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:b2561db47b96c6f4b0634fd8447e0d6f9114fe93956b864350f9a4b967b530e1

Observation 2ed13e14-0892-4b18-a2a6-46c16f8511df · inbound

Exploring and Mitigating Fawning Hallucinations in Large Language Models cites this paper.

Exploring and Mitigating Fawning Hallucinations in Large Language Models ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T13:11:15.385658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:11:15.385658Z digest=sha256:9c4d5099c7fc3cc3941b7a58e8f6bb999e0eae9e1d153f6394ba026abda91e63

Observation 03628607-c82e-43f1-b156-b6511d09dde8 · inbound

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs cites this paper.

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:29:44.265560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-26T09:52:52.603708Z digest=sha256:823cf61ad9ebf81d40b397305441d452a14087568f38262e727c21179455b76b

Observation 7d9809cf-d8ef-40c9-a3a9-8b10e5f846e8 · inbound

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs cites this paper.

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs ROSE Doesn't Do That: Boosting the Safety of Instruction-Tuned Large Language Models with Reverse Prompt Contrastive Decoding

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:55:35.054465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-07-01T07:07:45.177242Z digest=sha256:35b6815187758bfdfc2e82d05d4755c1309fe576ba26b5524f90e11b7ca7f64c