Pith. sign in

Paper Citation Record · LEDGER

Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2404.05880.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.05880 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:32:50.304538Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T22:26:18.073916Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2b14b808-5ddc-4f78-bcd0-a5c8461a1f6f · inbound

SEPS: A Separability Measure for Robust Unlearning in LLMs cites this paper.

SEPS: A Separability Measure for Robust Unlearning in LLMs Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:32:50.304538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:32:50.304538Z digest=sha256:7bc3b89a0c7ea5e9670d2a20df8824c839ee6a165efd23ee14acdd4bbd7e519e

Observation 61797ac5-a8bb-4bb9-a17d-ea10b8b0ed05 · inbound

Automating Evaluation of Diffusion Model Unlearning with (Vision-) Language Model World Knowledge cites this paper.

Automating Evaluation of Diffusion Model Unlearning with (Vision-) Language Model World Knowledge Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:38.651564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:38.651564Z digest=sha256:22fa7c8573316a76684825b48aa64f275b529c162549ec3b1283e8b0b302f85b

Observation 452431f8-91b4-4951-8e0e-7add4f59f70e · inbound

A Survey on Generative Model Unlearning: Fundamentals, Taxonomy, Evaluation, and Future Direction cites this paper.

A Survey on Generative Model Unlearning: Fundamentals, Taxonomy, Evaluation, and Future Direction Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 153

Resolution
unresolved
no resolver link, observed 2026-08-06T13:54:40.043778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:54:40.043778Z digest=sha256:057c3d606836a6bdbf960e5d45b060e9733185e64ef8beb90017b5d4b4e6619b

Observation 21975885-5cd2-41c3-b0a7-6523cf50f573 · inbound

SafeLLM: Unlearning Harmful Outputs from Large Language Models against Jailbreak Attacks cites this paper.

SafeLLM: Unlearning Harmful Outputs from Large Language Models against Jailbreak Attacks Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T18:05:32.075408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:05:32.075408Z digest=sha256:66ae34a28db0abb7672556090a70d5fdcd87abd60a7084fdbc93a9a1cce65aed

Observation e75cb372-fbec-4995-bb46-54f0ab7f6f36 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:49.054581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:49.054581Z digest=sha256:3694bd7c82688cdc1179430963ec81a503745652fa4b676cf56034a0f184127d

Observation 8a472273-706c-43bb-a14f-55dae21abc06 · inbound

Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting cites this paper.

Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:51:00.698323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T06:48:49.492536Z digest=sha256:97e74a8f91a281097414957f4f2c23873c7e82ef744f3bbd7ad28f51a6f63ad5

Observation 5460ab10-c69b-4656-8582-cead1e11ca9c · inbound

Exclusive Unlearning cites this paper.

Exclusive Unlearning Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:10:52.142037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:39:54.728106Z digest=sha256:9eea1c4909a7c44e8933d25f7d671a8a6fb7c4913d4fa9556a52123eb4391469

Observation c569a026-1139-43f3-8c96-e80baa2511a7 · inbound

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense cites this paper.

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:35:50.996117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T18:26:19.922383Z digest=sha256:e7b0223780a579e4875a7754df4ab960a1eb4c5cec52f0720c7e6b69ce507bd9

Observation 4ae492ea-516a-471d-b069-ef0fb166f5df · inbound

Jailbreaking Frontier Foundation Models Through Intention Deception cites this paper.

Jailbreaking Frontier Foundation Models Through Intention Deception Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:13.887943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T03:17:51.039062Z digest=sha256:2d88b5e29e8cf2e76c1c4ae6318c2ba6962872fe5f22b3ab4a76b173bbf21a09

Observation ace9d281-c844-48ec-80de-49b4352ca8ff · inbound

Fast Unlearning at Scale via Margin Self-Correction cites this paper.

Fast Unlearning at Scale via Margin Self-Correction Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:18.075413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T15:21:26.391287Z digest=sha256:829787cfe955be3ccfe9ed13613e9c3cb34ad342081b859a94c85981c1d78880