Pith. sign in

Paper Citation Record · LEDGER

BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2408.12798.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.12798 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:50:01.036983Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T03:26:29.013296Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04418d72-5551-4b88-ac7d-44c93d30cfa6 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.075764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:53fe1309b3fe15d04301408b3f9572865ee60d3a6c7c2c89a078c65c0ed291dd

Observation b99551ec-84d4-4d00-b52c-7d7eb0a5d704 · inbound

A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations cites this paper.

A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 235

Resolution
unresolved
no resolver link, observed 2026-08-09T00:50:01.036983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:50:01.036983Z digest=sha256:7b9a6e8c28ed9c1033ea04736047e5da9eddedf2dd5ae3480c67a29e7d80e736

Observation 67e57533-2f2c-4871-9fa0-2ecac2b42e68 · inbound

Exposing the Ghost in the Transformer: Abnormal Detection for Large Language Models via Hidden State Forensics cites this paper.

Exposing the Ghost in the Transformer: Abnormal Detection for Large Language Models via Hidden State Forensics BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:32:12.973209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T22:27:18.533162Z digest=sha256:fc6248ae9016badf959d55d0db7310643f529a6b59cd85e9311023c8e011e6fa

Observation 58b5a81b-e7b2-4d27-920b-f3f9343622ce · inbound

Daunce: Data Attribution through Uncertainty Estimation cites this paper.

Daunce: Data Attribution through Uncertainty Estimation BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:56:17.320422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:56:17.320422Z digest=sha256:30f733bf3dc6abacf2c157ed403c1eafd25a527f438700816ab664424acc4536

Observation e095e507-a3e2-4e29-8570-d1b6a2f9c7c0 · inbound

Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation cites this paper.

Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T05:45:56.195950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:45:56.195950Z digest=sha256:ee69a7e450c5ebe0e7359d72d1d152b4945f11146cb121ecdbfac46b0b90870b

Observation a403a9aa-9b42-41e3-b377-957f5d11445f · inbound

Architectural Backdoors in Deep Learning: A Survey of Vulnerabilities, Detection, and Defense cites this paper.

Architectural Backdoors in Deep Learning: A Survey of Vulnerabilities, Detection, and Defense BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:18.875445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:39:18.875445Z digest=sha256:54be00a65370166a938d0b1de1a3944e19a7b23cc79f815634384497a2dd43a2

Observation 99d0c563-eb4c-4a30-8c2f-656a0d540244 · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 174

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:49.508689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:49.508689Z digest=sha256:b11fc64e549cda9849e500c42b53b251131b0b6de9c27c2b7da43d6bdd228a3b

Observation ac94bc7f-60c4-45e8-9fbf-d93227e6ab76 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 168

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:07.072647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:07.072647Z digest=sha256:6601c42d6ad16913a74b18fe123a3da62a49e77b7386146519391a67dd4b200f

Observation ab287774-33e5-4f7b-b50c-296f93d4ccd6 · inbound

Backdoor Samples Detection Based on Perturbation Discrepancy Consistency in Pre-trained Language Models cites this paper.

Backdoor Samples Detection Based on Perturbation Discrepancy Consistency in Pre-trained Language Models BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T13:44:42.571849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:44:42.571849Z digest=sha256:f31c5351d778471e02c7b5476dfedb6ee3e19adfba4b25493b4c61ba58b361b7

Observation 6a232c11-1f9b-4ec7-adad-b637e28102f7 · inbound

SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems cites this paper.

SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:31:02.038525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:36:19.694339Z digest=sha256:c83c85269be6c2e1855ac700f678059a0064d1269909b2d997df538654e44496

Observation 04eada25-d38e-47ed-a11a-23936c5772c2 · inbound

On the Privacy of LLMs: An Ablation Study cites this paper.

On the Privacy of LLMs: An Ablation Study BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:25:48.856377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T18:25:05.586464Z digest=sha256:192c85f3b1124ddcc4c6c0a96ea660dce82212f3e9ecf275357a53ad14057cea

Observation 10ed5058-f3ef-4819-afde-805cb0407d6e · inbound

Backdooring Masked Diffusion Language Models cites this paper.

Backdooring Masked Diffusion Language Models BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.468459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T07:30:19.657903Z digest=sha256:fa0748a80ebaaf43855c60ebb5aedabb0f5c508e6f45e5c116ae44b6ae43e7a5

Observation 2375832c-2aef-49a3-86b2-95e0e8c2968b · inbound

Backdooring Masked Diffusion Language Models cites this paper.

Backdooring Masked Diffusion Language Models BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.206796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T18:00:05.311834Z digest=sha256:37404704414caddfa7511ce36aedef86e22d2ea5c75f8c6be288c38c36e99053

Observation 0f7590e0-be90-468a-984d-0efc18d2f298 · inbound

RogueMerge: Robust and Unified Attacks against LLM Model Merging cites this paper.

RogueMerge: Robust and Unified Attacks against LLM Model Merging BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:29.014856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T10:05:10.309553Z digest=sha256:406a6d172aae4c828526a698608ad27b7aebd7027ca74188adaebd52cc6f4b91

Observation 716d55bc-0077-4ed5-ac2c-ebbc88aa5579 · inbound

PathMark: Protecting Intellectual Property of Mixture-of-Expert LLMs via Path Watermarks cites this paper.

PathMark: Protecting Intellectual Property of Mixture-of-Expert LLMs via Path Watermarks BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T00:40:49.755070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:40:49.755070Z digest=sha256:ba591b3b82a3caee49f39c862844109395cc102ccecd2b143fae8107798aaaa6

Observation 4946c375-ab19-4105-a6ac-baff40de3650 · inbound

Defense Against LLM Backdoors using Critical Neuron Isolation Pruning cites this paper.

Defense Against LLM Backdoors using Critical Neuron Isolation Pruning BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:29:55.972004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:29:55.972004Z digest=sha256:86d60cd369920e63c93e29817fab8c89a9ad2a6c208c96b6a5253cdf5fae44fc

Observation 01fd38ed-a2d9-481b-a039-93c97a02448a · inbound

When Activation Oracles Learn Not to Read: Concept-Specific Blind Spots in Fine-Tuned Oracles cites this paper.

When Activation Oracles Learn Not to Read: Concept-Specific Blind Spots in Fine-Tuned Oracles BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T23:56:50.971987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T23:56:50.971987Z digest=sha256:0c2af39e3e4c264816e74ad01e8bba836b6bc7bdb7a88e210c8b4f73a0eed497

Observation cbdde26c-c822-4154-8f48-b5dcc0516983 · inbound

Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems cites this paper.

Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T00:51:17.041232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:51:17.041232Z digest=sha256:c5d99a31f20e56c7c2c849b4959cbef6795247c21ba3585d859a41d012db70a5

Observation dec966fe-f6bd-4bc7-b8d4-531ed879c860 · inbound

Hollow-LLM Attack: Computationally Trivial Weights in Zero-Knowledge Verification of LLM Inference cites this paper.

Hollow-LLM Attack: Computationally Trivial Weights in Zero-Knowledge Verification of LLM Inference BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T01:31:21.807777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:31:21.807777Z digest=sha256:e79104b676798842c32ff9362bc5f917671aea94555632820cc8f19ae4b06109