Pith. sign in

Paper Citation Record · LEDGER

On the Role of Attention Heads in Large Language Model Safety

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2410.13708.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.13708 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T18:12:33.522135Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T17:35:51.334325Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9cf928fe-45cb-49bf-8020-1197d95d1f33 · inbound

Reinforced Lifelong Editing for Language Models cites this paper.

Reinforced Lifelong Editing for Language Models On the Role of Attention Heads in Large Language Model Safety

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T18:12:33.522135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T18:12:33.522135Z digest=sha256:bbd74ca0d68b4b5ff4838288c5bb8a47714f0d5d78d9eff8686e412554271481

Observation 36b35852-d14a-4f02-82c6-ac02953c7d69 · inbound

The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions cites this paper.

The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions On the Role of Attention Heads in Large Language Model Safety

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T23:08:02.421540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:08:02.421540Z digest=sha256:428e728143c9f894f43bb0ae083204fb2475ad207f10d2c4bec59e379b66ee80

Observation 4a4f6db7-a64c-4f3a-ab37-4ce4e0fb6e08 · inbound

ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models cites this paper.

ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models On the Role of Attention Heads in Large Language Model Safety

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:33.227557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:31:33.227557Z digest=sha256:432ec921c3aeebbe1d25a2eb272ca73bf7fe3c993c2285ee731d1fe8cd1e5c5c

Observation ed3a0af8-9b98-495b-a1ef-b264678b1046 · inbound

To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems cites this paper.

To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems On the Role of Attention Heads in Large Language Model Safety

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:32:17.349169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T11:30:47.877793Z digest=sha256:0f49479cbebb70317d1d87dc03bbb5c1c8a9f2f3dd6cd8aa642f46982c6b2ca8

Observation 3ab5dd33-7b72-478c-8828-a72b8a526888 · inbound

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents cites this paper.

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents On the Role of Attention Heads in Large Language Model Safety

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:11:18.246898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:11:18.246898Z digest=sha256:098c8c115c3928a902e635a1845b405c071d6ddb9681a6ee863fa1760a728cc3

Observation 45bcdb12-ec21-4f9e-93f2-3216288e7406 · inbound

Boosting Parameter Efficiency in LLM-Based Recommendation through Sophisticated Pruning cites this paper.

Boosting Parameter Efficiency in LLM-Based Recommendation through Sophisticated Pruning On the Role of Attention Heads in Large Language Model Safety

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:54:18.141878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:54:18.141878Z digest=sha256:1db019545db406836a9655570bb78f23fe51b0bcb88e2b45aa5373a98b191830

Observation 838f11fa-4a06-497f-abe6-414a313cc94e · inbound

Soft Head Selection for Injecting ICL-Derived Task Embeddings cites this paper.

Soft Head Selection for Injecting ICL-Derived Task Embeddings On the Role of Attention Heads in Large Language Model Safety

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:41:59.765500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T02:41:52.906516Z digest=sha256:7f46053253588285778c6e22d41835e0bc39678466e034a70833c1daa05b1411

Observation 3f6b9687-55ac-49c7-8e14-5ecae468d82c · inbound

Correcting Prompt Dependence in LLM Benchmarks: A Bayesian Hierarchical Model with Embedding-Space Clustering cites this paper.

Correcting Prompt Dependence in LLM Benchmarks: A Bayesian Hierarchical Model with Embedding-Space Clustering On the Role of Attention Heads in Large Language Model Safety

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-04T11:19:51.206012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:19:51.206012Z digest=sha256:c7486efd80aae1c340096ba6b2a0f1f6b887daf9656fca221140f0c62b6b7064

Observation 84a6d79a-0654-4057-9f86-0d522e155444 · inbound

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems cites this paper.

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems On the Role of Attention Heads in Large Language Model Safety

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:16:03.952450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:07:31.602378Z digest=sha256:f931545fae2fcff7f8873bd1ace84a724cb72de97b5b53ee4bef70333b6f5440

Observation db8833b2-98aa-4b3b-a9f4-1031a28c3e12 · inbound

Why Do Large Language Models Generate Harmful Content? cites this paper.

Why Do Large Language Models Generate Harmful Content? On the Role of Attention Heads in Large Language Model Safety

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:21:04.490884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:31:13.545599Z digest=sha256:241a0284afa6e8ef608806dc457f7cedc241cb28c55f8a9b7cb73ac28ea2c5e0

Observation e78737db-90cd-47af-9522-f14ca3af9047 · inbound

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs cites this paper.

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs On the Role of Attention Heads in Large Language Model Safety

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T10:01:27.533964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T08:34:14.310656Z digest=sha256:537327111f8df57f83588fce57304a6c5bcb8492c28adf906d72132219afa433

Observation cc737df7-85e7-4416-ab3a-9c04ae0c2912 · inbound

Large Vision-Language Models Get Lost in Attention cites this paper.

Large Vision-Language Models Get Lost in Attention On the Role of Attention Heads in Large Language Model Safety

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:26:10.092015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T11:54:01.224588Z digest=sha256:562caa06b9fb40e45c2e3148bec6548af8638518008841697cc0a1f5375b1d07

Observation 30323868-1d1e-4566-898e-a2ae814bef17 · inbound

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models cites this paper.

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models On the Role of Attention Heads in Large Language Model Safety

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.457378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T13:25:40.986336Z digest=sha256:e4128bb357b4e2624ac522d63d74efb3bfe8618686e9db08c99b00006c36abc8

Observation 67aa9938-d5ba-4348-b004-945de94a1d1a · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models On the Role of Attention Heads in Large Language Model Safety

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:35:51.335775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:f9086f8e30c5eca5d0f329f74d75047b7c8a49b391692357706695269c0cbbcb

Observation 4edb1c6a-47b3-479b-8f9d-a0718c5ca62c · inbound

How Do LLMs Read Bug Reports? An Empirical Study of Attention in LLMs for Automated Program Repair cites this paper.

How Do LLMs Read Bug Reports? An Empirical Study of Attention in LLMs for Automated Program Repair On the Role of Attention Heads in Large Language Model Safety

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T01:17:54.450344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:17:54.450344Z digest=sha256:5bca5e6c12c09ca3307acd94441b69e76f0e28de05ac490c7276d2a418f45663