Pith. sign in

Paper Citation Record · LEDGER

Copy Suppression: Comprehensively Understanding an Attention Head

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2310.04625.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.04625 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:01:28.909500Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 034cb838-4cd1-480b-8e20-f2f30b88feeb · inbound

How to use and interpret activation patching cites this paper.

How to use and interpret activation patching Copy Suppression: Comprehensively Understanding an Attention Head

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:34:06.982041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T20:34:06.888683Z digest=sha256:f75d0f57fd5a5fd7c39f338556a6df5a86f8d4eedfb5bc9de78bd55390446ee7

Observation 1192216f-d4fc-4221-9c18-8203ebf96a88 · inbound

On Mechanistic Circuits for Extractive Question-Answering cites this paper.

On Mechanistic Circuits for Extractive Question-Answering Copy Suppression: Comprehensively Understanding an Attention Head

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T11:01:28.909500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:01:28.909500Z digest=sha256:e5efe9898565a96c5905253caa6d720b6b2bb756e0d429745e1a7061369c46af

Observation d07e6a75-ee99-41cb-86a3-0bb1d54231b9 · inbound

Transformers Don't Need LayerNorm at Inference Time: Scaling LayerNorm Removal to GPT-2 XL and the Implications for Mechanistic Interpretability cites this paper.

Transformers Don't Need LayerNorm at Inference Time: Scaling LayerNorm Removal to GPT-2 XL and the Implications for Mechanistic Interpretability Copy Suppression: Comprehensively Understanding an Attention Head

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:16.921232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:16.921232Z digest=sha256:44dee9d5dcee8738d447cd5f93f2c159100538dcff907f84b855ceb29c5300c5

Observation cca18d83-2420-492b-9947-913de669da1c · inbound

From Indirect Object Identification to Syllogisms: Exploring Binary Mechanisms in Transformer Circuits cites this paper.

From Indirect Object Identification to Syllogisms: Exploring Binary Mechanisms in Transformer Circuits Copy Suppression: Comprehensively Understanding an Attention Head

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:37:54.516239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:37:54.516239Z digest=sha256:87f66c51c54ed3361cde3c76e4731186e735e8158f8e28fd82341ba9aedd2602

Observation ff63a38a-b73b-4c95-a0fd-7d6a19f6b6b9 · inbound

Is One Layer Enough? Understanding Inference Dynamics in Tabular Foundation Models cites this paper.

Is One Layer Enough? Understanding Inference Dynamics in Tabular Foundation Models Copy Suppression: Comprehensively Understanding an Attention Head

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:09.871824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T12:41:22.959582Z digest=sha256:ec609ea13d39117f35e75d0cfa5b0cbf00e3ceee707e16ecd14f37eabc9ce123

Observation 43567132-71a8-45a6-ae20-2b72bd158fc0 · inbound

How LLMs Are Persuaded: A Few Attention Heads, Rerouted cites this paper.

How LLMs Are Persuaded: A Few Attention Heads, Rerouted Copy Suppression: Comprehensively Understanding an Attention Head

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:26:24.330150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T04:18:09.600353Z digest=sha256:8b66a5f82f98e3839582acd4169ff042cbda49e65e836b1b6ae44319c51cdb58

Observation a4bcb931-2e1c-4b54-a6b1-99a9121a3c2b · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Copy Suppression: Comprehensively Understanding an Attention Head

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:55.450507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:d117fedadc915e2e39f86b8c4ed7fb618a5ad745d5ede4cf696cbc97f332b28e

Observation 505525a7-bf34-4450-9a7d-6b7546a434e8 · inbound

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models cites this paper.

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models Copy Suppression: Comprehensively Understanding an Attention Head

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:30:25.657723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-25T06:27:42.024094Z digest=sha256:aa93c9e0cff544b20853612d0c0db6be0ae39f3a662c38b0a683f50b68df02f9

Observation 3e16cc81-1630-4491-8087-ff6dc4a56194 · inbound

Necessary, Decodable and Reversible, Yet Not Transferable: A Stress Test for Attention-Head Role Claims cites this paper.

Necessary, Decodable and Reversible, Yet Not Transferable: A Stress Test for Attention-Head Role Claims Copy Suppression: Comprehensively Understanding an Attention Head

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:47:27.720467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T19:33:19.875024Z digest=sha256:80d7ece3352e566fd35fc7253dccd9174acdf3e94ec2988b41f92d26ab9e6dfa

Observation ad7489bc-9000-4b55-8781-38f5fb6d001e · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models Copy Suppression: Comprehensively Understanding an Attention Head

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:35:51.342261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:d7343888d14129332a6d83c4ef7aa0c988b43de4e8a84b1417d0d4274c35bb19

Observation dd125648-1291-4896-b8a0-bec78cb275bc · inbound

Kernelized Linear Attention: Breaking the Capacity Wall with Symmetric Cones cites this paper.

Kernelized Linear Attention: Breaking the Capacity Wall with Symmetric Cones Copy Suppression: Comprehensively Understanding an Attention Head

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T18:06:54.358789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:06:54.358789Z digest=sha256:c6b0ebef4eba0c6dd8deadfc320233014ee8e628a20d853b20b0683960ab3d14