Pith. sign in

Paper Citation Record · LEDGER

Copy Suppression: Comprehensively Understanding an Attention Head

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2310.04625.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.04625 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:01:28.909500Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 034cb838-4cd1-480b-8e20-f2f30b88feeb · inbound

How to use and interpret activation patching cites this paper.

How to use and interpret activation patching Copy Suppression: Comprehensively Understanding an Attention Head

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:34:06.982041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T20:34:06.888683Z digest=sha256:3c91da63a664e0a1332acb9834fb508d512091206fbd4b90c69baa242bc5e092

Observation 1192216f-d4fc-4221-9c18-8203ebf96a88 · inbound

On Mechanistic Circuits for Extractive Question-Answering cites this paper.

On Mechanistic Circuits for Extractive Question-Answering Copy Suppression: Comprehensively Understanding an Attention Head

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T11:01:28.909500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:01:28.909500Z digest=sha256:57dc788aa2bd4467ac7f571676332de2b81ab4b08b6df546755e89d5c26b121b

Observation d07e6a75-ee99-41cb-86a3-0bb1d54231b9 · inbound

Transformers Don't Need LayerNorm at Inference Time: Scaling LayerNorm Removal to GPT-2 XL and the Implications for Mechanistic Interpretability cites this paper.

Transformers Don't Need LayerNorm at Inference Time: Scaling LayerNorm Removal to GPT-2 XL and the Implications for Mechanistic Interpretability Copy Suppression: Comprehensively Understanding an Attention Head

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:16.921232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:16.921232Z digest=sha256:0109290fb137fc05e71769d5cd1be41d9dd27684de37174e58e447b061f4f1eb

Observation cca18d83-2420-492b-9947-913de669da1c · inbound

From Indirect Object Identification to Syllogisms: Exploring Binary Mechanisms in Transformer Circuits cites this paper.

From Indirect Object Identification to Syllogisms: Exploring Binary Mechanisms in Transformer Circuits Copy Suppression: Comprehensively Understanding an Attention Head

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:37:54.516239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:37:54.516239Z digest=sha256:a1a1b76bad801dcb2d1e48d6f91e164a20e7e59704ec1c8ba3babb7cc89e8293

Observation ff63a38a-b73b-4c95-a0fd-7d6a19f6b6b9 · inbound

Is One Layer Enough? Understanding Inference Dynamics in Tabular Foundation Models cites this paper.

Is One Layer Enough? Understanding Inference Dynamics in Tabular Foundation Models Copy Suppression: Comprehensively Understanding an Attention Head

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:09.871824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T12:41:22.959582Z digest=sha256:1a948cccbf810a1104b3dac7f9250171425e262c461ccf172858ce675d8191cb

Observation 43567132-71a8-45a6-ae20-2b72bd158fc0 · inbound

How LLMs Are Persuaded: A Few Attention Heads, Rerouted cites this paper.

How LLMs Are Persuaded: A Few Attention Heads, Rerouted Copy Suppression: Comprehensively Understanding an Attention Head

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:26:24.330150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T04:18:09.600353Z digest=sha256:63ee0797d5bb6d5bfde9daa2b9b8d47844145a4900aa4ce09899514eda22acdd

Observation a4bcb931-2e1c-4b54-a6b1-99a9121a3c2b · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Copy Suppression: Comprehensively Understanding an Attention Head

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:55.450507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:e0e4fc3361720908bdca5957117dc54763062ededb943c4bc28844e8a4671a23

Observation 505525a7-bf34-4450-9a7d-6b7546a434e8 · inbound

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models cites this paper.

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models Copy Suppression: Comprehensively Understanding an Attention Head

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:30:25.657723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-25T06:27:42.024094Z digest=sha256:3e894c148166db861d09ddaa6aa836d6e20c4836b06235e39f1a5e9e85c26850

Observation 3e16cc81-1630-4491-8087-ff6dc4a56194 · inbound

Necessary, Decodable and Reversible, Yet Not Transferable: A Stress Test for Attention-Head Role Claims cites this paper.

Necessary, Decodable and Reversible, Yet Not Transferable: A Stress Test for Attention-Head Role Claims Copy Suppression: Comprehensively Understanding an Attention Head

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:47:27.720467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T19:33:19.875024Z digest=sha256:f56b13ffea0cdfcbf6c8ea08acbe812273c6cb8806806c5aecd9b43d28ce6934

Observation ad7489bc-9000-4b55-8781-38f5fb6d001e · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models Copy Suppression: Comprehensively Understanding an Attention Head

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:35:51.342261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:fb34ad8950c5ad6304fb8784242fe7b30d5fa7dd43db86cd5aba16879c3cbb2e

Observation dd125648-1291-4896-b8a0-bec78cb275bc · inbound

Kernelized Linear Attention: Breaking the Capacity Wall with Symmetric Cones cites this paper.

Kernelized Linear Attention: Breaking the Capacity Wall with Symmetric Cones Copy Suppression: Comprehensively Understanding an Attention Head

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T18:06:54.358789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:06:54.358789Z digest=sha256:c398d448872b788067cc213a47e968b71f122b65de2ae4cd6f88ed95a0332e11