Pith. sign in

Paper Citation Record · LEDGER

Attribution Patching Outperforms Automated Circuit Discovery

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2310.10348.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.10348 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T20:36:01.338244Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation db57e891-cce4-47b0-a567-c8cfa23c9035 · inbound

EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification cites this paper.

EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification Attribution Patching Outperforms Automated Circuit Discovery

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T20:36:01.338244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T20:36:01.338244Z digest=sha256:fc55517c75087dce80fe8d0d9826280a19888be55d849f84b71b4a8de67ed7a9

Observation 7a7a161d-b607-4ec2-b266-dd6549f76f18 · inbound

Mechanistic Unveiling of Transformer Circuits: Self-Influence as a Key to Model Reasoning cites this paper.

Mechanistic Unveiling of Transformer Circuits: Self-Influence as a Key to Model Reasoning Attribution Patching Outperforms Automated Circuit Discovery

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T22:56:40.292900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T22:56:40.292900Z digest=sha256:34bf4bd9078e1df26e27436e26fd6447aca10592266451b11de33f8bbbe0f5bf

Observation c03e999d-1783-4990-9473-2ab972e44569 · inbound

Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis cites this paper.

Pierce the Mists, Greet the Sky: Decipher Knowledge Overshadowing via Knowledge Circuit Analysis Attribution Patching Outperforms Automated Circuit Discovery

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:39:19.369293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:39:19.369293Z digest=sha256:40d4b6a9e45f01355ebe7d1bdfe6dbc5126877af8206e49e57a722957db66105

Observation ef6e9eec-cce5-4f92-9452-88191b5a20b1 · inbound

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective cites this paper.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Attribution Patching Outperforms Automated Circuit Discovery

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.964417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.964417Z digest=sha256:b2fcd4534b87768b2e7a3695b52616dc57f7100cf237b76f1f86229f8c955bbd

Observation 0707961c-acac-46e0-955d-bba952c91d39 · inbound

Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMs cites this paper.

Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMs Attribution Patching Outperforms Automated Circuit Discovery

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:53:04.164862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T08:52:45.818050Z digest=sha256:cfed17a39faf7f4a89fb18723aeb9fa68fce94f1d76f28c99f15294057c898cd

Observation b23b4ea6-c812-4b64-8cbd-3fec2f95ea8b · inbound

Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers cites this paper.

Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers Attribution Patching Outperforms Automated Circuit Discovery

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T18:00:51.713937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:00:51.713937Z digest=sha256:a3d188f9ab3775872a13e5da17070a9808431cc8d091de8b05d8852d94b99333

Observation a2f6727a-56a5-41a2-a1e0-0e65ba8efac6 · inbound

BlueGlass: A Framework for Composite AI Safety cites this paper.

BlueGlass: A Framework for Composite AI Safety Attribution Patching Outperforms Automated Circuit Discovery

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T17:46:21.376357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:46:21.376357Z digest=sha256:bf9bc6c1e98ac81019b454af1ff28790694d50485e571ec114c8b387573ac54d

Observation d7bab763-8f46-4d6c-845b-b05db4b9de79 · inbound

Mechanistic Interpretability as Statistical Estimation: A Variance Analysis cites this paper.

Mechanistic Interpretability as Statistical Estimation: A Variance Analysis Attribution Patching Outperforms Automated Circuit Discovery

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T13:23:59.112770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:23:59.112770Z digest=sha256:2bf88e2dae250f2f63a19839c7ff7ecb5ca96069f230945d07f458463b1ccc24

Observation cbec5caf-fb77-4414-b7bb-394327d94906 · inbound

Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings cites this paper.

Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings Attribution Patching Outperforms Automated Circuit Discovery

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:30:54.001813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:28:49.855280Z digest=sha256:d2031b2805ca679877d24721e70b60975e5ca6831ed63f4429a26e437840141e

Observation fdb650a9-56fc-4a02-ab8d-7e6f1438971e · inbound

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs cites this paper.

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs Attribution Patching Outperforms Automated Circuit Discovery

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:35:57.220849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:04:05.157103Z digest=sha256:efcd2627309f46a094016a7e8c20442b14f92636f06f3be151e2bb7f525ecb42

Observation 58f0e9e0-c5ed-4183-962d-bdbb17595cea · inbound

Pruning Unsafe Tickets: A Resource-Efficient Framework for Safer and More Robust LLMs cites this paper.

Pruning Unsafe Tickets: A Resource-Efficient Framework for Safer and More Robust LLMs Attribution Patching Outperforms Automated Circuit Discovery

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:13:29.808384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T09:10:55.001543Z digest=sha256:68e19db83f2f61130fbed9e2fbb37676dfe7282f88156fdaf728c4e0aa01412d

Observation 9090cf06-dafb-4255-92d9-74b1db3b499a · inbound

Instructions Shape Production of Language, not Processing cites this paper.

Instructions Shape Production of Language, not Processing Attribution Patching Outperforms Automated Circuit Discovery

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:12:09.336235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T03:09:02.902912Z digest=sha256:4f4b5ad2e3574664e3a28f7f3b973e5b813bad744c0bd46afe35df1d952abddc

Observation c7a0050e-1f41-45fc-9faf-766816ca56f5 · inbound

Instructions Shape Production of Language, not Processing cites this paper.

Instructions Shape Production of Language, not Processing Attribution Patching Outperforms Automated Circuit Discovery

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:02:58.115444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T21:02:02.135970Z digest=sha256:0af505f4d4867ed6c4dff0af311bc054664a6014174f5578cba382f9daeb96b5

Observation 5052f37d-e6f0-42e5-b17c-f71fa2b6d4df · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Attribution Patching Outperforms Automated Circuit Discovery

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:59:28.531952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T20:53:40.666929Z digest=sha256:ed4f5e25fd82b52e8f2dac06a297aff8c0e0ea0456cce12c6d4a4b328b63fc9a

Observation e403ca6d-c306-4b3f-92e3-aa5cbacb2793 · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Attribution Patching Outperforms Automated Circuit Discovery

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:59:45.138499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T04:59:11.877068Z digest=sha256:ad069f772206b57524bff492a03f0f40e2ba757051c90ab4f609a78546560845

Observation c0b4812f-8e4e-48d7-b8c0-6d68d4cc3b68 · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Attribution Patching Outperforms Automated Circuit Discovery

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:17:54.243464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:e30962549de719655bbe602e43a656b83cd6604d7eb37d24c62d44e19643547d

Observation 54ad83a9-3e6e-4ef2-be3c-1c125c6c37ca · inbound

Interaction Locality in Hierarchical Recursive Reasoning cites this paper.

Interaction Locality in Hierarchical Recursive Reasoning Attribution Patching Outperforms Automated Circuit Discovery

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:04:37.235516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:04:14.532933Z digest=sha256:667ec817809f563747a4be3d9339a6ecd0e682806bd68d0127062db0230c447a

Observation 3e0693ac-8eec-44c7-8b67-3ad4bf6fc732 · inbound

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet cites this paper.

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet Attribution Patching Outperforms Automated Circuit Discovery

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.506256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T07:50:04.813379Z digest=sha256:ce8c565fde230cf352fe6cd70093b71943051c3b24a1a0a7049a338b5e05c720

Observation 96b613b4-0e34-475c-9843-91f3fb3a9664 · inbound

Relational Rank Geometry in Transformers: Detecting and Steering Hidden-State Relation Frames cites this paper.

Relational Rank Geometry in Transformers: Detecting and Steering Hidden-State Relation Frames Attribution Patching Outperforms Automated Circuit Discovery

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:33:15.846677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T08:24:20.921284Z digest=sha256:484cd1bbdd9d020392e3251692931986e935899a0aef806fa582a83f6f85f998

Observation 3b77b5f2-0b89-4e62-897d-ebd818d3b7b6 · inbound

Quantifying the Agreement Between Data-Influence and Data-Similarity to Understand LLM Behavior cites this paper.

Quantifying the Agreement Between Data-Influence and Data-Similarity to Understand LLM Behavior Attribution Patching Outperforms Automated Circuit Discovery

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-26T08:49:14.991193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T08:45:34.884703Z digest=sha256:f49d9ca4094561e0f2c7af0669518fa0f745e8934126e3358b5a19646c4d9505

Observation 49997210-fd24-41d6-90c1-2982451447dd · inbound

Do Models Read What They Write? Causal Registers in Scratchpad Reasoning cites this paper.

Do Models Read What They Write? Causal Registers in Scratchpad Reasoning Attribution Patching Outperforms Automated Circuit Discovery

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:34:21.848322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T07:26:25.145919Z digest=sha256:915f195982823fd596b9150cf4dc7827f5c24a75c179490ce01b05d317d4ac15

Observation 3c9bb65e-df3f-464a-a5e6-93079a38dd38 · inbound

Validating Causal Abstraction Metrics on Simulated Complex Systems cites this paper.

Validating Causal Abstraction Metrics on Simulated Complex Systems Attribution Patching Outperforms Automated Circuit Discovery

Reference 165

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:27:18.567557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-02T19:24:02.616061Z digest=sha256:4b2f2d21301d727d6cacf559641e4979cd972115720dca0d5d468b63aa83758d

Observation 168627d5-b23e-4591-ad73-05745b4a230b · inbound

Faithfulness to Refusal: A Causal Audit of Neuron Selectors cites this paper.

Faithfulness to Refusal: A Causal Audit of Neuron Selectors Attribution Patching Outperforms Automated Circuit Discovery

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-07T15:43:53.897504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-07T15:35:54.665268Z digest=sha256:64f98d02aa2bb1c638bca7ddd49546ecb5a94a40aefd95921a90178439573a2f

Observation cedd31ae-05f5-426d-a39f-c557546e6972 · inbound

Verbalizable Representations Form a Global Workspace in Language Models cites this paper.

Verbalizable Representations Form a Global Workspace in Language Models Attribution Patching Outperforms Automated Circuit Discovery

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-01T23:15:30.815504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:15:30.815504Z digest=sha256:110cc579be7924328b9cb3121f65111727c46ff6fd4e0e880a9b633ffae218c3

Observation 32b301d7-4fd5-42de-af48-0da8275fbad9 · inbound

LAWFUL: Law-Aligned Witness for Faithful Use of Latents cites this paper.

LAWFUL: Law-Aligned Witness for Faithful Use of Latents Attribution Patching Outperforms Automated Circuit Discovery

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T00:47:36.133553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:47:36.133553Z digest=sha256:58bf4f6aeda24a395c5974cb8c5371c2ad70677713e6c4b6d6d7b9200423617b