Pith. sign in

Paper Citation Record · LEDGER

Self-Ablating Transformers: More Interpretability, Less Sparsity

As of 18 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2505.00509.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.00509 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:43:45.651719Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact1
  • verified fuzzy9
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4ea825f7-0835-4a2d-94b6-fbab139c7034 · outbound

This paper cites write newline.

Self-Ablating Transformers: More Interpretability, Less Sparsity write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.506675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.506675Z digest=sha256:a1d80317cd792e7542b0c916cf0414b7ddc51e427389d2f99286205f7bb94742

Observation 6af40797-922c-432f-b2ea-332e20faa2b1 · outbound

This paper cites Language Models Can Explain Neurons in Language Models , May 2023.

Self-Ablating Transformers: More Interpretability, Less Sparsity Language Models Can Explain Neurons in Language Models , May 2023

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:46.093263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.512439Z digest=sha256:e11a7fdaa2d76bb20902debea7a199f79a03ca37f0fa931d47628df005e5ce47

Observation a997d328-5dde-4958-bfa8-8085284a4226 · outbound

This paper cites GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow , March 2021.

Self-Ablating Transformers: More Interpretability, Less Sparsity GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow , March 2021

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.516143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.516143Z digest=sha256:0e2750bb89efa4ec29df2ab36ad993a9b94650f66ac9ef55acd7aae133d67889

Observation 03bbbbab-a3c2-401e-b729-5c48f1369363 · outbound

This paper cites Gradient Routing: Masking Gradients to Localize Computation in Neural Networks.

Self-Ablating Transformers: More Interpretability, Less Sparsity Gradient Routing: Masking Gradients to Localize Computation in Neural Networks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.520989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.520989Z digest=sha256:f7c081d92edac49f0f2ca092b43bb38efb3a9286cf80fbf724ef115f6ec7144d

Observation 49852d33-27fc-47d1-98ce-2b0c8598d3dd · outbound

This paper cites Mavor-Parker, Aengus Lynch, Stefan Heimersheim, and Adri \`a Garriga-Alonso.

Self-Ablating Transformers: More Interpretability, Less Sparsity Mavor-Parker, Aengus Lynch, Stefan Heimersheim, and Adri \`a Garriga-Alonso

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:46.081343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.525545Z digest=sha256:bbc9b0f1578b992d56306463e1646133ea833b10d65155f07ed2316cfbae1b68

Observation ea230a7e-0b3e-4bd8-a0d6-707adb01d225 · outbound

This paper cites TinyStories: How Small Can Language Models Be and Still Speak Coherent English?.

Self-Ablating Transformers: More Interpretability, Less Sparsity TinyStories: How Small Can Language Models Be and Still Speak Coherent English?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.529143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.529143Z digest=sha256:a5dd238f607906f33764fb8095150b1eb6232fdb4751f84571a1edd367a11e7b

Observation 28990087-6b1a-4341-9be6-fa54a9325584 · outbound

This paper cites Neuron to Graph: Interpreting Language Model Neurons at Scale.

Self-Ablating Transformers: More Interpretability, Less Sparsity Neuron to Graph: Interpreting Language Model Neurons at Scale

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.533558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.533558Z digest=sha256:5ce44f5f0374e740e13b0a2bcaaf708f00c4fbc2c7f082994f52add70f150b93

Observation 3655251c-b0ab-4b08-81bb-5cec5437e60f · outbound

This paper cites N2G: A Scalable Approach for Quantifying Interpretable Neuron Representations in Large Language Models.

Self-Ablating Transformers: More Interpretability, Less Sparsity N2G: A Scalable Approach for Quantifying Interpretable Neuron Representations in Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.537514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.537514Z digest=sha256:3af71b820bd3e3be84656aa5cd9f8db7414dda2bca09ff34940144abf998be6d

Observation c9449e0d-058e-4163-821a-0bfa8d659aa5 · outbound

This paper cites Stabilizing the Lottery Ticket Hypothesis.

Self-Ablating Transformers: More Interpretability, Less Sparsity Stabilizing the Lottery Ticket Hypothesis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.541038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.541038Z digest=sha256:87819bc44daf52fb9f5a7e11e84b2531eef8906bc57aa9e02310a3b86351e858

Observation d137550d-efa3-4484-a9b5-14f9bc90c2d4 · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Self-Ablating Transformers: More Interpretability, Less Sparsity The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.544672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.544672Z digest=sha256:3974e8d2a4e4eb39929080119e0164a181e3c2199273b70ef711481b1325b209

Observation 22536831-2b59-4e7e-bb65-266c137f10de · outbound

This paper cites Scaling and evaluating sparse autoencoders.

Self-Ablating Transformers: More Interpretability, Less Sparsity Scaling and evaluating sparse autoencoders

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.548149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.548149Z digest=sha256:e39b2cf75d4eb1e96bc3ada71e79ec86ac2fa779b221e3cab171d620e100fe5f

Observation 6e7901d4-411d-4bd2-86d9-c2dc3eaf38f7 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Self-Ablating Transformers: More Interpretability, Less Sparsity Gemini: A Family of Highly Capable Multimodal Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.551102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.551102Z digest=sha256:f7604f159fbcc3b3026f708a6420e9ad71c34a7aa6fa66e289912bdc0ceb11b5

Observation ed9e84da-f5f8-4d37-9d15-a0b2c8c5f419 · outbound

This paper cites FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI.

Self-Ablating Transformers: More Interpretability, Less Sparsity FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.553635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.553635Z digest=sha256:6b3388449747509044450ea3b0e01686b3c62d2af84d2f0ed127144079e1d954

Observation 2ee220c7-6eb6-4ed7-acbb-c9720dd2504f · outbound

This paper cites Openwebtext corpus.

Self-Ablating Transformers: More Interpretability, Less Sparsity Openwebtext corpus

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.556523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.556523Z digest=sha256:5d16f1e8c24255f80357466f089dbe5d4b72ec17763e02df8934c19edcd5851b

Observation 8cad3ac6-af2f-494a-8f5f-5cea8ad1eb35 · outbound

This paper cites Sparse autoencoders find highly interpretable features in language models.

Self-Ablating Transformers: More Interpretability, Less Sparsity Sparse autoencoders find highly interpretable features in language models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.559097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.559097Z digest=sha256:f8b809e011a1bb3b6a331c650235198111456070cf28ade62bdb4c4d86f9ce9d

Observation c4327362-cad3-4d04-9892-6e50a7664e36 · outbound

This paper cites Two sparsities are better than one: unlocking the performance benefits of sparse–sparse networks.

Self-Ablating Transformers: More Interpretability, Less Sparsity Two sparsities are better than one: unlocking the performance benefits of sparse–sparse networks

Reference 16

Resolution
verified exact
doi, observed 2026-08-16T04:43:45.694494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.562081Z digest=sha256:f3b4889232dbf309ee0a42a1bc8554fad8f72335445b95566643bd2b440d489b

Observation 269f3908-1896-4f84-9965-384c58b3db09 · outbound

This paper cites an unresolved cited work.

Self-Ablating Transformers: More Interpretability, Less Sparsity Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.565343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.565343Z digest=sha256:968ebc52e01f05437011dc8d339abe2ddcc7a6d54fbad88d68c8fdf661ffdd65

Observation ba88fd9f-15d1-471c-b721-c4f32dd2e5d3 · outbound

This paper cites Lazzaro, S.

Self-Ablating Transformers: More Interpretability, Less Sparsity Lazzaro, S

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:46.052630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.568550Z digest=sha256:ee5c5ac5c5e80efd31a52634cedde5797629e2261beb116ff17dbefaf2ad490b

Observation 6085f51d-8a80-4faa-9d84-051192654ce1 · outbound

This paper cites Lecun, L.

Self-Ablating Transformers: More Interpretability, Less Sparsity Lecun, L

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.571944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.571944Z digest=sha256:0f782194ca05fe720634e0fb47b70c5d3152fb47c0c0f086ce1582e4a4e2a141

Observation 1c4f1516-fa8e-4ad7-a72f-69555bf19ab9 · outbound

This paper cites Deep learning.

Self-Ablating Transformers: More Interpretability, Less Sparsity Deep learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.575516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.575516Z digest=sha256:de7e5c233d6d90f2c7348ea8ecb6f8d47542d96f87ec84d6d2d3f9dade69df4c

Observation 2cc3b821-cda7-435b-b31a-ca6d523c561c · outbound

This paper cites The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning.

Self-Ablating Transformers: More Interpretability, Less Sparsity The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.579184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.579184Z digest=sha256:4d71d716e260ea9e6b34c6e6a09f50795aaf44f68d7219d7959dff0cd516468d

Observation 6dd5c797-503a-474a-8f15-643e781b03d7 · outbound

This paper cites KAN: Kolmogorov-Arnold Networks.

Self-Ablating Transformers: More Interpretability, Less Sparsity KAN: Kolmogorov-Arnold Networks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.583081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.583081Z digest=sha256:800940b52d359e563411344a3d45d0aba5087133fa11a1a39536ff39aba30456

Observation 82319611-8522-4537-8833-a45a97063fe5 · outbound

This paper cites Majani, Ruth Erlanson, and Yaser Abu-Mostafa.

Self-Ablating Transformers: More Interpretability, Less Sparsity Majani, Ruth Erlanson, and Yaser Abu-Mostafa

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:46.040037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.586889Z digest=sha256:b15be0bb984cc40802c3736f955e73a9cafe493231e568aea6a2578307e4777e

Observation b8c3debc-0212-48a1-b138-862bd6c09a29 · outbound

This paper cites Transformer circuit evaluation metrics are not robust.

Self-Ablating Transformers: More Interpretability, Less Sparsity Transformer circuit evaluation metrics are not robust

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:46.028393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.590675Z digest=sha256:d8251c75485f9abca9051beb54c41f246df3f154770d844471f5a4da86efdf14

Observation 72949077-1963-4077-8d90-a55c1fad2c37 · outbound

This paper cites GPT-4o mini: advancing cost-efficient intelligence , 2024.

Self-Ablating Transformers: More Interpretability, Less Sparsity GPT-4o mini: advancing cost-efficient intelligence , 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:46.016687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.594181Z digest=sha256:d73eaa52c9faf020ff6bd9ccc46301439e8c197c33672f9b10688ae41edfe6ce

Observation a50be9b8-c0e8-4689-8843-2a24132ba8e9 · outbound

This paper cites GPT-4 Technical Report.

Self-Ablating Transformers: More Interpretability, Less Sparsity GPT-4 Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.597501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.597501Z digest=sha256:3d1f63fb0d2c89a709ad3ef3fcc08d755e919c1f69a2331da2043b87bf6eef8c

Observation 9965ed0e-3f4f-4d70-ba14-cadc8ab51d42 · outbound

This paper cites why should i trust you?.

Self-Ablating Transformers: More Interpretability, Less Sparsity why should i trust you?

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.601541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.601541Z digest=sha256:f36c1556ed564c09d9394b0c4f45b3281b0a81c6808c27c266580af651d67137

Observation 3b5abdd8-36a2-4c5f-b059-bf60fb302c5c · outbound

This paper cites Stop Explaining Black Box Machine Learning Models for High Stakes Decisions and Use Interpretable Models Instead.

Self-Ablating Transformers: More Interpretability, Less Sparsity Stop Explaining Black Box Machine Learning Models for High Stakes Decisions and Use Interpretable Models Instead

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.604891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.604891Z digest=sha256:ee9999151f7702f1875cfcbb691a2db85a8e66150018a67267adb5ea8425ad50

Observation 98a2693c-06e7-46f3-b96a-6520b5e8bfad · outbound

This paper cites SAEBench: A Comprehensive Benchmark for Sparse Autoencoders - Dec 2024 , January 2025.

Self-Ablating Transformers: More Interpretability, Less Sparsity SAEBench: A Comprehensive Benchmark for Sparse Autoencoders - Dec 2024 , January 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:46.004687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.609196Z digest=sha256:7645fb38f6590e4a87c10ad80a8b0097d346d9a7ce3e3e28570f2a0123455b82

Observation 8ecd6819-0c28-4f33-8f3a-c9c8fc1fea21 · outbound

This paper cites DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter.

Self-Ablating Transformers: More Interpretability, Less Sparsity DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.612842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.612842Z digest=sha256:b41e7dd3139d98c7db7eb5b72e4fe8dace35b889f76a048d9198e4c01bf60c57

Observation 7003ab81-db8a-4725-880b-0d51a484d87d · outbound

This paper cites The neural lasso: Local linear sparsity for interpretable explanations.

Self-Ablating Transformers: More Interpretability, Less Sparsity The neural lasso: Local linear sparsity for interpretable explanations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:45.992816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.615666Z digest=sha256:f388d92217d15ad16461b24f404caa647bb6dd789a17fb96f98114605177ebb9

Observation 369188da-d88d-4bae-b066-337bac871d4f · outbound

This paper cites Dropout: A simple way to prevent neural networks from overfitting.

Self-Ablating Transformers: More Interpretability, Less Sparsity Dropout: A simple way to prevent neural networks from overfitting

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.618279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.618279Z digest=sha256:4878629400cd19c2cdc3eccab385d287f36946e5b0686075137126ab584ba6b5

Observation 855a7c8c-0a35-4b55-88dd-0cbe1ec9fbd5 · outbound

This paper cites Codebook Features: Sparse and Discrete Interpretability for Neural Networks.

Self-Ablating Transformers: More Interpretability, Less Sparsity Codebook Features: Sparse and Discrete Interpretability for Neural Networks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.621423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.621423Z digest=sha256:c193439492a065cc519fbe24a876cbe71a888983e4b157a517bd9d7842c63e84

Observation 78e7904b-0e24-4f5a-ab81-63482c1df7ab · outbound

This paper cites Attention is all you need.

Self-Ablating Transformers: More Interpretability, Less Sparsity Attention is all you need

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.624315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.624315Z digest=sha256:b2d0ba2518fe3ffbdb235018f443eae842bb0996e992a53dacd396dcda9848e2

Observation 342cadae-d6d4-4391-a89c-025172654035 · outbound

This paper cites Interpretability in the wild: a circuit for indirect object identification in GPT -2 small.

Self-Ablating Transformers: More Interpretability, Less Sparsity Interpretability in the wild: a circuit for indirect object identification in GPT -2 small

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.628392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.628392Z digest=sha256:080ceaae447a393401a7fffb26aac2c3d97d9c45cc6b80febb9889e185826be7

Observation c37afdf1-3b83-4e62-b18c-2dae14fe1f99 · outbound

This paper cites MMLU -pro: A more robust and challenging multi-task language understanding benchmark.

Self-Ablating Transformers: More Interpretability, Less Sparsity MMLU -pro: A more robust and challenging multi-task language understanding benchmark

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.633209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.633209Z digest=sha256:abd1246a21a89b5bbc99cabcd56e9a058c907ea4f2223417c877118cec0821c4

Observation 18097027-a1fd-4a2c-bc16-d81cbbec50cc · outbound

This paper cites Understanding deep learning requires rethinking generalization.

Self-Ablating Transformers: More Interpretability, Less Sparsity Understanding deep learning requires rethinking generalization

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.637041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.637041Z digest=sha256:c488c70a9fa6276310b8ad8c494c5da39a8c70d5ad933cdfe48dbbbe6d4dea0f

Observation 578d8e96-1a14-493f-a267-199101320d09 · outbound

This paper cites Towards best practices of activation patching in language models: Metrics and methods.

Self-Ablating Transformers: More Interpretability, Less Sparsity Towards best practices of activation patching in language models: Metrics and methods

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:43:45.956445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T04:43:45.640924Z digest=sha256:75bdaf911bfee7a580e09618793609ca5a16cef26f1440707699293967634e02

Observation 70200a96-f2ef-4ca6-be3f-747401284603 · outbound

This paper cites @esa (Ref.

Self-Ablating Transformers: More Interpretability, Less Sparsity @esa (Ref

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.644152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.644152Z digest=sha256:6644da37442774207287121e57c0b3962cd491146b51af00de9b9f3db7a80ba2

Observation 60a58f32-d57d-4753-9079-9dd0f9ee2460 · outbound

This paper cites an unresolved cited work.

Self-Ablating Transformers: More Interpretability, Less Sparsity Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.648135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.648135Z digest=sha256:1f3b1ec395ef93403a4975d85a118378cea0802473c3e2547876fc7611283878

Observation 12c73708-74c6-4cb9-8ced-61dd33e6d50a · outbound

This paper cites an unresolved cited work.

Self-Ablating Transformers: More Interpretability, Less Sparsity Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:45.651719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:45.651719Z digest=sha256:7fbae53fc8e821d3ad85176839f2cfe2aed6d1cd344bfe925e3675bd118011a4

Pith citing papers

No inbound Pith citation observations are available.