Pith. sign in

Paper Citation Record · LEDGER

Transcoders Find Interpretable LLM Feature Circuits

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2406.11944.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.11944 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 40 of 40 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:06:05.025646Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

6
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 379428b2-cf50-43eb-abad-fc51edc418c7 · inbound

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions cites this paper.

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions Transcoders Find Interpretable LLM Feature Circuits

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T20:37:54.690268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:37:54.690268Z digest=sha256:b4ae780a89317d4cdfc3e74d45ade43ecaff97706ed7b8146299179a3e7c0729

Observation 0c22a108-5382-4e22-88ba-fb89956978d7 · inbound

InterPLM: Discovering Interpretable Features in Protein Language Models via Sparse Autoencoders cites this paper.

InterPLM: Discovering Interpretable Features in Protein Language Models via Sparse Autoencoders Transcoders Find Interpretable LLM Feature Circuits

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T21:17:17.200001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:17:17.200001Z digest=sha256:3c72f028378d87b960997d50cd7ed09800a9426a7d8cbd5591093826b8d5d29b

Observation 857d6d2a-c940-435b-af91-f7e7e6bbd5e5 · inbound

Interpretability in Parameter Space: Minimizing Mechanistic Description Length with Attribution-based Parameter Decomposition cites this paper.

Interpretability in Parameter Space: Minimizing Mechanistic Description Length with Attribution-based Parameter Decomposition Transcoders Find Interpretable LLM Feature Circuits

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T14:55:02.002736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T14:55:02.002736Z digest=sha256:2699e57ba27f936f499ca8ba58a97fd9ec3db9b0256ce94285dcee72bae672a8

Observation 0e814619-2107-4595-9062-a38af8b9a3cd · inbound

Transcoders Beat Sparse Autoencoders for Interpretability cites this paper.

Transcoders Beat Sparse Autoencoders for Interpretability Transcoders Find Interpretable LLM Feature Circuits

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T22:24:53.833482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:24:53.833482Z digest=sha256:6b2a625b9a81bead4eb9bb814be095d3e37a83565b7f29ee4bf0d3041d5ac986

Observation bd32d894-72f3-43de-a43d-bd0ce0c9275c · inbound

Partially Rewriting a Transformer in Natural Language cites this paper.

Partially Rewriting a Transformer in Natural Language Transcoders Find Interpretable LLM Feature Circuits

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T22:21:50.720744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:21:50.720744Z digest=sha256:dde04493411af46975e2dbafe95e39ee9673be25654e91c06cb96ad39c06c034

Observation bd06ef27-8348-4db3-9da3-6a3aeef9afdb · inbound

Perspectives for Direct Interpretability in Multi-Agent Deep Reinforcement Learning cites this paper.

Perspectives for Direct Interpretability in Multi-Agent Deep Reinforcement Learning Transcoders Find Interpretable LLM Feature Circuits

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T17:58:34.177121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:58:34.177121Z digest=sha256:649b8e72cad9c2fc4edde7cd5dd3623dad57b02334fa09a474aae9bbcd966b1f

Observation b7aadacf-6725-4519-b232-02aba34b30f3 · inbound

Analyze Feature Flow to Enhance Interpretation and Steering in Language Models cites this paper.

Analyze Feature Flow to Enhance Interpretation and Steering in Language Models Transcoders Find Interpretable LLM Feature Circuits

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T10:11:51.469627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:11:51.469627Z digest=sha256:58b5cf45cfaf6f554890c5718fcc5eae5ae7fa9c0bdafd78b52257fe109d0781

Observation 874c877f-280d-497d-bf95-e3984e979e07 · inbound

Sparse Autoencoders Do Not Find Canonical Units of Analysis cites this paper.

Sparse Autoencoders Do Not Find Canonical Units of Analysis Transcoders Find Interpretable LLM Feature Circuits

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T21:12:13.815125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:12:13.815125Z digest=sha256:18cfe7cd1ebc2eb27aa8d47cbfc554dbd8696e68da1367d00e6c9eff52ad326b

Observation 41b87ddb-d447-4b82-87c5-ef8e8d67fea8 · inbound

Scaling sparse feature circuit finding for in-context learning cites this paper.

Scaling sparse feature circuit finding for in-context learning Transcoders Find Interpretable LLM Feature Circuits

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T12:06:05.025646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T12:06:05.025646Z digest=sha256:c49deff33ee6cfcf0340912ddf3926fcb2b82232edad3b5cb6c5999f1d01412c

Observation a5bf3096-13ae-4fc6-b63d-faf210e7058f · inbound

Towards Understanding the Nature of Attention with Low-Rank Sparse Decomposition cites this paper.

Towards Understanding the Nature of Attention with Low-Rank Sparse Decomposition Transcoders Find Interpretable LLM Feature Circuits

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T05:21:33.927686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:21:33.927686Z digest=sha256:12d10b42cf3e7550e1cc8307ac691fb44aa150f15075d4cd9b45f38eefddcf2c

Observation 1bdd8fd1-31fb-440f-ac7e-f730a830fca3 · inbound

Incorporating Hierarchical Semantics in Sparse Autoencoder Architectures cites this paper.

Incorporating Hierarchical Semantics in Sparse Autoencoder Architectures Transcoders Find Interpretable LLM Feature Circuits

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:15.162228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:15.162228Z digest=sha256:c50c1a3aca788d7217b06c8527321d16ed67966ac35ca30ce93368386fe3acf1

Observation 73fed3ac-c32c-4424-987d-535437ef2b76 · inbound

Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? cites this paper.

Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? Transcoders Find Interpretable LLM Feature Circuits

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:21.072683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:28:21.072683Z digest=sha256:c62a281332d39fcea9ed3cc17cf873e9fc36e8aee923ed6df069904055da6ea6

Observation 305bf029-f788-413c-a67e-86e59700806e · inbound

RCStat: A Statistical Framework for using Relative Contextualization in Transformers cites this paper.

RCStat: A Statistical Framework for using Relative Contextualization in Transformers Transcoders Find Interpretable LLM Feature Circuits

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T18:40:37.650165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:40:37.650165Z digest=sha256:e2ca91feadd3a0a37a2c118f4bfdb4af0c34d9a8640e0639077c5b25b6bd21a7

Observation 28e4f11c-d675-4fef-b678-360e2fcb902d · inbound

Cross-Layer Discrete Concept Discovery for Interpreting Language Models cites this paper.

Cross-Layer Discrete Concept Discovery for Interpreting Language Models Transcoders Find Interpretable LLM Feature Circuits

Reference 2021

Resolution
malformed identifier
no resolver link, observed 2026-08-06T23:03:04.924942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:03:04.924942Z digest=sha256:1e7791c39086c93639c42ad6db43cb6e8a26dc6a1bde2bc53976a940c0dd77c9

Observation 85e77101-b6ba-4207-bf9b-8c9ccf225721 · inbound

Large Reasoning Models are not thinking straight: on the unreliability of thinking trajectories cites this paper.

Large Reasoning Models are not thinking straight: on the unreliability of thinking trajectories Transcoders Find Interpretable LLM Feature Circuits

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:12.849806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:13:12.849806Z digest=sha256:f2dda19b6f558bb8ae5c9847c1d1501121ce1594eb7a7868655ff4cf8db2358a

Observation 45d1f631-5cf5-46aa-bc47-32db5aa3e311 · inbound

BlueGlass: A Framework for Composite AI Safety cites this paper.

BlueGlass: A Framework for Composite AI Safety Transcoders Find Interpretable LLM Feature Circuits

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T17:46:17.561067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:46:17.561067Z digest=sha256:72beee9cb91ad8361d0af9870a867202a53a74f562a4b434d04e30dbd82709bc

Observation ae07df43-5c5b-4ada-88af-d3dc7ad647cd · inbound

Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation cites this paper.

Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation Transcoders Find Interpretable LLM Feature Circuits

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T18:47:45.265327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:47:45.265327Z digest=sha256:dbefde711227986e72c44b0b0b55707ae972e1ca955bcb0c1d2031a48bf429de

Observation 7dfd39f7-2da9-4227-b077-c9105219996a · inbound

How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding cites this paper.

How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding Transcoders Find Interpretable LLM Feature Circuits

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:36.127974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:36.127974Z digest=sha256:5fa5a6dee4c644578ce1472e17bb0aaf86389a412697d05bead30009f09c8161

Observation b8f9e34c-0dec-4fe2-bf92-56b08729be5e · inbound

Bridging Mechanistic Interpretability and Prompt Engineering with Gradient Ascent for Interpretable Persona Control cites this paper.

Bridging Mechanistic Interpretability and Prompt Engineering with Gradient Ascent for Interpretable Persona Control Transcoders Find Interpretable LLM Feature Circuits

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T12:29:39.563314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:29:39.563314Z digest=sha256:6dea753424d5db47b0a99105a8990323529acaf144de18595ce5fdc6a8e43778

Observation 02ca933a-765e-4aa0-854d-013d0dba4500 · inbound

Emotion Concepts and their Function in a Large Language Model cites this paper.

Emotion Concepts and their Function in a Large Language Model Transcoders Find Interpretable LLM Feature Circuits

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:35:57.943307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T18:03:52.210931Z digest=sha256:afb0c9c407b8a1d4c5b0c6f12195f264163dfbcff30df0d2b8b462fd8b41e829

Observation b1e3d4f8-a0ed-472e-ac93-16fc02ebc1dc · inbound

Understanding the Mechanism of Altruism in Large Language Models cites this paper.

Understanding the Mechanism of Altruism in Large Language Models Transcoders Find Interpretable LLM Feature Circuits

Reference 236

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:31:02.317778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-10T01:36:50.329664Z digest=sha256:4a18c9a535ec06165996d3f7f0db203b35626caede526f0d583c82740598866f

Observation cb8bc351-cdbe-4d14-a2bc-2b64568af97d · inbound

Decoding Alignment without Encoding Alignment: A critique of similarity analysis in neuroscience cites this paper.

Decoding Alignment without Encoding Alignment: A critique of similarity analysis in neuroscience Transcoders Find Interpretable LLM Feature Circuits

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-09T00:24:28.670130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-08T03:30:33.288849Z digest=sha256:db3419a130175cd69060f25ce2f55a54c974dab67b7569bc6efadb8d0513300f

Observation 564dfb56-d1c4-4d0b-a39d-2fd46d858fcb · inbound

From Token Lists to Graph Motifs: Weisfeiler-Lehman Analysis of Sparse Autoencoder Features cites this paper.

From Token Lists to Graph Motifs: Weisfeiler-Lehman Analysis of Sparse Autoencoder Features Transcoders Find Interpretable LLM Feature Circuits

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:11.055181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T09:41:01.775116Z digest=sha256:6d179007f8204c7c8aa1cb821159d397d78bdbb2823e22b0d978e0d776591fc0

Observation 2805d20e-ad21-4380-bcd2-276cf7b4d5fe · inbound

From Mechanistic to Compositional Interpretability cites this paper.

From Mechanistic to Compositional Interpretability Transcoders Find Interpretable LLM Feature Circuits

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:46:18.555932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-12T02:42:26.173782Z digest=sha256:67163d0e0b1fa84afbb95891979657c349a596e5f9f66501b033689e4a2ca6b8

Observation 53c7f48d-b347-45eb-8d35-4ce21ec90d9e · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Transcoders Find Interpretable LLM Feature Circuits

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:53:47.242717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-20T21:49:47.934339Z digest=sha256:a45d32edb108748ed7b04a600e8ff95cb1bb707001f4af8c369a3e0b75e4e221

Observation 86342e52-49f8-45e7-89d4-378414b619c6 · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Transcoders Find Interpretable LLM Feature Circuits

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:49:50.160301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T07:46:41.159688Z digest=sha256:dfdb907f9f37af2afb466178683652e82a9ae38bf663d2369c86ee2e10dad609

Observation e60af32d-7b8b-4a3a-9ba8-67e40602853a · inbound

Reading Task Failure Off the Activations: A Sparse-Feature Audit of GPT-2 Small on Indirect Object Identification cites this paper.

Reading Task Failure Off the Activations: A Sparse-Feature Audit of GPT-2 Small on Indirect Object Identification Transcoders Find Interpretable LLM Feature Circuits

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:44:42.756151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T07:42:48.816730Z digest=sha256:c10838fa7b9901361fd51498c757b650ba26c5aab06357e581514c6bbef57ce0

Observation 31f5b5fb-32e8-4f81-81f0-6a12c968de41 · inbound

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions cites this paper.

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions Transcoders Find Interpretable LLM Feature Circuits

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:34:48.510603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T15:17:37.904831Z digest=sha256:01c507442347ae5e551006f2c35cf441773500a801a4ca922dfa2061be18322a

Observation aec6e571-7567-496c-8720-770632b714cb · inbound

Polymorphism Is Rotation: Operational Mechanistic Interpretability from a Two-Layer Transformer to Pythia-70m cites this paper.

Polymorphism Is Rotation: Operational Mechanistic Interpretability from a Two-Layer Transformer to Pythia-70m Transcoders Find Interpretable LLM Feature Circuits

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:14:45.542319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T14:10:42.640805Z digest=sha256:9109ee1a151f2f74a87ceae86d9153c8b10358fd104c80f0cc9f68e7047fd224

Observation 3ac811f4-7e2f-4f79-bca5-30038b4f1f63 · inbound

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet cites this paper.

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet Transcoders Find Interpretable LLM Feature Circuits

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.459937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:50:04.813379Z digest=sha256:105d66c2b1bd2d9a58a8cf1cd818da4b50895e8dad7fb1dbd6867fb8dc97f47d

Observation 01004121-5282-4dc4-a141-4652f2c23ad2 · inbound

Sparsely gated tiny linear experts cites this paper.

Sparsely gated tiny linear experts Transcoders Find Interpretable LLM Feature Circuits

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:17:09.314503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T22:49:49.299925Z digest=sha256:75f47c314f8058792f04a4659c9b6693b405fb5c4d0d8e207cd7d77f36d178d6

Observation 6d595113-8780-425b-914e-f9aa7c96e2f2 · inbound

Interactions Between Crosscoder Features: A Compact Proofs Perspective cites this paper.

Interactions Between Crosscoder Features: A Compact Proofs Perspective Transcoders Find Interpretable LLM Feature Circuits

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:28.669471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T17:25:55.152823Z digest=sha256:c8dc56fcb9129ec9185f185d18aa5d27d1e37bf4e64a62ac202309084dfe9bf4

Observation 70c3af1d-2434-48eb-bc8a-6922bae50f2c · inbound

Mechanistic Interpretability for Neural Networks: Circuits, Sparse Features and Symbolic Reasoning cites this paper.

Mechanistic Interpretability for Neural Networks: Circuits, Sparse Features and Symbolic Reasoning Transcoders Find Interpretable LLM Feature Circuits

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-07-09T15:06:18.057814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-09T14:58:58.363330Z digest=sha256:e3bb54c50ce9cae45395f0382a8eeba51bf4c17362d8419ba72911cd3c7a792e

Observation 1dca91b4-4811-43b2-a629-62de8707da3d · inbound

Targeted Recovery of Weight-Space Mechanisms From Neural Networks cites this paper.

Targeted Recovery of Weight-Space Mechanisms From Neural Networks Transcoders Find Interpretable LLM Feature Circuits

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T10:39:47.713797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:39:47.713797Z digest=sha256:0f0955654657b3616cfdcba22210e1af107b1c8401163b0c0dfd28fd9ae86974

Observation b6f75f6e-e1b0-43d6-a9e3-fec9811898fd · inbound

Targeted Recovery of Weight-Space Mechanisms From Neural Networks cites this paper.

Targeted Recovery of Weight-Space Mechanisms From Neural Networks Transcoders Find Interpretable LLM Feature Circuits

Reference 156

Resolution
unresolved
no resolver link, observed 2026-08-02T10:40:01.415862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:40:01.415862Z digest=sha256:86e60cb6361ca63e40255a64051e2c020f4805d8624e84e721ca95cc8222bf28

Observation 9b9bd103-8203-42ef-adc8-eda0accbfd16 · inbound

Transcoders for Investigating Deception in Language Models cites this paper.

Transcoders for Investigating Deception in Language Models Transcoders Find Interpretable LLM Feature Circuits

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T01:06:13.602698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T01:06:13.602698Z digest=sha256:f65e66cdd3366ea7c80c706c2997807ce9fe79bea1a5863b10dc56c9db591484

Observation 914133ed-5878-495d-985e-60fd8a0ad5e6 · inbound

Verbalizable Representations Form a Global Workspace in Language Models cites this paper.

Verbalizable Representations Form a Global Workspace in Language Models Transcoders Find Interpretable LLM Feature Circuits

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T23:15:20.650998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:15:20.650998Z digest=sha256:aac08f9d80392d3625199af5b4af8e3c87fc889c12086ecbc67cbfbae5c3a2c5

Observation 54730b82-5245-4b7e-bd5e-c4fd189d1119 · inbound

Retrieval is Enough: Training-Free Interpretability with a Tool-Using Agent cites this paper.

Retrieval is Enough: Training-Free Interpretability with a Tool-Using Agent Transcoders Find Interpretable LLM Feature Circuits

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T20:58:51.267585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:58:51.267585Z digest=sha256:1e1c58fde452ccdd47e41bf073aeb6ec93cc29e6cd89f5c13d16c4219c71084e

Observation 9b6cba91-b16a-4c16-b976-b1f8874741f8 · inbound

Sparse Weight Decomposition for Efficient Circuit Extraction cites this paper.

Sparse Weight Decomposition for Efficient Circuit Extraction Transcoders Find Interpretable LLM Feature Circuits

Reference 2004

Resolution
unresolved
no resolver link, observed 2026-08-15T14:51:45.393010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:51:45.393010Z digest=sha256:239b2844490909fa7335e23dbd6e1e12ee82e812ba7d1a2531c961c4c87adbe1

Observation e2c8b7a4-14b2-435e-b6d2-227558dc75df · inbound

Where You Measure Decides What You Measure: Position Selection in Ablation-Based SAE Evaluation cites this paper.

Where You Measure Decides What You Measure: Position Selection in Ablation-Based SAE Evaluation Transcoders Find Interpretable LLM Feature Circuits

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T13:29:00.120325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:29:00.120325Z digest=sha256:902cb3e8762a866380bd09adafd070dfcee2ff6a59c1c7bf180e7e09835c481d