Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2406.11944.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:06:05.025646Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
6
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 379428b2-cf50-43eb-abad-fc51edc418c7 · inbound
A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions Transcoders Find Interpretable LLM Feature Circuits
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c22a108-5382-4e22-88ba-fb89956978d7 · inbound
InterPLM: Discovering Interpretable Features in Protein Language Models via Sparse Autoencoders Transcoders Find Interpretable LLM Feature Circuits
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 857d6d2a-c940-435b-af91-f7e7e6bbd5e5 · inbound
Interpretability in Parameter Space: Minimizing Mechanistic Description Length with Attribution-based Parameter Decomposition Transcoders Find Interpretable LLM Feature Circuits
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e814619-2107-4595-9062-a38af8b9a3cd · inbound
Transcoders Beat Sparse Autoencoders for Interpretability Transcoders Find Interpretable LLM Feature Circuits
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd32d894-72f3-43de-a43d-bd0ce0c9275c · inbound
Partially Rewriting a Transformer in Natural Language Transcoders Find Interpretable LLM Feature Circuits
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd06ef27-8348-4db3-9da3-6a3aeef9afdb · inbound
Perspectives for Direct Interpretability in Multi-Agent Deep Reinforcement Learning Transcoders Find Interpretable LLM Feature Circuits
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7aadacf-6725-4519-b232-02aba34b30f3 · inbound
Analyze Feature Flow to Enhance Interpretation and Steering in Language Models Transcoders Find Interpretable LLM Feature Circuits
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 874c877f-280d-497d-bf95-e3984e979e07 · inbound
Sparse Autoencoders Do Not Find Canonical Units of Analysis Transcoders Find Interpretable LLM Feature Circuits
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41b87ddb-d447-4b82-87c5-ef8e8d67fea8 · inbound
Scaling sparse feature circuit finding for in-context learning Transcoders Find Interpretable LLM Feature Circuits
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5bf3096-13ae-4fc6-b63d-faf210e7058f · inbound
Towards Understanding the Nature of Attention with Low-Rank Sparse Decomposition Transcoders Find Interpretable LLM Feature Circuits
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bdd8fd1-31fb-440f-ac7e-f730a830fca3 · inbound
Incorporating Hierarchical Semantics in Sparse Autoencoder Architectures Transcoders Find Interpretable LLM Feature Circuits
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73fed3ac-c32c-4424-987d-535437ef2b76 · inbound
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? Transcoders Find Interpretable LLM Feature Circuits
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 305bf029-f788-413c-a67e-86e59700806e · inbound
RCStat: A Statistical Framework for using Relative Contextualization in Transformers Transcoders Find Interpretable LLM Feature Circuits
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28e4f11c-d675-4fef-b678-360e2fcb902d · inbound
Cross-Layer Discrete Concept Discovery for Interpreting Language Models Transcoders Find Interpretable LLM Feature Circuits
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85e77101-b6ba-4207-bf9b-8c9ccf225721 · inbound
Large Reasoning Models are not thinking straight: on the unreliability of thinking trajectories Transcoders Find Interpretable LLM Feature Circuits
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45d1f631-5cf5-46aa-bc47-32db5aa3e311 · inbound
BlueGlass: A Framework for Composite AI Safety Transcoders Find Interpretable LLM Feature Circuits
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae07df43-5c5b-4ada-88af-d3dc7ad647cd · inbound
Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation Transcoders Find Interpretable LLM Feature Circuits
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dfd39f7-2da9-4227-b077-c9105219996a · inbound
How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding Transcoders Find Interpretable LLM Feature Circuits
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8f9e34c-0dec-4fe2-bf92-56b08729be5e · inbound
Bridging Mechanistic Interpretability and Prompt Engineering with Gradient Ascent for Interpretable Persona Control Transcoders Find Interpretable LLM Feature Circuits
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02ca933a-765e-4aa0-854d-013d0dba4500 · inbound
Emotion Concepts and their Function in a Large Language Model Transcoders Find Interpretable LLM Feature Circuits
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b1e3d4f8-a0ed-472e-ac93-16fc02ebc1dc · inbound
Understanding the Mechanism of Altruism in Large Language Models Transcoders Find Interpretable LLM Feature Circuits
Reference 236
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation cb8bc351-cdbe-4d14-a2bc-2b64568af97d · inbound
Decoding Alignment without Encoding Alignment: A critique of similarity analysis in neuroscience Transcoders Find Interpretable LLM Feature Circuits
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 564dfb56-d1c4-4d0b-a39d-2fd46d858fcb · inbound
From Token Lists to Graph Motifs: Weisfeiler-Lehman Analysis of Sparse Autoencoder Features Transcoders Find Interpretable LLM Feature Circuits
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2805d20e-ad21-4380-bcd2-276cf7b4d5fe · inbound
From Mechanistic to Compositional Interpretability Transcoders Find Interpretable LLM Feature Circuits
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 53c7f48d-b347-45eb-8d35-4ce21ec90d9e · inbound
WriteSAE: Sparse Autoencoders for Recurrent State Transcoders Find Interpretable LLM Feature Circuits
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 86342e52-49f8-45e7-89d4-378414b619c6 · inbound
WriteSAE: Sparse Autoencoders for Recurrent State Transcoders Find Interpretable LLM Feature Circuits
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e60af32d-7b8b-4a3a-9ba8-67e40602853a · inbound
Reading Task Failure Off the Activations: A Sparse-Feature Audit of GPT-2 Small on Indirect Object Identification Transcoders Find Interpretable LLM Feature Circuits
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 31f5b5fb-32e8-4f81-81f0-6a12c968de41 · inbound
ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions Transcoders Find Interpretable LLM Feature Circuits
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation aec6e571-7567-496c-8720-770632b714cb · inbound
Polymorphism Is Rotation: Operational Mechanistic Interpretability from a Two-Layer Transformer to Pythia-70m Transcoders Find Interpretable LLM Feature Circuits
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3ac811f4-7e2f-4f79-bca5-30038b4f1f63 · inbound
Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet Transcoders Find Interpretable LLM Feature Circuits
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 01004121-5282-4dc4-a141-4652f2c23ad2 · inbound
Sparsely gated tiny linear experts Transcoders Find Interpretable LLM Feature Circuits
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6d595113-8780-425b-914e-f9aa7c96e2f2 · inbound
Interactions Between Crosscoder Features: A Compact Proofs Perspective Transcoders Find Interpretable LLM Feature Circuits
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 70c3af1d-2434-48eb-bc8a-6922bae50f2c · inbound
Mechanistic Interpretability for Neural Networks: Circuits, Sparse Features and Symbolic Reasoning Transcoders Find Interpretable LLM Feature Circuits
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 1dca91b4-4811-43b2-a629-62de8707da3d · inbound
Targeted Recovery of Weight-Space Mechanisms From Neural Networks Transcoders Find Interpretable LLM Feature Circuits
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6f75f6e-e1b0-43d6-a9e3-fec9811898fd · inbound
Targeted Recovery of Weight-Space Mechanisms From Neural Networks Transcoders Find Interpretable LLM Feature Circuits
Reference 156
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b9bd103-8203-42ef-adc8-eda0accbfd16 · inbound
Transcoders for Investigating Deception in Language Models Transcoders Find Interpretable LLM Feature Circuits
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 914133ed-5878-495d-985e-60fd8a0ad5e6 · inbound
Verbalizable Representations Form a Global Workspace in Language Models Transcoders Find Interpretable LLM Feature Circuits
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54730b82-5245-4b7e-bd5e-c4fd189d1119 · inbound
Retrieval is Enough: Training-Free Interpretability with a Tool-Using Agent Transcoders Find Interpretable LLM Feature Circuits
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b6cba91-b16a-4c16-b976-b1f8874741f8 · inbound
Sparse Weight Decomposition for Efficient Circuit Extraction Transcoders Find Interpretable LLM Feature Circuits
Reference 2004
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2c8b7a4-14b2-435e-b6d2-227558dc75df · inbound
Where You Measure Decides What You Measure: Position Selection in Ablation-Based SAE Evaluation Transcoders Find Interpretable LLM Feature Circuits
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.