Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 36 inbound Pith citation observations for arXiv:2301.04709.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:19:44.215403Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
10
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation f2e48289-ffce-497a-9a7b-aa5db8fe4d98 · inbound
Localizing Model Behavior with Path Patching Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ac85d2ad-2034-4509-b7ff-a108d493aa27 · inbound
Linear Representations of Sentiment in Large Language Models Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 072600d1-7c4a-4cc7-98f8-874fb6884ae6 · inbound
Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e44f166a-e856-466f-9106-7dacbddeacb6 · inbound
Factored space models: Towards causality between levels of abstraction Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07fe9ccb-2b20-411f-ba91-3906b2393cb6 · inbound
Removing Spurious Correlation from Neural Network Interpretations Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a94c6dd8-759c-4b98-b19c-1e391c6f6b8e · inbound
What is causal about causal models and representations? Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d7cd494-b1de-4b9b-b4bd-199414ac0a7c · inbound
MIB: A Mechanistic Interpretability Benchmark Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22aea982-0987-44f7-ab5b-592ababe55dd · inbound
Evaluating Explanations: An Explanatory Virtues Framework for Mechanistic Interpretability -- The Strange Science Part I.ii Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 831539e5-aef9-43d1-9f5a-1fa53e94a015 · inbound
$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60e2b358-386b-40d1-b04e-197bf857f57f · inbound
Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8240af5a-2aed-4c27-985e-c4e89b35704f · inbound
Explaining Neural Networks with Reasons Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed653d77-400e-476f-b68f-b3871cd5a005 · inbound
Position: Mechanistic Interpretability Should Prioritize Feature Consistency in SAEs Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60b65955-83ca-44f0-98ca-bf7a822f87cd · inbound
How Do Transformers Learn Variable Binding in Symbolic Programs? Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a99928f2-fe36-4677-b909-72e3eae049fc · inbound
Identifying a Circuit for Verb Conjugation in GPT-2 Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88fa77e4-1605-455c-a09d-0b79149d539c · inbound
Identifiability in Causal Abstractions: A Hierarchy of Criteria Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e43d9ef-db7d-46cd-aa37-b05d7851210e · inbound
Can Interpretation Predict Behavior on Unseen Data? Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ecdeea4-ca1e-4230-8825-1855bc0ca1db · inbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7dee443-348d-4baa-98ad-9c9fe9f2a572 · inbound
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03fbbe87-c027-441e-a4b8-17876f48c024 · inbound
LLMs Should Not Yet Be Credited with Decision Explanation Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 38219848-7677-43dd-aed7-9a0cbec2407f · inbound
Bucketing the Good Apples: A Method for Diagnosing and Improving Causal Abstraction Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9395ec40-e26a-4326-91d3-68fa7c649b7c · inbound
From Mechanistic to Compositional Interpretability Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 147
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dcc8bb27-776b-4289-895d-4364afb81e32 · inbound
Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 540a2dd8-b6c1-4a46-b7d8-b72425f6549b · inbound
From Weight Perturbation to Feature Attribution for Explaining Fully Connected Neural Networks Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e0173868-2849-46b3-9fed-ba9a2a8aec7b · inbound
Causal Tongue-Tie: LLMs Can Encode Causal Direction, But Their Yes/No Outputs Fail to Express Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 82492bf6-f21a-4999-941f-563b442d697d · inbound
Probing LLMs for Syntactic Structure Beyond Universal Dependencies: A Minimalist Phase Account in English Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9dbb2565-b89d-4eed-9a79-7e1331e76973 · inbound
Probing LLMs for Syntactic Structure Beyond Universal Dependencies: A Minimalist Phase Account in English Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c934caf3-ffe0-40ac-961f-61d6a5a5394a · inbound
Temporal Preference Concepts and their Functions in a Large Language Model Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7b896aa9-f0cd-4ff4-8612-394cf17fddc8 · inbound
Temporal Preference Concepts and their Functions in a Large Language Model Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2455939-37e6-42cb-8548-afb76926aa3b · inbound
Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6204ff38-c9a2-4847-8cf6-6787df2f3870 · inbound
When Behavioral Safety Evaluation Fails: A Representation-Level Perspective Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 201b7313-9d69-4b6e-b126-58c5af926390 · inbound
XtrAIn: Training-Guided Occlusion for Feature Attribution Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 832bb100-a667-4f0d-adbb-bf5ed7e9d6c0 · inbound
Interpreting Neural Combinatorial Optimization via Evolving Programmatic Bottlenecks Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c6fe8413-40de-4ea3-aecc-856b9dbde6f4 · inbound
Steering Vision-Language Models with Joint Sparse Autoencoders Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 51b25163-ad74-4382-87a1-03216aaa2ef6 · inbound
Safety from Honesty in a Disinterested AI Predictor Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 890e0878-d980-4dd5-9b5e-32e7b8e7c48c · inbound
Safety from Honesty in a Disinterested AI Predictor Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 891db3d4-d0ba-45a8-b03d-7375297297fd · inbound
Are Single-Token Sparse Autoencoder Features Causally Necessary? Layer-Depth and SAE-Family Effects Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.