Pith. sign in

Paper Citation Record · LEDGER

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding

As of 16 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 3 inbound Pith citation observations for arXiv:2507.22928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22928 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:21:13.021612Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:23:50.436202Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 589de791-c76f-43cd-8f54-3e9053be49ce · outbound

This paper cites , " * write output.state after.block = add.period write newline.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.861030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.861030Z digest=sha256:fc8410b4e2bddb89f2f387188889553f2ca3a72ce6c68a316e9573bb8d430572

Observation 8cba9de4-e288-4b6a-a03b-22af76670bf0 · outbound

This paper cites write newline.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.865106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.865106Z digest=sha256:81762bf1ec2e21067931dddbda7ba7292beedc7ce20323e7a25ada7f6b956f10

Observation 106c5eff-deb8-4ab6-bc70-350596a3e86b · outbound

This paper cites Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.869648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.869648Z digest=sha256:e013491c756d8f3249443528feb123fefe108b7b7b52a5d7b3a10dcab301725c

Observation b03b5c8c-4166-4d74-8ec5-fd57b9eb772e · outbound

This paper cites Faithfulness Tests for Natural Language Explanations.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Faithfulness Tests for Natural Language Explanations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.873916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.873916Z digest=sha256:11936c9da3d4e6b469a68605276401f39b2797baf5681ebe615b1407acdc1ec5

Observation 78812c5a-0f7a-42f8-bf0a-e3bcb12a2722 · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.878718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.878718Z digest=sha256:64494cfd931349744b36751fb5afca6dd4f891998568ca0c29dc74dfe2aa6e3c

Observation b6a19327-babd-4d79-942c-b6ca9de3979d · outbound

This paper cites Mechanistic Interpretability for AI Safety -- A Review.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Mechanistic Interpretability for AI Safety -- A Review

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.882758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.882758Z digest=sha256:74a7563c2fc5a88701d3059de29b1f9d0ce36b46848b01c82d5f83e01680bdd6

Observation 13ba8ede-40fb-478f-9649-b849430f5504 · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.886688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.886688Z digest=sha256:0999d06b56f8fb5256f2920150c66251fe51f5f74f07b69c627b8d7e741b80e3

Observation b252c6dd-a085-447d-84a0-45df8c31833c · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:21:13.695349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:12.890429Z digest=sha256:9eed03a0fa4be1cc9ea9a352f3acdc2c02efff76dd8f36883771e53bd37c6e4a

Observation 61588e45-37d7-460e-a66d-4d184a866645 · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.894293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.894293Z digest=sha256:009a5fae305422cd75e175f21d4befd11e73edd044eaf72bf4f8d353fc61b6d3

Observation 459d7a01-4595-4e8a-b906-e9afa15921a1 · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.897903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.897903Z digest=sha256:9cb367ab20fcf2c2dfac3059c959554d906229f43eef3f44706e547d7c58ee8a

Observation def9c75a-a17d-475b-a93b-b8aef5fe897a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Training Verifiers to Solve Math Word Problems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.903109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.903109Z digest=sha256:18e7943f2ecd683d44f7c5b9184cbecde521c7329a23ce73bca3583ca0728ab2

Observation 168f8ef2-7449-45bf-80ee-9b7cb38b18a5 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.906793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.906793Z digest=sha256:0e0b5b557c947b388faf958afe64d273d2b71a9d9d6b6aeffcaa9657aad80fc7

Observation e6ca6c23-b91a-41c4-af27-5b27f1d34dc7 · outbound

This paper cites Sparse Autoencoders Reveal Temporal Difference Learning in Large Language Models.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Sparse Autoencoders Reveal Temporal Difference Learning in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.910702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.910702Z digest=sha256:017244f2af4b1b1634ad934216fcf85cb0eff03c537c92503c68cf9136072161

Observation 485cd831-2509-4d88-939e-77bf015a0b33 · outbound

This paper cites Tokenized SAEs: Disentangling SAE Reconstructions.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Tokenized SAEs: Disentangling SAE Reconstructions

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.914460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.914460Z digest=sha256:eb27e64b36fc351aa8810f91250a03a7914b532472d876adaafa037ca8b23144

Observation 7b685728-a967-4ea0-8ae6-1d06d751be4c · outbound

This paper cites How to think step-by-step: A mechanistic understanding of chain-of-thought reasoning.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding How to think step-by-step: A mechanistic understanding of chain-of-thought reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.918891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.918891Z digest=sha256:bc504887b66c0e25bdf2a70495addddc307fb95f9372e1b1c2535103606de790

Observation 35ad841d-6112-4508-badd-4358206a5a3d · outbound

This paper cites Toy Models of Superposition.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Toy Models of Superposition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.922606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.922606Z digest=sha256:79c9d663581a60499092f9ad7628d595f6287cbb77eeea2c5989e534e6eb5a6e

Observation d7dee443-348d-4baa-98ad-9c9fe9f2a572 · outbound

This paper cites Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.926237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.926237Z digest=sha256:6e47bc024f583c0a5b5879f262befa24a49929c039f96ea9001c2b9363d33493

Observation d950279e-504f-4308-987a-7645c75dc822 · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:21:13.677354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:12.929728Z digest=sha256:f3df696bf4779b13007ff0c82768d08aaa7f9864e47b18d91f204d1eb6552a7c

Observation 7b5b762f-7987-4fcb-863c-05c41b7b873e · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:21:13.666656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:12.933466Z digest=sha256:e28b8cdc3f2b3b24bf4469fcf52aebe2d66fc4beb241196c9789f6070b387588

Observation 25152b33-d5e2-4a7f-9ede-46a2cef824ca · outbound

This paper cites Localizing Model Behavior with Path Patching.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Localizing Model Behavior with Path Patching

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.936927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.936927Z digest=sha256:0a8d1f38b4a1656a88d86562c36042b16b1a947ad1f562ef5601ace4b6f03fd4

Observation e151c6df-458c-4509-a624-7cfb9d0b00bb · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:21:13.655664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:12.940718Z digest=sha256:1b9a45444dabfc2cc87d069b241d265c8ea1e5eade215b97d27a1331b292629a

Observation fb418409-32c5-4df7-8c80-eac30f3b189e · outbound

This paper cites How to use and interpret activation patching.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding How to use and interpret activation patching

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.944225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.944225Z digest=sha256:271f75433a840a0b391bdeb90d9929ec00c0113a2830ca7f31cfbef8d36f3d9b

Observation 1f992757-a0f2-4c94-8baf-fd3dd9143563 · outbound

This paper cites Evaluating Sparse Autoencoders on Targeted Concept Erasure Tasks.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Evaluating Sparse Autoencoders on Targeted Concept Erasure Tasks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.948006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.948006Z digest=sha256:905325499e6e261f9eab8b0b0eb5726a36f413213fa74a91c6ea46e3cd6f00a8

Observation 11333139-959c-4dae-8959-2d3e5f45c216 · outbound

This paper cites S.; Reid, M.; Matsuo, Y.; and Iwasawa, Y.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding S.; Reid, M.; Matsuo, Y.; and Iwasawa, Y

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.951668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.951668Z digest=sha256:14235e284ab2b131806635ee46616078606a5d0127475668c5a2a8e451acc9b4

Observation 6cc4ecbe-6206-473b-ac57-528640a35ad9 · outbound

This paper cites Leveraging LLMs for Hypothetical Deduction in Logical Inference: A Neuro-Symbolic Approach.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Leveraging LLMs for Hypothetical Deduction in Logical Inference: A Neuro-Symbolic Approach

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.955346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.955346Z digest=sha256:9474858032352d3867e63bf02d9cbc8e9b1e83f7dfc497347a5d4c3f6f7c2181

Observation 3eb7d61f-4e0e-4ed5-ae7c-019584a9d4df · outbound

This paper cites Is This the Subspace You Are Looking for? An Interpretability Illusion for Subspace Activation Patching.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Is This the Subspace You Are Looking for? An Interpretability Illusion for Subspace Activation Patching

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.958974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.958974Z digest=sha256:512c342998a8d5e9c95b2dea2605b62e8701524903cf2a4ff170dbd3e20fc1d4

Observation 697195d0-f591-491d-b3d1-f19c8100bd5c · outbound

This paper cites Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.963135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.963135Z digest=sha256:0da55cdc154480989fe2f25e6e830c1f555a4ea142c7d26350f92d511a022f90

Observation 3ef01f3c-ed5d-45e6-944d-fe60c2148131 · outbound

This paper cites Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.966829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.966829Z digest=sha256:8c936831a395f09ab78266b5016d1653117ecf80c942c4e23275d2a0ea5d7a53

Observation 54ef891a-87aa-411a-a459-061e9adbc338 · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.970470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.970470Z digest=sha256:9136f1bedc973c19adfa05119c847d420556bab20ed2710496da8f1f116df8b5

Observation 23463316-1fb9-47ce-a4bc-ba00203caec3 · outbound

This paper cites Analyzing (In)Abilities of SAEs via Formal Languages.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Analyzing (In)Abilities of SAEs via Formal Languages

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-15T18:21:13.222900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:12.974022Z digest=sha256:ee5da521fd838bb33936874f0b8e7f01822fe768e5c29d4f126de435f5ed0d40

Observation 0035dd49-a209-46dc-8421-0ac66948edcc · outbound

This paper cites Progress measures for grokking via mechanistic interpretability.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Progress measures for grokking via mechanistic interpretability

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.977751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.977751Z digest=sha256:e7fa30b5b122236037551d8b8c6d247ddd429a33e2b5c4f18743def63fb3528c

Observation c52bc77d-6ced-4579-a42f-72044479ff24 · outbound

This paper cites Making Reasoning Matter: Measuring and Improving Faithfulness of Chain-of-Thought Reasoning.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Making Reasoning Matter: Measuring and Improving Faithfulness of Chain-of-Thought Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.982072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.982072Z digest=sha256:95f888b9baf527a731f2496fa594500cf165d33f390ff7e5461e2ab425991e17

Observation ba748365-a235-45b5-804d-54fc01958b25 · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:12.985706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:12.985706Z digest=sha256:2c5e4dd67dc50ad18582b6d2b4419faeeced2f21fc9ae275c6907185e66a4bb6

Observation bb4680d5-af6c-4c09-95b2-3720a02b48df · outbound

This paper cites an unresolved cited work.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:21:13.631586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:12.989446Z digest=sha256:d7b37f000ce5a0eb8cf967bf42b3fce8b9284196126cd64564b94bc752e96d95

Observation e031bf3b-b3fc-4283-a6ba-816988a247eb · outbound

This paper cites The Probabilities Also Matter: A More Faithful Metric for Faithfulness of Free-Text Explanations in Large Language Models.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding The Probabilities Also Matter: A More Faithful Metric for Faithfulness of Free-Text Explanations in Large Language Models

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T18:21:13.124489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:12.994249Z digest=sha256:0a478a1a272bc5c40e113918c0f87db86cd22b9c49c849f228015417fd179c7c

Observation 3fe42dd0-8a06-4ff8-a738-a41e1a05585c · outbound

This paper cites Probing Language Models on Their Knowledge Source.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Probing Language Models on Their Knowledge Source

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-15T18:21:13.109366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:12.998112Z digest=sha256:c8633225f5622306228d901069d8023f023c98ca469b3c0ca58eb5e6aab4b020

Observation 38e2da06-e909-45a2-ae9a-51d3dd12b152 · outbound

This paper cites V.; Zhou, D.; et al.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding V.; Zhou, D.; et al

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:13.002452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:13.002452Z digest=sha256:69a40ae544943044f65abd9dabb70a0ffa28fdf56e47373d0fa7e90248612475

Observation 246d8e9a-f30e-48ba-a294-583f0470e090 · outbound

This paper cites A Reply to Makelov et al. (2023)'s "Interpretability Illusion" Arguments.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding A Reply to Makelov et al. (2023)'s "Interpretability Illusion" Arguments

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:13.006013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:13.006013Z digest=sha256:5834284f229c88a8fde56c2714ce0205b62af9a145ac515c6ed8011ff0467864

Observation 77d34e41-ec87-4778-9277-ae9e8c4fa7c1 · outbound

This paper cites Faithful Logical Reasoning via Symbolic Chain-of-Thought.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Faithful Logical Reasoning via Symbolic Chain-of-Thought

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:13.009786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:13.009786Z digest=sha256:190d537e5a089f1e2c5b3a2341e2945b75e312a8e76f8f50835e64d2831ad9d3

Observation c936717b-1635-4f66-9174-7e9e582abf1a · outbound

This paper cites Dissociation of Faithful and Unfaithful Reasoning in LLMs.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Dissociation of Faithful and Unfaithful Reasoning in LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:13.013526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:13.013526Z digest=sha256:1a2ee0496fb0620d1facb890374c8e9bac950771d0a7c30bfdc11762b332b81a

Observation 49c888dd-6035-4b3b-87a5-a271be3a721f · outbound

This paper cites Towards Faithful Natural Language Explanations: A Study Using Activation Patching in Large Language Models.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Towards Faithful Natural Language Explanations: A Study Using Activation Patching in Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:13.017376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:13.017376Z digest=sha256:f1d9d14aec76adab05643e6c672102f6b7b6a6e33d9eb198d8d7a53aac7a4152

Observation 65cb181a-6b30-493e-b970-ff4b516736f0 · outbound

This paper cites Towards Best Practices of Activation Patching in Language Models: Metrics and Methods.

How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding Towards Best Practices of Activation Patching in Language Models: Metrics and Methods

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:13.021612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:13.021612Z digest=sha256:42be023525b0b46a862dffb2530a5dfc6e1039dee9615559f2d5840b5da813c4

Pith citing papers

Observation d3757901-b926-4635-8ef6-a6577f73d58e · inbound

The Tell-Tale Norm: $\ell_2$ Magnitude as a Signal for Reasoning Dynamics in Large Language Models cites this paper.

The Tell-Tale Norm: $\ell_2$ Magnitude as a Signal for Reasoning Dynamics in Large Language Models How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:26:56.910598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-28T02:07:49.501480Z digest=sha256:db3e81bee86f1235e12c79ed7645499948bbcdfb6c1a2ae3c5d909a273e9a831

Observation bb601e90-db41-492e-a2db-e7cb53e512f7 · inbound

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces cites this paper.

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:31:21.457839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T22:22:52.690010Z digest=sha256:50804633198954bb9efd594f4c63e11ac0f2d9ba41af67d82dae7d987ed161de

Observation 33f4f543-12ee-479c-82d3-1435710c04de · inbound

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders cites this paper.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.436202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.436202Z digest=sha256:a3eafe10aee9dd822d305b2376b6dfdd067081713f9f79bca71bfc587bb2a1d8