Pith. sign in

Paper Citation Record · LEDGER

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

As of 9 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 3 inbound Pith citation observations for arXiv:2506.07406.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07406 v3

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:44:34.878071Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T22:20:52.947124Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved17
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7de31c2a-3748-41af-83e5-97b57e0084c2 · outbound

This paper cites Understanding intermediate layers using linear classifier probes.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Understanding intermediate layers using linear classifier probes

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.811248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.811248Z digest=sha256:567b17f6a205005ce96e316f2702a279ec5445707b4f078ebc28b12b1abf2b10

Observation 884b5900-8b90-48f0-9d6c-44d7f2449931 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.824559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.824559Z digest=sha256:22d97b73ae1d7617d5cc57a3ef629514965bf679bb568714f9011f627e182b52

Observation e292e361-3379-4728-9338-230cabb9bbb4 · outbound

This paper cites In-Context Learning Creates Task Vectors.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models In-Context Learning Creates Task Vectors

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.837497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.837497Z digest=sha256:d14bb95c1f05f53580ee6422f63a7113cfb95243f325bb6b32be63954f75a414

Observation 2e71ec9b-eeab-4355-9b5f-4fca8bccb89f · outbound

This paper cites RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations

Reference 10

Resolution
malformed identifier
no resolver link, observed 2026-08-07T05:44:34.840406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.840406Z digest=sha256:70f70cae444dd9bc37f263eb312ee60ededec840a258e07545160a158dd7cea6

Observation 99ca24ad-20bb-4290-94dc-230070a0d686 · outbound

This paper cites Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.843022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.843022Z digest=sha256:15c7140e51d9f89625e4c842777155218ac0acd1566375c2986ba9249b73d1ba

Observation 96bf3527-f6c5-476c-89b8-a478013b0911 · outbound

This paper cites Understanding deep image representations by inverting them.2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Understanding deep image representations by inverting them.2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.102039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:44:34.846181Z digest=sha256:264b371b732545e73beb8c5dd34b42fbfd0e1c91ece29229edc1e62578178a14

Observation 7465a8b7-834b-4ccf-a356-ba4da15f6817 · outbound

This paper cites Multifaceted Feature Visualization: Uncovering the Different Types of Features Learned By Each Neuron in Deep Neural Networks.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Multifaceted Feature Visualization: Uncovering the Different Types of Features Learned By Each Neuron in Deep Neural Networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.851804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.851804Z digest=sha256:979b3e1d76f7a10d7ec8acdc5324d58d027326d809cbede54b044149dfee45fd

Observation 3f4976f0-6dab-4d11-b9ec-2840a7e3155b · outbound

This paper cites The Linear Representation Hypothesis and the Geometry of Large Language Models.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models The Linear Representation Hypothesis and the Geometry of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.857742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.857742Z digest=sha256:59b65ee4f7cdd9f1b04c43fac5a0fc313cacfc96b38f47b5ef4b98a93fb7631b

Observation 8a525b8f-a385-4852-9f8b-4e0c8b325bc5 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Gemma 2: Improving Open Language Models at a Practical Size

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.862975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.862975Z digest=sha256:ad42594794e19fb980286aca4aad364b397265e25af47643a83505b97e9e5837

Observation 5f3ca20a-8fa8-4ff4-817c-02d1df676dd5 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.865749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.865749Z digest=sha256:bac2d36d28f718f5f564e012b2da53764101bc24abffe22281dcbd1d8082f1c9

Observation f2b18c6f-035b-4e61-bfb7-531a1aa1829f · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.868186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.868186Z digest=sha256:af2711ded4ebe04d4cc4d14d4e0f57e4d9fef05e287dac63a3268a9980ca5e17

Observation 841875b3-b3fd-4295-abc1-cd2f2616e098 · outbound

This paper cites the indirect object name in the prompt.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models the indirect object name in the prompt

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.094101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:44:34.870694Z digest=sha256:19f6d64abacd60377b2c0251ff0a544a896a39fbcbb95c9807971e021e3d01e5

Observation b2bb17f5-66e4-40da-9e8f-24620bb3cd3e · outbound

This paper cites Training is performed for 100,000 steps with a batch size of 2048 using the AdamW optimizer.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Training is performed for 100,000 steps with a batch size of 2048 using the AdamW optimizer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.078733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:44:34.875703Z digest=sha256:69e0639022184c678ed4658b4586c56c26ded84ae196adeda0e625f2a0eb7fb3

Observation 2e9583d3-d141-45cc-a038-8bc837b5cbc8 · outbound

This paper cites [City] is a city in the country of.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models [City] is a city in the country of

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.069695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:44:34.878071Z digest=sha256:ce63eab2e8d16a11663d8374aa1029e04415cf0b2bbbb6b5155737b85e500de3

Observation 767522fa-2384-4c3c-9976-8dde24a8ac93 · outbound

This paper cites Then, [B] and [A] went to the [PLACE]. [B] gave a [OBJECT] to.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Then, [B] and [A] went to the [PLACE]. [B] gave a [OBJECT] to

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.086347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:44:34.873298Z digest=sha256:f6fc44df8244bd6cc9ce93a6fdd3ee46297348a9e201dfdd42f9aa141ee57e7f

Observation 46645982-f0df-4415-955e-f13c988cb566 · outbound

This paper cites Scaling and evaluating sparse autoencoders.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Scaling and evaluating sparse autoencoders

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.827442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.827442Z digest=sha256:8e5d03f19275ba893f8e41291509cdb48fd12ca8826569f4fec1fb9400bcf3be

Observation 54d2460e-b0c6-4b0b-a2e0-f976a85526ea · outbound

This paper cites Mechanistic Interpretability for AI Safety -- A Review.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Mechanistic Interpretability for AI Safety -- A Review

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.814735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.814735Z digest=sha256:5fe21f1c119af030ab4848c1b53a2a8e6ffe7d581e9570843f931aa9c06de80a

Observation 441ed5b6-b3e9-411b-81aa-d2ea158f8d60 · outbound

This paper cites Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.849126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.849126Z digest=sha256:549d653f351bb4f60d2e8785f71ee87ddeb1c22352fb8b78e914e35da0c1574b

Observation 4cc61af8-21d7-4fe6-abc9-ff6dbf5e32d0 · outbound

This paper cites Alexander Pan, Lijie Chen, and Jacob Steinhardt.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Alexander Pan, Lijie Chen, and Jacob Steinhardt

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.854358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.854358Z digest=sha256:83355cd8234f50a6eda9317ad56b75505829f4d611a675d7e04d9fd4d18361a1

Observation 47020588-307b-4faa-a301-657d98bd9649 · outbound

This paper cites Open Problems in Mechanistic Interpretability.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Open Problems in Mechanistic Interpretability

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.860460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.860460Z digest=sha256:f9c30203cce7b2606de3c8e4faedf1415c16b3f8160deb2d0ea7046bf03e690c

Observation 989dcab7-d626-48ea-9f7b-930feae4c211 · outbound

This paper cites SelfIE: Self-Interpretation of Large Language Model Embeddings.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.821849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.821849Z digest=sha256:441ec34c9f5577684702f9c4f8b7e7edabfe159e8ec40ffb83e133e84b7d3d7e

Observation 15a5bb06-6583-43b2-9dbb-38e809088d8f · outbound

This paper cites Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.834001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.834001Z digest=sha256:8347416fab124ab2c1afcae30dedca4b76c824f1ecd48e1fb7a5a03514cf5015

Observation 905a8033-9350-41ff-9817-2c76ad7eb2e4 · outbound

This paper cites Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.110407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:44:34.818439Z digest=sha256:4a52f6b1982a002d654fe4f01ed7fa9bd5f9b6918aabe69479bc4f0f425af731

Observation 2a069249-8b0a-4bd1-96b1-37b4a6ef079c · outbound

This paper cites Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.830593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.830593Z digest=sha256:5fa931667fb8f62212d47f4bfd6e9ac47b6b2d7bb77b426b0ae1f1b2cf65035f

Pith citing papers

Observation fb51bfd6-9f2e-41b7-b344-ebd22b0de95f · inbound

Rep2Text: Decoding Full Text from a Single LLM Token Representation cites this paper.

Rep2Text: Decoding Full Text from a Single LLM Token Representation InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:08.538842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T23:15:45.828498Z digest=sha256:43669f2416d336321dfc1af5b29957f987e2a0539d3e8eec5de6f508bf6b755a

Observation 6fc72081-ca3a-4f5c-bb10-c4f8d34b83fb · inbound

Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy cites this paper.

Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T22:20:52.947124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T22:20:52.947124Z digest=sha256:e5fa50e1e3b766303989f00162366b336bac84e57d58488d100c4280a1c1a788

Observation d512d461-2c4e-43ce-bc65-3833159c71c0 · inbound

PRISM: Recovering Instruction Sets from Language Model Activations cites this paper.

PRISM: Recovering Instruction Sets from Language Model Activations InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-07-07T03:18:08.538842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T16:52:02.948457Z digest=sha256:4f326e54cbc6a8e93847c8e2b44c826c65aa5ccee41d929e26b85d4ae71dbf19