Pith. sign in

Paper Citation Record · LEDGER

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

As of 15 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 3 inbound Pith citation observations for arXiv:2506.07406.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07406 v3

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:44:34.878071Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T22:20:52.947124Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved17
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7de31c2a-3748-41af-83e5-97b57e0084c2 · outbound

This paper cites Understanding intermediate layers using linear classifier probes.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Understanding intermediate layers using linear classifier probes

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.811248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.811248Z digest=sha256:856585e4c15259bdfbd292c94eb53b38ac0239de6ecf7f4fd9bf355aa869a174

Observation 884b5900-8b90-48f0-9d6c-44d7f2449931 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.824559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.824559Z digest=sha256:f684220a5d7fa621661c37187edcce5b7ec4a4f10ae1a552fd9b9e6d3b193b0e

Observation e292e361-3379-4728-9338-230cabb9bbb4 · outbound

This paper cites In-Context Learning Creates Task Vectors.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models In-Context Learning Creates Task Vectors

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.837497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.837497Z digest=sha256:f1aa22938dfa47605ce071ac4ac7a97ab129f4cb4c5994ca46d4f6aefe5712ee

Observation 2e71ec9b-eeab-4355-9b5f-4fca8bccb89f · outbound

This paper cites RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations

Reference 10

Resolution
malformed identifier
no resolver link, observed 2026-08-07T05:44:34.840406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.840406Z digest=sha256:eed54fa61ba8b8386b8f568dc94932b47434c8dc130749c943f192b20cd75cc1

Observation 99ca24ad-20bb-4290-94dc-230070a0d686 · outbound

This paper cites Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.843022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.843022Z digest=sha256:2814448b35d0c8d9e26757525f1200f94764330bb69c5eee0e3b3d3cea2d6ed7

Observation 96bf3527-f6c5-476c-89b8-a478013b0911 · outbound

This paper cites Understanding deep image representations by inverting them.2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Understanding deep image representations by inverting them.2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.102039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:44:34.846181Z digest=sha256:3019e3a7190d932112a7fb36aeadf1c1fe4ff5a90242d5b18b0d899108015a14

Observation 7465a8b7-834b-4ccf-a356-ba4da15f6817 · outbound

This paper cites Multifaceted Feature Visualization: Uncovering the Different Types of Features Learned By Each Neuron in Deep Neural Networks.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Multifaceted Feature Visualization: Uncovering the Different Types of Features Learned By Each Neuron in Deep Neural Networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.851804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.851804Z digest=sha256:eecc0496bdd0c7ef971103573d4b5e3ae779f65c25de8591708de5303bed4610

Observation 3f4976f0-6dab-4d11-b9ec-2840a7e3155b · outbound

This paper cites The Linear Representation Hypothesis and the Geometry of Large Language Models.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models The Linear Representation Hypothesis and the Geometry of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.857742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.857742Z digest=sha256:1dc86378759d3ce50be705216b6415a31168094f6c0568345989fbf1cb4eece4

Observation 8a525b8f-a385-4852-9f8b-4e0c8b325bc5 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Gemma 2: Improving Open Language Models at a Practical Size

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.862975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.862975Z digest=sha256:ad42594794e19fb980286aca4aad364b397265e25af47643a83505b97e9e5837

Observation 5f3ca20a-8fa8-4ff4-817c-02d1df676dd5 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.865749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.865749Z digest=sha256:bac2d36d28f718f5f564e012b2da53764101bc24abffe22281dcbd1d8082f1c9

Observation f2b18c6f-035b-4e61-bfb7-531a1aa1829f · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.868186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.868186Z digest=sha256:af2711ded4ebe04d4cc4d14d4e0f57e4d9fef05e287dac63a3268a9980ca5e17

Observation 841875b3-b3fd-4295-abc1-cd2f2616e098 · outbound

This paper cites the indirect object name in the prompt.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models the indirect object name in the prompt

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.094101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:44:34.870694Z digest=sha256:039431df3cc5c018c67282179666af20aca0328610b336f52f196ee2b8344b85

Observation b2bb17f5-66e4-40da-9e8f-24620bb3cd3e · outbound

This paper cites Training is performed for 100,000 steps with a batch size of 2048 using the AdamW optimizer.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Training is performed for 100,000 steps with a batch size of 2048 using the AdamW optimizer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.078733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:44:34.875703Z digest=sha256:5e08198eeb38543739a89d67dea3ba34778ab24f1f02f8826710677efbcb463f

Observation 2e9583d3-d141-45cc-a038-8bc837b5cbc8 · outbound

This paper cites [City] is a city in the country of.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models [City] is a city in the country of

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.069695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:44:34.878071Z digest=sha256:0a9b19214c7e85dfc948d5efb130413efd114a8874fedca35079983f30e3fac6

Observation 767522fa-2384-4c3c-9976-8dde24a8ac93 · outbound

This paper cites Then, [B] and [A] went to the [PLACE]. [B] gave a [OBJECT] to.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Then, [B] and [A] went to the [PLACE]. [B] gave a [OBJECT] to

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.086347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:44:34.873298Z digest=sha256:372ed58c60ae4afe3d1bce2bd42c987cc61ee8f13ef76e03caab5bfda5cd7be4

Observation 46645982-f0df-4415-955e-f13c988cb566 · outbound

This paper cites Scaling and evaluating sparse autoencoders.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Scaling and evaluating sparse autoencoders

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.827442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.827442Z digest=sha256:8e5d03f19275ba893f8e41291509cdb48fd12ca8826569f4fec1fb9400bcf3be

Observation 54d2460e-b0c6-4b0b-a2e0-f976a85526ea · outbound

This paper cites Mechanistic Interpretability for AI Safety -- A Review.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Mechanistic Interpretability for AI Safety -- A Review

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.814735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.814735Z digest=sha256:5fe21f1c119af030ab4848c1b53a2a8e6ffe7d581e9570843f931aa9c06de80a

Observation 441ed5b6-b3e9-411b-81aa-d2ea158f8d60 · outbound

This paper cites Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.849126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.849126Z digest=sha256:7d2d6ac49d94a7e95d03743121134241469751b4d0ed9944d4ee350fa35ff446

Observation 4cc61af8-21d7-4fe6-abc9-ff6dbf5e32d0 · outbound

This paper cites Alexander Pan, Lijie Chen, and Jacob Steinhardt.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Alexander Pan, Lijie Chen, and Jacob Steinhardt

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.854358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.854358Z digest=sha256:83355cd8234f50a6eda9317ad56b75505829f4d611a675d7e04d9fd4d18361a1

Observation 47020588-307b-4faa-a301-657d98bd9649 · outbound

This paper cites Open Problems in Mechanistic Interpretability.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Open Problems in Mechanistic Interpretability

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.860460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.860460Z digest=sha256:f9c30203cce7b2606de3c8e4faedf1415c16b3f8160deb2d0ea7046bf03e690c

Observation 989dcab7-d626-48ea-9f7b-930feae4c211 · outbound

This paper cites SelfIE: Self-Interpretation of Large Language Model Embeddings.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models SelfIE: Self-Interpretation of Large Language Model Embeddings

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.821849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.821849Z digest=sha256:6b8abe80a84576a002a3356923c3f317c15c132aabb57607a466eefec1307f07

Observation 15a5bb06-6583-43b2-9dbb-38e809088d8f · outbound

This paper cites Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.834001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.834001Z digest=sha256:235640b56171d55a97fe0c41ae2bb8cabd88c015f183057e5d5a930bd0480ba5

Observation 905a8033-9350-41ff-9817-2c76ad7eb2e4 · outbound

This paper cites Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:44:35.110407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T05:44:34.818439Z digest=sha256:f331c9093b66f5ced628aadd803d6a0909b8ed6105d1918d68dfb5a58ecb13c3

Observation 2a069249-8b0a-4bd1-96b1-37b4a6ef079c · outbound

This paper cites Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space.

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:44:34.830593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:44:34.830593Z digest=sha256:d2a200a3ce4a9079b6df1395eaa277948ac1afaef027a803926bd38e6718099b

Pith citing papers

Observation fb51bfd6-9f2e-41b7-b344-ebd22b0de95f · inbound

Rep2Text: Decoding Full Text from a Single LLM Token Representation cites this paper.

Rep2Text: Decoding Full Text from a Single LLM Token Representation InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:08.538842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T23:15:45.828498Z digest=sha256:1222a672d4128ebece8bbe141b0e43dbb27f0882328958b4ad7478b3812dd13f

Observation 6fc72081-ca3a-4f5c-bb10-c4f8d34b83fb · inbound

Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy cites this paper.

Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T22:20:52.947124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T22:20:52.947124Z digest=sha256:e5fa50e1e3b766303989f00162366b336bac84e57d58488d100c4280a1c1a788

Observation d512d461-2c4e-43ce-bc65-3833159c71c0 · inbound

PRISM: Recovering Instruction Sets from Language Model Activations cites this paper.

PRISM: Recovering Instruction Sets from Language Model Activations InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-07-07T03:18:08.538842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-27T16:52:02.948457Z digest=sha256:583e11580acafe3ce911efe0ae7a83012252ce34f2ac1fd9e043ab26892f90ca