Pith. sign in

Paper Citation Record · LEDGER

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models

As of 14 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 1 inbound Pith citation observation for arXiv:2501.15054.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.15054 v2

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T14:43:14.801525Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:59:30.298880Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T18:59:33.160545Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c1f030de-aa82-4686-8c78-12fe3ae90242 · outbound

This paper cites Neel Nanda.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Neel Nanda

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:43:15.086106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T14:43:14.715218Z digest=sha256:78173d8fa6f6a511a21ad075b93b3f28ddbff97b4cfc383a148c2c5f9e9bca20

Observation 268e257b-7345-4887-b610-8562bc4a4508 · outbound

This paper cites A review of taxonomies of explainable artificial intelligence (xai) methods.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models A review of taxonomies of explainable artificial intelligence (xai) methods

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:43:15.053253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T14:43:14.726295Z digest=sha256:3660c73b65a7beec672c9eafa94d7fc8f343716c07cf4e911adda6974f33de93

Observation 6729c272-5017-4a3e-9ebd-bec9fcd2f384 · outbound

This paper cites Mechanistic Interpretability for AI Safety -- A Review.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Mechanistic Interpretability for AI Safety -- A Review

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.729779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.729779Z digest=sha256:a3c9445e6f25b9d4d9ab234a294acc0fc2269fd297f0a33bbf2ec59620557dd3

Observation 5d45590e-84b4-42de-b444-628aa8166c6b · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.741505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.741505Z digest=sha256:8ee8875c384ad86eb264c080696e98fbdef649ee680935d18c83925ea2b7d5ee

Observation 2ac73ebc-e5f0-4645-b3e4-e0a97897b17c · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.745081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.745081Z digest=sha256:c4d5cd860fbd41a6d47a3bd30e0ade51e06a0bb5947198b9a7a21cf2f4274c67

Observation 810515a5-3cf0-4547-ab2a-799b453335be · outbound

This paper cites Localizing Model Behavior with Path Patching.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Localizing Model Behavior with Path Patching

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.748803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.748803Z digest=sha256:b820f3506f4e91c043cdd261381e73d5dd75ab5b71d24b3d5934c783e6ce8c15

Observation 81f2b69b-d54c-4439-b631-24a4305e28c3 · outbound

This paper cites com/posts/AcKRB8wDpdaN6v6ru/interpreting-gpt-the-logit-lens.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models com/posts/AcKRB8wDpdaN6v6ru/interpreting-gpt-the-logit-lens

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:43:15.041894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T14:43:14.752451Z digest=sha256:f3a8cfbe18878c0f8e63ca709d3516fd745bd5d2502f2bf1e221be631762e619

Observation fc23120f-5abd-491a-84ea-8dcc3276034f · outbound

This paper cites Still No Lie Detector for Language Models: Probing Empirical and Conceptual Roadblocks.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Still No Lie Detector for Language Models: Probing Empirical and Conceptual Roadblocks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.759786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.759786Z digest=sha256:5701f0e92cefa25850ef4a0fc1fe5c2a58e775f7fd14546645930bfa4f6d7da3

Observation c8ba1d4b-b0f9-4cae-a9da-d4d1ec3d2dab · outbound

This paper cites Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.763946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.763946Z digest=sha256:de43cd4ebbb3f64bae4c2d92523fad8df11bdf32ab54cf16c9bb88054c82de06

Observation c7759ce2-4653-44f3-960e-f793910915cc · outbound

This paper cites Emergent Abilities of Large Language Models.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Emergent Abilities of Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.768096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.768096Z digest=sha256:8d61fe2a240aef769b69bb32b7c58e95fdb6fde7749c01d9f7c9a5e53cab0065

Observation 708a0b8d-41c3-4d26-896f-e0fd0c4471ad · outbound

This paper cites Progress measures for grokking via mechanistic interpretability.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Progress measures for grokking via mechanistic interpretability

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.771573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.771573Z digest=sha256:c3ce4283f222404f909e7d8fd17053841fd032b8900c075949d99cac0c597fb1

Observation 4382c841-0db5-4bfa-9031-dfca2a9091ca · outbound

This paper cites Hidden Progress in Deep Learning: SGD Learns Parities Near the Computational Limit.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Hidden Progress in Deep Learning: SGD Learns Parities Near the Computational Limit

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.775196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.775196Z digest=sha256:9d0729c53d74f05252a80c36b1fb26a2621e0779b48040c6f22f3ada9f700d8e

Observation e8a802a3-b2a4-4638-a89a-146c3af9cfb7 · outbound

This paper cites Locating and Editing Factual Associations in GPT.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Locating and Editing Factual Associations in GPT

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.779216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.779216Z digest=sha256:eb42b2f8f3269542f8dd32c4778c28589b55ea7fad377610a5fe748c32cd791d

Observation 860676fe-f804-4626-9fbf-632dacd16e3a · outbound

This paper cites A Survey of Machine Unlearning.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models A Survey of Machine Unlearning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.782996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.782996Z digest=sha256:45f93e0887744a8dbd4fe91ffed7983d87e88fe8a59ef482de514681d539f8b7

Observation 80f97076-180f-4218-8226-8aef73db19aa · outbound

This paper cites Circumventing interpretability: How to defeat mind-readers.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Circumventing interpretability: How to defeat mind-readers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.786443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.786443Z digest=sha256:a264d52e7116fbee92a5f06e057733795840ee7eb94fd18206d424eac182fc41

Observation 8fee2dd8-deeb-431c-ac2d-e6ecb3d8d85e · outbound

This paper cites Discovering Latent Knowledge in Language Models Without Supervision.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.789971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.789971Z digest=sha256:270dd63faa1a8ebe847c35d288463106e4f61e87387ea32f80d66b2fd68e9463

Observation 599be8c9-332f-4cc9-b99f-d2ca329015c0 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Representation Engineering: A Top-Down Approach to AI Transparency

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.793410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.793410Z digest=sha256:c996002a93390a97cdab613a0d819658f5c0f574b115eeceb507c29b671a4503

Observation 85c92d42-2d8c-43e8-9bf8-053b56e0a0dc · outbound

This paper cites AI Deception: A Survey of Examples, Risks, and Potential Solutions.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.797897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.797897Z digest=sha256:939f6002ed43718d7df0e3301fc86c8e48629e485879a8387af3ca5df4665bc2

Observation d41b6da4-6911-4cac-9ea4-44a829bda27e · outbound

This paper cites An overview of 11 proposals for building safe advanced AI.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models An overview of 11 proposals for building safe advanced AI

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.801525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.801525Z digest=sha256:c254446fe47c0898637670e9c6d1df04b41d152933391e3acf8af8f1ff41c15c

Observation cefe1f41-dc8d-4d9e-adfa-8be2419a5485 · outbound

This paper cites Neel Nanda.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Neel Nanda

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:43:15.064000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T14:43:14.722546Z digest=sha256:dd9f5eccb339b144e53c2616c446f0c628cd5d32c046bd0e43a5637829a0a831

Observation 6f5ceadc-785c-4c56-9ca3-dcfbfba038f4 · outbound

This paper cites Learning Syntax Without Planting Trees: Understanding Hierarchical Generalization in Transformers.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Learning Syntax Without Planting Trees: Understanding Hierarchical Generalization in Transformers

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.756025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.756025Z digest=sha256:00d529f5c323006592311f34a647290bf8be09de819ec69be05f086d06acfaad

Observation a2552a8a-796f-4dd6-bef8-6f489b7c0018 · outbound

This paper cites Lee Sharkey, Sid Black, and beren.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Lee Sharkey, Sid Black, and beren

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:43:15.074738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T14:43:14.718935Z digest=sha256:04d3c12261bdf7019a293d06b5fab9f2a1f4d1879f103c6e3d652f688ca57007

Observation 9015422d-b7ad-4ce4-8101-07aa05a192a1 · outbound

This paper cites Understanding intermediate layers using linear classifier probes.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models Understanding intermediate layers using linear classifier probes

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.733628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.733628Z digest=sha256:344ed9263de2b834e93bb922db41bf7e2179a16e0bace60bc8473c7db3822e90

Observation d45d8511-7f54-48d9-be16-a96a4dc080de · outbound

This paper cites X-Risk Analysis for AI Research.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models X-Risk Analysis for AI Research

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.702248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.702248Z digest=sha256:56637fadc62f923d1fe499752c3906f844e0825bbe6fe711c5e6dcdc8b0a857d

Observation fa06af59-535d-40a7-a8f9-58b83639b142 · outbound

This paper cites An Overview of Catastrophic AI Risks.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models An Overview of Catastrophic AI Risks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.706872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.706872Z digest=sha256:0e865cbeefcdafce3943732b787560b35e4f1ac11dcb724e9d35d0d450e5b173

Observation d51a07f8-32d9-4b56-a357-a4e17aba6e6f · outbound

This paper cites AI Alignment: A Comprehensive Survey.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models AI Alignment: A Comprehensive Survey

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.711050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.711050Z digest=sha256:a6e4d712aa72a1cad489f86e3731e185b4be6748317478499219025d88749023

Pith citing papers

Observation 43b03796-5d94-4077-addf-bed3abaabb52 · inbound

On the Effect of Uncertainty on Layer-wise Inference Dynamics cites this paper.

On the Effect of Uncertainty on Layer-wise Inference Dynamics Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:59:33.273968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-06T18:59:30.298880Z digest=sha256:c4ce0f717a5fee73adecb58a15230ab79dbdb29c1fd254d6800c1f693d4c585a