Pith. sign in

Paper Citation Record · LEDGER

Steering Large Language Model Activations in Sparse Spaces

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2503.00177.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.00177 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:21:30.044383Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7fed5dda-db73-4f3a-b954-95000fee73a0 · inbound

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models cites this paper.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.396550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.396550Z digest=sha256:40f5189dc023b5f6b7c95b8c072c28900f4887c9bf5a40a2ea31018d05288625

Observation a9a7c9c2-6377-43b2-ae6c-c27d0133efa0 · inbound

Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering cites this paper.

Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering Steering Large Language Model Activations in Sparse Spaces

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:37:15.808881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T11:35:01.836096Z digest=sha256:ebaded824274d71e76857c6fab8b3536581ea71a87ee075eff445a9ccebc05a6

Observation c2c3fa40-a74e-45e5-a30a-0642381fecf1 · inbound

Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering cites this paper.

Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering Steering Large Language Model Activations in Sparse Spaces

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:53:52.555270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:53:52.555270Z digest=sha256:04c5135f23c77470a02f39213bddd1bf4a3b681c4bf2d2e50b0c83fb661b1cc9

Observation da21b129-774f-4fb0-ba65-753705cbb947 · inbound

Resa: Transparent Reasoning Models via SAEs cites this paper.

Resa: Transparent Reasoning Models via SAEs Steering Large Language Model Activations in Sparse Spaces

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:45:42.688318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:45:42.688318Z digest=sha256:461dfa22941dcd68dac736d33e6b7e9f64b8c5a0ad883f5e0819297c509a0273

Observation 6ff86b81-6e43-4bf1-9a83-10780ba5bdf7 · inbound

Quantifying Conversation Drift in MCP via Latent Polytope cites this paper.

Quantifying Conversation Drift in MCP via Latent Polytope Steering Large Language Model Activations in Sparse Spaces

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T22:51:28.064772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:51:28.064772Z digest=sha256:cb3d2e50105633e4e3323f03d2b28a76e2a6cb9f6026276fd542108660458c54

Observation 1c478744-4bce-4874-a330-e8481626d9f6 · inbound

Can LLMs Lie? Investigation beyond Hallucination cites this paper.

Can LLMs Lie? Investigation beyond Hallucination Steering Large Language Model Activations in Sparse Spaces

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T10:55:31.191794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:55:31.191794Z digest=sha256:21787c29731074a9f436c89b4735d49b41259d4c912200cd1f6f2cd1b4eec69f

Observation 012dc5a2-03c8-46ba-a6f2-0072f61787d7 · inbound

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models cites this paper.

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:45:40.572737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T21:44:36.351517Z digest=sha256:8d63d33ce1d004ea05d518b45fe47d8f2a2846765e6c6c085b86567ad2c8c2bd

Observation f643abdf-37bf-4643-92a3-e221791edef7 · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:40:54.808987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:2454437a814ad1dd04ee68dbbce4bd7b164177155014360dd9a0bc7712c309c8

Observation 1a677b74-7079-44f6-a909-3ad05bc53d35 · inbound

Sparse Autoencoders are Capable LLM Jailbreak Mitigators cites this paper.

Sparse Autoencoders are Capable LLM Jailbreak Mitigators Steering Large Language Model Activations in Sparse Spaces

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T23:52:40.237099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:52:40.237099Z digest=sha256:80bd8dc38eb98afddd6e5aef1e852d69d51a659048f87062e4a41ce96ff1bb1d

Observation f56f0c2d-1217-4559-99f8-9b2a41b49403 · inbound

A Systematic Study of Training-Free Methods for Trustworthy Large Language Models cites this paper.

A Systematic Study of Training-Free Methods for Trustworthy Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:22:37.358896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T08:19:42.671690Z digest=sha256:d4568cea2781d306013effe7ff4001c904c50d29cc678b9e03003c737e9c0921

Observation 8a0bf699-6824-4266-8ef0-b48d1788f6d5 · inbound

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI cites this paper.

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI Steering Large Language Model Activations in Sparse Spaces

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:24.206288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T00:58:59.587622Z digest=sha256:dd8b18b1678be60a0bd9358a3e2306fec98d4fc0af3608cb8e1f655d5f75bc5e

Observation dbd43405-b993-4fcd-a586-968b2d8deedb · inbound

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI cites this paper.

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI Steering Large Language Model Activations in Sparse Spaces

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:25:45.211495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T23:18:57.280524Z digest=sha256:97f2316d2c29db6ee44ee7e72054634216df398b99eafd009d69318b0a9ef530

Observation faa2a36e-cccb-4985-b630-8bb727c0e3b9 · inbound

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models cites this paper.

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:42:21.056231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T05:38:31.094834Z digest=sha256:22e37190dc864dc5cba3f8a53f18264b92ecf1552b93ac67804240012bc99b0a

Observation e2d545c9-8b9c-48e3-bff0-18fe56995128 · inbound

Steered Generation via Gradient-Based Optimization on Sparse Query Features cites this paper.

Steered Generation via Gradient-Based Optimization on Sparse Query Features Steering Large Language Model Activations in Sparse Spaces

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:36:40.256994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T05:31:29.510639Z digest=sha256:589705d095590ca0651a232e24614398cd4238db59b4d75b1c841daa8e0fc545

Observation 095ecfb9-1be0-4759-ace8-4ee5190672b3 · inbound

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection cites this paper.

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection Steering Large Language Model Activations in Sparse Spaces

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:03:29.936584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T13:53:27.306664Z digest=sha256:8acb076e12d9c49938259df5a299a6fdf1ccd71faefce49ada72b38c86c13926

Observation e73107f4-1d3e-4921-89f9-3be0aa915c1f · inbound

Latent Reward Steering: An Adaptive Inference-Time Framework that Implicitly Promotes Cognitive Behaviors in Reasoning LLMs cites this paper.

Latent Reward Steering: An Adaptive Inference-Time Framework that Implicitly Promotes Cognitive Behaviors in Reasoning LLMs Steering Large Language Model Activations in Sparse Spaces

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:52:35.549158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T18:49:56.917505Z digest=sha256:8e0603a938547fe148a9c8d12c5331dcce0af4ff9d40b21b866d787a10a326f0

Observation 17e05bde-bc72-4c61-a1c7-78c18fbbd42c · inbound

Perplexity Can Miss SAE Feature Damage Under Quantization cites this paper.

Perplexity Can Miss SAE Feature Damage Under Quantization Steering Large Language Model Activations in Sparse Spaces

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:36:25.800030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T11:41:18.460538Z digest=sha256:b274f80f5e81593f29f841678cf220638df41be0bd493dd131cc4710d8899312

Observation 5de0e7ba-4a1d-41f6-9e97-026363e213ad · inbound

LLM Self-Recognition: Steering and Retrieving Activation Signatures cites this paper.

LLM Self-Recognition: Steering and Retrieving Activation Signatures Steering Large Language Model Activations in Sparse Spaces

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-28T01:41:29.197320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T01:41:03.518190Z digest=sha256:a1564f83fcc7cd2920d2a24e40f5ec65f19659d74a89de175643b4fe983571c6

Observation 046e59be-ba9f-46fb-b961-813ea7e5fc52 · inbound

Tapered Language Models cites this paper.

Tapered Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:09:44.002989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T09:11:20.341634Z digest=sha256:9eaab254f2e6f05653d0061b4de05de4ede56fa199685019e323d70ee52c486f

Observation 10bcbc85-fb11-43f5-9cd0-e858750647b2 · inbound

SparSEEty: Extracting Tokens from Sparsity-Exploiting LLM Serving Systems via Deterministic Side Channels cites this paper.

SparSEEty: Extracting Tokens from Sparsity-Exploiting LLM Serving Systems via Deterministic Side Channels Steering Large Language Model Activations in Sparse Spaces

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T04:21:30.044383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:21:30.044383Z digest=sha256:535accd710b872597f336c1ce9d8adb503f20ce84f9689058b2b60974d672b3d