Pith. sign in

Paper Citation Record · LEDGER

Steering Large Language Model Activations in Sparse Spaces

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2503.00177.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.00177 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:21:30.044383Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7fed5dda-db73-4f3a-b954-95000fee73a0 · inbound

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models cites this paper.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.396550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.396550Z digest=sha256:536f195a6bf933946017045a5468c1cfa025c4b6a9713fdd076cf2cab5e040ab

Observation a9a7c9c2-6377-43b2-ae6c-c27d0133efa0 · inbound

Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering cites this paper.

Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering Steering Large Language Model Activations in Sparse Spaces

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:37:15.808881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-19T11:35:01.836096Z digest=sha256:dbbd43550f45aafe64b5a29165f9af6b6fe126de63763691f2195ddbf5dafb90

Observation c2c3fa40-a74e-45e5-a30a-0642381fecf1 · inbound

Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering cites this paper.

Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering Steering Large Language Model Activations in Sparse Spaces

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:53:52.555270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:53:52.555270Z digest=sha256:64a857a7bbca0932f74946ff2e641630175ffbecb17434e9ddb6ae949db54da9

Observation da21b129-774f-4fb0-ba65-753705cbb947 · inbound

Resa: Transparent Reasoning Models via SAEs cites this paper.

Resa: Transparent Reasoning Models via SAEs Steering Large Language Model Activations in Sparse Spaces

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:45:42.688318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:45:42.688318Z digest=sha256:86c5abc69ca0ecc9b354a48db8a94745e81cb9c149443237651c5fb21da6239a

Observation 6ff86b81-6e43-4bf1-9a83-10780ba5bdf7 · inbound

Quantifying Conversation Drift in MCP via Latent Polytope cites this paper.

Quantifying Conversation Drift in MCP via Latent Polytope Steering Large Language Model Activations in Sparse Spaces

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T22:51:28.064772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:51:28.064772Z digest=sha256:b03312cf96a7486d33caa5c88fde5432f2f92ecb107b96563982fcbf2dfb6f11

Observation 1c478744-4bce-4874-a330-e8481626d9f6 · inbound

Can LLMs Lie? Investigation beyond Hallucination cites this paper.

Can LLMs Lie? Investigation beyond Hallucination Steering Large Language Model Activations in Sparse Spaces

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T10:55:31.191794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:55:31.191794Z digest=sha256:4a9563e65773833739621a52323393ce3e3a8597eaafdd3200c2f56a733f80f6

Observation 012dc5a2-03c8-46ba-a6f2-0072f61787d7 · inbound

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models cites this paper.

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:45:40.572737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-21T21:44:36.351517Z digest=sha256:980f7dad1c64659773431ea45c4e96cb889a78eec7d535c0d4f72d295b56036d

Observation f643abdf-37bf-4643-92a3-e221791edef7 · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:40:54.808987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:c5857982e44aa61d0ebf6e235bcff9a1e03cf9b5f34f344f1c3d2b3ef2d2bdf6

Observation 1a677b74-7079-44f6-a909-3ad05bc53d35 · inbound

Sparse Autoencoders are Capable LLM Jailbreak Mitigators cites this paper.

Sparse Autoencoders are Capable LLM Jailbreak Mitigators Steering Large Language Model Activations in Sparse Spaces

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T23:52:40.237099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:52:40.237099Z digest=sha256:e8b4045ebe33327b89ecd53ec0de9bc8a44d406d34394a2b4b1be6024cc15d8c

Observation f56f0c2d-1217-4559-99f8-9b2a41b49403 · inbound

A Systematic Study of Training-Free Methods for Trustworthy Large Language Models cites this paper.

A Systematic Study of Training-Free Methods for Trustworthy Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:22:37.358896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T08:19:42.671690Z digest=sha256:219f7266a511d3ecb162b6ccbf08d380f6f3a79b534bac7afb5a7acca6fc3713

Observation 8a0bf699-6824-4266-8ef0-b48d1788f6d5 · inbound

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI cites this paper.

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI Steering Large Language Model Activations in Sparse Spaces

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:24.206288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T00:58:59.587622Z digest=sha256:e8af69175351f43a37aa06fe8f183df8467e3fba6a54d396e4ea1e8221953e09

Observation dbd43405-b993-4fcd-a586-968b2d8deedb · inbound

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI cites this paper.

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI Steering Large Language Model Activations in Sparse Spaces

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:25:45.211495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T23:18:57.280524Z digest=sha256:e37664c7469fe1924e525fb3bf97220bb70d8f0aa0b6dbb17ed81a14802e2d2f

Observation faa2a36e-cccb-4985-b630-8bb727c0e3b9 · inbound

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models cites this paper.

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:42:21.056231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-13T05:38:31.094834Z digest=sha256:93682ffeae386fb8979e3a664c2ea6137d2c5e76c988d364f152996c8c623cec

Observation e2d545c9-8b9c-48e3-bff0-18fe56995128 · inbound

Steered Generation via Gradient-Based Optimization on Sparse Query Features cites this paper.

Steered Generation via Gradient-Based Optimization on Sparse Query Features Steering Large Language Model Activations in Sparse Spaces

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:36:40.256994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T05:31:29.510639Z digest=sha256:7d2433d6ec65e59702279fe346478544628102ba95b2b2d96b105b0cbcc2bbbe

Observation 095ecfb9-1be0-4759-ace8-4ee5190672b3 · inbound

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection cites this paper.

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection Steering Large Language Model Activations in Sparse Spaces

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:03:29.936584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T13:53:27.306664Z digest=sha256:e9803b4238ba67bfb61cc7c0a17f14bb5fb8c80ad5e0b06150930749cbcb40fe

Observation e73107f4-1d3e-4921-89f9-3be0aa915c1f · inbound

Latent Reward Steering: An Adaptive Inference-Time Framework that Implicitly Promotes Cognitive Behaviors in Reasoning LLMs cites this paper.

Latent Reward Steering: An Adaptive Inference-Time Framework that Implicitly Promotes Cognitive Behaviors in Reasoning LLMs Steering Large Language Model Activations in Sparse Spaces

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:52:35.549158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-28T18:49:56.917505Z digest=sha256:2304649a9f3a4fabe688fec0abdc3af8f1912b0385f8cd92a716cdd4d5047909

Observation 17e05bde-bc72-4c61-a1c7-78c18fbbd42c · inbound

Perplexity Can Miss SAE Feature Damage Under Quantization cites this paper.

Perplexity Can Miss SAE Feature Damage Under Quantization Steering Large Language Model Activations in Sparse Spaces

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:36:25.800030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T11:41:18.460538Z digest=sha256:366879d5f99ef92664c22b744239048bfa531c038c3d5b5c1078f30e53e51476

Observation 5de0e7ba-4a1d-41f6-9e97-026363e213ad · inbound

LLM Self-Recognition: Steering and Retrieving Activation Signatures cites this paper.

LLM Self-Recognition: Steering and Retrieving Activation Signatures Steering Large Language Model Activations in Sparse Spaces

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-28T01:41:29.197320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-28T01:41:03.518190Z digest=sha256:68001f6d12441bed69ca61e96ad8c9fb88b06ee41c2f02eee4992c42f14530d0

Observation 046e59be-ba9f-46fb-b961-813ea7e5fc52 · inbound

Tapered Language Models cites this paper.

Tapered Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:09:44.002989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T09:11:20.341634Z digest=sha256:32aa5a313be07dacd775fd90338319b28f1d43c582b75a00e3748dead3135126

Observation 10bcbc85-fb11-43f5-9cd0-e858750647b2 · inbound

SparSEEty: Extracting Tokens from Sparsity-Exploiting LLM Serving Systems via Deterministic Side Channels cites this paper.

SparSEEty: Extracting Tokens from Sparsity-Exploiting LLM Serving Systems via Deterministic Side Channels Steering Large Language Model Activations in Sparse Spaces

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T04:21:30.044383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:21:30.044383Z digest=sha256:970f1a0b12e8ddb1e8a4f33cd04351170feda4b0e24a353430ad4aa47309cf28