Pith. sign in

Paper Citation Record · LEDGER

Improving Activation Steering in Language Models with Mean-Centring

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2312.03813.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.03813 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:24:40.169417Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:27:56.553010Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 64ab4540-b665-49d7-9c5a-dbf3fa146168 · inbound

Refusal in Language Models Is Mediated by a Single Direction cites this paper.

Refusal in Language Models Is Mediated by a Single Direction Improving Activation Steering in Language Models with Mean-Centring

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:47:56.002005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T10:47:55.934081Z digest=sha256:659d368d329f12186a19465add5003245350a218bce726efbeb35a2169065c09

Observation 2e96da92-1044-4c69-b58c-6500163e9fcd · inbound

Denoising Concept Vectors with Sparse Autoencoders for Improved Language Model Steering cites this paper.

Denoising Concept Vectors with Sparse Autoencoders for Improved Language Model Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:30:21.140538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:30:21.140538Z digest=sha256:d5d12cff389622a5825378e0775028fd9342e239587c0e86d74cae322e0b324c

Observation a89742db-a8b8-4ee6-b2c5-cba54d8426c5 · inbound

Linear Spatial World Models Emerge in Large Language Models cites this paper.

Linear Spatial World Models Emerge in Large Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:27.899636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:27.899636Z digest=sha256:17fccc6bb88ccb53422f034608e941f4e1da357b0fa4253b98a0feb0ddd1925b

Observation ccda8ba5-f1c8-4351-b338-06c32a8b5839 · inbound

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models cites this paper.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.231617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.231617Z digest=sha256:e9a2616a10ecace99ba1c5c9689ae4e8c9c2bd753b59f9cbe81d149fe7727179

Observation cf554321-a7f5-4f89-8968-71a2fe8805f5 · inbound

Probing the Robustness of Large Language Models Safety to Latent Perturbations cites this paper.

Probing the Robustness of Large Language Models Safety to Latent Perturbations Improving Activation Steering in Language Models with Mean-Centring

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.924023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:14.924023Z digest=sha256:a7bdf1e5c8009ef419e5869cfeca9965c2918f82085b6aef39992c63705f73d6

Observation 8258fe95-752b-4992-b837-080c7a43eda1 · inbound

Balancing Stylization and Truth via Disentangled Representation Steering cites this paper.

Balancing Stylization and Truth via Disentangled Representation Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T23:59:14.682573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:59:14.682573Z digest=sha256:ae10025fc043bcd077a800caef305543bd444dc8ece7b151ba1af40a9488788d

Observation eadd3a0b-2ed8-4f62-9c0e-98dafa0b00e2 · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Improving Activation Steering in Language Models with Mean-Centring

Reference 162

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:05.005161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:05.005161Z digest=sha256:30c628d0529501a4039a7301490d0c1621052280a40f0dca287cbd94003545f3

Observation 1564b451-a3a0-4ac9-80cf-75dad5e35538 · inbound

Multimodal Function Vectors for Visual Relations cites this paper.

Multimodal Function Vectors for Visual Relations Improving Activation Steering in Language Models with Mean-Centring

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T12:44:41.880589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:44:41.880589Z digest=sha256:99d8bd817e053a36aac5cbbb5bd052ade00a9c3b4253ebd8d9b5899fb87d40ba

Observation 3eda15f8-7f19-4ffc-a0d3-6c46e23d0ab8 · inbound

Differential syntactic and semantic encoding in LLMs cites this paper.

Differential syntactic and semantic encoding in LLMs Improving Activation Steering in Language Models with Mean-Centring

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T12:03:27.087762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:03:27.087762Z digest=sha256:3affa1271d1bb55e70dd88525d47127ec6efcb2a1aee4e9a404a813e7d1b9d06

Observation e956099e-cb86-43b2-aa10-e979664e96df · inbound

The Cylindrical Representation Hypothesis for Language Model Steering cites this paper.

The Cylindrical Representation Hypothesis for Language Model Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:15:09.256284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T00:10:29.122196Z digest=sha256:3a79e77a88dfb8ab6af81bb2ff7151c64371ecf23427ea4913e83b7e7717f16a

Observation b5fa2390-e3a1-4908-8ad6-a6fb1ddaa531 · inbound

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes cites this paper.

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes Improving Activation Steering in Language Models with Mean-Centring

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:31:10.473529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T11:42:16.090162Z digest=sha256:522946391af957719262390316b096bc792b97e588fdc6aa4f77435ba01bca2a

Observation 59167a3b-6b18-4eae-89d0-5965bd3de790 · inbound

The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models cites this paper.

The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:11:09.048343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T10:11:08.494758Z digest=sha256:7004261a54bf82d9333b8e44b7823233d6f532a87e94ba38244f9c65fa5ba965

Observation 0592810a-e529-472f-aba0-f201b27989b0 · inbound

When is Your LLM Steerable? cites this paper.

When is Your LLM Steerable? Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.554416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T10:01:21.682285Z digest=sha256:db032218c470647191cddc7377b1455e66eaf35a02da85c777dfb182155ec7e9

Observation fdca373a-c120-4616-91e0-89efb0207b7a · inbound

Context Is King: How In-Context Specification Shapes the Geometry of Concepts cites this paper.

Context Is King: How In-Context Specification Shapes the Geometry of Concepts Improving Activation Steering in Language Models with Mean-Centring

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T15:17:59.762871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T15:17:59.762871Z digest=sha256:02687bb4b5421b86698e3f3da460c700817e3e19d5be7e694ee3cc813ccc30d9

Observation 13645654-ef1f-4b39-9d32-d190f42116db · inbound

Where Steering Signals Come From: Activation Source Selection in Activation Steering cites this paper.

Where Steering Signals Come From: Activation Source Selection in Activation Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T03:00:44.885003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:00:44.885003Z digest=sha256:ba6a57e69ea79c3747e6fbdafaec871b44098d96304a352a5a8f658f8bbfc040

Observation bb3bfebc-5642-4281-a2a0-d901d3320dfd · inbound

Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models cites this paper.

Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T00:39:50.627259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:39:50.627259Z digest=sha256:999e1b2c8b8cbb7600c6288a386421d1d03966658226592e8490a36d74a49cf6

Observation 3fddc452-18b3-4819-9178-43a5dec5c4ab · inbound

Subliminal Learning is Non-Semantic Distillation cites this paper.

Subliminal Learning is Non-Semantic Distillation Improving Activation Steering in Language Models with Mean-Centring

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T00:24:40.169417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:24:40.169417Z digest=sha256:f59fa9e8bc75b6f82bd7befe605c7db626a177dfb9ab2573ef82a985ce9f84e4