Pith. sign in

Paper Citation Record · LEDGER

Extracting Latent Steering Vectors from Pretrained Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2205.05124.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2205.05124 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T23:49:46.793197Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:27:56.567705Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c2aa1ea0-706e-4b52-a069-b786cb7efcf4 · inbound

Steering Llama 2 via Contrastive Activation Addition cites this paper.

Steering Llama 2 via Contrastive Activation Addition Extracting Latent Steering Vectors from Pretrained Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:37:21.234923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T20:37:20.408376Z digest=sha256:06bcc3d70f42065b62bade33031f0f21ad983bd9acfc81815036ad40e0046bda

Observation b29ef525-0abe-4fee-9502-e01eb1a420c9 · inbound

Probe-Free Low-Rank Activation Intervention cites this paper.

Probe-Free Low-Rank Activation Intervention Extracting Latent Steering Vectors from Pretrained Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T23:49:46.793197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T23:49:46.793197Z digest=sha256:cde850235c8e5569c556603c38ed24e672e6e6b14417941ef8d7d71bc7c0a82b

Observation 7c959b8a-0722-484a-b30c-ef1f2e56630b · inbound

Task-driven Layerwise Additive Activation Intervention cites this paper.

Task-driven Layerwise Additive Activation Intervention Extracting Latent Steering Vectors from Pretrained Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T16:46:58.489805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:46:58.489805Z digest=sha256:35fd0a59e66608bf7422640c7c171718f13095a6e396b16adedbafebee0bc29f

Observation 91217912-0b99-41be-a8ce-fef01a32ef9d · inbound

Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN cites this paper.

Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Extracting Latent Steering Vectors from Pretrained Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:17.998310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:17.998310Z digest=sha256:4d2f46fd95469445ad7c83091ed15ad2c9df2008abc721f70e3d9b6c049c4d11

Observation 774b271f-373e-430d-aa04-f20f3caf04e2 · inbound

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline cites this paper.

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline Extracting Latent Steering Vectors from Pretrained Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:58:19.396698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:58:19.396698Z digest=sha256:232d5580c2012f0ef373de403a423e0bacc082b701a666b806b90cbba6237f1b

Observation 1f35f1e4-2447-4a26-bde2-41621eebe10a · inbound

Improved Representation Steering for Language Models cites this paper.

Improved Representation Steering for Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:31.709484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:51:31.709484Z digest=sha256:da667590c213a9dc8e5679c438c6cbfa53c5ff03b7736b03fdf0c9d42a4f3821

Observation 82e4239a-a60c-4467-ab08-512ba4871739 · inbound

Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules cites this paper.

Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules Extracting Latent Steering Vectors from Pretrained Language Models

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.708567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:35:30.708567Z digest=sha256:ced8d72d702aa835956ff6b43c287d75d17a81ddfac63350826a5180aee6a2a0

Observation 7696cd5b-55a8-4088-b0f2-1a8750d958c1 · inbound

Fine-Grained Interpretation of Political Opinions in Large Language Models cites this paper.

Fine-Grained Interpretation of Political Opinions in Large Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:15.757086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:40:15.757086Z digest=sha256:bd53993187d46fe3aaa269058342fa6b5cf5ad30ac5d472dd057a402363cd5f0

Observation a6d630a0-3cb9-49e3-8836-bfb9274a7ece · inbound

Can Interpretation Predict Behavior on Unseen Data? cites this paper.

Can Interpretation Predict Behavior on Unseen Data? Extracting Latent Steering Vectors from Pretrained Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:10:30.409274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:10:30.409274Z digest=sha256:db56ac41037803415f474433bcb826b691586ff6cb3a79bf54c98c31ca5f9f69

Observation 724449c9-6fbd-4896-9ea8-2059a456fd76 · inbound

Simple Mechanistic Explanations for Out-Of-Context Reasoning cites this paper.

Simple Mechanistic Explanations for Out-Of-Context Reasoning Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:44.194532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:28:44.194532Z digest=sha256:37c2f8ba9550a11c5ca9122ddab67f3dad7c5070f7030d0c4881fc999a2ca5f5

Observation 0722ef15-5b41-414c-8a23-32de27c3fb68 · inbound

Quantifying Conversation Drift in MCP via Latent Polytope cites this paper.

Quantifying Conversation Drift in MCP via Latent Polytope Extracting Latent Steering Vectors from Pretrained Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T22:51:28.153771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:51:28.153771Z digest=sha256:e95ed1c53ded4606203b66f06c90787012036d24ff91e75cd031a1648d32cbd0

Observation 09fc1d60-2f5c-4ee1-a372-eaa35d99fcd3 · inbound

Steering Protein Language Models cites this paper.

Steering Protein Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:12.109738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:12.109738Z digest=sha256:8882123e340ae9c6cc02e17376f035ab76a2e1a81cc3f21ab93e92f01069fc3e

Observation 423175fe-4354-4288-a573-ac598925db28 · inbound

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation cites this paper.

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation Extracting Latent Steering Vectors from Pretrained Language Models

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:53:29.755378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T02:49:55.908142Z digest=sha256:dd00ac3c0a09c4eb598ef132deea67d2a077b80f91182620a9778a4e230dc0b6

Observation bc2ee5f1-a7f7-445b-b6c2-465bcad80d90 · inbound

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts cites this paper.

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts Extracting Latent Steering Vectors from Pretrained Language Models

Reference 127

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:06:07.199909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T18:48:04.046370Z digest=sha256:bb52a55d3addc9a6c85bcc750ee1a3a3978c1da9f1b376fe65aff88d6aedfb16

Observation 2bad3404-af5d-479e-bb04-d00a7430df55 · inbound

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis cites this paper.

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:10:37.052258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T18:52:01.624279Z digest=sha256:d50a5c618dc17df4875b0edf7b563326c89893846ad15bbf79677af96cd9f11c

Observation b9bb33ba-3b15-4641-8f08-991ce5cdfd33 · inbound

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis cites this paper.

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:57:40.631449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:57:40.631449Z digest=sha256:1f4f35aca47a5bbc7c2bc3649d5f6b8a142d3acb5ccc00ee73136628f5c2b4d9

Observation eb27ef93-dcfe-4a19-8f89-71fa180e83f2 · inbound

Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior cites this paper.

Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior Extracting Latent Steering Vectors from Pretrained Language Models

Reference 122

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:16:06.987805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T17:47:09.591001Z digest=sha256:df69c412ef0b20e365ffe6dabbcbf39a793b49b7da7888f6be48d4a9aa8437ee

Observation bb6281e2-a6f6-44e3-8656-bcda71a4e1e6 · inbound

DataDignity: Training Data Attribution for Large Language Models cites this paper.

DataDignity: Training Data Attribution for Large Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:26:10.297548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T11:53:19.594779Z digest=sha256:295516bcad24c41dcdf4e89517550875fec3fc0705d2b07c39e9d0897c198bef

Observation e5110b5a-2109-43c3-80dd-5407be0851d7 · inbound

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models cites this paper.

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:56:19.030279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T02:54:20.313718Z digest=sha256:20183b7627f8648db8b2395513fc73cb073af89a404c706030fce9e5a8b78923

Observation efce1c5c-78c1-400b-87d7-3b3849fa3471 · inbound

Steered Generation via Gradient-Based Optimization on Sparse Query Features cites this paper.

Steered Generation via Gradient-Based Optimization on Sparse Query Features Extracting Latent Steering Vectors from Pretrained Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:36:40.329750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T05:31:29.510639Z digest=sha256:db1f63587709fd146925360c64b98fec842e870208b7ee05d340df6cff981332

Observation a36bf0d3-956e-4a33-a1a5-6209fcb1ad45 · inbound

When is Your LLM Steerable? cites this paper.

When is Your LLM Steerable? Extracting Latent Steering Vectors from Pretrained Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.569211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T10:01:21.682285Z digest=sha256:1d9454a36cef72a5b11f24a91987a9507bc063323aca660c2d9c253cd53d5928

Observation 90f6a730-1046-4d1f-9ac6-5a9ea98ac2cd · inbound

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias cites this paper.

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias Extracting Latent Steering Vectors from Pretrained Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-14T02:33:34.084111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T02:33:34.084111Z digest=sha256:29209b597a73a28df19b58c20764f76edc3250d51294846be80a8f559c230668

Observation 71d3ad5d-1205-48ce-8f08-21638e79ba52 · inbound

Probabilistic Concept-Aware Steering for Trustworthy LLM Inference cites this paper.

Probabilistic Concept-Aware Steering for Trustworthy LLM Inference Extracting Latent Steering Vectors from Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T13:56:44.575890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:56:44.575890Z digest=sha256:623bc340ee305b717478340554fd1bb511ac52e7e6a6cd37a9df21479bf667a6