Pith. sign in

Paper Citation Record · LEDGER

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

As of 9 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 4 inbound Pith citation observations for arXiv:2505.14071.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14071 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:31.291568Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:45:42.708353Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T23:31:16.269375Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved7
  • parse uncertain4
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fe765d21-e912-4ff4-8f0f-2ef3bce18cb5 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:33.741602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:29.989830Z digest=sha256:5d744edb07c4d9a6fb8ebf66e22b215cda17b6027f7c93f440ab9b557e70e3d6

Observation 196a1e9a-71b0-4ac1-b608-7f780741b43c · outbound

This paper cites 2 OLMo 2 Furious.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models 2 OLMo 2 Furious

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.875453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:29.875453Z digest=sha256:21dfbcc403317cfee7dcfa3cb117f66f7a89997613b263a699ba60c796fb08e7

Observation 46a726e3-8d5e-4299-8b80-e784596effdd · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.003136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:30.578966Z digest=sha256:dd8eba6869e4d125f2ae3b482253da8741438d317b4d5aa2b209895591e9d565

Observation 59fc1de5-1f45-4072-b6a3-10035225c09d · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:33.544136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:30.148065Z digest=sha256:f0329c82ffb775b9fa83cd49b40c64f30932781bb85fd108354b21d1b677154b

Observation e70cef07-cd93-415f-be82-c2b78c35f96a · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.352863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:30.236996Z digest=sha256:66fb408806d0ad70a963f3b3e03c6b70b813b938829bb3bb5915e68d70a38b88

Observation a8625e75-cbaf-4db0-a92e-f31f3a07b8cf · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 6

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.191352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:30.410383Z digest=sha256:c55dc539497019f6e787bcafc787d1280d399ff5b6006a146ec81ea4ddaf0bec

Observation f0c5ad1c-ce68-4f24-8fc9-2e406c1fe6c2 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 8

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:32.855076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:30.765852Z digest=sha256:80bc9c67b0b7eee4ee2d203b39549739ceeb0f69254c69d9f035be0966364fa4

Observation 57c109e9-1f8c-4f8e-a646-79b430dfadd6 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.704442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:30.993716Z digest=sha256:f027f42b1d0ec4dbd15a909b559cd1c2d02c7d6dfae6d294a57c2ff643a4aeea

Observation 4f9dcd89-fe51-427e-a717-c1ad89b1420d · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.453474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:31.082517Z digest=sha256:4fa9a3d428e790b4794a228723dca495b5706ed2b94e7d4d90250ecad1d0db68

Observation 90024b35-b02d-4768-8692-0ea109b64263 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.136059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:31.188797Z digest=sha256:bf6b3d4aa0f81eb82dd67255cdb4529aacaaa36ace6bb4438debccf7916400da

Observation 8ab9e03c-c6f7-4819-8d0d-b6877b6567b3 · outbound

This paper cites INSTRUCTION:.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models INSTRUCTION:

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:42:31.726510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:42:31.291568Z digest=sha256:1ace2aff64e56fa8e64d8dcbdc554b5de82fed9dccfd6d150c4bb737ffdd8891

Observation 1853cc4c-7d87-4f40-9ec0-ad89cb879819 · outbound

This paper cites Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.746277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:29.746277Z digest=sha256:468e14bb739d047317fa89be501ed07b55d6828bf2862f7c8288d7d3642b3b7b

Pith citing papers

Observation e4e68646-ed60-4f23-a0a8-2c927f4350f9 · inbound

Resa: Transparent Reasoning Models via SAEs cites this paper.

Resa: Transparent Reasoning Models via SAEs Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:45:42.708353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:45:42.708353Z digest=sha256:f03bf5bd2c8ae761c175195a16ed418ca2485a4adca8c67481e2a8d1549482b3

Observation 2b2fd7ac-43cd-4335-9b36-25aaee936410 · inbound

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models cites this paper.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.425121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.425121Z digest=sha256:20260a683f35f1d0372f208580a0ecf6132218eddaca794f07deb4c4a75ef710

Observation 6092d87a-c1bf-4fb3-aa3c-e8ea842bf7e5 · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.272701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:3e5d2892c9f48ccea4642a432a39e37289c8dc523682d1950a21ca30710affc9

Observation 80f05176-4831-4e3c-a3bd-126c5f59d985 · inbound

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering cites this paper.

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T16:36:39.333944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:36:39.333944Z digest=sha256:83591b14e1bedb84179ec438751ee786d3a21446108fcdefefb17a6432baaf1e