Pith. sign in

Paper Citation Record · LEDGER

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

As of 19 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 4 inbound Pith citation observations for arXiv:2505.14071.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14071 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:31.291568Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:45:42.708353Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T23:31:16.269375Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved7
  • parse uncertain4
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fe765d21-e912-4ff4-8f0f-2ef3bce18cb5 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:33.741602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:29.989830Z digest=sha256:0c1f82b862b38d3190c1b42df5597fc4e82e5d9355dee5421222b4459363ed7e

Observation 196a1e9a-71b0-4ac1-b608-7f780741b43c · outbound

This paper cites 2 OLMo 2 Furious.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models 2 OLMo 2 Furious

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.875453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:29.875453Z digest=sha256:96fa8ed554dfc1941a9ff9ff38ae5ca5ad42bb57048ad796a0ce7e452fad8469

Observation 46a726e3-8d5e-4299-8b80-e784596effdd · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.003136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:30.578966Z digest=sha256:2dc6fc46100a2cf449c38693e8ab17148a8f2cf1313280bdc735c7816c41aca2

Observation 59fc1de5-1f45-4072-b6a3-10035225c09d · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:33.544136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:30.148065Z digest=sha256:2ddfedcb145d4134481ffd6471ad9a407ed8f88493722bdcdecf3c1db3979916

Observation e70cef07-cd93-415f-be82-c2b78c35f96a · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.352863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:30.236996Z digest=sha256:0e4b4521cc8c5d362b43bcdb577382bfe8a00eb44abc85f10da6ff4645fae0eb

Observation a8625e75-cbaf-4db0-a92e-f31f3a07b8cf · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 6

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.191352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:30.410383Z digest=sha256:0bff7d436facc512fdd9b817f6bdac1d57bd3939933ef7e5b42cda078fb6ff36

Observation f0c5ad1c-ce68-4f24-8fc9-2e406c1fe6c2 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 8

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:32.855076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:30.765852Z digest=sha256:482718b0c23fab52afa1e83b9f90c8e2113792f545763fa5a0e234f0ce75d377

Observation 57c109e9-1f8c-4f8e-a646-79b430dfadd6 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.704442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:30.993716Z digest=sha256:2e5ed8c0ef5d7595de1f5fdf88b0b4f7a10b0a42283429cd7a4438da83f2d7bb

Observation 4f9dcd89-fe51-427e-a717-c1ad89b1420d · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.453474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:31.082517Z digest=sha256:9579cb40ff3ce5eb7f2cf797b6ffd05e2a71f49914fb5ebe62037576049f5d2b

Observation 90024b35-b02d-4768-8692-0ea109b64263 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.136059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:31.188797Z digest=sha256:9d7224be05135fc83d43f7b519a6c349a234b17f0efa607a4d2f5501add9ed9f

Observation 8ab9e03c-c6f7-4819-8d0d-b6877b6567b3 · outbound

This paper cites INSTRUCTION:.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models INSTRUCTION:

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:42:31.726510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:42:31.291568Z digest=sha256:e1aa3294afea7e3b51cd4ce783dcf1147d85f58b9f5af7fbf3d120c6dec58375

Observation 1853cc4c-7d87-4f40-9ec0-ad89cb879819 · outbound

This paper cites Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.746277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:29.746277Z digest=sha256:c5e5c2001ee00b0af68bafa622279d8a64f0c2684f77c9449edb042ec4e313a8

Pith citing papers

Observation e4e68646-ed60-4f23-a0a8-2c927f4350f9 · inbound

Resa: Transparent Reasoning Models via SAEs cites this paper.

Resa: Transparent Reasoning Models via SAEs Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:45:42.708353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:45:42.708353Z digest=sha256:ea763565c8d82fc8ef8ce737171575a96fb3f2784774a7a7f72f335fdb996b5c

Observation 2b2fd7ac-43cd-4335-9b36-25aaee936410 · inbound

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models cites this paper.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.425121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.425121Z digest=sha256:e1d8ee274c5a001613640d664707bbfcef9f41ca2ccf2a592d7a4cf8731926ad

Observation 6092d87a-c1bf-4fb3-aa3c-e8ea842bf7e5 · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.272701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:e52e22c2f48697152bdf1ebc34a09ae8ad7f2cc66ab9f249d1b4ab8444b742b2

Observation 80f05176-4831-4e3c-a3bd-126c5f59d985 · inbound

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering cites this paper.

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T16:36:39.333944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:36:39.333944Z digest=sha256:96c361f376567a3e955c528acd80d77a6ed74de73454f413898427c217780942