Pith. sign in

Paper Citation Record · LEDGER

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

As of 8 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 5 inbound Pith citation observations for arXiv:2507.18043.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18043 v2

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:45:08.726791Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:15:45.641226Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact7
  • verified fuzzy17
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation b21dbeae-8672-45f1-bbd8-86efa39101db · outbound

This paper cites GPT-4 Technical Report.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.572148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.572148Z digest=sha256:92d696f173a9d90d31838a4b928adf909625bd035ceefdaa0e9c5b71d67ab55e

Observation d2a132a2-d174-4d2f-a068-fb19a78fa2e2 · outbound

This paper cites A diagnostic study of explainability techniques for text classification.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs A diagnostic study of explainability techniques for text classification

Reference 2

Resolution
verified exact
doi, observed 2026-08-06T14:45:08.801516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.575953Z digest=sha256:9263032d3a74bc252be3b28ffde3c35394efe4143043260ec9e3bb191e686054

Observation 7a2f0d49-3cf4-4906-bbe9-11933ca3da4b · outbound

This paper cites Madtp: Multimodal alignment-guided dynamic token pruning for accelerating vision-language transformer.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Madtp: Multimodal alignment-guided dynamic token pruning for accelerating vision-language transformer

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.468380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.579024Z digest=sha256:fa630d1a2dadadce565badf5be894a3e4e04677a13199d3ba3b5b991db697fc9

Observation 3201bf3e-025a-4982-93a6-c4fe254e9f0f · outbound

This paper cites Xprompt: Explaining large language model's generation via joint prompt attribution.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Xprompt: Explaining large language model's generation via joint prompt attribution

Reference 4

Resolution
verified exact
raw_fallback, observed 2026-08-06T14:45:09.267613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.582177Z digest=sha256:4b35076bf2a713bfa086fe2c11a65f884161c7b5a932870dd42969e975817eff

Observation 472e75fd-80c5-4a34-abf7-c13dfc1fc7ab · outbound

This paper cites ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.585297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.585297Z digest=sha256:bc6e98993e26abd971a6522dc353631465f8e52f48e3e7aa6ac8777dc490ef43

Observation f7946dc7-3893-4d32-9ef0-0b864639c556 · outbound

This paper cites Quantifying and Mitigating Unimodal Biases in Multimodal Large Language Models: A Causal Perspective.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Quantifying and Mitigating Unimodal Biases in Multimodal Large Language Models: A Causal Perspective

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:45:09.167867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.588593Z digest=sha256:59bf5de2e4be32c5f29cb4922f52d507a3033be602d3fe9d1eb308f65b7f0c1c

Observation 6a213b3b-a79b-4099-a489-265448e343e6 · outbound

This paper cites Covert, Scott Lundberg, and Su-In Lee.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Covert, Scott Lundberg, and Su-In Lee

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.459771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.591990Z digest=sha256:63ab6833ef4f39ea7d7263af0a1db4e36ffb5202fefb10e50c86feef29e331dd

Observation 7d61dc24-08f8-4ab9-9abf-82d767099187 · outbound

This paper cites The Llama 3 Herd of Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs The Llama 3 Herd of Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.594726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.594726Z digest=sha256:f59943095cbf040a2648c0819e20918eebc430beebb930c366761f4763e85e52

Observation 8f0f6360-46fd-4638-bf8b-25ea3bd72a8f · outbound

This paper cites Frustratingly easy test-time adaptation of vision-language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Frustratingly easy test-time adaptation of vision-language models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.451008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.597549Z digest=sha256:8bc1a75c35456ded827d6bd0dec45062b795f2b68c3fd3696150ebc8de460360

Observation 76e45a9b-8330-462b-908d-a21adb85346c · outbound

This paper cites Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.600409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.600409Z digest=sha256:6f81f017ad9248c3d9a2d48e7fdedb33eee8cde626e3742d458783652f250dc2

Observation 7ab32076-071c-4558-8c3c-479fce448f6c · outbound

This paper cites T oxi G en: A large-scale machine-generated dataset for adversarial and implicit hate speech detection.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs T oxi G en: A large-scale machine-generated dataset for adversarial and implicit hate speech detection

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.603323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.603323Z digest=sha256:5e20d2cc9467335bd8bd6d1ab8146f4524c8356df46b23b58f8ffe66baecd8b8

Observation 2da50cd6-d2d0-4cdc-9454-8fc2e2f81509 · outbound

This paper cites Measuring massive multitask language understanding.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Measuring massive multitask language understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.606012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.606012Z digest=sha256:b5194121ac42f881cbb24c67e9fc295cabf9ac62fde6de689b3dcf9ae900688e

Observation 04a805b3-bdef-4baa-b9c3-9a167c888880 · outbound

This paper cites Non-linear inference time intervention: Improving llm truthfulness.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Non-linear inference time intervention: Improving llm truthfulness

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.437357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.608618Z digest=sha256:76a5568f870f33a6e6ceb26312822af65861836163971bbff66a38683f28cc74

Observation d17db6cb-0ee8-4567-9910-42f85b2e7710 · outbound

This paper cites Lo RA : Low-rank adaptation of large language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Lo RA : Low-rank adaptation of large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.611121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.611121Z digest=sha256:42e52bdc04bbcd4e2ceb98fe172677e930cd6eb93ce85f3bd0b8e32f1d16f841

Observation 25c45511-6ade-4fc4-bf25-a37e90e5a1c9 · outbound

This paper cites an unresolved cited work.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.614115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.614115Z digest=sha256:b5631db1e1986b21aee7daa02dbcb929cb29aada0f6eebf9e713a0f5ba5f7226

Observation 43ee9a39-71cc-4eca-baab-d752563cdeb2 · outbound

This paper cites A unified understanding and evaluation of steering methods.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs A unified understanding and evaluation of steering methods

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.616809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.616809Z digest=sha256:9b6e8658778d8bd155b790bf962ce3b125615df0771a15448078b7a19eca5fde

Observation 3582b7c4-0072-4d99-95cb-f8f3687d2099 · outbound

This paper cites Guided integrated gradients: An adaptive path method for removing noise.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Guided integrated gradients: An adaptive path method for removing noise

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.415402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.619720Z digest=sha256:546fe7e9c9a7d0a611414009edf89f3eac78266320752f0de0683812366c6eb6

Observation efa84903-f5b9-434b-8227-86218e39290e · outbound

This paper cites Analyzing Finetuning Representation Shift for Multimodal LLMs Steering.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Analyzing Finetuning Representation Shift for Multimodal LLMs Steering

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:45:09.072091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.622861Z digest=sha256:fe036df682b22c93bc69f1c8d8bd55a22f69a05ee0fc9e76441b4bb7481f91ff

Observation 71fc0b72-392e-40eb-ad55-08afc769ee97 · outbound

This paper cites Mitigating object hallucinations in large vision-language models through visual contrastive decoding.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Mitigating object hallucinations in large vision-language models through visual contrastive decoding

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.405497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.625935Z digest=sha256:498d7b36dfa64a09d88e3a065d7d7ad4950c080ed442e14cc6c285fd683f64bb

Observation 2b00a94e-42cd-4214-a91e-cc483e2c660c · outbound

This paper cites Inference-time intervention: Eliciting truthful answers from a language model.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Inference-time intervention: Eliciting truthful answers from a language model

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.395690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.629137Z digest=sha256:3e74dd502dd03da22be4752614a4d9e3899ed44bdafcb2730ee36d7b23308aeb

Observation c3260a3d-e4c9-42e8-b920-f6b1b8176bde · outbound

This paper cites Learning without forgetting.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Learning without forgetting

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.385719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.631701Z digest=sha256:b9ecb4bb9535131188b6634abce7d13d9891cccf6198b245240dc6987abcae05

Observation cd2dd1af-1e0c-49bc-bd2a-2b4f748b7f53 · outbound

This paper cites T ruthful QA : Measuring how models mimic human falsehoods.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs T ruthful QA : Measuring how models mimic human falsehoods

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.634281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.634281Z digest=sha256:fabb981c985175184f8b29f6715d0ca1f270cdb2574bdb2ffa75a48b771489e2

Observation 7f7a930d-396f-4b96-8eb2-45c9931602cd · outbound

This paper cites A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.637173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.637173Z digest=sha256:72b92757fbe79d9b67eb4bc029c6553ac8fb57a57f291a56f03f6eb21d1dc7d6

Observation 22c972f3-8ab0-4eed-a648-6646e4c90cbe · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge, January 2024 a.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Llava-next: Improved reasoning, ocr, and world knowledge, January 2024 a

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.640083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.640083Z digest=sha256:f92a24278d6ff67c27c9016598dd026e17f73de2ea9ede84a1db9873cd359fda

Observation a7532f68-77b1-4b10-8712-fc3f60676dca · outbound

This paper cites Reducing Hallucinations in Vision-Language Models via Latent Space Steering.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Reducing Hallucinations in Vision-Language Models via Latent Space Steering

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.642780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.642780Z digest=sha256:7d1461b75707ed2cd01297997aabc5447b7a22e6b6e9097444a6376f1cee4983

Observation 4ff7d5fc-9ae2-44c6-bbe3-9bc4f78136a1 · outbound

This paper cites In-context vectors: Making in context learning more effective and controllable through latent space steering.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs In-context vectors: Making in context learning more effective and controllable through latent space steering

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.371802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.645791Z digest=sha256:7313959a32fbcf3543ee619f6a45bfc8b77b3a0eb67651682bda8da5d3c9c704

Observation 08649fc5-66b9-4b34-a3b9-caac353363b6 · outbound

This paper cites Gradient episodic memory for continual learning.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Gradient episodic memory for continual learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.648568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.648568Z digest=sha256:43aef3db1058e06a29110cff51f10dbd06c0047da966cd9fbc8e2d833c7b0984

Observation 307e13ad-cde6-43d4-bb0d-cd12def26bbb · outbound

This paper cites The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.651325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.651325Z digest=sha256:76945039fe0bfc78e2828f65ae17fab1ecc1d6f203b12d43944782645842b323

Observation c5d6b8ea-6ce1-4d61-a002-fdd8235e00d4 · outbound

This paper cites Faitheval: Can your language model stay faithful to context, even if ''the moon is made of marshmallows''.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Faitheval: Can your language model stay faithful to context, even if ''the moon is made of marshmallows''

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.357480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.654306Z digest=sha256:cddbdc6db77307225e1d392deed76ff8f94ae37846a2d6c1b4577fb08d50a087

Observation 90409f6c-d5da-4fb9-9a06-071da951e821 · outbound

This paper cites Risk-aware distributional intervention policies for language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Risk-aware distributional intervention policies for language models

Reference 30

Resolution
verified exact
raw_fallback, observed 2026-08-06T14:45:09.026204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.657077Z digest=sha256:b5fee28fba516bbd961b2e75d85765b0b0139a524d0ce3825808e0e56637b121

Observation c87a7114-afc0-463e-ad52-323906112331 · outbound

This paper cites Multi-Attribute Steering of Language Models via Targeted Intervention.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Multi-Attribute Steering of Language Models via Targeted Intervention

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.659788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.659788Z digest=sha256:22fb34e32c3115585822978741a843c8da556a497d70cae53da3de32c22ec88a

Observation bf484691-5e92-4699-a0c7-f686fe385641 · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Steering Llama 2 via Contrastive Activation Addition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.662726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.662726Z digest=sha256:1e47bd7ca1925d0e17d237c0742cf8b1003f60801ea64b672df915752f494c11

Observation 9e31f871-3e83-46f5-a1c1-0675b521466f · outbound

This paper cites Combining feature and instance attribution to detect artifacts.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Combining feature and instance attribution to detect artifacts

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.665561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.665561Z digest=sha256:02695ad7a898a0ec5eb971e305e4960a284c916b834dad97fd5484921e1393ef

Observation b3ab98e7-b1de-46dc-bcbc-c66f88fff68c · outbound

This paper cites Householder pseudo-rotation: A novel approach to activation editing in LLM s with direction-magnitude perspective.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Householder pseudo-rotation: A novel approach to activation editing in LLM s with direction-magnitude perspective

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.668433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.668433Z digest=sha256:3ec0e35028af16a94cee0f21241fbac29caab878dae1dac0ad41c3bf27ff050d

Observation f429a8e2-4c78-4c9c-bf63-850a4535e52a · outbound

This paper cites Rewarded soups: towards pareto-optimal alignment by interpolating weights fine-tuned on diverse rewards.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Rewarded soups: towards pareto-optimal alignment by interpolating weights fine-tuned on diverse rewards

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.348565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.671136Z digest=sha256:15e5e3ffb22649e2d1df94ec6efb515269fdfbdc1ab41d404060be13ea296e90

Observation d909c43d-62fc-492d-bd0f-027cb67db8d8 · outbound

This paper cites Steering llama 2 via contrastive activation addition.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Steering llama 2 via contrastive activation addition

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.674114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.674114Z digest=sha256:717a36689190b6b4d3347af194d74884566368ea0ee8781194a4717cf00c5f10

Observation 250a1e14-e315-43cd-a800-c86a6dff5b33 · outbound

This paper cites A consistent and efficient evaluation strategy for attribution methods.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs A consistent and efficient evaluation strategy for attribution methods

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.339433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.677039Z digest=sha256:c6445b833716bb0c598f97120d71e2a77e0cd67e6882f9fb8ace8a2a3bcedd87

Observation 9949bf89-c3a2-4818-b9ed-35abec54c628 · outbound

This paper cites Are vision-language transformers learning multimodal representations? a probing perspective.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Are vision-language transformers learning multimodal representations? a probing perspective

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.330728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.679731Z digest=sha256:c73be08a42eeb6553558740a48f9cefe0ba72d169bed400856fab218cea2cbe1

Observation 3e589c6d-1c9c-4c44-917a-7f39cbedc15f · outbound

This paper cites Smith, and Simon Shaolei Du.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Smith, and Simon Shaolei Du

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.321912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.682504Z digest=sha256:94cec831d31218d8f3dda32a6fbfab5951dfea8f10ccc6f2da315ef17b47329e

Observation d4b3c61d-32f4-4365-84f7-f216dfbea6e6 · outbound

This paper cites SmoothGrad: removing noise by adding noise.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs SmoothGrad: removing noise by adding noise

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.685271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.685271Z digest=sha256:9d4965909155ff036e8cef98c6d2cb245e4afdf188bdc5b428d3d954d5c9bf95

Observation fc9b94f9-54e9-4fac-98a5-39e8de287198 · outbound

This paper cites Efficient open-set test time adaptation of vision language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Efficient open-set test time adaptation of vision language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.312946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.688226Z digest=sha256:baa3dded2f72fb046cb581cb5795aef17048f80a7e34c07fe3dd2b4c7f9d749a

Observation c2d1f33e-2cb3-4dbf-803e-9a075293b848 · outbound

This paper cites LVP runing: An effective yet simple language-guided vision token pruning approach for multi-modal large language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs LVP runing: An effective yet simple language-guided vision token pruning approach for multi-modal large language models

Reference 42

Resolution
verified exact
doi, observed 2026-08-06T14:45:08.756999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.691139Z digest=sha256:c03e5828c0c9ece1b5e4d964e85b7f636afc4f8c8168f2b5a0df37dbd4c0564f

Observation a84a0dbb-863d-4fcb-9ac2-f169a35cd4f6 · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.693838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.693838Z digest=sha256:efe4d6404cb51983dc50aa07dc2585af53205f67c6e6ae4c65905e8437ce6086

Observation 1c2a945c-42f9-4d1e-b88d-d02ae2cd1484 · outbound

This paper cites Axiomatic attribution for deep networks.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Axiomatic attribution for deep networks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.304142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.696753Z digest=sha256:e95e84a8133950254e10603fa2029149bd1d6a036913677af21e1681f1c889df

Observation 07fda6dd-1818-49fb-8499-74aa319d0eaa · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Gemini: A Family of Highly Capable Multimodal Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.699397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.699397Z digest=sha256:fc36176863b697de3fb76fa9e5c9e25f3f12d89c54ebca51617e5c3c422b2d04

Observation d7aa9766-d119-414c-8595-a070a843e9ba · outbound

This paper cites Gemma 3 Technical Report.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Gemma 3 Technical Report

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.701915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.701915Z digest=sha256:1e6dc557cd450e1a6d20fbfa9a097b7c97393de4b033263ca7bfd052103b90c4

Observation 518427e3-aa99-4980-ba79-0eb85226a8c6 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Qwen2.5: A party of foundation models, September 2024

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.704780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.704780Z digest=sha256:9d826c30aed2904db0478458f83b1d4df619d2de53f333fb1e245956052784f2

Observation 5393a4fa-0e7e-457a-a5d9-1b62a0130e89 · outbound

This paper cites Steering Language Models With Activation Engineering.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Steering Language Models With Activation Engineering

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.707384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.707384Z digest=sha256:bbcd4cb332c8da5115ace942cd7dca3582f676e1173d0b316f716b4ab4ae8bc3

Observation 9641300d-4a91-494e-b546-e32db362dc7b · outbound

This paper cites Contrastive region guidance: Improving grounding in vision-language models without training.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Contrastive region guidance: Improving grounding in vision-language models without training

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.289794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.709928Z digest=sha256:867702bec50c5041f3d97f95684b5a4980441bfd02a503f0371bb732bd4e9e5e

Observation 6c4c1097-4f9f-4e3d-9636-c38df5f3b42c · outbound

This paper cites AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:45:08.856617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.712547Z digest=sha256:caeca97d1adc0f25cfbadfdb1b1799c47768763cb8e3ee42de0e9611479fe83f

Observation b4d6127e-65cb-4c24-9b3a-0da6255b599a · outbound

This paper cites Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.715420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.715420Z digest=sha256:45129610155e8284787d24ee204e6f82fcc637758532defd135e96338eabfbab

Observation 44b2cb02-fdb6-45b0-bfc3-3bae28d1c682 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.718348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.718348Z digest=sha256:5bdcd21927139245ba5de0043499a51bfa600d641ce0489c4111a0092c841766

Observation d841a76a-6ab2-4fe8-8b92-cefda9142c5a · outbound

This paper cites SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.721119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.721119Z digest=sha256:a72483c04c822e81a6b52008e636e3c8a311550253ac3071ef562b76801b8d4f

Observation 27dad7be-ea30-489b-bef2-e3dabe09eaff · outbound

This paper cites Bayesian Test-Time Adaptation for Vision-Language Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Bayesian Test-Time Adaptation for Vision-Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.723935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.723935Z digest=sha256:d0c2489797ff208b890d8d7c52bb582b28fd7d5a3e4274bfaf5d7fef0171d3b5

Observation e742f20d-6cef-48ab-b058-ac76114d6c0d · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Representation Engineering: A Top-Down Approach to AI Transparency

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.726791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.726791Z digest=sha256:e6ee7282a47e9d0fe084b77ad3f127c1b88d64177defe795b73c24dfca0cd4f2

Pith citing papers

Observation 678ca0c8-c38f-4fe7-bc54-6f713fff697f · inbound

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs cites this paper.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.641226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.641226Z digest=sha256:76420151b5568c8f98d6449404e42f3061add3426a5f9f11deab54b6bd768ef6

Observation 47dc51af-060a-47b1-a192-6fa63f5668a5 · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 220

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.983164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:6a1d5df9c72b03bab3ac1cc0c6f6c9bdc130f562742223dc294964880d0b5b1d

Observation d35ccd77-62a6-4cab-9aef-e20585a927c7 · inbound

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment cites this paper.

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.983164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:36:44.401045Z digest=sha256:20196d341641a23da260e8dda289dd98c5f947c7a36b422d1693e605c5ec9c63

Observation 04ddcd88-d3ec-46b6-b58f-1dbaebb94308 · inbound

Continuous Interpretive Steering for Scalar Diversity cites this paper.

Continuous Interpretive Steering for Scalar Diversity GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-13T00:17:01.983164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:08:02.296583Z digest=sha256:224672d7382ae6b0dc7a910cb02f71ddb2c6cf27cac3e480f837407eecd1d74c

Observation 82b611a0-a29f-46fc-9616-bd1bb95bfcb5 · inbound

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection cites this paper.

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.983164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T13:53:27.306664Z digest=sha256:380ae0a48e4a559ccfc00c5bd7a62c500d1613109bff7be82eb2e5ac069fae1e