Pith. sign in

Paper Citation Record · LEDGER

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

As of 10 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 5 inbound Pith citation observations for arXiv:2507.18043.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18043 v2

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:45:08.726791Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T21:15:45.641226Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact7
  • verified fuzzy17
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation b21dbeae-8672-45f1-bbd8-86efa39101db · outbound

This paper cites GPT-4 Technical Report.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.572148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.572148Z digest=sha256:92d696f173a9d90d31838a4b928adf909625bd035ceefdaa0e9c5b71d67ab55e

Observation d2a132a2-d174-4d2f-a068-fb19a78fa2e2 · outbound

This paper cites A diagnostic study of explainability techniques for text classification.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs A diagnostic study of explainability techniques for text classification

Reference 2

Resolution
verified exact
doi, observed 2026-08-06T14:45:08.801516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.575953Z digest=sha256:22c3f5174dcd1775233aca523cf67e204b52f294f9c42556ea651faafca7108b

Observation 7a2f0d49-3cf4-4906-bbe9-11933ca3da4b · outbound

This paper cites Madtp: Multimodal alignment-guided dynamic token pruning for accelerating vision-language transformer.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Madtp: Multimodal alignment-guided dynamic token pruning for accelerating vision-language transformer

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.468380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.579024Z digest=sha256:1ab01feee19346912a1e9cdede72c66b76872db314e07cb33226ccb8878be594

Observation 3201bf3e-025a-4982-93a6-c4fe254e9f0f · outbound

This paper cites Xprompt: Explaining large language model's generation via joint prompt attribution.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Xprompt: Explaining large language model's generation via joint prompt attribution

Reference 4

Resolution
verified exact
raw_fallback, observed 2026-08-06T14:45:09.267613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.582177Z digest=sha256:a13df85f5cb78e97d23cc62ca2354fdd51cc26d3e37f659b4a05d3d66473befe

Observation 472e75fd-80c5-4a34-abf7-c13dfc1fc7ab · outbound

This paper cites ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.585297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.585297Z digest=sha256:bc6e98993e26abd971a6522dc353631465f8e52f48e3e7aa6ac8777dc490ef43

Observation f7946dc7-3893-4d32-9ef0-0b864639c556 · outbound

This paper cites Quantifying and Mitigating Unimodal Biases in Multimodal Large Language Models: A Causal Perspective.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Quantifying and Mitigating Unimodal Biases in Multimodal Large Language Models: A Causal Perspective

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:45:09.167867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.588593Z digest=sha256:8061fdce50b85903be684d354cb81c5e02572e36ff17f1b7146f98ae7cfcbf59

Observation 6a213b3b-a79b-4099-a489-265448e343e6 · outbound

This paper cites Covert, Scott Lundberg, and Su-In Lee.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Covert, Scott Lundberg, and Su-In Lee

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.459771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.591990Z digest=sha256:621e52e563af8a39d1ded184931650510019d0498ad3042841edcddb61fb53a3

Observation 7d61dc24-08f8-4ab9-9abf-82d767099187 · outbound

This paper cites The Llama 3 Herd of Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs The Llama 3 Herd of Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.594726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.594726Z digest=sha256:4ec36f03c9935e94492dc6b5eeaa0992e86bf040131303771ae55048f57c209a

Observation 8f0f6360-46fd-4638-bf8b-25ea3bd72a8f · outbound

This paper cites Frustratingly easy test-time adaptation of vision-language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Frustratingly easy test-time adaptation of vision-language models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.451008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.597549Z digest=sha256:1aae4333f76571a01194f298991706a5c9635fb92276c2838242d277bf629dd5

Observation 76e45a9b-8330-462b-908d-a21adb85346c · outbound

This paper cites Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.600409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.600409Z digest=sha256:6f81f017ad9248c3d9a2d48e7fdedb33eee8cde626e3742d458783652f250dc2

Observation 7ab32076-071c-4558-8c3c-479fce448f6c · outbound

This paper cites T oxi G en: A large-scale machine-generated dataset for adversarial and implicit hate speech detection.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs T oxi G en: A large-scale machine-generated dataset for adversarial and implicit hate speech detection

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.603323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.603323Z digest=sha256:5e20d2cc9467335bd8bd6d1ab8146f4524c8356df46b23b58f8ffe66baecd8b8

Observation 2da50cd6-d2d0-4cdc-9454-8fc2e2f81509 · outbound

This paper cites Measuring massive multitask language understanding.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Measuring massive multitask language understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.606012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.606012Z digest=sha256:b5194121ac42f881cbb24c67e9fc295cabf9ac62fde6de689b3dcf9ae900688e

Observation 04a805b3-bdef-4baa-b9c3-9a167c888880 · outbound

This paper cites Non-linear inference time intervention: Improving llm truthfulness.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Non-linear inference time intervention: Improving llm truthfulness

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.437357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.608618Z digest=sha256:3a37e68b35bd46d3b57ec88fb32682c5cc8d87b297ada1dd019ac80e999030d7

Observation d17db6cb-0ee8-4567-9910-42f85b2e7710 · outbound

This paper cites Lo RA : Low-rank adaptation of large language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Lo RA : Low-rank adaptation of large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.611121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.611121Z digest=sha256:42e52bdc04bbcd4e2ceb98fe172677e930cd6eb93ce85f3bd0b8e32f1d16f841

Observation 25c45511-6ade-4fc4-bf25-a37e90e5a1c9 · outbound

This paper cites an unresolved cited work.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.614115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.614115Z digest=sha256:b5631db1e1986b21aee7daa02dbcb929cb29aada0f6eebf9e713a0f5ba5f7226

Observation 43ee9a39-71cc-4eca-baab-d752563cdeb2 · outbound

This paper cites A unified understanding and evaluation of steering methods.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs A unified understanding and evaluation of steering methods

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.616809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.616809Z digest=sha256:9b6e8658778d8bd155b790bf962ce3b125615df0771a15448078b7a19eca5fde

Observation 3582b7c4-0072-4d99-95cb-f8f3687d2099 · outbound

This paper cites Guided integrated gradients: An adaptive path method for removing noise.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Guided integrated gradients: An adaptive path method for removing noise

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.415402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.619720Z digest=sha256:5ba5757fdb17ce244c847124fe247e1bf731afe902f9c57f452389b69ae851e8

Observation efa84903-f5b9-434b-8227-86218e39290e · outbound

This paper cites Analyzing Finetuning Representation Shift for Multimodal LLMs Steering.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Analyzing Finetuning Representation Shift for Multimodal LLMs Steering

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:45:09.072091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.622861Z digest=sha256:f3fa77e1febdfe16b54e2888f0cac95aa649e3c2b243af70ef0bcf732a67cd92

Observation 71fc0b72-392e-40eb-ad55-08afc769ee97 · outbound

This paper cites Mitigating object hallucinations in large vision-language models through visual contrastive decoding.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Mitigating object hallucinations in large vision-language models through visual contrastive decoding

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.405497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.625935Z digest=sha256:f181a8b48cf4384ed0e9d0d17015f49ee6aa8e812839552df7b3ce182fc7409c

Observation 2b00a94e-42cd-4214-a91e-cc483e2c660c · outbound

This paper cites Inference-time intervention: Eliciting truthful answers from a language model.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Inference-time intervention: Eliciting truthful answers from a language model

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.395690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.629137Z digest=sha256:f9b0be2751578b7925df27aab6d970cd4b9252415f47f09c3b4fdac3b0fc745d

Observation c3260a3d-e4c9-42e8-b920-f6b1b8176bde · outbound

This paper cites Learning without forgetting.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Learning without forgetting

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.385719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.631701Z digest=sha256:c52dc1021a4d5b3976caceb18bd07ce65003d0b7e7b39277c1a71463530dc28b

Observation cd2dd1af-1e0c-49bc-bd2a-2b4f748b7f53 · outbound

This paper cites T ruthful QA : Measuring how models mimic human falsehoods.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs T ruthful QA : Measuring how models mimic human falsehoods

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.634281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.634281Z digest=sha256:fabb981c985175184f8b29f6715d0ca1f270cdb2574bdb2ffa75a48b771489e2

Observation 7f7a930d-396f-4b96-8eb2-45c9931602cd · outbound

This paper cites A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.637173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.637173Z digest=sha256:72b92757fbe79d9b67eb4bc029c6553ac8fb57a57f291a56f03f6eb21d1dc7d6

Observation 22c972f3-8ab0-4eed-a648-6646e4c90cbe · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge, January 2024 a.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Llava-next: Improved reasoning, ocr, and world knowledge, January 2024 a

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.640083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.640083Z digest=sha256:f92a24278d6ff67c27c9016598dd026e17f73de2ea9ede84a1db9873cd359fda

Observation a7532f68-77b1-4b10-8712-fc3f60676dca · outbound

This paper cites Reducing Hallucinations in Vision-Language Models via Latent Space Steering.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Reducing Hallucinations in Vision-Language Models via Latent Space Steering

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.642780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.642780Z digest=sha256:7d1461b75707ed2cd01297997aabc5447b7a22e6b6e9097444a6376f1cee4983

Observation 4ff7d5fc-9ae2-44c6-bbe3-9bc4f78136a1 · outbound

This paper cites In-context vectors: Making in context learning more effective and controllable through latent space steering.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs In-context vectors: Making in context learning more effective and controllable through latent space steering

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.371802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.645791Z digest=sha256:fd67f73575ca1d9b08da5722dc954825d9e70756afd44d252bf28ba858e1b485

Observation 08649fc5-66b9-4b34-a3b9-caac353363b6 · outbound

This paper cites Gradient episodic memory for continual learning.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Gradient episodic memory for continual learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.648568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.648568Z digest=sha256:43aef3db1058e06a29110cff51f10dbd06c0047da966cd9fbc8e2d833c7b0984

Observation 307e13ad-cde6-43d4-bb0d-cd12def26bbb · outbound

This paper cites The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.651325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.651325Z digest=sha256:76945039fe0bfc78e2828f65ae17fab1ecc1d6f203b12d43944782645842b323

Observation c5d6b8ea-6ce1-4d61-a002-fdd8235e00d4 · outbound

This paper cites Faitheval: Can your language model stay faithful to context, even if ''the moon is made of marshmallows''.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Faitheval: Can your language model stay faithful to context, even if ''the moon is made of marshmallows''

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.357480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.654306Z digest=sha256:ad86b649933840e65948e0e6d6a433cc59e0807d994ee19cf272cbd9dac446a0

Observation 90409f6c-d5da-4fb9-9a06-071da951e821 · outbound

This paper cites Risk-aware distributional intervention policies for language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Risk-aware distributional intervention policies for language models

Reference 30

Resolution
verified exact
raw_fallback, observed 2026-08-06T14:45:09.026204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.657077Z digest=sha256:620812f2894589f76873cc3662a0dc29bbeae946fa769edf08aeaa43612806fc

Observation c87a7114-afc0-463e-ad52-323906112331 · outbound

This paper cites Multi-Attribute Steering of Language Models via Targeted Intervention.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Multi-Attribute Steering of Language Models via Targeted Intervention

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.659788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.659788Z digest=sha256:22fb34e32c3115585822978741a843c8da556a497d70cae53da3de32c22ec88a

Observation bf484691-5e92-4699-a0c7-f686fe385641 · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Steering Llama 2 via Contrastive Activation Addition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.662726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.662726Z digest=sha256:1e47bd7ca1925d0e17d237c0742cf8b1003f60801ea64b672df915752f494c11

Observation 9e31f871-3e83-46f5-a1c1-0675b521466f · outbound

This paper cites Combining feature and instance attribution to detect artifacts.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Combining feature and instance attribution to detect artifacts

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.665561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.665561Z digest=sha256:02695ad7a898a0ec5eb971e305e4960a284c916b834dad97fd5484921e1393ef

Observation b3ab98e7-b1de-46dc-bcbc-c66f88fff68c · outbound

This paper cites Householder pseudo-rotation: A novel approach to activation editing in LLM s with direction-magnitude perspective.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Householder pseudo-rotation: A novel approach to activation editing in LLM s with direction-magnitude perspective

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.668433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.668433Z digest=sha256:3ec0e35028af16a94cee0f21241fbac29caab878dae1dac0ad41c3bf27ff050d

Observation f429a8e2-4c78-4c9c-bf63-850a4535e52a · outbound

This paper cites Rewarded soups: towards pareto-optimal alignment by interpolating weights fine-tuned on diverse rewards.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Rewarded soups: towards pareto-optimal alignment by interpolating weights fine-tuned on diverse rewards

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.348565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.671136Z digest=sha256:cb329e239f07cb8bf6b07ee86da5677b2ce00e9431fdbc9e792bc11aa17d1800

Observation d909c43d-62fc-492d-bd0f-027cb67db8d8 · outbound

This paper cites Steering llama 2 via contrastive activation addition.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Steering llama 2 via contrastive activation addition

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.674114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.674114Z digest=sha256:717a36689190b6b4d3347af194d74884566368ea0ee8781194a4717cf00c5f10

Observation 250a1e14-e315-43cd-a800-c86a6dff5b33 · outbound

This paper cites A consistent and efficient evaluation strategy for attribution methods.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs A consistent and efficient evaluation strategy for attribution methods

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.339433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.677039Z digest=sha256:9b79cbde2e0f446b93b8d5e2870bae42b08518d5e148779b6e22bb0c637347ad

Observation 9949bf89-c3a2-4818-b9ed-35abec54c628 · outbound

This paper cites Are vision-language transformers learning multimodal representations? a probing perspective.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Are vision-language transformers learning multimodal representations? a probing perspective

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.330728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.679731Z digest=sha256:3969d1f8d55aa2981dbe078b142cd5026adc86e1cb40eb854c888895d07d33c0

Observation 3e589c6d-1c9c-4c44-917a-7f39cbedc15f · outbound

This paper cites Smith, and Simon Shaolei Du.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Smith, and Simon Shaolei Du

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.321912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.682504Z digest=sha256:3531ab2f6bc53d603eaea8bc618c4586975fd570d73e408b8fb040d93c8cdb45

Observation d4b3c61d-32f4-4365-84f7-f216dfbea6e6 · outbound

This paper cites SmoothGrad: removing noise by adding noise.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs SmoothGrad: removing noise by adding noise

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.685271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.685271Z digest=sha256:9d4965909155ff036e8cef98c6d2cb245e4afdf188bdc5b428d3d954d5c9bf95

Observation fc9b94f9-54e9-4fac-98a5-39e8de287198 · outbound

This paper cites Efficient open-set test time adaptation of vision language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Efficient open-set test time adaptation of vision language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.312946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.688226Z digest=sha256:9b25f291e2f4e6c3fd518c17b81a5348739fa6d3b466c2d0baf893e8d0c12658

Observation c2d1f33e-2cb3-4dbf-803e-9a075293b848 · outbound

This paper cites LVP runing: An effective yet simple language-guided vision token pruning approach for multi-modal large language models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs LVP runing: An effective yet simple language-guided vision token pruning approach for multi-modal large language models

Reference 42

Resolution
verified exact
doi, observed 2026-08-06T14:45:08.756999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.691139Z digest=sha256:25800b83e2b4b2c5cfe466ee05999f27d0e9efa8afbae4bbc7dea3065fd6e0e7

Observation a84a0dbb-863d-4fcb-9ac2-f169a35cd4f6 · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.693838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.693838Z digest=sha256:efe4d6404cb51983dc50aa07dc2585af53205f67c6e6ae4c65905e8437ce6086

Observation 1c2a945c-42f9-4d1e-b88d-d02ae2cd1484 · outbound

This paper cites Axiomatic attribution for deep networks.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Axiomatic attribution for deep networks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.304142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.696753Z digest=sha256:da1f6b2263feb2a59d411793241d64edbed17f991d14d804688b63baa11e328c

Observation 07fda6dd-1818-49fb-8499-74aa319d0eaa · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Gemini: A Family of Highly Capable Multimodal Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.699397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.699397Z digest=sha256:fc36176863b697de3fb76fa9e5c9e25f3f12d89c54ebca51617e5c3c422b2d04

Observation d7aa9766-d119-414c-8595-a070a843e9ba · outbound

This paper cites Gemma 3 Technical Report.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Gemma 3 Technical Report

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.701915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.701915Z digest=sha256:1e6dc557cd450e1a6d20fbfa9a097b7c97393de4b033263ca7bfd052103b90c4

Observation 518427e3-aa99-4980-ba79-0eb85226a8c6 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Qwen2.5: A party of foundation models, September 2024

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.704780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.704780Z digest=sha256:9d826c30aed2904db0478458f83b1d4df619d2de53f333fb1e245956052784f2

Observation 5393a4fa-0e7e-457a-a5d9-1b62a0130e89 · outbound

This paper cites Steering Language Models With Activation Engineering.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Steering Language Models With Activation Engineering

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.707384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.707384Z digest=sha256:bbcd4cb332c8da5115ace942cd7dca3582f676e1173d0b316f716b4ab4ae8bc3

Observation 9641300d-4a91-494e-b546-e32db362dc7b · outbound

This paper cites Contrastive region guidance: Improving grounding in vision-language models without training.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Contrastive region guidance: Improving grounding in vision-language models without training

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:45:09.289794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.709928Z digest=sha256:473ad340b06c709b33b041622019292828d7e52e060c144f26504f786686d71c

Observation 6c4c1097-4f9f-4e3d-9636-c38df5f3b42c · outbound

This paper cites AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:45:08.856617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T14:45:08.712547Z digest=sha256:e3d0d16414e715979200a311278c70d715f416e556759ebf27276d8e7634e53c

Observation b4d6127e-65cb-4c24-9b3a-0da6255b599a · outbound

This paper cites Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Nullu: Mitigating Object Hallucinations in Large Vision-Language Models via HalluSpace Projection

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.715420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.715420Z digest=sha256:31020a37a71dd893339f49d981a94499d8971262e9cd045f5a7f8db2fd754997

Observation 44b2cb02-fdb6-45b0-bfc3-3bae28d1c682 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.718348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.718348Z digest=sha256:5bdcd21927139245ba5de0043499a51bfa600d641ce0489c4111a0092c841766

Observation d841a76a-6ab2-4fe8-8b92-cefda9142c5a · outbound

This paper cites SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.721119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.721119Z digest=sha256:75ba188de425ecc6d6877e21ef606cca059dbcd2926ce86f972af57ac1d5193d

Observation 27dad7be-ea30-489b-bef2-e3dabe09eaff · outbound

This paper cites Bayesian Test-Time Adaptation for Vision-Language Models.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Bayesian Test-Time Adaptation for Vision-Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.723935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.723935Z digest=sha256:d0c2489797ff208b890d8d7c52bb582b28fd7d5a3e4274bfaf5d7fef0171d3b5

Observation e742f20d-6cef-48ab-b058-ac76114d6c0d · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs Representation Engineering: A Top-Down Approach to AI Transparency

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:08.726791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:45:08.726791Z digest=sha256:e6ee7282a47e9d0fe084b77ad3f127c1b88d64177defe795b73c24dfca0cd4f2

Pith citing papers

Observation 678ca0c8-c38f-4fe7-bc54-6f713fff697f · inbound

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs cites this paper.

T2T-VICL: Cross-Task Visual In-Context Learning via Implicit Text-Driven VLMs GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T21:15:45.641226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:15:45.641226Z digest=sha256:76420151b5568c8f98d6449404e42f3061add3426a5f9f11deab54b6bd768ef6

Observation 47dc51af-060a-47b1-a192-6fa63f5668a5 · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 220

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.983164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:a4883b8c5cca53a9c9840171f0046771370b6cd99ca0bcbb39b03814eb89a7e3

Observation d35ccd77-62a6-4cab-9aef-e20585a927c7 · inbound

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment cites this paper.

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.983164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:36:44.401045Z digest=sha256:8bd96d9510ffad489f58ac8d98a6433eb1b85afd165bbbf8425141f7ee037c49

Observation 04ddcd88-d3ec-46b6-b58f-1dbaebb94308 · inbound

Continuous Interpretive Steering for Scalar Diversity cites this paper.

Continuous Interpretive Steering for Scalar Diversity GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-13T00:17:01.983164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:08:02.296583Z digest=sha256:b338a8893563268ac930203bcc26214bb0b0beaedd6b1add28d1ae2affcbe771

Observation 82b611a0-a29f-46fc-9616-bd1bb95bfcb5 · inbound

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection cites this paper.

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-13T00:17:01.983164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T13:53:27.306664Z digest=sha256:4cbf007ae77956bf7786be9846960fbac6e7d86f30caac0b99c296651c52fac3