Pith. sign in

Paper Citation Record · LEDGER

Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2404.19287.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.19287 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:04:01.363663Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T13:23:28.107852Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c45f3868-043e-4b3e-98c7-f58623422f7e · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 243

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:34.230016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:49b635e00c69946a01c8120bd62f100704957368df1ed7b623b2e287328ecbc7

Observation b9b58612-b4bb-4b69-a49d-d0ceda8543f4 · inbound

BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization cites this paper.

BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:01.363663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:01.363663Z digest=sha256:a9c0e25773b4bbc8086acc1c99ec0b1fe6d1951d8c5a0b414b0e1cd393496b2e

Observation ed9c31f9-b857-487a-bd06-383e5ae4aa40 · inbound

Dual-Path Stable Soft Prompt Generation for Domain Generalization cites this paper.

Dual-Path Stable Soft Prompt Generation for Domain Generalization Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:34.741589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:34.741589Z digest=sha256:c944c871f73d68894eff4644aa34343a344df0f99843b3864e9b0ce53bc1989d

Observation 48fb0b46-4211-47e3-ade7-8dedcfe9dfa8 · inbound

Coordinated Robustness Evaluation Framework for Vision-Language Models cites this paper.

Coordinated Robustness Evaluation Framework for Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:49.610244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:40:49.610244Z digest=sha256:2ee37e4a08985c28174dc845c15ed5a0611b9ca4f627f0e07f4b5c76b1c26f06

Observation 6cdb1c89-ffa7-413e-98b4-26347ee70ee5 · inbound

On the Feasibility of Poisoning Text-to-Image AI Models via Adversarial Mislabeling cites this paper.

On the Feasibility of Poisoning Text-to-Image AI Models via Adversarial Mislabeling Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:50.267956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:50.267956Z digest=sha256:c7782468cd28f225eaf86439ceda3b1fc9bb0ab30a279b6e60cc9e5ca9c66399

Observation d3863703-51e5-41df-9d6c-1d1c333341cb · inbound

Invisible Injections: Exploiting Vision-Language Models Through Steganographic Prompt Embedding cites this paper.

Invisible Injections: Exploiting Vision-Language Models Through Steganographic Prompt Embedding Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T11:54:43.081752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:54:43.081752Z digest=sha256:89cc1f98af2fafae0550bd6a81d06ed0270672de61ad808d244d70ff62067e9d

Observation 460e04dd-8b0b-4288-91bb-5d1f87039018 · inbound

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models cites this paper.

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:26:33.883060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T15:23:13.318310Z digest=sha256:c13ad714c7a2821e35619363f06c906e76bb1a757ddaf62cf329e37a64fa77dc

Observation 7178a73d-70d6-497f-a870-60ef53d4a74a · inbound

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models cites this paper.

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:30:39.418151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T21:28:59.898029Z digest=sha256:83da0c0c935e0f1f6b49adf75a17b3b6674aac28f2491799449a0464822d8d7c

Observation 25df41f2-ed47-49fe-b0ad-c51c755b6b34 · inbound

Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP cites this paper.

Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T07:47:37.588939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:47:37.588939Z digest=sha256:2e6ffb93731eff3523392003c7e25faaf10f308cc18b7ec18c9df713a08567b7

Observation 537782d0-b461-44ba-a0c9-ee9e00e5ff48 · inbound

Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patch for Infrared Vision-Language Models cites this paper.

Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patch for Infrared Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T19:48:11.625968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T19:43:29.058335Z digest=sha256:a43941f24bd7b5bae08e71436b338b466f34ff880751b319441b71fc5f27e00f

Observation 361705f3-d9fa-4e77-a6f4-173ccdcbdf28 · inbound

Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patch for Infrared Vision-Language Models cites this paper.

Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patch for Infrared Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T16:52:02.261328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:52:02.261328Z digest=sha256:fff711ea5adacaa98d344b4bae827a472fd2e1c6e7e065e42d18a46fa134afc8

Observation ca664a2a-660c-4bd0-8455-5fd486f25e56 · inbound

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning cites this paper.

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:19:27.931230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T20:16:30.070819Z digest=sha256:cda6e26fd484f280a40095f7ea3cd034a04f6f2f8254d27d1ecd685e2a7268d7

Observation ef32f8f9-1cf9-435e-9468-f114b735f816 · inbound

TAME: Test-Time Adversarial Prompt Tuning via Mixture-of-Experts for Vision-Language Models cites this paper.

TAME: Test-Time Adversarial Prompt Tuning via Mixture-of-Experts for Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:23:21.496636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T14:20:49.278545Z digest=sha256:0e8157bd715ed9f18bf7aed4baab1267cd2859382e142620537fdaf58dcd40b1

Observation 9e9b3f9d-0f8a-47cf-ae78-e3f053369026 · inbound

JECA^2: Judgment-Explanation Consistent Adversarial Attack against Forensic Vision-Language Models cites this paper.

JECA^2: Judgment-Explanation Consistent Adversarial Attack against Forensic Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:23:28.109593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T13:18:46.414179Z digest=sha256:82498558c7bc3d78948e80094718ed9f0179089a0dec8043dd6cb7494ba6bd7c