Pith. sign in

Paper Citation Record · LEDGER

Attacking Attention of Foundation Models Disrupts Downstream Tasks

As of 18 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2506.05394.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05394 v3

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:09:11.016820Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact3
  • verified fuzzy24
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7626cca1-6979-4450-a614-393ab109a233 · outbound

This paper cites Reveal of Vision Transformers Robustness against Adversarial Attacks.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Reveal of Vision Transformers Robustness against Adversarial Attacks

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:05.480828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:05.480828Z digest=sha256:ae023aaa392220ce76c1a12b6e1424efb3edba45dd3ec363707a2e1d580aa99f

Observation 11a91e3a-2a7f-40d2-bb0f-128da3243493 · outbound

This paper cites Are transformers more robust than cnns?Advances in neural information processing systems, 34:26831–26843, 2021.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Are transformers more robust than cnns?Advances in neural information processing systems, 34:26831–26843, 2021

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:17.858926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:05.545638Z digest=sha256:9cb6036766fb715b1437a5972bb2c4185512ffe17cf2bb135d532fc4dcd9220d

Observation 58f71eef-9dff-49a9-9b9e-bbd5a686e7ca · outbound

This paper cites Under- standing robustness of transformers for image classification.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Under- standing robustness of transformers for image classification

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:17.711964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:05.648923Z digest=sha256:d53f60c344c2aa7c3db05a0445ada6dfef1ef5321f08ca6cb0b91ba2fde42d1b

Observation 9daa0784-c5af-4ad8-a668-77c981a75d64 · outbound

This paper cites Language Models are Few-Shot Learners.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Language Models are Few-Shot Learners

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:05.717334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:05.717334Z digest=sha256:c9b03feb04bd69632dead5f13e43c6d65f92eedbfbb647ccb16d0adcbe5100e6

Observation da802625-2a0f-48fe-80da-a0791c43fc19 · outbound

This paper cites Towards evaluating the robustness of neural networks.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Towards evaluating the robustness of neural networks

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:17.603788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:05.788881Z digest=sha256:eb09c3509d527d29c1e286eb18d86f9ae6b0dfb3572134e38c07fcd456a2e39d

Observation 4fc3cd5e-2f00-4856-a911-12e536c0ae9d · outbound

This paper cites Poisoning Web-Scale Training Datasets is Practical.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Poisoning Web-Scale Training Datasets is Practical

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:05.889610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:05.889610Z digest=sha256:5ce553cda22b81d9c29cc1496e689dc9850dc907bb8a1e61a30e5d72425b5998

Observation 0e37b607-a716-46fc-91d4-2e2a3c05f109 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Emerg- ing properties in self-supervised vision transformers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:06.015566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:06.015566Z digest=sha256:25da4805ee90b43296aa0dc1432df9f184da39f05ef996215e0f1ae902b0703e

Observation dc601c4f-c72f-44ff-91bb-c0d20fb4fc71 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Attacking Attention of Foundation Models Disrupts Downstream Tasks BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:06.133020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:06.133020Z digest=sha256:3ab5027de9d71c1f28864e833530779c9fd2c63d91b7f8478c87e67a8e9961e9

Observation 2e9d71c7-e312-4234-a5d2-271fd261be2a · outbound

This paper cites Boosting adversarial at- tacks with momentum.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Boosting adversarial at- tacks with momentum

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:17.439580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:06.250022Z digest=sha256:0a154985fa4eee085096d6311df6cd3ec78c18f9fe8c085aada5a5dcef2220b5

Observation 793f91ff-fa44-43c0-a7d7-f9499318a575 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Attacking Attention of Foundation Models Disrupts Downstream Tasks An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:06.368933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:06.368933Z digest=sha256:40e9660ea2770b07eb676c477a2d473cb6df76e780adef898ef1142f2da1b153

Observation 56bb9b5a-5e2c-4e93-b17b-69f6e8bfe76e · outbound

This paper cites Adversarial examples for the openai clip in its zero-shot classification regime and their semantic gener- alization, 2021.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Adversarial examples for the openai clip in its zero-shot classification regime and their semantic gener- alization, 2021

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:17.304895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:06.470107Z digest=sha256:b1ef45ac514d9c0b523efbc6720d8b58e1f041cb19453977de7d9413ce8252f0

Observation f081d8eb-f314-4e95-a9ad-4cdba02ff7d0 · outbound

This paper cites Pixels still beat text: Attacking the openai clip model with text patches and adversarial pixel perturbations,.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Pixels still beat text: Attacking the openai clip model with text patches and adversarial pixel perturbations,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:17.199529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:06.574033Z digest=sha256:70cfbfe2f86c7183ee6459b81abfb33a9305839e9751446e1c043bb40db64608

Observation 5f77670b-aeb1-450d-814a-3b8be080d24f · outbound

This paper cites Patch-Fool: Are Vision Transformers Always Robust Against Adversarial Perturbations?.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Patch-Fool: Are Vision Transformers Always Robust Against Adversarial Perturbations?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:06.657206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:06.657206Z digest=sha256:cf80144c6e5f2165750db2a9b1c875867b63db3519bb3dca5338d96a42a4e704

Observation 00ef6cb9-51bf-409e-903a-0f7fd95a8826 · outbound

This paper cites Explaining and Harnessing Adversarial Examples.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Explaining and Harnessing Adversarial Examples

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:06.757462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:06.757462Z digest=sha256:db1be0ab2dccadf579a46dc8be5a99fb2e36f23861828a8555d51951dbc28c3e

Observation a933c4c3-6ee2-4534-bf22-366d31a37579 · outbound

This paper cites Are vision trans- formers robust to patch perturbations? InEuropean Con- ference on Computer Vision, pages 404–421.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Are vision trans- formers robust to patch perturbations? InEuropean Con- ference on Computer Vision, pages 404–421

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:17.064309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:06.800977Z digest=sha256:af945df4d780a9217bcb8bb057ac72af58a132ed716ea91ca66389b6a2d9acd7

Observation 23f999c7-dc3c-40cf-a482-00828550a3bc · outbound

This paper cites SA-Attack: Improving Adversarial Transferability of Vision-Language Pre-training Models via Self-Augmentation.

Attacking Attention of Foundation Models Disrupts Downstream Tasks SA-Attack: Improving Adversarial Transferability of Vision-Language Pre-training Models via Self-Augmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:06.917630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:06.917630Z digest=sha256:ead1e09276ca6faf7d646ee9574757e5f4cb4d4ef34488dba945410532d63db4

Observation 1dd8f011-0023-44fa-82e6-9b7d06bbcbfd · outbound

This paper cites Black-box adversarial attacks with limited queries and information.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Black-box adversarial attacks with limited queries and information

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:16.927343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:06.984387Z digest=sha256:690643f9c299526d2b8c7187f56fa02c9a292e35c4c8efc18a456dea82329e0f

Observation 6065268d-1d4a-47f2-bc12-2f7a41243596 · outbound

This paper cites Scal- ing up vision-language pretraining.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Scal- ing up vision-language pretraining

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:14.846145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:07.099614Z digest=sha256:f4ea387531ff6cb38aa5725eb420feea1120584cf5bfc6ba670f9ab968958a9e

Observation d1ce4c97-b8b6-4eee-b007-69eb45138b1e · outbound

This paper cites Exploring Adversarial Robustness of Vision Transformers in the Spectral Perspective.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Exploring Adversarial Robustness of Vision Transformers in the Spectral Perspective

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:09:11.913022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:07.214381Z digest=sha256:ddc4c84f7ae2f75390b5d9f8ed57b269a1423f42fa816b581b3739509440400a

Observation b3972a62-6d42-458f-b5e3-c7f924e66c48 · outbound

This paper cites Curved representation space of vision transformers.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Curved representation space of vision transformers

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:14.694350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:07.370476Z digest=sha256:6f7c4636d38fe13ee520a10f169f1b79926df6a9f3731103c5984e10964fc0f0

Observation 19adec9a-1c90-4a35-b8e9-1db62793f1f8 · outbound

This paper cites Segment any- thing.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Segment any- thing

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:07.484958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:07.484958Z digest=sha256:87e96b7a855dce09bc37a09a69c32581c58572d9f225917609871f1665478ced

Observation 65b4dde8-9768-407c-bf85-28772a52246f · outbound

This paper cites Benchmarking Robust Self-Supervised Learning Across Diverse Downstream Tasks.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Benchmarking Robust Self-Supervised Learning Across Diverse Downstream Tasks

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:09:11.626309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:07.644839Z digest=sha256:cce945cc05ef697b8d454987326cb0be1545c35b04e1a3b29877073891e986d3

Observation 3b7f22b2-8c01-4754-a1e4-c2c1e0d49ad8 · outbound

This paper cites Ad- versarial examples in the physical world.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Ad- versarial examples in the physical world

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:14.579557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:07.756757Z digest=sha256:931561f224b4fb4ad9ae30c5b2b6efe65c2e7869ca02af7a5c062e3bccbc2abb

Observation 0cd27d75-3b19-4e1d-aeb8-efd5757a6cba · outbound

This paper cites Align before fuse: Vision and language representation learn- ing with momentum distillation.Advances in neural infor- mation processing systems, 34:9694–9705, 2021.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Align before fuse: Vision and language representation learn- ing with momentum distillation.Advances in neural infor- mation processing systems, 34:9694–9705, 2021

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:07.862211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:07.862211Z digest=sha256:9eec289dc8db204e31e011ab944e3975f8d3b583551b4f7f64a96a661048c18a

Observation 9def767f-cd28-4f41-859c-5c79e7bb5712 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:14.369842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:07.989291Z digest=sha256:c97528b23ce11fca1c699d0402d709e702cbd71d94bac5e1fe0866073b54ace9

Observation eb46ec7c-4558-43b0-8573-cd926db7b156 · outbound

This paper cites Lawrence Zitnick, and Piotr Doll ´ar.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Lawrence Zitnick, and Piotr Doll ´ar

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:14.253854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:08.135547Z digest=sha256:46afe8151f9192753b6545be2c3d7262049407026457148ddca7f9f9cdb7e411

Observation 3eec59fc-367e-440a-9f6d-fc1d0caa59c7 · outbound

This paper cites Exploring the Relationship between Architecture and Adversarially Robust Generalization.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Exploring the Relationship between Architecture and Adversarially Robust Generalization

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:09:11.366802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:08.269622Z digest=sha256:ed694e4fd16b40fe84757afe6e09873c4c8b0ddfa18563eb0c1224d884ada768

Observation a5b6df67-9b01-44d6-99dc-aa2bf7619cd8 · outbound

This paper cites Decoupled Weight Decay Regularization.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Decoupled Weight Decay Regularization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:08.339731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:08.339731Z digest=sha256:9a2f6f83bb63c8d95284d847b4aaa9eac9d90575fec1d96527735deea151df39

Observation 570045b7-57d0-4948-8de4-394c01ea87f5 · outbound

This paper cites Set-level guidance at- tack: Boosting adversarial transferability of vision-language pre-training models.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Set-level guidance at- tack: Boosting adversarial transferability of vision-language pre-training models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:14.084024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:08.441542Z digest=sha256:bdb6ad9d3bcf9d981c2ce016069a9517fd842165eeceaceac0045079e3bbd640

Observation 8a34c410-6879-4e31-be9f-e0f7221af839 · outbound

This paper cites Towards Deep Learning Models Resistant to Adversarial Attacks.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Towards Deep Learning Models Resistant to Adversarial Attacks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:08.504906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:08.504906Z digest=sha256:fb93ab69e82444e262ebf1feb1e36fb6a15fe2d16d162f1227fbe4dacb4afdca

Observation 15af41e3-c67d-4f51-bbc2-dec40723f232 · outbound

This paper cites On the robustness of vision transformers to adversarial ex- amples.

Attacking Attention of Foundation Models Disrupts Downstream Tasks On the robustness of vision transformers to adversarial ex- amples

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:13.940599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:08.569013Z digest=sha256:e5bdfe5540f4fe7e067cf5112677ca89b183f23f77160c4bcff9966ecd702336

Observation 99ebc7d0-63ff-4303-8ca3-64ce3971cc23 · outbound

This paper cites Deepfool: a simple and accurate method to fool deep neural networks.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Deepfool: a simple and accurate method to fool deep neural networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:13.814685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:08.629534Z digest=sha256:7dd424a6777d864973b01729ba78368e8e6dd784fd4562711d5e3b5667b2b8fa

Observation 219af90e-baa5-4099-8c92-86ddd3e2877b · outbound

This paper cites On Improving Adversarial Transferability of Vision Transformers.

Attacking Attention of Foundation Models Disrupts Downstream Tasks On Improving Adversarial Transferability of Vision Transformers

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:08.714579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:08.714579Z digest=sha256:11804bd3b1826df45554034a3ba0dac9455b355d15e1e75038d13aa12495ccde

Observation cca0f566-c49a-43eb-9a85-eafc77709927 · outbound

This paper cites Reading Isn't Believing: Adversarial Attacks On Multi-Modal Neurons.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Reading Isn't Believing: Adversarial Attacks On Multi-Modal Neurons

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:08.838443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:08.838443Z digest=sha256:82fd3653fd577d32c3ddb06677442b9653e97eefd93cb1f945209091b5fe51d3

Observation 2b21263d-29d1-49a5-989e-b8ef3b6a3c7e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Attacking Attention of Foundation Models Disrupts Downstream Tasks DINOv2: Learning Robust Visual Features without Supervision

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:08.953482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:08.953482Z digest=sha256:2e351792913bb170c796e86cfa114dbda450ac5f55af2dcb7b56a3a04bb7e4a8

Observation 1dc4e3fc-2f26-4774-bf11-7723d3ed1b10 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Learning transferable visual models from natural language supervi- sion

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:09.044655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:09.044655Z digest=sha256:891a81a96fa7c5dcd6465e287eef766d4498cb9cc8e09a09cf35a367d09a6ee2

Observation 7836ddf6-d117-48ed-8c01-03583104991e · outbound

This paper cites Berg, and Li Fei-Fei.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Berg, and Li Fei-Fei

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:13.652122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:09.142432Z digest=sha256:c216876bab28934ffeda56a6e05b4ffff42a35ff64ac6d8c8c1e26667dd2eeb5

Observation a32c214e-383b-4d0b-966d-24e3e07131ed · outbound

This paper cites On the Adversarial Robustness of Vision Transformers.

Attacking Attention of Foundation Models Disrupts Downstream Tasks On the Adversarial Robustness of Vision Transformers

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:09.256860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:09.256860Z digest=sha256:c4746e9c1a70bac70768e77944e2ff2e1220b0666f16a23aae2ebee9e25c4f80

Observation a41a1064-c7b9-4c32-8ae1-1d44d16c9951 · outbound

This paper cites Cnn features off-the-shelf: an astound- ing baseline for recognition.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Cnn features off-the-shelf: an astound- ing baseline for recognition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:13.506985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:09.322501Z digest=sha256:c7a8262b7d058756112562bdd24cb4bb4e1a5da0c1925c160c8190f8e02eff9f

Observation 1334a06e-528a-4dad-a39e-23c81955865b · outbound

This paper cites Indoor segmentation and support inference from rgbd images.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Indoor segmentation and support inference from rgbd images

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:09.434559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:09.434559Z digest=sha256:232abf16d3535e558a82cd944c63fec49afe13298201ebcc7a7e1b4abc9ca4d6

Observation 1c1b17c9-ed47-4260-a236-d6b38c21b7b5 · outbound

This paper cites Adversarial risk and the dangers of eval- uating against weak attacks.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Adversarial risk and the dangers of eval- uating against weak attacks

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:13.358632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:09.768074Z digest=sha256:9956e3eff9c3f54911a8f2f714275f0217e27e375e61e2a6788ac587c8467f4f

Observation 2d089d22-30f3-447c-8fee-1fb4346a9d3e · outbound

This paper cites Benchmarking Zero-Shot Robustness of Multimodal Foundation Models: A Pilot Study.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Benchmarking Zero-Shot Robustness of Multimodal Foundation Models: A Pilot Study

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:09.880055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:09.880055Z digest=sha256:48d87e18b2ea5618368a2fa782bac18eec71461923fade2b4154da4aa4cb61b9

Observation ef17e8f2-5678-412c-bee2-392281cda67e · outbound

This paper cites Exploring Transferability of Multimodal Adversarial Samples for Vision-Language Pre-training Models with Contrastive Learning.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Exploring Transferability of Multimodal Adversarial Samples for Vision-Language Pre-training Models with Contrastive Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:10.053287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:10.053287Z digest=sha256:39d6cffed10b382811b2d928e7e0df65dbccf5456ef339be6fe1f7ffa52c950e

Observation 898cbbed-85a9-4753-a446-6d3e1550fe1e · outbound

This paper cites Towards transferable adversarial attacks on vision transformers.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Towards transferable adversarial attacks on vision transformers

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:13.118082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:10.169412Z digest=sha256:48a7c0581486423c74b808cc39f4c1d403b7e50588f372871dbd5b08ef720823

Observation e32fa98f-c5cb-40ef-accf-bc3e9b043359 · outbound

This paper cites Vision-language pre-training with triple contrastive learning.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Vision-language pre-training with triple contrastive learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:10.272598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:10.272598Z digest=sha256:f0e1766854ddbb8f6ef695764740788f18e18e21f465a64cb381bdbfbbfa8043

Observation 477ba9fa-9e72-480d-b95d-fef085abb6a7 · outbound

This paper cites an unresolved cited work.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:09:12.936539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:10.384701Z digest=sha256:e6a9a7512ef0f29fc9fea3ebf708f8883750f900fcab37be9851f59ca897f609

Observation 5d6e5fc0-836a-4f7c-a1c4-e3a3eb9f92e0 · outbound

This paper cites Florence: A New Foundation Model for Computer Vision.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Florence: A New Foundation Model for Computer Vision

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:10.491227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:10.491227Z digest=sha256:c243ae0a05c0c10e5da3587283c4d5a9f7af213b9f21f29ffa5bdf8e473275bd

Observation a43ef118-b3ca-41f4-bca2-47451371908a · outbound

This paper cites Towards adversarial attack on vision-language pre-training models.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Towards adversarial attack on vision-language pre-training models

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:12.772569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:10.669164Z digest=sha256:6dcd3cf7cd1408284f4c3323bd8b3ee149bd946862a549252c5629aaf29c9d1a

Observation 849749a2-0af8-4f96-897b-6277acf5c793 · outbound

This paper cites Transferable adversarial attacks on vision transform- ers with token gradient regularization.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Transferable adversarial attacks on vision transform- ers with token gradient regularization

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:12.638607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:10.768596Z digest=sha256:97dc919a0c36fc79e79dc26ac07d1c3291efa2290521617bb9df2b3a5ef92e3b

Observation 52842260-ca10-41b0-b35b-da2dc2307af3 · outbound

This paper cites Univer- sal adversarial perturbations for vision-language pre-trained models.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Univer- sal adversarial perturbations for vision-language pre-trained models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:12.387400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:10.895374Z digest=sha256:b453bbd900ffebcda94170d7915a0afe80a27c7eb76917c5c704704cdc2a39c8

Observation 2a16dad9-7676-4d23-94b8-a4dffcf7dbec · outbound

This paper cites Semantic under- standing of scenes through the ade20k dataset.International Journal of Computer Vision, 127(3):302–321, 2019.

Attacking Attention of Foundation Models Disrupts Downstream Tasks Semantic under- standing of scenes through the ade20k dataset.International Journal of Computer Vision, 127(3):302–321, 2019

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:09:12.194890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:09:11.016820Z digest=sha256:5fb50e1cbe193ce3f7e1d7b38ac25da87357bdaa149b85661ca5734a939b6f4a

Pith citing papers

No inbound Pith citation observations are available.