Pith. sign in

Paper Citation Record · LEDGER

Explainability for Vision Foundation Models: A Survey

As of 11 August 2026, this Paper Citation Record lists 100 of 270 outbound references and 1 inbound Pith citation observation for arXiv:2501.12203.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12203 v1

Coverage vector

measured 100 of 270 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:26:35.341966Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:22:35.055073Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T19:22:35.279421Z

Reference resolution

100 of 270 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved96
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 46a76fd3-68c3-4c5e-9e7c-a87dce9dddb6 · outbound

This paper cites LeCun, Y.

Explainability for Vision Foundation Models: A Survey LeCun, Y

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.903222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.903222Z digest=sha256:ec2e7470770ce43b317981bd0477d9d7966ff2d13c14616ca10e573223e1d539

Observation d63406fa-aab4-4879-92f9-184da08657ea · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.909395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.909395Z digest=sha256:c186f51c36a30c4bd6ba85aee0ed507be5c9df9ac30848766cd3534fbc019545

Observation 760e12d3-877c-447e-aa46-6124a4b170dc · outbound

This paper cites Wortsman, G.

Explainability for Vision Foundation Models: A Survey Wortsman, G

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.914088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.914088Z digest=sha256:8c152cfd73b3a79c9bae44d5df3589b26e152657c4b95dc77d86a1d55c9a7160

Observation 1ee4df89-838c-4d65-8761-3cdce60f2be7 · outbound

This paper cites Sauer, K.

Explainability for Vision Foundation Models: A Survey Sauer, K

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.919120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.919120Z digest=sha256:90c5bf7b075b3b2622e155377187a2520b20f0a15be4ab2ba226831d5e78cedc

Observation c0293a86-ed96-449f-9a5f-774ff61ff2d8 · outbound

This paper cites Castelvecchi, Can we open the black box of ai?, Nature News 538 (2016) 20.

Explainability for Vision Foundation Models: A Survey Castelvecchi, Can we open the black box of ai?, Nature News 538 (2016) 20

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.924046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.924046Z digest=sha256:ea8fa6b2cf06fb006e8097555ee903c8d1c53294a01dc25d1f84980dd2c5bf79

Observation cc7bda3c-ccc8-4213-b5cc-f20c0a4003be · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.928639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.928639Z digest=sha256:5d82ec3338f7c4f6247a600f85506d38d4b4023f2e23a35f958750ab5825c2c3

Observation 787a6a1f-6965-4c6e-bcdf-bf5a095f4fcb · outbound

This paper cites Stakeholders in Explainable AI.

Explainability for Vision Foundation Models: A Survey Stakeholders in Explainable AI

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.934124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.934124Z digest=sha256:24097ea4d9d8f16b8aabbc839956c205b97ee3afbf7c6a992e0c4279feacbaa3

Observation d7cf1af0-5e93-4fa5-88dd-9d4e96bd64f6 · outbound

This paper cites Rawal, J.

Explainability for Vision Foundation Models: A Survey Rawal, J

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.938809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.938809Z digest=sha256:eef159fcd736c6c37dad6b5727c27b46da9fc89081895e8fade3218d7847b3cb

Observation 1ad17d11-d531-4db2-aa49-5522af251bc3 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.942962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.942962Z digest=sha256:d87d0a5a96361a776c7055a1464ad9d226e1b088f762e6006b82d504a79b724d

Observation 2851572f-373e-410c-a0f8-33c77fee7757 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.947922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.947922Z digest=sha256:6b67dace09b8cf2f16b4bf6aae33c7812aa1d9a8dbaeba5e5acc380e17f550ef

Observation 1ca72ff5-7a2d-4988-a803-3893dd27cb3c · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.952939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.952939Z digest=sha256:3a6432b87ededcd8d647067529b76e3837abed32638ada671ef1a9e2f4e842f4

Observation f21a32ea-98cc-4e5b-8b86-4fe1f2f6c4b3 · outbound

This paper cites Explainable Artificial Intelligence for Autonomous Driving: A Comprehensive Overview and Field Guide for Future Research Directions.

Explainability for Vision Foundation Models: A Survey Explainable Artificial Intelligence for Autonomous Driving: A Comprehensive Overview and Field Guide for Future Research Directions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.957481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.957481Z digest=sha256:8f570ddb87b01f2be42a2bbdb94e9be9f4d28bd886666b7f4ec70dfe5f1549a3

Observation 0128b40a-02a5-4da1-a9bc-8db95c0bce87 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.962914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.962914Z digest=sha256:eaeb014d1bd6f98a7e3520822d5d2d8c80c0deaab9faa21a6c589b87a0843f2b

Observation 87d71178-eaea-4d30-8648-939dda8b83fd · outbound

This paper cites Nauta, J.

Explainability for Vision Foundation Models: A Survey Nauta, J

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.967347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.967347Z digest=sha256:8b08c4f5386aa5d9c5e35071267bbc266194f24c148b3cf4399945e0ad697b5a

Observation 3f0c512e-7398-40b3-8f62-52347f24f8dd · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.971674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.971674Z digest=sha256:07c85edae123318eb7850a173cc286d7572bc03e10c221073b73f4057fe2fa57

Observation f0fc3180-7b36-4b10-bb54-654b0362ccd2 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.978687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.978687Z digest=sha256:bc9306335aa8e38ff4102777e37c5d5611d05f622a325da3849d8699a35736ed

Observation c31996d1-c0de-43b1-98e7-eaabb427e772 · outbound

This paper cites Schneider, C.

Explainability for Vision Foundation Models: A Survey Schneider, C

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.983963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.983963Z digest=sha256:b704818be254d187565b063301e80108b7e195bc284620d57036ba5dac81eb46

Observation 2d8ec2f4-c0d9-4a83-b78c-f00e833e1f46 · outbound

This paper cites Brown, B.

Explainability for Vision Foundation Models: A Survey Brown, B

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.988058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.988058Z digest=sha256:74d2a1d19047125719816b91dec7a6c2b49b73485b9cc3138c1d0331567c2502

Observation d12d149b-826f-431c-b82f-13b39d839410 · outbound

This paper cites Radford, J.

Explainability for Vision Foundation Models: A Survey Radford, J

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.992501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.992501Z digest=sha256:82cf9db4d7fd094210a097ceff2df8303de509abd4c5fed57aee11e42cdc121d

Observation 769841ca-cab1-4c04-a7dd-d8ca945d679d · outbound

This paper cites A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT.

Explainability for Vision Foundation Models: A Survey A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:34.997097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:34.997097Z digest=sha256:177de218c01948250c379552583e3d35c1274eb036585763264d3772a8b9bf1e

Observation fdf7cc2c-c601-4b36-accc-9b17e8727958 · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

Explainability for Vision Foundation Models: A Survey Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.002379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.002379Z digest=sha256:1c276fe047483350a605c52a591a111c7741b00ede363abdd23055f53fe48f82

Observation 8f10c224-e42b-45ea-848c-66b7917b461a · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.006952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.006952Z digest=sha256:497c0041021f3c4479681e19c819f4ae7ced7a478790b72931dd555463676fd1

Observation 9f313169-d584-41bf-97ee-ad27d3bf89ef · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Explainability for Vision Foundation Models: A Survey An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.010949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.010949Z digest=sha256:07715415d10a1fdf54477aa878827e37e2e19ff61566f8c0ef9ad825056286fb

Observation 9e9ff115-da21-42ce-b928-8011f56b4686 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.015524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.015524Z digest=sha256:4ee81be6466fec00b0e30e4a3bbeb91552c928bd7f27dcb1bc285162a1dd42ec

Observation 0548c984-9199-4ae6-a919-6be5df0f4033 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.020274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.020274Z digest=sha256:62444edbcdebe37584fda83a94edef8ff716fb78d9994fbde4f0d6cccd907030

Observation 3fa46a93-892f-4034-87e4-3495de5b246e · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.024473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.024473Z digest=sha256:b5b9844753df4603565ba868d434d2ed175621506ac39675eec4e155eb7533df

Observation 3120ef8e-bf74-4606-8bf7-fe3191826ad9 · outbound

This paper cites Vaswani, N.

Explainability for Vision Foundation Models: A Survey Vaswani, N

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.028472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.028472Z digest=sha256:123ae62aa7cc321704679ab9551f36edadc05c9fe884bae0236eeb6af12bfd7a

Observation d29db030-b8b5-44b4-b0d3-0bd491a98ed1 · outbound

This paper cites Radford, J.

Explainability for Vision Foundation Models: A Survey Radford, J

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.032346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.032346Z digest=sha256:88459a8dbebf93cec31e442e76e8632770dacce69bd111259acae6ae6a425288

Observation e6a0b460-5ce3-4fb0-a173-ef0faac3047b · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Explainability for Vision Foundation Models: A Survey BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.036667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.036667Z digest=sha256:9fb0ad83a3c3e53130bce5a71ae8852f0d2e6686ee31ec7152458a57fc082588

Observation 2b8e487c-3c2c-46a6-889c-9ce7a0a813f7 · outbound

This paper cites Contextualized Perturbation for Textual Adversarial Attack.

Explainability for Vision Foundation Models: A Survey Contextualized Perturbation for Textual Adversarial Attack

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.041674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.041674Z digest=sha256:8a03b99286557745ca744c97ef305ca18e32b582483fc6b40d5584939e18ad29

Observation 920212cc-b8a0-436d-b63b-017ace97bc0a · outbound

This paper cites Ravichander, E.

Explainability for Vision Foundation Models: A Survey Ravichander, E

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.045938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.045938Z digest=sha256:a52f58f7d57e28e959ebe02441bcb4d6ccf8e013eaa49e84da9cf7c7946643a6

Observation cf3cb909-dbdf-467b-8743-67ccb3a8f12e · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.049987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.049987Z digest=sha256:30b2bbbd8fdfb21834f01004e198c9af0137329d4f9359b005239547698a6479

Observation f3a617b3-9de3-486b-bb59-2d2eafc228bb · outbound

This paper cites Rombach, A.

Explainability for Vision Foundation Models: A Survey Rombach, A

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.053708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.053708Z digest=sha256:a0d4604100026bb87861f52401def3ebcb8e0e8384ebe688c284904fde2d94f3

Observation 555e3e48-d565-4fe8-b20b-715531536233 · outbound

This paper cites Girdhar, A.

Explainability for Vision Foundation Models: A Survey Girdhar, A

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.057600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.057600Z digest=sha256:e07ce7d022d4f476c068fc31e64568ec7531841f4f28ce5c370e43d5948f862c

Observation 1fcf16d8-73f4-4ceb-af60-d72cecd0bb96 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.061295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.061295Z digest=sha256:b21fb83e0e1b4df57fde8cf59b7d2bf6de31dd853f13b0bd8b2bcdd6d95d7426

Observation 07cfb794-61fc-4754-8253-86eb61e4ca34 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Explainability for Vision Foundation Models: A Survey LLaMA: Open and Efficient Foundation Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.065201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.065201Z digest=sha256:c63bd0191ffb341bdde2473a72ca8975b1dbf0cf17eece89f8dede519ce54389

Observation d97b1657-1c4e-4df7-a809-6c889fd55c55 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.069343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.069343Z digest=sha256:b0a23a752831691b1b7f02ca9e0f268441d9234ed37490fb32c5198b0d759d3e

Observation 70dabecb-2649-42e5-a86d-aa44bff57721 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Explainability for Vision Foundation Models: A Survey SAM 2: Segment Anything in Images and Videos

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.073150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.073150Z digest=sha256:392263364eeb8656eece7defcb089463e5852caf81aed0b4f6b995dd9efefd7c

Observation 25f68a8f-1687-438c-97a7-927a3ea4edad · outbound

This paper cites GPT-4 Technical Report.

Explainability for Vision Foundation Models: A Survey GPT-4 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.077265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.077265Z digest=sha256:1ed7d956338ebcdfd9d8c10c02a436c1e67575a27610b23eeaedafa932b006a2

Observation a589e0b0-381e-4f7f-a32a-e2c920529d15 · outbound

This paper cites Pixtral 12B.

Explainability for Vision Foundation Models: A Survey Pixtral 12B

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.081419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.081419Z digest=sha256:dd33813eeaf7661a85cdb7443f4da9b2304e00afe943ab2630e2c85286c7c4fe

Observation e19000d3-a64d-479d-b364-1b9e18c8023e · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Explainability for Vision Foundation Models: A Survey LoRA: Low-Rank Adaptation of Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.085821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.085821Z digest=sha256:814a53978fbcbb3fb0b7b3776c7bb7dcd4adb8b8b1738de20d09b3ce7b813583

Observation 2058707e-af94-432d-aca3-562b18d9737d · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.089910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.089910Z digest=sha256:eeb530d22ce4141875f9508d429ad93f19b625a931f055c33ac0b93b38618092

Observation 7ad12f30-98b8-4d85-a2f6-85d3ff953da5 · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

Explainability for Vision Foundation Models: A Survey CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.093910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.093910Z digest=sha256:9203712eed4eaf2edabdff7868bb1ee5b907a624d4873e2f7480f4b2c10c0450

Observation 4bf28198-3d49-42af-8d14-801f31707961 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Explainability for Vision Foundation Models: A Survey Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.098046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.098046Z digest=sha256:2322e8fdc1285a82c249ee521c9f5c39f5e70729fed35fe1f2cc60dc7f08eb06

Observation 83aff805-12a7-4d9e-b9a4-e0743b781883 · outbound

This paper cites Betker, G.

Explainability for Vision Foundation Models: A Survey Betker, G

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.102701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.102701Z digest=sha256:d97874cc14b5ecdbaf9f9459392be70985ccfa888bc3be9de788e6fa45f310c8

Observation c87f9875-9c01-4905-a7ab-c179fdc7560f · outbound

This paper cites Alayrac, J.

Explainability for Vision Foundation Models: A Survey Alayrac, J

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.106458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.106458Z digest=sha256:b72a0ab967cbf0cb7c52390b97bd8a33404b36ea51850d9da7a5dc33494a29fa

Observation 482cf4ef-e6bf-4121-bd7a-2263f6d991ba · outbound

This paper cites GIT: A Generative Image-to-text Transformer for Vision and Language.

Explainability for Vision Foundation Models: A Survey GIT: A Generative Image-to-text Transformer for Vision and Language

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.110455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.110455Z digest=sha256:1bb504ec5912343c8d48855d8f5d8c5c61a906b03c913418b597d0e2cbaba459

Observation 719480d7-51a8-48da-9413-1f8899e16b09 · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Explainability for Vision Foundation Models: A Survey GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.114520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.114520Z digest=sha256:490d71cd469da9c29f0ee4fb008deefabec7326aa9397334ec378651713cce2c

Observation e4e4b0c0-dab0-4ed2-b30e-cb672a151ecd · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

Explainability for Vision Foundation Models: A Survey LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.118293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.118293Z digest=sha256:68a6d55146b183b678d9069edc6fdb9c16745936a0a33e1c19ffb5b5a4471752

Observation bbd38988-0781-4d64-b3c6-4a3aa1875a1d · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Explainability for Vision Foundation Models: A Survey MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.122142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.122142Z digest=sha256:43ddaaace2e9c9fa76c06c4a994fa5731c0ea47bc4300565caf34d2e06578e4e

Observation 9c639bd1-1bea-4d36-a773-fc828cb791ee · outbound

This paper cites mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections.

Explainability for Vision Foundation Models: A Survey mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.126201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.126201Z digest=sha256:3fbfae0075f711d3402f7524fa9b4cf9a1a19c106a9ca276ba70a77258d5add0

Observation aa32a758-4344-41df-bc45-03844e3022c7 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.130163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.130163Z digest=sha256:64fbf924073c1b267608cbed5251739656bf2cf8b9ff6c9cfce79cac1ac50183

Observation c883fd49-9bc7-4e6e-91c4-f2a6eaba401b · outbound

This paper cites Segment Anything.

Explainability for Vision Foundation Models: A Survey Segment Anything

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.134406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.134406Z digest=sha256:76b344ecbf0c2114cea22fe3e365901e7a4c18c44c0d555f7d37fbd263299619

Observation 295232f7-4a71-44e8-8189-1db30573fcef · outbound

This paper cites STAIR: Learning Sparse Text and Image Representation in Grounded Tokens.

Explainability for Vision Foundation Models: A Survey STAIR: Learning Sparse Text and Image Representation in Grounded Tokens

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.139053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.139053Z digest=sha256:b1dff0c139bfa6c832bb2a6acbc7a9bafe860cb5a5bd21e9d5bd093d5522f4ac

Observation 454df2c3-e13a-4f27-9095-0512a2a9d126 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

Explainability for Vision Foundation Models: A Survey VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.143411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.143411Z digest=sha256:b1be4892435ce2b9c51ec5cf0e1b7e5fc3a4410d49dd49bf6d9fadc4cbdb49c6

Observation 7230cb2c-1b2a-4140-aca0-8982a301648c · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.147315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.147315Z digest=sha256:514e2af3158b3d3d91745127fad41fda69067bc2414b6e4b5fc3ddb455f026b7

Observation 3b4cb5ca-cbb4-4440-8e06-857e28f1afc1 · outbound

This paper cites Poeta, G.

Explainability for Vision Foundation Models: A Survey Poeta, G

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.151066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.151066Z digest=sha256:62fc66ac4445e4806c44409fa0a64f6ce051a2279f12ca0879f6ff197a2253a4

Observation 2bc1fcb8-2024-4609-be15-f4ec89d4ffd5 · outbound

This paper cites Galton, Regression towards mediocrity in hereditary stature., The Journal of the Anthropological Institute of Great Britain and Ireland 15 (1886) 246–263.

Explainability for Vision Foundation Models: A Survey Galton, Regression towards mediocrity in hereditary stature., The Journal of the Anthropological Institute of Great Britain and Ireland 15 (1886) 246–263

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.154633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.154633Z digest=sha256:70f5139d192d61835712ecc37db783551dbfcc621243eb5e68edeeff69fbd403

Observation 5edbb53f-28cf-47fc-a8ec-eff13b6562cc · outbound

This paper cites McCullagh, Generalized Linear Models, Routledge, 2019.

Explainability for Vision Foundation Models: A Survey McCullagh, Generalized Linear Models, Routledge, 2019

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.158578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.158578Z digest=sha256:f7afb2a244596a92e3e086f8f1296aba594dfcb1e62863b676d953d0893eabcc

Observation 5deb0687-4235-4899-bd69-ffbcb2267cee · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.162114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.162114Z digest=sha256:9f67714c90563f275d9fbcc20d25895b43fa694eea20a1e2602e042b6cd27d0b

Observation d4cf70c5-1ba7-4be5-bff1-1232558f29bb · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.166335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.166335Z digest=sha256:c1f1ee0fd1bbd6d19d8e10e50040f738a3efd8e4863cfcc04337256511b78378

Observation 6383f8e9-b9db-4af5-b067-ecf638dd76d8 · outbound

This paper cites Why should I trust you?.

Explainability for Vision Foundation Models: A Survey Why should I trust you?

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.170075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.170075Z digest=sha256:6117690a60b8c9e38ca206cff97c908503e5548b2ff77c16a975ec7c179c7812

Observation 32c59e4a-4257-4775-9e09-6f38ac6ea1bf · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.173922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.173922Z digest=sha256:bf6351765ddc39e1e9f62907fade2263a2c6f2228394c7d5ba6f40c51d277489

Observation 390b381d-66e5-411f-8ad2-f1e210b9b046 · outbound

This paper cites RISE: Randomized Input Sampling for Explanation of Black-box Models.

Explainability for Vision Foundation Models: A Survey RISE: Randomized Input Sampling for Explanation of Black-box Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.177756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.177756Z digest=sha256:c11dbc7624fb46ff5ca2a3c81f204912d10f2aa62208bc371f75326f78c6b728

Observation b512ad95-a53d-4a75-9cb8-2eee47b4e98e · outbound

This paper cites Cortez, M.

Explainability for Vision Foundation Models: A Survey Cortez, M

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.182281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.182281Z digest=sha256:55026f0fd592008a912bd85f0784d184c1e0683965fed3326f146dc0657b8008

Observation 9f5f01e0-9cb3-43fc-b6c5-e211dbe3883f · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.186645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.186645Z digest=sha256:5977e67815c0fd830031c23da466dd47c4a31022c19f3e2593431d4a5a76c849

Observation 91dfb81e-e5cc-484d-8566-7dee6563f5d7 · outbound

This paper cites A Survey of Explainable AI in Deep Visual Modeling: Methods and Metrics.

Explainability for Vision Foundation Models: A Survey A Survey of Explainable AI in Deep Visual Modeling: Methods and Metrics

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.190633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.190633Z digest=sha256:a11b981206627526b20e9d1f1c5b12ffe81b42b352898581a685081d84488e95

Observation 44bea955-7108-44cb-858d-c0fed13cae44 · outbound

This paper cites Chatzimparmpas, R.

Explainability for Vision Foundation Models: A Survey Chatzimparmpas, R

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.195066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.195066Z digest=sha256:3ef69d7b64180c8750a11cb7a0187a50b4be13b9e806d3496a74f858a63b9739

Observation daa70789-b4b7-4cd1-9577-833bb4c3a9cf · outbound

This paper cites Saeed, C.

Explainability for Vision Foundation Models: A Survey Saeed, C

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.199472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.199472Z digest=sha256:b5e25141078a685400c80b5ba907baae7ad68506fea5518581525edac511aa50

Observation 705f3d6f-1367-47ba-bdb9-c63dc63757f9 · outbound

This paper cites Joshi, R.

Explainability for Vision Foundation Models: A Survey Joshi, R

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.203432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.203432Z digest=sha256:e71174dfbcd553a87b73be092b209daa594ef25f929de3bd058ab33dae5087b0

Observation 83443531-c551-4b1d-981e-2b9b2d5b5e35 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.207628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.207628Z digest=sha256:5881b76b4e384b70937d2d53a28a9ff5339c73e06fe7f018c9cdae86fc618599

Observation dc4a66e8-6bb6-4e74-b8fb-f7e06c8caeb4 · outbound

This paper cites Fumanal-Idocin, J.

Explainability for Vision Foundation Models: A Survey Fumanal-Idocin, J

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.212127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.212127Z digest=sha256:1dff63f6f19014db1195e8e9bbbe9176dd62b68d538499d268a0c211f90a9bef

Observation db1be952-ae85-4a22-a6a6-4c094771dbc4 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.216496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.216496Z digest=sha256:4ed983fc2d7cb7cf984b1872ca51f6e3f7abd29673aa0791bb65bc9abbc47e72

Observation b4afb7d4-7d55-4602-b98b-f60926213888 · outbound

This paper cites BBA: Bi-Modal Behavioral Alignment for Reasoning with Large Vision-Language Models.

Explainability for Vision Foundation Models: A Survey BBA: Bi-Modal Behavioral Alignment for Reasoning with Large Vision-Language Models

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.221054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.221054Z digest=sha256:7da016b565c0ec8ca6838c5bb8074432e8abbd25a6df1a805d4300cb90af62e5

Observation d02218ba-3ea5-422b-bddb-452759153a36 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.226035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.226035Z digest=sha256:4ce89d84a0e88613c6a499276787ff2477575da7982bcf9fda9131f4138564df

Observation 8486b9cf-c9de-4b4b-870e-e1ab10a57773 · outbound

This paper cites Interpretable Measures of Conceptual Similarity by Complexity-Constrained Descriptive Auto-Encoding.

Explainability for Vision Foundation Models: A Survey Interpretable Measures of Conceptual Similarity by Complexity-Constrained Descriptive Auto-Encoding

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.230069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.230069Z digest=sha256:b765155471cdc77aef48ed648d0f67fa8d033e81a196bb660e5a23bc49576dfa

Observation c4983aa1-06fe-4094-b52a-c24ec1b3179d · outbound

This paper cites CEIR: Concept-based Explainable Image Representation Learning.

Explainability for Vision Foundation Models: A Survey CEIR: Concept-based Explainable Image Representation Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.234514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.234514Z digest=sha256:cd6c14c8f42c5f3c2770174d5c0b7e80c9798d531c07077fb6a63d8b06f56e6a

Observation 3ec2d2ba-34f3-4b3f-b466-7b35a7821735 · outbound

This paper cites A ChatGPT Aided Explainable Framework for Zero-Shot Medical Image Diagnosis.

Explainability for Vision Foundation Models: A Survey A ChatGPT Aided Explainable Framework for Zero-Shot Medical Image Diagnosis

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.238540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.238540Z digest=sha256:93d735b7a7b9eb2bbe5361faddc488959cdb9fc5dc812d01e23b3aa39c22412e

Observation cf3693bd-d310-4d0c-8bc4-be7a704293f9 · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.242622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.242622Z digest=sha256:345bd00fd8abce98094d679e85efd7594ca6a99293d83a15ee4d5a50f987bae3

Observation e2f55cd9-a2bd-44a8-951b-106a88d9ef31 · outbound

This paper cites ChartThinker: A Contextual Chain-of-Thought Approach to Optimized Chart Summarization.

Explainability for Vision Foundation Models: A Survey ChartThinker: A Contextual Chain-of-Thought Approach to Optimized Chart Summarization

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.246209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.246209Z digest=sha256:ccd03dd37887328e23710265152ec23c4d7b694c7be5f02c92ca748879793a51

Observation 3eb7d7e3-f695-4597-8c35-e5c3d09f0dd4 · outbound

This paper cites Menon, C.

Explainability for Vision Foundation Models: A Survey Menon, C

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.250132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.250132Z digest=sha256:eb8f2d20d4812996e715cd300b8b60c2444c9bbc1aad0d17ec2e30c08caeff02

Observation 561563b2-c3d9-4d07-9d87-8bb4e3d4a333 · outbound

This paper cites Kazmierczak, E.

Explainability for Vision Foundation Models: A Survey Kazmierczak, E

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.254195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.254195Z digest=sha256:198599f23125b700a7e8b9377d52d4e986f636937d17432486269d0ff0e02eda

Observation 896c6ce8-1c20-4fd9-944b-f689fac4e894 · outbound

This paper cites Zhang, M.

Explainability for Vision Foundation Models: A Survey Zhang, M

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.258183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.258183Z digest=sha256:5262156ec7f5d5b5754b1a4a95f67bbc7160a2f6e06a9f1e9728a96b062b5978

Observation 6a896251-b22a-42f0-8219-c85acc8b221d · outbound

This paper cites Mannix, H.

Explainability for Vision Foundation Models: A Survey Mannix, H

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.262278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.262278Z digest=sha256:6cf41fdaa5ebd5b11c6e81b5d6b73ecf1232abf65e0719f4b7ee4c67fac583d4

Observation 67fc934f-0aae-4a4f-8576-fff36db3ef09 · outbound

This paper cites Chain of Thought Prompt Tuning in Vision Language Models.

Explainability for Vision Foundation Models: A Survey Chain of Thought Prompt Tuning in Vision Language Models

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.266818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.266818Z digest=sha256:2f9518482599dac58ad51e1481d49862b1ab2889c12078110116bd6193e11588

Observation a3bcf6c7-b2df-46ac-83b2-5e894646abd9 · outbound

This paper cites Visual Chain-of-Thought Diffusion Models.

Explainability for Vision Foundation Models: A Survey Visual Chain-of-Thought Diffusion Models

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.271280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.271280Z digest=sha256:bcb0307a4bad52db75d6aceaf1c5aa3133d92dece0e41c7cabff45fa9877870c

Observation b749aef6-e960-46a8-bf24-d94775189842 · outbound

This paper cites Measuring and Improving Chain-of-Thought Reasoning in Vision-Language Models.

Explainability for Vision Foundation Models: A Survey Measuring and Improving Chain-of-Thought Reasoning in Vision-Language Models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.275311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.275311Z digest=sha256:f68aec40d6baa386c7192c85585391e6cbe8bf2605b4d174a55e8031a89a50e0

Observation 218f416b-964d-4515-b656-1b062d0b76dc · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.279998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.279998Z digest=sha256:a51bb0e5ed8612e25ac4e962a16b74f6157335379a120960499f98669933e220

Observation 851d70a2-fa92-4509-8074-094b9d1b4390 · outbound

This paper cites Zheng, B.

Explainability for Vision Foundation Models: A Survey Zheng, B

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.290447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.290447Z digest=sha256:0cde226b69d8fbfaed8f3291f170bb2e54e4d1e032de506d2d7f2b7a65c26daa

Observation 756597f3-94ec-4b0f-b1a6-63ad4bf507b9 · outbound

This paper cites DECap: Towards Generalized Explicit Caption Editing via Diffusion Mechanism.

Explainability for Vision Foundation Models: A Survey DECap: Towards Generalized Explicit Caption Editing via Diffusion Mechanism

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.294786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.294786Z digest=sha256:a158fcd2b3cf5cab5a1ab479646224b82f7332b2fea68e4e9f725803c64fb0ce

Observation 07006743-8521-4dce-ac6e-454c6ca94d2d · outbound

This paper cites DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving.

Explainability for Vision Foundation Models: A Survey DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous Driving

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.300656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.300656Z digest=sha256:fe34708e45bdcb4a09e39e6be099a854e13b7332a72810cc0449985413857b6e

Observation 62b868df-f555-45a3-b895-b314e31d031e · outbound

This paper cites Navigating Hallucinations for Reasoning of Unintentional Activities.

Explainability for Vision Foundation Models: A Survey Navigating Hallucinations for Reasoning of Unintentional Activities

Reference 93

Resolution
verified exact
local_arxiv, observed 2026-08-10T17:26:37.795294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:26:35.305725Z digest=sha256:5ce65f3fa8c37e7c392592f01a0d2329a9f1da5245226495e60a19ebd0caf61c

Observation 68bb0ec4-245a-4f04-8f36-e3f0ec66f53d · outbound

This paper cites Dolphins: Multimodal Language Model for Driving.

Explainability for Vision Foundation Models: A Survey Dolphins: Multimodal Language Model for Driving

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.310664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.310664Z digest=sha256:36d510509548dd3d9734bcb46440c0be640b033fafe1668bf4feb02b0015f7eb

Observation 997f4aee-a7be-4335-938d-f5033ac03528 · outbound

This paper cites DriveGPT4: Interpretable End-to-end Autonomous Driving via Large Language Model.

Explainability for Vision Foundation Models: A Survey DriveGPT4: Interpretable End-to-end Autonomous Driving via Large Language Model

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.315329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.315329Z digest=sha256:21147fb30c74defc8b65d0e5b795119fe5519dbfc99604f7d5c112762a4450ef

Observation 101631e3-0413-4f93-aa71-b3eb0cea730d · outbound

This paper cites Image Translation as Diffusion Visual Programmers.

Explainability for Vision Foundation Models: A Survey Image Translation as Diffusion Visual Programmers

Reference 96

Resolution
verified exact
local_arxiv, observed 2026-08-10T17:26:37.748931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:26:35.320253Z digest=sha256:329304e96d933fc385254a8b3ab5f7a1139069678ad6bed9071d1269d76ed8df

Observation 631fe275-96d0-4ce8-9892-b5dd8961c0b0 · outbound

This paper cites Zhang, J.

Explainability for Vision Foundation Models: A Survey Zhang, J

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.324493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.324493Z digest=sha256:ae32d11d8ad8d75706ea5ecc8ee990a8a5ed6f264ad100e13ec33b26e0e13be6

Observation e19e70fa-14b6-4350-a152-75f3fe342b62 · outbound

This paper cites Multimodal and Explainable Internet Meme Classification.

Explainability for Vision Foundation Models: A Survey Multimodal and Explainable Internet Meme Classification

Reference 98

Resolution
verified exact
local_arxiv, observed 2026-08-10T17:26:37.731626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:26:35.328474Z digest=sha256:2cce2d71db6fbfa8cd5667617eaf93c6f067ba4a6e2df51604a2fbc58b40d6a6

Observation d98aaf30-6ffc-4bde-9035-18895bfd2bcf · outbound

This paper cites Advancing Large Multi-modal Models with Explicit Chain-of-Reasoning and Visual Question Generation.

Explainability for Vision Foundation Models: A Survey Advancing Large Multi-modal Models with Explicit Chain-of-Reasoning and Visual Question Generation

Reference 99

Resolution
verified exact
local_arxiv, observed 2026-08-10T17:26:37.713888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:26:35.332816Z digest=sha256:12cb6bd4caf0e5256e715a96c37362c087b31288f2cf0187c0608d08499f46cd

Observation b2ec6400-d0d2-4e4b-8063-0167203945b4 · outbound

This paper cites ExTraCT -- Explainable Trajectory Corrections from language inputs using Textual description of features.

Explainability for Vision Foundation Models: A Survey ExTraCT -- Explainable Trajectory Corrections from language inputs using Textual description of features

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.337370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.337370Z digest=sha256:47c8a5bb55e08cf65f411e5d38dc7c62079954986ae79ddf1c53978c071328d2

Observation cff610bb-6617-4b46-a85a-e16ea0223e5f · outbound

This paper cites an unresolved cited work.

Explainability for Vision Foundation Models: A Survey Unresolved cited work

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.341966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.341966Z digest=sha256:0bb9ab6d4d91c509b2162ca42c37a8efdffdc0fcf45314129615a4c880b5c426

Pith citing papers

Observation 016e3ed6-874c-47b5-8b5f-d8d8bdf2198a · inbound

Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs cites this paper.

Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs Explainability for Vision Foundation Models: A Survey

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:22:35.283207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:22:35.055073Z digest=sha256:09c48c87843d618ed64024c0848460c6c9ab7dea36e3b3d003249e774aba4ed6