Pith. sign in

Paper Citation Record · LEDGER

Zero-Shot Vision Encoder Grafting via LLM Surrogates

As of 9 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2505.22664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22664 v2

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:10:11.740399Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy37
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 260503fb-70c9-419e-9346-d45496269371 · outbound

This paper cites Understanding Inter- mediate Layers Using Linear Classifier Probes.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Understanding Inter- mediate Layers Using Linear Classifier Probes

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:24.280642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:03.570614Z digest=sha256:57557bdcb2b3779d2672aa8a6c918c7062eef17b7ecfacc3da5800441a1e916b

Observation a5d9bd1d-3cfd-476b-8415-2836dd75b673 · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Flamingo: a Visual Language Model for Few-Shot Learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.894969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:03.714649Z digest=sha256:adbde3ca8b53ccd809d9163906f56e066902e1aaae8adbaaa75c36658aebf3e5

Observation 9544eb4d-59cd-40dc-a31a-a6342a5cd088 · outbound

This paper cites Eliciting Latent Predictions from Transformers with the Tuned Lens.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Eliciting Latent Predictions from Transformers with the Tuned Lens

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:03.928054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:03.928054Z digest=sha256:8f5b23424291c862eeed0c5b51c849a9abe85f580c559bc82d7bd16305eaf6ed

Observation 362a970d-0fc7-457c-9579-d940edc8e5a0 · outbound

This paper cites PIQA: Reasoning about Physical Commonsense in Natural Language.

Zero-Shot Vision Encoder Grafting via LLM Surrogates PIQA: Reasoning about Physical Commonsense in Natural Language

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.407238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:04.074749Z digest=sha256:7ffa4e497f33dd871bdbe3fc858854d30a7f25a0c62b8d8b5580991fceab12be

Observation 0038ccda-7977-4d22-928b-75fbc3ad31fe · outbound

This paper cites GenQA: Generating Millions of Instructions from a Handful of Prompts.

Zero-Shot Vision Encoder Grafting via LLM Surrogates GenQA: Generating Millions of Instructions from a Handful of Prompts

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.258432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.258432Z digest=sha256:712b1e14be8e112df6c28e6249f0be59caa4efb4ede26737b3ff4a0a16442201

Observation 23ecf2b9-8082-4abd-9ce3-97b81ab43e7a · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

Zero-Shot Vision Encoder Grafting via LLM Surrogates InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.087790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:04.453895Z digest=sha256:20b4d444f3c3b3221d96b0a845aa9e3ec9500ad1ca396c81534f66ee6d1432c9

Observation 15b07d9f-64ed-41f7-9924-daad416ba5bc · outbound

This paper cites BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions.

Zero-Shot Vision Encoder Grafting via LLM Surrogates BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.825369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:04.575468Z digest=sha256:0e79db55c6e351b5b996c7f1978f45b7748a21655b003121f12eb75d7ec35552

Observation d5b5b0a8-89d3-4666-8664-009a8d53eee2 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.718039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.718039Z digest=sha256:88fa47f968e8988161b8633308bd5dd28f9b4f63608d4db611e753555f37a3ac

Observation 01fa7e38-1e5d-4ef9-93a5-901f89510876 · outbound

This paper cites The Llama 3 Herd of Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.898522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.898522Z digest=sha256:9fa52b05e6a1dc3bea04ee0d520e75f4114536f27a2344893eac06c01338e135

Observation 5b4a33fc-eb7a-4283-b220-365723fedb31 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.974166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.974166Z digest=sha256:55eac41af4f9761251ff374323dac34c80b10b0065f6df13ed8c2c85ce325982

Observation 551b002b-ce61-45cd-b6e5-ac28a7b6aded · outbound

This paper cites A Framework for Few-Shot Language Model Evaluation, 2024.

Zero-Shot Vision Encoder Grafting via LLM Surrogates A Framework for Few-Shot Language Model Evaluation, 2024

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.629018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:05.094039Z digest=sha256:0d167f6afffe527a1056c1d0cb2aea6a7e4b7251b775c0f26de012e239878d0e

Observation d1a63bc4-cf48-4255-bdb5-778f03175f2e · outbound

This paper cites The Unreasonable Ineffectiveness of the Deeper Layers.

Zero-Shot Vision Encoder Grafting via LLM Surrogates The Unreasonable Ineffectiveness of the Deeper Layers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.422088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:05.218644Z digest=sha256:5d09538aba2485a4e05f111ec3867c183c11acd655f7dbb44be4b12f0d2866d2

Observation b946c6a1-1a01-4cf3-a6f8-e9132388e955 · outbound

This paper cites VizWiz Grand Challenge: Answering Visual Questions from Blind People.

Zero-Shot Vision Encoder Grafting via LLM Surrogates VizWiz Grand Challenge: Answering Visual Questions from Blind People

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.118148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:05.387611Z digest=sha256:bb8d8dd0272772355cffcfeefa6577daf103cc844249ca1f52201464732efda8

Observation b82948d8-89bc-4133-8fea-961338574cef · outbound

This paper cites Word Embed- dings Are Steers for Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Word Embed- dings Are Steers for Language Models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.805678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:05.521073Z digest=sha256:5a40f3f91b05486f7f3383a6cfb903e3ef8c25c12d03c1d6e1b418d3d386ac15

Observation 59cedae0-882b-4c12-8981-88ae6f07fd93 · outbound

This paper cites Measur- ing Massive Multitask Language Understanding.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Measur- ing Massive Multitask Language Understanding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.475285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:05.648401Z digest=sha256:15d5291a3eff92d4df0053b9c5449cc9fbf0a742cd0557be34a7145da17373f7

Observation 23bde0c7-c291-4500-a284-599a88ae5017 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LoRA: Low-Rank Adaptation of Large Language Models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.140630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:05.846778Z digest=sha256:6c901bac943a61d4ca068c70e6afd1d3f8988086e195cd311d77c2772f63061e

Observation 649794ba-112a-4f81-bc14-897e75708ce8 · outbound

This paper cites GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering.

Zero-Shot Vision Encoder Grafting via LLM Surrogates GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.845832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:06.057813Z digest=sha256:e575c1d1b3d846af53fdc115298d7972305e77dc64641f72185f4efb24e1685a

Observation eb6c8812-05c5-41e1-95f8-1d5eee1132ef · outbound

This paper cites InternVL2: Better than the Best—Expanding Performance Boundaries of Open-Source Multimodal Mod- els with the Progressive Scaling Strategy.

Zero-Shot Vision Encoder Grafting via LLM Surrogates InternVL2: Better than the Best—Expanding Performance Boundaries of Open-Source Multimodal Mod- els with the Progressive Scaling Strategy

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.536709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:06.173724Z digest=sha256:e821974dbb6ee7809c8e520e7c9ea4c2faca53883bac66b262d1274a4cfd30af

Observation c7be7e24-b793-42cf-a693-fd05281938c6 · outbound

This paper cites A Diagram Is Worth a Dozen Images.

Zero-Shot Vision Encoder Grafting via LLM Surrogates A Diagram Is Worth a Dozen Images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.238302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:06.293534Z digest=sha256:caf7f42bdf70322e5434cb14c74661ea70f10d5519aa7eec13bee5e26bfa515b

Observation 51e4d9b8-f8e7-47e3-85ef-fb41ceeaaa29 · outbound

This paper cites Propulsion: Steering LLM with Tiny Fine-Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Propulsion: Steering LLM with Tiny Fine-Tuning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.880653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:06.462436Z digest=sha256:f706684b1c5c22e07382267027207b9be775545a36476e724b4b53d0e9677302

Observation dc51ade7-865f-494a-813a-e740ae092e60 · outbound

This paper cites SEED-Bench: Benchmarking Multi- modal LLMs with Generative Comprehension.

Zero-Shot Vision Encoder Grafting via LLM Surrogates SEED-Bench: Benchmarking Multi- modal LLMs with Generative Comprehension

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.640428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:06.605423Z digest=sha256:6f8fa430d62562c75bc14b15caaff282847801d5eb7b1edf2da400e03e62249a

Observation ef148f84-4a47-4d82-8ed6-8b250a7c0ea5 · outbound

This paper cites LLaV A-Next: Stronger Llms Supercharge Multimodal Capabilities in the Wild.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaV A-Next: Stronger Llms Supercharge Multimodal Capabilities in the Wild

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.339366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:06.744498Z digest=sha256:fa721953064c5c35d486e071f91638c9cfe3ccccaa68fe0877a7519ce8fc43a6

Observation 7832c4fa-c5e8-4888-b070-8974240f4229 · outbound

This paper cites LMMs-Eval: Accelerating the Development of Large Multimodal Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LMMs-Eval: Accelerating the Development of Large Multimodal Models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.028093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:06.928443Z digest=sha256:4553b80adcf3e5d6d2815c1de43ef4dabbc6713b670a98647b18295c4459fd7c

Observation c0410ddb-23dc-45dd-848c-d1419c58cf19 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaVA-OneVision: Easy Visual Task Transfer

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.072452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:07.072452Z digest=sha256:1c8b0b1290361298303ed46cb1cb74c9ea628c2bcf93394acc1808bdd6f4513f

Observation 0755a51f-eaac-4b00-9b4c-3239e119c8f5 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.689949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:07.202405Z digest=sha256:f628a8e0ce47a2c2fdfa7cb2c7cc4875c07667c9d298a5079fb867642b3a9c30

Observation 3e3532c1-4f67-4bb3-a78f-3d463911b73a · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Evaluating Object Hallucination in Large Vision-Language Models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.346769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:07.335320Z digest=sha256:dc898cf3dd86fa2dcd5391e7f2aa23f0548e99bb4c1dd7b757bad48b796b536c

Observation c9b6ce0f-78ed-4883-bbb3-dec8e561454a · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Improved Baselines with Visual Instruction Tuning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.009253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:07.486743Z digest=sha256:62ae88b90c9cc1c690de45886678630e313c9eca46cd0318b93387b5e23de40b

Observation 9dd22051-41b2-4fad-9ab1-324a36a2e708 · outbound

This paper cites LLaV A-Next: Im- proved Reasoning, Ocr, and World Knowledge.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaV A-Next: Im- proved Reasoning, Ocr, and World Knowledge

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.719998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:07.663602Z digest=sha256:8e234618447e2a81aaa2924ba6364890145d839882866cb9cf19b9366ec725e1

Observation b193148e-e3df-4f21-a0fd-15b2cdf9772e · outbound

This paper cites Visual Instruction Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Visual Instruction Tuning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.431218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:07.830359Z digest=sha256:9d2a4b8ab54e38c92d43d22dd179f9786e75dc80fd305ed8f6f0f4553e4c86ba

Observation c245efa1-3580-47e7-ba77-2bcc3340a6bd · outbound

This paper cites MMBench: Is Your Multi-modal Model an All-around Player? In ECCV, 2025.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MMBench: Is Your Multi-modal Model an All-around Player? In ECCV, 2025

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.038114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:07.935523Z digest=sha256:8f78bdb355728f5ba2018203b4b9e5a74936d63a8ca46ab0bdd336b4f74c6aea

Observation 54401c49-46c3-4af6-b992-1fdc4d582771 · outbound

This paper cites Chartqa: A Benchmark for Question Answering About Charts With Visual and Logical Reason- ing.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Chartqa: A Benchmark for Question Answering About Charts With Visual and Logical Reason- ing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.743903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:08.060825Z digest=sha256:7adbb21216935b372fb1b4b30d68e4f0f7480d8bfcf8c6f710c8d2d739ca71d3

Observation 4c64e431-059d-4041-88f4-0bf1cb177edd · outbound

This paper cites Docvqa: A Dataset for VQA on Document Images.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Docvqa: A Dataset for VQA on Document Images

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.373407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:08.282593Z digest=sha256:e71009fd3cc291598f8ff220052fa0f47e471091acd095b9ef2cb87a071b13a8

Observation a9d22fad-abc7-49d9-810f-a284bbfe4824 · outbound

This paper cites Infograph- icVQA.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Infograph- icVQA

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.063658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:08.400143Z digest=sha256:b1fca3d34229b191bba37dfbcab201044cf519431c96dacc37fac564cb265557

Observation b2bdd531-cbc6-4173-bb5f-290a138497d0 · outbound

This paper cites ShortGPT: Layers in Large Language Models are More Redundant Than You Expect.

Zero-Shot Vision Encoder Grafting via LLM Surrogates ShortGPT: Layers in Large Language Models are More Redundant Than You Expect

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:08.563542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:08.563542Z digest=sha256:3fc9c37724825058c4848cc2d0c6ccd52645cbca0f7a45c749e14b0b520e9406

Observation 6cdc855b-7bfd-4691-801c-425a5f9c46f5 · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.729807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:08.737225Z digest=sha256:aff5a30b70de4db971f7d559e7a9a751e2fd8adba9bce2b03983057cf27327b5

Observation 232146ff-0a88-4f50-855c-fb474f1916d3 · outbound

This paper cites interpreting GPT: the logit lens.

Zero-Shot Vision Encoder Grafting via LLM Surrogates interpreting GPT: the logit lens

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.404758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:08.896020Z digest=sha256:0b83d76cc80b8e92d6fb6bb442307206a955b6cf84a2418d97fca1bc3c81500b

Observation 70fe4b8e-5a8c-4c2e-ac06-9159051c07fa · outbound

This paper cites Learning Transferable Visual Models From Natural Language Super- vision.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Learning Transferable Visual Models From Natural Language Super- vision

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.037846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:09.084874Z digest=sha256:a27f6a2af226849a5734a5c499db265f61df9a8db09afb900398df7d2a6df2e5

Observation a54574e4-0195-4a87-b106-2c4102115b3f · outbound

This paper cites WinoGrande: An Adversarial Winograd Schema Challenge at Scale.

Zero-Shot Vision Encoder Grafting via LLM Surrogates WinoGrande: An Adversarial Winograd Schema Challenge at Scale

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:14.688023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:09.261308Z digest=sha256:b5bcf308811656348c17a6bcffbd0f805e4b8159c4f63cf9ffbc869b4def9280

Observation 919f7776-77bc-4d98-9e36-159147c403cf · outbound

This paper cites Open Problems in Mechanistic Interpretability.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Open Problems in Mechanistic Interpretability

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:09.455799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:09.455799Z digest=sha256:ee77f24348e726711bb5d75a24f5c9eb2aa6fad52bb6bb15a090460a0bcd5e81

Observation 78f61f07-327e-420c-a7aa-9e2c43699f45 · outbound

This paper cites Towards VQA Models That Can Read.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Towards VQA Models That Can Read

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:14.377820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:09.639043Z digest=sha256:b7f12742ccde11e456962acce912d1d579c20dcaf2ee9c5fb34ae3c8430f6483

Observation ae21e43a-ad87-459a-96a4-d61fe2ab610b · outbound

This paper cites Transformer Layers as Painters.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Transformer Layers as Painters

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:09.868806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:09.868806Z digest=sha256:1862c3bcb43b4450983829f513ab45da6f47361f28ee5dca79375d1ba9bb5f6b

Observation c2eeae5e-c64c-4045-9a0d-fe116ec56739 · outbound

This paper cites an unresolved cited work.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:14.074467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:10.023098Z digest=sha256:96bf6f354aa2595367e9748731dc0202bd4544525787b05466d11b6d7e7af6d4

Observation 8b80db9e-b15a-49f8-9293-6e6b9916aaad · outbound

This paper cites an unresolved cited work.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:13.774364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:10.154755Z digest=sha256:afa101f50fc173d24469fe15fee3e8a529c9d14dc6c9376305edd5597e606e2c

Observation a5233e1d-bf72-455f-9bde-cbcffb59e415 · outbound

This paper cites Qwen2.5 Technical Report.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen2.5 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.316669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.316669Z digest=sha256:2d460b2e7f80da8fa42a438f9c748d4fc5ea9244104856bbf60a5e178fdc35be

Observation 8cac2992-8ccd-4fe4-b59d-3433f20693fb · outbound

This paper cites Qwen3 Technical Report.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen3 Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.411364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.411364Z digest=sha256:08173e71e416a9d933ce60034fa0c45e3129424888f5df39d6fde9a8399b4e9a

Observation 9a9ab4b5-bf59-4450-99bf-fbc4ddfbb156 · outbound

This paper cites Cambrian- 1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Cambrian- 1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:13.491774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:10.548879Z digest=sha256:75f6f6da72a9cd5726e63c9db99f09d48fe3264fb34d0e206dd8e3021fd6eaad

Observation bbd238ab-8814-446f-beb0-b6b14efeb168 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Zero-Shot Vision Encoder Grafting via LLM Surrogates SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.711035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.711035Z digest=sha256:31f1a5d9a217ef51c5aa4c3b372160a26e559b20076588ff728ca02bb7836c2d

Observation 02bcb860-639e-41f6-b365-d8f131348935 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.917149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.917149Z digest=sha256:6c99d89d0bf887fb1ecce5007aa5e6ae9435ba72754296e1fe39cffef2c92ed5

Observation 90da24bf-cbc8-4556-ad96-05125bcc443d · outbound

This paper cites CogVLM: Visual Expert for Pretrained Lan- guage Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates CogVLM: Visual Expert for Pretrained Lan- guage Models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:13.221500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:11.082351Z digest=sha256:68505ff5b2bdf810b1ade4a696d3efdf612a22eed072f0f6ad20f37e3d40e5f9

Observation 5a1b519c-ef56-4100-9df7-56f6a70460de · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Inte- grated Capabilities.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MM-Vet: Evaluating Large Multimodal Models for Inte- grated Capabilities

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.806130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:11.239273Z digest=sha256:8775055844c79bfb7e48259891e0f1b93a7c9fa4efca4beb0e50de0e6a99631d

Observation 9f4c83da-0827-4fb2-b03c-2602aeeaa9b4 · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence? In ACL Anthology, 2019.

Zero-Shot Vision Encoder Grafting via LLM Surrogates HellaSwag: Can a Machine Really Finish Your Sentence? In ACL Anthology, 2019

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.491448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:11.437478Z digest=sha256:9025c8d782b16ea403f798cc2fa429804d845303855cea05a83f872a89b156f7

Observation 652f85e5-fdfb-400c-be5a-0ee75c053fec · outbound

This paper cites Sigmoid Loss for Language Image Pre- Training.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Sigmoid Loss for Language Image Pre- Training

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.138345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T13:10:11.570195Z digest=sha256:e2d918cbabd464df80ae6232116377d91e6712c3a13b122e42afe807481aa658

Observation f1dca64a-039e-41c8-b879-848cd520fce3 · outbound

This paper cites PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel.

Zero-Shot Vision Encoder Grafting via LLM Surrogates PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:11.740399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:11.740399Z digest=sha256:ab2e6a06293ca50487c913a816979c57749654f018f8f2b761119ff09396b444

Pith citing papers

No inbound Pith citation observations are available.