Pith. sign in

Paper Citation Record · LEDGER

Zero-Shot Vision Encoder Grafting via LLM Surrogates

As of 16 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2505.22664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22664 v2

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:10:11.740399Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy37
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 260503fb-70c9-419e-9346-d45496269371 · outbound

This paper cites Understanding Inter- mediate Layers Using Linear Classifier Probes.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Understanding Inter- mediate Layers Using Linear Classifier Probes

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:24.280642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:03.570614Z digest=sha256:1b230cca2e85379bfee35ca355827c15d9fa6f08f067d193bf8cb90892f3e757

Observation a5d9bd1d-3cfd-476b-8415-2836dd75b673 · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Flamingo: a Visual Language Model for Few-Shot Learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.894969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:03.714649Z digest=sha256:48e7b27010d2e609cdb7f339c1189efaed83d96502faf3022bdb6487fa69bfd7

Observation 9544eb4d-59cd-40dc-a31a-a6342a5cd088 · outbound

This paper cites Eliciting Latent Predictions from Transformers with the Tuned Lens.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Eliciting Latent Predictions from Transformers with the Tuned Lens

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:03.928054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:03.928054Z digest=sha256:580339fd0505ee7a2369e242947123f481f1c167b3d0a3654e3755b8ed346205

Observation 362a970d-0fc7-457c-9579-d940edc8e5a0 · outbound

This paper cites PIQA: Reasoning about Physical Commonsense in Natural Language.

Zero-Shot Vision Encoder Grafting via LLM Surrogates PIQA: Reasoning about Physical Commonsense in Natural Language

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.407238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:04.074749Z digest=sha256:446523b79ef0780476e18d2b2240b94549ed94eda79e97b3cc010e44475169bd

Observation 0038ccda-7977-4d22-928b-75fbc3ad31fe · outbound

This paper cites GenQA: Generating Millions of Instructions from a Handful of Prompts.

Zero-Shot Vision Encoder Grafting via LLM Surrogates GenQA: Generating Millions of Instructions from a Handful of Prompts

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.258432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.258432Z digest=sha256:88137b4bb49ff9777beb156e43c2cce4771dc7cdc0d0e0115447b31705302580

Observation 23ecf2b9-8082-4abd-9ce3-97b81ab43e7a · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

Zero-Shot Vision Encoder Grafting via LLM Surrogates InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:23.087790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:04.453895Z digest=sha256:f492ca6af1a1d731adaacbdc75bf1d6aa7d5ff5eb8b301d956c4bb93ac1dcae9

Observation 15b07d9f-64ed-41f7-9924-daad416ba5bc · outbound

This paper cites BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions.

Zero-Shot Vision Encoder Grafting via LLM Surrogates BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.825369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:04.575468Z digest=sha256:d6d7dcd64c2381628a7931c309671a5025ad840f82e7ba3087c1740fa56b9df7

Observation d5b5b0a8-89d3-4666-8664-009a8d53eee2 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.718039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.718039Z digest=sha256:6e0095537d5367141cda41004e135b4a98faa2b0fd6a121971f40d55e5582902

Observation 01fa7e38-1e5d-4ef9-93a5-901f89510876 · outbound

This paper cites The Llama 3 Herd of Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.898522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.898522Z digest=sha256:b59b7fd48474340781fa43bad56ac86e780368e6672cfc7c1f3668551d5d910e

Observation 5b4a33fc-eb7a-4283-b220-365723fedb31 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.974166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:04.974166Z digest=sha256:2420bc51f669819dfc6f62c87944a5bab049179d6d6cd362d714c3e7fd325a03

Observation 551b002b-ce61-45cd-b6e5-ac28a7b6aded · outbound

This paper cites A Framework for Few-Shot Language Model Evaluation, 2024.

Zero-Shot Vision Encoder Grafting via LLM Surrogates A Framework for Few-Shot Language Model Evaluation, 2024

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.629018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:05.094039Z digest=sha256:ffea289853db9107afc7f9d3afcff3dd8686ac9873192af2b47156a263263ec9

Observation d1a63bc4-cf48-4255-bdb5-778f03175f2e · outbound

This paper cites The Unreasonable Ineffectiveness of the Deeper Layers.

Zero-Shot Vision Encoder Grafting via LLM Surrogates The Unreasonable Ineffectiveness of the Deeper Layers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.422088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:05.218644Z digest=sha256:f2747e01007f9e99f5f979fc0f9c5c2e3af469bcbe3ef7ae1252756cb5fb48d1

Observation b946c6a1-1a01-4cf3-a6f8-e9132388e955 · outbound

This paper cites VizWiz Grand Challenge: Answering Visual Questions from Blind People.

Zero-Shot Vision Encoder Grafting via LLM Surrogates VizWiz Grand Challenge: Answering Visual Questions from Blind People

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:22.118148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:05.387611Z digest=sha256:b1d900af9554f591a9cbcc0b53945f4e27c69dbdf5f1060064fa45bff2ae1915

Observation b82948d8-89bc-4133-8fea-961338574cef · outbound

This paper cites Word Embed- dings Are Steers for Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Word Embed- dings Are Steers for Language Models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.805678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:05.521073Z digest=sha256:7be9314ab7ccd1a98fc9f9399f3a6c7b9d8169f96ad8fb46ce42fa38a4499658

Observation 59cedae0-882b-4c12-8981-88ae6f07fd93 · outbound

This paper cites Measur- ing Massive Multitask Language Understanding.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Measur- ing Massive Multitask Language Understanding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.475285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:05.648401Z digest=sha256:58855b6ed404177e6439512613a2237bc6b098a6925e70a1401b2f9f3277b29e

Observation 23bde0c7-c291-4500-a284-599a88ae5017 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LoRA: Low-Rank Adaptation of Large Language Models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:21.140630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:05.846778Z digest=sha256:1b1fc7ffbdb06d94c41f2c73dbbf68f0e6604d5c1fbf3c1baf07cdc793fb22c3

Observation 649794ba-112a-4f81-bc14-897e75708ce8 · outbound

This paper cites GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering.

Zero-Shot Vision Encoder Grafting via LLM Surrogates GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.845832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:06.057813Z digest=sha256:7fb6eafef8c5767fd6b35bfdcd5142513ae7004a6afc4b11a138b148b33128a4

Observation eb6c8812-05c5-41e1-95f8-1d5eee1132ef · outbound

This paper cites InternVL2: Better than the Best—Expanding Performance Boundaries of Open-Source Multimodal Mod- els with the Progressive Scaling Strategy.

Zero-Shot Vision Encoder Grafting via LLM Surrogates InternVL2: Better than the Best—Expanding Performance Boundaries of Open-Source Multimodal Mod- els with the Progressive Scaling Strategy

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.536709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:06.173724Z digest=sha256:833768cc2407aa593f00fda9876ea18c6d975b256df6dadeaf1d2fe6c0ff320a

Observation c7be7e24-b793-42cf-a693-fd05281938c6 · outbound

This paper cites A Diagram Is Worth a Dozen Images.

Zero-Shot Vision Encoder Grafting via LLM Surrogates A Diagram Is Worth a Dozen Images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:20.238302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:06.293534Z digest=sha256:00daa186b76c6152b11d795fd94b94c211c4fcf2780014726075b1eb1ec33d58

Observation 51e4d9b8-f8e7-47e3-85ef-fb41ceeaaa29 · outbound

This paper cites Propulsion: Steering LLM with Tiny Fine-Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Propulsion: Steering LLM with Tiny Fine-Tuning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.880653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:06.462436Z digest=sha256:e69790415b3e79901b693eca38cfef843320044009654c07a0b6fbeb40c12d95

Observation dc51ade7-865f-494a-813a-e740ae092e60 · outbound

This paper cites SEED-Bench: Benchmarking Multi- modal LLMs with Generative Comprehension.

Zero-Shot Vision Encoder Grafting via LLM Surrogates SEED-Bench: Benchmarking Multi- modal LLMs with Generative Comprehension

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.640428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:06.605423Z digest=sha256:c0c79e0e153fecf93b6af6d1d6841d7b57f1403e05ea4185c7bce4c36fe51290

Observation ef148f84-4a47-4d82-8ed6-8b250a7c0ea5 · outbound

This paper cites LLaV A-Next: Stronger Llms Supercharge Multimodal Capabilities in the Wild.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaV A-Next: Stronger Llms Supercharge Multimodal Capabilities in the Wild

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.339366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:06.744498Z digest=sha256:bd2ecd919ba908356f94cf501d2de99677575488d793a30c598dfccd7643354d

Observation 7832c4fa-c5e8-4888-b070-8974240f4229 · outbound

This paper cites LMMs-Eval: Accelerating the Development of Large Multimodal Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LMMs-Eval: Accelerating the Development of Large Multimodal Models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:19.028093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:06.928443Z digest=sha256:61d15b46ba6147ca33a21d7f0ec3468c4774b8adaea248a6a1b1cb6f0d5e0c6c

Observation c0410ddb-23dc-45dd-848c-d1419c58cf19 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaVA-OneVision: Easy Visual Task Transfer

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.072452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:07.072452Z digest=sha256:df2daae28889a9a9d41a90d410c9f446a50ac75f15696b89bc040853e4d023e5

Observation 0755a51f-eaac-4b00-9b4c-3239e119c8f5 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.689949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:07.202405Z digest=sha256:098be17422a9f290c3f7bb62cecfdee7c46e7f8ed75984a218dc85e381e35059

Observation 3e3532c1-4f67-4bb3-a78f-3d463911b73a · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Evaluating Object Hallucination in Large Vision-Language Models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.346769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:07.335320Z digest=sha256:65df30aae611d1811b6cd2be9525b00cf0300ce9c0114cae0d19b41d85741ffd

Observation c9b6ce0f-78ed-4883-bbb3-dec8e561454a · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Improved Baselines with Visual Instruction Tuning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:18.009253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:07.486743Z digest=sha256:1615596081d05951b4e0899db259dbbcef7a5bf7be9cc5bf101b0fe6cd2cb6b8

Observation 9dd22051-41b2-4fad-9ab1-324a36a2e708 · outbound

This paper cites LLaV A-Next: Im- proved Reasoning, Ocr, and World Knowledge.

Zero-Shot Vision Encoder Grafting via LLM Surrogates LLaV A-Next: Im- proved Reasoning, Ocr, and World Knowledge

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.719998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:07.663602Z digest=sha256:bba4c1544e6baff87b6f9b4a3f8312f43865832dd91284beb7f272249394b15f

Observation b193148e-e3df-4f21-a0fd-15b2cdf9772e · outbound

This paper cites Visual Instruction Tuning.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Visual Instruction Tuning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.431218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:07.830359Z digest=sha256:c9dee3e6996a2ded114ebc7950f677cc4338cf15300731e447f8674403eb0eaa

Observation c245efa1-3580-47e7-ba77-2bcc3340a6bd · outbound

This paper cites MMBench: Is Your Multi-modal Model an All-around Player? In ECCV, 2025.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MMBench: Is Your Multi-modal Model an All-around Player? In ECCV, 2025

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:17.038114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:07.935523Z digest=sha256:1e7c42731a41d5c5d96c4ada82081ebd114804ab4626a4319bbdec5142c4d778

Observation 54401c49-46c3-4af6-b992-1fdc4d582771 · outbound

This paper cites Chartqa: A Benchmark for Question Answering About Charts With Visual and Logical Reason- ing.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Chartqa: A Benchmark for Question Answering About Charts With Visual and Logical Reason- ing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.743903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:08.060825Z digest=sha256:80450983278b621f16c1bfb5d50216e3a9795e2e5d877030cb3333ab4a4bf4e6

Observation 4c64e431-059d-4041-88f4-0bf1cb177edd · outbound

This paper cites Docvqa: A Dataset for VQA on Document Images.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Docvqa: A Dataset for VQA on Document Images

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.373407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:08.282593Z digest=sha256:57e3c3f48116b636f5b534f2b5df3b5ff70cef8fd99f5e56777109d5d6d9a0ab

Observation a9d22fad-abc7-49d9-810f-a284bbfe4824 · outbound

This paper cites Infograph- icVQA.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Infograph- icVQA

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:16.063658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:08.400143Z digest=sha256:3617d7c8eb93c8af14d13d810f837f51cdab63eeedcc1d52fce0849276a97d38

Observation b2bdd531-cbc6-4173-bb5f-290a138497d0 · outbound

This paper cites ShortGPT: Layers in Large Language Models are More Redundant Than You Expect.

Zero-Shot Vision Encoder Grafting via LLM Surrogates ShortGPT: Layers in Large Language Models are More Redundant Than You Expect

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:08.563542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:08.563542Z digest=sha256:dcf7b72bf623df27fe554c36e4a800c4017187faba781fd7be12a7f043357b7b

Observation 6cdc855b-7bfd-4691-801c-425a5f9c46f5 · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.729807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:08.737225Z digest=sha256:bde92c0657a125ffd80c141085a2913dfc2cbc70a43e51e1135fe4c1fa398505

Observation 232146ff-0a88-4f50-855c-fb474f1916d3 · outbound

This paper cites interpreting GPT: the logit lens.

Zero-Shot Vision Encoder Grafting via LLM Surrogates interpreting GPT: the logit lens

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.404758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:08.896020Z digest=sha256:c76cc1ba498fa8c096323666f5a0b3b9e830fff24518b11e3993c6c7116c2905

Observation 70fe4b8e-5a8c-4c2e-ac06-9159051c07fa · outbound

This paper cites Learning Transferable Visual Models From Natural Language Super- vision.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Learning Transferable Visual Models From Natural Language Super- vision

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:15.037846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:09.084874Z digest=sha256:128fd5190f3b23032f89a5c232e5fa661343e53e2ddb7412a4645aa4819d9eb3

Observation a54574e4-0195-4a87-b106-2c4102115b3f · outbound

This paper cites WinoGrande: An Adversarial Winograd Schema Challenge at Scale.

Zero-Shot Vision Encoder Grafting via LLM Surrogates WinoGrande: An Adversarial Winograd Schema Challenge at Scale

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:14.688023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:09.261308Z digest=sha256:9daaa5b07ae946eac2c9dfe0a67bed3cd67b11801a57dbe3ea52464b82ca49df

Observation 919f7776-77bc-4d98-9e36-159147c403cf · outbound

This paper cites Open Problems in Mechanistic Interpretability.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Open Problems in Mechanistic Interpretability

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:09.455799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:09.455799Z digest=sha256:fe3c789758d2901846101e44d098ad0d26777a8a8122460bc550bd58dcf31595

Observation 78f61f07-327e-420c-a7aa-9e2c43699f45 · outbound

This paper cites Towards VQA Models That Can Read.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Towards VQA Models That Can Read

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:14.377820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:09.639043Z digest=sha256:9c9c4d1d77f561772cac9996fc5c8d0f89561718faf6f673a5eeb71598e21f96

Observation ae21e43a-ad87-459a-96a4-d61fe2ab610b · outbound

This paper cites Transformer Layers as Painters.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Transformer Layers as Painters

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:09.868806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:09.868806Z digest=sha256:ae398cc88df1769935a2fca4c9d2e79ffee9e92c015128ebdd8980f3e54719e6

Observation c2eeae5e-c64c-4045-9a0d-fe116ec56739 · outbound

This paper cites an unresolved cited work.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:14.074467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:10.023098Z digest=sha256:cd854eb43e794c597cb6b2b67f0c3f2cc9366d822b231cb6c7cda2f51914207b

Observation 8b80db9e-b15a-49f8-9293-6e6b9916aaad · outbound

This paper cites an unresolved cited work.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:13.774364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:10.154755Z digest=sha256:3d7ee74564ea951ef5ecb464ab1b0ae56c82b2a3316532e33249d0780c37362a

Observation a5233e1d-bf72-455f-9bde-cbcffb59e415 · outbound

This paper cites Qwen2.5 Technical Report.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen2.5 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.316669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.316669Z digest=sha256:3c78082636da2f674705602fc59c5ee486e77f765b3a8607c2eb9abb68a9aa29

Observation 8cac2992-8ccd-4fe4-b59d-3433f20693fb · outbound

This paper cites Qwen3 Technical Report.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen3 Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.411364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.411364Z digest=sha256:a7ea5d20ad049fda26becaa5f3bfd87faeb9f5eb0c97bacdd6aacfe6d2cdfbd0

Observation 9a9ab4b5-bf59-4450-99bf-fbc4ddfbb156 · outbound

This paper cites Cambrian- 1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Cambrian- 1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:13.491774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:10.548879Z digest=sha256:37eef140ea81771ae3b8a8b8fcb3d26e3c901e1b36d6885e220eaa7014837c40

Observation bbd238ab-8814-446f-beb0-b6b14efeb168 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Zero-Shot Vision Encoder Grafting via LLM Surrogates SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.711035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.711035Z digest=sha256:b604f9f3d506a395379784cbad7824f4ed92ea03b0d0873c9b6a3a2d93a5edf7

Observation 02bcb860-639e-41f6-b365-d8f131348935 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:10.917149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:10.917149Z digest=sha256:8a801aa704d195b540d4a5e14ec357e02f0d7a8220904bb5d397f4b8114118b1

Observation 90da24bf-cbc8-4556-ad96-05125bcc443d · outbound

This paper cites CogVLM: Visual Expert for Pretrained Lan- guage Models.

Zero-Shot Vision Encoder Grafting via LLM Surrogates CogVLM: Visual Expert for Pretrained Lan- guage Models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:13.221500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:11.082351Z digest=sha256:ae5842d577caba531fd2b3455f4bbdd5300d1c9932a074b3bfbe8d21c2eab8aa

Observation 5a1b519c-ef56-4100-9df7-56f6a70460de · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Inte- grated Capabilities.

Zero-Shot Vision Encoder Grafting via LLM Surrogates MM-Vet: Evaluating Large Multimodal Models for Inte- grated Capabilities

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.806130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:11.239273Z digest=sha256:709192e9853be48e13a8a17e41c1130aef466e7e9d50795a830e5b95fba49c74

Observation 9f4c83da-0827-4fb2-b03c-2602aeeaa9b4 · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence? In ACL Anthology, 2019.

Zero-Shot Vision Encoder Grafting via LLM Surrogates HellaSwag: Can a Machine Really Finish Your Sentence? In ACL Anthology, 2019

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.491448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:11.437478Z digest=sha256:2d82b3d2530b7d976079bc9a51255b09866eae261802b072d06801a18f0b9326

Observation 652f85e5-fdfb-400c-be5a-0ee75c053fec · outbound

This paper cites Sigmoid Loss for Language Image Pre- Training.

Zero-Shot Vision Encoder Grafting via LLM Surrogates Sigmoid Loss for Language Image Pre- Training

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.138345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T13:10:11.570195Z digest=sha256:4ff2e0a0f75b05a78f6bfdb4517ff0cb7b9411fb32476999d1dcdf40e43ba1bd

Observation f1dca64a-039e-41c8-b879-848cd520fce3 · outbound

This paper cites PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel.

Zero-Shot Vision Encoder Grafting via LLM Surrogates PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:11.740399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:11.740399Z digest=sha256:428aa14b2286d4cf606bb29e94600b48eb8b07ad845b43725c66e2023583c138

Pith citing papers

No inbound Pith citation observations are available.