Pith. sign in

Paper Citation Record · LEDGER

Vision Transformer Adapter for Dense Predictions

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2205.08534.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2205.08534 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 42 of 42 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:41:01.268204Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

204
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8b5962d8-85f5-41a6-9c72-5fe108e8be87 · inbound

Uncertainty in Real-Time Semantic Segmentation on Embedded Systems cites this paper.

Uncertainty in Real-Time Semantic Segmentation on Embedded Systems Vision Transformer Adapter for Dense Predictions

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-24T10:14:19.005267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T10:09:23.030973Z digest=sha256:87713f404233f738c20c3b8a471870e17e2c48132991aaf9e5f11114f7a763f1

Observation 2b68571f-9d3b-419b-add6-4fc097ce0ed1 · inbound

T2I-Adapter: Learning Adapters to Dig out More Controllable Ability for Text-to-Image Diffusion Models cites this paper.

T2I-Adapter: Learning Adapters to Dig out More Controllable Ability for Text-to-Image Diffusion Models Vision Transformer Adapter for Dense Predictions

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:47:50.412236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T22:47:50.355642Z digest=sha256:0f6080e064d3fe090b18b589569439e11ac18b3187bdfc06999ce9e4eb0a90f2

Observation 0bc9e5cd-1996-4d7a-84a9-fd770152ec25 · inbound

Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey cites this paper.

Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey Vision Transformer Adapter for Dense Predictions

Reference 189

Resolution
verified exact
arxiv_id, observed 2026-05-13T11:32:36.938095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T11:32:36.738536Z digest=sha256:40d652d16605def268592a5b4f94a6614c11ae01ad3b787d3c033bfd4d667570

Observation 963676bc-1edb-48fd-922d-7aef6b7701bd · inbound

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling cites this paper.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Vision Transformer Adapter for Dense Predictions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.268204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.268204Z digest=sha256:98fa988e4a858312a75944eab88190d973a1ea6d6090d9aaedbd7398423a2960

Observation 1c7f0a7f-72e7-404d-b54e-b92393250be5 · inbound

From Visuals to Vocabulary: Establishing Equivalence Between Image and Text Token Through Autoregressive Pre-training in MLLMs cites this paper.

From Visuals to Vocabulary: Establishing Equivalence Between Image and Text Token Through Autoregressive Pre-training in MLLMs Vision Transformer Adapter for Dense Predictions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T22:43:59.380438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T22:43:59.380438Z digest=sha256:aaac3744669ae6fd71483dcf596c913950ac5af02124bfffd3e5e97830b71432

Observation 77cdb184-211e-4059-a2f4-499d2cd81470 · inbound

Radar-Guided Polynomial Fitting for Metric Depth Estimation cites this paper.

Radar-Guided Polynomial Fitting for Metric Depth Estimation Vision Transformer Adapter for Dense Predictions

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:37:12.799127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T22:35:56.031755Z digest=sha256:430d09b660756941bfe0b0c2cb35db313ba411a79ad3746ed2164242558f7891

Observation 5483b062-d153-4a55-8294-5cf778965ae7 · inbound

Locality-Aware Zero-Shot Human-Object Interaction Detection cites this paper.

Locality-Aware Zero-Shot Human-Object Interaction Detection Vision Transformer Adapter for Dense Predictions

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:09.562800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:09.562800Z digest=sha256:1561e64672cf813b434d8e2377efb2a925c6ade8d883c3dc9ddfd7d1cc0753dd

Observation ce1d3cc0-b8a5-4566-beac-466ef12dbc95 · inbound

Data-Efficient Challenges in Visual Inductive Priors: A Retrospective cites this paper.

Data-Efficient Challenges in Visual Inductive Priors: A Retrospective Vision Transformer Adapter for Dense Predictions

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:10:59.140254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:10:59.140254Z digest=sha256:c97f23bc90dd52e1210111003d6eee466932ab5fa53582853428ec3c82c423d6

Observation b1b5eee8-d624-4a21-8933-82b1f20b639e · inbound

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models cites this paper.

Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models Vision Transformer Adapter for Dense Predictions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:21.100923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:02:21.100923Z digest=sha256:3e38a9db4b0f33a7c6ab4e3999355a58c42295f7a30508df68e132e0f1ac4b9d

Observation 6556d8c5-c82c-40a7-bc92-66645b3e94a6 · inbound

Mamba Guided Boundary Prior Matters: A New Perspective for Generalized Polyp Segmentation cites this paper.

Mamba Guided Boundary Prior Matters: A New Perspective for Generalized Polyp Segmentation Vision Transformer Adapter for Dense Predictions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:55:06.046279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:55:06.046279Z digest=sha256:cc8daa5b8d68534923a76c1a4555f28ea455f60e25ed7928ac003003435d4316

Observation 190292f5-0b9c-450d-b19e-a5b00cc01660 · inbound

SAILViT: Towards Robust and Generalizable Visual Backbones for MLLMs via Gradual Feature Refinement cites this paper.

SAILViT: Towards Robust and Generalizable Visual Backbones for MLLMs via Gradual Feature Refinement Vision Transformer Adapter for Dense Predictions

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:53:01.512069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:53:01.512069Z digest=sha256:590756cd539661ac5f196eff241597b4c7b7aaf02caf6acef069d5df836b0eb6

Observation 9167e7f9-2555-46a6-a149-6294af3d4f56 · inbound

Colorectal Cancer Tumor Grade Segmentation in Digital Histopathology Images: From Giga to Mini Challenge cites this paper.

Colorectal Cancer Tumor Grade Segmentation in Digital Histopathology Images: From Giga to Mini Challenge Vision Transformer Adapter for Dense Predictions

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:45:19.798495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:45:19.798495Z digest=sha256:9c62654dd64395fcde0cbc563883532a33d4ce52c7f4bcef1f2f6867837f3376

Observation 883b2de5-6e7e-4ce7-82b9-f04d83ff57dc · inbound

Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation cites this paper.

Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation Vision Transformer Adapter for Dense Predictions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:17:31.571904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:17:31.571904Z digest=sha256:a67839a8dad82c3e0669e41825ea03fea6c701b443504950d1aaef0f8243bf69

Observation ccee4867-5d63-4b24-8731-a556a30ae3c3 · inbound

Latest Object Memory Management for Temporally Consistent Video Instance Segmentation cites this paper.

Latest Object Memory Management for Temporally Consistent Video Instance Segmentation Vision Transformer Adapter for Dense Predictions

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:08:43.619315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:08:43.619315Z digest=sha256:861ad91d3dcf762e8b5ad0e8cf6df51f313a6e0ac18d2a34ec23e61a94061067

Observation fd56a0fa-cea4-4096-ba1e-76dd1f69161d · inbound

MPT: Motion Prompt Tuning for Micro-Expression Recognition cites this paper.

MPT: Motion Prompt Tuning for Micro-Expression Recognition Vision Transformer Adapter for Dense Predictions

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T21:05:13.366350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:05:13.366350Z digest=sha256:8f6d7668aac02a71da2928a24e8f3e1b4f328194df9af02f8930c4f0e1202c7a

Observation 32a08aa5-fe12-4efd-9136-73d8ac2d416e · inbound

AI-driven Remote Facial Skin Hydration and TEWL Assessment from Selfie Images: A Systematic Solution cites this paper.

AI-driven Remote Facial Skin Hydration and TEWL Assessment from Selfie Images: A Systematic Solution Vision Transformer Adapter for Dense Predictions

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T23:55:38.879111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:55:38.879111Z digest=sha256:d6d1a12d4302588949acd9b7f8c45e637ffd9eb70afa9ef7700e1464d7828c66

Observation 0f5a9b79-a8c9-47d7-9340-977239b167d7 · inbound

FoMo4Wheat: Toward reliable crop vision foundation models with globally curated data cites this paper.

FoMo4Wheat: Toward reliable crop vision foundation models with globally curated data Vision Transformer Adapter for Dense Predictions

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:17.147327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:17.147327Z digest=sha256:429bc609b57d32b07396ec5c5bcf5b4a60da1d043932ba6df1f1db253c2c0dc3

Observation f6a4c07a-596d-4536-8d02-7076b636a9a5 · inbound

Live(r) Die: Predicting Survival in Colorectal Liver Metastasis cites this paper.

Live(r) Die: Predicting Survival in Colorectal Liver Metastasis Vision Transformer Adapter for Dense Predictions

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T20:04:30.153556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:04:30.153556Z digest=sha256:0ccf403373b5ce85c09bc052aa46307f8220564e5e02b45c950e603099f24981

Observation 00b8faf1-c4cb-404e-85aa-1d7583512e92 · inbound

Delineate Anything Flow: Fast, Country-Level Field Boundary Detection from Any Source cites this paper.

Delineate Anything Flow: Fast, Country-Level Field Boundary Detection from Any Source Vision Transformer Adapter for Dense Predictions

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T21:40:17.427614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T21:38:46.583533Z digest=sha256:f75e84866d1dd8a35b7c5d0258973ce304c6bad5465026de59e2da4336063d95

Observation f71c6955-8b72-454b-bed0-fdd18bea0743 · inbound

Exploring the Rashomon Set for Concept-Based Models cites this paper.

Exploring the Rashomon Set for Concept-Based Models Vision Transformer Adapter for Dense Predictions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T20:33:47.381192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:33:47.381192Z digest=sha256:7250a857b6bcb19d54e586c925e06f9bbf8e322412957e8ac571ec652970e2f9

Observation 62c7246f-59ff-4671-9411-b9d281447a6f · inbound

Foundation Model-Driven Semantic Change Detection in Remote Sensing Imagery cites this paper.

Foundation Model-Driven Semantic Change Detection in Remote Sensing Imagery Vision Transformer Adapter for Dense Predictions

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:06:41.859261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T22:06:35.964891Z digest=sha256:dceea9792f3508aee843ac69ce5548faac3504641ce890462e66f6cf82d33a8f

Observation e38d90e6-9bcc-45e0-9421-03e1442bddba · inbound

Foundation Model-Driven Semantic Change Detection in Remote Sensing Imagery cites this paper.

Foundation Model-Driven Semantic Change Detection in Remote Sensing Imagery Vision Transformer Adapter for Dense Predictions

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T23:29:30.575182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:29:30.575182Z digest=sha256:c198f506ade2783d025c3d9850b92e6c1725e45a6825183d29640599902a1c16

Observation 2dce84b2-bdbf-4550-9557-7ae51dc15dd3 · inbound

DinoRADE: Full Spectral Radar-Camera Fusion with Vision Foundation Model Features for Multi-class Object Detection in Adverse Weather cites this paper.

DinoRADE: Full Spectral Radar-Camera Fusion with Vision Foundation Model Features for Multi-class Object Detection in Adverse Weather Vision Transformer Adapter for Dense Predictions

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:46:31.376203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:27:33.761096Z digest=sha256:189dbc23fc588d7d46d2559977d7835622f4f6f8ada1c63e528f13629be6b12e

Observation 086548d8-2bbb-4d99-b6a7-7198e7f9ba32 · inbound

Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition cites this paper.

Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition Vision Transformer Adapter for Dense Predictions

Reference 68

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:55:59.088465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:55:02.621221Z digest=sha256:c3dfd0e1200ee08d9ed491eb1a1e6ee9cc5fd7d338bda9f38da198ec4e066668

Observation e01d9289-e112-41f5-b4eb-833597ec8dc2 · inbound

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation cites this paper.

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation Vision Transformer Adapter for Dense Predictions

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:15:58.969449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:45:36.306400Z digest=sha256:283074b1806ca2c62dcbb6f8001e1d67647a8613616c8218708c5bb6d9532bda

Observation be9226d3-f0cd-467c-b247-47e919b40dac · inbound

HAMSA: Scanning-Free Vision State Space Models via SpectralPulseNet cites this paper.

HAMSA: Scanning-Free Vision State Space Models via SpectralPulseNet Vision Transformer Adapter for Dense Predictions

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:40:18.966766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T11:39:03.041265Z digest=sha256:2e04e66c7e73191ee4daad7a5c1af3b53df027d64bbab2d8cc8d244c4dcb5cf5

Observation 0f146482-5863-453a-89fa-851b8df1d7a1 · inbound

VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection cites this paper.

VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection Vision Transformer Adapter for Dense Predictions

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:31:07.259438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T21:38:15.877684Z digest=sha256:198c5b56cfceb4d2cf67ad4467ab2011be1b91a7b0b379b3c07744304dec4725

Observation b70ff739-7ea3-4c7a-b74a-2fbe89570b31 · inbound

VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection cites this paper.

VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection Vision Transformer Adapter for Dense Predictions

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:05:26.299830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T06:03:56.040612Z digest=sha256:6d7f0c47fce4298fbac2d1319fc214e162b981cbaf260a8a32bd33727aadd847

Observation d0d6cccf-86f6-43ea-8a8f-505af27a292d · inbound

MP-ISMoE: Mixed-Precision Interactive Side Mixture-of-Experts for Efficient Transfer Learning cites this paper.

MP-ISMoE: Mixed-Precision Interactive Side Mixture-of-Experts for Efficient Transfer Learning Vision Transformer Adapter for Dense Predictions

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:01:13.138381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T17:19:59.247074Z digest=sha256:e9e497ba55719f85e1295f29ad45ce96ed2aa558e92a642e1033ccdd711aaf36

Observation b199198f-ca0a-4001-bb3a-141bfafa1a9a · inbound

Beyond ViT Tokens: Masked-Diffusion Pretrained Convolutional Pathology Foundation Model for Cell-Level Dense Prediction cites this paper.

Beyond ViT Tokens: Masked-Diffusion Pretrained Convolutional Pathology Foundation Model for Cell-Level Dense Prediction Vision Transformer Adapter for Dense Predictions

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:24.452112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T00:58:28.390860Z digest=sha256:d645808368f2ff77f10f68f346fe273ebcc282ddc6116a94acc1e15485e30936

Observation 80107b2c-c9a6-49f5-b9d3-2255f56db1ac · inbound

Unleashing Vision Transformer Potential In Image Quality Assessment via Global-Local Adaptive Interaction cites this paper.

Unleashing Vision Transformer Potential In Image Quality Assessment via Global-Local Adaptive Interaction Vision Transformer Adapter for Dense Predictions

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:53:17.737755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T12:48:31.650362Z digest=sha256:14b2d8b10e94be29120f659c32f2ca045ab7ee170816c9cc433f7cc50f3fa00b

Observation 49988264-a3f0-4d14-a5f4-aedcf47f8883 · inbound

Selective, Regularized, and Calibrated: Harnessing Vision Foundation Models for Cross-Domain Few-Shot Semantic Segmentation cites this paper.

Selective, Regularized, and Calibrated: Harnessing Vision Foundation Models for Cross-Domain Few-Shot Semantic Segmentation Vision Transformer Adapter for Dense Predictions

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:48:05.834523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T06:45:35.591459Z digest=sha256:32200d9fd3a931c8afc6f6a545c4dac0b119a108a0393c0da317e5a9ab7e8441

Observation 50cde99e-7679-4066-9ebe-5d6fb953d490 · inbound

Adapting Prithvi-EO for Fallow Detection for Food-Water Nexus: ViT-Adapter Necks and Parameter-Efficient Backbone tuning of Geospatial Foundation Model cites this paper.

Adapting Prithvi-EO for Fallow Detection for Food-Water Nexus: ViT-Adapter Necks and Parameter-Efficient Backbone tuning of Geospatial Foundation Model Vision Transformer Adapter for Dense Predictions

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:18:03.458026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T09:38:51.799082Z digest=sha256:704d6cdaed37bb059105486a44957b94209391b41ed830eb9304bcd9bc332e09

Observation 19c0a0b8-7dea-422b-b587-a94cff80b226 · inbound

Timage: A Generative Text-in-Image Paradigm for Fine-Tuning Vision-Language Models cites this paper.

Timage: A Generative Text-in-Image Paradigm for Fine-Tuning Vision-Language Models Vision Transformer Adapter for Dense Predictions

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:39:29.883567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T17:51:49.198510Z digest=sha256:d72a970eb51e6c63c8706a816dc776f39e7b9b6f10537be6d62da1fe4050ec42

Observation 9f390e78-9dcf-43ab-90a8-8c91f4f4b9e9 · inbound

UNITY: Attention Flow Networks for Adaptive Conditioning in Diffusion cites this paper.

UNITY: Attention Flow Networks for Adaptive Conditioning in Diffusion Vision Transformer Adapter for Dense Predictions

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:49:31.034417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T17:33:37.227935Z digest=sha256:8053a1d0561f6a734cd2b6cedac737a5a1650569cc6c63fb7494bf319d063b9d

Observation bc2df1f0-09f7-4b92-b95e-eace0f4bbb29 · inbound

UNITY: Attention Flow Networks for Adaptive Conditioning in Diffusion cites this paper.

UNITY: Attention Flow Networks for Adaptive Conditioning in Diffusion Vision Transformer Adapter for Dense Predictions

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T13:09:35.871559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:09:35.871559Z digest=sha256:e32e6465bdd68a82688d6ac1bde9e30acbf381aaa09a1b2543ab3fc54b3859eb

Observation 5265217a-de6b-4d1f-9c81-5d4775edda06 · inbound

State Space Models Meet Remote Sensing: A Survey cites this paper.

State Space Models Meet Remote Sensing: A Survey Vision Transformer Adapter for Dense Predictions

Reference 197

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:30:07.544420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-25T21:18:05.054587Z digest=sha256:ac3b75ecace740bb751571e164a9416d6b810dd5b9717e0daf3d4da4faaa5298

Observation 5e823516-e69e-4310-a9f3-854997de7e19 · inbound

REAL-OW: Rehearsal-free Open World Object Detection with Low-Rank Adaptation and Dual-Stage Objectness Modeling cites this paper.

REAL-OW: Rehearsal-free Open World Object Detection with Low-Rank Adaptation and Dual-Stage Objectness Modeling Vision Transformer Adapter for Dense Predictions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T05:31:12.070781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:31:12.070781Z digest=sha256:9b6c325b13c7078b89a05dd0382840c4217035163e974e9ae20850f031b1e5fa

Observation 8cdfcd5f-23cf-4cd7-8f3e-4b36c732a3a0 · inbound

FBLayout: Optimizing Memory Layout for Efficient LLM Finetuning on Mobile GPUs cites this paper.

FBLayout: Optimizing Memory Layout for Efficient LLM Finetuning on Mobile GPUs Vision Transformer Adapter for Dense Predictions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T08:23:07.456234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:23:07.456234Z digest=sha256:852bc4594fe9061507b6bc1886eb138e0662ed98509ab4062d99f71385395175

Observation cb3d0ccc-9a1e-4929-8d21-5ef66be8553e · inbound

iFAN: Inference-Aware Learning for Plain Mask Transformers cites this paper.

iFAN: Inference-Aware Learning for Plain Mask Transformers Vision Transformer Adapter for Dense Predictions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T23:22:24.597250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:22:24.597250Z digest=sha256:038e2ac5f732034e21f79a241e16b2d3b8a3a154355fdb5d26345fad3c24cca2

Observation eb2548a9-2396-4638-ae4d-dc7f8104e90a · inbound

From Multi-Resolution Cells to Gigapixel Whole Slide Images Foundation Model for Computational Pathology cites this paper.

From Multi-Resolution Cells to Gigapixel Whole Slide Images Foundation Model for Computational Pathology Vision Transformer Adapter for Dense Predictions

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T17:47:37.173107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:47:37.173107Z digest=sha256:cfc14bb8275fcf5670ff84957179dd5c34d99fe717248be83517f13853c880f0

Observation 7ae84642-e6d8-421c-bdeb-4647fbafa78d · inbound

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation cites this paper.

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation Vision Transformer Adapter for Dense Predictions

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T14:32:59.260963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:32:59.260963Z digest=sha256:ad25fef971757987bd7e9f7ffff108f5a3a9bbd74f4a2a7dc40dd205f6230ffb