Pith. sign in

Paper Citation Record · LEDGER

Training data-efficient image transformers & distillation through attention

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 41 inbound Pith citation observations for arXiv:2012.12877.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2012.12877 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 41 of 41 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T05:01:50.977814Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:18:58.619569Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 62e97a20-5113-4187-b6ab-c46cfa621836 · inbound

Swin Transformer: Hierarchical Vision Transformer using Shifted Windows cites this paper.

Swin Transformer: Hierarchical Vision Transformer using Shifted Windows Training data-efficient image transformers & distillation through attention

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:27:56.932084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T19:27:56.785857Z digest=sha256:38079a7557d33bf790ec9b04f5cc4fe6ba712d0ed9dea088df2cb19c31fff0fa

Observation 22bc2d3f-a153-4b4c-9bd3-7594978e97e2 · inbound

Emerging Properties in Self-Supervised Vision Transformers cites this paper.

Emerging Properties in Self-Supervised Vision Transformers Training data-efficient image transformers & distillation through attention

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:04:51.630068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T14:04:51.458382Z digest=sha256:6f14ab0bb512d8da3a4bf8b519055aa910654eb63cccfca0520c8b28aa4ac15d

Observation 9193c30c-3234-47f1-b1c2-4371b20d77ea · inbound

BEiT: BERT Pre-Training of Image Transformers cites this paper.

BEiT: BERT Pre-Training of Image Transformers Training data-efficient image transformers & distillation through attention

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T11:50:11.559446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T11:50:11.476015Z digest=sha256:ef262704e9fd3fa5bfce9b5dda6e52ffe5eba1910f4e16032ee58703b1016395

Observation becc204f-26cd-46e1-9856-a034bda213a9 · inbound

Faster Segment Anything: Towards Lightweight SAM for Mobile Applications cites this paper.

Faster Segment Anything: Towards Lightweight SAM for Mobile Applications Training data-efficient image transformers & distillation through attention

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:41:43.513849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T22:41:43.411128Z digest=sha256:2fb3ffe7de29edc0d403e38d8c6bc90c3ed040e9ba1e70cfd0e8a6d27f95dcb6

Observation 4f4884d4-5a75-44c0-be61-65c2e4dc8715 · inbound

Vision Transformers Need Registers cites this paper.

Vision Transformers Need Registers Training data-efficient image transformers & distillation through attention

Reference 140

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T09:41:38.248058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-13T09:41:37.937046Z digest=sha256:f13b1c3077e95453b379586c811d648009ead0d6d9967cc4e220b31fba8836c2

Observation d7a21ffb-5f57-434f-91e8-61db8eaf14fc · inbound

TOAST: Transformer Optimization using Adaptive and Simple Transformations cites this paper.

TOAST: Transformer Optimization using Adaptive and Simple Transformations Training data-efficient image transformers & distillation through attention

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:48:23.068130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-23T19:46:07.124996Z digest=sha256:9020382bc1815c56cf63103a7ceb0c4cc626c015a762583d2340f9b3f9628aa9

Observation 6c646715-58ca-495e-966e-624831cc9755 · inbound

VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis cites this paper.

VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis Training data-efficient image transformers & distillation through attention

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T05:01:50.977814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:01:50.977814Z digest=sha256:4dacc468f0d93c8fab278d3b8bdfc833b3b8efc357fcfe4ebed4c086ece79bd9

Observation 1746af05-c3ca-4bba-957a-89c48a4bd120 · inbound

Cross-Layer Cache Aggregation for Token Reduction in Ultra-Fine-Grained Image Recognition cites this paper.

Cross-Layer Cache Aggregation for Token Reduction in Ultra-Fine-Grained Image Recognition Training data-efficient image transformers & distillation through attention

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T22:58:30.137033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:58:30.137033Z digest=sha256:4b918e6dc8222097c93e1c6f39d25f3f9606910672c5a1547aae356353a54bde

Observation 7164164d-1986-427c-98b1-cab1b9334db1 · inbound

Multiscaled Multi-Head Attention-based Video Transformer Network for Hand Gesture Recognition cites this paper.

Multiscaled Multi-Head Attention-based Video Transformer Network for Hand Gesture Recognition Training data-efficient image transformers & distillation through attention

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T22:44:42.366245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:44:42.366245Z digest=sha256:a483a9855faa22e4b97ec2c3b70a91fb22f08f40f79d6345feeb79bb1d56fc9a

Observation dbb95b0d-196f-4de1-af04-59fbb3be41a3 · inbound

Vulnerability-Aware Spatio-Temporal Learning for Generalizable Deepfake Video Detection cites this paper.

Vulnerability-Aware Spatio-Temporal Learning for Generalizable Deepfake Video Detection Training data-efficient image transformers & distillation through attention

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T22:38:00.672146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:38:00.672146Z digest=sha256:b64bb297c5b88835742d76ed4e2fc0259cc64997ae618dd5ecc8482ddd0a168a

Observation 1334db7d-b920-453c-87e7-8d2d0c66fc89 · inbound

Back Home: A Computer Vision Solution to Seashell Identification for Ecological Restoration cites this paper.

Back Home: A Computer Vision Solution to Seashell Identification for Ecological Restoration Training data-efficient image transformers & distillation through attention

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T21:28:04.557175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:28:04.557175Z digest=sha256:af6dc7ae7f31a11a14bb7603d6896aecc0da0f62a0c86ad2c808fbcbd9bb5d7d

Observation 73f363cd-db2a-4cfb-a526-1b56bdad42bd · inbound

SIM: Surface-based fMRI Analysis for Inter-Subject Multimodal Decoding from Movie-Watching Experiments cites this paper.

SIM: Surface-based fMRI Analysis for Inter-Subject Multimodal Decoding from Movie-Watching Experiments Training data-efficient image transformers & distillation through attention

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T13:07:33.320089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T13:07:33.320089Z digest=sha256:2774878b19d3061309ed7ddfe21253f2e282472ae7316b13d254c00a48a1023c

Observation ed3f86ac-51f9-49e3-a784-ef6b44ceaaf6 · inbound

Performance Analysis of Traditional VQA Models Under Limited Computational Resources cites this paper.

Performance Analysis of Traditional VQA Models Under Limited Computational Resources Training data-efficient image transformers & distillation through attention

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T18:10:38.319474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:10:38.319474Z digest=sha256:b46eaa0b5ffca63de28d978ce5167e478216166d088b92d30b91e65e24668d44

Observation 36f468b7-4ea3-481c-96c2-b99eae8add2c · inbound

Kolmogorov-Arnold Fourier Networks cites this paper.

Kolmogorov-Arnold Fourier Networks Training data-efficient image transformers & distillation through attention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T17:06:10.335008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:06:10.335008Z digest=sha256:000ad77e8c31f2a27b26880c1b4268d2e0fdbbd3df0f6ebfbc1cc1b28b308cd2

Observation b0c3e9b2-87dd-42ab-9730-9f4ffd65143f · inbound

From Pixels to Components: Eigenvector Masking for Visual Representation Learning cites this paper.

From Pixels to Components: Eigenvector Masking for Visual Representation Learning Training data-efficient image transformers & distillation through attention

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T15:53:51.527111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:53:51.527111Z digest=sha256:6ad530dc298ef7532143369956704aece3dca190e9c576aef770d0f31dff5a0f

Observation a372d044-985e-4053-b65d-f8443b42bcea · inbound

ViSIR: Vision Transformer Single Image Reconstruction Method for Earth System Models cites this paper.

ViSIR: Vision Transformer Single Image Reconstruction Method for Earth System Models Training data-efficient image transformers & distillation through attention

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T14:31:18.632986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:31:18.632986Z digest=sha256:e8572ed8846f4f9692aaac3d3584656c4bce9615e3e0dfdd8f89f4cbfd109866

Observation 40ce7000-3c56-400f-a776-40a3d1dba4b9 · inbound

Adaptive Camera Sensor for Vision Models cites this paper.

Adaptive Camera Sensor for Vision Models Training data-efficient image transformers & distillation through attention

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T02:17:24.507272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-23T02:17:02.071721Z digest=sha256:4fb1b21ff3d345b3dcf2650f8e2d14c1484ed52cfefcb14c3ee7bbdb67d9f270

Observation 20f06ed0-4c5f-43e0-8ae5-0951439e0639 · inbound

In Context Learning with Vision Transformers: Case Study cites this paper.

In Context Learning with Vision Transformers: Case Study Training data-efficient image transformers & distillation through attention

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:16.332752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:16.332752Z digest=sha256:a8113caaabd38d417cf2cbdc8838ebcc418a10a81677600dbeb7de825af89e21

Observation 47dea056-821e-4fc9-8d04-c317edad8254 · inbound

SAAT: Synergistic Alternating Aggregation Transformer for Image Super-Resolution cites this paper.

SAAT: Synergistic Alternating Aggregation Transformer for Image Super-Resolution Training data-efficient image transformers & distillation through attention

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:01:27.482186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:01:27.482186Z digest=sha256:6fa083203394b05030097156f25b2acbb70b6f755937bc889fbb54076ea8e8ab

Observation 6ef99801-722a-4663-9d16-3dbfef595049 · inbound

AdvMIM: Adversarial Masked Image Modeling for Semi-Supervised Medical Image Segmentation cites this paper.

AdvMIM: Adversarial Masked Image Modeling for Semi-Supervised Medical Image Segmentation Training data-efficient image transformers & distillation through attention

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:21.996010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:21.996010Z digest=sha256:957d78f2941fc8ca855c7b735a9ee4c3fbac25cf1fc39c2968e85251f667dfb6

Observation 62b3ebe3-3eb7-46b7-a31a-91ee9114fe7e · inbound

Fine-Grained Image Recognition from Scratch with Teacher-Guided Data Augmentation cites this paper.

Fine-Grained Image Recognition from Scratch with Teacher-Guided Data Augmentation Training data-efficient image transformers & distillation through attention

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:06.574467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:06.574467Z digest=sha256:536fafd121465245ebd62f9261417ec4f06fba116988f6a6ce3f4fe273f30e65

Observation 37c2ffe3-2a3e-4707-a57c-152aeee783fd · inbound

MoEcho: Exploiting Side-Channel Attacks to Compromise User Privacy in Mixture-of-Experts LLMs cites this paper.

MoEcho: Exploiting Side-Channel Attacks to Compromise User Privacy in Mixture-of-Experts LLMs Training data-efficient image transformers & distillation through attention

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T18:11:08.807726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:11:08.807726Z digest=sha256:8f6bf710cad49882b0b4dd491cd3fa91319b3e59c04b16044d9300a6908dd3b0

Observation 7bd38294-4c1e-47df-8a9e-73f282841ba6 · inbound

I-Segmenter: Integer-Only Vision Transformer for Efficient Semantic Segmentation cites this paper.

I-Segmenter: Integer-Only Vision Transformer for Efficient Semantic Segmentation Training data-efficient image transformers & distillation through attention

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T17:57:09.534743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:57:09.534743Z digest=sha256:40b49f549ab199e8931e8bdd9434d5490d146afc6d069bd75d32c7ff9aa6b4b5

Observation f1ac5414-7926-4d57-94c5-c2f7707c129a · inbound

Efficient Learned Image Compression Through Knowledge Distillation cites this paper.

Efficient Learned Image Compression Through Knowledge Distillation Training data-efficient image transformers & distillation through attention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T17:56:03.712567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:56:03.712567Z digest=sha256:bd25393f8bde16eda00c339d20a0653c72a490e0ad32163b43d0317ab23e5586

Observation 218be6df-a9d5-426b-96cc-2c4ecdfedf42 · inbound

Two-Dimensional Quantization for Geometry-Aware Audio Coding cites this paper.

Two-Dimensional Quantization for Geometry-Aware Audio Coding Training data-efficient image transformers & distillation through attention

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:20:29.175270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T18:16:51.486807Z digest=sha256:ae176768b81ebb2eacb4fbfd032a50c09c49157e678f785889c373f9ad2fbce5

Observation 7740c2e6-a228-4f4d-b26a-ce9a9323d902 · inbound

Learn from A Rationalist: Distilling Intermediate Interpretable Rationales cites this paper.

Learn from A Rationalist: Distilling Intermediate Interpretable Rationales Training data-efficient image transformers & distillation through attention

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T06:36:46.873497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T06:36:46.873497Z digest=sha256:00c16946aa76912858342e10059e3a9084697179fd84e838cb70b0ee9cc5266d

Observation b83b2b8c-464f-44ec-9bf1-3004bc87bcb4 · inbound

CORP: Closed-Form One-shot Representation-Preserving Structured Pruning for Transformers cites this paper.

CORP: Closed-Form One-shot Representation-Preserving Structured Pruning for Transformers Training data-efficient image transformers & distillation through attention

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:40:43.943667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T07:38:50.252917Z digest=sha256:f3171d9612bb232564d2e4bf3af2b1070358b786a7e14260e20d73c0bd99513b

Observation 353d65f8-ad6f-4557-a655-6018efb041fd · inbound

OneSearch-V2: The Latent Reasoning Enhanced Self-distillation Generative Search Framework cites this paper.

OneSearch-V2: The Latent Reasoning Enhanced Self-distillation Generative Search Framework Training data-efficient image transformers & distillation through attention

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:25:12.528985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T07:22:24.713032Z digest=sha256:a3741168a4d5516c95180c1a2ae11a28ed0daa67216e9110d54d4f51aad3b993

Observation 4f2f174a-a495-40a0-a533-c26af5f10259 · inbound

LAA-X: Unified Localized Artifact Attention for Quality-Agnostic and Generalizable Face Forgery Detection cites this paper.

LAA-X: Unified Localized Artifact Attention for Quality-Agnostic and Generalizable Face Forgery Detection Training data-efficient image transformers & distillation through attention

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:23:02.202120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T17:22:28.944874Z digest=sha256:6f3f98459eb05e687f39318a0f087ca58ef7d50b1fc6063e214eb58dd8fc76ed

Observation 9583f3eb-764b-45c1-bba0-8fd5e9ab8dd4 · inbound

Spectral Vision Transformer for Efficient Tokenization with Limited Data cites this paper.

Spectral Vision Transformer for Efficient Tokenization with Limited Data Training data-efficient image transformers & distillation through attention

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:57:28.189880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T06:53:51.211617Z digest=sha256:de3a4cef390ba9bd8b7acc463a4a7f3abdaf258f9fa545012e4f44eabcb4ba35

Observation 15d9f6af-1701-4190-8b13-c7b59526c20b · inbound

Architecture-Aware Explanation Auditing for Industrial Visual Inspection cites this paper.

Architecture-Aware Explanation Auditing for Industrial Visual Inspection Training data-efficient image transformers & distillation through attention

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:49:41.680483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T02:40:36.771439Z digest=sha256:e5da3db85f7ded84fe33851f5b44f83d01e6bb37fe2bc27ceac7ecac1887344b

Observation 5f842147-b676-4dfe-b834-f53a82b75a8c · inbound

Architecture-Aware Explanation Auditing for Industrial Visual Inspection cites this paper.

Architecture-Aware Explanation Auditing for Industrial Visual Inspection Training data-efficient image transformers & distillation through attention

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:03:46.333205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T21:02:52.559284Z digest=sha256:adb14527956b668095953ab6ff93459fe6cb9f1eeef2a586f04599fe4c9dd187

Observation 5728905f-5e00-4884-b9f3-120950d4e2a3 · inbound

Architecture-Aware Explanation Auditing for Industrial Visual Inspection cites this paper.

Architecture-Aware Explanation Auditing for Industrial Visual Inspection Training data-efficient image transformers & distillation through attention

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:25:46.309279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T21:37:34.236270Z digest=sha256:df25cea18659ffcc22d2ee6aa93efe9783944c0c306bd6ada1f6936779cb6da7

Observation 7e27a7ca-61ac-4184-ad7f-37a67bd8fca0 · inbound

FusionCell: Cross-Attentive Fusion of Layout Geometry and Netlist Topology for Standard-Cell Performance Prediction cites this paper.

FusionCell: Cross-Attentive Fusion of Layout Geometry and Netlist Topology for Standard-Cell Performance Prediction Training data-efficient image transformers & distillation through attention

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T08:24:03.482603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-21T08:23:09.232227Z digest=sha256:a42e683ef00cfe4b4f0b5195ce7fc6db8e9a7a9c6a5c11a7907964e47b10cc9d

Observation 2ce62072-1865-4d63-bd01-724e0d827e5c · inbound

Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers cites this paper.

Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers Training data-efficient image transformers & distillation through attention

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:13:48.728883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T18:08:42.533804Z digest=sha256:496dd92c74d04a32f2d547656db45b609e589e5deff269b977f3ba8ff817253b

Observation 2df3c329-abc4-43c0-8d65-bcd5f5058e0b · inbound

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory cites this paper.

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory Training data-efficient image transformers & distillation through attention

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:46:55.237538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T03:07:52.730713Z digest=sha256:12e3153e26bbfecf969e7956b6ca1148bced6147443cfff78930a14e0c69dd90

Observation 1f5f03a3-0507-4314-9f04-2c2d4b151301 · inbound

STAR: Rethinking MoE Routing as Structure-Aware Subspace Learning cites this paper.

STAR: Rethinking MoE Routing as Structure-Aware Subspace Learning Training data-efficient image transformers & distillation through attention

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T23:07:26.964860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T18:28:35.162934Z digest=sha256:324b748719c62e3ea80683a9d0a58108c2e51c1cd10e31e0c7a1ddcb04a76704

Observation d861370d-c0ef-4d25-aefa-00da6f255cc2 · inbound

Architectural Bias in Face Presentation Attack Detection: A Comparative Study of Vision Transformers and Convolutional Neural Networks cites this paper.

Architectural Bias in Face Presentation Attack Detection: A Comparative Study of Vision Transformers and Convolutional Neural Networks Training data-efficient image transformers & distillation through attention

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:18:58.622125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T00:45:33.196262Z digest=sha256:b4e1007af98066a7df82871b75eef195809f8b6c0a16e0a679cf24861cbc3c9e

Observation 71b52d96-24ac-4330-b9f9-3798b11fb4b7 · inbound

Gradient Smoothing: Coupling Layer-wise Updates for Improved Optimization cites this paper.

Gradient Smoothing: Coupling Layer-wise Updates for Improved Optimization Training data-efficient image transformers & distillation through attention

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:25:41.228666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-01T06:36:48.524846Z digest=sha256:216fd5ac7c8c9c48712ace81c84a9b0bed51a89b04494efcdcb1aa18bbe290aa

Observation c8b17841-f315-4858-893a-ed48ef104d10 · inbound

Multimodal Fusion for Fine-Grained Classification of Breast Fibroadenoma and Phyllodes Tumors cites this paper.

Multimodal Fusion for Fine-Grained Classification of Breast Fibroadenoma and Phyllodes Tumors Training data-efficient image transformers & distillation through attention

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:48:34.949976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-03T15:44:18.845387Z digest=sha256:24e983e534eb14cd3e00d078ded9266caf6e001f998737756705112d015ff835

Observation b82f85eb-2240-438b-8288-e0e6ce9dd8cd · inbound

The evolution of AI from image interpretation toward scientific inference in nanoparticle electron microscopy cites this paper.

The evolution of AI from image interpretation toward scientific inference in nanoparticle electron microscopy Training data-efficient image transformers & distillation through attention

Reference 119

Resolution
unresolved
no resolver link, observed 2026-07-14T12:06:28.342674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:06:28.342674Z digest=sha256:9c46e418bbc5a8f30d401a8434cc656b60c50abf70f0f86e2aa3f908be5c266f