Pith. sign in

Paper Citation Record · LEDGER

CvT: Introducing Convolutions to Vision Transformers

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2103.15808.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2103.15808 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:18:21.009470Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.785007Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e042909f-f0fb-4025-816c-8f3439786601 · inbound

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer cites this paper.

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer CvT: Introducing Convolutions to Vision Transformers

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:46:35.184858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T20:46:35.073600Z digest=sha256:f8e0d076d685723ab725b177e5b6b7af452d2a6300f87841c56576c7e187f8ba

Observation ad6687e0-d433-4f45-b37f-1f302491447b · inbound

Convolutional Vision Transformer for Cosmology Parameter Inference cites this paper.

Convolutional Vision Transformer for Cosmology Parameter Inference CvT: Introducing Convolutions to Vision Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T15:18:21.009470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:18:21.009470Z digest=sha256:757350e6b628a93dff34706ddd7fc260d15e624c73bfad92ef20c38d85c8e115

Observation c18b2d64-e096-4035-98aa-d7ef4360257b · inbound

Unified Local and Global Attention Interaction Modeling for Vision Transformers cites this paper.

Unified Local and Global Attention Interaction Modeling for Vision Transformers CvT: Introducing Convolutions to Vision Transformers

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T04:34:02.303884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:34:02.303884Z digest=sha256:b13cdeb4db32c94fd3d21ae15e1d92ac953ebe243953420eada839c082ac0a0b

Observation 3e529189-274c-41c4-9b48-e5ace37efa68 · inbound

Residual Transformer Fusion Network for Salt and Pepper Image Denoising cites this paper.

Residual Transformer Fusion Network for Salt and Pepper Image Denoising CvT: Introducing Convolutions to Vision Transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T23:01:17.102588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:01:17.102588Z digest=sha256:ee861be007be3fc79bebc0c4d5a519770cd86eaae74a62af1dfa85b4bc5e6ad8

Observation 348189fe-11a6-49c0-bba1-d12b81a82286 · inbound

JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment cites this paper.

JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment CvT: Introducing Convolutions to Vision Transformers

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T13:14:37.600229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T13:14:37.600229Z digest=sha256:3646c8b18bd9a788469b2f35267fbe566ea710f1aa2cda7c55dcc659bea8c458

Observation 342d9fec-7898-4ea9-95be-6a72925cb26a · inbound

Enhancing compact convolutional transformers with super attention cites this paper.

Enhancing compact convolutional transformers with super attention CvT: Introducing Convolutions to Vision Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:06:14.447164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:06:14.447164Z digest=sha256:1c81772003958ae2d7abb469057838983e6727cf79773c6c2e7ae1de83083733

Observation 07cf10e3-67cb-4e75-beef-b94f2606b45f · inbound

Single Domain Generalization in Diabetic Retinopathy: A Neuro-Symbolic Learning Approach cites this paper.

Single Domain Generalization in Diabetic Retinopathy: A Neuro-Symbolic Learning Approach CvT: Introducing Convolutions to Vision Transformers

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T11:24:06.954661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:24:06.954661Z digest=sha256:be96dff1e717253bc01060312fa3d2c366902c66df9fbbdec31dba28f226d2c5

Observation 3c04529a-b094-413f-a3a5-dce46cf245bc · inbound

Dual-attention ResNet outperforms transformers in HER2 prediction on DCE-MRI cites this paper.

Dual-attention ResNet outperforms transformers in HER2 prediction on DCE-MRI CvT: Introducing Convolutions to Vision Transformers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T09:55:52.457277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:55:52.457277Z digest=sha256:8f82c099bffc83a63595f576fbb2672cba18dfee6b47abe01904ffd54df48173

Observation f0367cbd-7e53-4350-8a75-ec95f82c3a95 · inbound

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents cites this paper.

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents CvT: Introducing Convolutions to Vision Transformers

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.786479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-05T11:32:36.356530Z digest=sha256:efbddb454b079118605919e0d574d1ce3c49fa3b621c954d3225ecf418959afa

Observation 422776b5-69fd-47eb-942a-37aec0eb6af9 · inbound

Advancing Vision Transformer with Enhanced Spatial Priors cites this paper.

Advancing Vision Transformer with Enhanced Spatial Priors CvT: Introducing Convolutions to Vision Transformers

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.655388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T05:22:21.264807Z digest=sha256:d643b9445d8e21e904b946b9bfe30db561fe352359cba3f999c5a8e5dfd8b08c