Pith. sign in

Paper Citation Record · LEDGER

DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2106.02034.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2106.02034 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:53:24.566993Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:36:24.028635Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eadc14ab-b024-49fc-b509-35fcd455a5a5 · inbound

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer cites this paper.

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:46:35.159941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T20:46:35.073600Z digest=sha256:0516440df2513eafe5a0665933078056d85450a51fd5947780b9d8d8ee008465

Observation fd173311-25a9-4ddf-b0af-d96955baa203 · inbound

Compact Vision Transformer by Reduction of Kernel Complexity cites this paper.

Compact Vision Transformer by Reduction of Kernel Complexity DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T16:48:48.087616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:48:48.087616Z digest=sha256:8b75119ba05aef7e7cb243c2f1c029edec6ae95986c4f6b9f46db8c88d0e8020

Observation 1cbac0c9-ca1c-4f56-adb7-065b17f59cd7 · inbound

CascadeFormer: A Family of Two-stage Cascading Transformers for Skeleton-based Human Action Recognition cites this paper.

CascadeFormer: A Family of Two-stage Cascading Transformers for Skeleton-based Human Action Recognition DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:13.649881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:24:13.649881Z digest=sha256:9fcc372b6cd0bc248468f93c918b5c91d6d9c3bcb291c2887409ee7194c71294

Observation 16108892-33f1-4062-9814-51b0cd456f28 · inbound

DC-DiT: Adaptive Compute and Elastic Inference for Visual Generation via Dynamic Chunking cites this paper.

DC-DiT: Adaptive Compute and Elastic Inference for Visual Generation via Dynamic Chunking DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:10:05.849862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T15:10:00.146694Z digest=sha256:5985b49f963326a063153c1c87df692a6fc8d32f2721bd6c44a4d68815ea3c2d

Observation e03e55f8-3645-4a8a-b5d1-50e7e61b4a65 · inbound

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs cites this paper.

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:24.030852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T14:04:43.270155Z digest=sha256:8ebc9230ef63aafd916577dd2f8b29ac472b4dd9be1cac503c489e1aa1c955ff

Observation fa90b629-bd8d-4f05-8132-fcc0ef69b948 · inbound

Foveation-Guided Dynamic Token Selection for Robust and Efficient Vision Transformers cites this paper.

Foveation-Guided Dynamic Token Selection for Robust and Efficient Vision Transformers DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T02:43:47.779443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T02:43:47.779443Z digest=sha256:2287b8db2571a6bf588aa89f915d15e935944f287b134128feb0a0c657bd452a

Observation 24449e0d-0655-4c2c-b96f-e4c7f936677e · inbound

Searching for Task-Specific Vision Paths: Evolutionary Block Pruning Across Vision-Language Models cites this paper.

Searching for Task-Specific Vision Paths: Evolutionary Block Pruning Across Vision-Language Models DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-01T19:13:27.889997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:13:27.889997Z digest=sha256:1160da6bb9f0bc9ec6a735e100dd08a3397521e546eb85fd74671329a8df7a09

Observation e95ad9f4-fe05-433b-9d6a-1e36fc9ffd25 · inbound

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models cites this paper.

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T23:46:51.535123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T23:46:51.535123Z digest=sha256:83defd16dac0a39e1df69cd154660c9d98dff2f598d1a53dfa658500209744db

Observation 5cf19484-a609-4667-beae-3740c334d917 · inbound

LaPrune: Controllable Differentiable Sparsity at Million Scale cites this paper.

LaPrune: Controllable Differentiable Sparsity at Million Scale DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T00:53:24.566993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:53:24.566993Z digest=sha256:4f10df291df686a5c03e3d89ed5cb4712ddf3cafa33ba0ab48d4eaa116bf2457