Pith. sign in

Paper Citation Record · LEDGER

SiT: Self-supervised vIsion Transformer

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2104.03602.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.03602 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:22:07.487328Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T12:36:22.499455Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dab44aef-1362-4631-98cb-7e3f6f4693dc · inbound

iBOT: Image BERT Pre-Training with Online Tokenizer cites this paper.

iBOT: Image BERT Pre-Training with Online Tokenizer SiT: Self-supervised vIsion Transformer

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-14T02:10:27.640003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T02:10:27.569787Z digest=sha256:27c01bf89dee6202e3b4b0b53fe8bead9c9283b27ef8d4619796d5e0a56d7db8

Observation ac7d8c13-6389-4aa4-8230-681d99bf37db · inbound

FlowAR: Scale-wise Autoregressive Image Generation Meets Flow Matching cites this paper.

FlowAR: Scale-wise Autoregressive Image Generation Meets Flow Matching SiT: Self-supervised vIsion Transformer

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T11:36:26.159137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:36:26.159137Z digest=sha256:13c688574138ccb4118620db081478a1dcae69a0b535e70ca4f3be5334f1a0f3

Observation ac797521-9667-46c4-ba6b-600ef5878036 · inbound

State-of-the-art AI-based Learning Approaches for Deepfake Generation and Detection, Analyzing Opportunities, Threading through Pros, Cons, and Future Prospects cites this paper.

State-of-the-art AI-based Learning Approaches for Deepfake Generation and Detection, Analyzing Opportunities, Threading through Pros, Cons, and Future Prospects SiT: Self-supervised vIsion Transformer

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T22:39:19.678345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:39:19.678345Z digest=sha256:ff5062405e0af850cf412518a1ccf3b060f37a64212d7f0e6770d440b75f6c82

Observation 3917b2b5-a13e-4bc7-8df5-c2f6354e4967 · inbound

Powerful Design of Small Vision Transformer on CIFAR10 cites this paper.

Powerful Design of Small Vision Transformer on CIFAR10 SiT: Self-supervised vIsion Transformer

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:58:02.162977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:58:02.162977Z digest=sha256:de0ede60d2eba0e544a71e69dde7d0ccc2228f3afd1676ff0f665aed73e83d0e

Observation ec637453-89cc-459b-88f8-9e10cb04b1c7 · inbound

ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers cites this paper.

ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers SiT: Self-supervised vIsion Transformer

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T20:12:58.118649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:12:58.118649Z digest=sha256:2b36ae4b390290457e254907ad067363f088793ddb3b820233a5936c7072662a

Observation fd901f05-5191-47fc-9d71-3c40c2d50250 · inbound

With Great Backbones Comes Great Adversarial Transferability cites this paper.

With Great Backbones Comes Great Adversarial Transferability SiT: Self-supervised vIsion Transformer

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T17:24:21.471935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:24:21.471935Z digest=sha256:7b5459a36c228a5fce16a792af01fbccd03c1b4cb61ac6808cc1c5bfbb403262

Observation 357d7c57-74e8-4922-8dd1-10141a3cb1e0 · inbound

Learning Counterfactually Decoupled Attention for Open-World Model Attribution cites this paper.

Learning Counterfactually Decoupled Attention for Open-World Model Attribution SiT: Self-supervised vIsion Transformer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:08.112520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:55:08.112520Z digest=sha256:3ed5b36efa0907c2fcbcbf954869c5cd043a0329c788a6fbe055711310667a90

Observation 21a3cb4a-5a8b-41fe-8572-5188631b2bcf · inbound

Self-Guided Masked Autoencoder cites this paper.

Self-Guided Masked Autoencoder SiT: Self-supervised vIsion Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:52.346880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:13:52.346880Z digest=sha256:0d952682e96f8b4ef3f92762e02cd61486ce7ba2affa77743e40df2d1e9f998b

Observation 5733ff07-5587-499d-8c14-dd5a3bb5a327 · inbound

PRPO: Paragraph-level Policy Optimization for Vision-Language Deepfake Detection cites this paper.

PRPO: Paragraph-level Policy Optimization for Vision-Language Deepfake Detection SiT: Self-supervised vIsion Transformer

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:36:22.502170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T12:34:39.312273Z digest=sha256:8831c2eb46ebee833af785f0b7d4fa84e4d73404fb1cd07bb7f7044778c9d66b

Observation 97fe6845-7b71-4f64-8a47-b26dd42b37bb · inbound

From pre-training to downstream performance: Does domain-specific pre-training make sense? cites this paper.

From pre-training to downstream performance: Does domain-specific pre-training make sense? SiT: Self-supervised vIsion Transformer

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:37:08.547116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T02:26:58.921154Z digest=sha256:40b5813c1e1a09a3037cb5903758911a0f6d3e7314cc5ff1c71917dd673d4cde

Observation 18301d1a-e6fc-4454-810f-66927207702d · inbound

NAE: Normalizing AutoEncoder cites this paper.

NAE: Normalizing AutoEncoder SiT: Self-supervised vIsion Transformer

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T00:22:07.487328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:22:07.487328Z digest=sha256:4f1241a819e951483bf330af9addb58a4cfcee8c750e7c05ed2ecc2169ec70b7