Pith. sign in

Paper Citation Record · LEDGER

ViTAR: Vision Transformer with Any Resolution

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2403.18361.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.18361 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:45:01.870482Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f7e03d8f-84fd-409f-9ca4-ba062b00cb5c · inbound

SeqPE: Transformer with Sequential Position Encoding cites this paper.

SeqPE: Transformer with Sequential Position Encoding ViTAR: Vision Transformer with Any Resolution

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:01.870482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:45:01.870482Z digest=sha256:83dc01d8260bbca30f13f63e31ea5fe7500b2561faab9e8710750caf181faf71

Observation 5f174a8f-2c61-4299-ba3f-d8c777647684 · inbound

Tile-Based ViT Inference with Visual-Cluster Priors for Zero-Shot Multi-Species Plant Identification cites this paper.

Tile-Based ViT Inference with Visual-Cluster Priors for Zero-Shot Multi-Species Plant Identification ViTAR: Vision Transformer with Any Resolution

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:17:45.218897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:17:45.218897Z digest=sha256:6e52fb97c2e684c075c656429e180751a597b6f4ef02902a0480b0a77dc09fa7

Observation bdf6c9be-9284-4ece-b053-17471d34d9bc · inbound

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations cites this paper.

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations ViTAR: Vision Transformer with Any Resolution

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:51:11.513644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T04:12:58.100610Z digest=sha256:f0ae18921f90fbb0b7ae0bd5cefa00687b3331538ef348d42626ffd78b0efda7

Observation c490c0e5-8ae0-4d63-b0df-af85a8adc81e · inbound

On What We Can Learn from Low-Resolution Data cites this paper.

On What We Can Learn from Low-Resolution Data ViTAR: Vision Transformer with Any Resolution

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:32:18.963370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T05:28:50.993737Z digest=sha256:5903335b6698c78b33c0339beabd132baf99890b5068e8ec9a33e0ed5d5a841b

Observation 0951599c-216e-44f1-9d26-5b1037c3ce6c · inbound

Weighted Reverse Convolution for Feature Upsampling cites this paper.

Weighted Reverse Convolution for Feature Upsampling ViTAR: Vision Transformer with Any Resolution

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:28:21.679569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T14:24:28.728963Z digest=sha256:207525a925acd3d96d78dbdba879655377fe49dd9bc824c51488201178efe312

Observation 52e5ca84-2b47-4969-b1ef-fd895338f01a · inbound

Weighted Reverse Convolution for Feature Upsampling cites this paper.

Weighted Reverse Convolution for Feature Upsampling ViTAR: Vision Transformer with Any Resolution

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:19:52.836537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T08:15:52.610554Z digest=sha256:d3eec27cdec0a38ce53ee91774c93a26ee92745e35a390904c3f6468ba7c94be

Observation 31a842b3-a9d4-41c2-89c2-71b8578d0212 · inbound

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions cites this paper.

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions ViTAR: Vision Transformer with Any Resolution

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T15:24:50.174266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T15:17:37.904831Z digest=sha256:be510eacc5517244281a866a9ad68e3ab45ef62b81bf4a3dc86566eaa7ecf36a