Pith. sign in

Paper Citation Record · LEDGER

ViTAR: Vision Transformer with Any Resolution

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2403.18361.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.18361 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:45:01.870482Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f7e03d8f-84fd-409f-9ca4-ba062b00cb5c · inbound

SeqPE: Transformer with Sequential Position Encoding cites this paper.

SeqPE: Transformer with Sequential Position Encoding ViTAR: Vision Transformer with Any Resolution

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:01.870482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:45:01.870482Z digest=sha256:32901b5d9e163078fab9b0a4e8d46c3df8370f0808d5d14d685fe0dbe2736ea3

Observation 5f174a8f-2c61-4299-ba3f-d8c777647684 · inbound

Tile-Based ViT Inference with Visual-Cluster Priors for Zero-Shot Multi-Species Plant Identification cites this paper.

Tile-Based ViT Inference with Visual-Cluster Priors for Zero-Shot Multi-Species Plant Identification ViTAR: Vision Transformer with Any Resolution

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:17:45.218897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:17:45.218897Z digest=sha256:6e52fb97c2e684c075c656429e180751a597b6f4ef02902a0480b0a77dc09fa7

Observation bdf6c9be-9284-4ece-b053-17471d34d9bc · inbound

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations cites this paper.

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations ViTAR: Vision Transformer with Any Resolution

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:51:11.513644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T04:12:58.100610Z digest=sha256:5817a6de9ec4d6b96fe10153919aae85b6fb2f5b8bc5bab86eeb93292756a31e

Observation c490c0e5-8ae0-4d63-b0df-af85a8adc81e · inbound

On What We Can Learn from Low-Resolution Data cites this paper.

On What We Can Learn from Low-Resolution Data ViTAR: Vision Transformer with Any Resolution

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:32:18.963370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T05:28:50.993737Z digest=sha256:d55455f41ac2151fbb50e4827ad653404ef07f9672dc882692e9248a5cc88ccb

Observation 0951599c-216e-44f1-9d26-5b1037c3ce6c · inbound

Weighted Reverse Convolution for Feature Upsampling cites this paper.

Weighted Reverse Convolution for Feature Upsampling ViTAR: Vision Transformer with Any Resolution

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:28:21.679569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T14:24:28.728963Z digest=sha256:536bed6e9ec4b1936eb02b1a930058ef02f36552d5b41192fd4369934b9e6c1b

Observation 52e5ca84-2b47-4969-b1ef-fd895338f01a · inbound

Weighted Reverse Convolution for Feature Upsampling cites this paper.

Weighted Reverse Convolution for Feature Upsampling ViTAR: Vision Transformer with Any Resolution

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:19:52.836537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T08:15:52.610554Z digest=sha256:229d5462a26a1a407bb1cd66ee73100762a32e4ccc2e80ccdd0c478c214e4aa1

Observation 31a842b3-a9d4-41c2-89c2-71b8578d0212 · inbound

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions cites this paper.

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions ViTAR: Vision Transformer with Any Resolution

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T15:24:50.174266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T15:17:37.904831Z digest=sha256:9048b3cc3179ab8af959b0d57744a8c34ff33c4e2ef9a0cbf7749531aee6a798