Pith. sign in

Paper Citation Record · LEDGER

Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2307.06304.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.06304 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:15:22.753088Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:37:14.474189Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cbf8093d-d53e-4fb6-af22-2200c3191c3a · inbound

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding cites this paper.

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:20:00.250289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T01:19:59.603343Z digest=sha256:b5b1833f96df374811a19839581ed6c2dfcb171b94d997d259ae66863e2d2f13

Observation a0eef25d-31f6-4a4f-ac30-1290f02b085f · inbound

Native-Resolution Image Synthesis cites this paper.

Native-Resolution Image Synthesis Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.753088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.753088Z digest=sha256:be5bdff2075582c7e3577cb728710effedecca0f392b8c5b098d45fe9c4811ad

Observation 4f99b5eb-d2a5-4734-9c52-5f71ef917b08 · inbound

Transition Matching: Scalable and Flexible Generative Modeling cites this paper.

Transition Matching: Scalable and Flexible Generative Modeling Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:46:25.589904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:46:25.589904Z digest=sha256:ea78e99f363cb8e426921964c20217a969761e068148d112e877b560de9613b6

Observation 755ef971-28dc-4cfa-bb36-474e27f4b3d3 · inbound

Kimi K2.5: Visual Agentic Intelligence cites this paper.

Kimi K2.5: Visual Agentic Intelligence Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:09:05.293386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:09:05.225767Z digest=sha256:57f3c0f776aea0dd272e41133adef204c0629d3209c90e1b2fc4adf2a039a0ce

Observation 748fd67e-8938-4509-85ac-eea615b63bbd · inbound

MegaScale-Omni: A Hyper-Scale, Workload-Resilient System for MultiModal LLM Training in Production cites this paper.

MegaScale-Omni: A Hyper-Scale, Workload-Resilient System for MultiModal LLM Training in Production Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:06:14.948067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:04:07.344134Z digest=sha256:b9bec3a043dbb6464bae03100106e618c50e43e00449b8d64840890b69a5bff9

Observation fd6e7296-5d9f-4715-ba3c-42f40a2ee28a · inbound

MultiMedVision: Multi-Modal Medical Vision Framework cites this paper.

MultiMedVision: Multi-Modal Medical Vision Framework Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:24.710887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:46:23.767997Z digest=sha256:029715557470b8f768db9ae8092c3434593692185ba47932e9d0b6d6b939ac02

Observation 9230ab26-e9a9-445e-8531-8ac85880edb2 · inbound

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models cites this paper.

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.475654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T21:58:53.702009Z digest=sha256:2ac572c5c5d91c915fbca2a6cd1f23ea86cb290b4c601e00025fbd9f284c8911

Observation e447ce7f-b94b-4391-95d7-c01f3e347390 · inbound

HorusEye: Language as Dynamic Attention for Emergency Visual Analysis cites this paper.

HorusEye: Language as Dynamic Attention for Emergency Visual Analysis Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T04:53:26.058308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:53:26.058308Z digest=sha256:2c5251cd49994b1b0fa1943d577d5d3778217794e9655fe7f5d20c771cff83d9