Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:36:24.024024Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 2 inbound Pith citation observations for arXiv:2505.09466.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:36:24.024024Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T03:24:59.868006Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
29 of 29 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation cd1b5b9e-6341-4f20-8b02-1ee2886cdf57 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Gomez, Łukasz Kaiser, and Illia Polosukhin
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fc6e9586-20f6-4370-8d02-f3441e01ada8 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 906fac3d-049f-44bf-a8fa-6c30ef0d9671 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Swin transformer: Hierarchical vision transformer using shifted windows
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ae4600c-3bde-4971-a901-a4350fc96269 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers An Introduction to Convolutional Neural Networks
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb579dfd-3b1a-47aa-ac69-d54769edda78 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers A survey of convolutional neural networks: analysis, applications, and prospects
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b6c9f95f-04e6-42ad-9528-9913bee3b495 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers End-to-end object detection with transformers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4370796-ab00-4037-9b60-b5e05f50bd18 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Convolutional sequence to sequence learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 39c52fd3-6693-4ce0-a5f2-8caf25e3de15 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Self-Attention with Relative Position Representations
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58d84363-93b7-4351-9ac8-95d3d762a5b4 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Exploring the limits of transfer learning with a unified text-to-text transformer
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85863d54-5b64-409d-ae17-b9000d32696f · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Roformer: Enhanced transformer with rotary position embedding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 456d29df-0870-4d50-8866-9e2f81bccdcf · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Contextual Position Encoding: Learning to Count What's Important
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db5f0865-fb75-4f5f-8d37-b8a964fb3f7d · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Rotary position embedding for vision transformer
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0ad9b095-7edb-4b29-98e7-a269f547f01a · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Tokens-to-token vit: Training vision transformers from scratch on imagenet
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34694718-ad46-4887-8389-89f3b8e1c3d0 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Training data-efficient image transformers & distillation through attention
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58b16f48-2b77-4894-b7af-dcb475bdc746 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Pyramid vision transformer: A versatile backbone for dense prediction without convolutions
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a3ed9bd-6661-466b-a91d-b28cbe50b83f · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Rethinking spatial dimensions of vision transformers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 783e7021-b934-4616-9483-6f54f9266815 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Cvt: Introducing convolutions to vision transformers
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18511900-5c0b-4433-b983-461ac0e4ec9c · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Incorporating convolution designs into visual transformers
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a91028e5-189b-4ce9-a799-a62cc3ead9b4 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Deformable DETR: Deformable Transformers for End-to-End Object Detection
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ba2335a-d9bb-441c-af8c-39a28af4863a · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers End-to-End Object Detection with Adaptive Clustering Transformer
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00f1041f-4053-4993-9d9c-2e15af692402 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Fast convergence of detr with spatially modulated co-attention
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 35ef02bd-7e58-4ddf-a11b-0b0d195462af · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Exploring plain vision transformer backbones for object detection
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d76d772f-3245-4030-b0a2-951a991b79cc · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Vision transformer adapter for dense predictions
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6526cece-c906-45b8-8a6d-d6f2228911d3 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b6d8ed79-3a7e-4815-9f99-b270636912b5 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Segmenter: Transformer for semantic segmentation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 85528404-788c-463f-b260-c52dd60bff93 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Segformer: Simple and efficient design for semantic segmentation with transformers
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48fe5ba8-7ced-43d6-bb40-f1525e6399d6 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Rethinking and improving relative position encoding for vision transformer
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b73f2579-21cd-4093-8a7b-a01cf6556784 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Lape: Layer-adaptive position embedding for vision transformers with independent layer normalization
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0729aed1-4fbc-4e0b-99b5-8fd1a9e4d851 · outbound
A 2D Semantic-Aware Position Encoding for Vision Transformers Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1516d0cf-18a8-4eb9-a0c6-551a561e9bde · inbound
Active Spatial Guidance: Eliminating Injected Positional Mechanisms in Vision Transformers A 2D Semantic-Aware Position Encoding for Vision Transformers
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e1fd5f79-d3dc-4291-bb49-048c01f8de92 · inbound
Out-of-Length Scene Text Recognition: A Two-Axis Diagnosis and a Training-Free Fix A 2D Semantic-Aware Position Encoding for Vision Transformers
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.