Pith. sign in

Paper Citation Record · LEDGER

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling

As of 10 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2502.02919.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.02919 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:41:01.420244Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e661855d-07eb-4557-b828-2c81dc1a4a5c · outbound

This paper cites Layer Normalization.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Layer Normalization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.239337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.239337Z digest=sha256:aca54a6bb029617b66ae299830c40eba7f00e2d085dcf75a0c1fe47e80eb8a68

Observation ec137b26-5003-4bd1-b0fd-29606f6bfc7d · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:02.009344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.246337Z digest=sha256:3bc2090ace190796459b105afb02146ad696ba162ec4be85ba6749c36d042e82

Observation ec6f87cd-03b7-44fe-8263-539446193a3d · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.251977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.251977Z digest=sha256:a237ca4fdd4c8762eed9b90cc89c6815b3adc9785f961c6c81f432403b7c5da4

Observation 16c3790a-fb14-424e-8ea9-bc6217ab2907 · outbound

This paper cites J.; Jin, R.; and Shou, M.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling J.; Jin, R.; and Shou, M

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:41:01.982712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.257588Z digest=sha256:2b5f00251c4c9d62ecec671d586a85d6656b225716278979fdd87a95ea98a2f4

Observation 6ee06156-ff64-425f-b760-4c1971901098 · outbound

This paper cites MMDetection: Open MMLab Detection Toolbox and Benchmark.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling MMDetection: Open MMLab Detection Toolbox and Benchmark

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.262696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.262696Z digest=sha256:52619d5d129eff506cd358d57b1b5e6f9986f976e45ac64b4705725ea8101aa4

Observation 963676bc-1edb-48fd-922d-7aef6b7701bd · outbound

This paper cites Vision Transformer Adapter for Dense Predictions.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Vision Transformer Adapter for Dense Predictions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.268204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.268204Z digest=sha256:cd1377d2901cddca95d3c757c8d057d886022a2f0c009964e3fdc2b6f3b6d0a2

Observation eefb3eef-0b7d-4a45-8348-3a7a6d85e486 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.966546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.274027Z digest=sha256:29d61fcfd91ac7894cf8c51751f3ecdd59489fa2b63188f533cbdc308025a34b

Observation 2a89c1d7-1ebc-4665-b812-c104c6b8916a · outbound

This paper cites Conditional Positional Encodings for Vision Transformers.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Conditional Positional Encodings for Vision Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.278794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.278794Z digest=sha256:38c9227eeb176b1c093927b267529001e8758aa71489b3a7b79101f2863067cb

Observation 0e20669a-ab27-43aa-bf4a-b58d34e2d938 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.949151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.284003Z digest=sha256:35257f896fa590d935aa7daafd7c603b38763e27c16a9a8cec2eca27e9fed130

Observation f6cab2d6-8f57-4e7f-9bce-75c1ea147823 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.288513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.288513Z digest=sha256:da7b2007862da41785230736d957ed439df686af1c6112b556265e814d51ec20

Observation 591e8f8a-cc38-47e5-9a9f-73e7ee6af321 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.293518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.293518Z digest=sha256:a066902c5eda72777b78afcee5b79a311f5a45a66f8bc18319680328a1393be4

Observation 717870e9-5ab3-4364-a5f6-c01970772fe0 · outbound

This paper cites Escaping the Big Data Paradigm with Compact Transformers.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Escaping the Big Data Paradigm with Compact Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.298875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.298875Z digest=sha256:8f0d104820666fed47182fa2ccde0e76ccea6acec986dc1b96759fac8c0a6c80

Observation 2068eef1-69cf-4add-b4fb-0f2d8febc9c1 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.303810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.303810Z digest=sha256:03d70452125a728aec833cd8401d5bf66a6dd6efac5fe364afe5d6544422caba

Observation f6034912-b7f0-4ccc-b819-7b146fb90ef1 · outbound

This paper cites Rotary Position Embedding for Vision Transformer.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Rotary Position Embedding for Vision Transformer

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.308538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.308538Z digest=sha256:b4d7193788459702960cd8653320e1f498af94e25155a628438e0b717d6dd075

Observation 0df307c7-248e-4f89-b52d-40080797030a · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.313341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.313341Z digest=sha256:b23388b32cf5fd4f36bdf9d5ab87faa288236a81205ae739996c56bdd23dea25

Observation de472f31-358d-4b32-9ad3-a6e256d5b830 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.318568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.318568Z digest=sha256:a64c31a39b1949e3aeb80d5901a1903cbf25c3878e65fa9d089ba3929e7ddb2a

Observation 21c0907a-0ba6-42d3-8c64-ad4e73036f87 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.323421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.323421Z digest=sha256:9f34246181ee4de3f03c671056ea975fe77a1458560c2abea00fd6e32be6d225

Observation 5a1140ae-6822-40c5-b119-1a01ee5eb784 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.328114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.328114Z digest=sha256:42210afa40804b582b089116dcf6476563f406cd6a82b4df14f49e0df557ac18

Observation 47b62d92-d59e-4398-9df3-b27b436c0e61 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.332753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.332753Z digest=sha256:acdc98ab760e20171aaf74ba48be2c0112e64700b9bd61d89837a1238f2d0895

Observation 14eaa9f1-5f2c-4801-9767-c27a832e012b · outbound

This paper cites Self-Attention with Relative Position Representations.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Self-Attention with Relative Position Representations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.337970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.337970Z digest=sha256:43f558b68da5521b119332dc234e8290380b0ec49cd21d0e2e394111bbf15f91

Observation f0fedcb6-1420-47ea-96b3-e9e676f782b9 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.343091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.343091Z digest=sha256:f39a51c6dcb7a2a9b1d00d1b200542da7c40fcc476435cfcf919fc9a4598916f

Observation 487896cd-b4f6-4c1d-b708-342e75481ac1 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.347656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.347656Z digest=sha256:319e359a5de131d88904c2164919404072797b857d4c9a3b0d295959871594a9

Observation e051b9ba-25d3-496d-825d-f5af358bf046 · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling N.; Kaiser, .; and Polosukhin, I

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.355182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.355182Z digest=sha256:1a0fcc6323f0c066778dcbb51c7f470f084555797ecfeeb1f69d6057c1556aec

Observation 85715c03-046f-4f1c-a2f1-ebb7ab802e61 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.815849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.360381Z digest=sha256:7f8380d08cfdbfafea71697ea8a2eb3ff247863bd1ed0118b0cfc27258bde0b0

Observation f91d2907-7005-45c4-af58-0edc020bdebb · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.797329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.365788Z digest=sha256:fe522dc46059dada20fc79da79099452726bde24bca369a8eccf0d76478d170c

Observation 81bbc29a-c45d-4554-9e30-7dffb78d04f7 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.779959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.370892Z digest=sha256:e584eb512f6aa80702170e6ad9e0cd9cfccc9ce8d5b75247d6b17e7f25c091e3

Observation d8969e2a-c39c-4bc5-a0a5-1f28b1d91e35 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.762467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.375825Z digest=sha256:751b604d518b6c466ba6c5fc4f2cecc03de6ab48a99259d2d429eb3d32827a81

Observation ee22e8d4-7482-4705-8de2-fd414c093775 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.742833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.380740Z digest=sha256:00c91f5b1e23387a0b96e3126bef61d0e5c65710c2d4581bdb1953a9e5db1c26

Observation 7903436a-3741-458a-9701-761b25010063 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.724563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.385331Z digest=sha256:9c1ab76d1450b47602245d7da34eb9429e475a756c94b657c866f758404a9f3e

Observation 4ecedc5b-3b8d-46ea-bf31-0f6c09762449 · outbound

This paper cites E.; Feng, J.; and Yan, S.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling E.; Feng, J.; and Yan, S

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:41:01.708349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.390265Z digest=sha256:467133e0e6c643d9aeeffb9350fa1fd1b872f3b7f16b2432f3c1e0c06a4e0e01

Observation 74b25d0e-f68f-4fb9-9f03-765cd1e38a1b · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.692463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.394845Z digest=sha256:2b2f6cd9d981980e649291f97d294b60f7b8e6a37f3db139754af8ce2ba53386

Observation b5f0ee3c-3f15-45f1-a2e3-e3ca2ab611fa · outbound

This paper cites H.; et al.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling H.; et al

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.399466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.399466Z digest=sha256:72f5466fb6c1c64582d9025bf9f208b88a89f558c1f5cf0bafc28dda3de86826

Observation de705d0b-0541-444a-ac28-c90a3343f01d · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.664832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.404419Z digest=sha256:9e91c683a0f4cb7a58fda7052a02354298ecb1e45d6130c3d5a41af48245d0f8

Observation f502618b-00cb-4b64-a0d0-66b3b71f479f · outbound

This paper cites Deformable DETR: Deformable Transformers for End-to-End Object Detection.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Deformable DETR: Deformable Transformers for End-to-End Object Detection

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.409443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.409443Z digest=sha256:9f229cd06df740ea10ca9d09b1a4cd1d9a7d7edb37cbc8d4beacf42b20c1f8b4

Observation 6cdc534d-677f-4fea-80a0-f0de778b9d9a · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling , " * write output.state after.block = add.period write newline

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.414853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.414853Z digest=sha256:db7b42d6d3d6e175a1d1c083613077501ff0f26728349dd4a5c4dc356a60b310

Observation 4efd5ca3-de0c-45f2-9cb1-559c22e23405 · outbound

This paper cites write newline.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling write newline

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.420244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.420244Z digest=sha256:ee9d2783f99ef2b0d6ebe1a444a8621edad9cbb1a0921c3169d51d771955d2a7

Pith citing papers

No inbound Pith citation observations are available.