Pith. sign in

Paper Citation Record · LEDGER

Token Pooling in Vision Transformers

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2110.03860.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2110.03860 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:23.781218Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:39:29.498864Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1ea5df7b-2946-42a2-a55a-76c068cfcb39 · inbound

Token Merging: Your ViT But Faster cites this paper.

Token Merging: Your ViT But Faster Token Pooling in Vision Transformers

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T20:52:10.491697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T20:52:10.065925Z digest=sha256:d77984683f7dc559d0be5944cfe0f07e3deacdf8b2a3bfe6e4a2f1cc702e4d7e

Observation 788a1717-40dd-4ee4-8aa0-f37e6cdbb4a8 · inbound

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation cites this paper.

Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation Token Pooling in Vision Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:23.781218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:23.781218Z digest=sha256:0eb9a66a1a35b018f2e6f30960362f57cd2ae0e149a93cf7eae905eedfa64da2

Observation ea941ac2-f62c-44be-937b-e125dc1e5b6e · inbound

D\'ej\`a Vu: Efficient Video-Language Query Engine with Learning-based Inter-Frame Computation Reuse cites this paper.

D\'ej\`a Vu: Efficient Video-Language Query Engine with Learning-based Inter-Frame Computation Reuse Token Pooling in Vision Transformers

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T00:26:12.625665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:26:12.625665Z digest=sha256:7dd445d0b619a51c21c52d0a4f4cbd3707767b6123eb8b6379469cc096c908c7

Observation accbf4d4-a0f0-4cef-9e4c-d95024abc0e1 · inbound

Geo-RepNet: Geometry-Aware Representation Learning for Surgical Phase Recognition in Endoscopic Submucosal Dissection cites this paper.

Geo-RepNet: Geometry-Aware Representation Learning for Surgical Phase Recognition in Endoscopic Submucosal Dissection Token Pooling in Vision Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:10.871048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:03:10.871048Z digest=sha256:593a707c60bb8883baf7c9dc0a2a96ca0e95a018c38d99aa504fdfa5515d97b1

Observation ba497e75-eca8-426b-9958-56a0f92e0bc5 · inbound

Training-free Token Reduction for Vision Mamba cites this paper.

Training-free Token Reduction for Vision Mamba Token Pooling in Vision Transformers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T16:16:49.332131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:16:49.332131Z digest=sha256:9f421c4e6dfb0d8e6dedf7726d8385d0f199fea59c5cd3cfc453462243fff601

Observation cd5d1a64-c8a4-45f3-b6f7-43293047e6dd · inbound

Why Training-Free Token Reduction Collapses: The Inherent Instability of Pairwise Scoring Signals cites this paper.

Why Training-Free Token Reduction Collapses: The Inherent Instability of Pairwise Scoring Signals Token Pooling in Vision Transformers

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:02:25.409065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T07:59:10.678150Z digest=sha256:7022519d76b333db12f8ab8276a47d102faf42952689ea8e0ae7865f37520747

Observation e50f609d-fa0e-4dbe-83f0-2c2191315f8a · inbound

Accelerating Vision Foundation Models with Drop-in Depthwise Convolution cites this paper.

Accelerating Vision Foundation Models with Drop-in Depthwise Convolution Token Pooling in Vision Transformers

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:04:41.468875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T07:03:37.388861Z digest=sha256:1b3b4506f794e10d766c9ab286c74e994be6fe6ff9c6b6d2d70ee9ecd0049a53

Observation 6e6994dc-e9df-4e12-8564-d5f1846809a0 · inbound

RAPID: Layer-Wise Redundancy-Aware Pruning and Importance-Driven Token Merging for Efficient ViT cites this paper.

RAPID: Layer-Wise Redundancy-Aware Pruning and Importance-Driven Token Merging for Efficient ViT Token Pooling in Vision Transformers

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.484522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T19:46:23.188930Z digest=sha256:a41df501e1c1bcf840141138b39e90faf8fba6a9d986134b13596b3bece706a8

Observation 687b7ebe-694a-4e53-b987-c7eb81cf76dd · inbound

Spatial-Aware Reduction Framework: Towards Efficient and Faithful Visual State Space Models cites this paper.

Spatial-Aware Reduction Framework: Towards Efficient and Faithful Visual State Space Models Token Pooling in Vision Transformers

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:39:29.501847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T17:53:38.503877Z digest=sha256:a227ff154863f9c8c75d02bd7fea4a6f72ff21e112f8005aaaac38a5b9cb6376

Observation 3d28a755-b435-451d-a2aa-e8fb81fd220e · inbound

DTM-Codec: Dynamic Token Masking for VFR Speech Coding with Efficient Boundary Selection cites this paper.

DTM-Codec: Dynamic Token Masking for VFR Speech Coding with Efficient Boundary Selection Token Pooling in Vision Transformers

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T02:14:11.489247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T02:10:08.691873Z digest=sha256:1374d4433eb36016b54143b3f5ddeed9b02c2d560164da858b73db3689f0d6a5