Pith. sign in

Paper Citation Record · LEDGER

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing

As of 10 August 2026, this Paper Citation Record lists 94 of 94 outbound references and 0 inbound Pith citation observations for arXiv:2502.00594.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.00594 v1

Coverage vector

measured 94 of 94 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:28:48.679484Z

measured 94 of 94 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

94 of 94 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved56
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5bbf912a-01ae-4e9d-bd57-95fc243458a6 · outbound

This paper cites Hiervl: Learning hierarchical video- language embeddings.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Hiervl: Learning hierarchical video- language embeddings

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.027458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.027458Z digest=sha256:bb24ca86e88114c2a9fac9605c4d8900ebed2ff6449b4cbb466289f10a67c140

Observation 9949af15-89e3-4f4a-b3af-6d505aaabfab · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing BEiT: BERT Pre-Training of Image Transformers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.077510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.077510Z digest=sha256:e091b1dd4cc5d93f0f079ae1fd3f8b9d38d0722970bf19524d8722bb4e861299

Observation f3c30f5f-6149-4898-bfc4-b0c156b7a00d · outbound

This paper cites Channel Vision Transformers: An Image Is Worth 1 x 16 x 16 Words.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Channel Vision Transformers: An Image Is Worth 1 x 16 x 16 Words

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.206021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.206021Z digest=sha256:b90b2be640a4c5dc9b2bac9c9f78ad6b0adba840fd148a9077aab95177f3d33d

Observation 97edd187-1858-446e-9b68-f4f24fc8d4f8 · outbound

This paper cites Hypermae: Modulating implicit neural representations for mae training.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Hypermae: Modulating implicit neural representations for mae training

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.211633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.211633Z digest=sha256:dc3c1ab9f85ba4fc5657da1412144086b06487f558a070430ef2340b278638a5

Observation ca69a3bd-37c7-4b68-9b87-5d6042ae02dd · outbound

This paper cites Prefix sums and their applications.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Prefix sums and their applications

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.216579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.216579Z digest=sha256:7b3d8f0b82b4be2308d930d0dcd2d8261165c8e05389e2a830d9a7450e85cd65

Observation 54ff10f1-55c4-4078-84ea-6e34e694eb1f · outbound

This paper cites Token Merging: Your ViT But Faster.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Token Merging: Your ViT But Faster

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.221551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.221551Z digest=sha256:da3a3c4fcc456e0ce26a9775bb640d1a182f9f3cd328ac027727bfaf9ff4cb7c

Observation efa9f274-ff20-437b-8586-5f06a6265aff · outbound

This paper cites High-performance large-scale image recognition without normalization.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing High-performance large-scale image recognition without normalization

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.227627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.227627Z digest=sha256:fb23bc8ef9aa5423d79917197770feb75656f02dd4d624b579f11860cae65fcd

Observation 1d1a7761-c3ff-4c41-8174-c5885e3f8f46 · outbound

This paper cites Jump cell painting dataset: morphological im- pact of 136,000 chemical and genetic perturbations.BioRxiv, pages 2023–03, 2023.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Jump cell painting dataset: morphological im- pact of 136,000 chemical and genetic perturbations.BioRxiv, pages 2023–03, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.232305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.232305Z digest=sha256:5d34fe667197370f2c75ec3dcee2cd0e86209ae6b60125eb87155ac043500988

Observation 12b4d707-183f-45f4-8c74-18a95e8010e6 · outbound

This paper cites Towards a general-purpose foundation model for computational pathology.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Towards a general-purpose foundation model for computational pathology

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.236234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.236234Z digest=sha256:d680b3322887cedd6e534fbe2230f2cfc1a2f163ff6eb7db890ad47d27210385

Observation f2d18df4-8a15-4d62-8759-a14a1e8fac3f · outbound

This paper cites Openmmlab semantic seg- mentation toolbox and benchmark, 2020.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Openmmlab semantic seg- mentation toolbox and benchmark, 2020

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.240362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.240362Z digest=sha256:4c4f2005255748ff9657b6807bd470ae44e8a93ba4bcf4aa0530bdf161a56ffe

Observation 3d490cf1-c8f7-47eb-90a5-b3c89bde4360 · outbound

This paper cites Randaugment: Practical automated data augmen- tation with a reduced search space.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Randaugment: Practical automated data augmen- tation with a reduced search space

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.244496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.244496Z digest=sha256:536b55eb1b1b86047df841e02ca37e55cbc4ae8aec44b3edbd814a3a0cd15207

Observation d0efb87a-c2bb-4a91-998d-2266d105f783 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.248751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.248751Z digest=sha256:5bc553e8a42dd5bc42a32127de590bbf4b2fd5c46238206107cfc24a92cdec33

Observation 172ea684-c459-409f-b2d5-d3328d2b95e3 · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.253626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.253626Z digest=sha256:739b9cd87dfd875007a229d861460cbdeeea9bfa2780098897e631655bb1d3a3

Observation e7b47e4f-9d2c-4748-959a-22054eaad46b · outbound

This paper cites Flashattention: Fast and memory-efficient exact at- tention with io-awareness.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Flashattention: Fast and memory-efficient exact at- tention with io-awareness

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.258168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.258168Z digest=sha256:595ca359a49a8048a508ebaac628c11c1affcfdba458daf12776a63907619f2e

Observation 48660da0-bd72-47a3-84cb-2b8a395f45e5 · outbound

This paper cites Vpn++: Rethinking video-pose embeddings for understand- ing activities of daily living.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Vpn++: Rethinking video-pose embeddings for understand- ing activities of daily living

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.262466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.262466Z digest=sha256:0af5a3fb72d96b03ce7b2409a2b9ff893231eb09fd86b1d4e480bde5ccd7df5d

Observation 17fe9fe6-b38e-44e3-80d9-36814e67a872 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Imagenet: A large-scale hierarchical image database

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.266716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.266716Z digest=sha256:de63f02d185b0bd3443e1c71317ad1849f98484e963110f961d42de23063a4f6

Observation 70ee05b0-8083-43ca-9169-8ed344a96e05 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.270938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.270938Z digest=sha256:40080be3d2205b0525193b816bc70f45250af480b3cff1df4f234a89f3334be5

Observation 9c06baab-2d78-4e7f-bd16-95b9cdfbed3c · outbound

This paper cites Learned representation-guided diffusion models for large-image generation.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Learned representation-guided diffusion models for large-image generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.275280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.275280Z digest=sha256:82d3f6f667090ead1a66e98daab1a61363728e1cc06f9520f2ff2b79ac74db5b

Observation b063ed24-cb3b-4042-bbb5-adc8cbbbc8de · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.279120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.279120Z digest=sha256:994e38059d97e0d274b5749ae92a8e832817e10c0e06bd92c2faab7aa87a4fdd

Observation 92f2adb6-0584-4bc7-992e-f56b1a030687 · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Efficiently Modeling Long Sequences with Structured State Spaces

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.284095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.284095Z digest=sha256:12ec9574b254b72f0c8387da713a154d81650aa0a6ca0dcff93478d0ea116cdc

Observation 252e95e9-9f57-4325-942f-f948285a6cf3 · outbound

This paper cites Combining recurrent, convolutional, and continuous-time models with linear state space layers.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Combining recurrent, convolutional, and continuous-time models with linear state space layers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.288506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.288506Z digest=sha256:8182653feaf62bb8b555af6e9eaf9f1899b2a2c2ffef2b470b9f03649bcffa77

Observation e8d547f6-8544-434a-bc7f-4b19283ac95a · outbound

This paper cites MambaVision: A Hybrid Mamba-Transformer Vision Backbone.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing MambaVision: A Hybrid Mamba-Transformer Vision Backbone

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.292793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.292793Z digest=sha256:1dccbd8825a306628dd23bfcf3f47d4766ff62f444e66026a26b546d62d3b567

Observation cf9d9376-a084-4029-ba2b-c7041c8873c6 · outbound

This paper cites Mask r-cnn.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Mask r-cnn

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.320508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.320508Z digest=sha256:79967a9f227036ab04dee3b791da3449a7b97887f1af4d1d236138c46d52a5ee

Observation 2e6616fc-4a8e-4411-acb1-f2c9ead3feb3 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Masked autoencoders are scalable vision learners

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.755200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.357924Z digest=sha256:7df98a6afeefa822bdf5d33428dd935497b035e4e5e5177528ad18f1796d2f6d

Observation dc1375a1-1019-4157-b9d7-7d69177970f7 · outbound

This paper cites Token Dropping for Efficient BERT Pretraining.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Token Dropping for Efficient BERT Pretraining

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.481444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.481444Z digest=sha256:8aafb2febc7493a3928a2daf8633b04c869347f0cf235abf9cabd632ebf5bb6b

Observation e4b05f06-9583-4134-a255-f0ec0367cd15 · outbound

This paper cites Deep networks with stochastic depth.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Deep networks with stochastic depth

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.742341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.604768Z digest=sha256:d34293655d511f3b2cfea0e9825bc5cb6c8c3a38bc5a2296803b0c2bdc1374ab

Observation 8fbfdd15-08c5-46b7-b7c2-432a62d6e3a7 · outbound

This paper cites LocalMamba: Visual State Space Model with Windowed Selective Scan.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing LocalMamba: Visual State Space Model with Windowed Selective Scan

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.687930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.687930Z digest=sha256:f5b21f352b1efc338439526ecfc3792b335e6f53b78598629bc0d245e50c8f92

Observation 55be8f63-ba7c-474e-93a3-01840838a081 · outbound

This paper cites Attention-based deep multiple instance learning.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Attention-based deep multiple instance learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.727992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.770585Z digest=sha256:04f2b16ddddb64e12c73941168f4cfaa8fd87aaf3d0b341ebb24010f29362803

Observation 90af7577-2010-4dae-b329-7255e9435e9c · outbound

This paper cites Object- centric diffusion for efficient video editing.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Object- centric diffusion for efficient video editing

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.713373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.845460Z digest=sha256:5e57f443e527ce246c2785bc58d11babc4802c91b08675b45c20939468f88bdf

Observation 0ca161cb-2a31-4368-b9ce-fc702d91c444 · outbound

This paper cites Si-mil: Taming deep mil for self-interpretability in gigapixel histopathology.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Si-mil: Taming deep mil for self-interpretability in gigapixel histopathology

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.699665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.850118Z digest=sha256:5b6f91697986396d6dbdf133b015dfdceded75d409b846bd13f48733c012744e

Observation eb207fd1-522b-433d-a3c4-4da3bb37feb3 · outbound

This paper cites Vision transformers inference acceleration based on adaptive layer normalization.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Vision transformers inference acceleration based on adaptive layer normalization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.685429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.854197Z digest=sha256:417706b5f9371a0925203e7f0972e17039512af07ad9dd1f3ff3632fcacd25a3

Observation 52b4b24a-9282-4863-8fa7-31496ac14be5 · outbound

This paper cites ViTally Consistent: Scaling Biological Representation Learning for Cell Microscopy.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing ViTally Consistent: Scaling Biological Representation Learning for Cell Microscopy

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.858959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.858959Z digest=sha256:f2d2f6a34b80b9d4cfd5306f0c86b728da576ab897e54a17aedb280e539c8354

Observation bb4e11b9-62b6-41ff-923d-59c9faa70eaf · outbound

This paper cites Videomamba: State space model for efficient video understanding.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Videomamba: State space model for efficient video understanding

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.671858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.863752Z digest=sha256:0fff92e695e2ffc8cf89eff6eb2b6ddff327849016773bd0268fafd1777f628b

Observation 30e19c5a-3904-4601-9d9e-797afdcd7e50 · outbound

This paper cites Exploring plain vision transformer backbones for object de- tection.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Exploring plain vision transformer backbones for object de- tection

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.658326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.868392Z digest=sha256:00636b63f37eb568e9653c635ca7cab1bab7963d6f1cf8a2c88f6d57b0e0d77a

Observation e85b7bc2-e70f-4390-900b-f2e2245e06e0 · outbound

This paper cites Not All Patches are What You Need: Expediting Vision Transformers via Token Reorganizations.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Not All Patches are What You Need: Expediting Vision Transformers via Token Reorganizations

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.872619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.872619Z digest=sha256:24a29b0c10e8e4a0dab6b0e01a77111b0048beff461b6515475de25043e3c9f0

Observation ee700de7-5fba-4293-9e27-74809cc3c35b · outbound

This paper cites Jamba: A Hybrid Transformer-Mamba Language Model.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Jamba: A Hybrid Transformer-Mamba Language Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.876559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.876559Z digest=sha256:6e5557b4330a36a120c00430ddc63d980279b64cdcfd689b8e1a4858356c3019

Observation d19f7cfd-63ad-491c-8bf9-137d9672bfce · outbound

This paper cites Microsoft coco: Common objects in context.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Microsoft coco: Common objects in context

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.880926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.880926Z digest=sha256:afb8ce90ced9ba7550e38a433c67fbeb24f201b3e932c4bbc271cd01f8dbc346

Observation 04f8b42e-6bcc-496c-ab62-829b83a80fa2 · outbound

This paper cites MAP: Unleashing Hybrid Mamba-Transformer Vision Backbone's Potential with Masked Autoregressive Pretraining.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing MAP: Unleashing Hybrid Mamba-Transformer Vision Backbone's Potential with Masked Autoregressive Pretraining

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.885072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.885072Z digest=sha256:68902ad5ddd633bacc3b6cca0159529e00765a06770d6e1437a389aaefd8f5fc

Observation c6c6fce7-36c3-4bf9-a647-0536860090b1 · outbound

This paper cites Vmamba: Visual state space model, 2024.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Vmamba: Visual state space model, 2024

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.635675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.889120Z digest=sha256:e255241b447d66eacdd8c0d09963314bf6367982477e884102c5a17c409daf0a

Observation 17a66bfa-ac0e-4496-b97d-42150f36b6b1 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Swin transformer: Hierarchical vision transformer using shifted windows

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.621712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.893304Z digest=sha256:a82ed218db0276b4218cb2eb07ef1f6e7773453552ae17158ccdaf904d14f531

Observation d50d473c-ba14-4b20-9645-d6b1ab339f06 · outbound

This paper cites A convnet for the 2020s.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing A convnet for the 2020s

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.897538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.897538Z digest=sha256:d42f2319625dbe3ad9f3603041ff158db4c8e65359a7b32feeabb45dd4b5eb91

Observation e754c47b-9635-44b9-b063-72f22461b372 · outbound

This paper cites Decoupled Weight Decay Regularization.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Decoupled Weight Decay Regularization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.902445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.902445Z digest=sha256:15fd9818ccaeba4e617146b8645af86e7312067c16179b60270ba6842a44e993

Observation fcaadab1-e1bc-4f30-9712-e984aaa06989 · outbound

This paper cites SGDR: Stochastic Gradient Descent with Warm Restarts.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing SGDR: Stochastic Gradient Descent with Warm Restarts

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.906753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.906753Z digest=sha256:25247e4a3d654f77cfa67e37ca8e560560471d8617f85f7d4b6edfef9174c976

Observation ce124bbb-00ac-4f99-b4bf-7bc51b946de2 · outbound

This paper cites Vim4path: Self-supervised vi- sion mamba for histopathology images.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Vim4path: Self-supervised vi- sion mamba for histopathology images

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.600004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.912064Z digest=sha256:3c29ecf8b3b2bd70b216ed11b6d45466095bcb3e971f39b2e95960b0453ca90f

Observation 2c515de7-6d4b-4e9e-9a9f-aae9157eb3e0 · outbound

This paper cites An Image is Worth More Than 16x16 Patches: Exploring Transformers on Individual Pixels.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing An Image is Worth More Than 16x16 Patches: Exploring Transformers on Individual Pixels

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-09T18:28:48.927713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.916243Z digest=sha256:57eb99c78aba3528df7b240e44f41ba424b2d1c104b61e9c32198bfd100eb507

Observation 7674ff40-4336-4b96-b777-0747e089fdad · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing DINOv2: Learning Robust Visual Features without Supervision

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.920635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.920635Z digest=sha256:1ecd31d8cf56ec0dc3808d9fe14acea147e9d38da3974d4e9c9cff92c2c7fdff

Observation 7d0a776c-ba42-44fe-8301-06be5fcc2179 · outbound

This paper cites EfficientVMamba: Atrous Selective Scan for Light Weight Visual Mamba.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing EfficientVMamba: Atrous Selective Scan for Light Weight Visual Mamba

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:47.924619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:47.924619Z digest=sha256:01beae47a256f13e3298a0a60feb5e9c803109a3b5ff2890e648de6242ab40fe

Observation 60f0f329-f820-4457-a0e5-1a6e6f75fbde · outbound

This paper cites Enhancing Feature Diversity Boosts Channel-Adaptive Vision Transformers.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Enhancing Feature Diversity Boosts Channel-Adaptive Vision Transformers

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-09T18:28:48.882474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:47.967988Z digest=sha256:322f2dee9253f25534d8216c6e265615b736d1587fb52dad06779c7a3c036223

Observation a70e1692-5ba0-477e-b8d7-3bc54dbbdb4c · outbound

This paper cites Per- ceptual grouping in contrastive vision-language models.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Per- ceptual grouping in contrastive vision-language models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.586010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.088516Z digest=sha256:d1cf4824790cdd610ecd0a34deb7d7e91e1eb47d1084eef5afe6228155e25701

Observation e7a8dcf1-097f-4d6f-8684-b5ec1cb6717c · outbound

This paper cites Dynamicvit: Efficient vision transformers with dynamic token sparsification.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Dynamicvit: Efficient vision transformers with dynamic token sparsification

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.163481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.163481Z digest=sha256:e29d58eb07549f2b41b99c3e856f60e177f7af29e0a188c377dd1a0f7a2c79b3

Observation 9c461ef3-cc6e-4b92-8c8a-3db74e1e26ab · outbound

This paper cites Autoregressive Pretraining with Mamba in Vision.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Autoregressive Pretraining with Mamba in Vision

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.358953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.358953Z digest=sha256:fb681c5780580b877c2c5ac6bd02aa3aae69979af0ec9aaffd34d8950c438331

Observation 1912697c-1ac6-4bc8-98ed-0c468cee33b2 · outbound

This paper cites Learning to Merge Tokens in Vision Transformers.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Learning to Merge Tokens in Vision Transformers

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.402338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.402338Z digest=sha256:b00b924beb3a265e3ddea05d418c23ff6ac9cc7ff1f8f3d69f1bba4e215b3a43

Observation c39a0be2-f050-43b9-b048-64ff5f04c902 · outbound

This paper cites TokenLearner: What Can 8 Learned Tokens Do for Images and Videos?.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing TokenLearner: What Can 8 Learned Tokens Do for Images and Videos?

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.407448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.407448Z digest=sha256:d0a2d9c5399e90ebcca4af161692a0e00dfc17d075a758fd4d080b72e449d75d

Observation 125869ff-5f78-417d-9af6-bbc9383231a2 · outbound

This paper cites GroupMamba: Efficient Group-Based Visual State Space Model.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing GroupMamba: Efficient Group-Based Visual State Space Model

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.412132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.412132Z digest=sha256:876653b121b3223743a451309fccde79092625c8ae0c7272ac4e8a033b823480

Observation 6f7e0aaa-3890-4453-82e8-089756588429 · outbound

This paper cites Famba-V: Fast Vision Mamba with Cross-Layer Token Fusion.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Famba-V: Fast Vision Mamba with Cross-Layer Token Fusion

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.416470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.416470Z digest=sha256:03f9f671cbed1d59299a57d0bf36b9719ab6e47291a9347b1fea241e94eed1bd

Observation ea069a7f-a15e-43ce-abfb-7f3d17f8f959 · outbound

This paper cites Simplified State Space Layers for Sequence Modeling.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Simplified State Space Layers for Sequence Modeling

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.421026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.421026Z digest=sha256:af4382e527cabed2bf38a344c8986534e880246f4d4df676ef092abec4e0d352

Observation 20c497cc-ecff-45ad-8e6c-d991d351b855 · outbound

This paper cites Training data-efficient image transformers & distillation through at- tention.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Training data-efficient image transformers & distillation through at- tention

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.565122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.425538Z digest=sha256:1629cba7b3f563d4b92b481d3a841892c9788e6985ee6763117700530f0fce0c

Observation f0c6ea6a-bdb4-423d-a2db-e4c906d18f08 · outbound

This paper cites Training data-efficient image transformers & distillation through at- tention.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Training data-efficient image transformers & distillation through at- tention

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.552685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.429965Z digest=sha256:4b7bb1dc024a92aed2e129a5892cadaaa51e4061d1acb77e6d4450b892a8c235

Observation 991bd086-c06d-4792-8654-b728adc15798 · outbound

This paper cites Attention is all you need.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Attention is all you need

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.434842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.434842Z digest=sha256:5bd7a8c0f61ab2e054d01c1d6f0e851fa0e362917f9813650d685f70c5050a04

Observation 0ad339ca-3122-4447-8af3-47c93253a370 · outbound

This paper cites Mamba-R: Vision Mamba ALSO Needs Registers.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Mamba-R: Vision Mamba ALSO Needs Registers

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.438869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.438869Z digest=sha256:061af0d5c3a61e6550dd831c0385814590187399954522a3256b6e686a49713a

Observation 2875a687-c8b4-475a-a1d0-892dae9ef6b2 · outbound

This paper cites Unified perceptual parsing for scene understand- ing.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Unified perceptual parsing for scene understand- ing

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.489438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.489438Z digest=sha256:83f498dc11422d80b2c457788d846f7840d4ba4dc8c986593ef3e57866e5a384

Observation 5e042235-bfed-423c-8b9e-1ff59bf6d8c8 · outbound

This paper cites A whole-slide foundation model for digital pathology from real-world data.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing A whole-slide foundation model for digital pathology from real-world data

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.523234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.537921Z digest=sha256:80a57db328f5ac6edb5c40127928633dbc58889ad6a291db14e1be3673418993

Observation 097394c1-05c2-46b5-b550-ebb2bcf64845 · outbound

This paper cites PlainMamba: Improving Non-Hierarchical Mamba in Visual Recognition.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing PlainMamba: Improving Non-Hierarchical Mamba in Visual Recognition

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.541812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.541812Z digest=sha256:94fb3323d559eeb7b65e770f11c8a990f41441f42a24743407ae592b40c71ee0

Observation d3a3a8bc-6e17-443d-9c8a-d14953886b67 · outbound

This paper cites Cutmix: Regu- larization strategy to train strong classifiers with localizable features.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Cutmix: Regu- larization strategy to train strong classifiers with localizable features

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.510913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.545751Z digest=sha256:b1a2d4a7b1f6972a4aa6c23092aea452d2bb2cf88b3dd1269843ff9174c7ba3d

Observation db50be05-09c9-4811-abbe-72743fcfd13d · outbound

This paper cites Exploring Token Pruning in Vision State Space Models.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Exploring Token Pruning in Vision State Space Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.549618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.549618Z digest=sha256:d8263d1df0094eee0a046de76f7751a3a5c3481ef8b555382dc4197fa7bbc6ce

Observation 2a9c3ef6-1368-4ee3-a3e7-26beef0231d8 · outbound

This paper cites Rethinking Token Reduction for State Space Models.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Rethinking Token Reduction for State Space Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.553647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.553647Z digest=sha256:dd5100f6c3b77a49ea79578e02958b5699328ef820bf809b8585f2965a0198e1

Observation c4ab4d19-b994-4c85-914d-a0b7f7315ca9 · outbound

This paper cites mixup: Beyond Empirical Risk Minimization.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing mixup: Beyond Empirical Risk Minimization

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.557693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.557693Z digest=sha256:d384453f993bcf1adb947985e5ca23cce58be9a6de9ffa208e6ec96f08a87ee1

Observation 96bd558b-6d88-40ff-8ea6-66a3befa02c5 · outbound

This paper cites Scene parsing through ade20k dataset.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Scene parsing through ade20k dataset

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.562442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.562442Z digest=sha256:49ae9950d9e869c301f20e367d74d7aa5ae3f3ac83c89e435f12c27158a5b7f2

Observation 884c7d76-4d70-46bb-a65f-9ada91e8cea6 · outbound

This paper cites Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-09T18:28:48.567077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:28:48.567077Z digest=sha256:b747879ad07664ad0f17edc51cd83f32a4d5dddd7315d423472fc150b212e5cf

Observation 72815423-f091-4727-9707-331e6092ac0c · outbound

This paper cites We closely followed the pre- training (Table 9), fine-tuning (Table 10), and linear- probing (Table 11) settings from the Masked Autoen- coders [24] codebase.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing We closely followed the pre- training (Table 9), fine-tuning (Table 10), and linear- probing (Table 11) settings from the Masked Autoen- coders [24] codebase

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.488891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.571996Z digest=sha256:5f17f3b0592a18b7aaebcc88b60737cc33abb8df04f5706d8cb7de7db0036e42

Observation c1b391d4-647e-48ef-952c-b126e0fae6c3 · outbound

This paper cites 2) We applied a scaling factor of 1 − mask ratio (75% masking by default) during fine-tuning and linear probing when pooling tokens before the SSM block.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing 2) We applied a scaling factor of 1 − mask ratio (75% masking by default) during fine-tuning and linear probing when pooling tokens before the SSM block

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.476026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.576472Z digest=sha256:b41dfb1715d7cc4d658a57e908696eceecd5cc3d3dbac1e359f8b79a3c032394

Observation 8f2b76bc-12f2-4a91-9726-bed24d7e4d83 · outbound

This paper cites mean pool in Fast- MaskVim.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing mean pool in Fast- MaskVim

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.462686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.580967Z digest=sha256:8f13ab76860c0dd0a685498097681d8fe8cf74c649f06622ebb8ab7d7d403b57

Observation 3677531d-e374-4dd1-a2af-83dfa366a714 · outbound

This paper cites In Table 13, we compare the performance of fine-tuning pre-trained FastMaskVim using alternate layer learning rate decay instead of per-layer decay as in the MAE codebase.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing In Table 13, we compare the performance of fine-tuning pre-trained FastMaskVim using alternate layer learning rate decay instead of per-layer decay as in the MAE codebase

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.449688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.585616Z digest=sha256:fe26c619d580c1d5adabe20da0288a4901a769deed95060ae2d4ba3ef3fdd916

Observation cad2c0ed-963b-4fdd-a011-b288a5b87dea · outbound

This paper cites Apply- ing the scaling factor results in an improvement of 0.3% compared to the default mean pooling in fine-tuning without multiplying by the scaling factor.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Apply- ing the scaling factor results in an improvement of 0.3% compared to the default mean pooling in fine-tuning without multiplying by the scaling factor

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.436446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.590168Z digest=sha256:9c450bd059b2823567b1cbd5b3c1c6c530289e80d5387259b09ea738f0d45920

Observation d008e347-c57b-4761-90cf-014178075cef · outbound

This paper cites In Table 15, we compare the linear probing performance of Fast- MaskVim with and without the scaling factor (0.25).

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing In Table 15, we compare the linear probing performance of Fast- MaskVim with and without the scaling factor (0.25)

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.423402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.594871Z digest=sha256:5e645cd02e58667af721aa8eabc59515a782596e30e102e9d5447a2529169cf7

Observation f5ae686b-e2e5-4250-86f0-02299296d374 · outbound

This paper cites an unresolved cited work.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Unresolved cited work

Reference 77

Resolution
parse uncertain
raw_fallback, observed 2026-08-09T18:28:49.410087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.599168Z digest=sha256:19c94b7b9b0efb99fb813c7b9b3022d6cb06e915327897d6c5b4863aaa6429e5

Observation 0066557f-f833-4bb4-aa90-bd5f325eafc7 · outbound

This paper cites In Table 16, we compare the performance of FastVim-S with a class token versus without a class token (default).

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing In Table 16, we compare the performance of FastVim-S with a class token versus without a class token (default)

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.396909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.603670Z digest=sha256:84bf0bf32c1d5080828a16eb9261986d4d4206d11207efff74416d6e78dd4337

Observation 39f5c2a8-6f78-4cfd-8eca-46fb7e5670d3 · outbound

This paper cites In Table 17, we empir- ically demonstrate the performance of FastVim trained with different combinations of input normalization and post-SSM normalization.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing In Table 17, we empir- ically demonstrate the performance of FastVim trained with different combinations of input normalization and post-SSM normalization

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.383837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.608160Z digest=sha256:0a8fda9380efa7827745042555405437a04f5bc5dca38341f9bf393444e8bf84

Observation ea6d773f-6309-4c77-a392-188473668fe5 · outbound

This paper cites In Table 18, we explore whether in Fig.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing In Table 18, we explore whether in Fig

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.371044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.612618Z digest=sha256:246e3649d04b2f37778ceee98f8b96313be3d32a9ee07ce7920309fa417de141

Observation f578d65f-0c8d-4930-a71f-173773202692 · outbound

This paper cites We followed the implementation details primarily from ChannelViT [3].

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing We followed the implementation details primarily from ChannelViT [3]

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.358146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.616956Z digest=sha256:535c5e3800ded9a5c5f579ba2cbdd7843663f011e5a730fe61f8cb202f32a357

Observation 3f58022d-ad0b-4c3c-83e0-9e36a20bc279 · outbound

This paper cites Channel- First with and without sorted HCS.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Channel- First with and without sorted HCS

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.346395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.621528Z digest=sha256:fc5936922104397125a6fdc0e7df92eba0ae132f7dbbd065c8d8682514e2010a

Observation c8d43edc-e6aa-41b9-9d7a-bf12f540d892 · outbound

This paper cites We then explore the effect of different pooling methods, such as max pooling [49] and attention pooling [28], as detailed in Table 20 on the JUMP-CP dataset.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing We then explore the effect of different pooling methods, such as max pooling [49] and attention pooling [28], as detailed in Table 20 on the JUMP-CP dataset

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.334222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.625803Z digest=sha256:adb0909dbc69cd1796ba9cbd4a3defc9826954e956aa0468400e953d024dbeab

Observation dad8146f-16ab-4f56-bb46-027a6f2d510c · outbound

This paper cites Now, we preliminarily explore pooling along two dimen- Table 20.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Now, we preliminarily explore pooling along two dimen- Table 20

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.322383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.630185Z digest=sha256:8cf1061e08e4c5f8e5326d3da069da8fe4c04eabc1a1f774d6f62ca1023961be

Observation bee2db5a-7fec-4ea4-a149-9b62938f443a · outbound

This paper cites an unresolved cited work.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:28:49.309789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.634655Z digest=sha256:9ec28aa25b7a6ac9da686ac2bf0f31498597759346931898d680217be4502e69

Observation 8d1121a7-1742-4580-adb0-a6c5775f1b32 · outbound

This paper cites an unresolved cited work.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:28:49.296577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.639100Z digest=sha256:79ac235eb3fdf2393374df896b3e43ff8065bb4bfab9cd959b316ea3360fda20

Observation 90da51ca-3337-4716-b396-7567125093c1 · outbound

This paper cites an unresolved cited work.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Unresolved cited work

Reference 87

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:28:49.283163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.643509Z digest=sha256:cde05e71ae4989a1b0423b45e8374fc0c37a1fe42f729f6d8d31f4157873862f

Observation cc8fc2a1-58e6-4821-8667-8a2fd380a361 · outbound

This paper cites an unresolved cited work.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Unresolved cited work

Reference 88

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:28:49.269793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.647976Z digest=sha256:42848f1685632f692ef8b361e5ee183f7b79c6d72dc03071b12e49a6cc35eca8

Observation b675a518-fb63-429c-a27c-fe1de2209a42 · outbound

This paper cites an unresolved cited work.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-09T18:28:49.256152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.652421Z digest=sha256:e042c7744f7e3f6f8cdef9971ac775879eaf116e5a25e61b1aafd91c1141ac5e

Observation 68b26e95-cc81-4cbd-b083-05aa12fc2641 · outbound

This paper cites In Table 22, we demonstrate the throughput improvement in FastChannelVim compared to ChannelVim.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing In Table 22, we demonstrate the throughput improvement in FastChannelVim compared to ChannelVim

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.242420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.656847Z digest=sha256:f9ab86b2f6778002f017d20bddafae2c35358265b8a9f094a46034ffa47f7c1a

Observation 505fef0c-ccf5-4083-b4f1-83bf717a8875 · outbound

This paper cites Here, we calcu- late the processing time for Forward SSM + Backward SSM in only one block (see Fig.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Here, we calcu- late the processing time for Forward SSM + Backward SSM in only one block (see Fig

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.227388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.661339Z digest=sha256:34a7b2c058b32a07ddb0146698eb79c93c54f1d611fbfa7ebc658b468819c7cc

Observation b83603b0-749f-4159-8994-013e7e0e667d · outbound

This paper cites We employed the AdamW optimizer with a weight decay of 0.01.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing We employed the AdamW optimizer with a weight decay of 0.01

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.213392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.665879Z digest=sha256:7c8f0c5db4e23fec9d0dc3e48fc5f223de6842c79131b1802d9734b78a1d4ff1

Observation 23b235ad-6ca2-48f2-bd7d-6068e50fbb1c · outbound

This paper cites We employed the AdamW optimizer with a weight decay of 0.05, with a total batch size of 64.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing We employed the AdamW optimizer with a weight decay of 0.05, with a total batch size of 64

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.199418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.670445Z digest=sha256:7ad328f2526c3457dcbf8e0507d1fc65d5b24dedd6ff2a3d37edac2ec4701dca

Observation 9e16d200-e858-4ce7-9c97-469ad51776c9 · outbound

This paper cites 2), we apply mean pooling to the tokens before performing the SSM scan.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing 2), we apply mean pooling to the tokens before performing the SSM scan

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.185072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.674860Z digest=sha256:ae7510fc9dde61a9390828dcf066936dc313bbf01c7378abf28bb7351813beca

Observation 63e740a5-14e1-4335-82eb-dd8fe6848f19 · outbound

This paper cites Model configurations for FastVim Model Layers Embedding dim.

Fast Vision Mamba: Pooling Spatial Dimensions for Accelerated Processing Model configurations for FastVim Model Layers Embedding dim

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T18:28:49.170789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T18:28:48.679484Z digest=sha256:c288ff59d7d296bf69075103e74bb3cfd2949dd878d12feb2254a662df71a3d7

Pith citing papers

No inbound Pith citation observations are available.