Pith. sign in

Paper Citation Record · LEDGER

Modality Agnostic Efficient Long Range Encoder

As of 19 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2507.19409.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19409 v1

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:59:25.515340Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

70 of 70 outbound references displayed

  • verified exact2
  • verified fuzzy49
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40c324da-9dd3-4fb6-a6e1-339eec8d53bd · outbound

This paper cites GQA: Training gener- alized multi-query transformer models from multi-head checkpoints.

Modality Agnostic Efficient Long Range Encoder GQA: Training gener- alized multi-query transformer models from multi-head checkpoints

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.504863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.215137Z digest=sha256:0e6c33d4e557937811feeed9a1fa27bb78c36efc520a9cabeb34f5e1e0fd8ee7

Observation d637edb3-9b63-4b0e-b8ee-fb77545f0f22 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Modality Agnostic Efficient Long Range Encoder Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.490859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.220029Z digest=sha256:edaff4d7ccef88dc5b6d11b690926c3586a799e7a9d73f7138373f19b9d219f8

Observation 2ebd3bc5-ce57-4ab1-a20c-71e9b69b887a · outbound

This paper cites Longformer: The Long-Document Transformer.

Modality Agnostic Efficient Long Range Encoder Longformer: The Long-Document Transformer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.224593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.224593Z digest=sha256:4fc747ce3941136ccc651181817a5e1603056684bbff1dad5b6d12034bcd1c8b

Observation 98a3b113-6bca-41bf-b8dd-b33dc0e8ff24 · outbound

This paper cites Token Merging: Your ViT But Faster.

Modality Agnostic Efficient Long Range Encoder Token Merging: Your ViT But Faster

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.229873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.229873Z digest=sha256:16c9651df0ce3b20ad4cd69a822604e7ebb1fd10e36ef507d602573769dfc977

Observation 1383d3df-aadf-4b0f-8eb1-7afe67b8464e · outbound

This paper cites Scaling Transformer to 1M tokens and beyond with RMT.

Modality Agnostic Efficient Long Range Encoder Scaling Transformer to 1M tokens and beyond with RMT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.234525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.234525Z digest=sha256:7d3fee9d94bf2215f01f46958f74a538e76a4de34a10108781e371d83898d5b4

Observation e2ea8a0c-c9a0-4e14-ad68-0f5fd9b360b9 · outbound

This paper cites Vggsound: A large-scale audio-visual dataset.

Modality Agnostic Efficient Long Range Encoder Vggsound: A large-scale audio-visual dataset

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.476645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.239060Z digest=sha256:b558ac19a4bad356fba93053f6e042230740800676ec2501ad8c3397414c2f83

Observation ce372dc8-eff8-473a-81cb-b13e67c1066e · outbound

This paper cites Classification of long sequential data using circular dilated convolutional neural networks.

Modality Agnostic Efficient Long Range Encoder Classification of long sequential data using circular dilated convolutional neural networks

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.462757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.244652Z digest=sha256:599309bd4093a536d20f8d39edc3ce4b2c36aacec2ef1c0382fa24d344708ab6

Observation 3a3eded5-c146-4be6-b11e-5a5788db84ae · outbound

This paper cites Generating Long Sequences with Sparse Transformers.

Modality Agnostic Efficient Long Range Encoder Generating Long Sequences with Sparse Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.248806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.248806Z digest=sha256:1bbddc334fb863017ffdc74644047ae347496aa74faf624afa7980bd51ea36fa

Observation 74a40dab-c92e-4313-884e-350194cfaae4 · outbound

This paper cites Rethinking attention with performers.

Modality Agnostic Efficient Long Range Encoder Rethinking attention with performers

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.449170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.253515Z digest=sha256:7b5d4e80891ae8c53576f44764389b9410f3a0fdedcc1d6ff2caf9b6e86f74e9

Observation 6fc60df1-913f-441c-ac32-a11e3238eccb · outbound

This paper cites Ran- daugment: Practical automated data augmentation with a reduced search space.

Modality Agnostic Efficient Long Range Encoder Ran- daugment: Practical automated data augmentation with a reduced search space

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.435665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.258004Z digest=sha256:a6c4627617f206d8904674267cc5ec7d049fae405de88949289cdecfd46f6c6f

Observation 472b63c3-65ce-451d-b6bf-37fd00998efa · outbound

This paper cites an unresolved cited work.

Modality Agnostic Efficient Long Range Encoder Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:59:26.422263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.263129Z digest=sha256:be4997a82f6cdefd13cab1099618f9f80fd5572389fd65395f767de95c2d0ebe

Observation 17a97211-7920-42c0-b8bd-8b9d13558edf · outbound

This paper cites FlashAttention-2: Faster attention with better parallelism and work partitioning.

Modality Agnostic Efficient Long Range Encoder FlashAttention-2: Faster attention with better parallelism and work partitioning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.408224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.267331Z digest=sha256:778b0d36550d15c7d657033ebb5c25270f853b060e84df095997ac75737cdc86

Observation 7f27df0e-d055-413b-b62c-5d5f15659eab · outbound

This paper cites Fu, Stefano Ermon, Atri Rudra, and Christopher R´e.

Modality Agnostic Efficient Long Range Encoder Fu, Stefano Ermon, Atri Rudra, and Christopher R´e

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.394228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.271441Z digest=sha256:e5ae1ec96aa701d1cab699072706a4f2d905ebe20ded9862d3f215412beaffad

Observation 1df443de-fc46-4f88-82e7-b5a1c91349ce · outbound

This paper cites The ucr time series classifica- tion archive, 2019.

Modality Agnostic Efficient Long Range Encoder The ucr time series classifica- tion archive, 2019

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.380830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.276528Z digest=sha256:91e9b66654ee4b3ea9cd51df31acc75a5dbcef9920f10ffe15c067a8ae5bc961

Observation 70b88857-625a-4ccc-a378-f427f406048f · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

Modality Agnostic Efficient Long Range Encoder DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.280766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.280766Z digest=sha256:d29fc65fef1ebe2011c769e4302b90699e73d6147786a25dd50b435c9f599116

Observation 9812f137-9952-4b9d-b4cb-9ff28c92b788 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Modality Agnostic Efficient Long Range Encoder Imagenet: A large-scale hierarchical image database

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.367324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.285702Z digest=sha256:e677fbfa9cee2e3deb1bf46d1e0a9dfbb721b4d5d37f03ec9dd129a2b020ea96

Observation a36df706-28d2-411a-9e10-fe7894b98148 · outbound

This paper cites fvcore: A collection of core libraries for com- puter vision.

Modality Agnostic Efficient Long Range Encoder fvcore: A collection of core libraries for com- puter vision

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.353386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.290213Z digest=sha256:5814e42fbc4a3347e1fed26cacdbd64eef1ec08e58b5249244a494979cf66f84

Observation e9652e75-ee4d-4152-9a9b-e8763c39efae · outbound

This paper cites Multiscale vision transformers.

Modality Agnostic Efficient Long Range Encoder Multiscale vision transformers

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.339847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.294296Z digest=sha256:544a0f904915196909f402051bbb17011fbf92060e682badd3aaaf9232428b70

Observation 2b5b81ca-432e-4d31-b44e-3f67d8965325 · outbound

This paper cites Hungry Hungry Hippos: Towards Language Modeling with State Space Models.

Modality Agnostic Efficient Long Range Encoder Hungry Hungry Hippos: Towards Language Modeling with State Space Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.298833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.298833Z digest=sha256:243013774edd6e408d11fd3230d095bf6aacb1a5be8b50cd3cd53b64762ee31e

Observation 76069de1-ba93-41d0-8855-dc3cca94f696 · outbound

This paper cites Dissecting Recall of Factual Associations in Auto-Regressive Language Models.

Modality Agnostic Efficient Long Range Encoder Dissecting Recall of Factual Associations in Auto-Regressive Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.303243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.303243Z digest=sha256:3262f2470619109dfc9a0b491ecf11886cae01d24cbc9da10e05d19d10cd4319

Observation 6a1df863-2b17-47de-9a55-b342cc01a06b · outbound

This paper cites Liu, David Har- wath, Leonid Karlinsky, Hilde Kuehne, and James R.

Modality Agnostic Efficient Long Range Encoder Liu, David Har- wath, Leonid Karlinsky, Hilde Kuehne, and James R

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.326851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.307614Z digest=sha256:4d327589b7983ff093ec174a4fe53de3abeb3890a18e363d0c9ab599182aa5da

Observation 832d6e1b-4a4b-44ba-9953-49dd3ec803f7 · outbound

This paper cites Choudhury, Saurabh M.

Modality Agnostic Efficient Long Range Encoder Choudhury, Saurabh M

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.313722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.311632Z digest=sha256:15e8cf07f85ccb694693047b75e6e2e3473c1cab1238c9168db7fb8b0717d6c8

Observation eca2bdb5-abcb-42e0-b621-7cea2e99dad0 · outbound

This paper cites Levit: a vision transformer in convnet’s clothing for faster inference.

Modality Agnostic Efficient Long Range Encoder Levit: a vision transformer in convnet’s clothing for faster inference

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.300518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.315688Z digest=sha256:39954e494a030cd2bb58a5f61c0e105066c16cbea96e71f6cfbd480ea5344a13

Observation 3598c43a-b2cc-4556-a459-144cd849f4f0 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Modality Agnostic Efficient Long Range Encoder Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.319687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.319687Z digest=sha256:4dcd08ca179313880171afc8ea8135e5df5bbb070115061df0755af01e83a000

Observation f192dcbc-ca27-40f6-8fa0-9b4667330089 · outbound

This paper cites Longt5: Efficient text-to-text transformer for long sequences.

Modality Agnostic Efficient Long Range Encoder Longt5: Efficient text-to-text transformer for long sequences

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.287274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.323970Z digest=sha256:78789818aa6b4b4e5d43bd6477440a4a45052a0e8bf5207b59c37312185c47da

Observation b4763ef0-e95b-4a4e-b44d-067276310a14 · outbound

This paper cites Flatten transformer: Vision transformer using focused linear atten- tion.

Modality Agnostic Efficient Long Range Encoder Flatten transformer: Vision transformer using focused linear atten- tion

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.273407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.327908Z digest=sha256:e3520afeb9bfc1556f264cfa7bbe038e01912d3a7c75e43a5fad68eee12fa983

Observation 6709e51d-7c88-4e42-be09-80c2fc9eafba · outbound

This paper cites Masked autoencoders are scalable vision learn- ers.

Modality Agnostic Efficient Long Range Encoder Masked autoencoders are scalable vision learn- ers

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.260553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.332340Z digest=sha256:303c6a8d1cd630a8662d5672ba43d40e1045b722d65ac62b97fe61bb5fbb4ac5

Observation f215666b-23e2-43db-b069-4e43ac9dcf0f · outbound

This paper cites Rethinking spatial dimensions of vision transformers.

Modality Agnostic Efficient Long Range Encoder Rethinking spatial dimensions of vision transformers

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.247518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.336473Z digest=sha256:32e34f97fcaf3b5f1749680761b8e6f04f09a361c214d829ac4bd242e83f9e97

Observation a773974a-63ca-4161-be25-957645d5e75d · outbound

This paper cites an unresolved cited work.

Modality Agnostic Efficient Long Range Encoder Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:59:26.233896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.340570Z digest=sha256:c37a19c794bfe123142da74b8ef8b07322230e8a1c4ad2b7a919616ba78072fd

Observation 38e4425a-a289-466f-888a-9a71f47c3090 · outbound

This paper cites Deep networks with stochastic depth.

Modality Agnostic Efficient Long Range Encoder Deep networks with stochastic depth

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.219846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.344819Z digest=sha256:6c57a5c868b59096d8747aab09d51aaa4c9c90fbdf55e497a89648f91cff7751

Observation b35ca5c3-3979-435c-8e28-49b8e3f78cfc · outbound

This paper cites Transformers: State-of-the-art machine learning for pytorch, tensorflow, and jax.

Modality Agnostic Efficient Long Range Encoder Transformers: State-of-the-art machine learning for pytorch, tensorflow, and jax

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.205532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.348878Z digest=sha256:7e43a299fe10f8fc2fdebd2b2cf70df75e210120ad9f3f65cff54a79a9fbee0f

Observation a8d10155-e7e8-4edf-b305-c51f4b269ddc · outbound

This paper cites An empirical survey of data augmentation for time series classification with neural networks.

Modality Agnostic Efficient Long Range Encoder An empirical survey of data augmentation for time series classification with neural networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.191554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.352830Z digest=sha256:319be2f011a047353c222db6f99662a2ad2c153feab911b75ad4a3a37fd0e15b

Observation 2e913e1d-44f5-4c50-98d0-0ec554898e92 · outbound

This paper cites Katharopoulos, A.

Modality Agnostic Efficient Long Range Encoder Katharopoulos, A

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.177627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.357165Z digest=sha256:544ce21cd9ad2ff7a327f0934c544a0fa25bbd299246b53a048cde27c1f1b955

Observation 09c1b765-c89c-42dc-8aba-31811d20d31c · outbound

This paper cites Reformer: The efficient transformer.

Modality Agnostic Efficient Long Range Encoder Reformer: The efficient transformer

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.164371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.361347Z digest=sha256:54dac053a68e4bca7f5910018d9b9570cafaa5753842b51e14e4bc93ef1f0936

Observation f2a17ed6-74b4-4c32-97ab-2dc84f820da3 · outbound

This paper cites Sequence parallelism: Long sequence training from sys- tem perspective.

Modality Agnostic Efficient Long Range Encoder Sequence parallelism: Long sequence training from sys- tem perspective

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.151210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.365969Z digest=sha256:0ab7b6d5be7a1d97c0ba4784dc7f59554e02d7a161d3c5ded76fe342c53efbef

Observation 8b3e01ac-e52f-415a-a6f3-2454d57caa58 · outbound

This paper cites Mvitv2: Im- proved multiscale vision transformers for classification and detec- tion.

Modality Agnostic Efficient Long Range Encoder Mvitv2: Im- proved multiscale vision transformers for classification and detec- tion

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.137162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.370122Z digest=sha256:4815580501f23a7a3f5a6bd9c6ab1baf920c3fb8d1f0c51e081355ae9af903b6

Observation bb17e37d-ac4b-43f8-8176-fa7524e45327 · outbound

This paper cites Efficientformer: Vi- sion transformers at mobilenet speed.

Modality Agnostic Efficient Long Range Encoder Efficientformer: Vi- sion transformers at mobilenet speed

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.123873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.374058Z digest=sha256:6b48242005a3d4fdec689e4f20474e92c4a8b49b1f6d71925dea191a73e0a17e

Observation f6ee13b4-36fc-43fa-b932-88a4db6f26c0 · outbound

This paper cites Ring Attention with Blockwise Transformers for Near-Infinite Context.

Modality Agnostic Efficient Long Range Encoder Ring Attention with Blockwise Transformers for Near-Infinite Context

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.378070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.378070Z digest=sha256:34146ed30da71ddec72a705df822746c8c99c49a45da17ea9458effa6a7f0925

Observation 7b8adc2b-ecdf-4c52-8683-59e0fadc48ae · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Modality Agnostic Efficient Long Range Encoder RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.382444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.382444Z digest=sha256:f9fa6a28cce22d397c7f8c8f6e67e835fa7b12aa6918fc9ca65bfe86edc00ecf

Observation 918ca2aa-219e-424f-8732-02491fe6a18d · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Modality Agnostic Efficient Long Range Encoder Swin transformer: Hierarchical vision transformer using shifted windows

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.111004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.386674Z digest=sha256:48b6f6c76381d342b418397212df612eb0c9c8af293d5337332038ef9c04656c

Observation 23d30d51-e82c-401f-9cf4-4f67d03a0e5d · outbound

This paper cites Sgdr: Stochastic gradient descent with warm restarts.

Modality Agnostic Efficient Long Range Encoder Sgdr: Stochastic gradient descent with warm restarts

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.097838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.390792Z digest=sha256:d5da7b1804630cc9aca8acdd993f82d397da87aab29ce11e59b21c97cc6223e7

Observation afaefcc2-3ff1-45a9-85a1-e3a223656ba1 · outbound

This paper cites Decoupled weight decay regular- ization.

Modality Agnostic Efficient Long Range Encoder Decoupled weight decay regular- ization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.083666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.394907Z digest=sha256:cad1e9d4c917aa818d9d3b59a7e1a611e02655c809588840fd18d30bd847daa4

Observation d6781f67-41c5-4bf6-b199-d4e5fac58704 · outbound

This paper cites Token pooling in vi- sion transformers for image classification.

Modality Agnostic Efficient Long Range Encoder Token pooling in vi- sion transformers for image classification

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.069938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.399303Z digest=sha256:8b48d5bc275da67d79f446985a527eff5f387afce7905bce2a88b2c9b96b9385

Observation ded4bb0b-b8c5-43c3-ad37-3e8c7564c964 · outbound

This paper cites Lo- cating and editing factual associations in gpt.

Modality Agnostic Efficient Long Range Encoder Lo- cating and editing factual associations in gpt

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.055908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.403328Z digest=sha256:6f77ebfe43b65db556be052974c10eb11835b3c795c326459ea42cfdc976af9a

Observation de3fd62f-775f-42dc-bcbe-5e061e48957c · outbound

This paper cites Scalable vision transformers with hierarchical pooling.

Modality Agnostic Efficient Long Range Encoder Scalable vision transformers with hierarchical pooling

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.042794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.407321Z digest=sha256:ee4571dc81ab507d311d19c3bdc69fd8b2893dd3406e3f2419e7fb673302faa2

Observation e5223b9e-3647-40da-982b-cb780d597be6 · outbound

This paper cites Investigating efficiently ex- tending transformers for long input summarization.

Modality Agnostic Efficient Long Range Encoder Investigating efficiently ex- tending transformers for long input summarization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.029374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.411439Z digest=sha256:99fdab7b15191d3f8ae115c35652f505484f8d7e20308805cb77dc860e86a1d5

Observation 43ac81ea-8578-4a32-91fe-38ac259db8fc · outbound

This paper cites Self-attention Does Not Need $O(n^2)$ Memory.

Modality Agnostic Efficient Long Range Encoder Self-attention Does Not Need $O(n^2)$ Memory

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.415528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.415528Z digest=sha256:9635e7028287942aead0feafe4ac66696f07009579d7c5b4cdcdb8cc1815441d

Observation ea68c8ca-08bd-443e-9ac4-bcea0df233a0 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Modality Agnostic Efficient Long Range Encoder Learning Transferable Visual Models From Natural Language Supervision

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.419888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.419888Z digest=sha256:9784458f375f4ea42d0c68a9efc93a4ce27966d37ffd77bd296922b4dc62693c

Observation cf490f3b-ba2d-4c84-bc08-b07d41e60581 · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

Modality Agnostic Efficient Long Range Encoder Fast Transformer Decoding: One Write-Head is All You Need

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.424102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.424102Z digest=sha256:09e6007a63e938875bf0cf5a4d726a8630168a3c6b500f74e45ef990ac5b0b9b

Observation 8e568b0c-f1df-468c-a9e3-1b02f0a5ae8c · outbound

This paper cites Dropout: A simple way to prevent neural networks from overfitting.

Modality Agnostic Efficient Long Range Encoder Dropout: A simple way to prevent neural networks from overfitting

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.015343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.428525Z digest=sha256:eb10a00882cb795f59557057745d1dec8917bddd7fd0d8d4abd44224092d02ab

Observation d6b1d789-7f89-4778-bc55-3f9a920933e9 · outbound

This paper cites Scaling Granite Code Models to 128K Context.

Modality Agnostic Efficient Long Range Encoder Scaling Granite Code Models to 128K Context

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.433055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.433055Z digest=sha256:e738943054019878e796600d8eb108e338e73a5c1707c103b421b2567a76bec2

Observation 8b34aa8f-5245-4abd-bf37-e9ac6d438833 · outbound

This paper cites Rethinking the inception architecture for computer vision.

Modality Agnostic Efficient Long Range Encoder Rethinking the inception architecture for computer vision

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.002442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.438136Z digest=sha256:63ce95fd798d79157bbc9115ab6235dc0ff55c8d34f0bd1b7af60f87412cebb5

Observation 27c25218-4a4c-4e58-b892-c55857594733 · outbound

This paper cites Sparse sinkhorn attention.

Modality Agnostic Efficient Long Range Encoder Sparse sinkhorn attention

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.988808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.442255Z digest=sha256:afb5526a53f4549d3f2d8f83c36bf6744adc1f2e54480473aff57758b9a2f551

Observation c568d9a2-a81d-4380-97b2-11a0af045455 · outbound

This paper cites Long range arena: A benchmark for efficient transformers.

Modality Agnostic Efficient Long Range Encoder Long range arena: A benchmark for efficient transformers

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.975646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.446268Z digest=sha256:e2ebc35eff3d33c709f3cc141fae54c63b6349e172d19ee08f9ef2b002b6a45e

Observation 9b3e551d-34f2-4704-b71c-0c858959ce64 · outbound

This paper cites Atten- tion is all you need.

Modality Agnostic Efficient Long Range Encoder Atten- tion is all you need

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.962476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.450471Z digest=sha256:221cb89fe15979368897172c1b3ffa7f536be16f8f78ff22f2bb94a0d73e41e5

Observation 27f9710a-aefa-4135-8eab-6c0349fa9bd2 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Modality Agnostic Efficient Long Range Encoder Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.454731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.454731Z digest=sha256:3bc588bc2c6a4e8fbf000f043d5b92d2ee58fcf10552275739e036f4bddf73e9

Observation d3cefd6b-9bfa-48fd-bdbe-9f1891eb5021 · outbound

This paper cites Huang, Krzysztof Choromanski, Valerii Likhosherstov, and Adrian Weller.

Modality Agnostic Efficient Long Range Encoder Huang, Krzysztof Choromanski, Valerii Likhosherstov, and Adrian Weller

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.948665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.459274Z digest=sha256:9048dbfd7ae0176a7ff77fb2322497044dad71b4b843de10169f09a0eb150709

Observation 205864dc-13d0-4511-9f0b-658740b3868d · outbound

This paper cites MeSHup: A Corpus for Full Text Biomedical Document Indexing.

Modality Agnostic Efficient Long Range Encoder MeSHup: A Corpus for Full Text Biomedical Document Indexing

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:59:25.605672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.464664Z digest=sha256:96c2e4f259a0e2b2427ae15251f7d0a21ba1d82e17c9da30ae7a11cab7485a47

Observation aed7673b-bfe3-4ea9-8b30-3d74e96e3ed5 · outbound

This paper cites Pytorch image models.

Modality Agnostic Efficient Long Range Encoder Pytorch image models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.934518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.469293Z digest=sha256:d2b3f4cb93a81b227eca2efe5659af497d478c0ebb4854a1faf4b9b3efb1332a

Observation e5fe8155-0781-420d-a003-e7eaaacebf1c · outbound

This paper cites Ef- fective long-context scaling of foundation models.

Modality Agnostic Efficient Long Range Encoder Ef- fective long-context scaling of foundation models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.920457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.473274Z digest=sha256:28a2a07fccfe7f8701d6fe1c62dd318d4d583de65c8002685c006856d1624a74

Observation 1da600db-4868-46d7-ac0d-f7ffa5f80bf4 · outbound

This paper cites A-vit: Adaptive tokens for efficient vision transformer.

Modality Agnostic Efficient Long Range Encoder A-vit: Adaptive tokens for efficient vision transformer

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.906773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.477372Z digest=sha256:ff2541f9f8efceb52019b6bf1f1d4981187e000d2ca3154625202d4817781798

Observation 65337461-5d63-4e10-b756-0c38205a0dbe · outbound

This paper cites MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers.

Modality Agnostic Efficient Long Range Encoder MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.481718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.481718Z digest=sha256:bd9372271f75ebae991b95c36772a4cfe6fc4f487e2c2ea988eaf32586e72edc

Observation b1a65efb-5ebd-44f3-bf4f-7a6b80d8256e · outbound

This paper cites Metaformer is actually what you need for vision.

Modality Agnostic Efficient Long Range Encoder Metaformer is actually what you need for vision

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.892725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.486110Z digest=sha256:37b994c6e4f20a1fb88d454831ce49485cc058990296de83989053f8ac36a361

Observation c4b9783c-bed3-400d-8f7c-dd4510e5d740 · outbound

This paper cites Cutmix: Regularization strategy to train strong classifiers with localizable features.

Modality Agnostic Efficient Long Range Encoder Cutmix: Regularization strategy to train strong classifiers with localizable features

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.877463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.490295Z digest=sha256:3eadebce93b944ab67eed4aa74e2bdfd5d642ff52de3abd676e68bb98df4cec7

Observation 1768df0d-4863-4051-8548-0f4155c2aea5 · outbound

This paper cites Big bird: Trans- formers for longer sequences.

Modality Agnostic Efficient Long Range Encoder Big bird: Trans- formers for longer sequences

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.863098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.494565Z digest=sha256:1672e42c9daae973224f6ad96c1c972699c72cebb7d81cd874dc2e71bd364fe5

Observation c6651615-4d64-4f29-b1cc-0edb604dc96f · outbound

This paper cites mixup: Beyond empirical risk minimization.

Modality Agnostic Efficient Long Range Encoder mixup: Beyond empirical risk minimization

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.849604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.498610Z digest=sha256:f45db00c46e9d1db1a4037cf1511d729bdd9909ae90f160cb939cb347fd7e513

Observation 39cd3b71-1fcd-483c-8f13-9eaaa2c1e444 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Modality Agnostic Efficient Long Range Encoder OPT: Open Pre-trained Transformer Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.502868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.502868Z digest=sha256:f484f3f0851a7437e92b6ae8d1ccb8797554b7ba67e0804779b8ee6363e0922a

Observation 9db88336-7389-4943-966a-6d4c20ab901c · outbound

This paper cites Adavit: Adaptive vision transformers for efficient image recognition.

Modality Agnostic Efficient Long Range Encoder Adavit: Adaptive vision transformers for efficient image recognition

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.835846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.507154Z digest=sha256:6f6fb7087734ef57d0383c0fa9dc31bbd28de86cf5c0be00db83ce9b0bbe2ca1

Observation 27b8126e-4ca5-4088-86e3-ad1eaa430b5f · outbound

This paper cites Long-short transformer: Efficient transformers for language and vision.

Modality Agnostic Efficient Long Range Encoder Long-short transformer: Efficient transformers for language and vision

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.822039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.511224Z digest=sha256:a7c0c7c55e9db4219e5fa392052cb2358d6a2ea1f3e4a38d97d790938138dc02

Observation feabdaa6-8dd1-42ba-857f-21e471ac9c98 · outbound

This paper cites Multiscale Audio Spectrogram Transformer for Efficient Audio Classification.

Modality Agnostic Efficient Long Range Encoder Multiscale Audio Spectrogram Transformer for Efficient Audio Classification

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:59:25.556777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T17:59:25.515340Z digest=sha256:32fdaa0dee38ad7694f765d9b68835a253be96b8df4006a60d5f02be6f97a11b

Pith citing papers

No inbound Pith citation observations are available.