Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T12:54:36.489014Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2510.01718.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T12:54:36.489014Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 73142ff2-4f01-4393-9534-36de299848ad · outbound
Accelerating Attention with Basis Decomposition Croci, Marcelo Gennari do Nascimento, Torsten Hoefler, and James Hensman
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 640eeb3a-151f-454f-9f2f-4da6f2fa88f4 · outbound
Accelerating Attention with Basis Decomposition Longformer: The Long-Document Transformer
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7045ed7-be93-4994-92ba-59ddb6e45bc7 · outbound
Accelerating Attention with Basis Decomposition Language Models are Few-Shot Learners
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00d8f182-ab52-47d9-8058-734c395a446f · outbound
Accelerating Attention with Basis Decomposition Linear least squares solutions by householder transformations
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a81a44c-5501-4bb1-a19d-95d6a2fb30ee · outbound
Accelerating Attention with Basis Decomposition u ker, Luisa Bentivogli, and Marcello Federico. Report on the 11th IWSLT evaluation campaign. In Marcello Federico, Sebastian St \
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5af29c4f-d56f-4c78-ac2c-3a8c2f8f77cb · outbound
Accelerating Attention with Basis Decomposition Generating Long Sequences with Sparse Transformers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce563b0-900f-4561-a90d-b07ee1fdde09 · outbound
Accelerating Attention with Basis Decomposition Rethinking attention with performers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20be47fb-d94d-4b85-8eb5-9b30cf963f76 · outbound
Accelerating Attention with Basis Decomposition FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21a61534-ad57-4eeb-a424-c151aadd6afb · outbound
Accelerating Attention with Basis Decomposition Flashattention: Fast and memory-efficient exact attention with io-awareness
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6941ade-a66e-46b3-b115-1816909e2f36 · outbound
Accelerating Attention with Basis Decomposition An image is worth 16x16 words: Transformers for image recognition at scale
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2bba1f1-1e01-4127-83a6-4f2bc1d452d3 · outbound
Accelerating Attention with Basis Decomposition Sparsegpt: Massive language models can be accurately pruned in one-shot
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6010a440-46a5-4e74-89ff-412c8e4fa407 · outbound
Accelerating Attention with Basis Decomposition OPTQ : Accurate quantization for generative pre-trained transformers
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9db1f6-3352-4f02-9ef6-f5ca335451d4 · outbound
Accelerating Attention with Basis Decomposition Strategies for applying low rank decomposition to transformer-based models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a057aba3-d32e-4e15-9bf0-b17222eaee1d · outbound
Accelerating Attention with Basis Decomposition SLTrain : a sparse plus low-rank approach for parameter and memory efficient pretraining
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be4e18b6-074b-42af-9377-38971a8d2709 · outbound
Accelerating Attention with Basis Decomposition Language model compression with weighted low-rank factorization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a63d98f-6da0-464d-9196-37cc203100a1 · outbound
Accelerating Attention with Basis Decomposition Lo RA : Low-rank adaptation of large language models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b676826-9b17-4b04-b090-538dc0d1f1d4 · outbound
Accelerating Attention with Basis Decomposition From Low Rank Gradient Subspace Stabilization to Low-Rank Weights: Observations, Theories, and Applications
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6a14013-6408-4d60-858d-e52a7e75f9e7 · outbound
Accelerating Attention with Basis Decomposition Exploring Low Rank Training of Deep Neural Networks
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 631783ef-d799-40a2-8984-3a39a3333220 · outbound
Accelerating Attention with Basis Decomposition Transformers are rnns: Fast autoregressive transformers with linear attention
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dedb3b3c-0143-447c-86ac-5a807bdcfb2e · outbound
Accelerating Attention with Basis Decomposition LORD: Low Rank Decomposition Of Monolingual Code LLMs For One-Shot Compression
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec62ac58-2159-4e00-978d-e4bd905ef6e1 · outbound
Accelerating Attention with Basis Decomposition Tenenholtz, Lester Mackey, and Nicolo Fusi
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f83de22-e96d-4476-9f62-7053a68070a3 · outbound
Accelerating Attention with Basis Decomposition Reformer: The efficient transformer
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31db1703-5b8b-4a93-b2ca-a68b152ecc44 · outbound
Accelerating Attention with Basis Decomposition Efficient memory management for large language model serving with pagedattention
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 643ba7d0-1136-467c-b008-08c2923f039d · outbound
Accelerating Attention with Basis Decomposition Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01fa899c-3f53-4895-be6e-7047fb963ed7 · outbound
Accelerating Attention with Basis Decomposition L o S parse: Structured compression of large language models based on low-rank and sparse approximation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f261d3aa-14a0-4a8d-9c5a-d529ae8ba1dc · outbound
Accelerating Attention with Basis Decomposition Relo RA : High-rank training through low-rank updates
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77cdf235-0bfd-4f03-99a5-b557d1825d12 · outbound
Accelerating Attention with Basis Decomposition MoDeGPT: Modular Decomposition for Large Language Model Compression
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83137743-3cb1-4ff3-b11a-6cf1c74487d4 · outbound
Accelerating Attention with Basis Decomposition Duquant: Distributing outliers via dual transformation makes stronger quantized LLM s
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f0d11ae-4f6e-4531-a7cc-a1a6c23fc160 · outbound
Accelerating Attention with Basis Decomposition Awq: Activation-aware weight quantization for on-device llm compression and acceleration
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ae5674c-ffb0-4b20-839e-613f321f3d60 · outbound
Accelerating Attention with Basis Decomposition DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07b3b64d-3b08-49e8-8f1c-d43f6e3b80c1 · outbound
Accelerating Attention with Basis Decomposition DeepSeek-V3 Technical Report
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0258c38e-9abe-4a8a-b13e-c7ebc7dac914 · outbound
Accelerating Attention with Basis Decomposition Dora: weight-decomposed low-rank adaptation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd5384e6-bb56-4bcc-8dd3-d47a66c0135b · outbound
Accelerating Attention with Basis Decomposition Eora: Fine-tuning-free compensation for compressed llm with eigenspace low-rank approximation, 2025
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ac60136-ac83-4676-a145-dda4bcac9790 · outbound
Accelerating Attention with Basis Decomposition Llm-pruner: On the structural pruning of large language models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 953b5e5b-d8ec-4d75-a517-f86fd2b5d80d · outbound
Accelerating Attention with Basis Decomposition Pi SSA : Principal singular values and singular vectors adaptation of large language models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dbea6ef-0b63-4017-9997-9c4ea74232fe · outbound
Accelerating Attention with Basis Decomposition Accelerating Sparse Deep Neural Networks
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d439b2ad-6e08-4225-bfbd-55466bb6f4e4 · outbound
Accelerating Attention with Basis Decomposition Dobi-svd: Differentiable svd for llm compression and some new perspectives
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f961544-b9b2-4dd2-8839-cc51cc6df62f · outbound
Accelerating Attention with Basis Decomposition Improving language understanding by generative pre-training
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c43b9212-5043-4bf1-9d40-21ba7e391419 · outbound
Accelerating Attention with Basis Decomposition Learning transferable visual models from natural language supervision
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 278dc369-fde2-4308-9668-e49d0114f923 · outbound
Accelerating Attention with Basis Decomposition Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31e01949-78d7-4159-beb9-0126df88ca27 · outbound
Accelerating Attention with Basis Decomposition Compressing large language models using low rank and low precision decomposition
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 342da8c0-a4e9-4b3f-9574-f2d413e0f6b4 · outbound
Accelerating Attention with Basis Decomposition ESPACE : Dimensionality reduction of activations for model compression
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e920ecbc-9788-49b6-8e95-0174b00261b3 · outbound
Accelerating Attention with Basis Decomposition Robust low-rank training via approximate orthonormal constraints
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ea1abe1-5963-4ead-a7f4-2d53663b6ac1 · outbound
Accelerating Attention with Basis Decomposition Low-rank lottery tickets: finding efficient low-rank neural networks via matrix differential equations
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93e826c3-8a4c-4eb0-a545-a663f681d28b · outbound
Accelerating Attention with Basis Decomposition Flashattention-3: Fast and accurate attention with asynchrony and low-precision
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 586d5a36-5a80-422b-b19e-72f3bccac425 · outbound
Accelerating Attention with Basis Decomposition The Truth is in There: Improving Reasoning in Language Models with Layer-Selective Rank Reduction
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 528f9003-faa7-48b4-80d5-90af2d766239 · outbound
Accelerating Attention with Basis Decomposition Roformer: Enhanced transformer with rotary position embedding, 2021
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18fd2129-8fbe-4609-949c-f15b7ecfa635 · outbound
Accelerating Attention with Basis Decomposition A simple and effective pruning approach for large language models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ae9a8c8-8529-4874-9453-f8d373e940e4 · outbound
Accelerating Attention with Basis Decomposition LLaMA: Open and Efficient Foundation Language Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e2ec9d2-3863-4724-a5b5-713fdfd1aa2f · outbound
Accelerating Attention with Basis Decomposition Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0136548-5b07-446f-8cd5-045e713c7f8b · outbound
Accelerating Attention with Basis Decomposition Attention is all you need
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bec3bf8d-1d0f-4873-9ac4-a9fea3897be1 · outbound
Accelerating Attention with Basis Decomposition Linformer: Self-Attention with Linear Complexity
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d362884c-1a04-4ac8-93cd-c40cc7ca3094 · outbound
Accelerating Attention with Basis Decomposition SVD-LLM: Truncation-aware Singular Value Decomposition for Large Language Model Compression
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c321b518-a77c-447f-97dd-88711359f07e · outbound
Accelerating Attention with Basis Decomposition S mooth Q uant: Accurate and efficient post-training quantization for large language models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a849161-2923-4d55-bd1f-e3a7efe8544a · outbound
Accelerating Attention with Basis Decomposition ASVD: Activation-aware Singular Value Decomposition for Compressing Large Language Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd338049-9301-4ef7-b7ea-92b6c9f09adf · outbound
Accelerating Attention with Basis Decomposition IncreLoRA: Incremental Parameter Allocation Method for Parameter-Efficient Fine-tuning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a45d9535-e250-4bd4-bb31-6d366b7a860a · outbound
Accelerating Attention with Basis Decomposition Adaptive budget allocation for parameter-efficient fine-tuning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1bce96b-694c-43aa-9b5e-4f059156fc26 · outbound
Accelerating Attention with Basis Decomposition OATS : Outlier-aware pruning through sparse and low rank decomposition
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc3d7516-ad5f-4dd5-a41c-de003089dfb9 · outbound
Accelerating Attention with Basis Decomposition Plug-and-play: An efficient post-training pruning method for large language models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f894493-e622-4418-a1e9-097317145de2 · outbound
Accelerating Attention with Basis Decomposition Pivoting Factorization: A Compact Meta Low-Rank Representation of Sparsity for Efficient Inference in Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dd34c8b-9224-430b-aadc-c57a7b72dc78 · outbound
Accelerating Attention with Basis Decomposition InRank: Incremental Low-Rank Learning
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f8f7a76-881c-47a7-8042-894e98116cea · outbound
Accelerating Attention with Basis Decomposition write newline
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d17f1a5-7d81-45eb-8f50-b2065d4371e8 · outbound
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93e27143-3258-49e6-9bd8-d9078579c549 · outbound
Accelerating Attention with Basis Decomposition Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07ee1195-c7ea-4742-86e3-af5f4901c69b · outbound
Accelerating Attention with Basis Decomposition u `:^!t )GeuwokcJ _ ]n?ICq .WT +BCBC &q=2
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.