Pith. sign in

Paper Citation Record · LEDGER

Parallelizing Linear Transformers with the Delta Rule over Sequence Length

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2406.06484.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.06484 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 39 of 39 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:25:42.751276Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.246124Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 45cbfaf9-f7d9-45b4-8537-f62bbf4cdf09 · inbound

Learning to (Learn at Test Time): RNNs with Expressive Hidden States cites this paper.

Learning to (Learn at Test Time): RNNs with Expressive Hidden States Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:20:12.362233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T05:20:12.134340Z digest=sha256:678d455cf4f401c7599e28baece65689298eeef18c6f7ddc9b2848b37f070dde

Observation a2b0d8cd-b9c5-450a-8bba-bde9b09a046c · inbound

LASP-2: Rethinking Sequence Parallelism for Linear Attention and Its Hybrid cites this paper.

LASP-2: Rethinking Sequence Parallelism for Linear Attention and Its Hybrid Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T12:25:42.751276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:25:42.751276Z digest=sha256:9bdb8f930a1bd22ceddfa60fdbf708e9a40736997a172ffa6a291f1e2622abab

Observation d4861c86-2a08-49fc-a7d7-aec6e08425eb · inbound

An Uncertainty Principle for Linear Recurrent Neural Networks cites this paper.

An Uncertainty Principle for Linear Recurrent Neural Networks Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T22:10:15.942223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:10:15.942223Z digest=sha256:1d5c516461a3e309cca14a449b94a8d7a647f7e59f03a7ef323e796ff98774b5

Observation 0ec81779-3758-4fce-8c0d-3256b2e0dd59 · inbound

Solving Empirical Bayes via Transformers cites this paper.

Solving Empirical Bayes via Transformers Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T20:21:58.415098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:21:58.415098Z digest=sha256:ea91ab208f5ef33f0eee00936a8bb04a26ee0f8cfb73d57f8fb603d134b0d727

Observation 18f01554-4540-4ce8-be3f-b9fbe424a5b2 · inbound

ModRWKV: Transformer Multimodality in Linear Time cites this paper.

ModRWKV: Transformer Multimodality in Linear Time Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.402647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.402647Z digest=sha256:3dd38336a8fbb2f73b46493cdf9fa59652987a040b9a0a9dafde64f71a2b5af6

Observation 02c654cb-d479-431e-b9f7-197bf653f528 · inbound

Understanding Transformer from the Perspective of Associative Memory cites this paper.

Understanding Transformer from the Perspective of Associative Memory Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:19:54.017185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:19:54.017185Z digest=sha256:96880cd6af095c0028a4df0a8dc8a2e5d45157dad0bb79bf6633711465768b08

Observation 23b43d5e-128a-4cf9-b784-405c7b5b86e2 · inbound

HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling cites this paper.

HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:43.404040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:43.404040Z digest=sha256:272990ab309ddf9ecf9df648bae846988f6487bd4bd94706c657770e419a06a8

Observation 232cce74-e538-4a71-8cb9-6ad92e78a480 · inbound

Scaling Reasoning without Attention cites this paper.

Scaling Reasoning without Attention Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:13:28.091568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:13:28.091568Z digest=sha256:9f09d5a0a7fac0498673e248a09db85525b8e9e241b5be5cda8a47ca85f50efc

Observation 62fd2402-dff1-48e8-b4f3-c9e6fdceaf46 · inbound

Test-Time Training Done Right cites this paper.

Test-Time Training Done Right Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:25:45.215790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T11:25:45.153563Z digest=sha256:b531b6a23ae6077cec093fdb00090e942e136f3374d40f57b12cc1e3b741424c

Observation 83eca67f-71cf-45bc-b181-ec21efb3df9d · inbound

Cartridges: Lightweight and general-purpose long context representations via self-study cites this paper.

Cartridges: Lightweight and general-purpose long context representations via self-study Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T06:04:35.156790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:04:35.156790Z digest=sha256:1d40478b007511ee6f3d0c2e68f3516e95f2cc912f932f5f012ebf861e152e49

Observation ba368b4c-83f6-4f2c-9b86-76d1891b91af · inbound

SeerAttention-R: Sparse Attention Adaptation for Long Reasoning cites this paper.

SeerAttention-R: Sparse Attention Adaptation for Long Reasoning Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T05:06:32.732274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:06:32.732274Z digest=sha256:3189d5064febe9226b8129874447f0be3eff48c00e89de69285125b8eacc43b6

Observation 97f09e7b-fda8-42f7-ab1c-af5db940abdd · inbound

pLSTM: parallelizable Linear Source Transition Mark networks cites this paper.

pLSTM: parallelizable Linear Source Transition Mark networks Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T01:09:03.309496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:09:03.309496Z digest=sha256:2b128f0e4378fe8c720c182a438c7b04df591a5da5c3b7b03b2ab567407c311d

Observation 1e1eccb2-dc13-48c8-b255-08e6364afbe2 · inbound

TPTT: Transforming Pretrained Transformers into Titans cites this paper.

TPTT: Transforming Pretrained Transformers into Titans Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.547568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.547568Z digest=sha256:9d732fbf0eabfa02dde4a8d99cf9a47f9e1b4d411161054a84e814a623bae31c

Observation e8f220fa-a809-4837-a241-12a35269490b · inbound

A Survey on Latent Reasoning cites this paper.

A Survey on Latent Reasoning Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:32.936567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:32.936567Z digest=sha256:08cd40ec4efdeb6e402906b1ee74a41228e771fdabfe4c4eefed39e26f066f82

Observation f7501579-84d1-44ef-9bf6-57230956cc11 · inbound

Elucidating the Design Space of Decay in Linear Attention cites this paper.

Elucidating the Design Space of Decay in Linear Attention Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T05:29:21.681430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:29:21.681430Z digest=sha256:e35763acdf9d727d7abe9d38e5d6a78641b16c85be69a8a966656f88ecce83e5

Observation 7356987f-d11c-4a9d-b54c-fcb96e8ae67a · inbound

A Vision Toward Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents cites this paper.

A Vision Toward Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 161

Resolution
unresolved
no resolver link, observed 2026-08-04T08:12:26.022787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:12:26.022787Z digest=sha256:6690d5e16f95da5f65bd743b3dc466243d7720b3421bef13192cac8fd018d437

Observation 2c7c4f8b-9510-468a-90ef-3938ee3967c0 · inbound

Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism cites this paper.

Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:05:48.075690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T03:05:06.069642Z digest=sha256:2828d1cfa1f0c7527254be3b13e50113f728b7dd1c927e5d66c209fac7942185

Observation 2ff72c95-29fb-42f6-bd33-9d57a928ab5a · inbound

Kimi Linear: An Expressive, Efficient Attention Architecture cites this paper.

Kimi Linear: An Expressive, Efficient Attention Architecture Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:10.947759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:2758dc258fe30304c55e7595a09932c57fbda95c4b9980405ec309abcab8c9c4

Observation 89921422-127d-466e-a755-21bce9b43d24 · inbound

LADY: Linear Attention for Autonomous Driving Efficiency without Transformers cites this paper.

LADY: Linear Attention for Autonomous Driving Efficiency without Transformers Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T06:42:11.959356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:42:11.959356Z digest=sha256:17a9e502535ab98f3b63912e48abc8101021b7304c0bb7fa38bc3aededdc1254

Observation d9ea9905-e22c-49f2-91f1-fd79b33813c1 · inbound

Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers cites this paper.

Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T15:22:55.477441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:22:55.477441Z digest=sha256:5e039141d6dea6891507a8df8abb3b59f70cff36a30392c9eeb7f52b33d616be

Observation 0d74fb44-14e2-4840-82e4-fa5e16a48c4a · inbound

ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training cites this paper.

ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:26:17.502452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T16:21:16.229770Z digest=sha256:c9fea6954a550007c6eb7b5488499a2e1432ca104bc3955198a6b60f1c25d5e3

Observation 6056f5d8-f335-4de9-8225-71322553850c · inbound

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space cites this paper.

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:48.610341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:39:50.311059Z digest=sha256:a1a03b9ea3647f2e0d84115731d5aa291ce6e1ff79ec308f4e92eae0560e6aa0

Observation 5767fc1c-ee17-4eb9-94b1-5624f8e8337b · inbound

In-Place Test-Time Training cites this paper.

In-Place Test-Time Training Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:49.017551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:07:47.174513Z digest=sha256:bd218d80459f22277f871b73e140f38e3945765f356d8f02d88f6ad251c9b784

Observation cefb2429-bd4c-4845-84e1-d1a29bca075d · inbound

COREY: Entropy-Guided Runtime Chunk Scheduling for Selective Scan Kernels cites this paper.

COREY: Entropy-Guided Runtime Chunk Scheduling for Selective Scan Kernels Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:06:03.419188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:10:02.445509Z digest=sha256:494311e790a424b402f63e0bbb0198b38beba8bf21951d3d71b9aa5ef13b2307

Observation 13e9f8fe-4a18-4c48-a4d5-589d3698fa85 · inbound

Adaptive Memory Decay for Log-Linear Attention cites this paper.

Adaptive Memory Decay for Log-Linear Attention Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:50:56.408116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:02:56.848785Z digest=sha256:a042b289cce3b7ea6c1feab5118ff4e8bbff5a536bfe54cc726952d1dacf6b14

Observation 92437f3f-a79a-45a7-b2b5-9048abe9d5a1 · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:59:28.639301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T20:53:40.666929Z digest=sha256:e01b07f9c7616df1a04471b9cd7e7eca93b01d2add320f715309fb6e2bb43b38

Observation 5f2fda31-0cc1-4eb1-b180-e08fd22bd39b · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:59:45.229904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T04:59:11.877068Z digest=sha256:d8bc2e22d5afd6d7512d85c4490c209ba6a84f72ed3fcbf48350e9efe67a7f8a

Observation acf0bbbd-8876-4b89-8c2b-64d76c440ad0 · inbound

Towards Understanding Self-Pretraining for Sequence Classification cites this paper.

Towards Understanding Self-Pretraining for Sequence Classification Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:33:58.947730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T05:29:58.809024Z digest=sha256:3e3137c6c65fac18e4a070ca46690931a04989b9a0cf75b384d6114d4a249c48

Observation 6b38ad3b-883b-47a6-9e75-d3ab1ffdac2c · inbound

Pretraining Recurrent Networks without Recurrence cites this paper.

Pretraining Recurrent Networks without Recurrence Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 140

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:26:56.635118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T02:09:01.018909Z digest=sha256:da6e96c9c4c1c6bb42bcde00b4ed200977a539b96a085ea392b73cf0f9df82f7

Observation 0bc8bce5-eb20-43ed-907b-64dc34790fea · inbound

Pretraining Recurrent Networks without Recurrence cites this paper.

Pretraining Recurrent Networks without Recurrence Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-02T12:20:56.082616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:20:56.082616Z digest=sha256:3cdcddc439ac24adc855b5b7825c6cb1c9487523f728a796947779736b37ca30

Observation ae2b1780-dc47-47ba-b973-84e95c2a6006 · inbound

Reversible Foundations: Training a 120B Sparse MoE through State-Preserving Scaling cites this paper.

Reversible Foundations: Training a 120B Sparse MoE through State-Preserving Scaling Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:17:08.659060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T22:55:09.477413Z digest=sha256:7093d95a0d97e01e7d097e5218c45e9039a7dc0423c81d88bb316bfb979ac812

Observation b240309e-de36-4b54-a789-cbaff4debd74 · inbound

Q-Delta: Beyond Key-Value Associative State Evolution cites this paper.

Q-Delta: Beyond Key-Value Associative State Evolution Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 91

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T23:07:26.542463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T18:30:51.523567Z digest=sha256:7c7a71f94f58d3a88824e41d09d2ffad7d5c511c2b26eb7d1942c548226cd241

Observation 19a33f76-fce7-4f0f-a010-4f98ee69db51 · inbound

UltraQuant: 4-bit KV Caching for Context-Heavy Agents cites this paper.

UltraQuant: 4-bit KV Caching for Context-Heavy Agents Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:39:30.395444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T17:49:02.835708Z digest=sha256:73c16fc002b6b771a7075e87a0e249c188c456be38278bc4fafd31755a4bb394

Observation 9ca11e76-f7f5-48d7-b5e9-c377b9e88a02 · inbound

ELiTeFormer: An Efficient Transformer for FPGAs cites this paper.

ELiTeFormer: An Efficient Transformer for FPGAs Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-12T00:55:09.690245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:55:09.690245Z digest=sha256:3b86fa49d695d32bafd6bb83491de63d5be5a702ec956f5804e933496469632d

Observation 43c01be1-5d43-46ce-bc59-8222c393ed08 · inbound

Sparse Delta Memory: Scaling the State of Linear RNNs through Sparsity cites this paper.

Sparse Delta Memory: Scaling the State of Linear RNNs through Sparsity Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 114

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T12:46:14.646570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-09T12:40:09.036905Z digest=sha256:efc5c5e6a5654e4a6d7efe43c8a7167c6321c9ac1894163028449160520b55af

Observation 4301c229-8a9b-405e-b51f-8591ae6b2b84 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 140

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.247325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:2e6364c8999ad0fb58510cd3c67a7e16466b30915237f97273c1024cdf3c30d2

Observation 9568a843-40ac-4db2-9004-f5155f3a2d5b · inbound

The Orthogonalized Read Is a Removable Training Scaffold for Recurrent Memory cites this paper.

The Orthogonalized Read Is a Removable Training Scaffold for Recurrent Memory Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T09:06:11.171935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:06:11.171935Z digest=sha256:a4d10948afd3951b7879fafcdf6707921bd50f90e0654966f53c01cf78ac44b7

Observation dfb0e7d7-7f23-40be-a02f-8603824b40bc · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:51.979801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:51.979801Z digest=sha256:0942e204b8b509a9993a115173ba1b60797e0fb2172c9ef7658fa0670d5d6c8d

Observation 47293eaf-457b-4b0a-be73-ed5dfe5943e0 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:48.805634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:48.805634Z digest=sha256:7c89afed0c2402ca106d68984e9d8ada9c02f5873a45d62510497d9ff18926c7