Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:44:49.061779Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 1 inbound Pith citation observation for arXiv:2505.17852.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:44:49.061779Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-18T15:25:21.814232Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T15:26:33.682527Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c9085359-cf6e-40c8-9f34-c1ce24afa187 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Demystify Mamba in Vision: A Linear Attention Perspective
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 452e639c-9a53-41c3-9be9-ab3616972d59 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5fb0433-8b3e-4338-bd1a-a288115f0777 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization FlashRNN: I/O-Aware Optimization of Traditional RNNs on modern hardware
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c591d36a-529e-4088-ba46-6240fde96dd9 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Aditya Rawal and Risto Miikkulainen
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9cc023b2-3ebc-42ce-bb83-c01f2f4541a5 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization LLaMA: Open and Efficient Foundation Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0da4f7d-486b-4e53-84fc-1612d11b0d46 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Unbiased Online Recurrent Optimization
Reference 1992
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d32bcea1-64c7-49e2-9041-ea2875a5bec2 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Suraj Srinivas Malladi, Xiang Wei, Josip Djolonga, and Dale Schuurmans
Reference 2005
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5a19286-959e-4c74-8927-59b1935e1d25 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization doi: 10.1145/2908812.2908941
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e49812c-ecd8-4cb3-817d-fde0fbffd9fd · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Efficient Transformers: A Survey
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3cf41fa-b619-4722-bf73-eaa3d9399cda · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization OPT: Open Pre-trained Transformer Language Models
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b04da9fd-e410-415c-8f3e-82eb4e7ac224 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Koopman-informed recurrent neural networks
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccf275d2-8c12-401d-a93b-52688f3993e9 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Simultaneous Computation and Memory Efficient Zeroth-Order Optimizer for Fine-Tuning Large Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e3decb9-8c5a-4957-b82d-50682f695791 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Textbooks Are All You Need
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec1e4f59-413a-46e9-8bfc-2cae6704f589 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Linear attention is (maybe) all you need (to understand transformer optimization)
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d958b103-bc00-4191-8d6c-a7e0973594d7 · outbound
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization Alex Graves, Greg Wayne, and Ivo Danihelka
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e24ab96-e287-4206-8acc-317294dcf2a4 · inbound
Low-rank surrogate modeling and stochastic zero-order optimization for training of neural networks with black-box layers Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.