Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.04084.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:49:39.588423Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T23:57:29.167982Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 51982e89-d015-491b-8279-7a1ff550281b · inbound
Training Dynamics of In-Context Learning in Linear Attention Provably learning a multi-head attention layer
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26894e0e-f0a5-48e1-9329-2652a5012270 · inbound
Attention Mechanism, Max-Affine Partition, and Universal Approximation Provably learning a multi-head attention layer
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d55574b-e675-4f26-ac77-97799f6e2fa2 · inbound
How Transformers Learn Regular Language Recognition: A Theoretical Study on Training Dynamics and Implicit Bias Provably learning a multi-head attention layer
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5fee61a-8b01-4b22-9bcd-6d619cbfed03 · inbound
Transformers Meet In-Context Learning: A Universal Approximation Theory Provably learning a multi-head attention layer
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56b1e1c8-31ae-4eb2-a06f-9c493c7e724e · inbound
A Theoretical Study of (Hyper) Self-Attention through the Lens of Interactions: Representation, Training, Generalization Provably learning a multi-head attention layer
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0a40042-207e-41fd-babe-981385d2535f · inbound
Tight Sample Complexity of Transformers Provably learning a multi-head attention layer
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.