Pith. sign in

Paper Citation Record · LEDGER

Selective Attention Improves Transformer

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.02703.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02703 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T22:34:08.054013Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T14:33:31.328935Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4580450f-51c6-4231-87ed-b6b090347ca0 · inbound

MUDDFormer: Breaking Residual Bottlenecks in Transformers via Multiway Dynamic Dense Connections cites this paper.

MUDDFormer: Breaking Residual Bottlenecks in Transformers via Multiway Dynamic Dense Connections Selective Attention Improves Transformer

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T22:34:08.054013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T22:34:08.054013Z digest=sha256:8f2f598ebb44179df951be6abf7d9d02f61917b18ba561deeb6d3afb10c4ade7

Observation ed200003-100b-40e0-9c1b-890f70af1cbe · inbound

StreamIndex: Memory-Bounded Compressed Sparse Attention via Streaming Top-k cites this paper.

StreamIndex: Memory-Bounded Compressed Sparse Attention via Streaming Top-k Selective Attention Improves Transformer

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:15:37.655983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T18:44:42.456111Z digest=sha256:e7f604776b6c5d3d4a4013be692e2450b1a0e90572fee40d8c9819c14db648ec

Observation ccd74e93-5200-4999-927e-203e6a584196 · inbound

Most Transformer Modifications Still Do Not Transfer at 1-3B: A 2020-2026 Update to Narang et al. (2021) with Downstream Evaluation and a Noise Floor cites this paper.

Most Transformer Modifications Still Do Not Transfer at 1-3B: A 2020-2026 Update to Narang et al. (2021) with Downstream Evaluation and a Noise Floor Selective Attention Improves Transformer

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:19:42.048320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T06:15:47.451870Z digest=sha256:24654c6d236b62a3a8097406b5d3522658be3d33510d9b8d92963ff2c82274d7

Observation 55e852be-ff82-45ff-87e9-4546c425d184 · inbound

Attention-based optimizer for symmetry finding cites this paper.

Attention-based optimizer for symmetry finding Selective Attention Improves Transformer

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:33:31.330444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T06:35:55.732968Z digest=sha256:8ca4b0a7c7eac52e5744e0f36071c077e1982629dbddcab68f1157e07eac912f

Observation 9d0894d7-d7f6-4bd8-9b9b-d9c088546a0d · inbound

Kernelized Linear Attention: Breaking the Capacity Wall with Symmetric Cones cites this paper.

Kernelized Linear Attention: Breaking the Capacity Wall with Symmetric Cones Selective Attention Improves Transformer

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T18:06:54.522323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:06:54.522323Z digest=sha256:76b962734319be43e71cd7da03b556abbb18a76a93ae456e784fb05ab05c6ad2

Observation 430744ad-7d24-4fa1-8623-4b32c6560d22 · inbound

Memory for Large Language Models cites this paper.

Memory for Large Language Models Selective Attention Improves Transformer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T02:37:54.185863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:37:54.185863Z digest=sha256:c9ba5e5537f4cfc10d37dad3853557bf4540da71e396c4cf05a4c8747b2a98ec