Pith. sign in

Paper Citation Record · LEDGER

Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2409.10559.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.10559 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:10:29.853631Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:06:44.923754Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 782a5028-34d8-4e92-8e4e-70c71505326b · inbound

KV Shifting Attention Enhances Language Modeling cites this paper.

KV Shifting Attention Enhances Language Modeling Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T10:10:29.853631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:10:29.853631Z digest=sha256:35170f4fcd11c1c5d2ceb9768eea09d5264c6f373340c64fb85fbca8db165f34

Observation 22a97b4d-9974-4b12-88db-2eab6e4ca4f0 · inbound

Rethinking Associative Memory Mechanism in Induction Head cites this paper.

Rethinking Associative Memory Mechanism in Induction Head Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:06:30.635308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:06:30.635308Z digest=sha256:3d54e78dcb5b5e2e40d54a51efdbee98884836c42f00dab8668620a140398d39

Observation b134bc0c-7e37-4fb7-824d-df2c5462488e · inbound

Learning Compositional Functions with Transformers from Easy-to-Hard Data cites this paper.

Learning Compositional Functions with Transformers from Easy-to-Hard Data Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:46:34.250512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:46:34.250512Z digest=sha256:dca7f2559fd95a08cb736eedfec005372777fa56d1016a439c887bb1fcde43f7

Observation e3c186f9-43e5-4d34-81ae-1abd813bc99b · inbound

A Survey on Latent Reasoning cites this paper.

A Survey on Latent Reasoning Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:22.935329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:22.935329Z digest=sha256:7eb5b416dd5b8ec894f7411af61d6da3c2af9b40fef2aaa9587db11a89002745

Observation 4b530587-a979-456d-87d1-5643f5b9a15a · inbound

Agentic Transformers Provably Learn to Search via Reinforcement Learning cites this paper.

Agentic Transformers Provably Learn to Search via Reinforcement Learning Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:42:49.925837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-28T23:26:28.158991Z digest=sha256:4f8d49eb0336de2ab15341b278d29ed7affb7444efc6e5250b796890e88fbd1d

Observation 6f20b3ba-50e7-41d0-b103-bfa5c7952806 · inbound

Why Muon Outperforms Adam: A Curvature Perspective cites this paper.

Why Muon Outperforms Adam: A Curvature Perspective Unveiling Induction Heads: Provable Training Dynamics and Feature Learning in Transformers

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:06:44.925493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-28T07:04:21.012269Z digest=sha256:6b0be2dbef4f4fdd714dec065a77bb7d2af525cb3f9eac5f54265ffa3c5d6206