Pith. sign in

Paper Citation Record · LEDGER

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model

As of 17 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2606.13280.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.13280 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T05:08:00.711576Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c7e90caf-61e0-4e2e-821b-938a02463cb5 · outbound

This paper cites [2]Anderson, T.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model [2]Anderson, T

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T05:08:00.711576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:761cd593ccc32f9493ab813aa44a857dd464f440ad9932f9a1f26b24188b0456

Observation b62d4437-18d7-4ab2-bc63-08029b1fb99d · outbound

This paper cites In2012 IEEE 53rd Annual Symposium on Foundations of Computer Science(2012), pp.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model In2012 IEEE 53rd Annual Symposium on Foundations of Computer Science(2012), pp

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T05:08:00.711576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:77bdb5f40b0c5ef021991a303ad02d9f49ea11fe94788944d1ae29708ecfef94

Observation 4e106056-835a-4c17-8d75-ecca0fb972cb · outbound

This paper cites Efficient and Minimax Optimal In-context Nonparametric Regression with Transformers.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model Efficient and Minimax Optimal In-context Nonparametric Regression with Transformers

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-03T16:38:40.632510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:5cea3da74039b7be0845dda2e6e4d66a0c2570e713ee2941a69407b413e98d73

Observation f1d4879a-a851-498b-a8aa-0c87d41b0441 · outbound

This paper cites Transformer Approximations from ReLUs.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model Transformer Approximations from ReLUs

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T16:38:40.640997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:dc622585f1be9b6f8106b69302f182b95076b40c471ce65e4308a92055a94d15

Observation ba9df328-6b4f-4b4d-b013-faa3b0d79172 · outbound

This paper cites an unresolved cited work.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T05:08:00.711576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:b2a65aa7e7f1c1e2c298456e6088096257d330acd6008fa6f7aab4bc1a5bff94

Observation 68c41094-657e-4173-97e4-2c0684c89151 · outbound

This paper cites Standard Transformers Achieve the Minimax Rate in Nonparametric Regression with $C^{s,\lambda}$ Targets.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model Standard Transformers Achieve the Minimax Rate in Nonparametric Regression with $C^{s,\lambda}$ Targets

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-29T02:24:15.834827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:6a7bc79e1946fb4fc7a94b16e00161a13889a4686d53107e9ca97cbf8383aa7c

Observation 64386fe3-6f1b-4140-be9e-ca41db70a115 · outbound

This paper cites T.Optimal convergence rates of deep neural networks in a classification setting.Electronic Journal of Statistics 17, 2 (2023), 3613 –.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model T.Optimal convergence rates of deep neural networks in a classification setting.Electronic Journal of Statistics 17, 2 (2023), 3613 –

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T05:08:00.711576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:3275b43c6c477b4f54c71e529b49e680df898cb67ff2d8fbb8a434aa1ed3aa6c

Observation 4ebdc87d-7ea4-4d16-9d6c-444e9340a4d6 · outbound

This paper cites InPro- ceedings of the 24th international conference on Machine learning(2007), pp.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model InPro- ceedings of the 24th international conference on Machine learning(2007), pp

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-27T05:08:00.711576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:389778280ed82e684fb0a7c7c408eaf74d489c74a41e84657756ce535855a0f9

Observation a032cd47-e354-49a2-ace3-488da6d12ce6 · outbound

This paper cites [60]Sanford, C., Hsu, D., and Telgarsky, M.Representational strengths and limitations of trans- formers.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model [60]Sanford, C., Hsu, D., and Telgarsky, M.Representational strengths and limitations of trans- formers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T05:08:00.711576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:06eb4cc58fdc6c78b9c5994346b88d6c8476545fafc5553c8b45361839cb769e

Observation eb2ef265-4f51-4b7a-b474-2096563d0769 · outbound

This paper cites Transformers for Learning on Noisy and Task-Level Manifolds: Approximation and Generalization Insights.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model Transformers for Learning on Noisy and Task-Level Manifolds: Approximation and Generalization Insights

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-03T16:38:40.641256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:ea7860e575699ec732660affa4fcb04c8636efabb36ea0ebb7b37336a984a9b3

Observation fca37527-c335-4aa9-b868-83d92a1fd0de · outbound

This paper cites In-context learning is provably bayesian inference: a generalization theory for meta-learning.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model In-context learning is provably bayesian inference: a generalization theory for meta-learning

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:38:40.638322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:7aa43f35069d4e56ccc52245b2fb39e1475eb9bdd01ce6b9971b2f38aed7f287

Observation 52811563-5949-4ce2-8cfa-793b11350af1 · outbound

This paper cites Large Language Models as Markov Chains.

Generalization Bounds for Transformer-Based Next-Token Prediction in a Language Model Large Language Models as Markov Chains

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T16:38:40.644732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T05:08:00.711576Z digest=sha256:0996dce1d16506a2e238a54b4460a9484ca31818be3e8a4741ed7092c8edbddb

Pith citing papers

No inbound Pith citation observations are available.