Pith. sign in

Paper Citation Record · LEDGER

Do pretrained Transformers Learn In-Context by Gradient Descent?

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2310.08540.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.08540 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:38.144629Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T00:33:52.628412Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a8d7fcd6-955c-4bce-8e16-aca8cdf94b8b · inbound

A Survey on In-context Learning cites this paper.

A Survey on In-context Learning Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T12:58:27.549625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T12:58:27.430374Z digest=sha256:ab1d0343f44fde7f5fc1c50d258b631dc5bfa573d730663e2ef269f29c8936b1

Observation 2109c450-790a-4dc0-b879-1194a1b2b11a · inbound

The Role of Diversity in In-Context Learning for Large Language Models cites this paper.

The Role of Diversity in In-Context Learning for Large Language Models Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:38.144629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:21:38.144629Z digest=sha256:e3b269a7de816a1238d2d53bd224a6b489a74a0e989f421385f7fc893563ec3d

Observation 40fd2e3a-7623-47cf-bc56-a9e2ef42d7db · inbound

Relational reasoning and inductive bias in transformers and large language models cites this paper.

Relational reasoning and inductive bias in transformers and large language models Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:32:17.057473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T11:31:37.942517Z digest=sha256:3df915123a04facb8918cb3d6106c7715cb331900377fca182cac290e123b7a8

Observation 43f50103-e083-430a-b5da-18b7258417fa · inbound

Transformers Meet In-Context Learning: A Universal Approximation Theory cites this paper.

Transformers Meet In-Context Learning: A Universal Approximation Theory Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T10:33:41.646967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:33:41.646967Z digest=sha256:ae2eabc771cc49e3ddf38e68e569bf62513899985cb7203d071cd0559b5f0d1f

Observation 4637800a-69f2-4bdf-a36b-dcc194d23490 · inbound

Transformers Don't In-Context Learn Least Squares Regression cites this paper.

Transformers Don't In-Context Learn Least Squares Regression Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.967182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.967182Z digest=sha256:a03c56a6aeb82fc2931545bb2fa44f72f5f27050800f8736da83ed6a3c88c7d8

Observation 29504811-0214-4f2c-b03a-4db3e8496101 · inbound

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective cites this paper.

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:12.470216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T08:06:25.436395Z digest=sha256:3d6ffedcc991dfc1f4d95ee87813e074d7fb7d27cc4a437c283d3132377e17f2

Observation fe6d1080-e341-46aa-9d55-944692354211 · inbound

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective cites this paper.

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-21T00:33:52.630143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T00:33:22.181383Z digest=sha256:ffad15f590cfc3dae9822d5210fe69d23e3b58a29993b11981e246b68e195b4e

Observation a4a73ce1-3014-44c2-855b-484563b2c91a · inbound

Mitigating Many-shot Jailbreak Attacks with One Single Demonstration cites this paper.

Mitigating Many-shot Jailbreak Attacks with One Single Demonstration Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T00:51:15.042381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T00:50:19.648405Z digest=sha256:6129bc31518d41d0bd37e7748864ef9ceca5cc21a5145e590f6bac3b470f795c