Pith. sign in

Paper Citation Record · LEDGER

Do pretrained Transformers Learn In-Context by Gradient Descent?

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2310.08540.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.08540 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:38.144629Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T00:33:52.628412Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a8d7fcd6-955c-4bce-8e16-aca8cdf94b8b · inbound

A Survey on In-context Learning cites this paper.

A Survey on In-context Learning Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T12:58:27.549625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T12:58:27.430374Z digest=sha256:57e244df08158d24b029aa065dac97ff3b3c6b88b662e4cd63df6e0ef7da164d

Observation 2109c450-790a-4dc0-b879-1194a1b2b11a · inbound

The Role of Diversity in In-Context Learning for Large Language Models cites this paper.

The Role of Diversity in In-Context Learning for Large Language Models Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:38.144629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:21:38.144629Z digest=sha256:64cd180400636f4d4c462d065d41026599899f6d19fd18b2265dc8455b45ee3a

Observation 40fd2e3a-7623-47cf-bc56-a9e2ef42d7db · inbound

Relational reasoning and inductive bias in transformers and large language models cites this paper.

Relational reasoning and inductive bias in transformers and large language models Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:32:17.057473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T11:31:37.942517Z digest=sha256:84fab0f3c442207f4da8258f6dc35a70723555027fd12a2557ab385b673b372a

Observation 43f50103-e083-430a-b5da-18b7258417fa · inbound

Transformers Meet In-Context Learning: A Universal Approximation Theory cites this paper.

Transformers Meet In-Context Learning: A Universal Approximation Theory Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T10:33:41.646967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:33:41.646967Z digest=sha256:f7143e459813f242465675b8647ccf489b8a2fc0b4ebc6cfd2110e338a80c2be

Observation 4637800a-69f2-4bdf-a36b-dcc194d23490 · inbound

Transformers Don't In-Context Learn Least Squares Regression cites this paper.

Transformers Don't In-Context Learn Least Squares Regression Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:03:14.967182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:03:14.967182Z digest=sha256:544fbf7d6500675801ef8a637e88e34d35fa7c1376dd01ed73ab9032393c0fb7

Observation 29504811-0214-4f2c-b03a-4db3e8496101 · inbound

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective cites this paper.

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:12.470216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T08:06:25.436395Z digest=sha256:2645e67d8d69cffefcf496d611eff37bd765f74404714c3b6d2f1de56dee2833

Observation fe6d1080-e341-46aa-9d55-944692354211 · inbound

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective cites this paper.

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-21T00:33:52.630143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T00:33:22.181383Z digest=sha256:65f787a778840b2ff4d6ec36e332bfbe636ceee48011d22964adae6c0be663f9

Observation a4a73ce1-3014-44c2-855b-484563b2c91a · inbound

Mitigating Many-shot Jailbreak Attacks with One Single Demonstration cites this paper.

Mitigating Many-shot Jailbreak Attacks with One Single Demonstration Do pretrained Transformers Learn In-Context by Gradient Descent?

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T00:51:15.042381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T00:50:19.648405Z digest=sha256:431696b66dcf4af0861ea24faa22b0b729399885ab15b5f89d3a534c734311c5