Pith. sign in

Paper Citation Record · LEDGER

ICR-RL: Deep Reinforcement Learning via In-Context Regression

As of 17 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 8 inbound Pith citation observations for arXiv:2509.11259.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.11259 v2

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T16:52:12.064508Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T07:13:43.014623Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 2b72894c-2a56-4bc3-9311-eafe04f6e3f3 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

ICR-RL: Deep Reinforcement Learning via In-Context Regression , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:09.673688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:09.673688Z digest=sha256:4746a84ef18716492c6e6449609b7ece7f182c36343523788559987ae87d6ad5

Observation 496532dc-0da0-4a29-aba8-95ee561e3b3d · outbound

This paper cites write newline.

ICR-RL: Deep Reinforcement Learning via In-Context Regression write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:09.775508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:09.775508Z digest=sha256:f11641e023079854930cfa76ce642094e194e34c9f7d5062835fe22a1b7a4c95

Observation 9a3a8fa4-8900-4225-b172-d861330856b8 · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:09.887019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:09.887019Z digest=sha256:e99009516c55670129eb2f274cc80c123e1c854a977e1cb0cd1333557b77b12d

Observation 5c0b4e36-5d23-4610-9c37-509124977563 · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:09.962310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:09.962310Z digest=sha256:5050c84b081ffb2091d028e416bf485eb56cb8c3ec9bee7eba9730d9c5b8725f

Observation ff5851c6-f5e3-419e-ab71-0f9e1465406c · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.094685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.094685Z digest=sha256:f6874c3e16d9276499b1b5494d475f6c836dfa9f8f7110fed44d1cd7ce097577

Observation 8c59eb3d-5533-4112-86a3-fdde9c9961b5 · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

ICR-RL: Deep Reinforcement Learning via In-Context Regression RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.146610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.146610Z digest=sha256:0620abeede23e00c893267b136edbf603a16786d5183d4990467dcfd61a1fdfc

Observation fcc74421-141d-4bda-9901-5c42464dc153 · outbound

This paper cites What Can Transformers Learn In-Context? A Case Study of Simple Function Classes.

ICR-RL: Deep Reinforcement Learning via In-Context Regression What Can Transformers Learn In-Context? A Case Study of Simple Function Classes

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.272420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.272420Z digest=sha256:336a70084a24f81cd658aa547702005f17f69fb8c9a133ddb672e380b50e9477

Observation e935cbf3-ccf0-480d-8a5d-001ad5269f55 · outbound

This paper cites H.; Tirumala, D.; Humplik, J.; Wulfmeier, M.; Tunyasuvunakool, S.; Siegel, N.

ICR-RL: Deep Reinforcement Learning via In-Context Regression H.; Tirumala, D.; Humplik, J.; Wulfmeier, M.; Tunyasuvunakool, S.; Siegel, N

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.383192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.383192Z digest=sha256:0e30e96d051f97ea4fa3c896ad668ee33696acf42deff5f523390649f9761c56

Observation 036d50af-e912-462c-97af-2daa2247084c · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.487706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.487706Z digest=sha256:c85a6c2d4db199eeb5db2f62ebf1cad41aa2ce768a3204a860f79a426dc74a83

Observation 2aee241a-d8bd-4389-b519-d7cca44779d8 · outbound

This paper cites B.; Müller, S.; Salinas, D.; and Hutter, F.

ICR-RL: Deep Reinforcement Learning via In-Context Regression B.; Müller, S.; Salinas, D.; and Hutter, F

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.645806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.645806Z digest=sha256:075df4e8be34e0d0266ed1267700bb48404bfbc606788c2389116c4224376d95

Observation 1a3020d1-61f3-480d-9681-8420f3623cb5 · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.769855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.769855Z digest=sha256:8a5d4ce9050435594fdcd995101072dea8027ef159d72c96009f70a1426219d9

Observation 75d6ce3f-96e3-4551-be7f-a37e1bbe9aec · outbound

This paper cites Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.872577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.872577Z digest=sha256:093115a13a20c2b7f9e4178cf4b889549145a6eca23a15020b8608f7f59a59a1

Observation f504aefd-9b12-4db4-96d7-18b4bb383dbc · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:10.960557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:10.960557Z digest=sha256:9bc9bd51bd0192651f4f2284408e8ecf5ce861cfa1981cc1209c616be7896a78

Observation d7e3c899-1bd2-42a4-8fa3-3b690bf1b4f7 · outbound

This paper cites A.; de Lope, J.; and Maravall, D.

ICR-RL: Deep Reinforcement Learning via In-Context Regression A.; de Lope, J.; and Maravall, D

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.078534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.078534Z digest=sha256:4e38a0071817825b3311810d37f1a7d0a6ce1aa10ffd5cabd29c9c807c59e002

Observation 4c3b3ae0-e54c-465e-ba78-d0ac495f815d · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Playing Atari with Deep Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.176652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.176652Z digest=sha256:5ee51d77d1f6e11b329be6cc95d51fcf18a487240b9f8d90c0ea15796bdfae8d

Observation ce82cd46-4887-4741-a499-1461d6c8645d · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.290929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.290929Z digest=sha256:ff9d940793338e220a87a3f02cd6fca060178022199f5f38651e6e57e4ede942

Observation b84a025a-b9c7-4146-9c52-4f1a4791a9bc · outbound

This paper cites Trust Region Policy Optimization.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Trust Region Policy Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.382075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.382075Z digest=sha256:d0e8635b9655262135b37a3466da114c51122941e3225b002249e462910ea5c8

Observation 5999da98-da5e-4025-9062-8f9b6fa7b76a · outbound

This paper cites Proximal Policy Optimization Algorithms.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Proximal Policy Optimization Algorithms

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.482356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.482356Z digest=sha256:6be7fefca306e59523b17e12db76f10d2a407cc95a7fb64d3be0115ef812b37a

Observation d3e64ac5-d90a-4b29-af37-801afbe7d3eb · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.590521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.590521Z digest=sha256:7683e431ace674a8a1d194f359433e759024fbd60e7a0d6988d11a4f24c612ee

Observation 102d1727-d5fc-4a89-b6e4-64a177f6f681 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.697175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.697175Z digest=sha256:c855221b3a23c22454988eced98657c2f97972b1f0dacef626d3f72c5b908c0c

Observation 85a0693c-5ac5-405e-a72b-87eff4d1fd22 · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.854655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.854655Z digest=sha256:d22f42c9661cf7c92f9469440868fdb96dee22e82ab63b71fc3e69c41f9835c7

Observation bfbf3361-ed07-4e7f-bfa7-b3181f31fc3d · outbound

This paper cites an unresolved cited work.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:11.964993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:11.964993Z digest=sha256:e2df6702b07967f9aceaa92db72b849f4ce65a4a84dd3d649dcb696295081046

Observation d9a598b0-dd3f-4ee7-9845-b6aa3d14f549 · outbound

This paper cites Fewer May Be Better: Enhancing Offline Reinforcement Learning with Reduced Dataset.

ICR-RL: Deep Reinforcement Learning via In-Context Regression Fewer May Be Better: Enhancing Offline Reinforcement Learning with Reduced Dataset

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T16:52:12.064508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:52:12.064508Z digest=sha256:a9231e952feea64705d04a9adabc5692e410a81dacfb5e903d176474e6ab64a5

Pith citing papers

Observation 94f0cc0e-ff34-4a9a-87ef-f0ea401b4ae2 · inbound

TabPFN-2.5: Advancing the State of the Art in Tabular Foundation Models cites this paper.

TabPFN-2.5: Advancing the State of the Art in Tabular Foundation Models ICR-RL: Deep Reinforcement Learning via In-Context Regression

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:19:40.488978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T04:14:44.792670Z digest=sha256:2f9556295ebbed8f31db2065b062644a8a8a2de9e0c6ef1550bf5353d1a667a8

Observation 01def5c4-06d9-4d4a-adea-2d77ee95aba9 · inbound

TabPFN-3: Technical Report cites this paper.

TabPFN-3: Technical Report ICR-RL: Deep Reinforcement Learning via In-Context Regression

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:19:40.488978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T06:05:16.199413Z digest=sha256:227373cdc8f246189c11bb296c47957c45edc6a8aba5e56803388d51b8dab274

Observation c437cc73-60fe-42b7-9068-49ab338a071a · inbound

TabPFN-3: Technical Report cites this paper.

TabPFN-3: Technical Report ICR-RL: Deep Reinforcement Learning via In-Context Regression

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:19:40.488978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T21:37:08.627791Z digest=sha256:affb67317112ecddc5e9ba9391bb14e6c628a5b1e6036a4840d9b6479d28143f

Observation e2d9e758-fa1e-46a2-aeeb-75656c78bd2c · inbound

TabQL: In-Context Q-Learning with Tabular Foundation Models cites this paper.

TabQL: In-Context Q-Learning with Tabular Foundation Models ICR-RL: Deep Reinforcement Learning via In-Context Regression

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:19:40.488978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T12:34:58.734670Z digest=sha256:f272a08472c11491b44b767ebded499992c6643951f163593236ae2c8039c413

Observation c6d5a20c-4444-4e07-93bd-7c6d90b29e72 · inbound

Reinforcement Learning Foundation Models Should Already Be A Thing cites this paper.

Reinforcement Learning Foundation Models Should Already Be A Thing ICR-RL: Deep Reinforcement Learning via In-Context Regression

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-07T02:19:40.488978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T21:50:00.974590Z digest=sha256:bcd54ecf7749b58e37f5be68d04c5747d521c5eb54038dedd0aa624538362b2d

Observation ad01a961-5430-42a4-bf63-5b3b982c27e5 · inbound

FlexTab: A Flexible Encoder-Decoder Architecture for In-Context Learning Across Diverse Tabular Tasks cites this paper.

FlexTab: A Flexible Encoder-Decoder Architecture for In-Context Learning Across Diverse Tabular Tasks ICR-RL: Deep Reinforcement Learning via In-Context Regression

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:19:40.488978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T07:22:32.638556Z digest=sha256:a18a47acfce336b1c05433c08a72ec88190cf2a184fdad4a53abbfc7ffd3a9bf

Observation 85578b5a-556b-4c2d-89af-de405305faab · inbound

FlexTab: A Flexible Encoder-Decoder Architecture for In-Context Learning Across Diverse Tabular Tasks cites this paper.

FlexTab: A Flexible Encoder-Decoder Architecture for In-Context Learning Across Diverse Tabular Tasks ICR-RL: Deep Reinforcement Learning via In-Context Regression

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:19:40.488978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-01T07:13:43.014623Z digest=sha256:fa3c523b38c80dfaf364c2f8ae9ff5bfdf1550041df452ed17af2f166dc976ab

Observation de65bdd6-f56a-47d1-a569-0b71999e5799 · inbound

Beyond IID: How General Are Tabular Foundation Models, Really? cites this paper.

Beyond IID: How General Are Tabular Foundation Models, Really? ICR-RL: Deep Reinforcement Learning via In-Context Regression

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-07T02:19:40.488978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T06:59:14.626274Z digest=sha256:8e0616cbfd98f28793bdbdada40a4653dc4acfc32f9be8dbd4a92ba5e3024ed5