Pith. sign in

Paper Citation Record · LEDGER

AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2308.03526.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.03526 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:58:35.028666Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8322f049-573f-4109-9dad-38c2c5e5b58b · inbound

Perspectives for Direct Interpretability in Multi-Agent Deep Reinforcement Learning cites this paper.

Perspectives for Direct Interpretability in Multi-Agent Deep Reinforcement Learning AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-09T17:58:35.028666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:58:35.028666Z digest=sha256:10e3031787f64d247e040ca1e8e4bc1b9fc9e106531cea2514a6ce9d92239dbc

Observation 9d314d07-1972-4440-a8c6-e9cf2a210f19 · inbound

SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data cites this paper.

SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:08.440100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T14:54:30.895137Z digest=sha256:be31c0afce183e888123dca258e816ecfef1eb6620ec0a56018e86f5359290ff

Observation b0865a38-1949-4e93-aabb-47cfd16615fb · inbound

SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data cites this paper.

SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:19:56.766956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T09:15:11.280343Z digest=sha256:cd55f727113fe83f685bbd56f352df7f0e8d9f6dfd140c8e6a09deb63c753712

Observation 4fad827a-6139-446a-b0a3-87fefec46150 · inbound

Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought cites this paper.

Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Reference 177

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:31:00.555793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T01:15:41.980346Z digest=sha256:864acc1ce23dc815a6a41e04c623ffc4a8cf8222ed195d628eea5a26d4e38476

Observation 44f484f7-a279-4bf5-b5da-ac48558db042 · inbound

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning cites this paper.

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Reference 206

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:35:58.645268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T01:13:34.836246Z digest=sha256:e27cef4e3a063400abcce8fb4f77b00825d2b69e6b49da7b2e9607032e029e4a

Observation dd2b2219-7c90-4fa4-bc23-6b0ed08787cc · inbound

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning cites this paper.

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Reference 209

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T23:29:13.131861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T23:26:35.072127Z digest=sha256:be2a338c4aaeef688eafcef0d68763de64d4d67c00d7101233a464f9bbf1d33a

Observation 19952726-9633-4991-b47b-7ae346c4c4a5 · inbound

Parametric Open Source Games cites this paper.

Parametric Open Source Games AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T02:08:54.915358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T01:59:35.625359Z digest=sha256:93afb9ac8f938230a4472600cb214cdb6e10e5c1c3680d6710c9b86204178c3a

Observation 0b420bd4-14bd-49e2-b0eb-2fe84b89c0f9 · inbound

Play Like Champions: Counterfactual Feedback Generation in Latent Space cites this paper.

Play Like Champions: Counterfactual Feedback Generation in Latent Space AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T19:47:18.859076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T19:41:39.901829Z digest=sha256:b486c687cdd98e9b9ee9ad49f22448bc3f784e02e99d6588ad1eff774847ff2f