Pith. sign in

Paper Citation Record · LEDGER

Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2306.09884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.09884 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:53:01.923105Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9099ace3-0b92-4b5a-88e7-8cdc8116c452 · inbound

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution cites this paper.

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:12:31.315387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-16T08:12:30.984870Z digest=sha256:6b34f1da3e605c8513ba6c2f234efdae7fc6f4f5b6216b591bbb5804a24b13ae

Observation c2b8b9d1-3d81-43d4-8c9f-0bc4ced60bf3 · inbound

Gymnasium: A Standard Interface for Reinforcement Learning Environments cites this paper.

Gymnasium: A Standard Interface for Reinforcement Learning Environments Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:29:49.581158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T17:29:49.186565Z digest=sha256:653aae22f15cdc4a432489729f7ccf1906270eb67078a6ff6d7ad2d1f12016de

Observation 211df60d-093c-47d9-8259-a334a3f6796d · inbound

Chargax: A JAX Accelerated EV Charging Simulator cites this paper.

Chargax: A JAX Accelerated EV Charging Simulator Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:53:01.923105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:53:01.923105Z digest=sha256:6d6d5b92e8f7d43df5e003aa6f7d8f02db31f97b9d3e71194cd3653d4d690f82

Observation f3969fdb-cf04-432e-909d-def8a1e27fb9 · inbound

CleanQRL: Lightweight Single-file Implementations of Quantum Reinforcement Learning Algorithms cites this paper.

CleanQRL: Lightweight Single-file Implementations of Quantum Reinforcement Learning Algorithms Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:41:01.885621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:41:01.885621Z digest=sha256:42c18c52b0fb378c5bcf51f89139d380d77975b39d8829ae29202071d2eb5192

Observation 148d73d9-d505-4656-9469-ff55e32f826b · inbound

Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks cites this paper.

Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T11:03:22.084295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:03:22.084295Z digest=sha256:ebd0d766940d9992a7a8a2bcc1e1d684ff2271c1bbf4f62f3f4b0bdde080c518

Observation 4f501582-ae8c-409c-abc9-2efc2ca501d9 · inbound

Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX cites this paper.

Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAX Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-04T12:54:39.273506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:54:39.273506Z digest=sha256:d0f1f3751de58b106d0b30e8f1d58c7c09d6d44adf8d76602ebeab9906c44c93

Observation 5888563a-2747-4fc3-b190-2c4515fb867d · inbound

Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments cites this paper.

Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T18:45:20.286849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T18:45:20.286849Z digest=sha256:417c77f00651a3eb6ce29b8070930878a2e31f20e08e592d48dd2a6faa7d02d0

Observation 5055a177-3d43-467d-a838-dad7540087ec · inbound

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation cites this paper.

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:12:55.256675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T20:12:07.920183Z digest=sha256:6c889bcc05abadf06babcf45292e9f8cf69d766616d116a27a5cdcec04761edb

Observation f54aa429-faff-4f7e-8b84-7923555950ee · inbound

Finding the Time to Think: Learning Planning Budgets in Real-Time RL cites this paper.

Finding the Time to Think: Learning Planning Budgets in Real-Time RL Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:29:51.943255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T05:08:19.504454Z digest=sha256:fbea5eee99f932506231caa2c41a80ac0f1967ae7b7b7e37af22297b77777652

Observation c26d9e7c-5aa4-4022-b140-395749ef39fb · inbound

Finding the Time to Think: Learning Planning Budgets in Real-Time RL cites this paper.

Finding the Time to Think: Learning Planning Budgets in Real-Time RL Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T09:34:34.919692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T09:26:31.944405Z digest=sha256:36b45c41ceb32de0eea8f9b0946e61b8b92f4f2c3957be2481f6bea87e811f45

Observation 7c1fb0cb-613b-40d1-9b18-05945d554523 · inbound

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy cites this paper.

Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:25:48.447624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T01:28:44.074862Z digest=sha256:87a30f9cb947766908488221076687810a91814630fe5866c194868b17272352

Observation 9de7e73d-3082-4698-90f8-6dd10cb56ec4 · inbound

PCBWorld: A Benchmark Environment for Engine-Grounded PCB Design Automation cites this paper.

PCBWorld: A Benchmark Environment for Engine-Grounded PCB Design Automation Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T21:15:39.082889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-08T21:09:25.719945Z digest=sha256:31c8e689e2a5a161e37de03a88ab06084c442622b3272136be5253450f180747