Pith. sign in

Paper Citation Record · LEDGER

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States

As of 14 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2607.18589.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.18589 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T15:02:10.268300Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7175be2c-5a82-4c74-83cd-8182aa26c547 · outbound

This paper cites Relational inductive biases, deep learning, and graph networks.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Relational inductive biases, deep learning, and graph networks

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.135079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.135079Z digest=sha256:3ca74d6edca4e5ca75565b05391527ec1e5bf99d5778517babb28b89174050b6

Observation 21c47137-978d-45ae-b1c3-b57e1e39931c · outbound

This paper cites Interpreting emergent planning in model-free reinforcement learning.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Interpreting emergent planning in model-free reinforcement learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.185967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.185967Z digest=sha256:89d973094db62111458f1194c93112191b1828b0a28ca022a969165564075d1f

Observation ad7b2732-9443-42a4-89f4-fa3688995c6d · outbound

This paper cites Cohen and Howard Eichenbaum.Memory, Amnesia, and the Hippocampal System.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Cohen and Howard Eichenbaum.Memory, Amnesia, and the Hippocampal System

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.266292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.266292Z digest=sha256:855958bcc586010441f81d2810e3239873e49daa745d50684e83943df43b75b0

Observation 88b8a5a6-bec6-4c7b-8845-698c480b8931 · outbound

This paper cites Emergent Response Planning in LLMs.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Emergent Response Planning in LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.349920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.349920Z digest=sha256:417cf709071c5b1f06b2c3bfbfebf9b94db05a90a4451d5fb624c08ad0644a25

Observation eda69215-df16-4d0e-b26e-cb5c3d912bac · outbound

This paper cites On the integration of space, time, and memory.Neuron, 95(5):1007–1018,.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States On the integration of space, time, and memory.Neuron, 95(5):1007–1018,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.430836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.430836Z digest=sha256:13d34483042d03ba7b609580ad7046dab47f8d984968a98686252bf93de7edcb

Observation 139b592d-c656-4a31-a783-721bdb8889f1 · outbound

This paper cites IM- PALA: Scalable distributed deep-RL with importance weighted actor-learner architectures.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States IM- PALA: Scalable distributed deep-RL with importance weighted actor-learner architectures

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.556389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.556389Z digest=sha256:5a3e1758e3f571c93a9a7d8ec2bf68533283280a7b44efe4096055c607e9f706

Observation 138a9800-8a70-47c1-85af-933b6add9294 · outbound

This paper cites Hybrid computing using a neural network with dynamic external memory.Nature, 538(7626): 471–476, 2016.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Hybrid computing using a neural network with dynamic external memory.Nature, 538(7626): 471–476, 2016

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.641623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.641623Z digest=sha256:bdccdf677e12990d3fc9f6b9b71cb996d034bee372c51e1bc4f32b73bbe1ca92

Observation f0211f95-3e38-422b-b997-ee3247d0b762 · outbound

This paper cites Highway and Residual Networks learn Unrolled Iterative Estimation.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Highway and Residual Networks learn Unrolled Iterative Estimation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.707423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.707423Z digest=sha256:a65e29a96f2903cb0f6d11b038f1fe68ad59c7186f90616dc04ebd42404d733d

Observation 65152fbd-010f-41d6-a396-a23689fac086 · outbound

This paper cites An investigation of model-free planning.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States An investigation of model-free planning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.755937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.755937Z digest=sha256:f535da31f8d764e2cf3f45e2a03b6d0d11e206b482b23beb7cab017ea3422356

Observation 8a999ab7-8275-4085-83e2-8241eccc7856 · outbound

This paper cites World Models.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States World Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.858752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.858752Z digest=sha256:0c9fd28fd28324ac46b7fd8187b0436b7a28cef800b50874b1efd13e43159632

Observation 2e6ceb7b-c584-4c50-bc92-d9583185caef · outbound

This paper cites Learning latent dynamics for planning from pixels.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Learning latent dynamics for planning from pixels

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.942708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.942708Z digest=sha256:088fddc02ad4cc4fe21917b835000c434d32eafe1e08cbb90f5aeecd4ef4b50b

Observation 85994764-c10a-4269-9d1a-861554acd880 · outbound

This paper cites Mastering Diverse Domains through World Models.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Mastering Diverse Domains through World Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.014108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.014108Z digest=sha256:cb5ff3c72aa9327ee63c50470d6a01578028871fede9fe030ee2cf971ad29031

Observation 18781938-fef4-4540-bbc8-5594c53a825e · outbound

This paper cites Evidence of Learned Look-Ahead in a Chess-Playing Neural Network.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Evidence of Learned Look-Ahead in a Chess-Playing Neural Network

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.102796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.102796Z digest=sha256:7567bc7e5c35663436a8d2107580bd1be4e0843f0eacb9d9f2a585344142e282

Observation dc309a59-8dd9-49f8-bf1f-fdb5ddcc6e27 · outbound

This paper cites An empirical exploration of recurrent network architectures.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States An empirical exploration of recurrent network architectures

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.196707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.196707Z digest=sha256:fe5c2ccbf77c12526506c2f839689958ccfb0bd12fe12e1be4ae93409c2e5593

Observation a313df4e-2275-461a-ae53-5a001b719130 · outbound

This paper cites Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.396594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.396594Z digest=sha256:ff0d62fe3d5f66d6efcf95c166c9664fa68d80383b3dee3d9291554cc855c6f1

Observation 9e2f8787-29d7-45c7-a28e-2339186de769 · outbound

This paper cites Neural rela- tional inference for interacting systems.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Neural rela- tional inference for interacting systems

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.466587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.466587Z digest=sha256:7fd56c16e560699253db22bf176b71d0dcc5f972900a2eb9401e1ed140356452

Observation 6e5ce3d9-56da-43db-aee4-ec5054d8097e · outbound

This paper cites Actor-critic algorithms.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Actor-critic algorithms

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.626712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.626712Z digest=sha256:b2e15dba51df29a8d02d79c3aaf071733ae2b33964cda12136e90c64936ea92a

Observation c44de500-899b-4947-9050-536b1ec00fa1 · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 18

Resolution
verified exact
doi, observed 2026-08-01T15:03:26.606056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-01T15:02:08.690441Z digest=sha256:8cd6ff0041eea27ac04bf9eb400ea65f0285062cdffe285b3be135c6ac817cb2

Observation 2d2ae8f8-48b3-4f16-976f-e7f402591358 · outbound

This paper cites Lambert, Brandon Amos, Omry Yadan, and Roberto Calandra.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Lambert, Brandon Amos, Omry Yadan, and Roberto Calandra

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.746027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.746027Z digest=sha256:7a000c77596c2a25346ee006cac1e8f914b674afda06cf24f72386624b39cf60

Observation 286f2150-bd29-4b40-b0b9-82e858668d62 · outbound

This paper cites Hopkins, David Bau, Fernanda Viégas, Hanspeter Pfister, and Martin Wattenberg.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Hopkins, David Bau, Fernanda Viégas, Hanspeter Pfister, and Martin Wattenberg

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.922477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.922477Z digest=sha256:10a5c6676749881a03812fdcb4bdbfa7d8b9e54b6b5f2359eda54c3b209110b8

Observation 569930b3-89d8-4396-bee1-3f84ae1ac784 · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.006973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.006973Z digest=sha256:f1672100d9be945121218cdc666b263abdeb9c2581132775da4cac4d5fd1b8d7

Observation ce5a85a2-29a5-436e-b039-69de08bd295f · outbound

This paper cites Puterman.Markov Decision Processes: Discrete Stochastic Dynamic Programming.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Puterman.Markov Decision Processes: Discrete Stochastic Dynamic Programming

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.076782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.076782Z digest=sha256:f2a3aa78631c61cf104b80fbe1c380668091263a28b029f642feda3d41826674

Observation cd3c98ed-e35a-43e1-b6a5-f08151da5aae · outbound

This paper cites Hopfield Networks is All You Need.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Hopfield Networks is All You Need

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.136169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.136169Z digest=sha256:b59769608db92c7a0a9afa3f606b85cde3701eed650aec11a3a53374b5303544

Observation 9edd9547-3fac-45e2-be9d-574496b522ba · outbound

This paper cites Graph networks as learnable physics engines for inference and control.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Graph networks as learnable physics engines for inference and control

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.218368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.218368Z digest=sha256:93e7cadd24b2a77ae8f9dfc991a37c13787f300b9d53b0bcdeaf332b08dfbed4

Observation ac27f4c3-5c54-43a9-801a-857e34f1ff64 · outbound

This paper cites Mastering Atari, Go, chess and shogi by planning with a learned model.Nature, 588:604–609, 2020.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Mastering Atari, Go, chess and shogi by planning with a learned model.Nature, 588:604–609, 2020

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.287298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.287298Z digest=sha256:77e96eff94ac6fb5bb9c62080775e70778410c97e83bc9b9a62ee196650fda92

Observation 6b0b7c34-eba7-403b-9b54-8390918c4124 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.336271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.336271Z digest=sha256:bedfc22be9c1ae0c554714a1c6e9fbf0ff8b1e9317bc1006a8accd83ce775270

Observation bf603cd5-7c34-4adb-af4c-6f0ea007a2c7 · outbound

This paper cites Convolutional LSTM network: A machine learning approach for pre- cipitation nowcasting.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Convolutional LSTM network: A machine learning approach for pre- cipitation nowcasting

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.394919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.394919Z digest=sha256:0a4f84b530177a5fd24649769c7db861cbb7ce47029bc9f0ff6cb9e7dc0e8d1c

Observation 9ebfd173-a3ee-426f-a151-d7e6ecc25506 · outbound

This paper cites Mastering the game of Go without human knowledge.Nature, 550(7676):354–359,.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Mastering the game of Go without human knowledge.Nature, 550(7676):354–359,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.474318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.474318Z digest=sha256:3fd4714be0b1642b0606b18b87f1c3d77fd0575c4a4c56d9c741e347ba02f479

Observation 153f4473-7cae-4ddd-a47c-295c64da6971 · outbound

This paper cites Stachenfeld, Matthew M.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Stachenfeld, Matthew M

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.615278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.615278Z digest=sha256:b616cecbe8a1d42871dedc15aa0a039918753f3f12f122eba7ea2a21a31a95e2

Observation 7426db5f-acd1-4232-b5e2-2e942bee1e69 · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.683089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.683089Z digest=sha256:0fa0d82f2bfc1e7be2094b6f963ab34ceb000d5211aff0cf2c521338e56b46b0

Observation 0c6ce0aa-c901-4507-8cae-3d0ea4318df8 · outbound

This paper cites Sutton and Andrew G.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Sutton and Andrew G

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.766727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.766727Z digest=sha256:d7812b26e661394815b5717520079bdfd6c7fd8bc942463dd321e90a56a19b2b

Observation 20b2035b-40d6-4146-b760-48eb02093e87 · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.547129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.547129Z digest=sha256:a717220381801e698ba06af97b2ef2ba0cf13ee919a4205194a79e0f1b128752

Observation 7772a4fc-a757-49ac-bcc4-5dbddaaccfeb · outbound

This paper cites Planning in a recurrent neural network that plays Sokoban,.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Planning in a recurrent neural network that plays Sokoban,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.878998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.878998Z digest=sha256:5b4150497f2652272677c3066b632f9fe0e7dab38593ed75e0deacf4ef2a1896

Observation bd602923-92c8-4cd6-92f7-a79a6cd34349 · outbound

This paper cites Path channels and plan extension kernels: A mechanistic description of planning in a Sokoban RNN.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Path channels and plan extension kernels: A mechanistic description of planning in a Sokoban RNN

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:10.036252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:10.036252Z digest=sha256:6a14e82173cb1c7c5624a7baee1b611d89c62a28bc84312584d35c3fe3fd21f6

Observation db1077a8-e2d7-49f4-a838-744d92e3954c · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:10.088740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:10.088740Z digest=sha256:1c65422f6f4713f0a08ca7dda86ed254837f15bac06018c6106c96ef55403fbc

Observation 95a08ea6-0310-4a8f-9221-8dfeafe319c4 · outbound

This paper cites Sutton, David McAllester, Satinder Singh, and Yishay Mansour.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Sutton, David McAllester, Satinder Singh, and Yishay Mansour

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.824297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.824297Z digest=sha256:ca34b96be67f5c42a532d03b6ac92b8dc6156e82162492e2d7887646d1901105

Observation b25fdc25-e292-4093-b69e-82f660107eeb · outbound

This paper cites Relational Deep Reinforcement Learning.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Relational Deep Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:10.268300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:10.268300Z digest=sha256:3ddd9759260d38a30aeeb6566d8ac97bccf7657d8b050b54c89ed9134b2d6288

Observation 501173ea-6396-4676-8443-578983c9e80c · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:10.199082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:10.199082Z digest=sha256:c52d3c457e8faaff8e125ab9305502962efabe18fef6f2310a2a7e5d08a27ba3

Observation 5eb28650-6644-43d2-b345-51c3ab44724b · outbound

This paper cites URL https://proceedings.mlr.press/v120/lambert20a.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States URL https://proceedings.mlr.press/v120/lambert20a

Reference 770

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.857631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.857631Z digest=sha256:467c83350f83cf41f1c2826cf5c18f43113f8a3b76a45e510f8c857e5b04c1f1

Observation 68313b1b-94db-4319-b0d4-28ef8993c0eb · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 1948

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:10.141206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:10.141206Z digest=sha256:6452b54a0a357162fa3a0dc9326a37c0db61285dbf6ce9d9613b58901ff19c16

Observation e1600f04-ab11-4765-8a3a-72b140fb7e4c · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:08.291294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:08.291294Z digest=sha256:ca902a39d541200e94a459485338a898b5ee8f730875ef330021857d02181632

Observation 5950a635-ce08-4ff7-b107-71c8205766f9 · outbound

This paper cites an unresolved cited work.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Unresolved cited work

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:07.489269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:07.489269Z digest=sha256:2dca6663d3099c5bc6173cdf13bdf809f236369e0f58fb6da856290b10269fe8

Observation a902f4d8-9798-4d43-8998-298549eba040 · outbound

This paper cites Planning in a recurrent neural network that plays Sokoban.

Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States Planning in a recurrent neural network that plays Sokoban

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T15:02:09.968163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:02:09.968163Z digest=sha256:b2f715349f7a3e584e4707813861bf0a765431156545f428f76ce643fc661d9b

Pith citing papers

No inbound Pith citation observations are available.