Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T19:37:29.933624Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 1 inbound Pith citation observation for arXiv:2506.15446.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T19:37:29.933624Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-19T20:37:36.030165Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T20:37:45.261549Z
93 of 93 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 47aec00e-6502-43af-8802-3139c8a455aa · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Deep reinforcement learning at the edge of the statistical precipice
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 399ac21f-4282-415f-829e-b21d30473c3c · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Hindsight experience replay
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c5e6fef-941e-41fc-9407-1dfda27694b3 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Rudder: Return decomposition for delayed rewards
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e4fac02-d175-476a-ab54-53c6b373f85b · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Optimal control of markov processes with incomplete state information
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45d2ae62-49ed-481a-a99a-32d37107d400 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Layer Normalization
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5958cc51-6da9-45ea-a146-5e56ed537992 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Reinforcement learning with long short-term memory
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d7598e5-dc2f-4475-b4e7-6791b449ef44 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Augmented world models facilitate zero-shot dynamics generalization from a single offline environment
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df418c7a-74d2-42fc-90e0-d4f9e9370ef7 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Successor features for transfer in reinforcement learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 807bbc2b-591e-476e-91d1-f08bbd0a2476 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f7f19c0-edcf-4ebc-8739-6b0b931a0ca5 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Universal Successor Features Approximators
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f20e5ef-211b-4ca3-8052-946d5f61aa52 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Language models are few-shot learners
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a51e508-d1cb-472c-b9eb-e962433713ee · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Acting optimally in partially observable stochastic domains
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18a9c2a8-f72a-47df-b564-5b3464cbbf7d · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cf032f4-5977-4097-86b2-29188d804dcc · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Quantifying generalization in reinforcement learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f1e26c-54b4-4b34-a55c-77caca9acb67 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Flashattention: Fast and memory-efficient exact attention with io-awareness
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 018f2ccd-ea2b-45d1-9a4a-544035a6d135 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Improving generalization for temporal difference learning: The successor representation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 917a2b4b-1c8a-4327-8789-f3a3e511695e · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Facing off world model backbones: Rnns, transformers, and s4
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fd46fdc-f1da-44db-a719-a2b49aabea3c · outbound
Zero-Shot Reinforcement Learning Under Partial Observability An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3a71902-fbe1-4608-abd0-7f0320dfcb80 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Finding structure in time
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b34e4d8-bf29-442c-80a1-6bf691cf607f · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Contrastive learning as goal-conditioned reinforcement learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73fc2854-b51f-4b46-831d-6332dc046698 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Generalization and Regularization in DQN
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 619b6e4f-3fe6-4971-9ecb-e1c500ca43f0 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Hyperbolic Discounting and Learning over Multiple Horizons
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fde3a438-0752-4baf-9aa8-c1f66459ba30 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability A minimalist approach to offline reinforcement learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f4615da-27d2-4969-8652-9ce997cf7745 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Off-policy deep reinforcement learning without exploration
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 868ea262-eba7-4a19-a0e2-33fb4a5319d3 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Amago: Scalable in-context reinforcement learning for adaptive agents
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 19a6fd89-7eff-4f64-8b40-6bb77964c9f3 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Amago-2: Breaking the multi-task barrier in meta-reinforcement learning with transformers
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 35a9efee-c19a-4029-87b3-3a21b3c93456 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Efficiently Modeling Long Sequences with Structured State Spaces
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a2375f2-8882-4cbb-b369-39a29f030b87 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability On the parameterization and initialization of diagonal state space models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e96b70a-c51c-4dac-b4a4-430d444f4b44 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability World Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cec1eb11-adb9-4954-b9fa-cd406450e963 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Dream to Control: Learning Behaviors by Latent Imagination
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4494390-9afd-4372-8eb9-a8cce066cecb · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Learning latent dynamics for planning from pixels
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e7f9fcb4-037f-4854-97dd-5005adc42819 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Mastering Atari with Discrete World Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 526d1092-4fa5-44ec-9586-d9fc7d2e24ab · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Mastering Diverse Domains through World Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6803eb5-81af-46ef-a626-26547d7140fe · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Contextual markov decision processes, 2015
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 958323c3-a7a9-4fc2-8b65-0c48db1a8552 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Array programming with numpy
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e5cdc20-671d-4396-9ac3-3f6f6fe3e72c · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Deep recurrent q-learning for partially observable mdps
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dffbaf06-fc6c-47b2-8540-f7d978c67333 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Memory-based control with recurrent neural networks
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20cc5211-af52-4287-a5dc-add1e0fdd9f5 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Long short-term memory
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b4a9fb5e-b0bf-410b-beb2-096d3fc161c5 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Matplotlib: A 2d graphics environment
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d63be35e-9830-451e-a8fa-8f4d1fd35ff8 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 81c701c8-fac6-4b18-a580-8f718be115a5 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Monotonic robust policy optimization with model discrepancy
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fc5dd288-84eb-4f1f-b2be-b8867e5896dc · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Planning and acting in partially observable stochastic domains
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6177a93-3c17-4a27-bb5f-5158d13e8ce0 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Morel: Model-based offline reinforcement learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e2565ba5-0f27-43d9-bbc7-5caaa321bd93 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Adam: A Method for Stochastic Optimization
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 316fda0c-1f91-4d51-92dd-1a97e660c5b5 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Actor-critic algorithms
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0aab0761-f32c-49a6-bd90-a5ccccec3cfa · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Stabilizing off-policy q-learning via bootstrapping error reduction
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation da367ce7-9459-4610-a7ca-ceacf3493761 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Stabilizing off-policy q-learning via bootstrapping error reduction
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 5cfd831a-980d-4ee2-9d6a-d92b81f6c87e · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Conservative Q-Learning for Offline Reinforcement Learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0067d450-c1b3-4c8f-b663-c6fc30b9e3a1 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Batch reinforcement learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 5e06f878-525f-472e-adad-f795481a4fb4 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Context-aware dynamics model for generalization in model-based reinforcement learning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e4664fa5-2a55-44d1-ae18-dab3ef0786dc · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9df9c00-9c55-42e5-beb8-3fba3fee2770 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Off-Policy Policy Gradient with State Distribution Correction
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1e78f30-178c-4e9e-ba34-f1a035390933 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Structured state space models for in-context reinforcement learning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fa96089b-dc21-49ce-8d75-5eb1e52daf8a · outbound
Zero-Shot Reinforcement Learning Under Partial Observability How Far I'll Go: Offline Goal-Conditioned Reinforcement Learning via $f$-Advantage Regression
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08750780-abcb-422c-beff-5bb4857cb255 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Robust Reinforcement Learning for Continuous Control with Model Misspecification
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6325162-4e18-4290-9751-450323571248 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability pandas: a foundational python library for data analysis and statistics
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3cd82194-0b93-42c2-bb5a-11296011faa1 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Memory-based deep reinforcement learning for pomdps
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4acd8763-474e-426d-b984-32f648861b0a · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Steps toward artificial intelligence
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a60f4992-812f-4624-9136-81254a918629 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Human-level control through deep reinforcement learning
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 323129d7-14c5-4449-bc98-43eccd4e5638 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability POPGym: Benchmarking Partially Observable Reinforcement Learning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49a2e96b-4457-41b0-942b-d83da35f6a7f · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Robust reinforcement learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c7d8530c-a575-4793-9688-dcc3157acc13 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Recurrent Model-Free RL Can Be a Strong Baseline for Many POMDPs
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b262d7c-34f8-46c8-a5b2-fae1dd7b0b27 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Robust control of markov decision processes with uncertain transition matrices
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68186d0b-af03-4957-908b-1541299b95e9 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Assessing Generalization in Deep Reinforcement Learning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60a194a3-7aac-4a71-81bc-1a8db8f6cf3d · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Stabilizing transformers for reinforcement learning
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e71ba0c2-2617-423e-a59a-b25b1ca40278 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Hiql: Offline goal-conditioned rl with latent states as actions
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8e683216-8398-4ecf-b387-9b5f5c20c041 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Foundation policies with hilbert representations
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 511fde79-6680-448b-840a-3c87033fd7c6 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Automatic differentiation in pytorch
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5e4fd5d-7023-4e3b-bee2-b7e4f92dca37 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Fast imitation via behavior foundation models
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 55abfbbe-c184-4cb9-8ee9-edf199a201e8 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Automatic Data Augmentation for Generalization in Deep Reinforcement Learning
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2227ed27-a674-4b16-a5a0-b4056ed2ff5c · outbound
Zero-Shot Reinforcement Learning Under Partial Observability EPOpt: Learning Robust Neural Network Policies Using Model Ensembles
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2c8a653-fcf0-4190-8bda-e8640f52f4f0 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Synthetic Returns for Long-Term Credit Assignment
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09b5a9d9-404b-451b-8af9-e16407be0569 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Reward-Free Curricula for Training Robust World Models
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81b78502-8f0f-4f5a-8a45-426ca853d253 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability High-resolution image synthesis with latent diffusion models
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4936e410-ef66-4b55-92fb-5520789b7837 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Python: a programming language for software integration and development
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 376148d4-a25c-4f5f-a7c9-9358faec5095 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Universal value function approximators
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 38fcec82-510a-4a84-9c06-39722501636c · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b01114b7-071b-4cea-91f1-65115677c446 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Reinforcement learning in markovian and non-markovian environments
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1ecd03db-429e-4f0b-95c7-02500b6b485e · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Trajectory-wise multiple choice learning for dynamics generalization in reinforcement learning
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 407b2b38-e276-4752-86a2-bfd964238e72 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Temporal credit assignment in reinforcement learning
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2f8887-3798-40b0-8beb-df1ae25b0f29 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability DeepMind Control Suite
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d4e7c52-71e6-4868-b5af-3b9062571e5d · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Learning Robot Soccer from Egocentric Vision with Deep Reinforcement Learning
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47b145d8-d437-4e1b-9037-cdcc5700b3f9 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Domain randomization for transferring deep neural networks from simulation to the real world
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2beca31-fc93-478e-9096-a643b663d7be · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Mujoco: A physics engine for model-based control
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 966fed24-1e51-4619-a0e3-1cf3138a950e · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Learning one representation to optimize all rewards
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e5f34fca-cc7e-4d0e-acd8-be2ac9b95955 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Does zero-shot reinforcement learning exist? In The Eleventh International Conference on Learning Representations, 2023
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation de40c939-25da-42d0-8f94-1af2e11fe73f · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Attention is all you need
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a13be115-0c55-47d0-ac2e-63fdc5df8801 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Active perception and reinforcement learning
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e5be8a5a-1179-4702-9d08-d29702676e7a · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Policy gradient critics
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2c43ebbf-cdcb-438d-a90c-89b5e2d43f8b · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Meta-gradient reinforcement learning with an objective discovered online
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4e439a89-65e0-422c-8c60-62163ed0f744 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Don't Change the Algorithm, Change the Data: Exploratory Data for Offline Reinforcement Learning
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 765afd43-5b4c-48f8-8968-bbf63bc2da0a · outbound
Zero-Shot Reinforcement Learning Under Partial Observability Learning deep neural network policies with continuous memory states
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 66ed7fcb-7b1b-4894-b12f-6225b0543ca4 · outbound
Zero-Shot Reinforcement Learning Under Partial Observability write newline
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50626c59-d028-4721-9d66-d21892112808 · inbound
When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited Zero-Shot Reinforcement Learning Under Partial Observability
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.