Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:32:58.885798Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2504.12568.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:32:58.885798Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
25 of 25 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b43e0d69-3662-49fe-a12a-070e4320112f · outbound
Evolutionary Policy Optimization Optuna: A Next-generation Hyperparameter Optimization Framework
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20f1554-c2e8-409b-abf5-044423659c62 · outbound
Evolutionary Policy Optimization The Arcade Learning Environment: An Evaluation Platform for General Agents
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1c60fa0-26be-4f0a-bc71-5487c7564c4c · outbound
Evolutionary Policy Optimization Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1ea9a75e-2d0e-4a74-887f-317202e05aaa · outbound
Evolutionary Policy Optimization Improving Exploration in Evolution Strategies for Deep Reinforcement Learning via a Population of Novelty-Seeking Agents
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cb7c1eb-051b-49b7-a2f5-143b300fc372 · outbound
Evolutionary Policy Optimization Effective Reinforcement Learning through Evolutionary Surrogate-Assisted Prescription
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7c9a5b4c-a97a-408c-b787-f362a8679664 · outbound
Evolutionary Policy Optimization Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc0cccf9-6d52-4c91-965a-83d179ab36c2 · outbound
Evolutionary Policy Optimization LeCun, B
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7834643c-c9e7-45ae-96f0-c036a218c2f8 · outbound
Evolutionary Policy Optimization Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 258092c3-9f51-4d33-97ca-a7403929e0c3 · outbound
Evolutionary Policy Optimization Evolving Deep Neural Networks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fcf6ed2-cde1-4d81-866b-802490fe6666 · outbound
Evolutionary Policy Optimization Asynchronous Methods for Deep Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64e880e7-d454-43be-ad6d-18209562856c · outbound
Evolutionary Policy Optimization Playing Atari with Deep Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b666516e-87ca-44fc-85df-08d78dbcdfdb · outbound
Evolutionary Policy Optimization Rusu, Joel Veness, Marc G
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0050898-77fa-435c-b008-9fb808fcd08a · outbound
Evolutionary Policy Optimization Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3a5fd67d-ea90-4b98-9529-2217ce5903cc · outbound
Evolutionary Policy Optimization Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25d2c83d-e0f4-486c-942a-89c1e86cb5ad · outbound
Evolutionary Policy Optimization Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba8af3ac-e005-46d7-b247-1ca2274ffc58 · outbound
Evolutionary Policy Optimization Evolution Strategies as a Scalable Alternative to Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f0588a6-41f2-43c1-a5ce-a6bb4de0b74f · outbound
Evolutionary Policy Optimization Trust Region Policy Optimization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0d862f1-4a49-4d51-a95d-0cda71fdf471 · outbound
Evolutionary Policy Optimization Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29aeddba-1058-45d2-8312-57c3cb8caf97 · outbound
Evolutionary Policy Optimization Optimal Advertising for Information Products
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4345154-42eb-4fa7-abc7-a9518981930b · outbound
Evolutionary Policy Optimization Deep Neuroevolution: Genetic Algorithms Are a Competitive Alternative for Training Deep Neural Networks for Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f7cf51d-543c-47a8-b916-37a21ed4ca51 · outbound
Evolutionary Policy Optimization Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b59266da-ddab-4f2d-ae8a-9ef361fb60b1 · outbound
Evolutionary Policy Optimization Sample Efficient Actor-Critic with Experience Replay
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f17263cc-c28f-4303-8056-6af0d767a132 · outbound
Evolutionary Policy Optimization Williams
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af23db3d-fdc9-40ad-aacb-e6d5f6d35e2f · outbound
Evolutionary Policy Optimization Nature 518 (2015), 529–533
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39309e09-c37e-414e-b44c-d727f10cf18c · outbound
Evolutionary Policy Optimization Proximal Policy Optimization Algorithms
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.