Pith. sign in

Paper Citation Record · LEDGER

Evolutionary Policy Optimization

As of 18 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2504.12568.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.12568 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:32:58.885798Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b43e0d69-3662-49fe-a12a-070e4320112f · outbound

This paper cites Optuna: A Next-generation Hyperparameter Optimization Framework.

Evolutionary Policy Optimization Optuna: A Next-generation Hyperparameter Optimization Framework

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.756420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.756420Z digest=sha256:fecd64b2de1e229740c8ae5d65c3fb543cb762a8d0ce25a238440b8649dd6959

Observation d20f1554-c2e8-409b-abf5-044423659c62 · outbound

This paper cites The Arcade Learning Environment: An Evaluation Platform for General Agents.

Evolutionary Policy Optimization The Arcade Learning Environment: An Evaluation Platform for General Agents

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.762671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.762671Z digest=sha256:37f43e01929cbb8cd747745fca4b1acc3310304dab6e5e1bac6bc11ccdce46ec

Observation c1c60fa0-26be-4f0a-bc71-5487c7564c4c · outbound

This paper cites an unresolved cited work.

Evolutionary Policy Optimization Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:32:59.395194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:32:58.767779Z digest=sha256:b800e2e1cbf0e42c904f93602e3724f6597fd75a02c9868f2151da3d28c404cf

Observation 1ea9a75e-2d0e-4a74-887f-317202e05aaa · outbound

This paper cites Improving Exploration in Evolution Strategies for Deep Reinforcement Learning via a Population of Novelty-Seeking Agents.

Evolutionary Policy Optimization Improving Exploration in Evolution Strategies for Deep Reinforcement Learning via a Population of Novelty-Seeking Agents

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.773221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.773221Z digest=sha256:0a8e8694c1e6ec9359cae3cd328cc974a583251c4eb37064f3b9f29e87de8bd2

Observation 2cb7c1eb-051b-49b7-a2f5-143b300fc372 · outbound

This paper cites Effective Reinforcement Learning through Evolutionary Surrogate-Assisted Prescription.

Evolutionary Policy Optimization Effective Reinforcement Learning through Evolutionary Surrogate-Assisted Prescription

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-16T12:32:59.141680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:32:58.778414Z digest=sha256:778df0f578296c391be0682ddb6f9752ccf7b0d16244ba620d02fa0bd196d54d

Observation 7c9a5b4c-a97a-408c-b787-f362a8679664 · outbound

This paper cites Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor.

Evolutionary Policy Optimization Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.784184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.784184Z digest=sha256:b2a22e50068f9981c358b18730be428e01ed6e82130e584043585c5b58d06e90

Observation fc0cccf9-6d52-4c91-965a-83d179ab36c2 · outbound

This paper cites LeCun, B.

Evolutionary Policy Optimization LeCun, B

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.790142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.790142Z digest=sha256:4811ea84c0dd882185387e51ceef84de4d4a1a1f23410290c420988f3815d978

Observation 7834643c-c9e7-45ae-96f0-c036a218c2f8 · outbound

This paper cites an unresolved cited work.

Evolutionary Policy Optimization Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.795130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.795130Z digest=sha256:26e7fcaad3f1bcf0aedbdea1a752bca6f5ea387cfd896f543a36a00c650385d0

Observation 258092c3-9f51-4d33-97ca-a7403929e0c3 · outbound

This paper cites Evolving Deep Neural Networks.

Evolutionary Policy Optimization Evolving Deep Neural Networks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.801681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.801681Z digest=sha256:bd08d274cdd8bc0833eb44e0abd5ce9f858d744437abe4fe658473e6b5e6a9e1

Observation 6fcf6ed2-cde1-4d81-866b-802490fe6666 · outbound

This paper cites Asynchronous Methods for Deep Reinforcement Learning.

Evolutionary Policy Optimization Asynchronous Methods for Deep Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.807134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.807134Z digest=sha256:73cf47b76f35878707a48fd64ed820d44ac92b4214f10981d40671971e09617e

Observation 64e880e7-d454-43be-ad6d-18209562856c · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Evolutionary Policy Optimization Playing Atari with Deep Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.812868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.812868Z digest=sha256:020d646e1422e086d79db9802b226413f4c1ce2f3d4e39f20756f3ce860bdef8

Observation b666516e-87ca-44fc-85df-08d78dbcdfdb · outbound

This paper cites Rusu, Joel Veness, Marc G.

Evolutionary Policy Optimization Rusu, Joel Veness, Marc G

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.818950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.818950Z digest=sha256:4feaabb2e4d53c1c4e4cba03c5cbe181b176eb9f881c6e4f8f1ae0a306712a7a

Observation f0050898-77fa-435c-b008-9fb808fcd08a · outbound

This paper cites an unresolved cited work.

Evolutionary Policy Optimization Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:32:59.287611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T12:32:58.829865Z digest=sha256:9c361a16bb94a3c35f5f3b58651f28ba4b82face0a8a8298b8f8b253822c0797

Observation 3a5fd67d-ea90-4b98-9529-2217ce5903cc · outbound

This paper cites an unresolved cited work.

Evolutionary Policy Optimization Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.834697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.834697Z digest=sha256:eb3201f7b9fce02150da2111d5ab427b4e8fc0ed9ad9fb1fd19b39d3475c59ed

Observation 25d2c83d-e0f4-486c-942a-89c1e86cb5ad · outbound

This paper cites an unresolved cited work.

Evolutionary Policy Optimization Unresolved cited work

Reference 15

Resolution
malformed identifier
no resolver link, observed 2026-08-16T12:32:58.839428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.839428Z digest=sha256:89b68879870867b6208494978992b499b550174547c3d9743bc914eda5fcd226

Observation ba8af3ac-e005-46d7-b247-1ca2274ffc58 · outbound

This paper cites Evolution Strategies as a Scalable Alternative to Reinforcement Learning.

Evolutionary Policy Optimization Evolution Strategies as a Scalable Alternative to Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.845268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.845268Z digest=sha256:7e1546e02d96058330b5f95574943c0d7d83e2a4cc137e1d214b3ec4765ece9b

Observation 5f0588a6-41f2-43c1-a5ce-a6bb4de0b74f · outbound

This paper cites Trust Region Policy Optimization.

Evolutionary Policy Optimization Trust Region Policy Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.851369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.851369Z digest=sha256:cee4d9c6b3da25763f08e677f26806828caefbfb57fa85b72b89a9a6a77941da

Observation f0d862f1-4a49-4d51-a95d-0cda71fdf471 · outbound

This paper cites an unresolved cited work.

Evolutionary Policy Optimization Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.856040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.856040Z digest=sha256:27d4e2727527f7b12203f4b8db274a71fdfe8cbe5647f9b6f87ef2c997807113

Observation 29aeddba-1058-45d2-8312-57c3cb8caf97 · outbound

This paper cites Optimal Advertising for Information Products.

Evolutionary Policy Optimization Optimal Advertising for Information Products

Reference 19

Resolution
malformed identifier
no resolver link, observed 2026-08-16T12:32:58.865853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.865853Z digest=sha256:202f22d4ff9dafc7aa149ad1f0c69cd8609e61690a36c0f8456a991e3cedd24e

Observation b4345154-42eb-4fa7-abc7-a9518981930b · outbound

This paper cites Deep Neuroevolution: Genetic Algorithms Are a Competitive Alternative for Training Deep Neural Networks for Reinforcement Learning.

Evolutionary Policy Optimization Deep Neuroevolution: Genetic Algorithms Are a Competitive Alternative for Training Deep Neural Networks for Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.871043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.871043Z digest=sha256:16972dee83ff1c12f767ad252a98b501715d5f4e0ec506f78d1dbdce4c4c0291

Observation 8f7cf51d-543c-47a8-b916-37a21ed4ca51 · outbound

This paper cites an unresolved cited work.

Evolutionary Policy Optimization Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.876318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.876318Z digest=sha256:9e49ef32b4e58d78ea034ecb0635c90f4aa7008a1d56a644af20072746a0b1ce

Observation b59266da-ddab-4f2d-ae8a-9ef361fb60b1 · outbound

This paper cites Sample Efficient Actor-Critic with Experience Replay.

Evolutionary Policy Optimization Sample Efficient Actor-Critic with Experience Replay

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.880894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.880894Z digest=sha256:0ab56413274ba2566a78df8c82a9ef84121613ebc2b30f2c3a1a1c03b5f5967c

Observation f17263cc-c28f-4303-8056-6af0d767a132 · outbound

This paper cites Williams.

Evolutionary Policy Optimization Williams

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.885798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.885798Z digest=sha256:70b5a37013c860a004735348691844fd8cf66a0d62ddb665db0983ea085435e7

Observation af23db3d-fdc9-40ad-aacb-e6d5f6d35e2f · outbound

This paper cites Nature 518 (2015), 529–533.

Evolutionary Policy Optimization Nature 518 (2015), 529–533

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.824420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.824420Z digest=sha256:24152b58626b866cdf0327382374102a5baab2343c7f21e742868620d466108a

Observation 39309e09-c37e-414e-b44c-d727f10cf18c · outbound

This paper cites Proximal Policy Optimization Algorithms.

Evolutionary Policy Optimization Proximal Policy Optimization Algorithms

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-16T12:32:58.861073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:32:58.861073Z digest=sha256:d37687daa4c82ba520cfe9dec5359614485da18a62389b9129a4e379c77bf0ce

Pith citing papers

No inbound Pith citation observations are available.