Pith. sign in

Paper Citation Record · LEDGER

A Survey of Reinforcement Learning for Optimization in Automation

As of 9 August 2026, this Paper Citation Record lists 100 of 108 outbound references and 0 inbound Pith citation observations for arXiv:2502.09417.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.09417 v1

Coverage vector

measured 100 of 108 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T21:36:33.708676Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 108 outbound references displayed

  • verified exact10
  • verified fuzzy52
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e1228a0a-e975-45b3-960b-cd68601ca37f · outbound

This paper cites an unresolved cited work.

A Survey of Reinforcement Learning for Optimization in Automation Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.190447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.190447Z digest=sha256:c637fab42b09688d772d818c2e4222bfbd1cdaa7f872428338645aeec3e48b82

Observation 49bbd95e-d178-4c57-b174-75b2ce8d4323 · outbound

This paper cites Human-level control through deep reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Human-level control through deep reinforcement learning,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.196234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.196234Z digest=sha256:a9ccf48edfbf2a0816fed88377c3c1578776cbb8f736b4b7c3bc563508d850d9

Observation 17c45300-98f1-4106-ab17-45ef09da20fb · outbound

This paper cites Deep reinforcement learning in smart manufacturing: A review and prospects,.

A Survey of Reinforcement Learning for Optimization in Automation Deep reinforcement learning in smart manufacturing: A review and prospects,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.201565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.201565Z digest=sha256:74cd844317752ecc388a1c32cc7aea3beba4a3c9a3145d4dff9fc0d7bc431157

Observation 02c5188b-dbf8-4d96-83f4-55d8b8b03173 · outbound

This paper cites Applications of reinforcement learning in energy systems,.

A Survey of Reinforcement Learning for Optimization in Automation Applications of reinforcement learning in energy systems,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.207234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.207234Z digest=sha256:8738a9a97e370d30c1ee9fe685ae7dd62ef87f8f96d8c5ce11cd30c27674a758

Observation 1ed9865d-c7da-4b76-89e6-ed0b164afe0d · outbound

This paper cites Reinforcement learning in robotics: A survey,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning in robotics: A survey,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.212197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.212197Z digest=sha256:c66b5a07e5e98c717697cd9171b630d01cb2a1e37e03b7700f77afdbf45b94f8

Observation be8b5044-338f-4630-9e31-11e3c918caf5 · outbound

This paper cites Reinforcement learning applied to production planning and control,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning applied to production planning and control,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.217564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.217564Z digest=sha256:f7bfb5e4425ff3a286b91f2764458c29535f5d941db72235e4c08d3b7d689ef2

Observation 4f3c26f1-ea9a-4db5-94b2-03e80f7a6e12 · outbound

This paper cites A review on reinforcement learning: Introduction and applications in industrial process control,.

A Survey of Reinforcement Learning for Optimization in Automation A review on reinforcement learning: Introduction and applications in industrial process control,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.223245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.223245Z digest=sha256:17a4842d9ecd50e53d5e303da9cff7eee087e3a44eaae09b0ec484e039c4aa2d

Observation 77ac5118-00f6-4e3a-bb4f-7f65f900aa6f · outbound

This paper cites Deep reinforcement learning for inventory control: A roadmap,.

A Survey of Reinforcement Learning for Optimization in Automation Deep reinforcement learning for inventory control: A roadmap,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.228590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.228590Z digest=sha256:7f1fb6b5c51c0230eb6f1c628075c1856754e46523df533ab24d0a59096b1d25

Observation 98f87576-114b-4095-b2d5-197db06aa285 · outbound

This paper cites Metaheuristics in combinatorial optimization: Overview and conceptual comparison,.

A Survey of Reinforcement Learning for Optimization in Automation Metaheuristics in combinatorial optimization: Overview and conceptual comparison,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.233323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.233323Z digest=sha256:4ef7ce67d6f4805ae744600174eb8e0873c6c98f2417488ee6de9dfc948e9d91

Observation 07806f6f-635b-41ba-b86c-0454111b9039 · outbound

This paper cites Deep Reinforcement Learning: An Overview.

A Survey of Reinforcement Learning for Optimization in Automation Deep Reinforcement Learning: An Overview

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.238458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.238458Z digest=sha256:dd6896203d38488ee3660bd6f2210d12bb812fde7bff4bb59ba302aebb90f84f

Observation c2c0afd4-2d13-4cd1-a148-bdb8649ddbd8 · outbound

This paper cites Deep reinforcement learning: A brief survey,.

A Survey of Reinforcement Learning for Optimization in Automation Deep reinforcement learning: A brief survey,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.243952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.243952Z digest=sha256:37b4526a789286662573ea457de1ed03da572707a3ce57039709555cbcefb9da

Observation 1f4390e4-f241-46f6-bf72-c8dc936162ed · outbound

This paper cites A deep reinforcement learning approach for chemical production scheduling,.

A Survey of Reinforcement Learning for Optimization in Automation A deep reinforcement learning approach for chemical production scheduling,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.249382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.249382Z digest=sha256:48de7490bf1fd6e76fa8565f9da8ee4129484e6f00a53a8ce8b642b46f5e0df8

Observation 092bbfb2-5b62-4176-9742-e3faa75bcac5 · outbound

This paper cites Intelligent scheduling of discrete automated production line via deep reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Intelligent scheduling of discrete automated production line via deep reinforcement learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.254402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.254402Z digest=sha256:d6d6779b604aa37e8d42a346db070f38854dee560b027bed267826ec6c864162

Observation 8f581a0b-572f-41da-aa4c-22dbd3f697de · outbound

This paper cites A reinforcement learning method to scheduling problem of steel production process,.

A Survey of Reinforcement Learning for Optimization in Automation A reinforcement learning method to scheduling problem of steel production process,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.259514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.259514Z digest=sha256:a893994e15f79848a6c3a8430e6afac75a8a3ccb33a4bea9d4e0ba900211da7f

Observation e85d2636-7391-4951-88ac-b4fe01ec0ec8 · outbound

This paper cites Distributional Reinforcement Learning for Scheduling of Chemical Production Processes.

A Survey of Reinforcement Learning for Optimization in Automation Distributional Reinforcement Learning for Scheduling of Chemical Production Processes

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:34.150530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.264933Z digest=sha256:e6fb129266697d880f8d1276bb7138a5d3ff4d47b30f3aee7ec1c2c822ac2ce6

Observation 50c6bcaf-e1fa-497e-8c4c-20836238db5d · outbound

This paper cites Reinforcement Learning for Multi-Product Multi-Node Inventory Management in Supply Chains.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement Learning for Multi-Product Multi-Node Inventory Management in Supply Chains

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:34.128913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.270606Z digest=sha256:dce58981488c36d6b4ad564c4e97a97e5d9300b5356c57d3ceefb6a04db561ff

Observation bd0797c2-3e08-459d-80e3-6980351a603c · outbound

This paper cites Reward shaping to improve the performance of deep reinforcement learning in perishable inventory management,.

A Survey of Reinforcement Learning for Optimization in Automation Reward shaping to improve the performance of deep reinforcement learning in perishable inventory management,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.276066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.276066Z digest=sha256:20ab67fe236ee6e3b269740a522fa2393d480ba7c3375a4fd7fee85a4d699ccf

Observation 1e673c8e-0458-415c-af6f-d97e8da66250 · outbound

This paper cites Cooperative multi-agent reinforcement learning for inventory man- agement,.

A Survey of Reinforcement Learning for Optimization in Automation Cooperative multi-agent reinforcement learning for inventory man- agement,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.280831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.280831Z digest=sha256:5e2b5e87d73a62b8eab139edaa08eaad1f393d5192313fe037e7fff72c86efe1

Observation 3173f6d7-6862-4fac-bd11-dad2466e3ed0 · outbound

This paper cites MARLIM: Multi-Agent Reinforcement Learning for Inventory Management.

A Survey of Reinforcement Learning for Optimization in Automation MARLIM: Multi-Agent Reinforcement Learning for Inventory Management

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.285873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.285873Z digest=sha256:dc155e252efd4736c7192815b1dee7816123b88958f7488bb750b1dfaa64909f

Observation 4ed62436-cd2f-4f9a-a31c-af7d2e58250a · outbound

This paper cites Reinforcement and deep reinforce- ment learning-based solutions for machine maintenance planning, scheduling policies, and optimization,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement and deep reinforce- ment learning-based solutions for machine maintenance planning, scheduling policies, and optimization,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.291435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.291435Z digest=sha256:af6b5f796545c2909a92b7106bc79fc479ae81e56e037c8d04d3ae275abd60a3

Observation dc6b4fd7-8851-4360-b2b2-040359d522b5 · outbound

This paper cites Reinforcement learning for dynamic condition-based maintenance of a system with individually repairable components,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning for dynamic condition-based maintenance of a system with individually repairable components,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.296552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.296552Z digest=sha256:4d799dbe6b1747621aefa37d04d63a5acd00cdc8bef3452557e018ebb2ca13f0

Observation 75427182-3295-43c2-9978-b96a43c1e503 · outbound

This paper cites Dynamic maintenance model for a repairable multi-component system using deep reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Dynamic maintenance model for a repairable multi-component system using deep reinforcement learning,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.301291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.301291Z digest=sha256:a9a1338e916662c1892ed93293d956c0e2bad5040b1017cbc1ad25237894eead

Observation 7d4924d7-3694-44d6-a8f0-1ff7f70a7980 · outbound

This paper cites Aircraft main- tenance check scheduling using reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Aircraft main- tenance check scheduling using reinforcement learning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.306212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.306212Z digest=sha256:58800b975598e33ef03dd8f634079f43041f4666d9ae71316410fb1246ddac36

Observation 1fff148a-67d5-4091-8c5c-df02609849ec · outbound

This paper cites Network maintenance planning via multi-agent reinforcement learn- ing,.

A Survey of Reinforcement Learning for Optimization in Automation Network maintenance planning via multi-agent reinforcement learn- ing,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.310870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.310870Z digest=sha256:b3259b4a545b801d3a3f5207a33d6d04e2bb37c7ebe25ab7651c0f8c2cf31830

Observation 6a0797b0-6f65-465a-9067-0a38339cc954 · outbound

This paper cites Reinforcement learning for statistical process control in manufacturing,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning for statistical process control in manufacturing,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.315478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.315478Z digest=sha256:f365ceb7126f15610387172e1b65e7252b1d644cf8104d134381fe711e2505ac

Observation 3faed340-6296-4a10-a72c-96f169277563 · outbound

This paper cites Explainable reinforcement learning in production control of job shop manufacturing system,.

A Survey of Reinforcement Learning for Optimization in Automation Explainable reinforcement learning in production control of job shop manufacturing system,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.320310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.320310Z digest=sha256:49ba101795233fe41ebf7818be76de7a282ac17ec7af69d0549c12898bfdc981

Observation 80ed8307-ae0e-4cb1-aae9-2471e4c02f33 · outbound

This paper cites Using process data to generate an optimal control policy via apprenticeship and reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Using process data to generate an optimal control policy via apprenticeship and reinforcement learning,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.325114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.325114Z digest=sha256:3429537081e55986047360467d71b9c51f3bd49a779e96e4c25387786ef228be

Observation 5da87fee-4c7d-4885-9322-685a70ebdbeb · outbound

This paper cites Reinforcement learning for process control with application in semiconductor manufacturing,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning for process control with application in semiconductor manufacturing,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.330402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.330402Z digest=sha256:be564a0604800eb9d15b7615796a6e0e859977a3e597dd5b5f74a9fb75b38853

Observation 4d6d80e1-92ca-4cbb-827a-2a2d409815cb · outbound

This paper cites Reinforcement learning for whole-building hvac control and demand response,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning for whole-building hvac control and demand response,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.335588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.335588Z digest=sha256:bb6a02cbfab4c41bb326529cb60d6d31829a1372a42fd2151b23693236f13241

Observation 42893808-7b3d-40aa-803c-21b0fb442131 · outbound

This paper cites Using meta reinforcement learning to bridge the gap between simulation and experiment in energy demand response,.

A Survey of Reinforcement Learning for Optimization in Automation Using meta reinforcement learning to bridge the gap between simulation and experiment in energy demand response,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.340446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.340446Z digest=sha256:2fbd5cb551cd07324de5990025353ad110eecd8a52ebda8c02401af6f1484278

Observation 64b59504-e520-4cbb-821e-c4a41865e3b8 · outbound

This paper cites Multiagent reinforce- ment learning for energy management in residential buildings,.

A Survey of Reinforcement Learning for Optimization in Automation Multiagent reinforce- ment learning for energy management in residential buildings,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.345556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.345556Z digest=sha256:f295adf383f8ee3761332475ea993d03f4bc1f05711de937f94c6adb4ef3c31b

Observation a2cfac70-59a5-4229-bec6-c0c34a896022 · outbound

This paper cites Deep reinforcement learning-based demand response for smart facilities energy management,.

A Survey of Reinforcement Learning for Optimization in Automation Deep reinforcement learning-based demand response for smart facilities energy management,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.350303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.350303Z digest=sha256:83965ca820f28e1f88910d4ffe242bfa0b047a33fb6970cbb15dbfd0a0d43586

Observation 1cf0221b-edd3-4756-bef8-a223f754e4ad · outbound

This paper cites Multi-agent deep rein- forcement learning based demand response for discrete manufacturing systems energy management,.

A Survey of Reinforcement Learning for Optimization in Automation Multi-agent deep rein- forcement learning based demand response for discrete manufacturing systems energy management,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:35.123388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.354947Z digest=sha256:5845c5b9a955bd5a125aad0d023e53d43b5770a01a48c13df12b172606e5e769

Observation 347e4ab7-b9d9-4049-b8e7-f82367b9c081 · outbound

This paper cites Testbed implementation of reinforcement learning-based demand response energy management system,.

A Survey of Reinforcement Learning for Optimization in Automation Testbed implementation of reinforcement learning-based demand response energy management system,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:35.108220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.359644Z digest=sha256:8d31542e0229254b34e50bcce0330d2cc93346d0e009e50814c255f6718ff357

Observation 99a2c2d1-40b7-4ffb-8db0-4c5c106cf971 · outbound

This paper cites Deep reinforcement learning for energy management in a microgrid with flexible demand,.

A Survey of Reinforcement Learning for Optimization in Automation Deep reinforcement learning for energy management in a microgrid with flexible demand,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:35.092659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.364648Z digest=sha256:6540d482273d2f58fc920711af4f39e4bd20a8107cf3d6df02958b8af3147398

Observation 7ccde4fa-612d-479f-9c90-c41d7797235d · outbound

This paper cites Energy management for microgrids using a reinforcement learning algorithm,.

A Survey of Reinforcement Learning for Optimization in Automation Energy management for microgrids using a reinforcement learning algorithm,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:35.076898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.369901Z digest=sha256:21927160a4da07391e5e8da122caf63b2664240360894cfd553e5f21e7a005b5

Observation 48958e56-0e98-44f3-9545-839af3dfbfaa · outbound

This paper cites Deep reinforcement learning- based energy management strategy for a microgrid with flexible loads,.

A Survey of Reinforcement Learning for Optimization in Automation Deep reinforcement learning- based energy management strategy for a microgrid with flexible loads,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:35.061494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.374847Z digest=sha256:d082ebc51b79121533d43416673290467ea322433483d23dcb2e32ebf5795d05

Observation 6bd8f1a8-8698-4f64-9c64-07cf32505de1 · outbound

This paper cites Energy management in microgrid based on deep rein- forcement learning with expert knowledge,.

A Survey of Reinforcement Learning for Optimization in Automation Energy management in microgrid based on deep rein- forcement learning with expert knowledge,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:35.045222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.379915Z digest=sha256:b7317a50b6481ec687e090d576e7df974d7136b20ddfeecd01c0d8fdd651cb3a

Observation 40c6fcb0-efbd-4dc9-9645-8c7c9f416923 · outbound

This paper cites Weather-aware data-driven microgrid energy manage- ment using deep reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Weather-aware data-driven microgrid energy manage- ment using deep reinforcement learning,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:35.026912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.385787Z digest=sha256:53eca503717ead31395870de8e7453afdc21e5347d2bb858fd45a1122876f1fd

Observation 11dd84d0-b1e5-44d9-8672-f173fb2df57e · outbound

This paper cites Intelligent multi-microgrid energy management based on deep neural network and model-free reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Intelligent multi-microgrid energy management based on deep neural network and model-free reinforcement learning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:35.010835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.390987Z digest=sha256:3f44ea4961a3c2bac361b08ba68e7eb55c9a81e4a2a73572ccb01cd0b086473c

Observation f61c2f09-028e-424e-ae15-6307a217de58 · outbound

This paper cites Reinforcement learning in sustainable energy and electric systems: A survey,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning in sustainable energy and electric systems: A survey,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.994185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.396006Z digest=sha256:8e32298321ad113c0dad5a693d7c5abb0866bafeb6a01586e134d4bdb1121ec0

Observation 1d31bb53-e907-4d2e-8514-a98310da4b04 · outbound

This paper cites Reinforcement learning and its applications in modern power and energy systems: A review,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning and its applications in modern power and energy systems: A review,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.976513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.400680Z digest=sha256:62a8c2da7d452a5248db5e84822c24a8668277b3aabc995a88c18619def2518e

Observation 30c87fc8-0488-4274-ad9f-8a1135833591 · outbound

This paper cites Reinforcement learning for selective key applications in power systems: Recent advances and future challenges,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning for selective key applications in power systems: Recent advances and future challenges,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.959177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.406012Z digest=sha256:d975b97ee2e6e7db55dc15fcc144be5e5c0e678cca3eda0bf564c866914ee037

Observation dff334e1-0be3-4ba1-bf81-5b18b5b9b0af · outbound

This paper cites A systematic study on reinforcement learning based applications,.

A Survey of Reinforcement Learning for Optimization in Automation A systematic study on reinforcement learning based applications,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.411370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.411370Z digest=sha256:73088cc48259412dd612a7cb2268193420aea072738e9e840dafaa32365318da

Observation 324fb30e-43f3-468e-a472-472e03a0c28c · outbound

This paper cites End-to-end deep reinforcement learning control for hvac systems in office buildings,.

A Survey of Reinforcement Learning for Optimization in Automation End-to-end deep reinforcement learning control for hvac systems in office buildings,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.932097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.416302Z digest=sha256:1928c568a62306a01dc1deb8332f4f486bc0d63e5c100673c688d3980b9f1128

Observation a9ebe5fc-d33a-4eeb-bd93-074165404b49 · outbound

This paper cites A review of reinforcement learn- ing applications to control of heating, ventilation and air conditioning systems,.

A Survey of Reinforcement Learning for Optimization in Automation A review of reinforcement learn- ing applications to control of heating, ventilation and air conditioning systems,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.918676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.421280Z digest=sha256:6e93dd7c8809f056a2f592e98434be59b48b177c7f3ce28882570fd6876f4fc5

Observation e0fcbf14-b415-459d-95fd-f187f1a7eff5 · outbound

This paper cites Safe hvac control via batch reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Safe hvac control via batch reinforcement learning,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.905324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.426203Z digest=sha256:c376a8a49ede7f325b9717cde5567569223f12c25c783b1808ac9fa568f5dea0

Observation a6327798-805b-4e04-9d6f-2739c85483b2 · outbound

This paper cites Study on the application of reinforcement learning in the operation optimization of hvac system,.

A Survey of Reinforcement Learning for Optimization in Automation Study on the application of reinforcement learning in the operation optimization of hvac system,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.891041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.431027Z digest=sha256:fa78d26d181eea88712f1f2b204833ed9501de1383f1b12bc7ef9788205449ba

Observation b1da68b9-e07e-446b-8f1a-7cf744b9d80b · outbound

This paper cites Experimental evalu- ation of model-free reinforcement learning algorithms for continuous hvac control,.

A Survey of Reinforcement Learning for Optimization in Automation Experimental evalu- ation of model-free reinforcement learning algorithms for continuous hvac control,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.875809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.436271Z digest=sha256:d8b6052c467d2eb5943998e90588e981ba9089c288cc52b9c970240268cb310c

Observation 5698b67f-3bc1-494b-88a5-0200a1c37d8e · outbound

This paper cites Robotic arm motion planning based on curriculum reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Robotic arm motion planning based on curriculum reinforcement learning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.860984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.441081Z digest=sha256:ccc804479fe160082c7b81cad80263e568c6f9321956903b63b891efde8f9c4d

Observation 70703579-d103-4c26-b4bf-d060840369dc · outbound

This paper cites Reinforcement Learning Based User-Guided Motion Planning for Human-Robot Collaboration.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement Learning Based User-Guided Motion Planning for Human-Robot Collaboration

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:34.090096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.445979Z digest=sha256:4fd027a61e189705a23ef28d8eb74b816bd9c36b40dba3db3f89584126b7b844

Observation 8833ca38-04e6-464c-b349-7c238eb01ae6 · outbound

This paper cites Reinforcement learning with prior policy guidance for motion planning of dual-arm free-floating space robot,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning with prior policy guidance for motion planning of dual-arm free-floating space robot,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.846502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.451666Z digest=sha256:e984043dbe159cbc5131cd93bea2f3727756dfd141a188998db5b491b1270c4c

Observation 7dc13f4f-04f4-4dd0-abad-6f1e55ef17bc · outbound

This paper cites Dext-Gen: Dexterous Grasping in Sparse Reward Environments with Full Orientation Control.

A Survey of Reinforcement Learning for Optimization in Automation Dext-Gen: Dexterous Grasping in Sparse Reward Environments with Full Orientation Control

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:34.065513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.456710Z digest=sha256:b4b0c9667c1a9308107e5a130e30498bc0eab8107a66d0b26d65cbeffeb1914c

Observation 6618b32d-3e96-4ce5-a2f6-11defec4054b · outbound

This paper cites Robotic grasping using deep reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Robotic grasping using deep reinforcement learning,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.831637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.461980Z digest=sha256:a47427789f6c4de913bb6d86db0fc90e8b78201d417e86a93e84e983ac8ed906

Observation 9fbf6898-67ec-4459-8a8b-a17efeaa3f37 · outbound

This paper cites Mrcdrl: Multi-robot coordination with deep reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Mrcdrl: Multi-robot coordination with deep reinforcement learning,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.817018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.467164Z digest=sha256:b483f65d216954e2cf81b26d027a5adf10299f2a3077486cf35560ffba6b29e5

Observation 304cb149-2520-4a71-b5f9-6dda9c93cd03 · outbound

This paper cites Towards pick and place multi robot coordination using multi-agent deep reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Towards pick and place multi robot coordination using multi-agent deep reinforcement learning,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.801866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.472557Z digest=sha256:37932dadafa3b3d7447194194d7819c94ded4ed57c4977b0c57d4e4c89b605c5

Observation e26e0375-1e66-4b32-93e4-defe3ed8697e · outbound

This paper cites Human-centered collaborative robots with deep reinforcement learn- ing,.

A Survey of Reinforcement Learning for Optimization in Automation Human-centered collaborative robots with deep reinforcement learn- ing,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.786998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.477393Z digest=sha256:eff5991294e95389c0fa945138649569b94e2042aab7705c2f7dadbc6873e393

Observation a5329fba-cc14-4147-9432-8827cc9de23a · outbound

This paper cites Explainable reinforcement learning for human-robot collaboration,.

A Survey of Reinforcement Learning for Optimization in Automation Explainable reinforcement learning for human-robot collaboration,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.771673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.482259Z digest=sha256:f3889b26ad988ad6e670c068f58f5729c810f1253a5f889525662e924be343d7

Observation e9c26b50-51b6-46d4-a109-54ad10613235 · outbound

This paper cites Real-world human-robot collaborative reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Real-world human-robot collaborative reinforcement learning,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.755910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.487077Z digest=sha256:a07912da464c42d207ada2f30f551899ff07df3dd71990c1dc08ef9e005937d8

Observation a3cebd9a-7306-41e0-980e-e3be2a21d1e0 · outbound

This paper cites Human-Robot Gym: Benchmarking Reinforcement Learning in Human-Robot Collaboration.

A Survey of Reinforcement Learning for Optimization in Automation Human-Robot Gym: Benchmarking Reinforcement Learning in Human-Robot Collaboration

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:34.043976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.492184Z digest=sha256:81fc64ebed5dfec61ba5b1d1ff3787d64cad3a204c310c90f92014500092e750

Observation 1d975f13-1b8d-4322-bc7c-b8b386f7c0f1 · outbound

This paper cites Towards safe human-robot collaboration using deep reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Towards safe human-robot collaboration using deep reinforcement learning,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.741221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.497527Z digest=sha256:31bfec997a35af7f970a67d4a792dab7a537f68b8759bb2334a6a54b6878db85

Observation 4d24db4b-5d3e-4ce4-808a-72c701bc2497 · outbound

This paper cites A framework and algorithm for human-robot collaboration based on multimodal reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation A framework and algorithm for human-robot collaboration based on multimodal reinforcement learning,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.726034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.502565Z digest=sha256:e97fed93e6e9b32c4aeb61d3d7d6cdd4cbafa7814cc285850765ce25fb4845e8

Observation 07085747-d61d-4340-a188-6e10b03d593c · outbound

This paper cites A survey of learning-based robot motion planning,.

A Survey of Reinforcement Learning for Optimization in Automation A survey of learning-based robot motion planning,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.711109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.507880Z digest=sha256:9ff5b06ab8151d18ca077a7143d0a2296128d1bf53771f92b90fa996b2673eda

Observation b5f86b26-6acf-48ce-aa1f-5f4094a570ea · outbound

This paper cites A survey on deep reinforcement learning algorithms for robotic manipulation,.

A Survey of Reinforcement Learning for Optimization in Automation A survey on deep reinforcement learning algorithms for robotic manipulation,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.696771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.512848Z digest=sha256:dd8cee6551389da41b2358ba32dcb79894e283392338e759ebe8c3b14bbd771b

Observation a6ade5fc-00bc-479d-a77f-8bf827705595 · outbound

This paper cites Reward shaping to learn natural object manipulation with an anthropomorphic robotic hand and hand pose priors via on- policy reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Reward shaping to learn natural object manipulation with an anthropomorphic robotic hand and hand pose priors via on- policy reinforcement learning,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.682569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.517671Z digest=sha256:afd1f20b6284106f932ed1a66f2bbb5b3051a1cb7209610041c1a02993cb1736

Observation 544cbe48-d6e9-47b8-8218-3278f40e9633 · outbound

This paper cites Enhancing robotic grasping of free-floating targets with soft actor-critic algorithm and tactile sensors: a focus on the pre-grasp stage,.

A Survey of Reinforcement Learning for Optimization in Automation Enhancing robotic grasping of free-floating targets with soft actor-critic algorithm and tactile sensors: a focus on the pre-grasp stage,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.667452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.522486Z digest=sha256:04017cfe57ce6c366725e2d9a73b77a84c9eb031ffb0b24fc256bdac7a5cfacd

Observation 4e1ef3d1-d871-46c9-9656-25a006853d42 · outbound

This paper cites Reinforcement learning for multi-robot system: A review,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning for multi-robot system: A review,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.651171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.528484Z digest=sha256:59057049f497e7d15b84aae03e99610535e59e304f82b1a5bcc2647eb80daba9

Observation ce4dd452-ff06-4335-8d09-119c9fbe2f4e · outbound

This paper cites Coordination of a multi robot system for pick and place using reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Coordination of a multi robot system for pick and place using reinforcement learning,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.633058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.534696Z digest=sha256:db213f6678967c34f51ca37ef709b4c9907c0f27677b4b44daf3f491652f5f66

Observation 501b8719-94cd-4d3c-bf9b-104c37ee22c8 · outbound

This paper cites an unresolved cited work.

A Survey of Reinforcement Learning for Optimization in Automation Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-07T21:36:34.617407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.540512Z digest=sha256:cdbcd7bc4bea0b6365bd187aec3a5c49d82f36e79979d9ed4d7ae6075c1a5917

Observation 16a6ebd3-002d-4a7a-ad9d-7b483a9f128d · outbound

This paper cites Adaptive coordination of multiple learning strategies in brains and robots,.

A Survey of Reinforcement Learning for Optimization in Automation Adaptive coordination of multiple learning strategies in brains and robots,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.601441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.546869Z digest=sha256:020852efb5fd9ae7fb9c89e93ae4a6e01bff295a4adc52e087a38c4ee83b27c0

Observation 550ce57c-396e-4014-87b6-9c82f2c81f7d · outbound

This paper cites Study of sample efficiency improvements for reinforcement learning algorithms,.

A Survey of Reinforcement Learning for Optimization in Automation Study of sample efficiency improvements for reinforcement learning algorithms,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.585485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.552190Z digest=sha256:30becf444446c3492625606b4e492ca39a0f6282f9dab207cd688fc887b8d698

Observation 9fdb04d5-b5d0-4e98-b5ec-1b0ff458c82b · outbound

This paper cites Measuring Progress in Deep Reinforcement Learning Sample Efficiency.

A Survey of Reinforcement Learning for Optimization in Automation Measuring Progress in Deep Reinforcement Learning Sample Efficiency

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:34.022989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.558901Z digest=sha256:94afd4c3721aace2e729c1b08e5c50e42fce6b778ee8a9b8156695f0c8742c07

Observation d842fd78-6b4b-4bf4-9c6a-6f5c545a560e · outbound

This paper cites Maximum Mutation Reinforcement Learning for Scalable Control.

A Survey of Reinforcement Learning for Optimization in Automation Maximum Mutation Reinforcement Learning for Scalable Control

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:34.001403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.565054Z digest=sha256:5e151d15fe025b96f15449cd453dd50f1f80136007feddfc1d4d9048fdee2531

Observation b3b6ec98-e4c1-4072-a26f-c0d1023d67c7 · outbound

This paper cites Sample efficient reinforcement learning method via high efficient episodic memory,.

A Survey of Reinforcement Learning for Optimization in Automation Sample efficient reinforcement learning method via high efficient episodic memory,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.569138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.570915Z digest=sha256:102a86aaaf8e521224006b1baa22977aa5d4d0dbfde75a532a7d0641c76b8ae3

Observation 1b547937-483b-4429-9ad4-4913ae6e7ab8 · outbound

This paper cites Efficient online reinforcement learning with offline data,.

A Survey of Reinforcement Learning for Optimization in Automation Efficient online reinforcement learning with offline data,

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.576696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.576696Z digest=sha256:d97c2517f19eb23ae7d296458d9e3ee1bf4a6d0f27fccd7f28406d5d10446e0f

Observation 5bc23ff2-040c-4f99-929c-b7fbd140bfed · outbound

This paper cites Breaking the sample size barrier in model-based reinforcement learning with a generative model,.

A Survey of Reinforcement Learning for Optimization in Automation Breaking the sample size barrier in model-based reinforcement learning with a generative model,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.539665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.582695Z digest=sha256:27aa95a4b089d4d6214eaea0022f8f5aa384a12849283a227c4718d4ada0e4e9

Observation 97b57b1d-a763-40ee-ad38-fa8bdd7c3e30 · outbound

This paper cites Elastic step ddpg: Multi-step reinforcement learning for improved sample efficiency,.

A Survey of Reinforcement Learning for Optimization in Automation Elastic step ddpg: Multi-step reinforcement learning for improved sample efficiency,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.522087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.587857Z digest=sha256:2e45dc4b32267b0ff6dfde00155687e7cfdd6ff76c3e9b3b238b59b1c32a513f

Observation ee8173eb-355a-4df7-848d-84d4bd778d6b · outbound

This paper cites Sample-efficient reinforcement learning via conservative model-based actor-critic,.

A Survey of Reinforcement Learning for Optimization in Automation Sample-efficient reinforcement learning via conservative model-based actor-critic,

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.505362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.593984Z digest=sha256:13a34e4b29718c21cdffe67ba8ead72b5b97d703b39c9bfa5ad7fa7f9ff79b8a

Observation 0ddf61cb-00a0-4b3c-8ec3-9aebb00f6f46 · outbound

This paper cites Safety robustness of reinforcement learning policies: A view from robust control,.

A Survey of Reinforcement Learning for Optimization in Automation Safety robustness of reinforcement learning policies: A view from robust control,

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.490108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.599188Z digest=sha256:63636aebde825444e972b1e8cf28469c85543cd1943b00ba40e71d33d8acfe94

Observation 2cc06383-7502-4214-af5e-5a22cf820ace · outbound

This paper cites Safe Reinforcement Learning with Dual Robustness.

A Survey of Reinforcement Learning for Optimization in Automation Safe Reinforcement Learning with Dual Robustness

Reference 80

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:33.979594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.604825Z digest=sha256:976360da1179ec3360dd0ae3d0b4befb5d6d22d5e85c99716b5c24447b49101d

Observation d9c98fea-5183-47eb-b67d-026444cd6663 · outbound

This paper cites On the Robustness of Safe Reinforcement Learning under Observational Perturbations.

A Survey of Reinforcement Learning for Optimization in Automation On the Robustness of Safe Reinforcement Learning under Observational Perturbations

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.610142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.610142Z digest=sha256:dd75f779914f42bc6c442fb108427022c0f4a41a258f0059b7377dd292248446

Observation 5457f1a7-3bd7-47a5-bc78-5d901f1baab8 · outbound

This paper cites Safe reinforcement learning using robust control barrier functions,.

A Survey of Reinforcement Learning for Optimization in Automation Safe reinforcement learning using robust control barrier functions,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.475887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.616906Z digest=sha256:cff813aeb5c0489bde44d13e69884df7d780c5ee98f3146fc66c668ff31af1d0

Observation 9c128383-c9dc-4c0d-8e04-0710731d6ad2 · outbound

This paper cites Safe reinforcement learning using robust action governor,.

A Survey of Reinforcement Learning for Optimization in Automation Safe reinforcement learning using robust action governor,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.461578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.621836Z digest=sha256:0a45c6659f267372b4a3962a505bf48342972ece754564effe341d1edd5fe0f0

Observation 226086d9-4c0c-49b2-9579-c358e5e98367 · outbound

This paper cites Safe reinforcement learning using robust mpc,.

A Survey of Reinforcement Learning for Optimization in Automation Safe reinforcement learning using robust mpc,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.626931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.626931Z digest=sha256:5393e3249ef810d386cac09ae9be5d090f2a040bb1e55e65d13a118996417f7f

Observation 0d91ed9e-980d-42b9-8eca-41f78d83e390 · outbound

This paper cites Optimal Transport Perturbations for Safe Reinforcement Learning with Robustness Guarantees.

A Survey of Reinforcement Learning for Optimization in Automation Optimal Transport Perturbations for Safe Reinforcement Learning with Robustness Guarantees

Reference 85

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:33.939102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.631729Z digest=sha256:c8cac2401a98fa1e3414194d41bb481add4b6ce78774dc742de6f44faa9b948b

Observation dff55c0a-093b-4d0a-90fe-d9344f031119 · outbound

This paper cites Task-agnostic safety for rein- forcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Task-agnostic safety for rein- forcement learning,

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.438028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.637192Z digest=sha256:28c1b4fb891a5c51b6f1fd77fb954569d5a21af7e2d343ad1b885c41164c5d44

Observation 29b3ce88-d0b3-4d8a-9a68-69d81ff0f06d · outbound

This paper cites Falsification-based robust ad- versarial reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Falsification-based robust ad- versarial reinforcement learning,

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.423999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.641895Z digest=sha256:f55c62ac624457cc1be415e0fc931a802fc541c6fc348840088157a1002eb602

Observation 503dd278-7361-4b56-8667-d90b26627b8f · outbound

This paper cites A Survey on Interpretable Reinforcement Learning.

A Survey of Reinforcement Learning for Optimization in Automation A Survey on Interpretable Reinforcement Learning

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.646549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.646549Z digest=sha256:ae24723a21dd4e7a8f381947e5c11c0af6c55e9e8a09bc33915d560567d4adff

Observation 36328faf-762e-4051-a9a9-6e03687c19f5 · outbound

This paper cites Interpretable Model-based Hierarchical Reinforcement Learning using Inductive Logic Programming.

A Survey of Reinforcement Learning for Optimization in Automation Interpretable Model-based Hierarchical Reinforcement Learning using Inductive Logic Programming

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-08-07T21:36:33.900439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.651576Z digest=sha256:7f4b68f61ff6b0080b43cb44e451dd989340f0b2d32450bea91e29dd9e0098a0

Observation 039c8826-59dd-45a1-86d0-63e374e46491 · outbound

This paper cites There is no Accuracy-Interpretability Tradeoff in Reinforcement Learning for Mazes.

A Survey of Reinforcement Learning for Optimization in Automation There is no Accuracy-Interpretability Tradeoff in Reinforcement Learning for Mazes

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.656651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.656651Z digest=sha256:de1535e33f12bdd56cd56da9615956c9c43d13996fcd162c867609202c5cbf8f

Observation 060a3d21-8994-4ba3-9c77-72db014205a7 · outbound

This paper cites What do rein- forcement learning models measure? interpreting model parameters in cognition and neuroscience,.

A Survey of Reinforcement Learning for Optimization in Automation What do rein- forcement learning models measure? interpreting model parameters in cognition and neuroscience,

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.407812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.661666Z digest=sha256:143791cca9209d54fef9c08a003ec5c7d528147ffcf31e4389d1fe7bfdce2747

Observation 3f9ff627-597a-4669-a12c-dd8d83ed650e · outbound

This paper cites Reinforcement learning interpretation methods: A survey,.

A Survey of Reinforcement Learning for Optimization in Automation Reinforcement learning interpretation methods: A survey,

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.390612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.666689Z digest=sha256:6b265b588621122383543053be666e76e78aa2f2f4df5d84beffb08303952138

Observation b4292e42-6ae8-4514-a3d1-19e206c1950a · outbound

This paper cites Self- supervised discovering of interpretable features for reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Self- supervised discovering of interpretable features for reinforcement learning,

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.374926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.671778Z digest=sha256:23cc05f0e45aca38badbc5aa5e266afcc57d41f821222749e6865b9f65d1cbee

Observation 9f8389df-804e-4f18-8799-c8c09b7af67c · outbound

This paper cites Learning sparse evidence-driven interpretation to understand deep reinforcement learning agents,.

A Survey of Reinforcement Learning for Optimization in Automation Learning sparse evidence-driven interpretation to understand deep reinforcement learning agents,

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.358993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.676726Z digest=sha256:61e6d73a70d10e40164d26ac735fd4cada8cfa2a6d3f783df4c71cbe2ad9483f

Observation cc37d66b-bf1a-460c-824d-7a63bbd33147 · outbound

This paper cites Meta- learning in neural networks: A survey,.

A Survey of Reinforcement Learning for Optimization in Automation Meta- learning in neural networks: A survey,

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.342061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.681931Z digest=sha256:f7668914760df32db7f18bac2377c8ac5b19582fcdd837067e7df33f5355d6b0

Observation a5e3b93f-c857-4a58-bed6-1a4378ac9865 · outbound

This paper cites Learning action translator for meta reinforcement learning on sparse-reward tasks,.

A Survey of Reinforcement Learning for Optimization in Automation Learning action translator for meta reinforcement learning on sparse-reward tasks,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.325714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.687168Z digest=sha256:8491c02fa7cad9e37b5433aa3e550a004ac8b639948a2762418a90bb7b39ac89

Observation 126117cc-c59f-4078-8cbd-3923a748c2a5 · outbound

This paper cites Curriculum learning for reinforcement learning domains: A framework and survey,.

A Survey of Reinforcement Learning for Optimization in Automation Curriculum learning for reinforcement learning domains: A framework and survey,

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T21:36:33.692403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:36:33.692403Z digest=sha256:6833b6504ae631c1a18d190780a4e8e8ca162eb28e0220b617bbbbf4f60b6ca2

Observation 808b3c83-1777-459a-a754-3ada3a3aa015 · outbound

This paper cites Effective reinforcement learning using transfer learning,.

A Survey of Reinforcement Learning for Optimization in Automation Effective reinforcement learning using transfer learning,

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.299553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.697555Z digest=sha256:76eedd6a19f1c70d9f2e4b930e2732abb85c07293c5f3bf6b8316ac691b7c940

Observation 1268171f-12d1-4267-902c-662832d31dfe · outbound

This paper cites Multi-source transfer learning for deep model-based reinforcement learning,.

A Survey of Reinforcement Learning for Optimization in Automation Multi-source transfer learning for deep model-based reinforcement learning,

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.284441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.702827Z digest=sha256:631912bc3922e265fa3dcee344942293b517b7901b649a5a764023e3bcfa69fa

Observation 1f0bd06a-6bff-4527-8419-76ff699b49a9 · outbound

This paper cites Efficient meta reinforcement learning for preference-based fast adaptation,.

A Survey of Reinforcement Learning for Optimization in Automation Efficient meta reinforcement learning for preference-based fast adaptation,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T21:36:34.266401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T21:36:33.708676Z digest=sha256:2fdda4bf21afcf45501d337977e9fc1277979706fb0a95618b6b7c6bf698803b

Pith citing papers

No inbound Pith citation observations are available.