Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T16:30:09.564252Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 83 of 83 outbound references and 4 inbound Pith citation observations for arXiv:2411.13410.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T16:30:09.564252Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:49:31.040589Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T21:52:10.203651Z
83 of 83 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation dba0a613-d316-4f9d-b94a-a0cbe8c90065 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Reinforcement learning in healthcare: A survey
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ece67a6e-c4cb-43ec-8111-f4434523d3c8 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Survey on reinforcement learning for language processing
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d533b9b4-7d0f-4a14-823f-c0f7d05e0879 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Recent advances in reinforcement learning in finance.Mathematical Finance, 33(3):437–503, 2023
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0a1a4b4-6abd-4a3b-8a96-033b2393db73 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 912ce50e-cab6-4bf8-aea1-874afb8b87fd · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback A review on interactive reinforcement learning from human social feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 66592c6e-5ce4-451a-8cb7-a8ab2e6d2eaf · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Assessing Generalization in Deep Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8601ff7b-f855-4927-9ca3-e85480ff2dfb · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Latent exploration for reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b3a785be-45c8-44f4-9892-1199dc287144 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Reinforcement learning: A tutorial survey and recent advances
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 779a3c90-4ab6-443f-9b0b-c1e44a71992a · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback A survey of reinforcement learning from human feedback
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0af62018-f84e-4e15-9d07-0ce3d9c9c4aa · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback The RL/LLM Taxonomy Tree: Reviewing Synergies Between Reinforcement Learning and Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d1a0ee0-36f0-4ead-9e5e-827218367201 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback A conceptual framework for externally-influenced agents: An assisted reinforcement learning review
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 46328215-0d1b-4d7a-9e74-0cef46f6e311 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Language as an abstraction for hierarchical deep reinforcement learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c50d11f1-13a8-4bcc-8131-573e4ad2724f · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Beating Atari with Natural Language Guided Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a70e08c7-be92-40b6-8ab9-ab4a73542ce1 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback The arcade learning environment: An evaluation platform for general agents
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a960bfc-8359-4e6c-a399-c0d6a5c680fa · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Human Instruction-Following with Deep Reinforcement Learning via Transfer-Learning from Text
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d691cd-61f9-46eb-a063-a0b105de59da · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b09890bf-e265-4160-8ecb-8e33408dd0e4 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Deep reinforcement learning for instruction following visual navigation in 3d maze-like environments
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 29ab2f82-cc7c-4ad6-81bd-1e3022f00a89 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Learning to follow directions in street view
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a596834c-90ed-4a55-9c85-5aae50cc25ce · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Language Instructed Reinforcement Learning for Human-AI Coordination
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1a5feb8-1933-4169-828f-fe6ac8c498af · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Ask Your Humans: Using Human Instructions to Improve Generalization in Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e1b405a-1831-4b81-9140-de686d30f9ce · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Meta-reinforcement learning via language instructions
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ee46c1be-a28f-4b45-8c27-9803dbbf5d49 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Guiding multi-step rearrangement tasks with natural language instructions
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7a4ab7f4-e985-4296-ac73-e358b99baebd · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Interactive language: Talking to robots in real time
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b1aa724-3810-4a16-9f40-3c5c97943700 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Correcting Robot Plans with Natural Language Feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76f80baf-89aa-4716-99f8-0bdb7696f8ee · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Yell At Your Robot: Improving On-the-Fly from Language Corrections
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb15ca3e-2973-418d-abc0-06094a7bcec3 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Continual learning for instruction following from realtime feedback
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 39ed1bf9-07e2-4e03-9a01-45f86e252ca8 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Incorporating Voice Instructions in Model-Based Reinforcement Learning for Self-Driving Cars
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c60460ae-1e47-43a1-80d3-6f615e7e111e · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Correct me if i’m wrong: Using non-experts to repair reinforcement learning policies
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8e969b6e-c6f5-4969-8125-a5ab2c57e84c · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Natural language specification of reinforcement learning policies through differentiable decision trees
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 256d1d44-d0e8-43af-93a3-40baeece233f · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Ella: Exploration through learned language abstraction
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b95c84dd-6179-4bff-b851-c4812e8457e5 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback How to talk so ai will learn: Instructions, descriptions, and autonomy
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5f637c27-ecbd-438f-9027-04e95f7a3d5f · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Human-in-the-loop reinforcement learning in continuous-action space
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 847d25fa-3f81-41ec-9bfd-71f1fa8c9a97 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Toward human-in-the-loop ai: Enhancing deep reinforcement learning via real-time human guidance for autonomous driving
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c4e0947a-1354-4e51-8642-cb1697bc091f · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Deep reinforcement learning with interactive feedback in a human–robot environment
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 169bccfd-1b74-4778-ab33-6c0468be2041 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Deploying Offline Reinforcement Learning with Human Feedback
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01f09b5c-e070-4ab6-974b-cbe59bf40f02 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Widening the Pipeline in Human-Guided Reinforcement Learning with Explanation and Context-Aware Data Augmentation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1b028d16-ebc7-447e-bf10-22de4c3437d7 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback The Expertise Problem: Learning from Specialized Feedback
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c270c114-4b55-4360-b264-6e99c2474bc8 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Interactive reinforcement learning with bayesian fusion of multimodal advice
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5d13a663-c34d-4dc8-831a-7a98553e55f5 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Policy shaping: Integrating human feedback with reinforcement learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a23ad482-10bd-455c-95c3-3f305bb90669 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Learning from unreliable human action advice in interactive reinforcement learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d05e02e6-2cc2-4618-92db-d43c4a9ed02a · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Improving deep reinforcement learning in minecraft with action advice
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2866a9f4-fa85-4bd3-b11e-4db6dd8d526a · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Advice-guided reinforce- ment learning in a non-markovian environment
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 727e6e35-83aa-47d0-b91f-38fc10d6fabd · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Few-shot preference learning for human-in-the-loop rl
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c49ed6b2-728a-4f45-8e93-1f04123c2b18 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Effect of human guidance and state space size on interactive reinforcement learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 029f2d92-f262-459b-9e30-d5e6821d0d38 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Interactive reinforcement learning from demon- stration and human evaluative feedback
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 21e5c316-b4eb-46b5-90db-7bfbba879e22 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback HAIM-DRL: Enhanced Human-in-the-loop Reinforcement Learning for Safe and Efficient Autonomous Driving
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2957aa46-575b-4285-b102-bf2f97ef7607 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Tag: Teacher-advice mechanism with gaussian process for reinforcement learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8a5396a2-821d-41da-9f83-8e1d8d724ee5 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Guiding Pretraining in Reinforcement Learning with Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b17fa341-a4f4-4a7e-8b96-6fc8542b7e17 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Reinforcement Learning from LLM Feedback to Counteract Goal Misgeneralization
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac41e7ba-232f-4042-9c8d-586f17887e94 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Mutual Enhancement of Large Language and Reinforcement Learning Models through Bi-Directional Feedback Mechanisms: A Planning Case Study
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 512c7b6e-1e6b-4c6e-a772-002df1cccb31 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback LaGR-SEQ: Language-Guided Reinforcement Learning with Sample-Efficient Querying
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2af37483-66ce-4409-ab30-92b282968b82 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Reward Design with Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a787be6d-525b-4114-a43a-a193252556e8 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Enabling Intelligent Interactions between an Agent and an LLM: A Reinforcement Learning Approach
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95c2393c-dd4b-42f8-a8b0-f98d151fdfe4 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Instruction-Following Agents with Multimodal Transformer
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbc52b58-0e77-42c8-8f21-4950f0d39a63 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Eureka: Human-Level Reward Design via Coding Large Language Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a7c8c65-2cd9-4f15-984a-9c543c75ebd3 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Inherently explainable reinforcement learning in natural language
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a01f3d9c-208a-4b6e-8cb6-1fc9066e7e21 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Multi-granularity Knowledge Transfer for Continual Reinforcement Learning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b96f90a6-1417-4d05-a595-cd4b91c98aa0 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Motif: Intrinsic Motivation from Artificial Intelligence Feedback
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67361466-daf3-4a45-9412-d24e78ee2ada · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Building Open-Ended Embodied Agent via Language-Policy Bidirectional Adaptation
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 88801a12-9867-4e80-9afe-0cf8a0b72c20 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback LLM Augmented Hierarchical Agents
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13809432-f3f1-4db0-a8a9-ac6aba334e4f · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Exploiting Contextual Structure to Generate Useful Auxiliary Tasks
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b4c23225-a55b-4689-8944-cd924a87917b · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback WorldCoder, a Model-Based LLM Agent: Building World Models by Writing Code and Interacting with the Environment
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 360aec97-ce14-46ed-a39b-77667414aa3b · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Natural Language Reinforcement Learning
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd84c2cd-df40-4453-9b7c-fb1d5bb418df · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback LLM-Grounder: Open-Vocabulary 3D Visual Grounding with Large Language Model as an Agent
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdb33758-2b85-4dee-a750-e3d7e3b4d9bf · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Interactive planning using large language models for partially observable robotics tasks
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 79446f17-6732-4ccc-8603-307651e85257 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Grounding language to entities and dynamics for generalization in reinforcement learning
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa73d3af-1194-4100-ada5-28e8e6a8577d · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback An End-to-End Approach to Natural Language Object Retrieval via Context-Aware Deep Reinforcement Learning
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ac0ebf85-eb39-406a-8664-8a63ac6ca6ac · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Read, watch, and move: Reinforce- ment learning for temporally grounding natural language descriptions in videos
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 35a08499-919a-45bd-bbd6-f3f4ab59f91b · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback FollowNet: Robot Navigation by Following Natural Language Directions with Deep Reinforcement Learning
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afc2cbac-3051-456f-b4b0-b5c356d60cc8 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Toward collaborative reinforcement learning agents that commu- nicate through text-based natural language
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 724d9902-8733-4cd3-9b5d-200ac32f8df3 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Language Understanding for Text-based Games Using Deep Reinforcement Learning
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d3c351d-cb10-41c4-bb72-f09f8bdbe6a2 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Towards End-to-End Learning for Dialog State Tracking and Management using Deep Reinforcement Learning
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7c04b73-a88e-4a6e-8b1a-be83bb9bcb87 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Online learning of task-driven object-based visual attention control
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5733aa7f-01a7-484e-abd7-f2a68823abae · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Visual navigation with spatial attention
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b076e219-691e-4e3f-b6c9-29155909665a · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback An Actor-Critic-Attention Mechanism for Deep Reinforcement Learning in Multi-view Environments
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 27dd67a0-3297-43e6-86c8-849f1a11c98c · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Deep reinforcement learning for the dynamic and uncertain vehicle routing problem
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0845676b-cf2f-4ef2-b5fd-3e371c79aff2 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Optimizing attention for sequence modeling via reinforcement learning
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 88eef297-5419-4bc3-a90a-18c1be411750 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Attention-based curiosity-driven exploration in deep reinforcement learning
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9a4af098-f314-4e1a-ba64-8eab09b5eff4 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Towards interpretable reinforcement learning using attention augmented agents
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddf6d89f-d111-44ab-b51c-c240eb8f4f33 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Dynamic interaction between reinforcement learning and attention in multidimensional environments
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e8d1eed7-bf4d-44e6-9e14-4f2f75b986e6 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Machine versus Human Attention in Deep Reinforcement Learning Tasks
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 795964a7-88a7-48ac-97f3-32d491dffb94 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback Rt-2: Vision-language-action models transfer web knowledge to robotic control
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9725d750-cd2b-41ce-a8e4-df3bf1b13276 · outbound
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback MLANet: Multi-Level Attention Network with Sub-instruction for Continuous Vision-and-Language Navigation
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9faa6ac4-da25-47ec-a49f-093ca02ed54f · inbound
Large Language Model Agent: A Survey on Methodology, Applications and Challenges A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8f4ded18-1345-48f0-99d2-3ef83c081391 · inbound
MARCO: Meta-Reflection with Cross-Referencing for Code Reasoning A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 693244ae-1063-432e-a529-52d44d492358 · inbound
LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2af1b1a6-5461-4764-9987-518e640f6b4c · inbound
Evaluation of LLMs for mathematical problem solving A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.