Domain-Independent Optimistic Initialization for Reinforcement Learning

Marlos C. Machado; Michael Bowling; Sriram Srinivasan

arxiv: 1410.4604 · v1 · pith:2VUKCFTInew · submitted 2014-10-16 · 💻 cs.LG · cs.AI

Domain-Independent Optimistic Initialization for Reinforcement Learning

Marlos C. Machado , Sriram Srinivasan , Michael Bowling This is my paper

classification 💻 cs.LG cs.AI

keywords initializationoptimisticapproachdomainlearningmustreinforcementcommon

0 comments

read the original abstract

In Reinforcement Learning (RL), it is common to use optimistic initialization of value functions to encourage exploration. However, such an approach generally depends on the domain, viz., the scale of the rewards must be known, and the feature representation must have a constant norm. We present a simple approach that performs optimistic initialization with less dependence on the domain.

This paper has not been read by Pith yet.

Domain-Independent Optimistic Initialization for Reinforcement Learning

discussion (0)