REVIEW 1 cited by
Causally Correct Partial Models for Reinforcement Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In reinforcement learning, we can learn a model of future observations and rewards, and use it to plan the agent's next actions. However, jointly modeling future observations can be computationally expensive or even intractable if the observations are high-dimensional (e.g. images). For this reason, previous works have considered partial models, which model only part of the observation. In this paper, we show that partial models can be causally incorrect: they are confounded by the observations they don't model, and can therefore lead to incorrect planning. To address this, we introduce a general family of partial models that are provably causally correct, yet remain fast because they do not need to fully model future observations.
Forward citations
Cited by 1 Pith paper
-
Parameter Estimation using Reinforcement Learning Causal Curiosity: Limits and Challenges
Systematic analysis of Causal Curiosity in a simulated robotic manipulator shows high accuracy in single-factor and high-granularity settings, but frequent failures when multiple causal factors vary simultaneously.
Discussion (0). Continue with ORCID to comment.