REVIEW 2 cited by
Resolving Spurious Correlations in Causal Models of Environments via Interventions
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Causal models bring many benefits to decision-making systems (or agents) by making them interpretable, sample-efficient, and robust to changes in the input distribution. However, spurious correlations can lead to wrong causal models and predictions. We consider the problem of inferring a causal model of a reinforcement learning environment and we propose a method to deal with spurious correlations. Specifically, our method designs a reward function that incentivizes an agent to do an intervention to find errors in the causal model. The data obtained from doing the intervention is used to improve the causal model. We propose several intervention design methods and compare them. The experimental results in a grid-world environment show that our approach leads to better causal models compared to baselines: learning the model on data from a random policy or a policy trained on the environment's reward. The main contribution consists of methods to design interventions to resolve spurious correlations.
Forward citations
Cited by 2 Pith papers
-
Parameter Estimation using Reinforcement Learning Causal Curiosity: Limits and Challenges
Systematic analysis of Causal Curiosity in a simulated robotic manipulator shows high accuracy in single-factor and high-granularity settings, but frequent failures when multiple causal factors vary simultaneously.
-
Efficient and Generalizable Environmental Understanding for Visual Navigation
Adding an auxiliary next-state prediction loss to EmbCLIP substantially improves object and point navigation in RoboTHOR and Habitat and boosts supervised vision-and-language navigation baselines.
Discussion (0). Continue with ORCID to comment.