Pith. sign in

REVIEW 3 cited by

Reinforcement Learning with Efficient Active Feature Acquisition

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2011.00825 v1 pith:GOM7HXZ2 submitted 2020-11-02 cs.LG cs.AI

classification cs.LGcs.AI
keywords acquisitioninformationcostmedicalactiveagentdomainefficient
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Solving real-life sequential decision making problems under partial observability involves an exploration-exploitation problem. To be successful, an agent needs to efficiently gather valuable information about the state of the world for making rewarding decisions. However, in real-life, acquiring valuable information is often highly costly, e.g., in the medical domain, information acquisition might correspond to performing a medical test on a patient. This poses a significant challenge for the agent to perform optimally for the task while reducing the cost for information acquisition. In this paper, we propose a model-based reinforcement learning framework that learns an active feature acquisition policy to solve the exploration-exploitation problem during its execution. Key to the success is a novel sequential variational auto-encoder that learns high-quality representations from partially observed states, which are then used by the policy to maximize the task reward in a cost efficient manner. We demonstrate the efficacy of our proposed framework in a control domain as well as using a medical simulator. In both tasks, our proposed method outperforms conventional baselines and results in policies with greater cost efficiency.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Stochastic Encodings for Active Feature Acquisition

    cs.LG 2025-08 conditional novelty 7.0 of 10

    SEFA, a supervised latent-variable model with stochastic encoders and a gradient-based acquisition score, outperforms RL and mutual-information baselines on active feature acquisition benchmarks.

  2. PRECISE-AS: Personalized Reinforcement Learning for Efficient Point-of-Care Echocardiography in Aortic Stenosis Diagnosis

    cs.CV 2025-09 conditional novelty 5.0 of 10

    An RL-driven video acquisition policy for echocardiography keeps aortic stenosis classification at 80.6% balanced accuracy while acquiring only 47% of the videos on average.

  3. To Measure or Not: A Cost-Sensitive, Selective Measuring Environment for Agricultural Management Decisions with Reinforcement Learning

    cs.LG 2025-01 conditional novelty 5.0 of 10

    A cost-sensitive reinforcement learning environment shows an agent can learn when to pay for crop measurements to guide nitrogen fertilization in winter wheat.

Pith tools