Pith. sign in

REVIEW 1 cited by

Amortized Active Causal Induction with Deep Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.16718 v1 pith:FIKIHDNF submitted 2024-05-26 cs.LG cs.AI

classification cs.LGcs.AI
keywords designcausalpolicyamortizedactivedatagraphintervention
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present Causal Amortized Active Structure Learning (CAASL), an active intervention design policy that can select interventions that are adaptive, real-time and that does not require access to the likelihood. This policy, an amortized network based on the transformer, is trained with reinforcement learning on a simulator of the design environment, and a reward function that measures how close the true causal graph is to a causal graph posterior inferred from the gathered data. On synthetic data and a single-cell gene expression simulator, we demonstrate empirically that the data acquired through our policy results in a better estimate of the underlying causal graph than alternative strategies. Our design policy successfully achieves amortized intervention design on the distribution of the training environment while also generalizing well to distribution shifts in test-time design environments. Further, our policy also demonstrates excellent zero-shot generalization to design environments with dimensionality higher than that during training, and to intervention types that it has not been trained on.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Goal-Oriented Sequential Bayesian Experimental Design for Causal Learning

    cs.LG 2025-07 conditional novelty 6.0 of 10

    GO-CBED trains a transformer policy to choose intervention sequences that maximize expected information gain on a user-specified causal query, using a variational bound with normalizing-flow posteriors, and reports ga...

Pith tools