Pith. sign in

REVIEW 1 cited by

Reinforcement Learning for Causal Discovery without Acyclicity Constraints

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.13448 v4 pith:NPQ6NPTG submitted 2024-08-24 cs.LG stat.MEstat.ML

classification cs.LGstat.MEstat.ML
keywords dagsacyclicitycausallearningspaceconstraintsdiscoveryalias
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recently, reinforcement learning (RL) has proved a promising alternative for conventional local heuristics in score-based approaches to learning directed acyclic causal graphs (DAGs) from observational data. However, the intricate acyclicity constraint still challenges the efficient exploration of the vast space of DAGs in existing methods. In this study, we introduce ALIAS (reinforced dAg Learning wIthout Acyclicity conStraints), a novel approach to causal discovery powered by the RL machinery. Our method features an efficient policy for generating DAGs in just a single step with an optimal quadratic complexity, fueled by a novel parametrization of DAGs that directly translates a continuous space to the space of all DAGs, bypassing the need for explicitly enforcing acyclicity constraints. This approach enables us to navigate the search space more effectively by utilizing policy gradient methods and established scoring functions. In addition, we provide compelling empirical evidence for the strong performance of ALIAS in comparison with state-of-the-arts in causal discovery over increasingly difficult experiment conditions on both synthetic and real datasets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Causal Discovery via Bayesian Optimization

    cs.LG 2025-01 conditional novelty 6.0 of 10

    DrBO uses Bayesian optimization with low-rank graph embeddings and trained surrogate models to recover causal DAGs, reaching lower structural error than prior score-based methods on several benchmarks.

Pith tools