Pith. sign in

REVIEW 1 cited by

Learning What to Defer for Maximum Independent Sets

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2006.09607 v2 pith:GWVRY2EN submitted 2020-06-17 cs.LG stat.ML

classification cs.LGstat.ML
keywords learningnumbersolutiondefergraphsindependentlarge-scalemaximum
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Designing efficient algorithms for combinatorial optimization appears ubiquitously in various scientific fields. Recently, deep reinforcement learning (DRL) frameworks have gained considerable attention as a new approach: they can automate the design of a solver while relying less on sophisticated domain knowledge of the target problem. However, the existing DRL solvers determine the solution using a number of stages proportional to the number of elements in the solution, which severely limits their applicability to large-scale graphs. In this paper, we seek to resolve this issue by proposing a novel DRL scheme, coined learning what to defer (LwD), where the agent adaptively shrinks or stretch the number of stages by learning to distribute the element-wise decisions of the solution at each stage. We apply the proposed framework to the maximum independent set (MIS) problem, and demonstrate its significant improvement over the current state-of-the-art DRL scheme. We also show that LwD can outperform the conventional MIS solvers on large-scale graphs having millions of vertices, under a limited time budget.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Nonlocal Monte Carlo via Reinforcement Learning

    cs.LG 2025-08 conditional novelty 6.0 of 10

    A reinforcement-learning-trained policy for selecting nonlocal cluster moves improves a Monte Carlo solver for hard 4-SAT benchmarks over simulated annealing.

Pith tools