Pith. sign in

REVIEW 2 cited by

Self-Improved Learning for Scalable Neural Combinatorial Optimization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.19561 v3 pith:RJV2PFI3 submitted 2024-03-28 cs.LG cs.AI

classification cs.LGcs.AI
keywords combinatorialmethodoptimizationproblemlarge-scalemodelneuralself-improved
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The end-to-end neural combinatorial optimization (NCO) method shows promising performance in solving complex combinatorial optimization problems without the need for expert design. However, existing methods struggle with large-scale problems, hindering their practical applicability. To overcome this limitation, this work proposes a novel Self-Improved Learning (SIL) method for better scalability of neural combinatorial optimization. Specifically, we develop an efficient self-improved mechanism that enables direct model training on large-scale problem instances without any labeled data. Powered by an innovative local reconstruction approach, this method can iteratively generate better solutions by itself as pseudo-labels to guide efficient model training. In addition, we design a linear complexity attention mechanism for the model to efficiently handle large-scale combinatorial problem instances with low computation overhead. Comprehensive experiments on the Travelling Salesman Problem (TSP) and the Capacitated Vehicle Routing Problem (CVRP) with up to 100K nodes in both uniform and real-world distributions demonstrate the superior scalability of our method.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Rethinking Efficiency in Neural Combinatorial Optimization: Batched Preference Optimization with Mamba

    cs.LG 2026-02 unverdicted novelty 6.0 of 10

    ECO uses supervised warm-up plus iterative batched DPO on a Mamba backbone to reach top neural performance on TSP and CVRP while lowering memory growth and raising throughput.

  2. Hierarchical Learning-based Graph Partition for Large-scale Vehicle Routing Problems

    cs.LG 2025-02 conditional novelty 6.0 of 10

    A hierarchical graph-partition framework with K local two-way re-partition levels improves neural CVRP solvers at scales up to 10,000 customers.

Pith tools