Pith. sign in

REVIEW

MetaOptimize: A Framework for Optimizing Step Sizes and Other Meta-parameters

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.02342 v6 pith:YYQ3B5J7 submitted 2024-02-04 cs.LG cs.AImath.OC

classification cs.LGcs.AImath.OC
keywords learningmetaoptimizesizesstepmeta-parameterstrainingintroducemachine
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We address the challenge of optimizing meta-parameters (hyperparameters) in machine learning, a key factor for efficient training and high model performance. Rather than relying on expensive meta-parameter search methods, we introduce MetaOptimize: a dynamic approach that adjusts meta-parameters, particularly step sizes (also known as learning rates), during training. More specifically, MetaOptimize can wrap around any first-order optimization algorithm, tuning step sizes on the fly to minimize a specific form of regret that considers the long-term impact of step sizes on training, through a discounted sum of future losses. We also introduce lower-complexity variants of MetaOptimize that, in conjunction with its adaptability to various optimization algorithms, achieve performance comparable to those of the best hand-crafted learning rate schedules across diverse machine learning tasks.

Discussion (0). Continue with ORCID to comment.

Pith tools