REVIEW 3 cited by
A Mean-Field Optimal Control Formulation of Deep Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Recent work linking deep neural networks and dynamical systems opened up new avenues to analyze deep learning. In particular, it is observed that new insights can be obtained by recasting deep learning as an optimal control problem on difference or differential equations. However, the mathematical aspects of such a formulation have not been systematically explored. This paper introduces the mathematical formulation of the population risk minimization problem in deep learning as a mean-field optimal control problem. Mirroring the development of classical optimal control, we state and prove optimality conditions of both the Hamilton-Jacobi-Bellman type and the Pontryagin type. These mean-field results reflect the probabilistic nature of the learning problem. In addition, by appealing to the mean-field Pontryagin's maximum principle, we establish some quantitative relationships between population and empirical learning problems. This serves to establish a mathematical foundation for investigating the algorithmic and theoretical connections between optimal control and deep learning.
Forward citations
Cited by 3 Pith papers
-
Scalable Dynamic Optimal Transport via Distributed Linearized ADMM
Distributed linearized ADMM with exact proximal steps solves dynamic optimal transport stably near vanishing densities and achieves ~6–7× speedup on 12 cores.
-
Neural Dynamics on Complex Networks
A graph neural network integrated over continuous time learns the differential equations governing networked systems and predicts their future states, outperforming several temporal-graph baselines on simulated dynamics.
-
Deep Learning Theory Review: An Optimal Control and Dynamical Systems Perspective
A review that frames neural networks as dynamical systems, SGD as stochastic dynamics, and training as mean-field optimal control to unify deep learning theory.
Discussion (0). Continue with ORCID to comment.