REVIEW 2 cited by
An operator preconditioning perspective on training in physics-informed machine learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this paper, we investigate the behavior of gradient descent algorithms in physics-informed machine learning methods like PINNs, which minimize residuals connected to partial differential equations (PDEs). Our key result is that the difficulty in training these models is closely related to the conditioning of a specific differential operator. This operator, in turn, is associated to the Hermitian square of the differential operator of the underlying PDE. If this operator is ill-conditioned, it results in slow or infeasible training. Therefore, preconditioning this operator is crucial. We employ both rigorous mathematical analysis and empirical evaluations to investigate various strategies, explaining how they better condition this critical operator, and consequently improve training.
Forward citations
Cited by 2 Pith papers
-
Regularized dynamical parametric approximation of stiff evolution problems
Regularized implicit Runge-Kutta integrators for nonlinear parametrizations such as neural networks are shown to have global errors bounded by the non-parametric integrator's error plus the size of the residuals.
-
Discrete energy as an exact label-free training objective for finite-element surrogates
For linear elastostatics, training a surrogate by minimizing discrete potential energy is exactly equivalent to supervised stiffness-norm regression, with identical gradients and no reference solutions.
Discussion (0). Continue with ORCID to comment.