Pith. sign in

REVIEW 1 cited by

Largest Eigenvalues of the Conjugate Kernel of Single-Layered Neural Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2201.04753 v1 pith:NFUEDZBY submitted 2022-01-13 math.PR cs.LG

classification math.PRcs.LG
keywords randomlargestmatrixneuralfunctionasymptoticconjugatedistribution
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

This paper is concerned with the asymptotic distribution of the largest eigenvalues for some nonlinear random matrix ensemble stemming from the study of neural networks. More precisely we consider $M= \frac{1}{m} YY^\top$ with $Y=f(WX)$ where $W$ and $X$ are random rectangular matrices with i.i.d. centered entries. This models the data covariance matrix or the Conjugate Kernel of a single layered random Feed-Forward Neural Network. The function $f$ is applied entrywise and can be seen as the activation function of the neural network. We show that the largest eigenvalue has the same limit (in probability) as that of some well-known linear random matrix ensembles. In particular, we relate the asymptotic limit of the largest eigenvalue for the nonlinear model to that of an information-plus-noise random matrix, establishing a possible phase transition depending on the function $f$ and the distribution of $W$ and $X$. This may be of interest for applications to machine learning.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Eigenvalue distribution of the Neural Tangent Kernel in the quadratic scaling

    math.PR 2025-08 conditional novelty 7.0 of 10

    The limiting eigenvalue distribution of the two-layer NTK in the quadratic scaling n/(dp) tends to a Marchenko-Pastur map applied to a deterministic measure depending on the activation and output weights.

Pith tools