Pith. sign in

REVIEW 1 cited by

Neural Network Layer Matrix Decomposition reveals Latent Manifold Encoding and Memory Capacity

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.05968 v1 pith:UOFRU4BM submitted 2023-09-12 cs.LG cs.NEphysics.bio-ph

classification cs.LGcs.NEphysics.bio-ph
keywords layermatrixdecompositioneverytheoremcapacitycontinuousdataset
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We prove the converse of the universal approximation theorem, i.e. a neural network (NN) encoding theorem which shows that for every stably converged NN of continuous activation functions, its weight matrix actually encodes a continuous function that approximates its training dataset to within a finite margin of error over a bounded domain. We further show that using the Eckart-Young theorem for truncated singular value decomposition of the weight matrix for every NN layer, we can illuminate the nature of the latent space manifold of the training dataset encoded and represented by every NN layer, and the geometric nature of the mathematical operations performed by each NN layer. Our results have implications for understanding how NNs break the curse of dimensionality by harnessing memory capacity for expressivity, and that the two are complementary. This Layer Matrix Decomposition (LMD) further suggests a close relationship between eigen-decomposition of NN layers and the latest advances in conceptualizations of Hopfield networks and Transformer NN models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Interpretable QSPR Modeling using Recursive Feature Machines and Multi-scale Fingerprints

    q-bio.BM 2024-11 reject novelty 5.0 of 10

    Using Recursive Feature Machines with a custom hybrid fingerprint yields lower solubility prediction errors than graph neural networks on ESOL and FreeSolv, while also producing feature-importance scores.

Pith tools