REVIEW 4 cited by
Formation of Representations in Neural Networks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Understanding neural representations will help open the black box of neural networks and advance our scientific understanding of modern AI systems. However, how complex, structured, and transferable representations emerge in modern neural networks has remained a mystery. Building on previous results, we propose the Canonical Representation Hypothesis (CRH), which posits a set of six alignment relations to universally govern the formation of representations in most hidden layers of a neural network. Under the CRH, the latent representations (R), weights (W), and neuron gradients (G) become mutually aligned during training. This alignment implies that neural networks naturally learn compact representations, where neurons and weights are invariant to task-irrelevant transformations. We then show that the breaking of CRH leads to the emergence of reciprocal power-law relations between R, W, and G, which we refer to as the Polynomial Alignment Hypothesis (PAH). We present a minimal-assumption theory proving that the balance between gradient noise and regularization is crucial for the emergence of the canonical representation. The CRH and PAH lead to an exciting possibility of unifying major key deep learning phenomena, including neural collapse and the neural feature ansatz, in a single framework.
Forward citations
Cited by 4 Pith papers
-
Heterosynaptic Circuits Are Universal Gradient Machines
A two-signal (heterosynaptic) synaptic update rule reduces to preconditioned gradient descent at its stationary point, provided consistency scores across neurons share a sign, unifying Hebbian, anti-Hebbian and hetero...
-
Grounding Functional Similarity by Invariance-Aware Model Stitching
FuLA, a task-agnostic stitching objective that aligns intermediate features through the frozen end network, is claimed to be a more reliable functional similarity metric than task-based stitching.
-
Ubiquity of Emergent Hebbian Dynamics in Regularized Learning
L2 weight decay generically makes many learning rules look Hebbian near stationarity, and added noise can make them look anti-Hebbian.
-
Self-Assembly of a Biologically Plausible Learning Circuit
A four-synapse bidirectional learning circuit trained with local heterosynaptic rules matches backpropagation accuracy and is claimed to self-assemble from random connectivity.
Discussion (0). Continue with ORCID to comment.