Pith. sign in

REVIEW 1 cited by

Forward Thinking: Building and Training Neural Networks One Layer at a Time

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1706.02480 v1 pith:AAESLMTU submitted 2017-06-08 stat.ML cs.LG

classification stat.MLcs.LG
keywords networksforwardtimedatadeeplayerlayersneural
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present a general framework for training deep neural networks without backpropagation. This substantially decreases training time and also allows for construction of deep networks with many sorts of learners, including networks whose layers are defined by functions that are not easily differentiated, like decision trees. The main idea is that layers can be trained one at a time, and once they are trained, the input data are mapped forward through the layer to create a new learning problem. The process is repeated, transforming the data through multiple layers, one at a time, rendering a new data set, which is expected to be better behaved, and on which a final output layer can achieve good performance. We call this forward thinking and demonstrate a proof of concept by achieving state-of-the-art accuracy on the MNIST dataset for convolutional neural networks. We also provide a general mathematical formulation of forward thinking that allows for other types of deep learning problems to be considered.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. An optimal control approach for neural network architecture adaptation with a posteriori error estimation

    cs.LG 2026-07 conditional novelty 7.0 of 10

    The paper derives a posteriori error estimates for neural network depth adaptation by formulating training as an optimal control problem and using dual weighted residuals to insert layers where error is highest.

Pith tools