Pith. sign in

REVIEW 1 cited by

Interpreting the Residual Stream of ResNet18

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.05340 v1 pith:GWX2OXTM submitted 2024-07-07 cs.CV cs.LG

classification cs.CVcs.LG
keywords streamfeatureresidualblockscaleinputoutputdnns
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A mechanistic understanding of the computations learned by deep neural networks (DNNs) is far from complete. In the domain of visual object recognition, prior research has illuminated inner workings of InceptionV1, but DNNs with different architectures have remained largely unexplored. This work investigates ResNet18 with a particular focus on its residual stream, an architectural mechanism which InceptionV1 lacks. We observe that for a given block, channel features of the stream are updated along a spectrum: either the input feature skips to the output, the block feature overwrites the output, or the output is some mixture between the input and block features. Furthermore, we show that many residual stream channels compute scale invariant representations through a mixture of the input's smaller-scale feature with the block's larger-scale feature. This not only mounts evidence for the universality of scale equivariance, but also presents how the residual stream further implements scale invariance. Collectively, our results begin an interpretation of the residual stream in visual object recognition, finding it to be a flexible feature manager and a medium to build scale invariant representations.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Interpretable Company Similarity with Sparse Autoencoders

    cs.CL 2024-12 conditional novelty 6.0 of 10

    Sparse autoencoder features from Llama 3.1's internal representation of company descriptions yield clusters whose members' stock returns move together more than clusters from SIC codes, BISC, or standard embeddings.

Pith tools