Pith. sign in

REVIEW 4 cited by

Verifying the Union of Manifolds Hypothesis for Image Data

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2207.02862 v3 pith:AOJ632MX submitted 2022-07-06 stat.ML cs.AIcs.LG

classification stat.MLcs.AIcs.LG
keywords datahypothesisintrinsicliesmanifoldsuniondimensionimage
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep learning has had tremendous success at learning low-dimensional representations of high-dimensional data. This success would be impossible if there was no hidden low-dimensional structure in data of interest; this existence is posited by the manifold hypothesis, which states that the data lies on an unknown manifold of low intrinsic dimension. In this paper, we argue that this hypothesis does not properly capture the low-dimensional structure typically present in image data. Assuming that data lies on a single manifold implies intrinsic dimension is identical across the entire data space, and does not allow for subregions of this space to have a different number of factors of variation. To address this deficiency, we consider the union of manifolds hypothesis, which states that data lies on a disjoint union of manifolds of varying intrinsic dimensions. We empirically verify this hypothesis on commonly-used image datasets, finding that indeed, observed data lies on a disconnected set and that intrinsic dimension is not constant. We also provide insights into the implications of the union of manifolds hypothesis in deep learning, both supervised and unsupervised, showing that designing models with an inductive bias for this structure improves performance across classification and generative modelling tasks. Our code is available at https://github.com/layer6ai-labs/UoMH.

Discussion (0). Sign in to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Hessian Geometry of Latent Space in Generative Models

    cs.LG 2025-06 reject novelty 6.0 of 10

    A method to recover the Hessian/Fisher metric of generative model latent spaces is validated on Ising and TASEP and applied to claim fractal phase transitions in diffusion models.

  2. Sparse Autoencoders, Again?

    cs.LG 2025-06 conditional novelty 6.0 of 10

    VAEase gates the VAE decoder input by the encoder's variance, combining sparse-autoencoder adaptive sparsity with a hyperparameter-free loss; a global-minimizer theorem says active latent dimensions recover per-manifo...

  3. Diffusion Sampling Path Tells More: An Efficient Plug-and-Play Strategy for Sample Filtering

    cs.CV 2025-05 conditional novelty 6.0 of 10

    CFG-Rejection filters low-quality diffusion samples early using the accumulated norm of the classifier-free guidance vector, improving quality scores without external reward models.

  4. Computing Optimal Transport Maps and Wasserstein Barycenters Using Conditional Normalizing Flows

    stat.ML 2025-05 conditional novelty 6.0 of 10

    A conditional normalizing flow method that solves the primal optimal transport problem and computes Wasserstein-2 barycenters as weighted averages of maps from a shared latent distribution.

Pith tools