Pith. sign in

REVIEW 3 cited by

Maven: A Multimodal Foundation Model for Supernova Science

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.16829 v1 pith:KS2DG73H submitted 2024-08-29 astro-ph.HE astro-ph.IMcs.LG

classification astro-ph.HEastro-ph.IMcs.LG
keywords mavenmodeldatanumberobservedsupernovaesyntheticcommon
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A common setting in astronomy is the availability of a small number of high-quality observations, and larger amounts of either lower-quality observations or synthetic data from simplified models. Time-domain astrophysics is a canonical example of this imbalance, with the number of supernovae observed photometrically outpacing the number observed spectroscopically by multiple orders of magnitude. At the same time, no data-driven models exist to understand these photometric and spectroscopic observables in a common context. Contrastive learning objectives, which have grown in popularity for aligning distinct data modalities in a shared embedding space, provide a potential solution to extract information from these modalities. We present Maven, the first foundation model for supernova science. To construct Maven, we first pre-train our model to align photometry and spectroscopy from 0.5M synthetic supernovae using a constrastive objective. We then fine-tune the model on 4,702 observed supernovae from the Zwicky Transient Facility. Maven reaches state-of-the-art performance on both classification and redshift estimation, despite the embeddings not being explicitly optimized for these tasks. Through ablation studies, we show that pre-training with synthetic data improves overall performance. In the upcoming era of the Vera C. Rubin Observatory, Maven serves as a Rosetta Stone for leveraging large, unlabeled and multimodal time-domain datasets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Microlensing Detection and Inference via Learned Bayes Factors

    astro-ph.IM 2026-07 conditional novelty 6.0 of 10

    A unified transformer-based pipeline detects 99.9% of recoverable simulated microlensing events and outperforms literature hard cuts in the short-duration finite-source regime with amortized neural posterior inference.

  2. AstroM$^3$: A self-supervised multimodal model for astronomy

    astro-ph.IM 2024-11 conditional novelty 5.0 of 10

    A trimodal CLIP-style model trained on 21,440 variable stars with light curves, spectra, and metadata improves photometry classification and unsupervised subtype discovery.

  3. From stellar light to astrophysical insight: automating variable star research with machine learning

    astro-ph.IM 2025-07 unverdicted

    An invited review of machine learning for automated variable star research, covering data cleaning, variability classification, stellar parameter inference, and foundation models.

Pith tools