Pith. sign in

REVIEW 7 cited by

Parthenon -- a performance portable block-structured adaptive mesh refinement framework

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2202.12309 v3 pith:MRHGH25I submitted 2022-02-24 cs.DC astro-ph.IM

Parthenon -- a performance portable block-structured adaptive mesh refinement framework

classification cs.DC astro-ph.IM
keywords parthenoncpusexascaleframeworkmeshperformanceportableprogramming
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

On the path to exascale the landscape of computer device architectures and corresponding programming models has become much more diverse. While various low-level performance portable programming models are available, support at the application level lacks behind. To address this issue, we present the performance portable block-structured adaptive mesh refinement (AMR) framework Parthenon, derived from the well-tested and widely used Athena++ astrophysical magnetohydrodynamics code, but generalized to serve as the foundation for a variety of downstream multi-physics codes. Parthenon adopts the Kokkos programming model, and provides various levels of abstractions from multi-dimensional variables, to packages defining and separating components, to launching of parallel compute kernels. Parthenon allocates all data in device memory to reduce data movement, supports the logical packing of variables and mesh blocks to reduce kernel launch overhead, and employs one-sided, asynchronous MPI calls to reduce communication overhead in multi-node simulations. Using a hydrodynamics miniapp, we demonstrate weak and strong scaling on various architectures including AMD and NVIDIA GPUs, Intel and AMD x86 CPUs, IBM Power9 CPUs, as well as Fujitsu A64FX CPUs. At the largest scale on Frontier (the first TOP500 exascale machine), the miniapp reaches a total of $1.7\times10^{13}$ zone-cycles/s on 9,216 nodes (73,728 logical GPUs) at ~92% weak scaling parallel efficiency (starting from a single node). In combination with being an open, collaborative project, this makes Parthenon an ideal framework to target exascale simulations in which the downstream developers can focus on their specific application rather than on the complexity of handling massively-parallel, device-accelerated AMR.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Observational signatures of thermonuclear electron-capture supernovae -- Ne II line strengthening and color evolution as traces of the explosion mechanism

    astro-ph.SR 2026-06 unverdicted novelty 7.0

    Synthetic observables from tECSN models show slower early red-color decline due to higher Ti/Cr and a late-time 12.8 μm Ne II line that strengthens over time, unlike comparable CO deflagration models.

  2. Supersonic Motion in the Driving Region of M82

    astro-ph.GA 2026-07 conditional novelty 6.5

    XRISM velocity dispersion of M82's hot wind cannot be produced by free-wind models alone and requires supersonic (Mach 1.7–3.1) small-scale motions that may channel energy into B-fields and cosmic rays.

  3. BlackHoleWeather -- Spin-coupled chaotic cold accretion across the meso-scale: Morphology and thermodynamics

    astro-ph.GA 2026-05 unverdicted novelty 6.0

    Hybrid SMBH spin model in CCA simulations shows cold gas reservoir independent of spin prescription while turbulence controls angular momentum coherence, spin evolution rate, and jet reorientation.

  4. XMAGNET -- Stir before serving: a Lagrangian perspective on mixing-driven condensation in the intracluster medium

    astro-ph.GA 2026-05 unverdicted novelty 6.0

    Lagrangian tracers show mixing with low-entropy seeds drives most condensation in cluster cores; magnetic fields cause earlier divergence, higher vorticity, lower Mach numbers, and slower cold-cloud motion via tension.

  5. BlackHoleWeather -- Spin-coupled chaotic cold accretion across the meso scale: Variability and kinematics

    astro-ph.GA 2026-05 unverdicted novelty 5.0

    Driven turbulence interrupts meso-scale accretion continuity, dropping radial accretion rates by 2-3 orders of magnitude and jet-axis reorientation rates by two orders of magnitude relative to decaying-turbulence controls.

  6. SACRA-K: A Performance-Portable Numerical Relativity Code with Kokkos

    astro-ph.HE 2026-07 accept novelty 4.0

    A Kokkos-based C++ port of the SACRA numerical relativity code achieves ~10x speedup on GPU/APU over the Fortran CPU version while preserving waveform accuracy, pi-symmetry, and second-order convergence.

  7. GRMHD and GRRT Simulations of Black Hole Accretion: Flares, Precession, and Complex Spacetimes

    astro-ph.HE 2026-06 unverdicted novelty 4.0

    Simulations of accreting black holes in standard and complex spacetimes indicate that magnetic geometry, quantum corrections, and binary dynamics influence flares, precession, photon rings, and multi-wavelength variab...