REVIEW 7 cited by
Performance Portable Monte Carlo Particle Transport on Intel, NVIDIA, and AMD GPUs
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
OpenMC is an open source Monte Carlo neutral particle transport application that has recently been ported to GPU using the OpenMP target offloading model. We examine the performance of OpenMC at scale on the Frontier, Polaris, and Aurora supercomputers, demonstrating that performance portability has been achieved by OpenMC across all three major GPU vendors (AMD, NVIDIA, and Intel). OpenMC's GPU performance is compared to both the traditional CPU-based version of OpenMC as well as several other state-of-the-art CPU-based Monte Carlo particle transport applications. We also provide historical context by analyzing OpenMC's performance on several legacy GPU and CPU architectures. This work includes some of the first published results for a scientific simulation application at scale on a supercomputer featuring Intel's Max series "Ponte Vecchio" GPUs. It is also one of the first demonstrations of a large scientific production application using the OpenMP target offloading model to achieve high performance on all three major GPU platforms.
Forward citations
Cited by 7 Pith papers
-
Transformers with Physics-Informed Encodings and Simulation-Based Inference for Robust Detection of Eccentric Binary Black Holes in Pulsar Timing Array Data
Physics-informed Transformer encodings plus conditional normalizing flows yield sharper, better-calibrated posteriors for eccentric BBHs in white-noise PTA data than physics-agnostic SBI baselines.
-
Training Language Models to Use Prolog as a Tool
GRPO can teach a 3B language model to emit executable Prolog, but the highest-accuracy models often hardcode answers instead of reasoning in Prolog, producing an accuracy–auditability trade-off.
-
Mask-GCG: Are All Tokens in Adversarial Suffixes Necessary for Jailbreak Attacks?
Mask-GCG uses learnable masks to prune a minority of low-impact tokens from GCG attack suffixes, slightly improving speed while showing most tokens are necessary.
-
Multilingual Training and Evaluation Resources for Vision-Language Models
Releases regenerated multilingual training data and translated benchmarks for VLMs in five languages and demonstrates consistent benefits from multilingual training over English-only baselines.
-
Scientific-Intention Driven Embodied Intelligent Solar Telescope: Conceptual Design
A three-layer AI-agent design for intention-driven autonomous solar telescopes is proposed; only the precision temperature-control prototype was tested.
-
HySafe-AI: Hybrid Safety Architectural Analysis Framework for AI Systems: A Case Study
A hybrid FMEA/FTA safety-analysis framework for foundation-model-based autonomous driving, illustrated on a GenAD and GAIA-2 style reference architecture.
-
Evaluating LLM Agent Collusion in Double Auctions
LLM sellers in a simulated double auction collude more when they can communicate, and urgency from an authority figure sustains collusion even when an overseer monitors them.
Discussion (0). Sign in to comment.