Pith. sign in

REVIEW 1 cited by

Parallel implementation of the Density Matrix Renormalization Group method achieving a quarter petaFLOPS performance on a single DGX-H100 GPU node

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.07411 v1 pith:WRRHX5FP submitted 2024-07-10 physics.chem-ph cond-mat.str-elphysics.comp-ph

classification physics.chem-phcond-mat.str-elphysics.comp-ph
keywords performanceimplementationactivearchitecturescompareddensitydgx-h100dmrg
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We report cutting edge performance results for a hybrid CPU-multi GPU implementation of the spin adapted ab initio Density Matrix Renormalization Group (DMRG) method on current state-of-the-art NVIDIA DGX-H100 architectures. We evaluate the performance of the DMRG electronic structure calculations for the active compounds of the FeMoco and cytochrome P450 (CYP) enzymes with complete active space (CAS) sizes of up to 113 electrons in 76 orbitals [CAS(113, 76)] and 63 electrons in 58 orbitals [CAS(63, 58)], respectively. We achieve 246 teraFLOPS of sustained performance, an improvement of more than 2.5x compared to the performance achieved on the DGX-A100 architectures and an 80x acceleration compared to an OpenMP parallelized implementation on a 128-core CPU architecture. Our work highlights the ability of tensor network algorithms to efficiently utilize high-performance GPU hardware and shows that the combination of tensor networks with modern large-scale GPU accelerators can pave the way towards solving some of the most challenging problems in quantum chemistry and beyond.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Advantages of density in tensor network geometries for gradient based training

    quant-ph 2024-12 conditional novelty 6.0 of 10

    Densely connected tensor network geometries train to lower infidelity than sparse ones on random quantum states, and a new leaf-contraction trick reduces memory while improving training.

Pith tools