Pith. sign in

REVIEW 4 cited by

The Landscape and Challenges of HPC Research and LLMs

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.02018 v3 pith:VYABQQQV submitted 2024-02-03 cs.LG

classification cs.LG
keywords languagemodelstaskscomputinghigh-performancellmsresearchtechniques
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recently, language models (LMs), especially large language models (LLMs), have revolutionized the field of deep learning. Both encoder-decoder models and prompt-based techniques have shown immense potential for natural language processing and code-based tasks. Over the past several years, many research labs and institutions have invested heavily in high-performance computing, approaching or breaching exascale performance levels. In this paper, we posit that adapting and utilizing such language model-based techniques for tasks in high-performance computing (HPC) would be very beneficial. This study presents our reasoning behind the aforementioned position and highlights how existing ideas can be improved and adapted for HPC tasks.

Discussion (0). Sign in to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CelloAI: Leveraging Large Language Models for HPC Software Development in High Energy Physics

    cs.SE 2025-08 conditional novelty 6.0 of 10

    A locally hosted RAG-based coding assistant improves kernel retrieval and porting coverage for HEP codebases, though no tested LLM correctly ports the hardest kernels.

  2. Evaluating the Efficacy of LLM-Based Reasoning for Multiobjective HPC Job Scheduling

    cs.DC 2025-05 conditional novelty 6.0 of 10

    ReAct-style LLM schedulers can balance multiple HPC scheduling objectives on 10-100 job workloads, though cloud API latency makes them unsuitable for real-time deployment.

  3. HARGO: Heterogeneity-Aware Reward-Guided Optimization for RL Post-Training of LLMs on HPC Tasks

    cs.LG 2026-07 conditional novelty 5.0 of 10

    Confidence-modulated per-response advantage weighting (HARGO) improves GRPO-style RL post-training on four heterogeneous HPC tasks, leading WinRate, data-race F1, and PLP similarity at 0.5B.

  4. LLM4VV: Evaluating Cutting-Edge LLMs for Generation and Evaluation of Directive-Based Parallel Programming Model Compiler Tests

    cs.SE 2025-07 conditional novelty 5.0 of 10

    In a six-model comparison, DeepSeek-Coder-33B generated the most passable directive-based compiler tests (Pass@1=0.434) and Qwen2.5-Coder-32B judged test validity best (F1=0.735).

Pith tools