Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:52:48.879174Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 8 inbound Pith citation observations for arXiv:2508.20258.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:52:48.879174Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T05:17:05.015451Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T22:29:00.455930Z
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e4fa1617-73b4-46dd-b6ea-ab8848447d71 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Learning to optimize halide with tree search and random programs
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f89de765-a9d3-488f-8d16-760f3e165eae · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31f951f8-3696-4014-9fb8-f68bd63548ff · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Heterogeneous-computing interface for portability (HIP)
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 292641d1-778e-4417-9bdd-35500a258163 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization ROCprofiler-SDK: Application profiling, tracing, and performance analysis
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c01fd3fa-c52e-4862-996a-2125ba499365 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization GPU Kernel Scientist: An LLM-Driven Framework for Iterative Kernel Optimization
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c3ce8d3-a7b1-4892-adc0-6bd824b8e534 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Opentuner: An extensible framework for program autotuning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04a52275-2c27-4c13-ba48-933fee9af7c4 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Intelliperf: LLM-powered autonomous GPU performance engineer, July 2025
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b683da53-5549-4f81-b824-1b0157d8bf3c · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Kevin: Multi-Turn RL for Generating CUDA Kernels
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 657f3e46-93ea-4dd0-915e-200af479c25d · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Learning to Optimize Tensor Programs
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e09d121e-42cf-4a1a-8c3e-a243a38349c0 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Multi-Head Attention: Collaborate Instead of Concatenate
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7213ee2c-2d07-4ec6-bff6-6769b9ca088a · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Flashattention: Fast and memory-efficient exact attention with io-awareness
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0e5c09b-8fae-4560-ba02-65543371751f · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Flex Attention: A Programming Model for Generating Optimized Attention Kernels
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36a59c8e-378c-472c-9ca2-d2b8f488d599 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 36c229e2-23ef-4c6a-a53b-c3b69a30d2e3 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization An integrated gpu power and performance model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2b7a7be9-8b5b-4b02-a5d0-a31b7fc2cc6b · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01ff1895-c04d-4c8b-9061-9db7b4b22854 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Joshi, Hanna Moazam, Heather Miller, Matei Zaharia, and Christopher Potts
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cd925acc-6a93-4223-9eb5-4146574b6d0e · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c4d4751-0e71-4a38-821b-cef0392ce7fb · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Competition-level code generation with alphacode
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9da3f234-6622-41cf-8e98-07c8b900f353 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Rigorous evaluation of computer processors with statistical model checking
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 58f2089d-5bf0-4f28-b3a6-5a3b26175670 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization KernelBench: Can LLMs Write Efficient GPU Kernels?
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6446b73b-9261-4bfb-a52b-e267b74cac26 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Quarch: A question-answering dataset for ai agents in computer architecture
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6a1a9ed2-b8e6-4477-9869-f92f4785ce49 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Lean Attention: Hardware-Aware Scalable Attention Mechanism for the Decode-Phase of Transformers
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f06d7f6-6316-44d0-b099-f14a41ba61e9 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Learning Performance-Improving Code Edits
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 830203f4-87ea-4b39-a89d-9a4d09ff2d8a · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization 11.1 amd instincttm mi300 series modular chiplet package–hpc and ai accelerator for exa-class systems
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 702965d7-d447-42e8-a886-19971e214975 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization MLPerf Power: Benchmarking the Energy Efficiency of Machine Learning Systems from Microwatts to Megawatts for Sustainable AI
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af9f8efc-6079-4649-8d54-4b1b0d9bc7db · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Measuring energy and power with papi
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3ebd13be-26ad-4035-89b2-a6dfffaa1053 · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Clint Whaley, Antoine Petitet, and Jack J
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8c3b6360-621b-4fdb-bb46-51fe55499cbd · outbound
SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Gonzalez, Ion Stoica, and Koushik Sen
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 88e47590-849a-4654-858c-22b63208049d · inbound
CUDA-L2: Surpassing cuBLAS Performance for Matrix Multiplication through Reinforcement Learning SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bcbc9cb-2ef2-4442-ad32-d721ebd61e73 · inbound
Towards Automated Kernel Generation in the Era of LLMs SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86fcc89c-f9cc-45dc-8373-fdd0747458cf · inbound
LEO: Tracing GPU Stall Root Causes via Cross-Vendor Backward Slicing SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 59802901-4e8e-49ce-8369-9e723d36cdee · inbound
LEO: Tracing GPU Stall Root Causes via Cross-Vendor Backward Slicing SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0670648b-7238-4aa5-936d-26be8e3c7ff1 · inbound
Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 460e180a-3037-4469-8328-8c697cb0b7bf · inbound
Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a565637-34fa-4324-8f73-ba9102e76de5 · inbound
SpecGen: Accelerating Agentic Kernel Optimization with Speculative Generation SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 030a03f7-669e-4e8b-b5a6-e982ee3547f3 · inbound
JAXBench: Benchmarking Autonomous TPU Kernel Optimization SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.