Pith. sign in

Paper Citation Record · LEDGER

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

As of 17 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 8 inbound Pith citation observations for arXiv:2508.20258.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.20258 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:52:48.879174Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T05:17:05.015451Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T22:29:00.455930Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4fa1617-73b4-46dd-b6ea-ab8848447d71 · outbound

This paper cites Learning to optimize halide with tree search and random programs.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Learning to optimize halide with tree search and random programs

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.395532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.764789Z digest=sha256:bc39eaee830df66ce3f9b2665a6e8f744814461e3dc806f9bc32189c2cdab4b2

Observation f89de765-a9d3-488f-8d16-760f3e165eae · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.769505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.769505Z digest=sha256:99a0de259800cc4f9438ca8aaaa48978e8a3071a217a45de91cae8550f906043

Observation 31f951f8-3696-4014-9fb8-f68bd63548ff · outbound

This paper cites Heterogeneous-computing interface for portability (HIP).

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Heterogeneous-computing interface for portability (HIP)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.382299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.774765Z digest=sha256:abdecd0055322be54ba6b0f6e86cf4bd0b2ff0eb83eb3abb5cd87b1d22cc16ca

Observation 292641d1-778e-4417-9bdd-35500a258163 · outbound

This paper cites ROCprofiler-SDK: Application profiling, tracing, and performance analysis.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization ROCprofiler-SDK: Application profiling, tracing, and performance analysis

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.367684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.778592Z digest=sha256:1d51d46d15703e9e521f6eefabb6373611241e5da9ca37efe948aed1b310505c

Observation c01fd3fa-c52e-4862-996a-2125ba499365 · outbound

This paper cites GPU Kernel Scientist: An LLM-Driven Framework for Iterative Kernel Optimization.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization GPU Kernel Scientist: An LLM-Driven Framework for Iterative Kernel Optimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.782723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.782723Z digest=sha256:a875c74c96b154eb5fa30a74819d6ace775b320b7e43648bcfded4dee3915c98

Observation 4c3ce8d3-a7b1-4892-adc0-6bd824b8e534 · outbound

This paper cites Opentuner: An extensible framework for program autotuning.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Opentuner: An extensible framework for program autotuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.787433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.787433Z digest=sha256:180ad6e18ff80cd7bb73a15ea97a2c1a1243c3e84edf0a4fca62d4a586b514bf

Observation 04a52275-2c27-4c13-ba48-933fee9af7c4 · outbound

This paper cites Intelliperf: LLM-powered autonomous GPU performance engineer, July 2025.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Intelliperf: LLM-powered autonomous GPU performance engineer, July 2025

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.354652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.792055Z digest=sha256:4db6a8e2882f154a26106abf45c568272973a16e4ea451fe6eed75cc3fe9a80a

Observation b683da53-5549-4f81-b824-1b0157d8bf3c · outbound

This paper cites Kevin: Multi-Turn RL for Generating CUDA Kernels.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Kevin: Multi-Turn RL for Generating CUDA Kernels

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.796751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.796751Z digest=sha256:f7934fdedd176b6e0d7bf1c7618774f86340288eda2e4c8a455c1202395628c7

Observation 657f3e46-93ea-4dd0-915e-200af479c25d · outbound

This paper cites Learning to Optimize Tensor Programs.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Learning to Optimize Tensor Programs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.801185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.801185Z digest=sha256:e89747becf8badca13f74c21d39bc43e9f14ebbe2833fa6362739b86949ceebe

Observation e09d121e-42cf-4a1a-8c3e-a243a38349c0 · outbound

This paper cites Multi-Head Attention: Collaborate Instead of Concatenate.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Multi-Head Attention: Collaborate Instead of Concatenate

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.805535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.805535Z digest=sha256:4ddfa7c1f749f94d7baf8a4536ec075e31511b9c8d7ccaf7dee059d65d265ce7

Observation 7213ee2c-2d07-4ec6-bff6-6769b9ca088a · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Flashattention: Fast and memory-efficient exact attention with io-awareness

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.809913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.809913Z digest=sha256:2df698e3c7884d168c3ee931ddd0711c138db827acf5702e5908d00c91fa953e

Observation f0e5c09b-8fae-4560-ba02-65543371751f · outbound

This paper cites Flex Attention: A Programming Model for Generating Optimized Attention Kernels.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Flex Attention: A Programming Model for Generating Optimized Attention Kernels

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.813709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.813709Z digest=sha256:d57469879c6c690f9239c21dabe68b919f3d28b3b775b5b820ddc42c975d1267

Observation 36a59c8e-378c-472c-9ca2-d2b8f488d599 · outbound

This paper cites an unresolved cited work.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:52:49.334947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.818335Z digest=sha256:a8557b17813661d137ff93550239332496309452267c9a12e78fd5e0d4e63c27

Observation 36c229e2-23ef-4c6a-a53b-c3b69a30d2e3 · outbound

This paper cites An integrated gpu power and performance model.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization An integrated gpu power and performance model

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.322336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.821972Z digest=sha256:0976a409b9be9857e08b646c0299360b9271bd68d9560a2fbb85a91a82b4c584

Observation 2b7a7be9-8b5b-4b02-a5d0-a31b7fc2cc6b · outbound

This paper cites Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.825748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.825748Z digest=sha256:f28f170d0dbdf4b394abb90cd373b6c2787e6b564c457d1fabf453865eb2734c

Observation 01ff1895-c04d-4c8b-9061-9db7b4b22854 · outbound

This paper cites Joshi, Hanna Moazam, Heather Miller, Matei Zaharia, and Christopher Potts.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Joshi, Hanna Moazam, Heather Miller, Matei Zaharia, and Christopher Potts

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.308726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.829659Z digest=sha256:902280f43a4cece6647a2bf2fbb56955ace1fa25f144d74fd645de68acaec46c

Observation cd925acc-6a93-4223-9eb5-4146574b6d0e · outbound

This paper cites CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.833300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.833300Z digest=sha256:beee99be556f5f67aaf8a6f03adf44f04f014ea7d3eabb40bcc45a9b880f300b

Observation 3c4d4751-0e71-4a38-821b-cef0392ce7fb · outbound

This paper cites Competition-level code generation with alphacode.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Competition-level code generation with alphacode

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.837373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.837373Z digest=sha256:13d810c77706b97bd214888a88e229a470b2e500b9b636d3c689a001301b02dd

Observation 9da3f234-6622-41cf-8e98-07c8b900f353 · outbound

This paper cites Rigorous evaluation of computer processors with statistical model checking.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Rigorous evaluation of computer processors with statistical model checking

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.282434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.840961Z digest=sha256:ede2f450a06955b31c1475cbc22d994a0be8afc5b4271492e70b3314fa8a9d53

Observation 58f2089d-5bf0-4f28-b3a6-5a3b26175670 · outbound

This paper cites KernelBench: Can LLMs Write Efficient GPU Kernels?.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization KernelBench: Can LLMs Write Efficient GPU Kernels?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.844809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.844809Z digest=sha256:5061f61a24cf7dae9b5b29f156acce0d6cc5057f16ade1156e5966f4ce62ad45

Observation 6446b73b-9261-4bfb-a52b-e267b74cac26 · outbound

This paper cites Quarch: A question-answering dataset for ai agents in computer architecture.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Quarch: A question-answering dataset for ai agents in computer architecture

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.256650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.848915Z digest=sha256:3023659f05e1169a0bb36e59268ca16e906c7c5a255ab00d00ae3e7732acfec8

Observation 6a1a9ed2-b8e6-4477-9869-f92f4785ce49 · outbound

This paper cites Lean Attention: Hardware-Aware Scalable Attention Mechanism for the Decode-Phase of Transformers.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Lean Attention: Hardware-Aware Scalable Attention Mechanism for the Decode-Phase of Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.852560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.852560Z digest=sha256:b5694fe8c2115b244859ae5b29a95b5e131b082e01dc3cf29b909948aab0d090

Observation 4f06d7f6-6316-44d0-b099-f14a41ba61e9 · outbound

This paper cites Learning Performance-Improving Code Edits.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Learning Performance-Improving Code Edits

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.857043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.857043Z digest=sha256:9a008127bec37dea2b79901f8b151a4d359557436725f110bc2a12f04f52e10f

Observation 830203f4-87ea-4b39-a89d-9a4d09ff2d8a · outbound

This paper cites 11.1 amd instincttm mi300 series modular chiplet package–hpc and ai accelerator for exa-class systems.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization 11.1 amd instincttm mi300 series modular chiplet package–hpc and ai accelerator for exa-class systems

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.243144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.861873Z digest=sha256:9530cd3f66cbfa909a05d5f5d67e0914dd013f9d7cac6e4e61631c5d42ad5c55

Observation 702965d7-d447-42e8-a886-19971e214975 · outbound

This paper cites MLPerf Power: Benchmarking the Energy Efficiency of Machine Learning Systems from Microwatts to Megawatts for Sustainable AI.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization MLPerf Power: Benchmarking the Energy Efficiency of Machine Learning Systems from Microwatts to Megawatts for Sustainable AI

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T16:52:48.865474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:52:48.865474Z digest=sha256:512c56dab5e0ae8f86faa2b17bc2329a94e71b9339e3fecbb699c324ee6b2982

Observation af9f8efc-6079-4649-8d54-4b1b0d9bc7db · outbound

This paper cites Measuring energy and power with papi.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Measuring energy and power with papi

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.229627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.870250Z digest=sha256:44413e230d0da8b4027848c63ec51d694c56479ddcc872e4d6b8edee7e4ddebb

Observation 3ebd13be-26ad-4035-89b2-a6dfffaa1053 · outbound

This paper cites Clint Whaley, Antoine Petitet, and Jack J.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Clint Whaley, Antoine Petitet, and Jack J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.209755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.874337Z digest=sha256:a749186f4898cb5514466db0c1569ad5084850b041cdd2b041e7319560f06a56

Observation 8c3b6360-621b-4fdb-bb46-51fe55499cbd · outbound

This paper cites Gonzalez, Ion Stoica, and Koushik Sen.

SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization Gonzalez, Ion Stoica, and Koushik Sen

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:52:49.191567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:52:48.879174Z digest=sha256:ccc4180bc4e6f64f55ad02ed5cc0478b149b72ba7dd01ac5f65f6fde796845f5

Pith citing papers

Observation 88e47590-849a-4654-858c-22b63208049d · inbound

CUDA-L2: Surpassing cuBLAS Performance for Matrix Multiplication through Reinforcement Learning cites this paper.

CUDA-L2: Surpassing cuBLAS Performance for Matrix Multiplication through Reinforcement Learning SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T19:04:28.510867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:04:28.510867Z digest=sha256:1dbceb3cfd488af617eea07e1ef1489dffff60f886c4e9eca9ca75c1073210bd

Observation 1bcbc9cb-2ef2-4442-ad32-d721ebd61e73 · inbound

Towards Automated Kernel Generation in the Era of LLMs cites this paper.

Towards Automated Kernel Generation in the Era of LLMs SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T08:49:50.385438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:49:50.385438Z digest=sha256:3b4338fc517d188707b99bd7fcbc58bbf4c7fd12442a805d894d768dc0844f1b

Observation 86fcc89c-f9cc-45dc-8373-fdd0747458cf · inbound

LEO: Tracing GPU Stall Root Causes via Cross-Vendor Backward Slicing cites this paper.

LEO: Tracing GPU Stall Root Causes via Cross-Vendor Backward Slicing SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T01:04:50.649360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T00:55:09.017037Z digest=sha256:ff95b4c9596e8fc131b7806fc45d0c4a5c7c3f5a5a43df7762adb4375f373d6b

Observation 59802901-4e8e-49ce-8369-9e723d36cdee · inbound

LEO: Tracing GPU Stall Root Causes via Cross-Vendor Backward Slicing cites this paper.

LEO: Tracing GPU Stall Root Causes via Cross-Vendor Backward Slicing SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T15:49:36.610007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:49:36.610007Z digest=sha256:86fa1839ad3f286df5090d78ca09cdf90fe9928825bd68ea5767d4bc3485c0bd

Observation 0670648b-7238-4aa5-936d-26be8e3c7ff1 · inbound

Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts cites this paper.

Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:35:07.141292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T23:33:47.477482Z digest=sha256:2148707aeed93931f2207ddc1d0f0575386838cea779f91324072be145eb0b9b

Observation 460e180a-3037-4469-8328-8c697cb0b7bf · inbound

Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts cites this paper.

Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T05:17:05.015451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:17:05.015451Z digest=sha256:7b07a5079f949eb20ec6ba772adbb265906300fa2af1af6e7c1b0c4e8a416218

Observation 4a565637-34fa-4324-8f73-ba9102e76de5 · inbound

SpecGen: Accelerating Agentic Kernel Optimization with Speculative Generation cites this paper.

SpecGen: Accelerating Agentic Kernel Optimization with Speculative Generation SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T22:29:00.457665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-26T23:35:32.175768Z digest=sha256:33727b8f230e90c081a43b75ad695f82b8f2f654723acc227f5da26ce7a50bbe

Observation 030a03f7-669e-4e8b-b5a6-e982ee3547f3 · inbound

JAXBench: Benchmarking Autonomous TPU Kernel Optimization cites this paper.

JAXBench: Benchmarking Autonomous TPU Kernel Optimization SwizzlePerf: Hardware-Aware LLMs for GPU Kernel Performance Optimization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T13:46:47.513187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:46:47.513187Z digest=sha256:22166da75bf953a94e6913d6295e2767d29ef8e1ba9912493386b0ad3800b20c