arXiv cs.CLSeptember 14, 2026
AMDKernelVault: Large-Scale Datasets and Agentic Training for AMD GPU Kernel Optimization
Excerpt
arXiv:2609.12471v1 Announce Type: new Abstract: We introduce AMDKernelVault, an open HIP and Triton kernel corpus and training framework for recent AMD CDNA GPUs. Existing LLM-based kernel agents are largely CUDA/NVIDIA-centric and often depend on repeated frontier-LLM calls for generation, reflection, and optimization. To address this gap, we develop HIPKernelGen and TritonKernelGen, agent-driven pipelines that transform PyTorch references into HIP or Triton kernels, compile and validate candid