AI as a Compiler: Compiling Triton kernels without the Triton compiler
Researchers propose using large language models (LLMs) to replace the conventional optimizing and lowering pipeline in compiler backends, demonstrating an LLM agent can translate Triton kernels directly into PTX with improved performance (0.83x-3.34x) compared to autotuned Triton.
Save an API key to vote.