OpenAI Blog

Introducing Triton: Open-source GPU programming for neural networks

Read the original on OpenAI Blog →

We’re releasing Triton 1. 0, an open-source Python-like programming language which enables researchers with no CUDA experience to write highly efficient GPU code—most of the time on par with what an expert would be able to produce.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at OpenAI Blog.

arXiv AI
4d ago

AI as a Compiler: Compiling Triton kernels without the Triton compiler

The paper explores using large language models (LLMs) to replace traditional compiler backends, a process termed AI lowering. An LLM agent translates Triton kernels directly into NVIDIA PTX, achieving 0.83x–3.34x the performance of autotuned Triton on a variety of GPUs and ML kernels. The study also extends a PTX verifier to support modern GPU features, highlighting the potential for AI compilers to reduce engineering effort for new hardware.

By Fran\c{c}ois Costa, Charly Castes, Thomas Bourgeat, Azalia Mirhoseini