GPU Programming with Triton: Accelerate AI training and inference
Paperback
$69.99
Premium Members save an extra 10% and all Members collect stamps to save with Rewards. 10 stamps = $5.
Get the eBook free when you register your print book at Manning.
Until recently, writing GPU kernels for LLM training and inference meant learning low-level programming tools like CUDA and C++. Triton, an open source, Python-based DSL created by OpenAI, bridges the gap between high-level machine learning frameworks and low-level GPU programming. Triton is built into PyTorch 2 and backed by NVIDIA, Intel, AMD, and Red Hat.
In this book, you'll learn how to work within the Triton ecosystem, fr...
Until recently, writing GPU kernels for LLM training and inference meant learning low-level programming tools like CUDA and C++. Triton, an open source, Python-based DSL created by OpenAI, bridges the gap between high-level machine learning frameworks and low-level GPU programming. Triton is built into PyTorch 2 and backed by NVIDIA, Intel, AMD, and Red Hat.
In this book, you'll learn how to work within the Triton ecosystem, fr...


