Loading...
Triton Programming — all 26 problems
Write GPU kernels with OpenAI Triton. Start from vector addition and build up to fused elementwise ops, row reductions, softmax, LayerNorm, tiled matrix multiplication, autotuning, fused dropout, and a simplified attention kernel — all running on a real NVIDIA GPU.