Tag
The PyTorchCon North America 2026 features a Kernel Engineering track focused on compilers, optimization, custom kernels, and systems work, with a keynote by Mark Saroufim.
Jetha Chan disassembled the Attention kernel for the SM100 data center GPU, identified and fixed inefficiencies, resulting in significant performance improvements.