training-kernel

Tag

Cards List
#training-kernel

@cursor_ai: We're open-sourcing Mixture-of-Kittens (MoK), our MoE training megakernel for NVL72s. It fuses all Mixture-of-Experts c…

X AI KOLs Timeline · 2026-08-04 Cached

Cursor is open-sourcing Mixture-of-Kittens (MoK), a fused MoE training megakernel for NVIDIA NVL72 that runs up to 2.37x faster than the strongest public baselines.

0 favorites 0 likes
#training-kernel

Flash-MSA: Accelerating Million-Token Training with Sparse Attention Kernels

Hacker News Top · 2026-07-12 Cached

Introduces Flash-MSA, the first performant open-source training kernels for MiniMax Sparse Attention on Hopper and Blackwell GPUs, enabling efficient million-token training.

0 favorites 0 likes
← Back to home

Submit Feedback