nvl72

Tag

Cards List
#nvl72

NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut

NVIDIA Blog · 6d ago Cached

NVIDIA's Vera Rubin NVL72 system debuts with leading performance in MLPerf Inference v6.1, delivering up to 3.7x better throughput than GB300 NVL72 on benchmarks like DeepSeek-R1 and Qwen3-VL, emphasizing system performance, scaling efficiency, and software optimizations.

0 favorites 0 likes
#nvl72

Mixture-of-Kittens: our open-source MoE megakernel for NVL72s (25 minute read)

TLDR AI · 2026-08-05 Cached

Cursor is open-sourcing Mixture-of-Kittens (MoK), a production MoE training megakernel for NVL72s that fuses communication and computation, delivering a 1.41x end-to-end training throughput improvement for their Composer model.

0 favorites 0 likes
#nvl72

@cursor_ai: We're open-sourcing Mixture-of-Kittens (MoK), our MoE training megakernel for NVL72s. It fuses all Mixture-of-Experts c…

X AI KOLs Timeline · 2026-08-04 Cached

Cursor is open-sourcing Mixture-of-Kittens (MoK), a fused MoE training megakernel for NVIDIA NVL72 that runs up to 2.37x faster than the strongest public baselines.

0 favorites 0 likes
#nvl72

SpaceX AI Sat V1 peak power spec has been raised to ~250kW (battery-assisted), with average power of ~160kW. Will be able to handle an NVL72 Ruben rack.

Reddit r/singularity · 2026-07-16 Cached

SpaceX raised the peak power spec of its AI Satellite V1 to ~250kW (battery-assisted) with ~160kW average, enabling it to handle an NVL72 Ruben rack.

0 favorites 0 likes
#nvl72

@haoailab: Can Attention-FFN Disaggregation still win on the newest rack-scale GPU systems? We built FastAFD, an open-source AFD r…

X AI KOLs Timeline · 2026-07-13 Cached

FastAFD is an open-source serving system for Attention-FFN Disaggregation of MoE models on Blackwell NVL72, achieving 1.35-1.45× per-GPU decode throughput improvement over colocated MoE serving.

0 favorites 0 likes
← Back to home

Submit Feedback