b200

Tag

Cards List
#b200

@jessiedong_: split-K matrix multiplications are known for giving different answers even when the code and inputs are the same. but h…

X AI KOLs Timeline · 2026-09-02 Cached

The article describes experiments on split-K matrix multiplications, revealing that answer variability depends on block layouts and split counts, with tests on a B200 GPU showing output changes under different conditions.

0 favorites 0 likes
#b200

@Suhail: I looked at 13 different providers for even 1 node of B200/B200s while I wait for my order to get delivered. Zero avail…

X AI KOLs Following · 2026-08-02 Cached

Suhail reports zero availability of B200 GPUs across 13 providers, with prices heading toward $6.50-7/gpu/hr and warns that inference will get more expensive.

0 favorites 0 likes
#b200

GLM-5.2 on 8xB200: the deployment math nobody spells out - NVFP4 + 2x TP=4 replicas should beat TP=8 by ~2x. Full config guidance inside.

Reddit r/LocalLLaMA · 2026-07-07

The article provides the optimal deployment configuration for GLM-5.2 on 8xB200 nodes, showing that NVFP4 with two TP=4 replicas achieves roughly 2x throughput over FP8 TP=8, with detailed performance data and caveats.

0 favorites 0 likes
#b200

@elliotarledge: Claude Fable 5 [max] on KernelBench-Hard. The main kernel that impressed me was a B200 fp8 GEMM: it HAND WROTE raw SM10…

X AI KOLs Timeline · 2026-07-03 Cached

Claude Fable 5 achieves top results on KernelBench-Hard by hand-writing PTX code for B200 fp8 GEMM, outperforming other models and reaching 44-59% of peak performance on compute-bound shapes.

0 favorites 0 likes
← Back to home

Submit Feedback