memory-requirements

Tag

Cards List
#memory-requirements

Kimi K3 weights drop today. We're deploying on A100s, H200s and B300s this week and the A100 math is already rough

Reddit r/LocalLLaMA · 2026-07-27

Kimi K3 weights are being released today. The model has 2.8T parameters, MoE with 896 experts, 1M context, vision, and MXFP4 quantization. Deployment requires multiple nodes for A100s and H200s, but fits in single B300 node. Benchmarks for tok/s, ttft, and cost per M token across GPU configs are expected by end of week.

0 favorites 0 likes
#memory-requirements

I mapped which local LLMs actually fit each RAM tier, 8 to 128GB (open dataset)

Reddit r/LocalLLaMA · 2026-07-01

An open dataset on GitHub maps which local LLMs fit various RAM tiers (8GB to 128GB), providing memory sizing rules, per-tier model lists, and Ollama commands, with a JSON API for programmatic access.

1 favorites 1 likes
← Back to home

Submit Feedback