moe-experts

Tag

Cards List
#moe-experts

ExLlamav3 Recent Updates : CPU offload, GLM-5.3-FLASH, Qwen3.8-Flash, SC Quants ++

Reddit r/LocalLLaMA · 3d ago

ExLlamav3 has released major updates including CPU offload for MoE experts, support for new AI models like GLM-5.3-Flash and Qwen-3.8-Flash, and various performance optimizations.

0 favorites 0 likes
← Back to home

Submit Feedback