exllama3

Tag

Cards List
#exllama3

To the dozens of 3x 3090 Local LLM people - I found our current best fit

Reddit r/LocalLLaMA ↗ · 2026-09-20

The author finds that running Qwen 3.8 Next Flash on Exllama3 at 3.05 bpw on 3x 3090 GPUs delivers exceptional performance and quality for local LLM usage, outperforming other quantizations.

0 favorites 0 likes
← Back to home

Submit Feedback