amd-7900xtx

Tag

Cards List
#amd-7900xtx

dual 7900 xtx - some guy made a pretty optimized fork of lamacpp optimized for this setup Qwen 3.8 Q8 at 82 tokens / seconds decode

Reddit r/LocalLLaMA ↗ · 2026-09-18

A GitHub fork of llama.cpp optimized for dual AMD 7900 XTX GPUs, significantly improving decode speed for the Qwen 3.8 Q8 model to 82 tokens per second.

0 favorites 0 likes
← Back to home

Submit Feedback