amd-r9700

Tag

Cards List
#amd-r9700

Qwen3.8-Flash-Next turns 4xR9700 into a local AI powerhouse! 120 t/s TG and 12k t/s PP single request with optimized vLLM

Reddit r/LocalLLaMA · yesterday

The article reports that the Qwen3.8-Flash-Next model achieves 120 tokens/second generation speed and 12k tokens/second prefill on a system with 4x AMD R9700 GPUs using optimized vLLM and a custom Docker image.

0 favorites 0 likes
← Back to home

Submit Feedback