Buy recommendations on a thight Budget to aid my RX 6800

Reddit r/LocalLLaMA Tools

Summary

This post discusses budget GPU options (Radeon VII vs two P100s) for LLM inference with an RX 6800, focusing on VRAM vs speed tradeoffs for MoE models.

So after a few hours of reserach, im torn between getting either a radeon vii or 2 p100 (both options for roughly 240€). The Radeon would give me 32gb of vram and fast inferference, while the 2 p100 would give me a total of 48gb, but roughly about 30% slower inference, if my estimate is correct. Are there Valid reasons to go for more VRAM or will it simply go unused? Are my numbers off or did i make a mistake? Been wondering if the additional vram is more usefull for MoE Models at q8? Are there other Bigger MoE models besides qwen and gemma that are worht a look where i might profit of more vram? What are your recommendations? thankfull for any Input
Original Article

Similar Articles

4xR9700, 2xMi210 or 4x4080S 32G

Reddit r/LocalLLaMA

The user is comparing GPU options like R9700, Mi210, and 4080S to achieve 128GB VRAM for running multiple AI models in parallel, considering factors like cost, performance, and compatibility.