egpu

Tag

Cards List
#egpu

3 days benchmarking most llama.cpp flags on my weird 40gb vram laptop + tb4 egpu setup. Got +70% generation, +40% prefill, 60k more context, and filed a bug in llama around MTP. What I learned.

Reddit r/LocalLLaMA · 5d ago

The author benchmarked llama.cpp flags on a hybrid GPU setup with an RTX 4090 laptop and AMD XTX 7900 eGPU, achieving 70% faster generation, 40% faster prefill, and discovering a bug related to MTP in multi-GPU configurations.

0 favorites 0 likes
#egpu

New strix halo box: GMKtec EVO-X3, superior cooling to avoid thermal throttling, $3,600

Reddit r/LocalLLaMA · 2026-07-06

GMKtec launches the EVO-X3 mini PC featuring AMD Ryzen AI Max 395, advanced cooling, USB4, and OCuLink port for eGPU, priced at $3,600.

0 favorites 0 likes
#egpu

Scrambling to max StrixHalo (+NVLink dual eGPU 3090 mod)

Reddit r/LocalLLaMA · 2026-05-22

A user details their modding and benchmarking of an AMD Strix Halo system with dual RTX 3090 eGPUs and NVLink, finding improvements in LLM inference speed for dense models, especially with vLLM, and discusses power efficiency trade-offs.

0 favorites 0 likes
#egpu

You can do CUDA inference on an Apple Silicon Mac with PCI Passthrough

Reddit r/LocalLLaMA · 2026-05-08 Cached

This article explores the feasibility of using an external NVIDIA RTX 5090 GPU with an Apple Silicon Mac via Thunderbolt for CUDA inference and gaming, covering methods like tinygrad eGPU drivers and PCI passthrough to a Linux VM.

0 favorites 0 likes
← Back to home

Submit Feedback