llama-swap

Tag

Cards List
#llama-swap

RetroCraft - Qwen 3.8 27B Q8, one shot with exact performance data on dual 3090s.

Reddit r/LocalLLaMA · 5d ago

A demonstration of Qwen 3.8 27B Q8 model's performance on dual NVIDIA 3090 GPUs, creating a Minecraft clone named RetroCraft in a single prompt with detailed token processing and generation speeds.

0 favorites 0 likes
#llama-swap

@iluciddreaming: Played with local LLMs for two months. Extensively tested various open-source models using Windows 11 + llama.cpp + llama-swap. Here is my final report card: Hardware: i7-13700 + 64GB RAM + RTX 4070. The best combination currently is gemm…

X AI KOLs Timeline · 2026-06-15 Cached

After two months of local LLM testing, the author finds that the combination of gemma-4-12B-it-QAT and MTP assistance performs best in speed and usability, with hardware i7-13700 + 64GB RAM + RTX 4070.

0 favorites 0 likes
#llama-swap

@leopardracer: https://x.com/leopardracer/status/2055341758523883631

X AI KOLs Timeline · 2026-05-15 Cached

A user shares their experience setting up a dual-GPU local AI lab with RTX 4080 Super and 5060 Ti, running Qwen 3.6 models via llama.cpp and llama-swap to reduce API costs and enable unrestricted experimentation.

0 favorites 0 likes
← Back to home

Submit Feedback