rtx-3090

Tag

Cards List
#rtx-3090

RTX 3090 EBay Pricing is Crazy!!

Reddit r/LocalLLaMA · 2026-06-06

The author observes that used RTX 3090 GPUs are being sold on eBay for $1,300-$1,500, higher than a new 3090 Ti purchased 5 years ago, and questions why people buy older used GPUs at such high prices for AI rigs.

0 favorites 0 likes
#rtx-3090

@TheAhmadOsman: You should buy an RTX 3090 and learn how to run models locally The elite don’t want you to know this but running local …

X AI KOLs Following · 2026-05-24 Cached

A tweet recommending users buy an RTX 3090 to run AI models locally, claiming it is easy, performant, and affordable.

0 favorites 0 likes
#rtx-3090

Scrambling to max StrixHalo (+NVLink dual eGPU 3090 mod)

Reddit r/LocalLLaMA · 2026-05-22

A user details their modding and benchmarking of an AMD Strix Halo system with dual RTX 3090 eGPUs and NVLink, finding improvements in LLM inference speed for dense models, especially with vLLM, and discusses power efficiency trade-offs.

0 favorites 0 likes
#rtx-3090

@TheAhmadOsman: Gentle reminder that all you need to start with Local AI is: - 2x RTX 3090s (pick up for $700-$900 on r/hardwareswap) -…

X AI KOLs Timeline · 2026-05-19 Cached

A reminder that two RTX 3090s and open-source models like Qwen 3.6 27B or Gemma 4 31B can run powerful local AI agents, comparable to Opus 4.5, using tools like Claude Code and self-hosted SearXNG.

0 favorites 0 likes
#rtx-3090

Benchmarking the new b9200 update: Optimizing Qwen 3.6 27B mtp for Hermes Agent on a single RTX 3090

Reddit r/LocalLLaMA · 2026-05-18

Benchmarking the b9200 update of llama.cpp with optimized flags for Qwen 3.6 27B MTP on a single RTX 3090 shows significant performance gains, especially in prompt processing speed, for agentic workflows.

0 favorites 0 likes
#rtx-3090

@Snixtp: https://x.com/Snixtp/status/2055734339346768225

X AI KOLs Timeline · 2026-05-16 Cached

A user benchmarks the MTP variant of Qwen3.6 27B against the normal version on a single RTX 3090 using llama.cpp, finding MTP offers up to 2.37x faster generation at long contexts (32k-64k) but with slower prefill and no concurrency support yet.

0 favorites 0 likes
#rtx-3090

Finding the 4x 3090 Sweet Spot

Reddit r/LocalLLaMA · 2026-05-15

A user shares power limit testing on a 4x RTX 3090 setup running Qwen3.6-27B with vLLM, finding 220W as the sweet spot for peak efficiency with minimal throughput loss.

0 favorites 0 likes
#rtx-3090

@pupposandro: PFlash now run @poolsideai's Laguna-XS.2 (33B-A3B MoE) on a single RTX 3090. - 111 tok/s decode @ short ctx - 128K TTFT…

X AI KOLs Following · 2026-05-14 Cached

PFlash now supports running @poolsideai's Laguna-XS.2 (33B-A3B MoE) on a single RTX 3090, achieving 111 tok/s decode and 5.4x faster prefill than llama.cpp, with NIAH passes up to 131K context.

0 favorites 0 likes
#rtx-3090

@Snixtp: More efficiency tests on a single 3090 TL;DR: - I tested 8 local LLMs on a single RTX 3090, power limit from 100W to 45…

X AI KOLs Following · 2026-05-08

The article presents benchmark results for 8 local LLMs on an RTX 3090, showing that power efficiency peaks around 225W, with diminishing returns at maximum power.

0 favorites 0 likes
← Previous
← Back to home

Submit Feedback