@vmiss33: I installed Hermes Agent on Windows, and set it up with GPT 5.5. I gave it one of @above_spec's amazing twitter threads…
Summary
The user shares a report on successfully running the Qwen3.6 35B A3B model on Windows using Hermes Agent and an 8GB VRAM GPU.
View Cached Full Text
Cached at: 05/13/26, 12:19 PM
I installed Hermes Agent on Windows, and set it up with GPT 5.5. I gave it one of @above_spec’s amazing twitter threads on Qwen3.6 35B A3B model running on 8GB of VRAM and told it to get it working annnnd… @NousResearch https://t.co/yPNJDptR9m
Similar Articles
Qwen3.8 flash next + exllamav3 + hermes is amazing
User shares their positive experience using Qwen3.8 model with ExLlamaV3 on 6x3090 GPUs, achieving 80-120 tokens per second, and controlling Hermes agent via Matrix for daily use.
@svpino: Hermes with Gemma 4 or Qwen 3.5 is literally the best combo you can run locally on your computer. You've got to give th…
Developer claims Hermes fine-tunes of Gemma 4 and Qwen 3.5 deliver the best local LLM performance, suggesting they rival paid BigAI models.
@analogalok: I just got Gemma 4 26B A4B MoE model running fully locally with Hermes agent on an 8GB RTX 4060 and it's now backtestin…
A developer demonstrates running Gemma 4 26B MoE model locally on an 8GB RTX 4060 with Hermes agent to fully automate backtesting of trading strategies, highlighting the growing capability of local LLMs as autonomous agents.
@ivanfioravanti: Hermes Agent + Computer Use by @trycua is pretty cool! Looking at Hermes interacting with apps and windows is mind blow…
A tweet showcases Hermes Agent using computer control on a Mac with MiniMax M3 and Reachy Mini, demonstrating AI interacting with apps.
@sudoingX: this is a laptop running a 31b parameter model at 99% gpu autonomously through hermes agent, 15 tok/s sustained, 22.8 o…
A 31B parameter model runs locally on a laptop via Hermes agent at 15 tok/s, using 22.8 GB VRAM and 94 W power, highlighting fully autonomous, private AI inference without cloud dependencies.