26b

Tag

Cards List
#26b

DiffusionGemma 26b on a 4090 at up to 475t/s... and some thoughts...

Reddit r/LocalLLaMA · 2026-06-18

A user shares their experience running DiffusionGemma 26B on a 4090 GPU via vLLM, achieving up to 475t/s but noting drawbacks like single-user limitation, lower accuracy, and short context, concluding it's not worth using over the regular 26B model.

0 favorites 0 likes
#26b

DiffusionGemma 26B A4B results on my 5090

Reddit r/LocalLLaMA · 2026-06-11

This post presents benchmark results and tuning parameters for running DiffusionGemma 26B A4B GGUF models on an RTX 5090 GPU, showing up to 44% speedup via optimized temperature settings and quantization choices.

0 favorites 0 likes
#26b

Thoughts on Gemma4 12b vs 26a4b, which one is better?

Reddit r/LocalLLaMA · 2026-06-08

Discussion comparing Gemma4 12b and 26a4b variants, focusing on creative tasks like writing and chatting.

0 favorites 0 likes
#26b

@0x0SojalSec: SUPER GEMMA 4 26B UNCENSORED GGUF v2 IS INSANE, - 0/100 refusals (actually uncensored) - Fixed all the tool-call + toke…

X AI KOLs Following · 2026-06-07 Cached

Super Gemma 4 26B Uncensored GGUF v2 is a community fine-tuned model offering uncensored responses with zero refusals, improved speed, and fixed tool-calling, optimized for local inference on llama.cpp and vLLM.

0 favorites 0 likes
← Back to home

Submit Feedback