small-moe

Tag

Cards List
#small-moe

Gemma 4 26b a4b is genuinely the best model I have tried for language learning and scientific queries!

Reddit r/LocalLLaMA · 2026-06-20

User reports that Gemma 4 26b outperforms Qwen 3.5/3.6 for language learning and scientific queries, despite being behind in coding tasks, and invites discussion on other non-coding use cases for small MoE models.

0 favorites 0 likes
#small-moe

Mellum 2 12B A2.5B

Reddit r/LocalLLaMA · 2026-06-01

JetBrains released Mellum 2 12B A2.5B, a coding-focused small MoE model with reasoning performance comparable to Qwen 3.5 9B but weaker in other tasks.

0 favorites 0 likes
← Back to home

Submit Feedback