qwen-3

Tag

Cards List
#qwen-3

GPT-6.1 Sol is cheap. We made it 77% cheaper by never letting it write code

Reddit r/AI_Agents ↗ · 20h ago

A practical test demonstrates that using GPT-6.1 Sol as an orchestrator without write access and Qwen 3.8 27B as workers reduces costs by 77% but increases execution time for small tasks in AI agent setups.

0 favorites 0 likes
#qwen-3

What is Qwen 3.8 Next Engram usage?

Reddit r/LocalLLaMA ↗ · 2026-08-28

This article examines the Engram component in the Qwen 3.8 Next AI model, detailing it as a 51B-parameter Zipfian cache with highly skewed access patterns, and evaluates optimization through frequency pruning for compression.

0 favorites 0 likes
#qwen-3

We quantized Qwen 3.8 27B and compared the quants on an RTX 6000

Reddit r/LocalLLaMA ↗ · 2026-08-23

The team quantized Qwen 3.8 27B into various GGUF formats and benchmarked them on an RTX 6000, finding similar performance across quants with AD-Q6_K recommended for safety.

0 favorites 0 likes
#qwen-3

Claude sonnet 4.6 was really good at estimating the future qwen 3.8 27b performance

Reddit r/LocalLLaMA ↗ · 2026-08-20

A user shared how Claude 3.5 Sonnet accurately estimated the future performance of Qwen 3.8 27B by extrapolating from earlier model differences, with benchmarks matching closely.

0 favorites 0 likes
#qwen-3

Un modello 100% locale sul tuo smartphone!

Reddit r/artificial ↗ · 2026-07-05

A developer shares their success in fine-tuning Qwen 3 models (1.5B and 4B) for local use on smartphones, with a downloadable APK that works offline, and plans for a Windows version.

0 favorites 0 likes
#qwen-3

A barebones CPU-only inference engine for Qwen 3, written from scratch in pure C

Reddit r/LocalLLaMA ↗ · 2026-06-28

A minimal CPU-only inference engine for Qwen 3 models implemented from scratch in pure C.

0 favorites 0 likes
#qwen-3

@zhixianio: Finished testing, feeling quite surprised, not sure if I'm using it wrong. Feel free to provide counterexamples. Here are my results: On M5 Max, pitting this community fine-tuned gemma-4-12B-coder (llama.cpp) against my daily driver Qwen3.6-35B-…

X AI KOLs Timeline ↗ · 2026-06-15 Cached

The user tested the community fine-tuned gemma-4-12B-coder against Qwen3.6-35B-A3B MoE on three programming tasks, finding that gemma performed poorly on complex stateful programs, while Qwen 35B remained robust.

0 favorites 0 likes
#qwen-3

Gemma 4 31B's competence surprised me

Reddit r/LocalLLaMA ↗ · 2026-06-09

A user shares anecdotal findings that Gemma 4 31B outperforms Qwen 3.6 models and matches Opus 4.7 in understanding and refactoring messy academic code, highlighting a benchmark (SciCode) where Gemma excels.

0 favorites 0 likes
← Back to home

Submit Feedback