qwen-models

Tag

Cards List
#qwen-models

Qwen, where's the small stuff? (1B/2B/4B)

Reddit r/LocalLLaMA ↗ · 5d ago

The article questions why Qwen hasn't released new small-scale models (1B/2B/4B), which affects accessibility and development for hardware-limited users, and mentions a similar issue with Google's Gemma series.

0 favorites 0 likes
#qwen-models

How long can I expect to wait until the local ~30B A3B frontier catches up to GLM 5.3 Flash quality?

Reddit r/LocalLLaMA ↗ · 2026-09-25

The user queries whether local AI models around 30B parameters can achieve GLM 5.3 Flash quality within a year, given current hardware constraints like 16GB RAM and 8GB VRAM.

0 favorites 0 likes
#qwen-models

Putting the question before the context took my local Qwen from 89% to 100% on a decision benchmark, and from ~400 ms to ~80 ms

Reddit r/LocalLLaMA ↗ · 2026-09-21

Switching the order of question and context in prompts for local Qwen models improved accuracy from 89% to 100% and reduced latency from ~400 ms to ~80 ms on a decision benchmark.

0 favorites 0 likes
#qwen-models

What's the best model you have running on strix halo 128GB?

Reddit r/LocalLLaMA ↗ · 2026-09-16

The author shares their experience running local AI models on a Framework desktop with Strix Halo and 128GB unified memory, preferring Qwen models for coding, and asks for recommendations on better hardware utilization.

0 favorites 0 likes
#qwen-models

GPT Live clone on an RTX 3060

Reddit r/LocalLLaMA ↗ · 2026-09-11

A developer tested their local voice assistant project, Fulloch, which uses Qwen models on an RTX 3060 to replicate GPT Live functionality, showing impressive performance with open-source tools.

0 favorites 0 likes
#qwen-models

EnvCraft: Synthesizing Executable Environments in Agentic RL for Claw-like Agent

arXiv cs.AI ↗ · 2026-09-10 Cached

EnvCraft is an automated framework for synthesizing executable environments to address the scarcity in Agentic RL training, showing significant performance gains on claw-like and general tool-use benchmarks using Qwen models.

0 favorites 0 likes
#qwen-models

@Darkolorin: Today we are releasing our speculative decoding implementation in Uzu. Biggest release since inception of our lab. Init…

X AI KOLs Timeline ↗ · 2026-09-03 Cached

Release of a speculative decoding implementation in Uzu, initially supporting Qwen3.6 27B with upcoming support for Qwen3.8 27B and Muse Glimmer.

0 favorites 0 likes
#qwen-models

Qwen3.8-Flash-Next-NVFP4 vs Qwen3.8-27B-FP Test Results

Reddit r/LocalLLaMA ↗ · 2026-08-31

This article presents detailed test results comparing the performance of Qwen3.8-Flash-Next-NVFP4 and Qwen3.8-27B-FP8 AI models across various tasks, highlighting that Flash-Next is faster with fewer failures but struggles with multi-step symbolic work.

0 favorites 0 likes
#qwen-models

dgx sparks and new models my tests and results

Reddit r/LocalLLaMA ↗ · 2026-08-29

This article presents test results for AI models like DeepSeek V4 Flash and Qwen3.8 on NVIDIA DGX Sparks hardware, detailing performance metrics, context lengths, and benchmark scores with operational insights.

0 favorites 0 likes
#qwen-models

Qwen3.8-27B vs Qwen3.8-Flash-Next smaller quant?

Reddit r/LocalLLaMA ↗ · 2026-08-29

A user compares Qwen3.8-27B and Qwen3.8-Flash-Next models for intelligence and coding performance with 128GB RAM, seeking advice on which is better.

0 favorites 0 likes
#qwen-models

@VraserX: Qwen3.8-Flash reportedly needs around ONE NINTH the training cost of Qwen3.7-Plus. This is the trend I think people und…

X AI KOLs Following ↗ · 2026-08-28 Cached

A tweet reports that the Qwen3.8-Flash model requires about one-ninth the training cost of the Qwen3.7-Plus model, highlighting a trend of decreasing AI training expenses.

0 favorites 0 likes
#qwen-models

@no_stp_on_snek: Check out Buun's work, he cookin.

X AI KOLs Following ↗ · 2026-08-19 Cached

A user highlights Buun's work on optimizing AI models, achieving high-speed inference of Qwen 3.6 on a single 3090 GPU and developing DFlash2 for Qwen 3.8.

0 favorites 0 likes
#qwen-models

Birds Don't Fly Like Planes. Neither Does AI. (4 minute read)

TLDR AI ↗ · 2026-08-19 Cached

The article compares local AI models like Qwen3.8-27B with cloud models, showing that smaller models can achieve similar performance through different reasoning processes, with trade-offs in speed and token usage.

0 favorites 0 likes
#qwen-models

Qwen3.8-27B vs Qwen3.6-27B writing ray-tracers in BASIC

Reddit r/LocalLLaMA ↗ · 2026-08-16

A hobbyist compares Qwen3.8 and Qwen3.6 AI models in generating ray-tracing code in BASIC, finding that Qwen3.8 iterates to better results independently.

0 favorites 0 likes
#qwen-models

Survival of the Fitted: Qwen3.6-27B’s Jacobian lens reads and steers Qwen3.8-27B with zero refitting [R]

Reddit r/MachineLearning ↗ · 2026-08-15

This article tests the transferability of a Jacobian interpretability lens from Qwen3.6-27B to Qwen3.8-27B, finding that it can read and steer the newer model with zero refitting for specific tasks.

0 favorites 0 likes
#qwen-models

peculiar-ragdoll/Qwen-Sharp-Chat-Templates

Hugging Face Models Trending ↗ · 2026-08-10 Cached

This article describes a drop-in chat template fix for Qwen AI models that improves accuracy, reduces token usage, and speeds up responses for knowledge work and coding tasks.

0 favorites 0 likes
#qwen-models

Models and Quants quality test results - the chessboard svg (Qwen3.6 27B/35B-A3B/Zaya1)

Reddit r/LocalLLaMA ↗ · 2026-05-12

Community testers evaluate quantized versions of Qwen3.6, ZAYA1, and other models for SVG chessboard generation accuracy using local inference frameworks like MLX.

0 favorites 0 likes
← Back to home

Submit Feedback