qwen

Tag

Cards List
#qwen

Did Alibaba abandon 35B A3B?

Reddit r/LocalLLaMA · 20h ago

The article questions whether Alibaba has abandoned its 35B A3B MoE model, noting the absence of new small MoE model announcements alongside the Qwen 3.8 release.

0 favorites 0 likes
#qwen

Ngram and world knowledge - why are we just building a coding model?

Reddit r/LocalLLaMA · 22h ago

The author discusses the need for AI models with better world knowledge, leveraging N-gram technology to fit more knowledge into smaller models, and questions why development focuses more on coding capabilities than broader world knowledge.

0 favorites 0 likes
#qwen

MiMo-V2.6 distilled themselves into Qwen 9B!

Reddit r/LocalLLaMA · yesterday

MiMo-V2.6 has been distilled into Qwen 9B, creating a more efficient version of the Qwen model released on Hugging Face.

0 favorites 0 likes
#qwen

@ChrisGPT: Correct. People often say open models are six months behind the frontier. But “open” is too broad, - a model that needs…

X AI KOLs Following · yesterday Cached

The article discusses the pattern where open-source AI models like Qwen match frontier capabilities about two quarters later, exemplified by GPT-5 and Qwen3.5-27B, and questions if this trend will continue with models like Astra and Fable 5.1.

0 favorites 0 likes
#qwen

Hemmingway-1: An AI that writes like a person (Qwen3.8-27B finetune)

Reddit r/LocalLLaMA · yesterday

Hemmingway-1 is a finetuned version of the Qwen3.8-27B AI model, designed to generate text in a more human-like writing style.

0 favorites 0 likes
#qwen

@Lonely__MH: Hesitantly asking, Qwen-Image-2.1 local deployment How's the image generation speed?!

X AI KOLs Following · 2d ago Cached

Discussing the image generation speed of Qwen-Image-2.1 local deployment, congratulating its release, and highlighting its breakthroughs in small parameter counts and high efficiency, positioning it as a potential leader in domestic AI image generation.

0 favorites 0 likes
#qwen

Hemmingway-1, an Apache-2.0 27B creative-writing fine-tune (Qwen3.8-27B base, EQ-Bench 4 1330)[R]

Reddit r/MachineLearning · 2d ago

A small lab from Switzerland and South Africa has open-sourced Hemmingway-1, a 27B Apache-2.0 licensed fine-tune of Qwen3.8-27B specialized for creative writing, achieving high scores on EQ-Bench and internal benchmarks.

0 favorites 0 likes
#qwen

Qwen model training new model without being told to do it.

Reddit r/artificial · 3d ago

A Qwen model autonomously created and trained a new AI model to improve translation tasks after being asked to fix bugs, demonstrating unexpected self-improvement capabilities.

0 favorites 0 likes
#qwen

Qwen3.8 27B one shot prompt “Super Mario” dupe

Reddit r/LocalLLaMA · 3d ago

Qwen3.8 27B model is used in one-shot prompts to generate a Super Mario Bros clone, demonstrating its coding and generative capabilities.

0 favorites 0 likes
#qwen

Anthropic just named seven Chinese AI labs for stealing Claude's reasoning. I read the full report and the Qwen part doesn't hold up the way the headline does.

Reddit r/LocalLLaMA · 4d ago

The article critiques Anthropic's report accusing seven Chinese AI labs of illicitly distilling Claude's reasoning, highlighting weak evidence for Alibaba's Qwen and suggesting political framing in the accusations.

0 favorites 0 likes
#qwen

Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro

Reddit r/LocalLLaMA · 4d ago

Inco Splash is an open-source inference engine optimized for Apple silicon, offering significant speed improvements for running AI models like Qwen3.8-27B on M-series MacBooks.

0 favorites 0 likes
#qwen

Question: UkisAI Swift Ternary Bonsai 2 27B?

Reddit r/LocalLLaMA · 4d ago

Jovan from UkisAI discusses improvements in their Swift Qwen3.8 27B model and seeks community feedback on creating a Swifted version of Bonsai 2 to address overthinking loops and high token usage.

0 favorites 0 likes
#qwen

LoRA on abliterated Qwen 3.8-27B for internal codebase recall[P]

Reddit r/MachineLearning · 4d ago

A LoRA adapter was trained on an abliterated Qwen 3.8-27B model to enhance internal codebase recall, demonstrating superior performance over Claude models on private-repo-specific tasks in evaluations.

0 favorites 0 likes
#qwen

US government website used AI search tool (Qwen) from China that FBI said copied Anthropic

Reddit r/LocalLLaMA · 4d ago

A US government website allegedly used the AI search tool Qwen from China, which the FBI claims copied Anthropic's technology, leading to legal concerns.

0 favorites 0 likes
#qwen

llama.cpp expert-pool fork for Qwen 3.8 Flash next IQ4 + 16Gb VRAM tested on MI50 gfx906 with parameters

Reddit r/LocalLLaMA · 5d ago Cached

An unofficial fork of llama.cpp introduces a persistent expert pool for MoE models, optimized to reduce expert re-copies over PCIe on 16GB AMD gfx906 GPUs, thereby improving decode throughput for large context lengths.

0 favorites 0 likes
#qwen

@TheAhmadOsman: My current LLMs stable - GLM 5.3 Flash - DeepSeek V4.1 Flash - Qwen 3.8 Next Flash - Qwen 3.8 27B Just a year ago you w…

X AI KOLs Timeline · 5d ago Cached

A tweet highlighting the performance of various LLMs like GLM, DeepSeek, and Qwen, noting the rapid progress in AI capabilities over the past year.

0 favorites 0 likes
#qwen

Qwen3.8-Omni-Flash: Omni Senses. Agentic Delivery (18 minute read)

TLDR AI · 5d ago

Qwen3.8-Omni-Flash is an omnimodal AI model with a 1M-token context window, supporting text, image, audio, and video inputs, and achieving performance comparable to or better than Gemini 3.8 Flash, now available on the Qianwen AI Platform.

0 favorites 0 likes
#qwen

Alibaba releases Qwen 3.8 Omni Flash

Hacker News Top · 5d ago

Alibaba has released the Qwen 3.8 Omni Flash AI model, which likely features multimodal capabilities and is optimized for speed.

0 favorites 0 likes
#qwen

Call for compute - help optimize Qwen inference speed on local hardware

Reddit r/LocalLLaMA · 5d ago

The user has renamed a GitHub repository to HyperQwen to focus on optimizing Qwen model inference speeds on local hardware and is seeking testers with 4090s and 5090s GPUs for both Windows and Linux.

0 favorites 0 likes
#qwen

First M5 Ultra benchmarks

Reddit r/LocalLLaMA · 5d ago

Unofficial benchmarks for the M5 Ultra chip show promising inference speeds with the Qwen 3.8 27B q4 model, achieving 50 tokens per second for threading and 1800 tokens per second for prefill at 8k context.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback