A hunch: Qwen3.8-27B's general knowledge got pruned (good, if true)

Reddit r/LocalLLaMA News

Summary

The author shares a hunch that Qwen3.8-27B has pruned general knowledge to improve coding and agentic skills, based on reduced knowledge of a specific German town compared to earlier Qwen models.

I'm always testing an image prompt with a picture of a historic place in my hometown – a small but well known 250,000 people town in Germany. I'll just ask the model, in which City this photo has been taken. With the 3.6 generation of both the 27B and the 35B A3B variants, the models sometimes got the right answer and sometimes they didn't. So the signal for this particular knowledge was already weak. The 35B variant got it right more often but at least, the models reasoning showed my City most of the times, even if it hallucinated the wrong final answer. Both models could be easily nudged to the right answer with a few hints and then produced some little extra insight about the history or scene and its surroundings, that was mostly true. Qwen3.8-27B on the other hand barely knows the city at all and has absolutely no clue about related popular, historic facts regarding the scenery or the surrounding buildings. Nudging isn't very fruitful as well and if told the real name of the city, reasoning shows, that the model only agrees, because the user says so. I have the feeling, that Qwen labs maybe pruned useless general knowledge for more coding knowledge and agentic skill. All models ud q4_k_xl variants, image-min-tokens 2048, with and without reasoning. Anyone else with this feeling? Disclaimer: My hunch could be very well absolute bullshit. Sample size way to low and methodically sloppy af.
Original Article

Similar Articles

Qwen3.8-27B took a serious hit to *knowledge* vs 3.6

Reddit r/LocalLLaMA

The author finds that Qwen3.8-27B has weaker knowledge recall compared to Qwen3.6 based on personal benchmarks and offline tests, suggesting it may not be suitable for airgapped knowledge retrieval.

Qwen 3.8 27B's existence raises questions

Reddit r/singularity

The article questions how the relatively small Qwen 3.8 27B model achieves high intelligence, raising doubts about current scaling laws and the efficiency of large model parameters.

Qwen3.6-27B-GGUF is here!

Reddit r/LocalLLaMA

Community GGUF release of Qwen’s 27B hybrid-architecture model with 262k context, multimodal inputs, tool calling and "Thinking Preservation" for agentic coding.

Qwen 3.8 27B - Fantastic German capabilities

Reddit r/LocalLLaMA

The author is highly impressed with Qwen 3.8 27B's German translation capabilities, noting its superior vocabulary, sentence structure, and punctuation compared to other frontier models like GPT-5.6 and Fable 5.