Tag
This article compares Qwen 3.8 to 3.6, noting that Qwen 3.8 reduces reasoning loops on low settings and includes a preserve_thinking parameter to avoid redundant reasoning.
Google's Gemma 4 31B IT model now has a chat template fix that preserves thinking and improves null handling, reasoning preservation, and input validation.