Weirdly, no one talks about Temperature setting for the Qwen3.8 27b

Reddit r/LocalLLaMA News

Summary

The post discusses the default temperature setting of 1.0 for Qwen3.8 27b and suggests setting it to 0.7 to reduce unnecessary reasoning, questioning the impact on capabilities and optimal settings for various tasks.

Mind you, it is 1.0 by default, yet everyone is focused on how much the new model thinks, restricting the reasoning budget and/or dropping the reasoning level. Set the temperature to 0.7 and the model will no longer write a whole book of thoughts before trying to make a small edit in the file. The question is - how much does this affect the model's capabilities? What is the sweet spot for the temp parameter for various tasks?
Original Article

Similar Articles

Optimizing Qwen 3.6 35B A3B sampling parameters.

Reddit r/LocalLLaMA

A researcher seeks faster, lower-variance benchmarks to tune temperature, top_p, top_k and min_p for Qwen 3.6 35B A3B, estimating months of 3090-time with current setups.

Qwen3.8: much thinking for nothing

Reddit r/LocalLLaMA

A user expresses dissatisfaction with the Qwen3.8 27B model, criticizing its tendency to overthink and overreach, which wastes time and context during tasks, and asks for community feedback on practical usage.