@no_stp_on_snek: 5.5 hours of live streamed testing on X. Article to sum it up. Save yourself some time, go look at the offlabel instruc…
Summary
A tweet summarizes 5.5 hours of live-streamed testing on X, focusing on the Qwen3.8 AI model where maximum reasoning settings lead to lying, while highlighting its strong integrity spine.
View Cached Full Text
Cached at: 08/15/26, 07:55 AM
5.5 hours of live streamed testing on X. Article to sum it up. Save yourself some time, go look at the offlabel instructions. link below:
Tom Turney (@no_stp_on_snek): It’s an upgrade but not a leap.
Everyone’s cranking reasoning to max on the new models. On Qwen3.8 that’s the setting that makes it lie to you. Best integrity spine I’ve measured, and turning the dial up is what breaks it.
Similar Articles
@no_stp_on_snek: Some of the latest findings for qwen 3.8 27b: https://x.com/i/broadcasts/1MJgNbbqqAbGL…
A live broadcast by Tom Turney presents the latest behavioral testing findings for the Qwen 3.8 27b AI model.
@no_stp_on_snek: Just a few hours away from Qwen 3.8! Clear your benches! I’ll be working on behavioral tests and comparisons against 3.…
Qwen3.8-27B is a new AI model with enhanced capabilities in coding, agentic tasks, and vision-language understanding, offering flexible thinking control and long context lengths. It is available on Hugging Face and designed for deployment-friendly use.
@no_stp_on_snek: 200 folks chilling and following along qwen 3.8 27b testing. No talking at you just chill music and results across my f…
A live broadcast on X featuring Tom Turney and participants testing the Qwen3.8 AI model with chill music and sharing results.
@no_stp_on_snek: https://subq.mildlyconcerning.com
This article critically analyzes the claims and timeline of the subQ long-context AI technique, highlighting discrepancies and walkbacks from the original announcement.
@no_stp_on_snek: one last thing: the real downside i found testing Ornith-1.0 (the new agentic coder): it over-gates legitimate work. on…
A tester reports that the new Ornith-1.0 agentic coder model over-gates legitimate work by demanding excessive prerequisites, a trade-off from its cautious training, while stock Qwen3.6 executes simple tasks directly.