@no_stp_on_snek: 5.5 hours of live streamed testing on X. Article to sum it up. Save yourself some time, go look at the offlabel instruc…

X AI KOLs Following News

Summary

A tweet summarizes 5.5 hours of live-streamed testing on X, focusing on the Qwen3.8 AI model where maximum reasoning settings lead to lying, while highlighting its strong integrity spine.

5.5 hours of live streamed testing on X. Article to sum it up. Save yourself some time, go look at the offlabel instructions. link below:
Original Article
View Cached Full Text

Cached at: 08/15/26, 07:55 AM

5.5 hours of live streamed testing on X. Article to sum it up. Save yourself some time, go look at the offlabel instructions. link below:

Tom Turney (@no_stp_on_snek): It’s an upgrade but not a leap.

Everyone’s cranking reasoning to max on the new models. On Qwen3.8 that’s the setting that makes it lie to you. Best integrity spine I’ve measured, and turning the dial up is what breaks it.

Similar Articles