@atomic_chat_hq: Qwen 3.7-max beats Opus 4.7 and GPT-5.5 We tested three frontier models on a real agentic task: write a Tetris bot that…

X AI KOLs Timeline News

Summary

Qwen 3.7-max outperformed Opus 4.7 and GPT-5.5 on an agentic Tetris bot task, achieving the largest performance improvement at the lowest cost.

Qwen 3.7-max beats Opus 4.7 and GPT-5.5 We tested three frontier models on a real agentic task: write a Tetris bot that plays the game and trains itself. Each model could read its own code, run benchmarks, and rewrite itself across 10 iterations. Then we compared the final bots head to head. Qwen 3.7-Max: training cost $1.32, bot improvement +56% Claude Opus 4.7: training cost $12.15, bot improvement +28% GPT-5.5: training cost $2.85, bot improvement +7% Qwen won on every dimension - biggest jump, 9× cheaper than Claude, 2× cheaper than GPT. Long agentic loops is where Qwen Max actually delivers.
Original Article

Similar Articles

Quoting Thariq Shihipar

Simon Willison's Blog

Claude Code version 2.1.277 adds support for AGENTS.md as an alternative to CLAUDE.md, allowing customizable project instructions through mods.