Grok-4.5 on par with gpt-5.5-xhigh in coding at half the cost
Summary
Grok-4.5 achieves coding performance comparable to GPT-5.5-xhigh while costing half as much.
Similar Articles
@rohanpaul_ai: Grok 4.6 beat GPT-5.6 Sol on agentic loop efficiency, spending $13.11 versus $20.18 across the same 3 builds. Grok 4.6'…
The article reports an experiment comparing Grok 4.6 and GPT-5.6 Sol on agentic loop efficiency for coding tasks, showing Grok 4.6 is more cost-effective with fewer model calls and effective prompt caching.
Grok 4.6 Edges Out GPT 5.6 Sol Pro On SimpleBench
Grok 4.6 reportedly outperforms GPT 5.6 Sol Pro on the SimpleBench benchmark, signaling a notable shift in AI model capabilities.
@elonmusk: Grok Build
Grok 4.5 with Grok Build achieved #1 on the SWE-Atlas-QnA benchmark with a score of 84, matching GPT-5.6 Codex and outperforming other coding setups.
@gregpr07: Grok 4.6 is within 1 point of Opus 5 at ~40% lower cost. Across 7 runs, it solved 105/106 hard browser tasks. A little …
Grok 4.6 achieves near-Opus 5 performance at a 40% lower cost, solving 105/106 hard browser tasks, indicating potential for further advancement with reinforcement learning.
We made Grok 4.5, GPT-5.5, and Claude build the same apps
This article benchmarks Grok 4.5, GPT-5.5, Claude Opus 4.8, and Claude Fable 5 by having each model build three interactive apps (3D Rubik's Cube, particle gravity sandbox, Breakout game) from a single prompt, comparing their one-shot coding capabilities.