@Youssofal_: Qwen 3.8 27B vs Fable, Kimi K3, GPT 5.6 Sol Pro on Flappy Bird. Prompt: “Make the ultimate adorable cute and beautiful …
Summary
This article compares the performance of various AI models, including Qwen 3.8 27B, Fable, Kimi K3, and GPT 5.6 Sol Pro, in generating a Flappy Bird game in HTML, highlighting differences in execution time and quality.
View Cached Full Text
Cached at: 08/26/26, 03:34 AM
Qwen 3.8 27B vs Fable, Kimi K3, GPT 5.6 Sol Pro on Flappy Bird.
Prompt: “Make the ultimate adorable cute and beautiful flappy bird game in HTML.”
Qwen 3.8 27B Xhigh INT8:
- Video 1 (Puff)
- 8m 11s
- 160 TPS (2x3090)
Fable 5 Max:
- Video 2 (Flappy Fluff)
- 20m 21s
Kimi K3 Max Fireworks Fast:
- Video 3 (Fluffy Flap!)
- 4m 40s
- 111.3 TPS
GPT 5.6 Sol Pro:
- Video 4 (Flappy Boom)
- 19m 7s
Overview:
Personally, I think Fable is the best. It feels extremely polished, sharp and stands out with small things like the dizzy effect on death. However, 27B is freakishly close. Never would I have guessed 2 months ago this would have come from a local model. 27B actually beats Fable on certain aspects like the 3D tilting ui buttons on death.
Kimi K3 felt the best to play mechanics wise. It felt smooth with the perfect balance of difficulty and ease to play. Also, it was the only one that had music.
5.6 Sol Pro was hilariously bad in all aspects. I tried 5.6 xhigh multiple times and the outputs were even worse.
Similar Articles
(Interactive)OpenCode Racing Game Comparison Qwen3.6 35B vs Qwen3.5 122B vs Qwen3.5 27B vs Qwen3.5 4B vs Gemma 4 31B vs Gemma 4 26B vs Qwen3 Coder Next vs GLM 4.7 Flash
An informal benchmark comparing 8 AI models (Qwen3.6 35B, Qwen3.5 series, Gemma 4 series, GLM 4.7 Flash) in creating racing games via OpenCode/Playwright MCP, testing their coding agent capabilities and documenting various implementation quirks.
ArenaAI Kimi K3 comparison to Fable and Sol
A comparison video of the AI models Kimi K3, Fable, and Sol on ArenaAI.
GPT-5.6 Sol vs Fable 5 - mobile app design
The article compares the performance of GPT-5.6 Sol and Fable 5 AI models in mobile app design tasks using the same prompt, asking readers which model they prefer.
Qwen-3.8-27B, Nemotron-3.5-Lightning-30B-A3B, Ornith-1.5-35B-A3B, Muse-Glimmer-30B oQ8e comparison
A comparison of multiple AI models including Qwen, Nemotron, Ornith, and Muse-Glimmer on benchmarks, with Ornith performing well and TielCoder showing potential in coding tasks.
@rohanpaul_ai: Qwen 3.8 Max just built better 3D physics scenes than Fable 5 while costing about 7x less to run. Test was done by @ato…
A test by atomic.chat shows Qwen 3.8 Max outperforming Fable 5 at generating self-contained 3D physics scenes while costing about 7x less per run.