How Far Have We Come? Comparing LLMs: Sonnet 4 vs. GPT-5.5

Reddit r/ArtificialInteligence Models

Summary

The article compares Claude Sonnet 4 and GPT-5.5 by generating code for a Flappy Bird game, demonstrating significant progress in LLM capabilities over 1.5 years.

Claude Sonnet 4 vs Sonnet 5.5 medium effort for both. Free plan used for the outputs. In just a span of 1.5 years, the models have made an incredible amount of progress. Prompt Given Description: Create a simple side-scrolling game called "Flappy Bird." The main character is a small cartoon bird that automatically moves forward. The player taps or clicks to make the bird flap and stay in the air, trying to avoid hitting randomly spaced pipes. Game Features to Include: Colorful 2D background (sky, clouds, maybe a cityscape). Flappy bird character that flaps upward on tap/click and falls due to gravity. Rows of pipes with gaps, moving leftward across the screen. Score increases by 1 each time the bird passes through a gap. Game ends if the bird touches a pipe or the ground, then shows final score and a “Play Again” button. Fun sound effects for flapping, passing pipes, and crashing. Art Style: Cartoonish and cheerful. The bird looks cute and friendly. Pipes and backgrounds are bright, simple, and easy to understand. Controls: Tap or click = Flap (bird moves up for a moment). No other controls. Optional Extras: Show the top score. Simple animation for bird’s wings and pipe entry/exit. "Get Ready" message before game starts. Real-life Example: Think of the actual Flappy Bird mobile game that became super popular—fast to play, hard to master, and a little bit addictive. You tap your phone to keep the bird in the air and dodge green pipes. Actionable Tips: Focus on smooth, responsive tapping for controls. Keep graphics light and not cluttered. Make the first run a bit easier so players don’t get frustrated.
Original Article

Similar Articles