Where's the raccoon with the ham radio? (ChatGPT Images 2.0)

Simon Willison's Blog Models

Summary

OpenAI released ChatGPT Images 2.0, claiming a GPT-3-to-GPT-5 leap; Simon Willison benchmarks it with a "Where's Waldo"-style raccoon-and-ham-radio prompt against gpt-image-1, Google Nano Banana 2 and Pro, showing mixed hide-and-seek success.

No content available
Original Article
View Cached Full Text

Cached at: 04/22/26, 02:07 AM

# Where’s the raccoon with the ham radio? (ChatGPT Images 2.0) Source: [https://simonwillison.net/2026/Apr/21/gpt-image-2/](https://simonwillison.net/2026/Apr/21/gpt-image-2/) 21st April 2026 OpenAI[released ChatGPT Images 2\.0 today](https://openai.com/index/introducing-chatgpt-images-2-0/), their latest image generation model\. On[the livestream](https://www.youtube.com/watch?v=sWkGomJ3TLI)Sam Altman said that the leap from gpt\-image\-1 to gpt\-image\-2 was equivalent to jumping from GPT\-3 to GPT\-5\. Here’s how I put it to the test\. My prompt: > `Do a where's Waldo style image but it's where is the raccoon holding a ham radio` #### gpt\-image\-1 First as a baseline here’s what I got from the older gpt\-image\-1 using ChatGPT directly: [![There's a lot going on, but I couldn't find a raccoon.](https://static.simonwillison.net/static/2026/image_crop_1402x1122_w1402_q0.3.jpg)](https://static.simonwillison.net/static/2026/chatgpt-image-1-ham-radio.png) I wasn’t able to spot the raccoon—I quickly realized that testing image generation models on Where’s Waldo style images \(Where’s Wally in the UK\) can be pretty frustrating\! I tried[getting Claude Opus 4\.7](https://claude.ai/share/bd6e9b88-29a9-420b-8ac1-3ac5cebac215)with its new higher resolution inputs to solve it but it was convinced there was a raccoon it couldn’t find thanks to the instruction card at the top left of the image: > **Yes — there’s at least one raccoon in the picture, but it’s very well hidden**\. In my careful sweep through zoomed\-in sections, honestly, I couldn’t definitively spot a raccoon holding a ham radio\. \[\.\.\.\] #### Nano Banana 2 and Pro Next I tried Google’s Nano Banana 2,[via Gemini](https://gemini.google.com/share/3775db96c576): [![Busy Where's Waldo-style illustration of a park festival with crowds of people, tents labeled "FOOD & DRINK", "CRAFT FAIR", "BOOK NOOK", "MUSIC FEST", and "AMATEUR RADIO CLUB - W6HAM" (featuring a raccoon in a red hat at the radio table), plus a Ferris wheel, carousel, gazebo with band, pond with boats, fountain, food trucks, and striped circus tents](https://static.simonwillison.net/static/2026/gemini-ham-radio-small.jpg)](https://static.simonwillison.net/static/2026/nano-banana-2-ham-radio.jpg) That one was pretty obvious, the raccoon is in the “Amateur Radio Club” booth in the center of the image\! Claude said: > Honestly, this one wasn’t really hiding — he’s the star of the booth\. Feels like the illustrator took pity on us after that last impossible scene\. The little “W6HAM” callsign pun on the booth sign is a nice touch too\. I also tried Nano Banana Pro[in AI Studio](https://aistudio.google.com/app/prompts?state=%7B%22ids%22:%5B%221sGU5A7mrngkfLfSEU84xaV1DhtOTnS--%22%5D,%22action%22:%22open%22,%22userId%22:%22106366615678321494423%22,%22resourceKeys%22:%7B%7D%7D&usp=sharing)and got this, by far the worst result from any model\. Not sure what went wrong here\! [![The raccoon is larger than everyone else, right in the middle of the image with an ugly white border around it.](https://static.simonwillison.net/static/2026/nano-banana-pro-ham-radio-small.jpg)](https://static.simonwillison.net/static/2026/nano-banana-pro-ham-radio.jpg) #### gpt\-image\-2 With the baseline established, let’s try out the new model\. I used an updated version of my[openai\_image\.py](https://github.com/simonw/tools/blob/main/python/openai_image.py)script, which is a thin wrapper around the[OpenAI Python](https://github.com/openai/openai-python)client library\. Their client library hasn’t yet been updated to include`gpt\-image\-2`but thankfully it doesn’t validate the model ID so you can use it anyway\. Here’s how I ran that: ``` OPENAI_API_KEY="$(llm keys get openai)" \ uv run https://tools.simonwillison.net/python/openai_image.py \ -m gpt-image-2 \ "Do a where's Waldo style image but it's where is the raccoon holding a ham radio" ``` Here’s what I got back\. I don’t*think*there’s a raccoon in there—I couldn’t spot one, and neither could Claude\. [![Lots of stuff, a ham radio booth, many many people, a lake, but maybe no raccoon?](https://static.simonwillison.net/static/2026/gpt-image-2-default.jpg)](https://static.simonwillison.net/static/2026/gpt-image-2-default.png) The[OpenAI image generation cookbook](https://github.com/openai/openai-cookbook/blob/main/examples/multimodal/image-gen-models-prompting-guide.ipynb)has been updated with notes on`gpt\-image\-2`, including the`outputQuality`setting and available sizes\. I tried setting`outputQuality`to`high`and the dimensions to`3840x2160`—I believe that’s the maximum—and got this—a 17MB PNG which I converted to a 5MB WEBP: ``` OPENAI_API_KEY="$(llm keys get openai)" \ uv run 'https://raw.githubusercontent.com/simonw/tools/refs/heads/main/python/openai_image.py' \ -m gpt-image-2 "Do a where's Waldo style image but it's where is the raccoon holding a ham radio" \ --quality high --size 3840x2160 ``` [![Big complex image, lots of detail, good wording, there is indeed a raccoon with a ham radio.](https://static.simonwillison.net/static/2026/image-fc93bd-q100.jpg)](https://static.simonwillison.net/static/2026/image-fc93bd-q100.webp) That’s pretty great\! There’s a raccoon with a ham radio in there \(bottom left, quite easy to spot\)\. The image used 13,342 output tokens, which are charged at $30/million so a total cost of around[40 cents](https://www.llm-prices.com/#ot=13342&ic=5&cic=1.25&oc=10&sel=gpt-image-2-image)\. #### Takeaways I think this new ChatGPT image generation model takes the crown from Gemini, at least for the moment\. Where’s Waldo style images are an infuriating and somewhat foolish way to test these models, but they do help illustrate how good they are getting at complex illustrations combining both text and details\. #### Update: asking models to solve this is risky rizaco[on Hacker News](https://news.ycombinator.com/item?id=47852835#47853561)asked ChatGPT to draw a red circle around the raccoon in one of the images in which I had failed to find one\. Here’s an animated mix of their result and the original image: ![The circle appears around a raccoon with a ham radio who is definitely not there in the original image!](https://static.simonwillison.net/static/2026/ham-radio-cheat.gif) Looks like we definitely can’t trust these models to usefully solve their own puzzles\!

Similar Articles

Lost Cat

YouTube AI Channels

OpenAI demonstrates ChatGPT Images, a new feature that allows users to generate images from text descriptions in any size.

New AI image generator BEATS EVERYTHING

YouTube AI Channels

OpenAI releases ChatGPT Images 2.0, a new image model that decisively beats Google’s Nano Banana Pro on 11 real-world tests featuring anime posters, UI screenshots, brand boards, and data infographics with consistently readable text and accurate layouts.

ChatGPT Images — Chameleon

YouTube AI Channels

OpenAI released ChatGPT Images 2.0, enabling users to generate entire video frames for storytelling.