Google announced Gemini 3.6 Flash with improved coding efficiency and lower token costs, alongside Gemini 3.5 Flash Lite and a cybersecurity-focused model, while still developing Gemini 3.5 Pro and hinting at Gemini 4.
<p>Google announced a significant evolution of its AI models at I/O in May with the release of Gemini 3.5 Flash, and it's not slowing down. The company <a href="https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/">revealed three new AI models today</a>, including its first version of Gemini geared toward cybersecurity. However, none of the new models is the delayed Gemini 3.5 Pro, which was supposed to launch in June.</p>
<p>Gemini 3.5 Flash, which was the <a href="https://arstechnica.com/google/2026/05/google-announces-agent-optimized-gemini-3-5-flash-and-a-do-anything-model-called-omni/">star of the show at I/O</a>, has already been deprecated. In its place, developers and users will find Gemini 3.6 Flash. Google makes the usual claims about this model—it's marginally more capable and better at coding, and it has great multimodal features.</p>
<p>Google says the changes to 3.6 Flash were made in response to user feedback on the 3.5 release. In general, Gemini 3.5 Flash didn't appear to live up to Google's promises around code generation. Perhaps that is simply a consequence of Google's intense focus on efficiency as businesses have started to fret over the cost of AI tokens.</p><p><a href="https://arstechnica.com/google/2026/07/google-reveals-faster-and-cheaper-gemini-3-6-flash-says-3-5-pro-is-still-in-testing/">Read full article</a></p>
<p><a href="https://arstechnica.com/google/2026/07/google-reveals-faster-and-cheaper-gemini-3-6-flash-says-3-5-pro-is-still-in-testing/#comments">Comments</a></p>
# Google announces Gemini 3.6 Flash and cybersecurity AI, teases 3.5 Pro and Gemini 4
Source: [https://arstechnica.com/google/2026/07/google-reveals-faster-and-cheaper-gemini-3-6-flash-says-3-5-pro-is-still-in-testing/](https://arstechnica.com/google/2026/07/google-reveals-faster-and-cheaper-gemini-3-6-flash-says-3-5-pro-is-still-in-testing/)
Google announced a significant evolution of its AI models at I/O in May with the release of Gemini 3\.5 Flash, and it’s not slowing down\. The company[revealed three new AI models today](https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/), including its first version of Gemini geared toward cybersecurity\. However, none of the new models is the delayed Gemini 3\.5 Pro, which was supposed to launch in June\.
Gemini 3\.5 Flash, which was the[star of the show at I/O](https://arstechnica.com/google/2026/05/google-announces-agent-optimized-gemini-3-5-flash-and-a-do-anything-model-called-omni/), has already been deprecated\. In its place, developers and users will find Gemini 3\.6 Flash\. Google makes the usual claims about this model—it’s marginally more capable and better at coding, and it has great multimodal features\.
Google says the changes to 3\.6 Flash were made in response to user feedback on the 3\.5 release\. In general, Gemini 3\.5 Flash didn’t appear to live up to Google’s promises around code generation\. Perhaps that is simply a consequence of Google’s intense focus on efficiency as businesses have started to fret over the cost of AI tokens\.
[https://cdn.arstechnica.net/wp-content/uploads/2026/07/gemini-3-6-flash__evals__figure.png](https://cdn.arstechnica.net/wp-content/uploads/2026/07/gemini-3-6-flash__evals__figure.png)[https://cdn.arstechnica.net/wp-content/uploads/2026/07/gemini-3-6-flash__evals__figure.png](https://cdn.arstechnica.net/wp-content/uploads/2026/07/gemini-3-6-flash__evals__figure.png)[](https://cdn.arstechnica.net/wp-content/uploads/2026/07/gemini-3-6-flash__evals__figure.png)
Credit: Google
Credit: Google
In the DeepSWE test for coding, 3\.6 Flash jumps to 49 percent versus 37 percent for 3\.5 Flash\. The new model now supports computer use as a standard feature in the Gemini API, too\. The OSWorld test for computer use shows a modest boost to 83 percent from 3\.5’s 78\.4 percent score\. Efficiency was a big focus for Gemini 3\.5 Flash, and Google says that effort has been amped up with 3\.6\. Even with small benchmark gains, Gemini 3\.6 Flash uses about 17 percent fewer tokens\.
In agentic workflows \(like the one below\), Gemini 3\.6 Flash should complete tasks more accurately, in fewer steps, and with fewer tokens\. That could save developers \(and Google\) a lot of money\. The new model has a lower API cost, at $1\.50/1M input tokens and $7\.50/1M output tokens\. It was $1\.50 and $9, respectively, for 3\.5 Flash\.
Token efficiency in Gemini 3\.6 Flash\.
Token efficiency in Gemini 3\.6 Flash\.
Google is not done with the 3\.5 branch yet, though\. It has also released Gemini 3\.5 Flash Lite and 3\.5 Flash Cyber\. The new Flash Lite is Google’s most efficient modern AI, hitting an impressive 350 tokens per second\. The company claims this model is ideal for scaling agentic systems without breaking the bank\. Based on benchmark numbers, the new Flash Lite is almost on par with frontier models from about a year ago, but it’s cheap\. Pricing is set at $0\.30/1M input tokens and $2\.50/1M output tokens, though that is slightly higher than the previous 3\.1 Flash Lite \($0\.25 and $1\.50\)\.
Google DeepMind released Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber, focusing on efficiency, coding, and cybersecurity, but the anticipated Gemini 3.5 Pro was not included due to internal delays.
Google introduces Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber models, offering improved token efficiency, lower latency, and better performance for agentic workflows, with 3.6 Flash reducing output token usage by 17% and showing gains in coding and knowledge tasks.
Google introduces Gemini 3.5 Flash Cyber, a lightweight cybersecurity model fine-tuned for automated vulnerability discovery and patching, offering cost-efficient performance competitive with larger models.
Google has released Gemini 3 Flash, a fast, cost-effective AI model that combines Pro-grade reasoning with Flash-level speed for tasks like coding, complex analysis, and agentic workflows.