Key takeaways
What happened
8 Flash. The budget model is supposed to beat its predecessor by a wide margin in coding. 8 Flash Cyber. 5 Pro and Gemini 4 depends on who you ask. New Deepmind head Koray Kavukcuoglu made it clear that Google isn't just chasing price-performance optimization and still wants to lead on raw capability too. 1 benchmark for long-horizon software engineering tasks. 3%).
Why it matters
Google also says the new model has improved at 3D generation. 8 Flash reportedly built from a single prompt in Google's AI coding tool Antigravity. The textures were generated with Google's Nano Banana image model. 7 Flash. 50. 00. 8 Flash would remain far cheaper on a per-token basis than the top models from OpenAI and Anthropic. 8 Flash runs extra reasoning steps on complex tasks and calls tools iteratively.
The model "works harder," Google says, which means higher token consumption that partly offsets the lower per-token price. 7 Flash. 8 Flash is available to developers through Google AI Studio, Google Antigravity, and Android Studio. Businesses can access it through Gemini Enterprise. Consumers get it in the Gemini app, in Google Search's AI Mode, and for paying subscribers, in Google Sheets. 7 Flash at 56.
What to watch
6 at medium reasoning, both also scoring 59. The Intelligence Index gains come mainly from stronger performance on agentic benchmarks like tool use and coding tasks, according to Artificial Analysis. 58 per task as the cheapest model at its intelligence level. 40, even though the per-token price stayed the same. 7 Flash for efficiency-focused workloads. 5 minutes. 2 minutes). At low reasoning levels, the time drops to about 48 seconds.
5 Flash Cyber, isn't publicly available. Google distributes it through the Fairwind Program to government agencies, critical infrastructure operators, and software maintainers. The model has less restrictive safety settings than the standard version because it's built for defensive cybersecurity work. 6%). 8 percent while costing much less. The model also appears more resilient against prompt injection attacks. 5 percent. 8 percent. 8 percent, and Opus 5 with its additional security options scores even lower within the Claude ecosystem.


.gif)
