Google just rolled out three new Gemini models designed to make AI agents faster and cheaper to run in production. The focus isn’t just on better chatbot answers it’s about agents that can code, reason, use tools, handle multimodal inputs, and operate without becoming prohibitively expensive.
Gemini 3.6 Flash is built to improve coding, reasoning, tool use, and multimodal performance while consuming fewer output tokens than Gemini 3.5 Flash. Gemini 3.5 Flash‑Lite is Google’s fastest and most affordable 3.5‑class model, capable of generating around 350 output tokens per second. Meanwhile, Gemini 3.5 Flash Cyber is tailored for security work, helping detect and patch software vulnerabilities through Google’s CodeMender system.
The first two models are already available via the Gemini API, Google AI Studio, and the Gemini app, while Flash Cyber will initially be offered to trusted partners. The bigger takeaway is that AI agents aren’t just getting faster they’re becoming cheaper to ac...
Suggested Credits
Tags, Events, and Projects