Google DeepMind has launched three new AI models—Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber—designed for greater token efficiency, speed, and cost-effectiveness. Gemini 3.6 Flash is priced at $1.50 per million input tokens and $7.50 per million output tokens, with Gemini 3.5 Flash-Lite even more affordable at $0.30/$2.50. These models significantly reduce token consumption, with Gemini 3.6 Flash cutting up to 65% of token use on complex engineering tasks compared to previous versions, enhancing speed and reducing costs for enterprises. The models support up to 1 million tokens in input and 64,000 tokens in output, improving benchmarking scores in long-horizon engineering, machine learning, and knowledge work efforts. The new Flash Cyber model is fine-tuned for cybersecurity, targeting vulnerabilities and integrating with Google’s CodeMender platform, though it’s available only to governments and trusted partners. Google is still testing the more powerful Gemini 3.5 Pro, expected to be announced soon. These AI advancements emphasize efficient, autonomous agent systems for enterprise applications, though all models remain proprietary and accessible solely through Google’s API services on a commercial licensing basis.
Back