New Gemini AI Models Enhance Efficiency and Performance for Developers
The Gemini team has launched three new AI models: 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber, aimed at improving efficiency, latency, and reliability for AI agents. These models offer cost-effective solutions for various tasks, including coding, data analysis, and cybersecurity.
Gemini 3.6 Flash improves on the previous 3.5 Flash model by enhancing token efficiency, reducing output tokens by 17%, and lowering costs to $1.50 per million input tokens and $7.50 per million output tokens.
3.6 Flash demonstrates better performance across various use cases, including financial data analysis and code migrations, while also incorporating advanced safety safeguards against misuse.
Gemini 3.5 Flash-Lite is designed for high-throughput tasks, running at 350 output tokens per second, and is priced at $0.30 per million input tokens and $2.50 per million output tokens. It outperforms its predecessor, 3.1 Flash-Lite, in speed and quality.
3.5 Flash-Lite supports efficient scaling for agentic systems and can be configured for low-latency or higher thinking level tasks.
Gemini 3.5 Flash Cyber is tailored for identifying and addressing cybersecurity vulnerabilities, available exclusively to governments and trusted partners through a limited-access pilot program. It utilizes multiple agents to produce comprehensive reports and aims to enhance software security measures.