Google Slashes AI Agent Costs With Three New Models
Google DeepMind has released three new AI models: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber, designed to make AI agents faster and more cost-efficient at scale. Gemini 3.6 Flash cuts token costs by up to 65% on long-horizon engineering tasks.
Gemini 3.5 Flash-Lite is priced at just $0.30 per million input tokens, while Gemini 3.6 Flash comes in at $1.50 per million input tokens. Despite the older Gemini 3.1 Flash-Lite remaining the cheapest option, the newer 3.5 Flash-Lite runs twice as fast, offering enterprises better value for speed-sensitive workloads.
Google commoditizes inference costs to make agentic AI deployments economically viable at enterprise scale.
