Google Ships Gemini 3.6 Flash and 3.5 Flash-Lite, Betting on Cheap Inference

Two new lightweight Gemini models target speed and lower API costs, aimed squarely at high-volume agent workloads.

Google released Gemini 3.6 Flash and Gemini 3.5 Flash-Lite, a pair of lightweight models optimized for speed and lower API costs, as reported by @techcodebee and @ivke2006. Both were flagged across multiple developer briefings as top stories, and the reason is not glamour — it's arithmetic.

Unlock the full briefing

Get every story in today's briefing, the full archive, and the daily AI intelligence brief.

All stories today

Full archive

Daily brief

Cancel anytime. Payments powered by Stripe.