This week's releases cut costs and put safety policy in your hands
Gemini 2.5 Flash-Lite and T5Gemma, and The Deploy Log for this week. Every price and claim read off the publisher's own page, with a call on each.
Google DeepMind
Gemini 2.5 Flash-Lite
DEPLOYif you run high-volume translation or classification
WHAT SHIPPEDThe stable version of Gemini 2.5 Flash-Lite is now generally available at $0.10 per 1M input tokens and $0.40 per 1M output tokens, with a 1 million-token context window and optional reasoning.
THE CATCHAudio input pricing is 40% lower than the preview, and the preview alias is removed on August 25th, so you must switch to gemini-2.5-flash-lite in code.
IF YOU RUNlatency-sensitive tasks like translation and classification at large request volumes
SOURCEGoogle DeepMind, Gemini 2.5 Flash-Lite is now stable and generally available
Founding memberships
Full access, free, for a year.
Read every edition
What follows is the rest of this edition, with The Deploy Log for this week, and The pattern.
Free with a signup. No card.