This week's releases cut costs and put safety policy in your hands

Gemini 2.5 Flash-Lite and T5Gemma, and The Deploy Log for this week. Every price and claim read off the publisher's own page, with a call on each.

Google DeepMind

Gemini 2.5 Flash-Lite

DEPLOYif you run high-volume translation or classification

WHAT SHIPPEDThe stable version of Gemini 2.5 Flash-Lite is now generally available at $0.10 per 1M input tokens and $0.40 per 1M output tokens, with a 1 million-token context window and optional reasoning.

THE CATCHAudio input pricing is 40% lower than the preview, and the preview alias is removed on August 25th, so you must switch to gemini-2.5-flash-lite in code.

IF YOU RUNlatency-sensitive tasks like translation and classification at large request volumes

SOURCEGoogle DeepMind, Gemini 2.5 Flash-Lite is now stable and generally available

Founding memberships

Full access, free, for a year.

It becomes $50 a year. The date is not announced. Claim it when you sign up.

Claim your free year

Read every edition

What follows is the rest of this edition, with The Deploy Log for this week, and The pattern.

Free with a signup. No card.

Already a member? Sign in