Gemini 2.5 Flash-Lite
The stable version of Gemini 2.5 Flash-Lite is now generally available at $0.10 per 1M input tokens and $0.40 per 1M output tokens, with a 1 million-token context window and optional reasoning.
13 deployments dated October 27 to November 2, 2025 are in The Deploy Log: 2 leads, 6 in Models and APIs, 4 in Tools and products and 1 in Robotics, hardware and chips, from 6 publishers. They include Gemini 2.5 Flash-Lite from Google DeepMind, T5Gemma from Google DeepMind, Aardvark from OpenAI, Granite 4.0 Nano from Hugging Face, Streaming datasets from Hugging Face and 8 others. Every item sourced to the publisher's own page.
The stable version of Gemini 2.5 Flash-Lite is now generally available at $0.10 per 1M input tokens and $0.40 per 1M output tokens, with a 1 million-token context window and optional reasoning.
Google DeepMind released pretrained and instruction-tuned T5Gemma encoder-decoder checkpoints from T5 Small to 9B, including an unbalanced 9B encoder with 2B decoder.
OpenAI announced Aardvark, an agentic security researcher powered by GPT-5 that identified 92% of known and synthetically-introduced vulnerabilities in benchmark testing and is now in private beta.
IBM released Granite 4.0 Nano models from 350M to 1.5B parameters under Apache 2.0, trained on over 15T tokens with native support on vLLM, llama.cpp, and MLX.
Hugging Face improved streaming datasets with 100x fewer startup requests, 10x faster data resolution, and up to 2x faster streaming speed, outrunning local SSDs when training on 64xH100 with 256 workers.
huggingface_hub reached v1.0 with 113.5 million monthly downloads, powering access to over 2 million public models, 0.5 million public datasets, and 1 million public Spaces.
OpenAI released gpt-oss-safeguard-120b and 20b, open-weight reasoning models for safety classification under Apache 2.0, downloadable from Hugging Face.
OpenAI updated ChatGPT's default model to reduce undesired responses in sensitive conversations by 65-80%, working with over 170 mental health experts.
Vercel BotID Deep Analysis detected a 500% traffic spike from 40-45 new browser profiles cycling through proxy nodes and blocked the bot network in 5 minutes with no manual intervention.
Vercel Functions now support Bun in Public Beta, with internal testing showing Bun reduced average latency by 28% in CPU-bound Next.js rendering workloads compared to Node.js.
Vercel microfrontends are generally available, serving nearly 1 billion routing requests per day with over 250 teams deploying, priced at $250 per additional project per month and $2 per million routing requests.
Anthropic expanded Claude for Financial Services with an Excel add-in, new connectors, and 6 Agent Skills, topping the Finance Agent benchmark at 55.3% accuracy.
Waymo is advancing its Driver for snowier weather, having amassed tens of thousands of miles in snowy conditions and using a 6th-generation Driver informed by over 100 million autonomous miles.
The edition every row came from arrives in your inbox, free.
Join free