68 rows dated November 2025 are in The Deploy Log: 11 leads, 28 in Models and APIs, 20 in Tools and products, 2 in Robotics, hardware and chips, 6 in Breakthroughs and 1 in Funding and acquisitions, from 9 publishers. They include Claude Opus 4.5 from Anthropic, FLUX.2 from Hugging Face, FLUX.2 Pro on AI Gateway from Vercel, GPT-5.1-Codex-Max from OpenAI, Gemini 3 from Google DeepMind and 63 others. Every item sourced to the publisher's own page.
Claude Opus 4.5 is available today on Anthropic's apps, API, and all three major cloud platforms at $5 per million input tokens and $25 per million output tokens. It leads 7 out of 8 programming languages on SWE-bench Multilingual and scores a 10.6% jump over Sonnet 4.5 on Aider Polyglot.
FLUX.2 is a new open image generation model from Black Forest Labs with a 32B parameter DiT and a single Mistral Small 3.1 text encoder. It supports multiple reference images and runs on 24GB GPUs with 4-bit quantization.
FLUX.2 Pro from Black Forest Labs is now available via Vercel's AI Gateway with no other provider accounts required. Set model to bfl/flux-2-pro in the AI SDK to generate images up to 4MP with multi-reference input.
GPT-5.1-Codex-Max is available in Codex today for the CLI, IDE extension, cloud, and code review. It is the first model natively trained to operate across multiple context windows through compaction, working over millions of tokens in a single task.
Gemini 3 Pro is available today in the Gemini app, AI Studio, Vertex AI, and Google Antigravity. It tops the LMArena Leaderboard with a score of 1501 Elo and scores 76.2% on SWE-bench Verified.
Waymo is introducing fully autonomous driving in five new cities: Miami, Dallas, Houston, San Antonio, and Orlando. Operations start today in Miami, and will begin in the remaining four cities over the coming weeks, ahead of opening doors to riders next year.
AWS and OpenAI announced a multi-year, strategic partnership that provides AWS infrastructure to run and scale OpenAI's core AI workloads starting immediately. The $38 billion agreement gives OpenAI access to AWS compute comprising hundreds of thousands of NVIDIA GPUs, with the ability to expand to tens of millions of CPUs.
Moonshot AI's Kimi K2 Thinking and Kimi K2 Thinking Turbo are now available on Vercel's AI Gateway with no other provider accounts required. Kimi K2 Thinking handles up to 200 to 300 sequential tool calls.
The stable version of Gemini 2.5 Flash-Lite is now generally available at $0.10 per 1M input tokens and $0.40 per 1M output tokens, with a 1 million-token context window and optional reasoning.
Google DeepMind released pretrained and instruction-tuned T5Gemma encoder-decoder checkpoints from T5 Small to 9B, including an unbalanced 9B encoder with 2B decoder.
Shopping research in ChatGPT is rolling out today on mobile and web for logged-in users on Free, Go, Plus, and Pro plans, powered by a version of GPT-5 mini trained with reinforcement learning specifically for shopping tasks.
OVHcloud is now a supported Inference Provider on the Hugging Face Hub, offering serverless access to open-weight models like gpt-oss, Qwen3, DeepSeek R1, and Llama with pay-per-token pricing starting at €0.04 per million tokens.
Claude Sonnet 4.5, Haiku 4.5, and Opus 4.1 are now available in public preview in Microsoft Foundry, and Microsoft's Agent Mode in Excel includes an option to use Claude in preview.
Google Antigravity is available in public preview at no charge with generous rate limits on Gemini 3 Pro usage, and includes access to Gemini 3, Claude Sonnet 4.5, and OpenAI's GPT-OSS.
Nano Banana Pro, built on Gemini 3 Pro, generates images with legible text in multiple languages, blends up to 14 images, maintains consistency of up to 5 people, and supports 2K and 4K resolution.
Hugging Face TRL now officially integrates with RapidFire AI to accelerate fine-tuning and post-training experiments, with internal benchmarks showing approximately 16-24x higher experimentation throughput than sequential config comparison.
Google launched AI image verification in the Gemini app using SynthID, letting users upload an image and ask if it was generated or edited by Google AI.
AnyLanguageModel is a Swift package that provides a drop-in replacement for Apple's Foundation Models framework, supporting local and remote LLM providers including MLX, llama.cpp, and cloud APIs.
Anthropic reports a Chinese state-sponsored group manipulated Claude Code to attempt infiltration into roughly thirty global targets, with AI performing 80 to 90% of the campaign.
Group chats roll out to all logged-in users on ChatGPT Free, Go, Plus and Pro plans globally over the coming days, with responses powered by GPT-5.1 Auto.
Maryland will deploy Claude across multiple state agencies to connect families with benefits and help caseworkers verify more than 150,000 documents each month.
OpenAI says the New York Times is demanding 20 million private ChatGPT conversations and that OpenAI is fighting the demand while developing client-side encryption.
OpenAI launches OpenAI for Ireland with an SME Booster programme in 2026, a Dogpatch Labs partnership, and a three-year Patch partnership for founders aged 16 to 21.
Anthropic open-sources an automated evaluation showing Claude Sonnet 4.5 scores 94% on political even-handedness, compared with GPT-5 at 89% and Llama 4 at 66%.
GPT-5.1 Instant and Thinking roll out to all users, with warmer tone, adaptive reasoning, and faster simple tasks. API release includes no-reasoning mode, 24-hour prompt caching, and new apply_patch and shell tools.
OpenAI announced more than 1 million business customers around the world are directly using OpenAI, with more than 7 million total ChatGPT for Work seats, up 40% in just 2 months.
CRED built Cleo, an AI conversational companion powered by OpenAI models including GPT-4.0, GPT-5, and o3, and in three months since launch Cleo reached a 98% resolution accuracy rate with a 14 percentage point improvement in CSAT scores.
OpenAI released IndQA, a new benchmark that evaluates AI systems on Indian culture and languages, spanning 2,278 questions across 12 languages and 10 cultural domains, created in partnership with 261 domain experts.
Cognizant will deploy Claude to up to 350,000 employees globally, combining Claude with agentic tooling, engineering platforms, and industry blueprints to accelerate enterprise AI adoption and internal transformation.
OpenAI announced Aardvark, an agentic security researcher powered by GPT-5 that identified 92% of known and synthetically-introduced vulnerabilities in benchmark testing and is now in private beta.
IBM released Granite 4.0 Nano models from 350M to 1.5B parameters under Apache 2.0, trained on over 15T tokens with native support on vLLM, llama.cpp, and MLX.
Hugging Face improved streaming datasets with 100x fewer startup requests, 10x faster data resolution, and up to 2x faster streaming speed, outrunning local SSDs when training on 64xH100 with 256 workers.
huggingface_hub reached v1.0 with 113.5 million monthly downloads, powering access to over 2 million public models, 0.5 million public datasets, and 1 million public Spaces.
Vercel AI Gateway now offers image-only models including Black Forest Labs FLUX.2 Flex, FLUX.2 Pro, FLUX.1 Kontext Max, FLUX.1 Kontext Pro, FLUX 1.1 Pro Ultra, FLUX 1.1 Pro, FLUX.1 Fill Pro, and Google Imagen 4.0 Generate, Fast Generate, and Ultra Generate.
Vercel Streamdown 1.6 is now available with memoization, LRU caching, optimized string operations, removal of regexes, lazy-loaded Code Blocks, Mermaid, and Math components, a rebuilt code highlighting system, a custom markdown renderer, and Static Mode.
Grok 4.1 Fast Reasoning and Grok 4.1 Fast Non-Reasoning are now accessible via Vercel's AI Gateway with no other provider accounts required, and both have a 2M context window.
Vercel Agent investigations are now included in Observability Plus, adding 10 investigations to every billing cycle at no extra cost to the subscription.
Vercel launched improvements to the Firewall UI, including an updated Overview page for DDoS attacks and a new Traffic page to drill into top sources of traffic by IPs, request paths, JA4 digests, ASN, and user agents.
Nano Banana Pro (Gemini 3 Pro Image) is now available on Vercel's AI Gateway with no other provider accounts required, supporting higher resolution and multi-image input limits.
BotID Deep Analysis, Vercel's advanced bot protection system, is free for all Pro and Enterprise customers from November 5 to January 15, 2026, and uses thousands of telemetry points for real-time client-side checks.
Vercel Edge Config Reads and Writes are moving from package-based to per-unit pricing on the Pro plan, with reads at $0.000003 per read and writes at $0.01 per write.
Vercel Skew Protection max age can now persist for the entire lifetime of your deployments, removing the previous limits of 12 hours on Pro and 7 days on Enterprise.
Vercel introduced the Sandbox CLI, a command-line interface for managing isolated compute environments built on the Docker CLI model, supporting Node.js node22 and Python python3.13 workloads.
Vercel announced a Snowflake integration for v0 that lets you connect v0 to Snowflake, ask questions about your data, and build data-driven Next.js applications that deploy directly to Snowflake, with data never leaving your Snowflake environment.
Vercel BotID Deep Analysis detected a 500% traffic spike from 40-45 new browser profiles cycling through proxy nodes and blocked the bot network in 5 minutes with no manual intervention.
Vercel Functions now support Bun in Public Beta, with internal testing showing Bun reduced average latency by 28% in CPU-bound Next.js rendering workloads compared to Node.js.
Vercel microfrontends are generally available, serving nearly 1 billion routing requests per day with over 250 teams deploying, priced at $250 per additional project per month and $2 per million routing requests.
Google Research deployed a lightweight linear regression model that predicts EV charging port availability and reduces bad predictions by approximately 20% in morning peak times and approximately 40% in evening peak times.
Google Research introduced an end-to-end speech-to-speech translation model that enables real-time translation in the original speaker's voice with only a 2-second delay, now available in Google Meet on servers and as an on-device feature for Pixel 10 devices.
Google introduced generative UI, where AI models create interactive interfaces on the fly, rolling out in the Gemini app as dynamic view and in AI Mode in Google Search.
Google Quantum AI introduces Decoded Quantum Interferometry, a quantum algorithm that converts certain optimization problems into decoding problems and could solve OPI examples with a few million quantum operations versus over 10^23 classical operations.