Mac Studio M5 Ultra
Mac Studio with M5 Ultra scales to a 36-core CPU, up to an 80-core GPU, and 512GB of unified memory with 1.2TB/s bandwidth, delivering up to 4.3x the peak AI compute performance of M3 Ultra and 9.8x more than M1 Ultra.
The Deploy Log | Month by month
59 rows dated August 2026 are in The Deploy Log: 12 leads, 13 in Models and APIs, 31 in Tools and products, 2 in Robotics, hardware and chips and 1 in Breakthroughs, from 13 publishers. They include Mac Studio M5 Ultra from Apple, ChatGPT for Teens from OpenAI, Replit Free Mode from OpenAI, ChatGPT Business Premium seats from OpenAI, GPT-5.6 Sol Ultrafast from OpenAI and 54 others. Every item sourced to the publisher's own page.
Mac Studio with M5 Ultra scales to a 36-core CPU, up to an 80-core GPU, and 512GB of unified memory with 1.2TB/s bandwidth, delivering up to 4.3x the peak AI compute performance of M3 Ultra and 9.8x more than M1 Ultra.
ChatGPT for Teens places users estimated under 18 or stating age 13 to 17 into a protected experience with Study Mode, responsible homework reminders, quizzes, learning visualizations, Study Hours, break reminders, and parent controls including Quiet Hours and safety notifications.
Replit introduces Free Mode powered by GPT-5.6 Luna, letting users get answers, suggestions, feedback, and analysis in seconds without consuming usage, with routing to GPT-5.6 Sol for tasks requiring more advanced reasoning.
Premium seats are now available on ChatGPT Business at $125 per user per month, or $100 per user per month billed annually, with 5x more usage than Standard seats and no five-hour usage limit.
GPT-5.6 Sol on Ultrafast mode runs up to 14x faster than Standard processing and generates up to 750 output tokens per second, launching first in the OpenAI API and powered by Cerebras.
Gemini 3.7 Flash is available at $0.75 per 1M input tokens and $3.75 per 1M output tokens through the end of the year, half the original 3.6 Flash cost per million tokens.
GPT-5.6 Sol in ChatGPT now gives more focused answers and makes factual errors 68% less common than GPT-5.5 Instant on financial, medical, and legal prompts. A new slider controls how much thought each response gets.
Grok Imagine Image 2.0 Preview from xAI is available on AI Gateway. It follows detailed instructions closely and plans typography and layout together, so dense visuals like infographics, posters, and title screens hold their structure and small text stays legible.
Shieldstral is a 3B open-weights multimodal safety classifier under Apache 2.0 that matches or outperforms open guard models up to 7x its size on text safety, refusal detection, policy adaptability, and multimodal benchmarks. It runs on a single 16GB NVIDIA GPU.
OpenAI is giving 100,000 researchers at selected academic institutions free access to frontier models including GPT-5.6 Sol Pro, starting with 10,000 researchers this summer and expanding through 2027, with access already available at the Institute for Advanced Study and École normale supérieure.
GPT-5.6 Luna now costs 80% less and GPT-5.6 Terra 20% less, with API pricing at $0.20 per million input tokens and $1.20 per million output tokens for Luna, and $2 per million input tokens and $12 per million output tokens for Terra starting July 30.
Gemini Robotics ER 2 is now publicly available via the Gemini API, Google AI Studio, and in private preview on Gemini Enterprise Agent Platform, with 57.4% accuracy on progress classification and 91.3% accuracy with 0.96s mean absolute distance on moment-finding.
Anthropic is opening 10,000 seats for scientists to access Claude subscriptions free and at discounted rates for one year, with standard seats free and premium seats with 5x usage limits at $15 per month.
The GPT-5.6 model family, including Sol, Terra, and Luna, is now available in Kiro, and testing found that on Terminal-Bench 2.1, GPT-5.6 Terra completed successful tasks in Kiro at roughly 82% cost reduction.
OpenAI is expanding ChatGPT for Teachers to 55 additional school systems, bringing it to over 100,000 more educators and staff.
Stability AI ships a DAW plugin for macOS AU and VST3 plus an enhanced StableAudio.com web app, both powered by commercially-safe models where users own outputs and can distribute them freely.
Mistral Agentic Search improves accuracy up to 3x on FinanceBench (26.7% to 86%) and cuts p90 latency up to 39.6% and token use by one-third.
ChatGPT Ads expands to 31 European markets, including Germany, France, Spain, Italy, Sweden, Norway, Denmark, the Netherlands, and Austria. Ads show only on Free and Go plans; Plus, Pro, and Enterprise remain ad-free.
GPT-5.6-Cyber completes 95.0% of advanced cybersecurity requests compared with 1.5% for GPT-5.6 Sol, and is available through Daybreak Red access.
Google DeepMind introduced SL2T, a sign-language-to-text model powering sign-to-text dictation in Gboard and Live Transcribe on Pixel 11, starting with ASL to English.
OpenAI expanded its ChatGPT ads pilot to the UK, Mexico, Brazil, Japan, and South Korea, with ads on Free and Go tiers; Plus, Pro, Business, Enterprise, and Education remain ad-free.
OpenAI made Daybreak Blue and Daybreak Red available through Amazon Bedrock, giving defenders access to GPT-5.6 Sol and purpose-trained cyber models in AWS environments.
Meta released Muse Glimmer, a 30B multimodal model under Apache 2.0, with day-0 support in transformers, llama.cpp, vLLM, and Inference Endpoints.
Baseten is now a supported Inference Provider on the Hugging Face Hub, launching support for conversational and text-generation tasks with open-weight LLMs such as Kimi K3, DeepSeek V4 Flash, and GLM-5.2.
Anthropic reduced biology-related fallbacks on Claude Fable 5 by about 85% across product surfaces, so users see far fewer switches to Opus 5 on everyday health and educational questions.
GLM 5.3 Flash from Z.ai is now available on AI Gateway as zai/glm-5.3-flash, a multimodal coding model with a 1M-token context window that accepts text and image inputs and supports function calling, structured output, and streaming.
Hy4 Preview from Tencent is now available on AI Gateway, an open-source Mixture-of-Experts model with 770B total parameters and 49B active per token, serving a context window of 1M tokens.
Ling 3.0 Flash Fin from Inclusion AI is now available on AI Gateway and free to use through September 25, with a 256K-token context window, up to 32K output tokens, and support for reasoning and function calling.
Qwen 3.8 Flash from Alibaba is now available on AI Gateway as alibaba/qwen3.8-flash, taking text and images as input with a 1 million token context window and up to 65k tokens in a response.
Speed Insights now has a free tier available on every plan for any number of projects, including 10,000 events per team every 30 days, with the paid product renamed Speed Insights Plus.
Vercel Connect is now generally available on all plans and in v0, replacing stored long-lived provider secrets with short-lived scoped tokens requested at runtime, with 100+ preset connectors and pricing at $3 per 1,000 token requests and $0.95 per 1,000 triggers on Pro.
Vercel Sandbox now runs globally starting with four regions: iad1, sfo1, cle1, and cdg1, with region selection on all plans and failover regions configurable for Pro and Enterprise teams.
Gemini-based data classification in Google Drive is now in open beta, automating file labeling to enforce DLP policies at scale.
Google Drive now supports direct previewing of client-side encrypted PDF and image files in beta, eliminating unnecessary downloads.
Google Workspace announced updates including Meet hardware UI, Chat member list controls, and data import for Microsoft Teams and OneDrive.
Apple's new Mac mini with M6 delivers up to 4x faster AI performance, 2x faster storage and graphics, and 40 percent faster CPU. M5 Pro supports up to 64GB unified memory. Pre-orders start today, availability September 22.
Vercel introduces Deployment Storage billed at $0.10 per GB per month with Hobby teams including up to 10 GB, keeping deployment files available for inspection and instant rollback in seconds with no rebuild.
Fish Audio models arrive on AI Gateway with every model free for 30 days through September 18, including text-to-speech normally priced at $15.00 per million characters and speech-to-text at $0.36 per hour of audio.
GLM 5.3 from Z.ai is available on AI Gateway with improvements over GLM 5.2 at complex software engineering and multi-step agent tasks while producing fewer output tokens, keeping a 1M token context window and 128K maximum output.
Cloudflare's Bot Preference Sync, available to all customers from Free to Enterprise, updates robots.txt to match your AI bot configuration for Search, Agent, and Training traffic.
DeepSeek V4 Flash Vision Experimental is now on AI Gateway, accepting images alongside text, with tool use, reasoning, and caching unchanged.
DeepSeek V4 Pro now runs on updated weights on AI Gateway, used by default when you call deepseek/deepseek-v4-pro.
Exa web search is now free on AI Gateway through August 31, and it is now the default web search for eve agents.
Grok 4.6 from SpaceXAI is now available on AI Gateway with a 500K token context window and text and image inputs.
You can now integrate 100+ services through Vercel Connect from the CLI, passing the service name to create a connector and attach it to your project.
Vercel Managed Images introduces versioned, open-source base images, with new sandboxes defaulting to vercel/sandbox/universal:latest on Ubuntu 26.04 with Node.js 24 and Python 3.14.
Vercel CDN now supports Encrypted Client Hello (ECH) for domains managed by Vercel DNS, encrypting the SNI in the TLS handshake.
Cloudflare made Certificate Transparency Monitoring generally available, now filtering out Cloudflare-issued certificates to reduce noise for over 650,000 customer domains.
Agent Plugins 1.0.0 is publicly available as an open, vendor-neutral standard for packaging Agent Skills and MCP servers into portable plugins, supported at launch across ChatGPT and Codex, Cursor, GitHub Copilot, Kiro, and VS Code.
Vercel Sandbox default quotas on Pro and Enterprise plans rose from 2,000 to 10,000 concurrent sandboxes and from 200 or 400 vCPUs per minute to a dynamic rate up to 5,000 vCPUs per minute.
The new v0 API is generally available, giving programmatic headless access to v0's app-building agent that generates an app, starts a dev server in a Vercel Sandbox, and returns a preview URL you can embed.
Notion now lets you share context with Custom Agents directly from the Share menu, without clicking over to the agent's settings.
DeepSeek V4 Flash now runs on updated weights by default on AI Gateway, scoring 82.7 on Terminal-Bench, up 25.8 points from 56.9 in the April preview, with no change to the model ID or your code.
Vercel deployments are now up to 7 seconds faster end to end, with up to 5 seconds of fixed platform overhead removed from every build and up to 2 seconds more saved from the newest Vercel CLI.
Grok Voice Think Fast 2.0 from xAI is now available on AI Gateway as a speech-to-speech voice model that reasons in parallel with speech and fires tool calls sooner with fewer reasoning tokens.
Sign in with ChatGPT adds your ChatGPT account as an authentication option for Vercel, available when adding the Vercel plugin to ChatGPT and when signing in to Vercel or v0.
NVIDIA's first custom CPU, Vera, is shipping at scale with 88 custom Olympus cores, 1.2TB/s memory bandwidth, and up to 1.8x faster per-core performance on agentic AI workloads. AWS received its first Vera CPU server and Vera Rubin GPU.
GeForce NOW now supports Firefox, delivering RTX-powered performance at up to 1440p and 120 fps for Ultimate members, plus 12 new games this week.
Skala 1.1, trained on 2.5x more data than its predecessor, achieves a weighted average error of 2.8 kcal/mol on GMTKN55 and is now available in CP2K with integrations underway for Psi4, FHI-aims, ORCA, and VASP.
The edition every row came from arrives in your inbox, free.
Join free