Kimi K3 on Bedrock
Moonshot AI's Kimi K3 is available on Amazon Bedrock, the first open model to reach 2.8 trillion parameters, with a 1-million-token context window and a 2.5x improvement in scaling efficiency over Kimi K2.
deployedbyai
Every deployment an edition of deployedbyai has carried, 745 rows so far. Every item sourced to the publisher's own page.
A row enters The Deploy Log when an edition carries it. Each one names the publisher, the thing that shipped and the number that moved, and links to the page it came from, so you can open the source and check the number against it.
Moonshot AI's Kimi K3 is available on Amazon Bedrock, the first open model to reach 2.8 trillion parameters, with a 1-million-token context window and a 2.5x improvement in scaling efficiency over Kimi K2.
Amazon SageMaker HyperPod Inference Gateway is a Kubernetes-native, GPU-aware routing system that deploys as a single EKS managed addon and reduces first-token latency by up to 82%.
OpenAI introduced Astra for Law, combining GPT-6 Astra with a legal search index covering more than 230 million URLs and 26 partner-built plugins, passing the Vals AI Legal Research Bench overall correctness check on 54.0% of questions versus 38.7% for GPT-6 Astra with web search alone.
GPT-6 Astra, rolling out to the OpenAI API and through Microsoft Azure and Amazon Bedrock at $10 per million input tokens and $50 per million output, the same headline price as Anthropic's Claude Fable 5.1, which shipped the same week. Cache reads are priced separately, and Fast mode costs 2x.
Waymo began welcoming first public riders in Denver, San Diego, and Tampa on September 1, 2026, marking 14 cities with fully autonomous trips. Tens of thousands in each city have signed up.
Mac Studio with M5 Ultra scales to a 36-core CPU, up to an 80-core GPU, and 512GB of unified memory with 1.2TB/s bandwidth, delivering up to 4.3x the peak AI compute performance of M3 Ultra and 9.8x more than M1 Ultra.
ChatGPT for Teens places users estimated under 18 or stating age 13 to 17 into a protected experience with Study Mode, responsible homework reminders, quizzes, learning visualizations, Study Hours, break reminders, and parent controls including Quiet Hours and safety notifications.
Replit introduces Free Mode powered by GPT-5.6 Luna, letting users get answers, suggestions, feedback, and analysis in seconds without consuming usage, with routing to GPT-5.6 Sol for tasks requiring more advanced reasoning.
Premium seats are now available on ChatGPT Business at $125 per user per month, or $100 per user per month billed annually, with 5x more usage than Standard seats and no five-hour usage limit.
GPT-5.6 Sol on Ultrafast mode runs up to 14x faster than Standard processing and generates up to 750 output tokens per second, launching first in the OpenAI API and powered by Cerebras.
Gemini 3.7 Flash is available at $0.75 per 1M input tokens and $3.75 per 1M output tokens through the end of the year, half the original 3.6 Flash cost per million tokens.
GPT-5.6 Sol in ChatGPT now gives more focused answers and makes factual errors 68% less common than GPT-5.5 Instant on financial, medical, and legal prompts. A new slider controls how much thought each response gets.
Grok Imagine Image 2.0 Preview from xAI is available on AI Gateway. It follows detailed instructions closely and plans typography and layout together, so dense visuals like infographics, posters, and title screens hold their structure and small text stays legible.
Shieldstral is a 3B open-weights multimodal safety classifier under Apache 2.0 that matches or outperforms open guard models up to 7x its size on text safety, refusal detection, policy adaptability, and multimodal benchmarks. It runs on a single 16GB NVIDIA GPU.
OpenAI is giving 100,000 researchers at selected academic institutions free access to frontier models including GPT-5.6 Sol Pro, starting with 10,000 researchers this summer and expanding through 2027, with access already available at the Institute for Advanced Study and École normale supérieure.
GPT-5.6 Luna now costs 80% less and GPT-5.6 Terra 20% less, with API pricing at $0.20 per million input tokens and $1.20 per million output tokens for Luna, and $2 per million input tokens and $12 per million output tokens for Terra starting July 30.
Gemini Robotics ER 2 is now publicly available via the Gemini API, Google AI Studio, and in private preview on Gemini Enterprise Agent Platform, with 57.4% accuracy on progress classification and 91.3% accuracy with 0.96s mean absolute distance on moment-finding.
Claude Opus 5 is available today at half the price of Claude Fable 5. It more than doubles Opus 4.8 on Frontier-Bench v0.1 and scores 10.2 percentage points higher on organic chemistry tasks.
Free premium Claude access for verified US K-12 educators with a library of teaching skills, Learning Commons standards mapping across all 50 states, and connections to tools like ASSISTments, Brisk Teaching, Canva Education, and MagicSchool.
Moonshot AI's open-source K3 model with a 1M-token context window and native visual understanding for text, image, and video inputs, now available on AI Gateway with no markup and no platform fee.
GitHub Actions API and UI now report "2,500+" when workflow run query results exceed 2,500 records.
Claude Opus 5.5, GPT-6 Sol, GPT-6 Luna, and Grok 4.7 are available to Copilot plans.
Google Calendar on the web now displays up to three time zones, up from two.
LiteLLM v1.100.3 released with 7445 commits to main since this release.
New vercel/vcr-action/login action authenticates with GitHub OIDC and pushes container images to VCR.
Sandbox memory usage data now available in dashboard and CLI, reporting average, P75, and P95 memory.
GameLift Servers is now available in 5 new regions and 8 Local Zones.
Amazon RDS for MySQL Extended Support adds minor versions 5.7.44-rds.20260902 and 8.0.46-rds.20260908.
Amazon RDS for PostgreSQL now supports post-quantum TLS key exchange.
Azure Instant Access for VM restore points is generally available.
Model Armor prompt injection and jailbreak detection filters are available in Melbourne with data residency enforcement.
langchain-openai 1.6.6 released with a fix that raises on error events in the stream path.
The edition every row came from arrives in your inbox, free.
Join free