Frontier models work, but supervision tools lag
Claude Opus 4.6, GPT-5.3-Codex and OpenAI Frontier, and The Deploy Log for this week. Every price and claim read off the publisher's own page, with a call on each.
Anthropic
Claude Opus 4.6
DEPLOYif you run long agentic coding tasks
WHAT SHIPPEDClaude Opus 4.6 is available today on claude.ai, the API, and all major cloud platforms at $5/$25 per million tokens. It adds a 1M token context window in beta, leads Terminal-Bench 2.0 and Humanity's Last Exam, and beats GPT-5.2 by around 144 Elo points on GDPval-AA.
THE CATCHThe 1M context window is beta, and Opus 4.6 can add cost and latency on simpler tasks because it thinks more deeply by default. Anthropic recommends dialing effort down from high to medium when the model overthinks.
IF YOU RUNa team doing multi-step coding, code review, or long-context retrieval, this model's 76% on the 8-needle 1M MRCR v2 versus Sonnet 4.5's 18.5% is the number that matters
SOURCEAnthropic, Introducing Claude Opus 4.6
Founding memberships
Full access, free, for a year.
Read every edition
What follows is the rest of this edition, with The Deploy Log for this week, and The pattern.
Free with a signup. No card.