Frontier models work, but supervision tools lag

Claude Opus 4.6, GPT-5.3-Codex and OpenAI Frontier, and The Deploy Log for this week. Every price and claim read off the publisher's own page, with a call on each.

Anthropic

Claude Opus 4.6

DEPLOYif you run long agentic coding tasks

WHAT SHIPPEDClaude Opus 4.6 is available today on claude.ai, the API, and all major cloud platforms at $5/$25 per million tokens. It adds a 1M token context window in beta, leads Terminal-Bench 2.0 and Humanity's Last Exam, and beats GPT-5.2 by around 144 Elo points on GDPval-AA.

THE CATCHThe 1M context window is beta, and Opus 4.6 can add cost and latency on simpler tasks because it thinks more deeply by default. Anthropic recommends dialing effort down from high to medium when the model overthinks.

IF YOU RUNa team doing multi-step coding, code review, or long-context retrieval, this model's 76% on the 8-needle 1M MRCR v2 versus Sonnet 4.5's 18.5% is the number that matters

SOURCEAnthropic, Introducing Claude Opus 4.6

Founding memberships

Full access, free, for a year.

It becomes $50 a year. The date is not announced. Claim it when you sign up.

Claim your free year

Read every edition

What follows is the rest of this edition, with The Deploy Log for this week, and The pattern.

Free with a signup. No card.

Already a member? Sign in