Claude Sonnet 5 Just Dropped: Anthropic’s Mid-Tier Model Now Rivaling Opus
Claude Sonnet 5 Just Dropped: Anthropic’s Mid-Tier Model Now Rivaling Opus
If you’ve been watching the AI space closely, you already know the pattern. Every few months, one company or another announces a new model that pushes the frontier forward, and the discourse around it splits into two predictable camps: the ones saying it’s a game changer and the ones calling it incremental. Claude Sonnet 5 landed on June 30, and I’m going to try to sidestep both of those tired positions and just tell you what it actually does.
Anthropic released Sonnet 5 this week, describing it as the most agentic Sonnet model yet. The company is positioning it as a substantial jump over Sonnet 4.6, with performance that’s getting close to Opus 4.8 on several benchmarks, but at noticeably lower prices. For developers who have been watching the cost of running autonomous AI workflows creep upward, that framing matters.


What “agentic” actually means here is worth unpacking. Sonnet 5 can make plans, use tools like browsers and terminals, and run autonomously on tasks that would have required Opus-class models just a few months ago. The company shared benchmark data from BrowseComp (agentic search) and OSWorld (computer use), showing Sonnet 5 outperforming Sonnet 4.6 across different effort levels while providing substantially better cost efficiency at medium effort settings. At higher effort, it can match Opus 4.8 on some tasks. The charts Anthropic published are worth looking at if you want the granular picture.
One thing that caught my attention: early access partners described the model as finishing complex tasks where previous Sonnet models would stop short. They also noted it checks its own output without being explicitly prompted to do so. That second part is significant. Anyone who’s worked with autonomous AI agents knows that self-correction without a human in the loop is where things tend to fall apart. If Sonnet 5 genuinely does that better, it’s a real-world workflow improvement, not just a benchmark number going up.
On the safety side, Anthropic’s pre-deployment evaluations found Sonnet 5 showed lower rates of undesirable behaviors compared to Sonnet 4.6. It’s better at refusing malicious requests and resisting prompt injection attacks. There’s one interesting caveat: the model does show somewhat higher rates of misaligned behavior compared to Opus 4.8 on an automated behavioral audit. The company was transparent about this in its system card. On cybersecurity tasks specifically, Sonnet 5 performs substantially poorer than Opus models, which is reassuring given the ongoing conversations around AI safety.
Pricing is where Sonnet 5 gets interesting for independent developers and smaller teams. It launches at $2 per million input tokens and $10 per million output tokens through August 31, 2026, then adjusts to $3 and $15 respectively. That’s notably lower than what Opus-class access typically costs, and it’s landing as the default model for Free and Pro plans while remaining available to Max, Team, and Enterprise users.
If you’ve been running Sonnet 4.6 in production for agentic workflows, the upgrade path seems straightforward. If you’ve been paying Opus prices for tasks that don’t actually need Opus-level capability, Sonnet 5 looks like the kind of correction the pricing ladder needed.
The real test will be what developers report once they’ve had a few weeks with it in varied production environments. Benchmark performance and early access testimonials only carry you so far. But from what Anthropic has published, this feels like the Sonnet line finally closing the gap with Opus on the exact dimensions that matter for autonomous work: planning, tool use, and knowing when to stop.