Claude Sonnet 5 Release Notes: Enhanced Agentic Capabilities and Pricing
Claude Sonnet 5 delivers higher autonomy for agentic workflows
Claude Sonnet 5 is designed as the most agentic model in the Sonnet family, capable of autonomous planning, browser and terminal tool use, and complex multi-step execution. It narrows the performance gap between the mid-tier Sonnet models and the high-end Opus-class models, specifically targeting improvements in reasoning, coding, and knowledge work.
Key Performance Gains
Sonnet 5 provides a strict improvement over Sonnet 4.6 across agentic search (BrowseComp) and computer use (OSWorld-Verified) evaluations. While Opus 4.8 remains the superior choice for maximum accuracy, Sonnet 5 allows developers to balance cost and performance by adjusting "effort levels."
Early access partners report significant improvements in follow-through, including:
- Software Engineering: Handling sustained coding and debugging in "messy" technical contexts.
- End-to-End Automation: Completing multi-part tasks (e.g., updating Salesforce tiers and sending announcements) without stalling.
- Autonomous Bug Fixing: Writing reproducing tests and implementing fixes in a single pass without explicit prompting.
- Brownfield Code: Tracing failures to root causes in legacy codebases rather than patching symptoms.
Safety and Cybersecurity Guardrails
Sonnet 5 demonstrates a lower rate of undesirable behaviors, hallucinations, and sycophancy compared to Sonnet 4.6. It is specifically optimized to resist prompt injection attacks and malicious requests in agentic contexts.
Cybersecurity Limitations
Anthropic explicitly states that Sonnet 5 has a substantially lower ability to perform dangerous cybersecurity tasks than Opus 4.8 or Mythos 5. In evaluations involving the development of software exploits for Firefox 147, Sonnet 5 failed to develop any full working exploits. Due to a slight increase in partial success rates over Sonnet 4.6—attributed to general intelligence gains—Anthropic has enabled real-time cyber safeguards by default.
Pricing and Availability
Claude Sonnet 5 is available across all plans (Free, Pro, Max, Team, and Enterprise) and is the default model for Free and Pro users. It is also integrated into Claude Code and the Claude Platform.
Token Pricing
| Period | Input Tokens (per 1M) | Output Tokens (per 1M) |
|---|---|---|
| Through Aug 31, 2026 | $2 | $10 |
| After Aug 31, 2026 | $3 | $15 |
Note on Tokenization: Sonnet 5 uses an updated tokenizer. This results in the same input mapping to roughly 1.0–1.35× more tokens than previous versions, which impacts the effective cost per request.
Community Analysis and Counterpoints
While Anthropic positions Sonnet 5 as a cost-effective alternative to Opus, community feedback from Hacker News highlights several critical trade-offs:
Cost-Performance Efficiency
Multiple users argue that Sonnet 5 is not Pareto optimal at higher effort levels.
"The cost per task chart is telling me that I should never use Sonnet 5 above medium effort level - Opus always performs better for a given cost."
Agentic vs. Assisted Development
Some developers report a decline in "assisted" development (where the human leads) as models are optimized for "agentic" development (where the model leads).
"I have found that the more models are optimized for fully agentic development, the worse they get at assisted development and often start doing too much despite very strict/specific instructions."
Comparative Benchmarks
Independent testers have noted varying results compared to other frontier models:
- GLM-5.2: Some users report Sonnet 5 is comparable to GLM-5.2 in coding and agentic ability, though it may be faster and less verbose. Others claim it is less price-performant than GLM-5.2.
- Reliability: Some reports suggest Sonnet 5 can "spin" and waste tokens on complex HPC kernels where Opus previously succeeded.
Sources
Related
- Dispatch
- Dispatch
- Dispatch
- Dispatch
- Dispatch