Anthropic · Claude Sonnet 5 · June 2026

Claude Sonnet 5 — near-Opus capability with intelligence and speed in balance

  • Million-token context
  • Extra-long single replies
  • Tunable depth tiers
  • Stronger tools and coding
Claude Sonnet 5

What is Claude Sonnet 5?

Claude Sonnet 5 is Anthropic’s Sonnet released in June 2026. It’s built for scaled agents: planning, browser and terminal tools, and driving to a result. Anthropic reports clear gains over Claude Sonnet 4.6 on many agent evals, and at some depth settings it can approach Claude Opus 4.8. When you need the Opus ceiling, look at Claude Opus 5; for single ultra-long, ultra-hard jobs, compare Claude Fable 5. On iMini Agent, pick Claude Sonnet 5 to use it.

Vendor
Anthropic
Released
June 2026
Context
1M tokens
Max output
128K tokens
Deep thinking
Tunable tiers
Best as
High-frequency agent & everyday coding primary

What’s new versus Claude Sonnet 4.6

Agent completion, cost–performance range, and tool collaboration—what to check before making Sonnet your primary.

Large jump in agent capability

Reasoning, tool use, coding, and knowledge work all lift versus 4.6; at the right depth, some jobs approach Opus 4.8 completion.

Longer chains finished autonomously

Complex jobs stop mid-way less often and self-check outputs more often—suited to multi-step tool flows as an everyday primary.

Wider cost–performance range

Multiple depth tiers cover efficient through high-completion choices, so one product line can tune by task weight without sending everything to Opus.

Steadier agent-scenario behavior

Official alignment and safety evals show fewer improper behaviors and stronger resistance to injection and malicious requests—suited as a Sonnet primary for scaled deploy.

Official evaluations and safety audits

Results Anthropic published with Claude Sonnet 5, compared with Claude Sonnet 4.6 and Claude Opus 4.8—covering software engineering, multidisciplinary reasoning, computer use, and safety audits.

Claude Sonnet 5 vs Sonnet 4.6 and Opus 4.8 multi-task evaluation table
Multi-task overview. SWE-bench Pro 63.2% and Terminal-Bench 2.1 80.4% both clearly beat Claude Sonnet 4.6 (58.1% / 67.0%); GDPval-AA v2 at 1618 nearly matches Claude Opus 4.8 at 1615—the key row for whether Sonnet 5 can take former Opus-tier work at mid-tier cost.
Behavioral alignment audit: Claude Sonnet 5 misalignment score
Behavioral alignment audits count deception, induced misuse, and related misalignment—lower is better. Claude Sonnet 5 scores 2.53, below Claude Sonnet 4.6 at 2.89; read this alongside capability scores before putting it in automation.
Firefox 147 exploit-development evaluation: Claude Sonnet 5 results
Firefox 147 exploit development measures whether a model can write real usable browser attack code. Claude Sonnet 5 full-exploit success is 0.0%, register control only 13.2%—far below Claude Mythos 5 at 88.4%. Anthropic uses this to show offensive cyber capability is heavily suppressed on Sonnet 5.

Three common workflows

Claude Sonnet 5 fits high-frequency, scalable work that still needs reliable tool collaboration.

Day-to-day coding & tool agents

For routine feature work, test repair, and browser/terminal tool collaboration. Use it as the team’s everyday primary; leave ceiling long-horizon jobs for Opus 5.

Knowledge work & material synthesis

For research summaries, table work, and constrained conclusions. Focus on stable format and checkable evidence.

Scaled multi-step automation

For support assist, internal processes, and repetitive multi-step jobs. Pin success criteria and control latency via depth tiers.

How to choose vs Claude Opus 5 and Claude Opus 4.8

All three are Claude workhorses. The differences are ceiling tier, high-frequency cost, and whether you need the current flagship.

DimensionClaude Sonnet 5Claude Opus 5Claude Opus 4.8
Context1M tokens1M tokens1M tokens
ReasoningTunable depth tiersDeep thinking built in, tunableUsually high depth, tunable
Coding / agentsScaled agents & everyday codingFlagship primary for long-horizon coding & multi-step agentsOpus 4.8 ceiling
Long tool runsBetter as a high-frequency primaryStronger goal holding and completionSteady Opus-tier collaboration
Speed postureThroughput and scale firstCompletion quality firstQuality and experience in balance
Prefer whenNew high-frequency agents and everyday codingNew long-horizon coding and knowledge-work flagshipStill bound to 4.8 workflows, or as a safety fallback

Using Claude Sonnet 5 on iMini

Free quota, model naming, key specs, and how it differs from Opus.

Use Claude Sonnet 5 free on iMini Agent

Million-token context, tunable depth, built as a scaled-agent primary.

Open Claude Sonnet 5

No Anthropic API key required.