AI Model Profile

Claude Sonnet 5

Claude Sonnet 5 is Anthropic's speed-and-intelligence balance in the Claude 5 line, with a 1M-token context window and 128K maximum output.

Family
Claude
Release date
Unknown
Status
Current
Context window
1M tokens
Output limit
128K tokens
API
yes
Open weights
no
Local/self-hosted
no
Pricing
$3 per million input tokens, $15 per million output tokens. Introductory pricing of $2 per million input and $10 per million output applies through 31 August 2026; confirm the current rate on the official pricing page after that date.
Verification
verified

Verification & Sources

Status
Verified
Source links
2
Freshness
Verified July 25, 2026
Last verified
July 25, 2026
Last updated
July 25, 2026

Key source checks

Suggest a correction

Form submissions, correction notes, score details, URLs, and analytics events may be stored for editorial review, spam prevention, product improvement, and follow-up. Do not submit secrets, unreleased financials, private customer data, or regulated personal data through these forms.

Benchmark Caveat

Benchmarks and provider capability notes are directional, not universal rankings. Results can shift with prompts, tool use, latency targets, pricing tier, safety filters, context length, and the workload mix a real team runs.

See the linked official model, docs, model-card, or pricing source for provider-published capability notes.

Best for

Teams that want near-frontier Claude capability with faster latency and lower cost than Opus 5.

Skip if

Skip if you need Opus 5's higher capability ceiling, or if Haiku 4.5 meets the task at lower cost.

Strengths

Anthropic lists Sonnet 5 with a 1M-token context window, 128K max output, adaptive thinking, and fast comparative latency. On the Message Batches API it supports up to 300K output tokens using the output-300k-2026-03-24 beta header.

Weaknesses

Reliable knowledge cutoff is January 2026, earlier than Opus 5. Extended thinking is not supported.

Agent suitability

Useful for agent workflows when the provider supports tool use, long context, structured outputs, or workflow-specific APIs.

Kingy AI take

Use this as a source-backed shortlist candidate, not a universal ranking. Re-check official provider docs and run a task-specific trial before production adoption.

Coding notes

Use official docs and live evals before selecting this model for production coding workflows.

Reasoning notes

Provider capability notes are useful but should be validated on representative prompts and tools.

Creative notes

Use a small creative test set before standardizing outputs for brand, media, or customer-facing work.

Research notes

Track release notes and model lifecycle notices because availability and aliases can change.

API pricing notes

Check the official pricing page before budget decisions; Kingy does not freeze token, credit, or subscription prices in model cards.

License notes

Commercial/API terms apply unless the linked official source states otherwise.

Hardware requirements

Cloud/API model; local hardware requirements are not published as a self-hosted path.