Anthropic
Claude Sonnet 4.6
Anthropic's coding + agent workhorse. 1M-token context.
ReasoningLong contextCode
$ / 1M input
$3.000
What you pay per million prompt tokens.
$ / 1M output
$15.000
What you pay per million completion tokens.
Blended (70 / 30)
$6.600
Typical chat workload mix.
Ratings & benchmarks
Snapshot 2026-04-28Overall
4.8 / 5
Composite of intelligence + reliability.
Output speed
65 t/s
Output tokens per second under typical load.
Time to first token
1.5 s
Lower is better. Reasoning models naturally stretch this.
Value
3.5 / 5
Intelligence per dollar at typical mix.
Context window
1,000,000 tokens