Anthropic
Claude Sonnet 4.6
The record carries a second output ceiling for this model — 300,000 tokens, marked beta — and says nothing about what unlocks it. Until it does, size a long-output job against the standard limit.
The published record
Every number below was read from the provider's own documentation on the date shown. benchr does not restate a figure it has not seen published.
- API identifier
claude-sonnet-4-6- Context window
- 1MtokensSource
- Maximum output
- 64KtokensSource
- Input
- $3.00per 1M tokensSource
- Output
- $15.00per 1M tokensSource
- Cached input
- $0.3per 1M tokensSource
- Released
- February 17, 2026Source
- License
- Proprietary
Record verified July 30, 2026
What the record says
No fast-mode tier (fast mode is Opus-only). Cache-hit input $0.30 confirmed on the official pricing page June 12, 2026; official tentative retirement floor: not sooner than February 17, 2027. Benchmark values read from the official Sonnet 4.6 System Card (Table 2.1.A). SWE-bench Verified 79.6% averaged over 25 trials (80.2% with a stated prompt modification). Anthropic flags possible AIME 2025 contamination. Anthropic reports MMMLU, not plain MMLU.
What benchr has documented it doing
Capabilities in the benchr ledger that name this model. Documented means a provider says it works and benchr recorded where. Verified means benchr ran it.
2 capabilities name this model. All 2 are officially supported by the provider; benchr has run 0 of them itself.
Which one to use
The rest of the family, with the two numbers that usually decide it. The current model is marked.
| Models | Input | Output | Context window |
|---|---|---|---|
| Claude Haiku 4.5 | $1.00 | $5.00 | 200K |
| Claude Sonnet 5 | $2.00 | $10.00 | 1M |
| Claude Sonnet 4.6 This page | $3.00 | $15.00 | 1M |
| Claude Opus 4.7 | $5.00 | $25.00 | 1M |
| Claude Opus 4.8 | $5.00 | $25.00 | 1M |
| Claude Opus 5 | $5.00 | $25.00 | 1M |
| Claude Fable 5 | $10.00 | $50.00 | 1M |
| Claude Fable 5.1 | $10.00 | $50.00 | 1M |
benchr's read
Worth it for
- Production default
- Cost-effective coding
- Bulk content tasks
- Daily-driver API workloads
Look elsewhere if
- You need frontier-grade reasoning
Prices, limits, and identifiers above are provider facts. This section is benchr's judgement about them.
What benchr has written about it
Pieces that name this model, newest first. The ones written about this model come before the ones that mention it in passing.
What is not on this page
Stated rather than filled in.
- No first-token or tokens-per-second figure is recorded. benchr has not measured it and the provider does not publish one.
- benchr has not run any capability on this model itself. Everything in the capability section is documentation, not a test result.