Claude Sonnet 4.6 with extended thinking enabled for tougher coding, planning, and multi‑tool tasks. Ideal for long‑horizon agent workflows and complex problem solving.
Added Feb 17, 2026
Context Window
1.0M
Max Output
128.0K
Avg output tokens (7d)
1.2K tokens
Input Price (Auto)
$3.00/1M
Output Price (Auto)
$15.00/1M
Cache Read (Auto)
$0.30/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
30.1
Coding Index
63.0
Agentic Index
31.8
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
20.1%
Better than 41% of models compared
Harvey LAB-AA
Legal agentic work criterion pass rate
86.0%
Better than 44% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
1063 Elo
Better than 54% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1220 Elo
Better than 63% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
15.8%
Better than 59% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
80.0%
Better than 87% of models compared
MLCR-AA
Medical long-context reasoning
24.4%
Better than 77% of models compared
Reasoning
HLE
Humanity's Last Exam
33.6%
Better than 81% of models compared
IFBench
Instruction-following benchmark
56.6%
Better than 67% of models compared
CritPt
Research-level physics reasoning
3.1%
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
3.0%
Better than 53% of models compared
SciCode
Python programming for scientific computing
50.1%
Better than 53% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
40.9%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
48.4%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
87.5%
Better than 84% of models compared
Terminal-Bench Hard (legacy)
Agentic coding and terminal use
53.0%
Better than 96% of models compared
T²-Bench Telecom (legacy)
Conversational AI agents in dual-control scenarios
75.7%
Better than 67% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
80.0%
Better than 87% of models compared
GDPval-AA (unversioned / legacy)
Economically valuable tasks
36.0%
Last updated Oct 1, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Claude Sonnet 4.6 Thinking with similar models from the same provider or model family.
Claude Sonnet 5
anthropic/claude-sonnet-5Anthropic's newest Sonnet model, built for stronger coding, agentic workflows, long-context reasoning, computer use, and practical knowledge work.
Claude Sonnet 5 Thinking
anthropic/claude-sonnet-5:thinkingClaude Sonnet 5 with adaptive thinking enabled for complex coding, planning, and agentic work.
Claude Sonnet Latest
anthropic/claude-sonnet-latestCompatibility alias that routes to the newest version of Claude Sonnet. Currently routes to Claude Sonnet 5 Thinking.
Claude Sonnet 4.6
anthropic/claude-sonnet-4.6Claude Sonnet 4.6 is Anthropic's most capable Sonnet yet — a full upgrade across coding, computer use, long-context reasoning, agent planning, and design. Supports up to a 1M-token context window at Sonnet pricing.
Claude Opus 5.5
anthropic/claude-opus-5.5Claude Opus 5.5 is Anthropic's flagship model for agentic coding, computer use, complex knowledge work, and long-running tasks, with improved efficiency and more natural communication.
Claude Fable 5.1
anthropic/claude-fable-5.1Claude Fable 5.1 improves on Fable 5 across agentic coding, long-running workflows, front-end and visual code generation, finance, analysis, and knowledge work, with more concise plans and summaries. Anthropic retains prompts and outputs for 30 days; Zero Data Retention is not available.