Anthropic Claude Opus 4.8 with thinking enabled.
Added May 28, 2026
Context Window
1.0M
Max Output
128.0K
Avg output tokens (7d)
4.5K tokens
Input Price (Auto)
$5.00/1M
Output Price (Auto)
$25.00/1M
Cache Read (Auto)
$0.50/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
41.8
Coding Index
74.3
Agentic Index
41.9
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
45.6%
Better than 61% of models compared
Harvey LAB-AA
Legal agentic work criterion pass rate
91.1%
Better than 65% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
1321 Elo
Better than 72% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1438 Elo
Better than 80% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
22.8%
Better than 78% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
77.7%
Better than 79% of models compared
MLCR-AA
Medical long-context reasoning
45.6%
Better than 86% of models compared
Reasoning
HLE
Humanity's Last Exam
48.7%
Better than 95% of models compared
IFBench
Instruction-following benchmark
62.2%
Better than 73% of models compared
CritPt
Research-level physics reasoning
20.9%
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
21.7%
Better than 76% of models compared
SciCode
Python programming for scientific computing
54.4%
Better than 74% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
48.8%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
39.3%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
92.0%
Better than 93% of models compared
Terminal-Bench Hard (legacy)
Agentic coding and terminal use
58.3%
Better than 98% of models compared
T²-Bench Telecom (legacy)
Conversational AI agents in dual-control scenarios
94.4%
Better than 93% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
77.7%
Better than 79% of models compared
GDPval-AA (unversioned / legacy)
Economically valuable tasks
46.9%
Last updated Oct 1, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Claude Opus 4.8 Thinking with similar models from the same provider or model family.
Claude Opus 5.5
anthropic/claude-opus-5.5Claude Opus 5.5 is Anthropic's flagship model for agentic coding, computer use, complex knowledge work, and long-running tasks, with improved efficiency and more natural communication.
Claude Opus 5
anthropic/claude-opus-5Claude Opus 5 is Anthropic's flagship model for advanced coding, long-running agentic tasks, research, and complex knowledge work.
Claude Opus 4.8
anthropic/claude-opus-4.8Anthropic Claude Opus 4.8 with text, image, and file inputs and a 1M-token context window.
Claude 4.7 Opus
anthropic/claude-opus-4.7Claude Opus 4.7 is a major upgrade for advanced software engineering, long-running complex tasks, and high-resolution vision understanding.
Claude 4.7 Opus Thinking
anthropic/claude-opus-4.7:thinkingClaude Opus 4.7 with thinking enabled (default budget: 16k tokens).
Claude Opus Latest
anthropic/claude-opus-latestCompatibility alias that routes to the newest version of Claude Opus. Currently routes to Claude Opus 5.