Model
Model explorer

Claude Opus 4.5 (High)

CLOSED
Anthropic · Claude 4 family · released Nov 24, 2025

Efficiency-focused flagship; API default (high) effort; record 80.9% SWE-bench Verified with major price cut vs Opus 4.1.

ReasoningCodingVisionFunction callingTool useAgentic
2261.5
Elo · rank #53
Parameters
Undisclosed
Active params
Context
200K tokens
Architecture
Dense Transformer (undisclosed configuration)
License
Proprietary
Languages
API price (in/out)
$5 / $25
Modalities
text · vision
Benchmark results
Bar shows position within the tracked field; marker = field best
AIMEMath92.8%#33
best: GPT-5.2 · 100.0%
ARC-AGI-1Reasoning80.0%#10
best: GPT-6 Astra · 98.5%
ARC-AGI-2Reasoning37.6%#17
best: GPT-6 Astra · 95.0%
CyberGymCoding50.6%#12
best: GLM-5.3 · 84.5%
GPQA DiamondReasoning87.0%#51
best: GPT-6 Astra · 96.0%
Humanity's Last ExamReasoning43.2%#20
best: Claude Opus 5 · 64.7%
LiveCodeBenchCoding83.7%#33
best: DeepSeek-V4-Pro (Think Max) · 93.5%
MMLU-ProKnowledge87.3%#11
best: Claude Fable 5 · 91.5%
MMMLUKnowledge90.8%#5
best: Gemini 3.1 Pro · 92.6%
MMMUVision80.7%#23
best: Claude Fable 5 · 89.3%
OSWorld-VerifiedAgents66.3%#19
best: Claude Fable 5 · 85.0%
SWE-bench ProCoding52.0%#47
best: Claude Fable 5.1 · 81.2%
SWE-bench VerifiedCoding80.9%#11
best: Claude Opus 5 · 96.0%
τ²-Bench RetailAgents88.9%#3
best: Claude Sonnet 4.6 · 91.7%
τ²-Bench TelecomAgents98.2%#6
best: Claude Opus 4.6 · 99.3%
Terminal-Bench 2.0Coding59.3%#45
best: Gemini 3.8 Flash · 89.4%
Terminal-BenchCoding59.3%#1
best: this model · 59.3%
API price $5/$25 · each benchmark row carries its own source badge (see methodology)