Model
Model explorer

GPT-5.5

CLOSED
OpenAI · GPT-5.5 family · released Apr 23, 2026

Codename Spud; ChatGPT/Codex launched Apr 23, 2026 with API following ~1 day later once safeguards were finalized.

ReasoningCodingVisionFunction callingTool useAgentic
2634.5
Elo · rank #23
Parameters
Undisclosed
Active params
Context
1.05M tokens
Architecture
Undisclosed architecture (presumed dense transformer, reasoning-tuned); codename "Spud"; Thinking + Pro tiers
License
Proprietary
Languages
API price (in/out)
$5 / $30
Modalities
text · vision
Benchmark results
Bar shows position within the tracked field; marker = field best
ARC-AGI-2Reasoning85.0%#5
best: GPT-6 Astra · 95.0%
BrowseCompAgents84.4%#9
best: Kimi K3 · 91.2%
CyberGymCoding81.8%#4
best: GLM-5.3 · 84.5%
best: GPT-5.6 Sol · 89.0%
best: GPT-5.6 Sol · 83.0%
GDPvalAgents84.9%#2
best: Seed 2.1 Pro · 87.9%
GPQA DiamondReasoning93.6%#8
best: GPT-6 Astra · 96.0%
Humanity's Last ExamReasoning41.4%#25
best: Claude Opus 5 · 64.7%
LiveCodeBenchCoding85.3%#23
best: DeepSeek-V4-Pro (Think Max) · 93.5%
OSWorld-VerifiedAgents78.7%#7
best: Claude Fable 5 · 85.0%
SWE-bench ProCoding58.6%#20
best: Claude Fable 5.1 · 81.2%
SWE-bench VerifiedCoding82.6%#8
best: Claude Opus 5 · 96.0%
τ²-Bench TelecomAgents98.0%#9
best: Claude Opus 4.6 · 99.3%
Terminal-Bench 2.0Coding82.7%#15
best: Gemini 3.8 Flash · 89.4%
Run it locally
Closed weights — available via API only. No local deployment.
Input / M tok
$5
Output / M tok
$30
API price $5/$30 · each benchmark row carries its own source badge (see methodology)