Benchmarks
Benchmarks

FrontierCode (Extended)

Codingunit % · normalized over [0, 75]

FrontierCode v1.1 Extended set — Cognition's agentic coding benchmark built from real open-source pull requests, scored on the wider Extended split rather than the 150-task Main set. Not comparable with the Main-set scores filed under `frontiercode`.

#ModelSourceScoreNormalized
Score distribution
7 tracked results across the normalization window
075
Score vs. parameters
Open-weights models, log-x params
No open-weights models with disclosed parameter counts have a score here yet.