LabsCognition
Cognition
Track Cognition model scores across public AI benchmarks including FrontierCode, Terminal-Bench 2.1, and SWE-bench ML. Each result is shown one benchmark at a time, with source links and evaluation dates — no blended score or composite ranking. 2 models tracked, spanning SWE.
Models 2
Progress by benchmark
Show progress on
Single benchmark only
This view shows FrontierCode (weighted score (Main)) only. Other benchmarks use different metrics and are not directly comparable.
Progress matrix
Scores are not normalised across benchmarks. Each column uses its own metric. Compare columns independently.