FrontierCode 1.1 Extended

Cognition's 150-task Extended subset of the FrontierCode 1.1 software-engineering benchmark.

  • Not in index
  • Coding
ModelsModels with a result.7
Top resultBest result on this test.N/A
Top-3 spreadPoints from first to third.N/A
YearYear of release.N/A

Result and price

020406080100$1$3$10$30$100RESULTOUTPUT PRICE PER 1M TOKENS · LOG SCALE

Results

7 results
#ModelResultThe score from the source.
GPT-6 Astra64.5%
Grok 4.661.3%
GPT-5.6 Sol60.6%
GPT-5.6 Luna55.1%