Artificial Analysis Terminal-Bench v4.0

An independently evaluated Terminal-Bench v4.0 result from Artificial Analysis, one of the ten components of its Intelligence Index v4.3.

ModelsModels with a result.23
Top resultBest result on this test.63.6%Claude Sonnet 5.5
Top-3 spreadPoints from first to third.4.5 pts
YearYear of release.2026

Result and price

020406080100$0.3$1$3$10$30$100RESULTOUTPUT PRICE PER 1M TOKENS · LOG SCALE

Results

23 results
#ModelResultThe score from the source.
07GPT-6 Sol43.9%
08GLM-5.341.9%
15Grok 4.725.8%
17Kimi K312.6%
17GPT-6 Luna12.6%
21Inkling1.0%