GPT-6 Luna (low)

GPT-6 Luna (low) is a reasoning model from OpenAI. It has 2 scored results in 1 category. It has no index.

  • Reasoning
  • Input: file · image · text
partialShows if the model has enough results for an index.
IndexOverall score out of 100.UnrankedMedian 50.7
CoverageShare of the index weight with results.15%
SpeedOutput tokens per second.57/sMedian 62/s
First tokenSeconds to the first output token.3.0 sMedian 1.6 s
Input / 1MUS dollars per 1M input tokens.$0.1Median $0.6
Output / 1MUS dollars per 1M output tokens.$0.5Median $2.5
ContextMaximum tokens in one request.1.05MMedian 262K
EloLMArena rating and rank.N/AMedian 1427

Details

From OpenRouter
API IDThe model ID on OpenRouter.
openai/gpt-6-luna
Max outputMaximum output tokens per request.
128K
Knowledge cutoffLast date of training data.
N/A
ToolsTool calls on OpenRouter.
Yes
JSON outputOutput follows a given JSON schema.
Yes
Cache read / 1MUS dollars per 1M cached tokens.
$0.01
WeightsLink to the published weights.
N/A
RetiresDate OpenRouter removes the model.
N/A

Capabilities

Score per category, out of 100.
Out of 100
AgenticMulti-step tasks with tools.None countedN/A—
CodingCode writing and repair.None countedN/A—
ReasoningLogic problems and puzzles.2 counted68.7—
MultimodalTasks with images and text.None countedN/A—
KnowledgeFacts and expert knowledge.None countedN/A—
MultilingualTasks in many languages.None countedN/A—
InstructionTasks with strict rules in the prompt.None countedN/A—
MathMath problems.None countedN/A—

Results

2 counted
BenchmarkThe test name.ResultThe score from the source.PlacePlace among results on this benchmark.
jevBench15Reasoning40.5%34 / 103
jevBench14Reasoning35.1%33 / 89

Sources

BenchLM benchmark aggregationCC BY-NC 4.0 · Data from BenchLM.aiOpenRouter, collected directlyNo licence stated

More from OpenAI

All models
GPT-6 Astra82.1GPT-6.1 Sol79.9GPT-5.5 Pro77.3GPT-6 Sol76.3GPT-5.4 Pro75.8