Release Date: Sep 22, 2026

Developer OpenAI Β πŸ‡ΊπŸ‡Έ
Context Window 1M
Max Output Tokens 128k
Token Costs (in/out) $0.10/0.50
Weights Private
Input Modalities

Accuracy

58.45 % Β± 1.11

Cost / Test (Vals Index)

$ 0.419

Latency

26 min 30 s

Fallback Rate 0.00%
Refusal Rate 0.05%
Vals Index
BenchmarksAccuracyRankings

0.0%

Β±1.11
20/65

0.0%

Β±4.42
21/68

0.0%

Β±2.81
11/65

0.0%

Β±0.23
34/68

0.0%

Β±3.19
36/68

0.0%

Β±2.30
35/98

0.0%

Β±1.95
36/100

0.0%

Β±2.66
14/16

0.0%

Β±4.82
13/40

0.0%

Β±3.38
26/85

0.0%

Β±3.21
21/28

0.0%

Β±3.38
15/103

0.0%

Β±2.59
14/16

0.0%

Β±8.87
16/33

0.0%

Β±0.50
17/52

0.0%

Β±1.72
22/73
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : OpenAI
Temperature: Default
Top P: Default
Top K: Default
Max Output Tokens: 128,000
Reasoning Effort: max

Updates

Sep 22, 2026

We evaluated OpenAI’s new GPT-6 Luna across our benchmark suite.

No fallback models were used; refusals and provider policy blocks are counted as failed tasks. Terminal-Bench 2.1 saw 1 refusal (0.38% of tasks), which did not change its score.

GPT-6 Luna is priced at $0.10 per million input tokens and $0.50 per million output tokens ($0.20 / $0.75 for long-context requests). The model has a 1M-token context window and 128k max output tokens. Evaluations were run with reasoning effort set to β€œmax”.

Congrats to the OpenAI team on the release!