GPT 5.1 Codex Max

Release Date: Dec 4, 2025

Developer OpenAIย ๐Ÿ‡บ๐Ÿ‡ธ
Context Window 400k
Max Output Tokens 128k
Token Costs (in/out) $1.25/10.00
Weights Private
Input Modalities

Accuracy

42.38 %

Avg. Cost (In/Out)

$ 1.25 / $ 10.00

Latency

178 min 55 s

Vals Index
BenchmarksAccuracyRankings

0.0%

ยฑ3.84
52/84

0.0%

ยฑ7.00
22/62

0.0%

ยฑ1.05
45/138
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : OpenAI
Temperature: Default
Top P: Default
Top K: Default
Max Output Tokens: 128,000
Reasoning Effort: high

Updates

Dec 4, 2025

Weโ€™ve evaluated GPT 5.1 Codex Max on our coding benchmarks. It boasts a +9.5% performance boost on VibeCodeBench (#3), +1% performance on SWE-bench Verified (#4), and a slight regression on Terminal-Bench.

This is not the fastest model in the shed. Itโ€™s 3x slower on SWE-bench Verified, 10x on VCB. This is a model for when you need the absolute best, not when you need it quickly.

Correlation isnโ€™t causationโ€ฆ โ€ฆbut the more suffixes we tack onto these models, the longer they seem to take. Excited to test 5.25 Codex Max Ultra High Supreme Edition

The model was run with reasoning high and verbosity medium through OpenAIโ€™s API. Results on IOI and LCB will be released soon!