Dec 11, 2025
GPT 5.2 Tops Vals Index
GPT 5.2 is the new state of the art on our Vals Index, showing strong performance across domains, especially coding.
Most impressive was the model setting a new state-of-the-art on Vibe Code Bench by an extremely large margin, from 24.6% to 41.31%. It also got first on our IOI, Terminal-Bench, and SWE-bench Verified.
This performance improvement does come with an increased cost across the board - the model is priced at 14, compared to 10 for its predecessor. It also tends to see increased token usage, especially on longer running agentic tasks. With the βhighβ or the new βxhighβ reasoning modes enabled, you may also see very long response times - for particularly tricky questions, it would think for more than 30 minutes.
Overall, the model will be a powerhouse for users seeking strong performance and reliability, particularly on complex reasoning and coding tasks.