Mar 17, 2026
GPT 5.4 Nano evaluated on our full benchmark suite
We evaluated GPT 5.4 Nano across our full benchmark suite.
- GPT 5.4 Nano ranks #18 on both our Vals Index and #16 on the Vals Multimodal Index.
- Its performance on Vibe Code Bench was impressive, coming in at #10, and beating out both Gemini 3 Flash (12/25) and Claude Haiku 4.5 (Thinking)
- On SWE-bench Verified, it scored 69.0%, comparable to GPT 5 Codex (69.4%) at a 20x lower cost.
- At ~$0.05 per test on the index, the model offers strong cost-effectiveness for its tier.
Evaluations were run via the official OpenAI API using βhighβ reasoning.