OpenAI's precusor to ChatGPT-4o. Great on English text and code, with significant improvements on text in non-English languages.
Context Window
128.0K
Max Output
16.4K
Avg output tokens (7d)
45 tokens
Input Price (Auto)
$2.50/1M
Output Price (Auto)
$10.00/1M
Cache Read (Auto)
$1.25/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
3.9
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
52.1%
Better than 26% of models compared
HLE
Humanity's Last Exam
2.3%
Better than 0% of models compared
IFBench
Instruction-following benchmark
36.0%
Better than 27% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
28.9%
Better than 36% of models compared
AA-LCR
Long context reasoning evaluation
41.0%
Better than 40% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
Terminal-Bench Hard
Agentic coding and terminal use
8.3%
Better than 42% of models compared
LiveCodeBench
Contamination-free coding benchmark
31.7%
Better than 38% of models compared
Math
AIME
American Invitational Mathematics Examination
11.7%
Better than 37% of models compared
Math-500
Diverse mathematical problem solving benchmark
79.5%
Better than 45% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
23.7%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
58.4%
Last updated Sep 7, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GPT-4o (2024-08-06) with similar models from the same provider or model family.
GPT-4o (2024-11-20)
openai/gpt-4o-2024-11-20OpenAI's precusor to ChatGPT-4o. Great on English text and code, with significant improvements on text in non-English languages.
GPT 6 Astra Pro
openai/gpt-6-astra-proGPT 6 Astra in Pro reasoning mode. Uses additional model work for difficult tasks, with higher latency and token usage at the same per-token rates. Reasoning effort remains independently configurable.
GPT 6 Astra
openai/gpt-6-astraOpenAI's most capable model for the hardest end-to-end work, with state-of-the-art performance in computer use, browsing, software engineering, science, and professional tasks. Astra is designed to carry long, multistep workflows across code, browsers, and professional software through to a finished result.
GPT 5.6 Luna
openai/gpt-5.6-lunaGPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume chat, classification, lightweight agentic workflows, and latency-sensitive reasoning tasks.
GPT 5.6 Luna Pro
openai/gpt-5.6-luna-proGPT-5.6 Luna with Pro reasoning mode. Pro mode performs more model work for difficult tasks; reasoning effort remains independently configurable and defaults to medium.
GPT 5.6 Sol
openai/gpt-5.6-solGPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, agentic workflows, command-line work, and multi-step professional tasks.