Optimized for speed and ideal for applications requiring low latency. The fastest and most efficient GPT-5 variant.
Added Aug 7, 2025
Context Window
400.0K
Max Output
128.0K
Avg output tokens (7d)
1.5K tokens
Input Price (Auto)
$0.050/1M
Output Price (Auto)
$0.40/1M
Cache Read (Auto)
$0.0050/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
13.7
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
67.6%
Better than 45% of models compared
HLE
Humanity's Last Exam
9.5%
Better than 52% of models compared
IFBench
Instruction-following benchmark
67.6%
Better than 81% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
36.5%
Better than 44% of models compared
AA-LCR
Long context reasoning evaluation
45.0%
Better than 43% of models compared
Coding
Terminal-Bench Hard
Agentic coding and terminal use
12.1%
Better than 48% of models compared
LiveCodeBench
Contamination-free coding benchmark
78.9%
Better than 91% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
83.7%
Better than 80% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
78.0%
Better than 59% of models compared
Last updated Sep 7, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare GPT 5 Nano with similar models from the same provider or model family.
GPT 5.4 Nano
openai/gpt-5.4-nanoGPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks.
GPT 4.1 Nano
openai/gpt-4.1-nanoCheapest model in the GPT-4.1 series. Huge context window with fast throughput and low latency.
GPT 6 Astra Pro
openai/gpt-6-astra-proGPT 6 Astra in Pro reasoning mode. Uses additional model work for difficult tasks, with higher latency and token usage at the same per-token rates. Reasoning effort remains independently configurable.
GPT 6 Astra
openai/gpt-6-astraOpenAI's most capable model for the hardest end-to-end work, with state-of-the-art performance in computer use, browsing, software engineering, science, and professional tasks. Astra is designed to carry long, multistep workflows across code, browsers, and professional software through to a finished result.
GPT 5.6 Luna
openai/gpt-5.6-lunaGPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume chat, classification, lightweight agentic workflows, and latency-sensitive reasoning tasks.
GPT 5.6 Luna Pro
openai/gpt-5.6-luna-proGPT-5.6 Luna with Pro reasoning mode. Pro mode performs more model work for difficult tasks; reasoning effort remains independently configurable and defaults to medium.