Celeris 1 is a diffusion language model built for ultra-low-latency classification, extraction, judging, query rewriting, and other short structured responses.
Added Jul 25, 2026
Context Window
8.2K
Max Output
8.2K
Avg output tokens (7d)
61 tokens
Input Price (Auto)
$2.00/1M
Output Price (Auto)
$6.00/1M
Cache Read (Auto)
$1.00/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
7.6
Coding Index
14.4
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
63.1%
Better than 38% of models compared
HLE
Humanity's Last Exam
6.8%
Better than 44% of models compared
AA-LCR
Long context reasoning evaluation
38.0%
Better than 38% of models compared
Coding
SciCode
Python programming for scientific computing
21.6%
Better than 1% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
78.0%
Better than 59% of models compared
Last updated Sep 7, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Celeris 1 with similar models from the same provider or model family.
DeepSeek V4 Flash Vision Exp Uncensored
deepseek/deepseek-v4-flash-vision-exp-uncensoredAn uncensored variant of the experimental vision-enabled DeepSeek V4 Flash model for chat, image understanding, reasoning, coding, and tool use, with a 524K context window.
TheDrummer/Artemis v1.1
TheDrummer/Artemis-v1.1TheDrummer's Artemis v1.1 is a Gemma 4 31B fine-tune for creative writing, expressive dialogue, and roleplay, with optional thinking and a 262K context window.
Synth 2.5 Flash Preview
synth-2.5-flashSynth 2.5 Flash Preview is a low-cost text model designed for role-play, character dialogue, and collaborative storytelling.
Synth 2.5 Pro Preview
synth-2.5-proSynth 2.5 Pro Preview is a low-cost text model designed for role-play, character dialogue, and collaborative storytelling.
GPT 6 Astra Pro
openai/gpt-6-astra-proGPT 6 Astra in Pro reasoning mode. Uses additional model work for difficult tasks, with higher latency and token usage at the same per-token rates. Reasoning effort remains independently configurable.
GPT 6 Astra
openai/gpt-6-astraOpenAI's most capable model for the hardest end-to-end work, with state-of-the-art performance in computer use, browsing, software engineering, science, and professional tasks. Astra is designed to carry long, multistep workflows across code, browsers, and professional software through to a finished result.