Celeris 1 is a diffusion language model built for ultra-low-latency classification, extraction, judging, query rewriting, and other short structured responses.
Added Jul 25, 2026
Context Window
8.2K
Max Output
8.2K
Input Price (Auto)
$2.00/1M
Output Price (Auto)
$6.00/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
6.3
Coding Index
14.4
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
0.4%
Better than 7% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
153 Elo
Better than 12% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
288 Elo
Better than 17% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
1.2%
Better than 9% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
38.0%
Better than 35% of models compared
Reasoning
HLE
Humanity's Last Exam
6.8%
Better than 41% of models compared
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
0.0%
Better than 16% of models compared
SciCode
Python programming for scientific computing
21.6%
Better than 3% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
78.0%
Better than 59% of models compared
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
63.1%
Better than 38% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
38.0%
Better than 35% of models compared
Last updated Oct 1, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Celeris 1 with similar models from the same provider or model family.
Claude Opus 5.5
anthropic/claude-opus-5.5Claude Opus 5.5 is Anthropic's flagship model for agentic coding, computer use, complex knowledge work, and long-running tasks, with improved efficiency and more natural communication.
Nex N2.5 Mini
nex-agi/nex-n2.5-miniNex N2.5 Mini is an open-source multimodal agentic model for coding and long-horizon workflows. It supports image input, structured output, configurable reasoning, a 262K-token context window, and up to 235K output tokens.
Nex N2.5 Pro
nex-agi/nex-n2.5-proNex N2.5 Pro is an open-source multimodal agentic model for coding, tool use, and long-horizon workflows. It supports image input, structured output, configurable reasoning, a 262K-token context window, and up to 235K output tokens.
GPT 6 Sol
openai/gpt-6-solGPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional work, reasoning, coding, and agentic workflows.
GPT 6 Sol Pro
openai/gpt-6-sol-proGPT-6 Sol Pro uses the same underlying model as GPT-6 Sol with Pro reasoning mode enabled for higher-quality responses on complex tasks.
Qwen 3.8 27B Hemingway
qwen/qwen3.8-27b-hemmingwayQwen 3.8 27B Hemingway is an open-weight NVFP4 multimodal creative finetune for long-form prose, character dialogue, storytelling, and roleplay.