Grok 4.6 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Added Aug 12, 2026
Context Window
500.0K
Max Output
500.0K
Avg output tokens (7d)
1.8K tokens
Input Price (Auto)
$2.00/1M
Output Price (Auto)
$6.00/1M
Cache Read (Auto)
$0.50/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
44.3
Coding Index
76.8
Agentic Index
53.0
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
66.7%
Better than 96% of models compared
AutomationBench-AA Tasks Completed
Fully completed workflows without guardrail violations
32.7%
Better than 60% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
1540 Elo
Better than 88% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1609 Elo
Better than 92% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
17.0%
Better than 61% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
80.3%
Better than 88% of models compared
MLCR-AA
Medical long-context reasoning
12.2%
Better than 38% of models compared
Reasoning
HLE
Humanity's Last Exam
42.9%
Better than 91% of models compared
CritPt
Research-level physics reasoning
17.1%
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
21.2%
Better than 75% of models compared
SciCode
Python programming for scientific computing
56.5%
Better than 85% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
48.2%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
34.3%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
94.9%
Better than 99% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
80.3%
Better than 88% of models compared
GDPval-AA (unversioned / legacy)
Economically valuable tasks
55.5%
Last updated Oct 2, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Grok 4.6 with similar models from the same provider or model family.
Grok 4.7
x-ai/grok-4.7Grok 4.7 is SpaceXAI's flagship reasoning model for coding, long-running agentic tasks, and knowledge work. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active, and requests above 200k input tokens use higher long-context rates. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.5
x-ai/grok-4.5Grok 4.5 is SpaceXAI's frontier reasoning model for coding, knowledge work, and STEM. It supports text, image, and file inputs with tool calling and structured outputs. Reasoning is always active. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Build 0.1
x-ai/grok-build-0.1Grok Build 0.1 is SpaceXAI's fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding agents, tool use, and multi-step development tasks. Currently in early access. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok Latest
x-ai/grok-latestCompatibility alias that routes to the newest Grok model. Currently routes to Grok 4.6. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.3
x-ai/grok-4.3Grok 4.3 is SpaceXAI's reasoning model for text and image inputs, built for agentic workflows, instruction following, factual accuracy, long-document analysis, and deep research. Reasoning is always active and requests above 200k total tokens are charged at the higher long-context rate. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.
Grok 4.20
x-ai/grok-4.20SpaceXAI's Grok 4.20 flagship release with tool calling, multimodal input support, and a 2M-token context window. Content policy rejections can still be charged: SpaceXAI may pass through a $0.05 moderation-failure fee, or a $0.055 usage-guidelines violation fee, depending on which rejection upstream returns.