Grok 4.3
X.AI·
LLMsflagshipcloud
Overview
Reasoning-first flagship with a 1M-token context window and native video input. Agentic-focused release at lower pricing than Grok 4.1.
Capabilities and innovations
1M context windowNative video inputReasoning-first
Benchmarks
| Benchmark | Score | Source |
|---|---|---|
| GPQA Diamond | 90.1% | independent |
| MMLU-Pro | 85.84% | unsourced |
- Humanity’s Last Exam (no tools): independent evaluation pending
- SWE-bench Pro: independent evaluation pending
- Terminal-Bench 2.x: independent evaluation pending
- MMMU-Pro: independent evaluation pending
API pricing
$1.25 per million input tokens, $2.5 per million output tokens (USD, cheapest OpenRouter endpoint, checked 2026-06-08)
Links
More from X.AI
Data curated by AI Model Timeline. See the methodology for admission criteria, benchmark eras and source priority.