AI Model Timeline

Tracking the accelerating release frequency of frontier AI models.

Grok 4.3

X.AI·

LLMsflagshipcloud

Overview

Reasoning-first flagship with a 1M-token context window and native video input. Agentic-focused release at lower pricing than Grok 4.1.

Capabilities and innovations

1M context windowNative video inputReasoning-first

Benchmarks

BenchmarkScoreSource
GPQA Diamond90.1%independent
MMLU-Pro85.84%unsourced
  • Humanity’s Last Exam (no tools): independent evaluation pending
  • SWE-bench Pro: independent evaluation pending
  • Terminal-Bench 2.x: independent evaluation pending
  • MMMU-Pro: independent evaluation pending

API pricing

$1.25 per million input tokens, $2.5 per million output tokens (USD, cheapest OpenRouter endpoint, checked 2026-06-08)

Links

More from X.AI

Data curated by AI Model Timeline. See the methodology for admission criteria, benchmark eras and source priority.