AI model maker
xAI
33 models on Artificials, 16 with published test results.
Where xAI stands
Its best model’s rank among all models tested on each skill area’s main test, the same tests model pages use. A longer bar means a better rank.
- OverallGrok 4.6 on Epoch Capabilities IndexTop 8%
- CodingGrok 4.7 (xhigh) on LiveBench coding#35 of 63
- MathGrok 4.6 (xhigh) on Mock AIME 2024–2025Top 8%
- KnowledgeGrok 4.7 (xhigh) on SimpleQA VerifiedTop 24%
- AgentsGrok 4.6 on LiveBench agentic codingTop 34%
- LanguageGrok 4.6 on LiveBench languageTop 27%
- Data analysisGrok 4.7 (xhigh) on LiveBench data analysisTop 45%
- ReasoningGrok 4.6 (high) on GPQA DiamondTop 5%
On the capability index
Its 10 models on the Epoch Capabilities Index, each with its rank among all 274 models. The top score is 167.3, held by Claude Opus 5.5 from Anthropic.
Every test
Its best result on each published test, and that result’s rank among every model tested. Scores from different tests are never combined.
Every model
Newest first. The price is per 1M tokens, blended three input tokens to one output token: xAI’s own listing, or else the middle price across hosts.
| Model | Released | Context | Price | Capability index |
|---|---|---|---|---|
| Grok Imagine Video 1.5 Lite | Oct 1, 2026 | 1K tokens | No paid listing | Not on the index |
| Grok 4.7 | Sep 21, 2026 | 500K tokens | $3.00 | 153.5 |
| Grok 4.6 | Aug 12, 2026 | 500K tokens | $3.00 | 156.4 |
| Grok Imagine Image 2.0 | Aug 7, 2026 | 66K tokens | No paid listing | Not on the index |
| Grok 4.5 | Jul 8, 2026 | 500K tokens | $3.00 | 153.9 |
| Grok Imagine Video 1.5 | May 30, 2026 | 1K tokens | No paid listing | Not on the index |
| Grok Latest | May 3, 2026 | 500K tokens | $3.00 | Not on the index |
| Grok 4.3 | Apr 17, 2026 | 1M tokens | $1.56 | 149.2 |
| Grok Build 0.1 | Apr 16, 2026 | 256K tokens | $1.25 | Not on the index |
| Grok Imagine Image Quality | Apr 3, 2026 | 16K tokens | No paid listing | Not on the index |
| Grok 4.20 Multi-Agent | Mar 10, 2026 | 2M tokens | $1.56 | Not on the index |
| Grok 4.20 | Mar 9, 2026 | 2M tokens | $1.77 | 152.0 |
| Grok 4.20 (0309 non reasoning) | Mar 9, 2026 | 1M tokens | $1.56 | Not on the index |
| Grok 4.20 (0309 reasoning) | Mar 9, 2026 | 1M tokens | $1.56 | Not on the index |
| Grok 4.20 (beta 0309 non reasoning) | Mar 9, 2026 | 2M tokens | $3.00 | Not on the index |
Models, prices, context windows and release dates from models.dev (MIT). Test results from LiveBench (Apache-2.0), Epoch AI (CC BY 4.0), Datacurve DeepSWE, compiled by Epoch AI (CC BY 4.0), Cursor CursorBench, compiled by Epoch AI (CC BY 4.0) and Arena (CC BY 4.0). A model counts here when xAI sells it, several hosts list it or a test covers it. Each figure is the source’s own; nothing is combined into a new score.