Grok 4.7 is xAI's most capable model for coding and knowledge work, pitched as twice as fast at half the price of comparable frontier models. It scores 71.0% on DeepSWE v1.1 and 46.3% on CursorBench 4.0. Pricing is $2/$6 per million tokens below 200K prompt tokens and doubles above that. It accepts text and images with a 500K-token context window; a faster variant exists only inside Cursor and Grok Build.
| Benchmark | Score | Type | Recorded |
|---|---|---|---|
| Humanity's Last Exam | 43.1 | accuracy | today |
| LCR | 76.7 | accuracy | today |
| TerminalBench | 25.8 | accuracy | today |
| SciCode | 57.4 | accuracy | today |
Updated today
Predecessors
Grok 4.6