Artificial Analysis benchmarks Grok 4.7, finds high intelligence but verbose output
Grok 4.7 scores 46 on Artificial Analysis Intelligence Index, well above average, but generates 2.5x more tokens than comparable models.
What to know
- Grok 4.7 achieves an intelligence score of 46 on Artificial Analysis's benchmark, nearly double the median of 24, placing it among leading models.
- The model generates 2.5x more tokens than comparable models during evaluation, indicating significant inefficiency and verbosity in output generation.
- Grok 4.7 pricing is competitive with peers at $2.00 per 1M input tokens and $6.00 per 1M output tokens, both at or below median rates.
“Pricing for Grok 4.7 (xhigh) is $2.00 per 1M input tokens (moderately priced, median: $2.00) and $6.00 per 1M output tokens (moderately priced, median: $10.00).”
Artificial Analysis, Benchmarking firm · Artificial Analysis ↗
Artificial Analysis AI benchmarking firmGrok (xAI) AI model subject
How it unfolded 2 developments, newest first · click a bar or a number to jump articlesposts
-
2
Hacker News shares Grok 4.7 analysis with community
The Artificial Analysis benchmark was shared on Hacker News, surfacing the detailed evaluation to the developer community. The submission was titled 'Grok 4.7 Intelligence, Performance and Price Analysis' and gained 11 points.
-
1
Artificial Analysis publishes Grok 4.7 benchmark evaluation
Artificial Analysis released a comprehensive intelligence, performance, and price analysis of Grok 4.7 (xhigh). The model scored 46 on their Intelligence Index, placing it well above average. The analysis found the model notably slow and very verbose, generating 240M tokens during evaluation versus a median of 94M.
“Grok 4.7 (xhigh) is amongst the leading models in intelligence and reasonably priced when comparing to other models of similar price. It's also notably slow and very verbose.”
— Artificial Analysis -
first by Artificial Analysis, 4d ago
-