Artificial Analysis
#29of 115
Score 27.2. Best run: x-ai/grok-build-0.1.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
xAI
Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place Grok Build 0.1, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.
Read Oct 5, 2026
#68of 126
on the LLM leaderboard, with a consensus score of 48 across 2 leaderboards.
Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.
By leaderboard
Artificial Analysis
#29of 115
Score 27.2. Best run: x-ai/grok-build-0.1.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
LiveBench
#30of 36
Score 67.8. Best run: Grok Build 0.1.
Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.
LMArena
Not measured by LMArena.
Blind head-to-head human votes, style controlled.
Berkeley BFCL
Not measured by Berkeley BFCL.
Function calling and tool use, single turn to multi turn and web search.
Price and context
$1.25
OpenRouter: $1.00 in, $2.00 out Context window 256K tokens. LiveBench cost per solved task $0.024.
Nearby
Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.