Artificial Analysis
#26of 115
Score 28.5. Best run: openai/gpt-5.2-codex.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
OpenAI
Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place GPT-5.2-Codex, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.
Read Oct 5, 2026
#37of 126
on the LLM leaderboard, with a consensus score of 61 across 2 leaderboards.
Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.
By leaderboard
Artificial Analysis
#26of 115
Score 28.5. Best run: openai/gpt-5.2-codex.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
LiveBench
#17of 36
Score 74.0. Best run: GPT 5.2 Codex.
Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.
LMArena
Not measured by LMArena.
Blind head-to-head human votes, style controlled.
Berkeley BFCL
Not measured by Berkeley BFCL.
Function calling and tool use, single turn to multi turn and web search.
Price and context
$4.81
OpenRouter: $1.75 in, $14 out. Vercel AI Gateway: $1.75 in, $14 out Context window 400K tokens. LiveBench cost per solved task $0.187.
Nearby
Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.