Artificial Analysis
#51of 51
Score 3.8. Best run: google/gemma-3-12b-it.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place Gemma 3 12B, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.
Read Oct 1, 2026
#75of 85
on the LLM leaderboard, with a consensus score of 25 across 3 leaderboards.
Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.
By leaderboard
Artificial Analysis
#51of 51
Score 3.8. Best run: google/gemma-3-12b-it.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
LiveBench
Not measured by LiveBench.
Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.
LMArena
#63of 79
Score 1342. Best run: gemma-3-12b-it.
Blind head-to-head human votes, style controlled.
Berkeley BFCL
#26of 37
Score 30.43%. Best run: Gemma-3-12b-it (Prompt).
Function calling and tool use, single turn to multi turn and web search.
Price and context
$0.075
OpenRouter: $0.050 in, $0.15 out Context window 131K tokens.
Nearby
Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.