Artificial Analysis
Not measured by Artificial Analysis.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
Alibaba
Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place qwen3-30b-a3b-instruct-2507, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.
Read Oct 1, 2026
#53of 85
on the LLM leaderboard, with a consensus score of 43 across 2 leaderboards.
Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.
By leaderboard
Artificial Analysis
Not measured by Artificial Analysis.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
LiveBench
Not measured by LiveBench.
Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.
LMArena
#55of 79
Score 1383. Best run: qwen3-30b-a3b-instruct-2507.
Blind head-to-head human votes, style controlled.
Berkeley BFCL
#20of 37
Score 41.39%. Best run: Qwen3-30B-A3B-Instruct-2507 (FC).
Function calling and tool use, single turn to multi turn and web search.
Price and context
–
No gateway lists it.
Nearby
Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.