Artificial Analysis
#12of 51
Score 34.3. Best run: deepseek/deepseek-v4-flash-0731.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
Deepseek
Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place DeepSeek V4 Flash 0423, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.
Read Oct 1, 2026
#42of 85
on the LLM leaderboard, with a consensus score of 49 across 3 leaderboards.
Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.
By leaderboard
Artificial Analysis
#12of 51
Score 34.3. Best run: deepseek/deepseek-v4-flash-0731.
Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.
LiveBench
#31of 34
Score 65.5. Best run: DeepSeek V4 Flash.
Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.
LMArena
#33of 79
Score 1439. Best run: deepseek-v4-flash-thinking.
Blind head-to-head human votes, style controlled.
Berkeley BFCL
Not measured by Berkeley BFCL.
Function calling and tool use, single turn to multi turn and web search.
Price and context
$0.095
OpenRouter: $0.079 in, $0.16 out. Vercel AI Gateway: $0.076 in, $0.15 out Context window 1.0M tokens. LiveBench cost per solved task $0.016.
Nearby
Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.