Skip to content

Deepseek

DeepSeek V3.1 Terminus benchmarks, ranks and price

Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place DeepSeek V3.1 Terminus, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.

Updated Oct 1, 2026Full leaderboardSources and method

Read Oct 1, 2026

#60of 85

on the LLM leaderboard, with a consensus score of 40 across 2 leaderboards.

Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.

By leaderboard

Where each leaderboard places DeepSeek V3.1 Terminus

Artificial Analysis

#38of 51

Score 14.8. Best run: deepseek/deepseek-v3.1-terminus.

Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.

LiveBench

Not measured by LiveBench.

Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.

LMArena

#44of 79

Score 1417. Best run: deepseek-v3.1-terminus-thinking.

Blind head-to-head human votes, style controlled.

Berkeley BFCL

Not measured by Berkeley BFCL.

Function calling and tool use, single turn to multi turn and web search.

Price and context

$0.45

OpenRouter: $0.30 in, $1.00 out. Vercel AI Gateway: $0.27 in, $1.00 out Context window 164K tokens.

Nearby

The models either side of DeepSeek V3.1 Terminus

  1. #57Kimi K2.7 Code41
  2. #58qwen3-32b41
  3. #59Grok 4.341
  4. #60DeepSeek V3.1 Terminus40
  5. #61R138
  6. #62GPT-5.4 Nano38
  7. #63Qwen3 235B A22B Thinking 250735

Questions about DeepSeek V3.1 Terminus

How good is DeepSeek V3.1 Terminus?
DeepSeek V3.1 Terminus ranks #60 of 85 models on the Rank.ai LLM leaderboard, with a consensus score of 40 out of 100 across 2 leaderboards.
Is DeepSeek V3.1 Terminus better than Grok 4.3?
Not on the consensus: Grok 4.3 scores 41 to DeepSeek V3.1 Terminus's 40. Compare them source by source below, since the leaderboards test different things.
Is DeepSeek V3.1 Terminus better than R1?
On the consensus, yes: 40 to 38.
How much does DeepSeek V3.1 Terminus cost?
$0.27 per million input tokens and $1.00 per million output tokens on Vercel AI Gateway, the cheaper listing we read. Blended at three input tokens to one output, that is $0.45 per million.
Is DeepSeek V3.1 Terminus good for coding?
It scores 43.5 on the Artificial Analysis Coding Index, placing it #35 for coding among the models on this leaderboard.

Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.

  • Your grade out of 100How often AI names you, cites your site, and how high it ranks you.
  • Who gets namedEvery competitor in the answers, most named first.
  • The pages AI readsThe sources behind each answer.
  • Three fixesWhat to fix first, with a brief for the first article.