Skip to content

Mistral

Mistral Small 3 benchmarks, ranks and price

Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place Mistral Small 3, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.

Updated Oct 5, 2026Full leaderboardSources and method

Read Oct 5, 2026

#118of 126

on the LLM leaderboard, with a consensus score of 22 across 2 leaderboards.

Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.

By leaderboard

Where each leaderboard places Mistral Small 3

Artificial Analysis

#104of 115

Score 6.7. Best run: mistralai/mistral-small-24b-instruct-2501.

Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.

LiveBench

Not measured by LiveBench.

Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.

LMArena

#107of 115

Score 1274. Best run: mistral-small-24b-instruct-2501.

Blind head-to-head human votes, style controlled.

Berkeley BFCL

Not measured by Berkeley BFCL.

Function calling and tool use, single turn to multi turn and web search.

Price and context

$0.057

OpenRouter: $0.050 in, $0.080 out Context window 33K tokens.

Nearby

The models either side of Mistral Small 3

  1. #115granite-3.1-8b-instruct23
  2. #116GPT-423
  3. #117Llama 3.1 70B Instruct22
  4. #118Mistral Small 322
  5. #119amazon-nova-micro-v1.022
  6. #120Phi 421
  7. #121Llama 3.1 8B Instruct20

Questions about Mistral Small 3

How good is Mistral Small 3?
Mistral Small 3 ranks #118 of 126 models on the Rank.ai LLM leaderboard, with a consensus score of 22 out of 100 across 2 leaderboards.
Is Mistral Small 3 better than Llama 3.1 70B Instruct?
Not on the consensus: Llama 3.1 70B Instruct scores 22 to Mistral Small 3's 22. Compare them source by source below, since the leaderboards test different things.
Is Mistral Small 3 better than amazon-nova-micro-v1.0?
On the consensus, yes: 22 to 22.
How much does Mistral Small 3 cost?
$0.050 per million input tokens and $0.080 per million output tokens on OpenRouter, the cheaper listing we read. Blended at three input tokens to one output, that is $0.057 per million.

Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.

  • Your grade out of 100How often AI names you, cites your site, and how high it ranks you.
  • Who gets namedEvery competitor in the answers, most named first.
  • The pages AI readsThe sources behind each answer.
  • Three fixesWhat to fix first, with a brief for the first article.