Skip to content

Meta

llama-4-scout-17b-16e-instruct benchmarks, ranks and price

Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place llama-4-scout-17b-16e-instruct, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.

Updated Oct 1, 2026Full leaderboardSources and method

Read Oct 1, 2026

#70of 85

on the LLM leaderboard, with a consensus score of 30 across 2 leaderboards.

Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.

By leaderboard

Where each leaderboard places llama-4-scout-17b-16e-instruct

Artificial Analysis

Not measured by Artificial Analysis.

Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.

LiveBench

Not measured by LiveBench.

Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.

LMArena

#66of 79

Score 1323. Best run: llama-4-scout-17b-16e-instruct.

Blind head-to-head human votes, style controlled.

Berkeley BFCL

#29of 37

Score 28.13%. Best run: Llama-4-Scout-17B-16E-Instruct (FC).

Function calling and tool use, single turn to multi turn and web search.

Price and context

–

No gateway lists it.

Nearby

The models either side of llama-4-scout-17b-16e-instruct

  1. #67Mistral Small 432
  2. #68gpt-oss-120b31
  3. #69o3 Mini High30
  4. #70llama-4-scout-17b-16e-instruct30
  5. #71Mistral Large 3 251228
  6. #72phi-428
  7. #73Gemma 3 27B27

Questions about llama-4-scout-17b-16e-instruct

How good is llama-4-scout-17b-16e-instruct?
llama-4-scout-17b-16e-instruct ranks #70 of 85 models on the Rank.ai LLM leaderboard, with a consensus score of 30 out of 100 across 2 leaderboards.
Is llama-4-scout-17b-16e-instruct better than o3 Mini High?
Not on the consensus: o3 Mini High scores 30 to llama-4-scout-17b-16e-instruct's 30. Compare them source by source below, since the leaderboards test different things.
Is llama-4-scout-17b-16e-instruct better than Mistral Large 3 2512?
On the consensus, yes: 30 to 28.

Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.

  • Your grade out of 100How often AI names you, cites your site, and how high it ranks you.
  • Who gets namedEvery competitor in the answers, most named first.
  • The pages AI readsThe sources behind each answer.
  • Three fixesWhat to fix first, with a brief for the first article.