Skip to content

ibm

granite-3.1-8b-instruct benchmarks, ranks and price

Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place granite-3.1-8b-instruct, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.

Updated Oct 1, 2026Full leaderboardSources and method

Read Oct 1, 2026

#78of 85

on the LLM leaderboard, with a consensus score of 24 across 2 leaderboards.

Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.

By leaderboard

Where each leaderboard places granite-3.1-8b-instruct

Artificial Analysis

Not measured by Artificial Analysis.

Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.

LiveBench

Not measured by LiveBench.

Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.

LMArena

#77of 79

Score 1208. Best run: granite-3.1-8b-instruct.

Blind head-to-head human votes, style controlled.

Berkeley BFCL

#30of 37

Score 27.10%. Best run: Granite-3.1-8B-Instruct (FC).

Function calling and tool use, single turn to multi turn and web search.

Price and context

–

No gateway lists it.

Nearby

The models either side of granite-3.1-8b-instruct

  1. #75Gemma 3 12B25
  2. #76llama-3.1-nemotron-ultra-253b-v124
  3. #77amazon-nova-pro-v1.024
  4. #78granite-3.1-8b-instruct24
  5. #79llama-3.1-8b-instruct24
  6. #80gpt-oss-20b23
  7. #81amazon-nova-micro-v1.023

Questions about granite-3.1-8b-instruct

How good is granite-3.1-8b-instruct?
granite-3.1-8b-instruct ranks #78 of 85 models on the Rank.ai LLM leaderboard, with a consensus score of 24 out of 100 across 2 leaderboards.
Is granite-3.1-8b-instruct better than amazon-nova-pro-v1.0?
Not on the consensus: amazon-nova-pro-v1.0 scores 24 to granite-3.1-8b-instruct's 24. Compare them source by source below, since the leaderboards test different things.
Is granite-3.1-8b-instruct better than llama-3.1-8b-instruct?
On the consensus, yes: 24 to 24.

Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.

  • Your grade out of 100How often AI names you, cites your site, and how high it ranks you.
  • Who gets namedEvery competitor in the answers, most named first.
  • The pages AI readsThe sources behind each answer.
  • Three fixesWhat to fix first, with a brief for the first article.