Skip to content

OpenAI

gpt-4.1-2025-04-14 benchmarks, ranks and price

Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place gpt-4.1-2025-04-14, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.

Updated Oct 1, 2026Full leaderboardSources and method

Read Oct 1, 2026

#37of 85

on the LLM leaderboard, with a consensus score of 52 across 2 leaderboards.

Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.

By leaderboard

Where each leaderboard places gpt-4.1-2025-04-14

Artificial Analysis

Not measured by Artificial Analysis.

Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.

LiveBench

Not measured by LiveBench.

Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.

LMArena

#46of 79

Score 1414. Best run: gpt-4.1-2025-04-14.

Blind head-to-head human votes, style controlled.

Berkeley BFCL

#14of 37

Score 53.96%. Best run: GPT-4.1-2025-04-14 (FC).

Function calling and tool use, single turn to multi turn and web search.

Price and context

$3.50

Vercel AI Gateway: $2.00 in, $8.00 out

Nearby

The models either side of gpt-4.1-2025-04-14

  1. #34Claude Haiku 4.553
  2. #35Kimi K2.653
  3. #36qwen3-235b-a22b-instruct-250752
  4. #37gpt-4.1-2025-04-1452
  5. #38Inkling xHigh51
  6. #39Command A50
  7. #40Qwen3.5 397B A17B50

Questions about gpt-4.1-2025-04-14

How good is gpt-4.1-2025-04-14?
gpt-4.1-2025-04-14 ranks #37 of 85 models on the Rank.ai LLM leaderboard, with a consensus score of 52 out of 100 across 2 leaderboards.
Is gpt-4.1-2025-04-14 better than qwen3-235b-a22b-instruct-2507?
Not on the consensus: qwen3-235b-a22b-instruct-2507 scores 52 to gpt-4.1-2025-04-14's 52. Compare them source by source below, since the leaderboards test different things.
Is gpt-4.1-2025-04-14 better than Inkling xHigh?
On the consensus, yes: 52 to 51.
How much does gpt-4.1-2025-04-14 cost?
$2.00 per million input tokens and $8.00 per million output tokens on Vercel AI Gateway, the cheaper listing we read. Blended at three input tokens to one output, that is $3.50 per million.

Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.

  • Your grade out of 100How often AI names you, cites your site, and how high it ranks you.
  • Who gets namedEvery competitor in the answers, most named first.
  • The pages AI readsThe sources behind each answer.
  • Three fixesWhat to fix first, with a brief for the first article.