Skip to content

Qwen

Qwen3.7 Plus benchmarks, ranks and price

Where Artificial Analysis, LiveBench, LMArena and Berkeley BFCL place Qwen3.7 Plus, what it costs through OpenRouter and Vercel AI Gateway, and the models either side of it on the consensus.

Updated Oct 1, 2026Full leaderboardSources and method

Read Oct 1, 2026

#28of 85

on the LLM leaderboard, with a consensus score of 59 across 2 leaderboards.

Source: Artificial Analysis, LiveBench, LMArena and Berkeley BFCL, combined by Rank.ai.

By leaderboard

Where each leaderboard places Qwen3.7 Plus

Artificial Analysis

#24of 51

Score 25.2. Best run: qwen/qwen3.7-plus.

Intelligence Index: a composite of reasoning, knowledge, maths and coding evaluations.

LiveBench

Not measured by LiveBench.

Contamination-limited questions refreshed each release: reasoning, maths, coding, data analysis, language.

LMArena

#22of 79

Score 1460. Best run: qwen3.7-plus.

Blind head-to-head human votes, style controlled.

Berkeley BFCL

Not measured by Berkeley BFCL.

Function calling and tool use, single turn to multi turn and web search.

Price and context

$0.56

OpenRouter: $0.32 in, $1.28 out. Vercel AI Gateway: $0.40 in, $1.60 out Context window 1M tokens.

Nearby

The models either side of Qwen3.7 Plus

  1. #25Qwen3.7 Max61
  2. #26Claude Sonnet 4.660
  3. #27GPT 5.2 (2025-12-11) High59
  4. #28Qwen3.7 Plus59
  5. #29DeepSeek V4 Pro 042358
  6. #30deepseek-v3.258
  7. #31grok-4-070956

Questions about Qwen3.7 Plus

How good is Qwen3.7 Plus?
Qwen3.7 Plus ranks #28 of 85 models on the Rank.ai LLM leaderboard, with a consensus score of 59 out of 100 across 2 leaderboards.
Is Qwen3.7 Plus better than GPT 5.2 (2025-12-11) High?
Not on the consensus: GPT 5.2 (2025-12-11) High scores 59 to Qwen3.7 Plus's 59. Compare them source by source below, since the leaderboards test different things.
Is Qwen3.7 Plus better than DeepSeek V4 Pro 0423?
On the consensus, yes: 59 to 58.
How much does Qwen3.7 Plus cost?
$0.32 per million input tokens and $1.28 per million output tokens on OpenRouter, the cheaper listing we read. Blended at three input tokens to one output, that is $0.56 per million.
Is Qwen3.7 Plus good for coding?
It scores 55.9 on the Artificial Analysis Coding Index, placing it #24 for coding among the models on this leaderboard.

Enter your website. In about two minutes, Rank.ai asks ChatGPT, Claude and Gemini 12 questions your customers ask and grades how often they name you.

  • Your grade out of 100How often AI names you, cites your site, and how high it ranks you.
  • Who gets namedEvery competitor in the answers, most named first.
  • The pages AI readsThe sources behind each answer.
  • Three fixesWhat to fix first, with a brief for the first article.