Licensed capability-per-dollar rankings with exact model cohort, price basis, evidence, and freshness gates.
LiveBench model cost efficiency
45 models ranked by LiveBench cost per successful task.
DeepSeek V4 Flash is out in front at $0.016.
45 modelschecked Jun 25, 2026, 11:59 PM UTCdaily88 failed refresheshow we measure this →CSV / JSON
DeepSeek V4 Flash
$0.016LiveBench cost per successful task
- Dearest published row
- $1.573
- Middle of the board
- $0.24
GPT 5.6 Sol xHigh
Ranks exact model configurations by official cost per successful task for the complete frozen LiveBench workload.
Some explanation inputs are unavailable Missing: publicFrozenEvidence, comparablePriorRanking.
- Published rank
- #29
- Score
- 0.355
- Sample size
- 7
100% confidence · 64% component coverage · as of Jun 25, 2026, 11:59 PM UTC
Component ledger
| Component | Value | Weight | Contribution | Evidence | Status |
|---|---|---|---|---|---|
| livebench average input tokens | — | 0% | 0 | 1 | available |
| livebench average output tokens | — | 0% | 0 | 1 | available |
| livebench cost per question USD | — | 0% | 0 | 1 | available |
| livebench cost per successful task USD | — | 100% | 0.355 | 1 | available |
| livebench input price per million USD | — | 0% | 0 | 1 | available |
| livebench output price per million USD | — | 0% | 0 | 1 | available |
| livebench overall score | — | 0% | 0 | 1 | available |
| metrics | Structured | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
| rankingDirection | ascending | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
| release | 2026-06-25 | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
| sourceModelId | gpt-5.6-sol-xhigh | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
new since prior snapshot
no_prior_published_snapshot
No component moved since the prior snapshot.
Source trail
LiveBench average input tokenslivebench average input tokens · observed Jun 25, 2026, 11:59 PM UTC86,795 tokens100% confidence
- Source
- LiveBench official model benchmark
- Grade
- A
- Freshness
- fresh
- Locator
{"formula":"(sum task cost / sum question count) / (overall score / 100)","costAsset":"https://livebench.ai/cost_2026_06_25.csv","scoreAsset":"https://livebench.ai/table_2026_06_25.csv","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}
LiveBench average output tokenslivebench average output tokens · observed Jun 25, 2026, 11:59 PM UTC7,161 tokens100% confidence
- Source
- LiveBench official model benchmark
- Grade
- A
- Freshness
- fresh
- Locator
{"formula":"(sum task cost / sum question count) / (overall score / 100)","costAsset":"https://livebench.ai/cost_2026_06_25.csv","scoreAsset":"https://livebench.ai/table_2026_06_25.csv","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}
LiveBench cost per questionlivebench cost per question USD · observed Jun 25, 2026, 11:59 PM UTC$0.28100% confidence
- Source
- LiveBench official model benchmark
- Grade
- A
- Freshness
- fresh
- Locator
{"formula":"(sum task cost / sum question count) / (overall score / 100)","costAsset":"https://livebench.ai/cost_2026_06_25.csv","scoreAsset":"https://livebench.ai/table_2026_06_25.csv","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}
LiveBench cost per successful tasklivebench cost per successful task USD · observed Jun 25, 2026, 11:59 PM UTC$0.36100% confidence
- Source
- LiveBench official model benchmark
- Grade
- A
- Freshness
- fresh
- Locator
{"formula":"(sum task cost / sum question count) / (overall score / 100)","costAsset":"https://livebench.ai/cost_2026_06_25.csv","scoreAsset":"https://livebench.ai/table_2026_06_25.csv","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}
LiveBench input pricelivebench input price per million USD · observed Jun 25, 2026, 11:59 PM UTC$5.00100% confidence
- Source
- LiveBench official model benchmark
- Grade
- A
- Freshness
- fresh
- Locator
{"formula":"(sum task cost / sum question count) / (overall score / 100)","costAsset":"https://livebench.ai/cost_2026_06_25.csv","scoreAsset":"https://livebench.ai/table_2026_06_25.csv","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}
LiveBench output pricelivebench output price per million USD · observed Jun 25, 2026, 11:59 PM UTC$30.00100% confidence
- Source
- LiveBench official model benchmark
- Grade
- A
- Freshness
- fresh
- Locator
{"formula":"(sum task cost / sum question count) / (overall score / 100)","costAsset":"https://livebench.ai/cost_2026_06_25.csv","scoreAsset":"https://livebench.ai/table_2026_06_25.csv","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}
LiveBench overall scorelivebench overall score · observed Jun 25, 2026, 11:59 PM UTC79.663 percent100% confidence
- Source
- LiveBench official model benchmark
- Grade
- A
- Freshness
- fresh
- Locator
{"formula":"(sum task cost / sum question count) / (overall score / 100)","costAsset":"https://livebench.ai/cost_2026_06_25.csv","scoreAsset":"https://livebench.ai/table_2026_06_25.csv","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}
More in Model efficiency
Capability relative to token price, context, latency, and throughput.
Intelligence points per weighted inference dollar at public API prices.
Usable context capacity relative to published input-token pricing.