Loading company intelligence…
Loading the latest published snapshot and its supporting evidence.
Loading the latest published snapshot and its supporting evidence.
Exact claude-opus-5-xhigh-effort inference configuration in the official LiveBench 2026-06-25 release.
A current roll-up of published ranks, normalized facts, primary evidence, and entity relationships. No missing value is inferred.
| Metric | Value | As of | Evidence | Confidence |
|---|---|---|---|---|
| model_capabilityLiveBench Agentic Coding score | 61.31 percentderived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_capabilityLiveBench Coding score | 83.18 percentderived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_capabilityLiveBench Data Analysis score | 77.88 percentderived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_capabilityLiveBench IF score | 67.52 percentderived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_capabilityLiveBench Language score | 87.32 percentderived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_capabilityLiveBench Mathematics score | 94.8 percentderived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_capabilityLiveBench overall score | 80.31 percentderived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_capabilityLiveBench Reasoning score | 90.17 percentderived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_efficiencyLiveBench average input tokens | 1,248 tokensmeasured | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_efficiencyLiveBench average output tokens | 11,110 tokensmeasured | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_efficiencyLiveBench cost per question | $0.39derived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_efficiencyLiveBench cost per successful task | $0.49derived | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_efficiencyLiveBench input price | $5.00reported | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
| model_efficiencyLiveBench output price | $25.00reported | Jun 25, 2026 | LiveBench official model benchmark ↗Grade A | 100% |
Every currently published position, joined to its snapshot, confidence, and methodology version.
| Rank | Leaderboard | Score | Movement | Confidence | As of | Method | Open explanation |
|---|---|---|---|---|---|---|---|
| #3 | LiveBench Model Capabilitylivebench-model-capability | 80.31 | — | 100% | Jun 25, 2026 | vlivebench-2026-06-25-official-v1-livebench-model-capability-v1 | Explain → |
| #35 | LiveBench Model Cost Efficiencylivebench-model-efficiency | 0.49 | — | 100% | Jun 25, 2026 | vlivebench-2026-06-25-official-v1-livebench-model-efficiency-v1 | Explain → |
Ranks exact model configurations by the mean of seven official LiveBench category averages.
Some explanation inputs are unavailable Missing: publicFrozenEvidence, comparablePriorRanking.
| Component | Value | Weight | Contribution | Evidence | Status |
|---|---|---|---|---|---|
| livebench agentic coding score | — | 0% | 0 | 1 | available |
| livebench coding score | — | 0% | 0 | 1 | available |
| livebench data analysis score | — | 0% | 0 | 1 | available |
| livebench instruction following score |
15 published observations for Claude Opus 5 xHigh Effort. Open a row for the source, evidence location, extraction method, and quality state.
| Metric | Value | Kind | As of | Confidence | Evidence |
|---|---|---|---|---|---|
| LiveBench overall scorelivebench-overall-score | 80.3129percent | derived | Jun 25, 2026 | 100% | View source
|
| LiveBench cost per successful tasklivebench-cost-per-successful-task-usd | 0.487USD · USD_per_successful_task | derived | Jun 25, 2026 | 100% | View source
|
| — |
| 0% |
| 0 |
| 1 |
| available |
| livebench language score | — | 0% | 0 | 1 | available |
|---|
| livebench mathematics score | — | 0% | 0 | 1 | available |
|---|
| livebench overall score | — | 100% | 80.313 | 1 | available |
|---|
| livebench reasoning score | — | 0% | 0 | 1 | available |
|---|
| metrics | Structured | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| rankingDirection | descending | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| release | 2026-06-25 | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| sourceModelId | claude-opus-5-xhigh-effort | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
no_prior_published_snapshot
No comparable component drivers were published.
{"scoreAsset":"https://livebench.ai/table_2026_06_25.csv","aggregation":"mean_of_category_means","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}{"scoreAsset":"https://livebench.ai/table_2026_06_25.csv","aggregation":"mean_of_category_means","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}{"scoreAsset":"https://livebench.ai/table_2026_06_25.csv","aggregation":"mean_of_category_means","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}{"scoreAsset":"https://livebench.ai/table_2026_06_25.csv","aggregation":"mean_of_category_means","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}{"scoreAsset":"https://livebench.ai/table_2026_06_25.csv","aggregation":"mean_of_category_means","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}{"scoreAsset":"https://livebench.ai/table_2026_06_25.csv","aggregation":"mean_of_category_means","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}{"scoreAsset":"https://livebench.ai/table_2026_06_25.csv","aggregation":"mean_of_category_means","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}{"scoreAsset":"https://livebench.ai/table_2026_06_25.csv","aggregation":"mean_of_category_means","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}{"formula":"(sum task cost / sum question count) / (overall score / 100)","costAsset":"https://livebench.ai/cost_2026_06_25.csv","scoreAsset":"https://livebench.ai/table_2026_06_25.csv","categoryAsset":"https://livebench.ai/categories_2026_06_25.json"}| LiveBench Language scorelivebench-language-score | 87.3243percent | derived | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench cost per questionlivebench-cost-per-question-usd | 0.3911USD · USD_per_question | derived | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench IF scorelivebench-instruction-following-score | 67.5248percent | derived | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench average input tokenslivebench-average-input-tokens | 1,248tokens | measured | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench overall scorelivebench-overall-score | 80.3129percent | derived | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench Agentic Coding scorelivebench-agentic-coding-score | 61.3133percent | derived | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench Reasoning scorelivebench-reasoning-score | 90.173percent | derived | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench Mathematics scorelivebench-mathematics-score | 94.8percent | derived | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench output pricelivebench-output-price-per-million-usd | 25USD · USD_per_million_tokens | reported | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench Data Analysis scorelivebench-data-analysis-score | 77.8797percent | derived | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench input pricelivebench-input-price-per-million-usd | 5USD · USD_per_million_tokens | reported | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench average output tokenslivebench-average-output-tokens | 11,110tokens | measured | Jun 25, 2026 | 100% | View source
|
|---|
| LiveBench Coding scorelivebench-coding-score | 83.175percent | derived | Jun 25, 2026 | 100% | View source
|
|---|