Loading company intelligence…
Loading the latest published snapshot and its supporting evidence.
Loading the latest published snapshot and its supporting evidence.
Exact Qwen3-235B-A22B-Instruct-2507 (FC) configuration in the reviewed BFCL v4 leaderboard.
A current roll-up of published ranks, normalized facts, primary evidence, and entity relationships. No missing value is inferred.
| Metric | Value |
|---|---|
| model_tool_useBFCL evaluation total cost | $2.50reported |
| model_tool_useBFCL irrelevance-detection accuracy | 81.73 percentreported |
| model_tool_useBFCL live accuracy | 68.91 percentreported |
| model_tool_useBFCL mean latency | 2.57 secondsmeasured |
| model_tool_useBFCL memory accuracy | 23.87 percentreported |
| model_tool_useBFCL multi-turn accuracy | 45.38 percentreported |
| model_tool_useBFCL non-live AST accuracy | 37.4 percentreported |
| model_tool_useBFCL overall accuracy | 47.99 percentreported |
| model_tool_useBFCL p95 latency | 6.27 secondsmeasured |
| model_tool_useBFCL relevance-detection accuracy | 87.5 percentreported |
| model_tool_useBFCL web-search accuracy | 54 percentreported |
Every currently published position, as of Apr 13, 2026 · 100% confidence · methodology vbfcl-v4-2026-04-13-official-v1-bfcl-tool-use-configurations-v1.
| Rank | Leaderboard | Score | Explain |
|---|---|---|---|
| #31 | BFCL Tool-Use Configurations | 47.99 | Explain BFCL Tool-Use Configurations rank → |
Ranks exact model-plus-inference configurations by official UC Berkeley BFCL overall accuracy.
Some explanation inputs are unavailable Missing: publicFrozenEvidence, comparablePriorRanking.
100% confidence · 58% component coverage · as of Apr 13, 2026, 11:59 PM UTC
| Component | Value | Weight | Contribution | Evidence | Status |
|---|---|---|---|---|---|
| benchmarkVersion | v4 | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
| bfcl evaluation total cost USD | — | 0% | 0 | 1 | available |
| bfcl irrelevance detection accuracy | — | 0% | 0 | 1 | available |
| bfcl latency mean seconds |
11 published observations for Qwen3-235B-A22B-Instruct-2507 (FC), all at 100% confidence, as of Apr 13, 2026. Open a row for the source, evidence location, extraction method, and quality state.
| Metric | Value | Kind | Evidence |
|---|---|---|---|
| BFCL multi-turn accuracybfcl-multi-turn-accuracy | 45.38percent | reported | View source
|
| BFCL web-search accuracybfcl-web-search-accuracy | 54percent | reported | View source
|
| — |
| 0% |
| 0 |
| 1 |
| available |
| bfcl latency p95 seconds | — | 0% | 0 | 1 | available |
|---|
| bfcl live accuracy | — | 0% | 0 | 1 | available |
|---|
| bfcl memory accuracy | — | 0% | 0 | 1 | available |
|---|
| bfcl multi turn accuracy | — | 0% | 0 | 1 | available |
|---|
| bfcl non live ast accuracy | — | 0% | 0 | 1 | available |
|---|
| bfcl overall accuracy | — | 100% | 47.99 | 1 | available |
|---|
| bfcl relevance detection accuracy | — | 0% | 0 | 1 | available |
|---|
| bfcl web search accuracy | — | 0% | 0 | 1 | available |
|---|
| interactionMode | FC | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| metrics | Structured | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| organization | Qwen | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| rankingDirection | descending | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| reviewedAsOf | 2026-04-13 | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| sourceConfiguration | Qwen3-235B-A22B-Instruct-2507 (FC) | Unavailable | — | 0 | unavailableno_public_frozen_evidence |
|---|
| upstreamRank | 31 | Unavailable | 31 | 0 | unavailableno_public_frozen_evidence |
|---|
no_prior_published_snapshot
No component moved since the prior snapshot.
{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}{"csv":"https://gorilla.cs.berkeley.edu/data_overall.csv","repository":"https://github.com/ShishirPatil/gorilla/tree/main/berkeley-function-call-leaderboard","leaderboard":"https://gorilla.cs.berkeley.edu/leaderboard.html","rowIdentity":"Model","rankingColumn":"Overall Acc"}| BFCL overall accuracybfcl-overall-accuracy | 47.99percent | reported | View source
|
|---|
| BFCL non-live AST accuracybfcl-non-live-ast-accuracy | 37.4percent | reported | View source
|
|---|
| BFCL mean latencybfcl-latency-mean-seconds | 2.57seconds | measured | View source
|
|---|
| BFCL relevance-detection accuracybfcl-relevance-detection-accuracy | 87.5percent | reported | View source
|
|---|
| BFCL live accuracybfcl-live-accuracy | 68.91percent | reported | View source
|
|---|
| BFCL memory accuracybfcl-memory-accuracy | 23.87percent | reported | View source
|
|---|
| BFCL p95 latencybfcl-latency-p95-seconds | 6.27seconds | measured | View source
|
|---|
| BFCL evaluation total costbfcl-evaluation-total-cost-usd | 2.5USD · USD | reported | View source
|
|---|
| BFCL irrelevance-detection accuracybfcl-irrelevance-detection-accuracy | 81.73percent | reported | View source
|
|---|