Berkeley Function Calling Leaderboard v4 leaderboard
9 ranked models · higher is better
View accessible chart data
| Model | Rank | Provider | Score |
|---|---|---|---|
| Qwen3.7-Max | #1 | Alibaba Cloud | 75% |
| Qwen3.7-Plus | #2 | Alibaba Cloud | 72.9% |
| LFM2.5-8B-A1B | #3 | LiquidAI | 49.7% |
| Mellum2-12B-A2.5B-Thinking | #4 | JetBrains | 45.6% |
| Mellum2-12B-A2.5B-Instruct | #5 | JetBrains | 44.2% |
| ZAYA1-8B | #6 | Zyphra | 39.2% |
| MiniCPM5-1B | #7 | OpenBMB | 25.1% |
| LFM2.5-VL-450M | #8 | LiquidAI | 21.1% |
| LFM2.5-230M | #9 | LiquidAI | 21.0% |

