| Microsoft AI | MAI-Thinking-1 | 85.0% | Source for MAI-Thinking-1 on MMLU-Pro Original (opens in a new tab) | 70003 |
| Google (DeepMind) | Gemma 4 26B A4B Instruct (thinking mode) | 82.6% | Source for Gemma 4 26B A4B Instruct (thinking mode) on MMLU-Pro Original (opens in a new tab) | 35001 |
| Google (DeepMind) | Gemma 4 31B Instruct (thinking mode) | 85.2% | Source for Gemma 4 31B Instruct (thinking mode) on MMLU-Pro Original (opens in a new tab) | 35002 |
| Google (DeepMind) | Gemma 4 E2B Instruct (thinking mode) | 60.0% | Source for Gemma 4 E2B Instruct (thinking mode) on MMLU-Pro Original (opens in a new tab) | 35003 |
| Google (DeepMind) | Gemma 4 E4B Instruct (thinking mode) | 69.4% | Source for Gemma 4 E4B Instruct (thinking mode) on MMLU-Pro Original (opens in a new tab) | 35004 |
| Google (DeepMind) | Gemma 4 12B Instruct (thinking mode) | 77.2% | Source for Gemma 4 12B Instruct (thinking mode) on MMLU-Pro Original (opens in a new tab) | 35005 |
| Alibaba (Tongyi) | Qwen3.7-Plus | 88.5% | Source for Qwen3.7-Plus on MMLU-Pro Original (opens in a new tab) | 130007 |
| Alibaba (Tongyi) | Qwen3.7-Max | 89.6% | Source for Qwen3.7-Max on MMLU-Pro Original (opens in a new tab) | 130006 |
| DeepSeek AI | DeepSeek-V4-Flash (non-thinking) | 83.0% | Source for DeepSeek-V4-Flash (non-thinking) on MMLU-Pro Original (opens in a new tab) | 110001 |
| DeepSeek AI | DeepSeek-V4-Flash (max) | 86.2% | Source for DeepSeek-V4-Flash (max) on MMLU-Pro Original (opens in a new tab) | 110001 |
| DeepSeek AI | DeepSeek-V4-Flash (high) | 86.4% | Source for DeepSeek-V4-Flash (high) on MMLU-Pro Original (opens in a new tab) | 110001 |
| DeepSeek AI | DeepSeek-V4-Pro (max) | 87.5% | Source for DeepSeek-V4-Pro (max) on MMLU-Pro Original (opens in a new tab) | 110002 |
| DeepSeek AI | DeepSeek-V4-Pro (high) | 87.1% | Source for DeepSeek-V4-Pro (high) on MMLU-Pro Original (opens in a new tab) | 110002 |
| DeepSeek AI | DeepSeek-V4-Pro (non-thinking) | 82.9% | Source for DeepSeek-V4-Pro (non-thinking) on MMLU-Pro Original (opens in a new tab) | 110002 |
| Alibaba (Tongyi) | Qwen3.6-35B-A3B | 85.2% | Source for Qwen3.6-35B-A3B on MMLU-Pro Original (opens in a new tab) | 130005 |
| Alibaba (Tongyi) | Qwen3.6-Plus | 88.5% | Source for Qwen3.6-Plus on MMLU-Pro Original (opens in a new tab) | 130004 |
| Mistral AI | Mistral Small 4 (high) | 78.0% | Source for Mistral Small 4 (high) on MMLU-Pro Original (opens in a new tab) | 90004 |
| Mistral AI | Mistral Small 4 (none) | 73.5% | Source for Mistral Small 4 (none) on MMLU-Pro Original (opens in a new tab) | 90004 |
| NVIDIA | NVIDIA Nemotron 3 Super 120B-A12B | 83.7% | Source for NVIDIA Nemotron 3 Super 120B-A12B on MMLU-Pro Original (opens in a new tab) | 60002 |
| Alibaba (Tongyi) | Qwen3.5-397B-A17B | 87.8% | Source for Qwen3.5-397B-A17B on MMLU-Pro Original (opens in a new tab) | 130003 |
| Moonshot AI | Kimi K2.5 (thinking) | 87.1% | Source for Kimi K2.5 (thinking) on MMLU-Pro Original (opens in a new tab) | 120001 |
| Alibaba (Tongyi) | Qwen3-Max-Thinking | 85.7% | Source for Qwen3-Max-Thinking on MMLU-Pro Original (opens in a new tab) | 130002 |
| Z.ai | GLM-4.7 | 84.3% | Source for GLM-4.7 on MMLU-Pro Original (opens in a new tab) | 150001 |
| NVIDIA | NVIDIA Nemotron 3 Nano 30B-A3B | 78.3% | Source for NVIDIA Nemotron 3 Nano 30B-A3B on MMLU-Pro Original (opens in a new tab) | 60001 |