Skip to main content
Benchmark Registry
Models
Benchmarks
Organizations
Compare
Search the registry
Search
MMLU-Pro
Benchmark
MMLU-Pro
Evaluated by
Qwen Team
Release date
June 3, 2024
Version
MMLU-Pro Original
Metric
Accuracy
Search models
Search
Results
2 results
Rows per page
50
100
500
Apply
Latest
History
All providers
Alibaba (Tongyi)
DeepSeek AI
Google (DeepMind)
Microsoft AI
Mistral AI
Moonshot AI
NVIDIA
Z.ai
MMLU-Pro Original results
Provider
Model
Score
Source
Registry No.
Mistral AI
Mistral Small 4
(high)
78.0%
Source
for Mistral Small 4 (high) on MMLU-Pro Original
(opens in a new tab)
90004
Mistral AI
Mistral Small 4
(none)
73.5%
Source
for Mistral Small 4 (none) on MMLU-Pro Original
(opens in a new tab)
90004
Previous
Page
1
of
1
Next