Skip to main content
Benchmark Registry
ModelsBenchmarksOrganizations

Humanity's Last Exam

Benchmark
Humanity's Last Exam
Evaluated by
Center for AI Safety, Scale AI
Release date
January 24, 2025
Version
Humanity's Last Exam Full set — No tools
Metric
Humanity's Last Exam Accuracy

Results

2 results

LatestHistory
All providersAlibaba (Tongyi)AnthropicDeepSeek AIGoogle (DeepMind)Moonshot AINVIDIAOpenAI
Humanity's Last Exam Full set — No tools results
ProviderModelScoreSourceRegistry No.
DeepSeek AIDeepSeek-V4-Flash (max)34.8%Source for DeepSeek-V4-Flash (max) on Humanity's Last Exam Full set — No tools (opens in a new tab)110001
DeepSeek AIDeepSeek-V4-Pro (max)37.7%Source for DeepSeek-V4-Pro (max) on Humanity's Last Exam Full set — No tools (opens in a new tab)110002
Page 1 of 1

© 2026 Densa Labs

Toggle between the data update date and the application build time.
Legal
Color theme