Skip to main content
Benchmark Registry
ModelsBenchmarksOrganizationsCompare

Terminal-Bench

Benchmark
Terminal-Bench
Evaluated by
Terminal-Bench Team
Release date
November 7, 2025
Version
Terminal-Bench 2.0
Metric
Terminal-Bench Accuracy

Results

2 results

LatestHistory
All providersAnthropicCursorDeepSeek AIGoogle (DeepMind)Meta AI (originally Facebook AI Research)Microsoft AIMiniMaxMistral AIMoonshot AINVIDIAOpenAIZ.ai
Terminal-Bench 2.0 results
ProviderModelScoreSourceRegistry No.
Mistral AIDevstral 232.6%Source for Devstral 2 on Terminal-Bench 2.0 (opens in a new tab)90002
Mistral AIDevstral Small 222.5%Source for Devstral Small 2 on Terminal-Bench 2.0 (opens in a new tab)90003
Page 1 of 1

© 2026 Densa Labs

Toggle between the data update date and the application build time.
Legal
Color theme