Skip to main content
Benchmark Registry
Models
Benchmarks
Organizations
Search the registry
Search
CharXiv
Benchmark
CharXiv
Evaluated by
CharXiv Authors
Release date
June 26, 2024
Version
CharXiv Reasoning — No tools
Metric
CharXiv Accuracy
Search models
Search
Results
4 results
Rows per page
50
100
500
Apply
Latest
History
All providers
Alibaba (Tongyi)
Anthropic
Google (DeepMind)
Meta AI (originally Facebook AI Research)
Moonshot AI
OpenAI
Thinking Machines Lab
CharXiv Reasoning — No tools results
Provider
Model
Score
Source
Registry No.
Google (DeepMind)
Gemini 3.8 Flash
(no tools)
86.2%
Source
for Gemini 3.8 Flash (no tools) on CharXiv Reasoning — No tools
(opens in a new tab)
30004
Google (DeepMind)
Gemini 3.7 Flash
84.5%
Source
for Gemini 3.7 Flash on CharXiv Reasoning — No tools
(opens in a new tab)
30010
Google (DeepMind)
Gemini 3.6 Flash
85.2%
Source
for Gemini 3.6 Flash on CharXiv Reasoning — No tools
(opens in a new tab)
30007
Google (DeepMind)
Gemini 3.5 Flash
84.2%
Source
for Gemini 3.5 Flash on CharXiv Reasoning — No tools
(opens in a new tab)
30006
Previous
Page
1
of
1
Next