Skip to main content
Benchmark Registry
ModelsBenchmarksOrganizations

CharXiv

Benchmark
CharXiv
Evaluated by
CharXiv Authors
Release date
June 26, 2024
Version
CharXiv Reasoning — No tools
Metric
CharXiv Accuracy

Results

4 results

LatestHistory
All providersAlibaba (Tongyi)AnthropicGoogle (DeepMind)Meta AI (originally Facebook AI Research)Moonshot AIOpenAIThinking Machines Lab
CharXiv Reasoning — No tools results
ProviderModelScoreSourceRegistry No.
Google (DeepMind)Gemini 3.8 Flash (no tools)86.2%Source for Gemini 3.8 Flash (no tools) on CharXiv Reasoning — No tools (opens in a new tab)30004
Google (DeepMind)Gemini 3.7 Flash84.5%Source for Gemini 3.7 Flash on CharXiv Reasoning — No tools (opens in a new tab)30010
Google (DeepMind)Gemini 3.6 Flash85.2%Source for Gemini 3.6 Flash on CharXiv Reasoning — No tools (opens in a new tab)30007
Google (DeepMind)Gemini 3.5 Flash84.2%Source for Gemini 3.5 Flash on CharXiv Reasoning — No tools (opens in a new tab)30006
Page 1 of 1

© 2026 Densa Labs

Toggle between the data update date and the application build time.
Legal
Color theme