Rabdos AI

Rabdos Vizi Bench

27 problems · Updated on August 6, 2026

For reasoning over scientific and mathematical visualizations.

Sample visualization from the Rabdos Vizi Bench evaluation set
Sample visualization
Leaderboard, ranked by average score.
RankProviderModelReasoning effortRubricStrict
1AnthropicClaude Opus 5max32.7%7/27 (25.9%)
2AnthropicClaude Fable 5max24.6%6/27 (22.2%)
3OpenAIGPT-5.6 Solmax12.4%2/27 (7.4%)
4GoogleGemini 3.6 Flashhigh9.4%2/27 (7.4%)
5AlibabaQwen 3.8 Maxmax5.6%0/27 (0.0%)
6xAIGrok 4.5xhigh4.2%0/27 (0.0%)
7Moonshot AIKimi K3max3.7%0/27 (0.0%)
8MetaMuse Spark 1.1xhigh1.9%0/27 (0.0%)

Interested in evaluating your model on Rabdos Vizi Bench? Get in touch →