Devstral Small (Jul '25)
Released
Benchmarks
Independent benchmark scores — composite indices for reasoning, coding, and math, plus individual eval scores where available.
Global rank#406 of 597 LLMs
TierEfficient
Intelligence Index9.1
Math Index29.3
Reasoning & knowledge
MMLU-Pro
62%
GPQA Diamond
41%
Humanity's Last Exam
4%
Long-context reasoning
19%
Coding
LiveCodeBench
25%
SciCode
24%
Agentic & tool use
Terminal-Bench Hard
6%
τ²-Bench Telecom
28%
Math & instruction following
AIME 2025
29%
IFBench
35%
See full leaderboard →Benchmarks via Artificial Analysis · View on AA
Available routes
No routes currently available — Devstral Small (Jul '25) isn't routed through the Opper gateway right now. It's tracked here for its release history.
Contact us about this model →