Codestral 22B vs Qwen2.5-Coder 32B
Two open-weight models, head to head — task-by-task quality, memory needs, speed and license. Scores are 0-100 from public evaluations; speed and VRAM are computed for a common 12GB card.
Verdict
The Qwen2.5-Coder 32B is the stronger all-rounder (85 vs 79 average), but pick Codestral 22B for coding and Qwen2.5-Coder 32B for reasoning. Licenses differ: Codestral 22B is MNPL, Qwen2.5-Coder 32B is Apache-2.0.
Specs
| Codestral 22B | Qwen2.5-Coder 32B | |
|---|---|---|
| Parameters | 22.2B | 32B |
| Context | 32K | 128K |
| VRAM (Q4) | 15 GB | 22 GB |
| Speed on RTX 4070 | 5.9 tok/s | 2.7 tok/s |
| License | MNPL | Apache-2.0 |
| Released | 2024-05 | 2024-11 |
Task-by-task quality
coding
90
92
reasoning
74
82
math
76
84
chat
76
82
multilingual
74
82
agent
82
86
Average7985
Codestral 22BQwen2.5-Coder 32B
Can your GPU run them?
Frequently asked
Is Codestral 22B or Qwen2.5-Coder 32B better?
Overall the Qwen2.5-Coder 32B edges it on average quality (85 vs 79 across six task areas), but it's use-case dependent — Codestral 22B leads on coding, while Qwen2.5-Coder 32B leads on reasoning.
Which is better for coding, Codestral 22B or Qwen2.5-Coder 32B?
Qwen2.5-Coder 32B — it scores 92/100 for coding vs 90/100.
Which needs more VRAM?
Qwen2.5-Coder 32B is larger (32B vs 22.2B) and needs more VRAM. At Q4 they need about 15 GB and 22 GB respectively.
Which is faster?
On a 12GB RTX 4070 at Q4, Codestral 22B runs at ~5.9 tok/s and Qwen2.5-Coder 32B at ~2.7 tok/s.
More model comparisons
Run the interactive advisor
Auto-detect your exact hardware and get personalised picks, speed & memory.