Skip to content

Codestral 22B vs Qwen2.5-Coder 32B

Two open-weight models, head to head — task-by-task quality, memory needs, speed and license. Scores are 0-100 from public evaluations; speed and VRAM are computed for a common 12GB card.

Verdict

The Qwen2.5-Coder 32B is the stronger all-rounder (85 vs 79 average), but pick Codestral 22B for coding and Qwen2.5-Coder 32B for reasoning. Licenses differ: Codestral 22B is MNPL, Qwen2.5-Coder 32B is Apache-2.0.

Specs

Codestral 22BQwen2.5-Coder 32B
Parameters22.2B32B
Context32K128K
VRAM (Q4)15 GB22 GB
Speed on RTX 40705.9 tok/s2.7 tok/s
LicenseMNPLApache-2.0
Released2024-052024-11

Task-by-task quality

coding
90
92
reasoning
74
82
math
76
84
chat
76
82
multilingual
74
82
agent
82
86
Average7985
Codestral 22BQwen2.5-Coder 32B

Can your GPU run them?

Frequently asked

Is Codestral 22B or Qwen2.5-Coder 32B better?

Overall the Qwen2.5-Coder 32B edges it on average quality (85 vs 79 across six task areas), but it's use-case dependent — Codestral 22B leads on coding, while Qwen2.5-Coder 32B leads on reasoning.

Which is better for coding, Codestral 22B or Qwen2.5-Coder 32B?

Qwen2.5-Coder 32B — it scores 92/100 for coding vs 90/100.

Which needs more VRAM?

Qwen2.5-Coder 32B is larger (32B vs 22.2B) and needs more VRAM. At Q4 they need about 15 GB and 22 GB respectively.

Which is faster?

On a 12GB RTX 4070 at Q4, Codestral 22B runs at ~5.9 tok/s and Qwen2.5-Coder 32B at ~2.7 tok/s.

More model comparisons

Run the interactive advisor
Auto-detect your exact hardware and get personalised picks, speed & memory.