Yi 1.5 34B vs Qwen2.5 32B
Two open-weight models, head to head — task-by-task quality, memory needs, speed and license. Scores are 0-100 from public evaluations; speed and VRAM are computed for a common 12GB card.
Verdict
The Qwen2.5 32B is the stronger all-rounder (85 vs 81 average), but pick Yi 1.5 34B for chat and Qwen2.5 32B for multilingual. Both are Apache-2.0.
Specs
| Yi 1.5 34B | Qwen2.5 32B | |
|---|---|---|
| Parameters | 34.4B | 32.5B |
| Context | 4K | 128K |
| VRAM (Q4) | 23 GB | 22 GB |
| Speed on RTX 4070 | 2.6 tok/s | 2.7 tok/s |
| License | Apache-2.0 | Apache-2.0 |
| Released | 2024-05 | 2024-09 |
Task-by-task quality
coding
80
86
reasoning
83
86
math
82
85
chat
84
86
multilingual
80
87
agent
76
82
Average8185
Yi 1.5 34BQwen2.5 32B
Can your GPU run them?
Frequently asked
Is Yi 1.5 34B or Qwen2.5 32B better?
Overall the Qwen2.5 32B edges it on average quality (85 vs 81 across six task areas), but it's use-case dependent — Yi 1.5 34B leads on chat, while Qwen2.5 32B leads on multilingual.
Which is better for coding, Yi 1.5 34B or Qwen2.5 32B?
Qwen2.5 32B — it scores 86/100 for coding vs 80/100.
Which needs more VRAM?
Yi 1.5 34B is larger (34.4B vs 32.5B) and needs more VRAM. At Q4 they need about 23 GB and 22 GB respectively.
Which is faster?
On a 12GB RTX 4070 at Q4, Yi 1.5 34B runs at ~2.6 tok/s and Qwen2.5 32B at ~2.7 tok/s.
More model comparisons
Run the interactive advisor
Auto-detect your exact hardware and get personalised picks, speed & memory.