modelbenchmark.io

Qwen3 32B

alibaba-qwen3-32b

Compare

Qwen instruction model for multilingual chat, reasoning, and tool use

Specification

most-agreed values

Context
33K
Max output
33K
Released
2025-04-30
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.25
Output
$0.55

Quality

4 benchmarks · 2 sources

Composite
41st
percentile of 334 scored models
Rank
197 / 334
±0.08 sd
Evidence
4 × 2
3 or more benchmarks, from 2 or more sources
Effort range
one configuration only
coding33rd1/3 bench
reasoning38th2/4 bench
math44th1/7 bench

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
GPQA diamond59.9 ±2.8default · 2 runsEpoch AI2026-08-30
OTIS Mock AIME 2024-202545.0 ±5.1default · 2 runsEpoch AI2026-08-30
Aider polyglot40.0default · aider diffAider2025-05-08
Chess Puzzles3.0 ±1.0default · 2 runsEpoch AI2026-08-28

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
digitalocean$0.25$0.5533K33K
gradient_ai · gradient-ai131K41K

Other listings of this model

this page is one of them

The corpus files Qwen3 32B under several keys. This page is the alibaba-qwen3-32b listing. The full record — every host, every price and every benchmark score — is on the main Qwen3 32B page.