modelbenchmark.io

qwen3-4b

Alibaba · qwen3-4b

Compare

Specification

most-agreed values

Context
33K
Max output
33K
Released
Knowledge cutoff
Retires
Open weights
Input
Output

Price

US dollars per million tokens · most-agreed

Input
$0.08
Output
$0.24

Quality

2 benchmarks · 1 source

Composite
25th
percentile of 334 scored models
Rank
251 / 334
±0.13 sd
Evidence
2 × 1
too little evidence to place with confidence
Effort range
one configuration only
reasoning27th2/4 bench

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
GPQA diamond48.0 ±2.7default · 2 runsEpoch AI2026-08-28
Chess Puzzles1.0 ±1.0defaultEpoch AI2026-08-28

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nebius$0.08$0.2433K33K
fireworks_ai$0.20$0.2041K41K

Other listings of this model

same model, different key

More from Alibaba

most-hosted first