modelbenchmark.io

Qwen3 8B

Alibaba · qwen3-8b

Compare

Qwen instruction model for multilingual chat, reasoning, and tool use

Specification

alibaba listing

Context
131K
Max output
8K
Released
2025-04
Knowledge cutoff
2025-04
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · alibaba list price

Input
$0.18
Output
$0.70
Cache read
$0.235
Cache write
$0.20

Quality

3 benchmarks · 1 source

Composite
44th
percentile of 334 scored models
Rank
188 / 334
±0.10 sd
Evidence
3 × 1
one source, or fewer than 3 benchmarks
Effort range
one configuration only
reasoning36th2/4 bench
math61st1/7 bench

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
OTIS Mock AIME 2024-202556.1 ±6.5defaultEpoch AI2026-08-27
GPQA diamond56.8 ±2.9defaultEpoch AI2026-08-27
Chess Puzzles5.0 ±2.2defaultEpoch AI2026-08-27

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 10 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
alibabavendor$0.18$0.70131K8K
llamagate · llamagate$0.04$0.1433K8K
siliconflow$0.06$0.06131K131K
siliconflow-cn$0.06$0.06131K131K
alibaba-cn$0.072$0.287131K8K
kilo$0.117$0.455131K8K
openrouter$0.117$0.455131K8K
fireworks_ai$0.20$0.2041K41K
pioneer$0.20$0.20$0.20$0.2041K41K
nano-gpt$0.47$0.47$0.23541K33K

Price history

append-only observations · a listing writes a row only when its price moves

DateHostField$/M
2026-09-02openrouterInput$0.117
2026-09-06openrouterInput$0.117
2026-09-02openrouterOutput$0.455
2026-09-06openrouterOutput$0.455
2026-09-05nano-gpt · teeCache read$0.11
2026-09-09nano-gpt · teeCache read$0.11
2026-09-05nano-gpt · teeInput$0.11
2026-09-09nano-gpt · teeInput$0.11
2026-09-05nano-gpt · teeOutput$0.45
2026-09-09nano-gpt · teeOutput$0.45

Other listings of this model

same model, different key

More from Alibaba

most-hosted first