modelbenchmark.io

DeepSeek R1 Distill Qwen 1.5B

DeepSeek · deepseek-r1-distill-qwen-1-5b

Compare

Qwen instruction model for multilingual chat, reasoning, and tool use

Specification

most-agreed values

Context
33K
Max output
16K
Released
2025-01-01
Knowledge cutoff
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
free
Output
free

Quality

3 benchmarks · 1 source

Composite
17th
percentile of 334 scored models
Rank
277 / 334
±0.10 sd
Evidence
3 × 1
one source, or fewer than 3 benchmarks
Effort range
one configuration only
reasoning15th2/4 bench
math19th1/7 bench

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
OTIS Mock AIME 2024-202521.4 ±5.3defaultEpoch AI2026-08-28
Chess Puzzles0.0 ±0.0defaultEpoch AI2026-08-28
GPQA diamond33.6 ±2.1defaultEpoch AI2026-08-28

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 4 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
alibaba-cnfreefree33K16K
nscale$0.09$0.09
fireworks_ai$0.10$0.10131K131K
together_ai$0.18$0.18

More from DeepSeek

most-hosted first