modelbenchmark.io

DeepSeek R1 Distill Qwen 14B

DeepSeek · deepseek-r1-distill-qwen-14b

Compare

Qwen instruction model for multilingual chat, reasoning, and tool use

Specification

most-agreed values

Context
33K
Max output
16K
Released
2025-01-01
Knowledge cutoff
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.144
Output
$0.431

Quality

4 benchmarks · 1 source

Composite
48th
percentile of 334 scored models
Rank
175 / 334
±0.07 sd
Evidence
4 × 1
one source, or fewer than 3 benchmarks
Effort range
one configuration only
reasoning24th2/4 bench
math79th2/7 bench

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
MATH level 587.1 ±0.7defaultEpoch AI2025-03-11
OTIS Mock AIME 2024-202550.6 ±6.0defaultEpoch AI2026-08-28
GPQA diamond44.7 ±2.9defaultEpoch AI2025-03-10
Chess Puzzles1.0 ±1.0defaultEpoch AI2026-08-28

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 6 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nscale$0.07$0.07
alibaba-cn$0.144$0.43133K16K
novita$0.15$0.1533K16K
novita-ai$0.15$0.1533K16K
fireworks_ai$0.20$0.20131K131K
together_ai$1.60$1.60

More from DeepSeek

most-hosted first