Skip to content

Qwen3 8B

ModelOpen sourceActive
by Alibaba (Qwen) · family “qwen3b”

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,... (description from the OpenRouter listing)

in: textout: textReasoningTool useStructured output
Epoch Capabilities Index
136.2 #130 of 268
90% CI 130.1 – 138.0
Listing ↗
Released
28 Apr 2025
Context window
131K tokens
Max output
8K tokens
Input $ / 1M tokens
$0.12
Output $ / 1M tokens
$0.46
Cached input $ / 1M
Unknown
Each benchmark shown separately with its own source

Benchmark results (6)

BenchmarkDomainScorevs best recordedSettingRunSource
LMCAagents8.8%
13%
——External ↗
DTBenchreasoning59.7%
60%
——External ↗
Chess Puzzlesgames5.0% ±2.2
7%
—27 Aug 2026Eval log ↗
Fiction.LiveBenchlong-context62.1%
64%
——External ↗
OTIS Mock AIME 2024-2025math56.1% ±6.5
56%
—27 Aug 2026Eval log ↗
GPQA diamondscience56.8% ±2.9
59%
—27 Aug 2026Eval log ↗

Source: Epoch AI Benchmarking Hub (CC BY 4.0). “External” rows are leaderboard results Epoch collects from third parties. Best reported setting per benchmark is shown.

API list price over time

Price history

$ per 1M tokens

13 recorded prices on 1 Oct 2025 (OpenRouter listing + Internet Archive snapshots); re-read on every data refresh, most recently 10h ago. Steps show when the price changed.

Events