DeepSeek-R1
Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
A language reasoning model release from DeepSeek, released with reinforcement-learning training and open model weights.
Pricing
Official site · Sep 8, 2026, 1:02 PM
Input / 1M
$0.56
Output / 1M
$1.68
Cached input / 1M
$0.07
Per 1M tokens; cache-miss input is USD 0.56, cache-hit input is USD 0.07, and output is USD 1.68.
DeepSeek's R1 launch documentation identifies the API model name as deepseek-reasoner. The current official pricing page lists deepseek-reasoner (currently DeepSeek-V3.1 Thinking Mode) at the stated per-1M-token rates, effective from 2025-09-05 16:00 UTC.
Model Intelligence
Arena ranking
13
Benchmarkable
Yes
Model level
release
Intelligence Index
13.1
Math Index
76
MMLU Pro
0.85
GPQA
0.81
HLE
0.16
LiveCodeBench
0.77
MATH-500
0.98
AIME
0.89
AIME 2025
0.76
IFBench
0.4
LCR
0.56
TerminalBench Hard
0.16
TAU2
0.37
Recent stories
0 linked stories
No linked stories yet.