DeepSeek R1
163,840 token context with transparent <think> delimiters showing reasoning over retrieved documents. MIT license enables fine-tuning on domain-specific retrieval tasks and full model customization.
Model Information
- Provider
- DeepSeek
- License
- Open Source
- Input Price per 1M
- $0.30
- Output Price per 1M
- $1.20
- Context Window
- 164K
- Release Date
- 2025-01-20
- Model Name
- deepseek-r1
- Total Evaluations
- 810
Performance Record
Wins164 (20.2%)
Losses540 (66.7%)
Ties106 (13.1%)
Wins
Losses
Ties
Performance Overview
ELO ratings by dataset
DeepSeek R1's ELO performance varies across different benchmark datasets, showing its strengths in specific domains.
DeepSeek R1 - ELO by Dataset
Detailed Metrics
Dataset breakdown
Performance metrics across different benchmark datasets, including accuracy and latency percentiles.
SciFact
ELO 142825.6% WR69W-139L-62T
Quality Metrics
- Correctness
- 4.93
- Faithfulness
- 4.97
- Grounding
- 4.93
- Relevance
- 5.00
- Completeness
- 4.83
- Overall
- 4.93
Latency Distribution
- Mean
- 14826ms
- Min
- 7765ms
- Max
- 33129ms
PG
ELO 131218.9% WR51W-210L-9T
Quality Metrics
- Correctness
- 4.93
- Faithfulness
- 4.93
- Grounding
- 4.90
- Relevance
- 4.97
- Completeness
- 4.60
- Overall
- 4.87
Latency Distribution
- Mean
- 23334ms
- Min
- 12280ms
- Max
- 85633ms
MSMARCO
ELO 117916.3% WR44W-191L-35T
Quality Metrics
- Correctness
- 4.73
- Faithfulness
- 4.77
- Grounding
- 4.77
- Relevance
- 4.87
- Completeness
- 4.37
- Overall
- 4.70
Latency Distribution
- Mean
- 16654ms
- Min
- 9675ms
- Max
- 31255ms
Compare Models
See how it stacks up
Compare DeepSeek R1 with other top llms to understand the differences in performance, accuracy, and latency.