DeepSeek·released 2025-01-201 source
DeepSeek-R1-Distill-Qwen-14B benchmark scores: 5 benchmarks tracked. 0% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 5 benchmarks·best result 1481 on Codeforces rating (self-reported)·0 of 5 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| Codeforces rating | 1481rating | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| MATH-500 | 93.9pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| AIME | 69.7pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| GPQA diamond | 59.1accuracy (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| LiveCodeBench | 53.1pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
Claims drawn from cited facts, not live model generation.
This model originates from DeepSeek and was released in 2025.
2 cited facts
This model is scored on 5 benchmarks. It holds no current top score. Because the status is self-reported, read these numbers as vendor claims rather than independently confirmed results.
3 cited facts
The model trails the leader by an average of 121.6 points on its benchmarks, which corresponds to a wide gap well behind the leaders. It holds the top score on 0 benchmarks, indicating it has not yet achieved genuine category leadership.
2 cited facts
1 facts cross-checked across data sources: 1 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: curated_benchmark_scores, curated_capability_claims, curated_models
DeepSeek-R1-Distill-Qwen-14B is an AI model developed by DeepSeek, released 2025-01-20. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
DeepSeek-R1-Distill-Qwen-14B has recorded scores on 5 benchmarks, each shown with its evidence status.
DeepSeek-R1-Distill-Qwen-14B has recorded scores on 5 benchmarks — AIME, Codeforces rating, GPQA diamond, LiveCodeBench, MATH-500. The full table above shows each score with its evidence status.
0 of 5 recorded scores (0%) are independently reproduced rather than self-reported by the lab.
DeepSeek-R1-Distill-Qwen-14B has 5 tracked claims: 5 self-reported.