DeepSeek·released 2025-01-201 source
DeepSeek-R1-Distill-Qwen-1.5B benchmark scores: 5 benchmarks tracked. 0% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 5 benchmarks·best result 954 on Codeforces rating (self-reported)·0 of 5 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| Codeforces rating | 954rating | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| MATH-500 | 83.9pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| GPQA diamond | 33.8accuracy (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| AIME | 28.9pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| LiveCodeBench | 16.9pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
Claims drawn from cited facts, not live model generation.
Developed by DeepSeek, this model was released in 2025 with 1.78 billion parameters, making it compact in scale. This size favors efficiency and lower operating cost, at the price of reduced headroom compared with larger, more expensive models.
4 cited facts
This model is benchmarked across five tracked evaluations. It currently holds no top score on any of them. Because the results are self-reported and not independently reproduced, treat the numbers as vendor claims rather than verified results.
3 cited facts
With an average gap of 249.52 points behind the SOTA leader, this model trails by a wide margin and sits well behind the front of the pack. It holds the top score on none of the tracked benchmarks, meaning it has no category leadership yet.
2 cited facts
2 facts cross-checked across data sources: 2 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: curated_benchmark_scores, huggingface_models
DeepSeek-R1-Distill-Qwen-1.5B is an AI model developed by DeepSeek, released 2025-01-20. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
DeepSeek-R1-Distill-Qwen-1.5B has recorded scores on 5 benchmarks, each shown with its evidence status.
DeepSeek-R1-Distill-Qwen-1.5B has recorded scores on 5 benchmarks — AIME, Codeforces rating, GPQA diamond, LiveCodeBench, MATH-500. The full table above shows each score with its evidence status.
0 of 5 recorded scores (0%) are independently reproduced rather than self-reported by the lab.
DeepSeek-R1-Distill-Qwen-1.5B has 5 tracked claims: 5 self-reported.