DeepSeek·released 2025-01-201 source
DeepSeek-R1-Distill-Qwen-7B benchmark scores: 5 benchmarks tracked. 0% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 5 benchmarks·best result 1189 on Codeforces rating (self-reported)·0 of 5 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| Codeforces rating | 1189rating | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| MATH-500 | 92.8pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| AIME | 55.5pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| GPQA diamond | 49.1accuracy (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
| LiveCodeBench | 37.6pass@1 (%) | self-reported· optimizedT1 | 2025-01-22DeepSeek AI |
Claims drawn from cited facts, not live model generation.
This model originates from DeepSeek and was released in 2025.
2 cited facts
The model is currently scored on 5 tracked benchmarks. It holds no current top score among them. Because these figures are self-reported, readers should treat them as vendor claims rather than independently verified results.
3 cited facts
The model trails the benchmark leader by an average of 188.16 points, representing a wide gap that reads as well behind the leaders. It holds the top score on zero benchmarks, indicating it has not yet achieved category leadership.
2 cited facts
1 facts cross-checked across data sources: 1 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: curated_benchmark_scores, curated_models
DeepSeek-R1-Distill-Qwen-7B is an AI model developed by DeepSeek, released 2025-01-20. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
DeepSeek-R1-Distill-Qwen-7B has recorded scores on 5 benchmarks, each shown with its evidence status.
DeepSeek-R1-Distill-Qwen-7B has recorded scores on 5 benchmarks — AIME, Codeforces rating, GPQA diamond, LiveCodeBench, MATH-500. The full table above shows each score with its evidence status.
0 of 5 recorded scores (0%) are independently reproduced rather than self-reported by the lab.
DeepSeek-R1-Distill-Qwen-7B has 5 tracked claims: 5 self-reported.