DeepSeek·released 2024-01-251 source
DeepSeek Coder 1.3B benchmark scores: 4 benchmarks tracked. 0% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 4 benchmarks·best result 6.6 on Winogrande (self-reported)·0 of 4 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| Winogrande | 6.6accuracy (%) | self-reported· optimizedT1 | 2024-01-25Qwen2.5-Coder Technical Report |
| GSM8K | 4.4accuracy (%) | self-reported· optimizedT1 | 2024-01-25Qwen2.5-Coder Technical Report |
| MMLU | 1.07accuracy (%) | self-reported· optimizedT1 | 2024-01-25Qwen2.5-Coder Technical Report |
| ARC AI2 | 0.53accuracy (%) | self-reported· optimizedT1 | 2024-01-25Qwen2.5-Coder Technical Report |
Claims drawn from cited facts, not live model generation.
This model was developed by DeepSeek and released in 2024.
2 cited facts
Its 4 scored benchmarks form a minimal record — enough to appear in comparisons, too few to support a profile of strengths. First place is absent, which mostly means other entrants currently sit higher under their harnesses — a position, not an indictment. Because its scores are self-reported — meaning they are vendor-claimed and not yet independently confirmed — these results should be read as claims rather than verified findings.
3 cited facts
The model trails the leader by an average of 83.91 points—a wide gap that places it well behind the front-of-pack—and holds the top score on none of the benchmarks, indicating no category leadership.
2 cited facts
1 facts cross-checked across data sources: 1 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: epoch_benchmarks
DeepSeek Coder 1.3B is an AI model developed by DeepSeek, released 2024-01-25. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
DeepSeek Coder 1.3B has recorded scores on 4 benchmarks, each shown with its evidence status.
DeepSeek Coder 1.3B has recorded scores on 4 benchmarks — Winogrande, MMLU, ARC AI2, GSM8K. The full table above shows each score with its evidence status.
0 of 4 recorded scores (0%) are independently reproduced rather than self-reported by the lab.
DeepSeek Coder 1.3B has 4 tracked claims: 4 self-reported.