Mistral AI·released 2023-10-101 source
Mistral 7B v0.1 benchmark scores: 10 benchmarks tracked. 0% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 10 benchmarks·best result 75.2 on TriviaQA (unverified)·0 of 10 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| TriviaQA | 75.2accuracy (%) | unverified· optimizedT1 | 2023-10-10Gemma: Open Models Based on Gemini Research and Technology |
| HellaSwag | 74.67accuracy (%) | unverified· optimizedT1 | 2023-10-10Mixtral of Experts |
| OpenBookQA | 73.07accuracy (%) | self-reported· optimizedT1 | 2023-10-10Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone |
| ARC AI2 | 71.47accuracy (%) | self-reported· optimizedT1 | 2023-10-10Nemotron-4 15B Technical Report |
| PIQA | 66accuracy (%) | self-reported· optimizedT1 | 2023-10-10Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone |
| GSM8K | 54.4accuracy (%) | unverified· optimizedT1 | 2023-10-10Stanford HELM |
| Winogrande | 50.6accuracy (%) | unverified· optimizedT1 | 2023-10-10Training Compute-Optimal Large Language Models |
| MMLU | 50accuracy (%) | unverified· optimizedT1 | 2023-10-10Mixtral of Experts |
| BBH | 41.47accuracy (%) | unverified· optimizedT1 | 2023-10-10Gemma: Open Models Based on Gemini Research and Technology |
| ANLI | 20.65accuracy (%) | self-reported· optimizedT1 | 2023-10-10Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone |
Claims drawn from cited facts, not live model generation.
This model was developed by Mistral AI and released in 2023.
2 cited facts
This model is scored on 10 tracked benchmarks but holds no current top score; because these results are self-reported, they should be read as vendor claims rather than independently verified numbers.
3 cited facts
With an average gap of 23.79 points from the leader and no benchmark where it holds the top score, the model sits well behind the frontier.
2 cited facts
1 facts cross-checked across data sources: 1 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: epoch_benchmarks
Mistral 7B v0.1 is an AI model developed by Mistral AI, released 2023-10-10. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
Mistral 7B v0.1 has recorded scores on 10 benchmarks, each shown with its evidence status.
Mistral 7B v0.1 has recorded scores on 10 benchmarks — PIQA, OpenBookQA, Winogrande, TriviaQA, MMLU, ARC AI2, and 4 more. The full table above shows each score with its evidence status.
0 of 10 recorded scores (0%) are independently reproduced rather than self-reported by the lab.
Mistral 7B v0.1 has 10 tracked claims: 4 self-reported, 6 unverified.