Qwen·released 2025-07-291 source
Qwen3-30B-A3B-Instruct (Jul 2025) benchmark scores: 5 benchmarks tracked. 0% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 5 benchmarks·best result 62.18 on OTIS Mock AIME 2024-2025 (unverified)·0 of 5 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 62.18accuracy (%) | unverifiedT2 | 2025-07-29 |
| DTBench | 45.33accuracy (%) | unverifiedT1 | 2025-07-29https://conceptualreasoning.ai/dtbench |
| GPQA diamond | 40.82accuracy (%) | unverifiedT2 | 2025-07-29 |
| LMCA | 26.32accuracy (%) | unverifiedT1 | 2025-07-29https://conceptualreasoning.ai/lmca |
| Chess Puzzles | 0accuracy (%) | unverifiedT2 | 2025-07-29 |
Claims drawn from cited facts, not live model generation.
This model was developed by Qwen. This model was released in 2025.
2 cited facts
This model is tracked against 5 benchmarks. It currently holds no top score on any of them. Its record is unverified, meaning no outside evaluation has either confirmed or disputed the reported numbers.
3 cited facts
On average this model trails the best reported score by 52.41 points, a wide gap on a disclosed harness that places it well behind the leaders by the plain band reading and should not be taken as a capability claim. It holds the top score on 0 of its benchmarks, meaning there is no category in which it currently leads.
2 cited facts
1 facts cross-checked across data sources: 1 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: epoch_benchmarks
Qwen3-30B-A3B-Instruct (Jul 2025) is an AI model developed by Qwen, released 2025-07-29. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
Qwen3-30B-A3B-Instruct (Jul 2025) has recorded scores on 5 benchmarks, each shown with its evidence status.
Qwen3-30B-A3B-Instruct (Jul 2025) has recorded scores on 5 benchmarks — GPQA diamond, OTIS Mock AIME 2024-2025, DTBench, LMCA, Chess Puzzles. The full table above shows each score with its evidence status.
0 of 5 recorded scores (0%) are independently reproduced rather than self-reported by the lab.
Qwen3-30B-A3B-Instruct (Jul 2025) has 5 tracked claims: 5 unverified.