Qwen·released 2025-04-281 source
Qwen3-30B-A3B benchmark scores: 8 benchmarks tracked. 0% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 8 benchmarks·best result 75.3 on Lech Mazur Writing (unverified)·0 of 8 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| Lech Mazur Writing | 75.3accuracy (%) | unverifiedT1 | 2025-04-28lechmazur/writing Github repository |
| OTIS Mock AIME 2024-2025 | 62.74accuracy (%) | unverifiedT2 | 2025-04-29 |
| GPQA diamond | 48.91accuracy (%) | unverifiedT2 | 2025-04-29 |
| Fiction.LiveBench | 40.6accuracy (%) | unverifiedT1 | 2025-04-28Fiction.live leaderboard |
| DTBench | 33.78accuracy (%) | unverifiedT1 | 2025-04-28https://conceptualreasoning.ai/dtbench |
| WeirdML | 29.75accuracy (%) | unverifiedT1 | 2025-04-28WeirdML Leaderboard |
| LMCA | 18.61accuracy (%) | unverifiedT1 | 2025-04-28https://conceptualreasoning.ai/lmca |
| Chess Puzzles | 0accuracy (%) | unverifiedT2 | 2025-04-29 |
Claims drawn from cited facts, not live model generation.
This model was developed by Qwen and released in 2025.
2 cited facts
This model is scored on 8 tracked benchmarks, and it currently holds no top score on any of them. Its record is unverified: no outside evaluation has either confirmed or disputed the reported numbers, so they should not be read as settled results.
3 cited facts
Under a disclosed harness, it trails the average state-of-the-art by 50.39 points — a score gap, not a capability claim — and that wide gap reads as well behind the leaders. It holds the top score on 0 benchmarks, meaning no genuine category leadership yet.
2 cited facts
5 facts cross-checked across data sources: 5 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: epoch_benchmarks, openrouter_models
Qwen3-30B-A3B is an AI model developed by Qwen, released 2025-04-28. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
Qwen3-30B-A3B has recorded scores on 8 benchmarks, each shown with its evidence status.
Qwen3-30B-A3B has recorded scores on 8 benchmarks — Lech Mazur Writing, GPQA diamond, OTIS Mock AIME 2024-2025, DTBench, WeirdML, LMCA, and 2 more. The full table above shows each score with its evidence status.
0 of 8 recorded scores (0%) are independently reproduced rather than self-reported by the lab.
Qwen3-30B-A3B has 8 tracked claims: 8 unverified.