Qwen·released 2026-02-251 source
Qwen 3.5 Flash (hosted 35B-A3B) benchmark scores: 7 benchmarks tracked. 100% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 7 benchmarks·best result 84.43 on OTIS Mock AIME 2024-2025 (reproduced)·7 of 7 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 84.43accuracy (%) | reproduced· optimizedT1 | 2026-02-25Epoch AI |
| GPQA diamond | 76.43accuracy (%) | reproduced· optimizedT1 | 2026-02-25Epoch AI |
| SimpleQA Verified | 19.76accuracy (%) | reproducedT1 | 2026-02-25Epoch AI |
| Chess Puzzles | 16.88accuracy (%) | reproducedT1 | 2026-02-25Epoch AI |
| FrontierMath-2025-02-28-Private | 10.89accuracy (%) | reproduced· optimizedT1 | 2026-02-25Epoch AI |
| Mystery Game Puzzles | 6.37accuracy (%) | reproducedT1 | 2026-02-25Epoch AI |
| FrontierMath-Tier-4-2025-07-01-Private | 0accuracy (%) | reproducedT1 | 2026-02-25Epoch AI |
Claims drawn from cited facts, not live model generation.
This model was developed by Qwen and released in 2026.
2 cited facts
This model is tracked across seven benchmarks, and it currently holds no top score on any of them. Because the scores have been independently reproduced, the record can be trusted as verified rather than vendor self-report.
3 cited facts
Relative to disclosed benchmark leaders, this model trails by an average of 38.9 points under the harness's scoring, a wide gap that places it well behind the front-of-pack. It also holds the top score on none of the tracked benchmarks, confirming no current category leadership.
2 cited facts
1 facts cross-checked across data sources: 1 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: epoch_benchmarks
Qwen 3.5 Flash (hosted 35B-A3B) is an AI model developed by Qwen, released 2026-02-25. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
Qwen 3.5 Flash (hosted 35B-A3B) has recorded scores on 7 benchmarks, each shown with its evidence status.
Qwen 3.5 Flash (hosted 35B-A3B) has recorded scores on 7 benchmarks — GPQA diamond, OTIS Mock AIME 2024-2025, Chess Puzzles, FrontierMath-2025-02-28-Private, SimpleQA Verified, FrontierMath-Tier-4-2025-07-01-Private, and 1 more. The full table above shows each score with its evidence status.
7 of 7 recorded scores (100%) are independently reproduced rather than self-reported by the lab.
Qwen 3.5 Flash (hosted 35B-A3B) has 7 tracked claims: 7 independently reproduced.