Qwen·released 2026-02-241 source
Qwen3.5-35B-A3B benchmark scores: 5 benchmarks tracked. 0% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 5 benchmarks·best result 77.95 on GPQA diamond (unverified)·0 of 5 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| GPQA diamond | 77.95accuracy (%) | unverifiedT2 | 2026-02-24 |
| OTIS Mock AIME 2024-2025 | 69.97accuracy (%) | unverifiedT2 | 2026-02-24 |
| DTBench | 66.67accuracy (%) | unverifiedT1 | 2026-02-24https://conceptualreasoning.ai/dtbench |
| LMCA | 34.67accuracy (%) | unverifiedT1 | 2026-02-24https://conceptualreasoning.ai/lmca |
| Chess Puzzles | 5.3accuracy (%) | unverifiedT2 | 2026-02-24 |
Claims drawn from cited facts, not live model generation.
This model was developed by Qwen and released in 2026, placing it among the newer generation of models from that team. No parameter count is disclosed here, so its scale can only be described in terms of origin and vintage rather than raw size.
4 cited facts
This model is scored on 5 tracked benchmarks, yet it holds no current top score on any of them. Its results have been neither confirmed nor disputed by outside evaluation, so its numbers remain unverified and should be read as claims rather than established outcomes.
3 cited facts
On average it trails the benchmark leader by 36.43 points under the disclosed harness, a wide gap that reads as well behind the leaders rather than at the frontier. It holds the top score on 0 benchmarks, meaning there is no category leadership yet on these benchmarks.
2 cited facts
1 facts cross-checked across data sources: 1 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: epoch_benchmarks
Qwen3.5-35B-A3B is an AI model developed by Qwen, released 2026-02-24. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
Qwen3.5-35B-A3B has recorded scores on 5 benchmarks, each shown with its evidence status.
Qwen3.5-35B-A3B has recorded scores on 5 benchmarks — GPQA diamond, OTIS Mock AIME 2024-2025, DTBench, LMCA, Chess Puzzles. The full table above shows each score with its evidence status.
0 of 5 recorded scores (0%) are independently reproduced rather than self-reported by the lab.
Qwen3.5-35B-A3B has 5 tracked claims: 5 unverified.