Google DeepMind·released 2025-03-121 source
Gemma 3 27B benchmark scores: 9 benchmarks tracked. 44% of its scores are independently reproduced — the rest are self-reported or unverified.
Leads on 0 of 9 benchmarks·best result 79.9 on Lech Mazur Writing (unverified)·4 of 9 independently reproduced
| Benchmark | Score | Evidence status | Measured |
|---|---|---|---|
| Lech Mazur Writing | 79.9accuracy (%) | unverifiedT1 | 2025-03-12lechmazur/writing Github repository |
| MATH level 5 | 74.04accuracy (%) | reproduced· optimizedT1 | 2025-03-12Epoch AI |
| GeoBench | 52accuracy (%) | unverifiedT1 | 2025-03-12GeoBench leaderboard |
| Fiction.LiveBench | 33.3accuracy (%) | unverifiedT1 | 2025-03-12Fiction.live leaderboard |
| GPQA diamond | 25.25accuracy (%) | reproduced· optimizedT1 | 2025-03-12Epoch AI |
| OTIS Mock AIME 2024-2025 | 22.14accuracy (%) | reproduced· optimizedT1 | 2025-03-12Epoch AI |
| Aider polyglot | 4.9accuracy (%) | unverified· optimizedT1 | 2025-03-12Aider LLM Leaderboards |
| Chess Puzzles | 0accuracy (%) | reproducedT1 | 2025-03-12Epoch AI |
| CritPt | 0accuracy (%) | unverifiedT2 | 2025-03-12 |
Claims drawn from cited facts, not live model generation.
This model originates from Google DeepMind. It was released in 2025.
2 cited facts
This model is scored on 9 tracked benchmarks and currently holds no top score; because these results have been independently reproduced, the record can be treated as verified.
3 cited facts
With an average gap of 50.37 points behind the leader, this model is well behind the frontier rather than competitive with top scores. It holds the top score on none of the tracked benchmarks, meaning it has no category leadership yet.
2 cited facts
1 facts cross-checked across data sources: 1 single-source.
Source facts, citations, and refresh stamp for this record.
Sources: epoch_benchmarks
Gemma 3 27B is an AI model developed by Google DeepMind, released 2025-03-12. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
Gemma 3 27B has recorded scores on 9 benchmarks, each shown with its evidence status.
Gemma 3 27B has recorded scores on 9 benchmarks — Lech Mazur Writing, GPQA diamond, MATH level 5, OTIS Mock AIME 2024-2025, Chess Puzzles, Aider polyglot, and 3 more. The full table above shows each score with its evidence status.
4 of 9 recorded scores (44%) are independently reproduced rather than self-reported by the lab.
Gemma 3 27B has 9 tracked claims: 4 independently reproduced, 5 unverified.