Zhipu AI·released 2026-01-191 source
Not yet benchmarked
Consensus: LiteLLM · Cross-check: models.dev, OpenRouter · See every model’s pricing →
Claims drawn from cited facts, not live model generation.
Developed by Zhipu AI and released in 2026, this model is a large system with roughly 31.22 billion parameters. That scale gives it substantial headroom for demanding tasks, though it carries higher computational expense than more compact alternatives.
4 cited facts
This model is not yet benchmarked on any tracked benchmark, and therefore holds no current top score.
2 cited facts
The average gap to the SOTA leader is not available, and this model holds the top score on none of the tracked benchmarks. Thus, there is no evidence of genuine category leadership on the tracked leaderboards.
3 cited facts
6 facts cross-checked across data sources: 2 single-source, 4 disputed.
Source facts, citations, and refresh stamp for this record.
Sources: huggingface_models, litellm_prices, modelsdev_models, openrouter_models
GLM-4.7-Flash is an AI model developed by Zhipu AI, released 2026-01-19. Its benchmark record below tags every score as independently reproduced, self-reported, or unverified.
Listed API pricing: $0 per million input tokens, $0 per million output tokens (prices disputed across sources). See the pricing block for the full breakdown.