Quick answer: GLM-4.5 is Zhipu AI's June 2025 model scoring 72.9% LiveCodeBench, 84.6% MMLU-Pro, 77.8% BFCL-v3, and 8.32% HLE. Proprietary — the standard tier in the GLM-4.5 family.
Where GLM-4.5 leads
Where it lags
Best for: Zhipu AI platform coding and knowledge tasks; BFCL function-calling pipelines; Jun 2025 GLM tier.
GLM-4.5 is the standard tier in Zhipu AI's June 2025 GLM-4.5 generation, above the GLM-4.5 Air variant. Its 72.9% LiveCodeBench outperforms GLM-4.7 (no LiveCodeBench data available) while MMLU-Pro (84.6%) is also higher than GLM-4.7 (no data).
| Field | Value |
|---|---|
| Organization | Zhipu AI |
| License | Proprietary |
| Release date | June 2025 |
| Modality | Text only |
Available via Zhipu AI API (bigmodel.cn). Refer to Zhipu pricing.
| Benchmark | Score | Source | Date |
|---|---|---|---|
| LiveCodeBench | 72.9% | Benchgen evaluation | 2025-06 |
| MMLU-Pro | 84.6% | Benchgen evaluation | 2025-06 |
| BFCL-v3 | 77.8% | Benchgen evaluation | 2025-06 |
| Humanity's Last Exam | 8.32% | Benchgen evaluation | 2025-06 |
| Model | LiveCodeBench | MMLU-Pro | BFCL-v3 | License |
|---|---|---|---|---|
| GLM-4.5 | 72.9% | 84.6% | 77.8% | Proprietary |
| GLM-4.5 Air | 70.7% | 81.4% | 76.4% | Proprietary |
| GLM-5.2 | — | — | — | Proprietary |
GLM-4.5 vs Air: higher LiveCodeBench (72.9% vs 70.7%), higher MMLU-Pro (84.6% vs 81.4%). Use GLM-4.5 for quality; Air for cost optimization.
Specs from Zhipu AI's GLM-4.5 release (June 2025) and Benchgen evaluations. Last updated 2026-07-24.