GLM-4.5-MoE-355B-A32B-0715
GLM-4.5-MoE-355B-A32B-0715 is an AI model published by Zhipu AI, released on 2025-07-28, for Reasoning model, with 355B parameters, and 128K context length, requiring about 710 GB storage, with a 1340.50 score on Creative Writing.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Model basics
Open source & experience
Official resources
API details
| Type | Condition | Input | Output |
|---|---|---|---|
| Text | - | ¥0.800/ 1M tokens | ¥2.00/ 1M tokens |
| Type | TTL | Write | Read |
|---|---|---|---|
| Text | - | ¥0.0000/ 1M tokens | ¥0.400/ 1M tokens |
| Text | - | ¥0.0000/ 1M tokens Cache = write | — |
| Text | - | — | ¥0.400/ 1M tokens Cache = hit |
“—” means the modality is not billed in that direction, or the vendor has not published a price for it.
Benchmark Results
GLM-4.5 currently shows benchmark results led by MATH-500 (3 / 45, score 98.20), AIME 2024 (14 / 62, score 91), MMLU Pro (34 / 134, score 84.60). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.
Coding and Software Engineer
2 evaluationsWriting and Creative Capabilities
1 evaluationsPublisher
Model Overview
GLM-4.5-MoE-355B-A32B-0715 is an AI model published by Zhipu AI, released on 2025-07-28, for Reasoning model, with 355B parameters, and 128K context length, requiring about 710 GB storage, with a 1340.50 score on Creative Writing.
DataLearner on WeChat
Follow DataLearner on WeChat for AI model updates and research notes.
