Z.ai: GLM 4.7 Flash
Zhipu's flagship GLM-4. Strong on Chinese-language benchmarks, with frontier-class reasoning at competitive Chinese pricing.
Z.ai: GLM 4.7 Flash is a efficient AI model from z-ai. It costs $0.060 per million input tokens and $0.400 per million output tokens (blended $0.162/M), with a 202,752-token context window.
Profile inherited from upstream GLM-4 Plus ↗ — this is a hosted variant of the same open-weights model.
- Strong on Chinese benchmarks
- Tool use
- 128K context
- Tightly integrated with Zhipu cloud
- Chinese-first production chat
- Multilingual agents
Benchmarks
More from z-ai
See all 6 →Frequently asked questions
How much does Z.ai: GLM 4.7 Flash cost?
Z.ai: GLM 4.7 Flash costs $0.060 per million input tokens and $0.400 per million output tokens, for a blended reference rate of $0.162 per million tokens.
What is Z.ai: GLM 4.7 Flash's context window?
Z.ai: GLM 4.7 Flash supports up to 202,752 tokens of context in a single request.
What is Z.ai: GLM 4.7 Flash best for?
Z.ai: GLM 4.7 Flash is well suited to Strong on Chinese benchmarks, Tool use and 128K context.
Who makes Z.ai: GLM 4.7 Flash?
Z.ai: GLM 4.7 Flash is developed and served by z-ai.