Z.AI: GLM 5.2 Fast
z-ai/glm-5.2-fast
GLM-5.2-Fast-Preview is the high-speed version of Zhipu AI’s flagship model, GLM-5.2. It supports an ultra-long context window of 1 million tokens and matches the capabilities of the standard GLM-5.2 edition, excelling in logical reasoning, long-text comprehension, and code generation. Through inference acceleration optimizations, its output throughput (TPS) reaches 1.5 to 2 times that of the standard version, significantly boosting output speed. It is ideally suited for scenarios where speed is critical, such as real-time conversations, multi-turn agent interactions, and streaming code generation.
Model specifications
- Input
- text
- Output
- text
- Context
- 1,000,000 tokens
- Max output
- 128,000 tokens
- Input price