GLM 4.7 Flash
Z.AI
z-ai/glm-4.7-flashGLM-4.7-Flash is a 30B-class SOTA model from Z.ai that balances performance and efficiency, offering a faster, cheaper alternative to the full GLM-4.7 model. It targets general reasoning and coding tasks where speed and cost matter more than maximum capability.
Cost rate
0.1x
Context
203K
Released
Jan 19, 2026
Input
Text
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
Software & IT Services
#156 · top 40%
Mathematical
#160 · top 43%
Business, Management, & Finance
#171 · top 44%
Legal & Government
#179 · top 49%
Performance
Median latency and throughput measured across recent requests.
Throughput104 tok/s
Latency