Z.ai: GLM 4.7 Flash
z-ai/glm-4.7-flash
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
Modalities
Text
In / out price
₩105 / ₩702
1M
Context
202.75K
Released
Jan 19, 2026
Providers
Several places host the same model. Requests are spread by priority and health, and a local GPU always goes first when one is available.
| Provider | Context | Quantization | Max output | Priority | State |
|---|---|---|---|---|---|
| OpenRouter | 202.75K | — | — | 50 | unknown |