Models

16 models

Z.ai: GLM 5.3 FlashTextImageVideo
11.95T tokens weekly

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

by z-aiAug 26, 20261.31M context₩132 / ₩439 · 1M

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

by z-aiAug 26, 20261.05M context₩263 / ₩878 · 1M
2.24T tokens weekly

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

by z-aiAug 19, 20261.31M context₩2,457 / ₩7,722 · 1M
2.19T tokens weekly

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

by z-aiJun 17, 20261.05M context₩1,695 / ₩5,328 · 1M

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

by z-aiJun 17, 2026256K contextFree / Free · 1M

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

by z-aiApr 8, 2026204.8K context₩1,695 / ₩5,328 · 1M
Z.ai: GLM 5V TurboImageTextVideo

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

by z-aiApr 2, 2026202.75K context₩2,106 / ₩7,020 · 1M

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

by z-aiMar 15, 2026202.75K context₩2,106 / ₩7,020 · 1M

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

by z-aiFeb 12, 2026204.8K context₩1,053 / ₩3,370 · 1M

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

by z-aiJan 19, 2026202.75K context₩105 / ₩702 · 1M

GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...

by z-aiDec 22, 2025204.8K context₩702 / ₩3,071 · 1M
Z.ai: GLM 4.6VImageTextVideo

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

by z-aiDec 9, 2025131.07K context₩527 / ₩1,580 · 1M

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

by z-aiSep 30, 2025204.8K context₩965 / ₩3,861 · 1M

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

by z-aiAug 11, 202565.54K context₩1,053 / ₩3,159 · 1M

GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

by z-aiJul 26, 2025131.07K context₩1,053 / ₩3,861 · 1M

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

by z-aiJul 26, 2025131.07K context₩228 / ₩1,492 · 1M