GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
Modalities
TextImage
In / out price
₩1,053 / ₩3,159
1M
Context
6.55万
リリース日
2025/08/11
このモデルのベンチマークデータはありません。