Qwen: Qwen3 VL 32B Instruct
qwen/qwen3-vl-32b-instruct
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
Modalities
TextImage
In / out price
₩183 / ₩730
1M
Context
131.07K
Released
Oct 23, 2025
Apps
Apps that use this model the most.
| Apps | tokens |
|---|---|
| Hermes Agent | 11.32T |
| Zazen (Freebuff fork) | 3.16T |
| Cline | 2.56T |
| Kilo Code | 2.44T |
| pi | 1.59T |
| OpenClaw | 1.05T |
| Codex | 938.81B |
| DeepSeek Harness | 847.26B |
| Portkey AI | 639.39B |
| HighLevel | 375.1B |
| ISEKAI ZERO | 364.56B |
| Framer | 348.63B |
| Hello Minds, powered by Ethoswarm | 316.97B |
| ZCode | 279.83B |
| Descript | 253.44B |