Qwen: Qwen3 VL 32B Instruct

qwen/qwen3-vl-32b-instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Modalities
TextImage
In / out price
₩183 / ₩730
1M
Context
131.07K
Released
Oct 23, 2025

Apps

Apps that use this model the most.

Appstokens
Hermes Agent11.32T
Zazen (Freebuff fork)3.16T
Cline2.56T
Kilo Code2.44T
pi1.59T
OpenClaw1.05T
Codex938.81B
DeepSeek Harness847.26B
Portkey AI639.39B
HighLevel375.1B
ISEKAI ZERO364.56B
Framer348.63B
Hello Minds, powered by Ethoswarm316.97B
ZCode279.83B
Descript253.44B