Qwen: Qwen3 VL 32B Instruct

qwen/qwen3-vl-32b-instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Modalities
TextImage
In / out price
₩183 / ₩730
1M
Context
131.07K
Released
Oct 23, 2025

Providers

Several places host the same model. Requests are spread by priority and health, and a local GPU always goes first when one is available.

ProviderContextQuantizationMax outputPriorityState
OpenRouter131.07K50unknown