Qwen: Qwen3 VL 32B Instruct

qwen/qwen3-vl-32b-instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Modalities
TextImage
In / out price
₩183 / ₩730
1M
Context
131.07K
Released
Oct 23, 2025

Pricing

Prices are quoted in Korean won. Changes to the exchange rate or margin show up immediately.

Input (prompt)₩183 / 1M
Output (completion)₩730 / 1M
Cache read₩18 / 1M
Cache write

1,000 input tokens plus 1,000 output tokens costs about ₩0.9126.