Qwen: Qwen3 VL 8B Instruct

qwen/qwen3-vl-8b-instruct

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

Modalities
ImageText
In / out price
₩205 / ₩799
1M
Context
262.14K
Released
Oct 15, 2025

Apps

Apps that use this model the most.

Appstokens
Hermes Agent11.32T
Zazen (Freebuff fork)3.16T
Cline2.56T
Kilo Code2.44T
pi1.59T
OpenClaw1.05T
Codex938.81B
DeepSeek Harness847.26B
Portkey AI639.39B
HighLevel375.1B
ISEKAI ZERO364.56B
Framer348.63B
Hello Minds, powered by Ethoswarm316.97B
ZCode279.83B
Descript253.44B