Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
模态
图像文本
输入 / 输出价格
₩205 / ₩799
1M
上下文
26.21万
发布日期
2025年10月15日
供应商
多个平台托管同一个模型。请求将根据优先级和健康状况进行分配,如果本地 GPU 可用,则始终优先使用。
| 供应商 | 上下文 | 量化 | 最大输出 | 优先级 | 状态 |
|---|---|---|---|---|---|
| OpenRouter | 26.21万 | — | — | 50 | unknown |