The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

Modalities
TextImageVideo
In / out price
₩114 / ₩456
1M
Context
10 लाख
रिलीज़
26 फ़र॰ 2026

Providers

कई स्थान एक ही model को होस्ट करते हैं। Requests को प्राथमिकता और स्वास्थ्य के आधार पर वितरित किया जाता है, और यदि उपलब्ध हो तो local GPU को हमेशा प्राथमिकता दी जाती है।

ProviderContextQuantizationMax outputPriorityState
OpenRouter10 लाख50unknown