A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...

Modalities
Text
In / out price
₩33 / ₩53
1M
Context
1.31 लाख
रिलीज़
19 जुल॰ 2024

Providers

कई स्थान एक ही model को होस्ट करते हैं। Requests को प्राथमिकता और स्वास्थ्य के आधार पर वितरित किया जाता है, और यदि उपलब्ध हो तो local GPU को हमेशा प्राथमिकता दी जाती है।

ProviderContextQuantizationMax outputPriorityState
OpenRouter1.31 लाख50unknown