Модели

433 моделей

OpenAI: GPT-6 AstraФайлИзображениеТекст

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

от openai5 сент. 2026 г.контекст 1,05 млн17 550 ₩ / 87 750 ₩ · 1M
OpenAI: GPT-6 Astra (batch)ФайлИзображениеТекст

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

от openai5 сент. 2026 г.контекст 1,05 млн8 775 ₩ / 43 875 ₩ · 1M
OpenAI: GPT-6 Astra ProФайлИзображениеТекст

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

от openai5 сент. 2026 г.контекст 1,05 млн17 550 ₩ / 87 750 ₩ · 1M
OpenAI: GPT-6 Astra Pro (batch)ФайлИзображениеТекст

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

от openai5 сент. 2026 г.контекст 1,05 млн8 775 ₩ / 43 875 ₩ · 1M

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

от inclusionai5 сент. 2026 г.контекст 262,14 тыс.Бесплатно / Бесплатно · 1M
Qwen: Qwen3.8 Max (0902)ТекстИзображениеВидео

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

от qwen4 сент. 2026 г.контекст 1 млн3 510 ₩ / 10 530 ₩ · 1M
Meta: Muse Spark 1.3 ContributorТекстИзображениеВидео

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...

от meta3 сент. 2026 г.контекст 1,05 млн176 ₩ / 351 ₩ · 1M
Meta: Muse Spark 1.3ТекстИзображениеВидео

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...

от meta3 сент. 2026 г.контекст 1,05 млн2 194 ₩ / 7 459 ₩ · 1M
Google: Gemini 3.8 FlashТекстИзображениеВидео

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

от google3 сент. 2026 г.контекст 1,05 млн1 316 ₩ / 6 581 ₩ · 1M
Google: Gemini 3.8 Flash (batch)ТекстИзображениеВидео

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

от google3 сент. 2026 г.контекст 1,05 млн658 ₩ / 3 291 ₩ · 1M
Anthropic: Claude Fable 5.1ТекстИзображениеФайл

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

от anthropic2 сент. 2026 г.контекст 1 млн17 550 ₩ / 87 750 ₩ · 1M
Anthropic: Claude Fable 5.1 (batch)ТекстИзображениеФайл

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

от anthropic2 сент. 2026 г.контекст 1 млн8 775 ₩ / 43 875 ₩ · 1M

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

от inception1 сент. 2026 г.контекст 260 тыс.70 ₩ / 263 ₩ · 1M

Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

от ibm-granite1 сент. 2026 г.контекст 131,07 тыс.176 ₩ / 263 ₩ · 1M
10,97 трлн токенов в неделю

Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...

от tencent28 авг. 2026 г.контекст 1,05 млн1 464 ₩ / 4 389 ₩ · 1M

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

от inclusionai28 авг. 2026 г.контекст 262,14 тыс.105 ₩ / 316 ₩ · 1M

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

от inclusionai28 авг. 2026 г.контекст 262,14 тыс.Бесплатно / Бесплатно · 1M
Z.ai: GLM Flash LatestТекстИзображениеВидео

This model always redirects to the latest model in the GLM Flash family.

от ~z-ai27 авг. 2026 г.контекст 1,31 млн132 ₩ / 439 ₩ · 1M
Qwen: Qwen3.8 FlashТекстИзображениеВидео

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

от qwen27 авг. 2026 г.контекст 1 млн263 ₩ / 825 ₩ · 1M
Z.ai: GLM 5.3 FlashТекстИзображениеВидео
11,95 трлн токенов в неделю

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

от z-ai26 авг. 2026 г.контекст 1,31 млн132 ₩ / 439 ₩ · 1M
Z.ai: GLM 5.3 Flash (batch)ТекстИзображениеВидео

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

от z-ai26 авг. 2026 г.контекст 1,05 млн263 ₩ / 878 ₩ · 1M
Meta: Muse Spark 1.2 ContributorТекстИзображениеВидео

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

от meta22 авг. 2026 г.контекст 1,05 млн176 ₩ / 351 ₩ · 1M
DeepSeek: DeepSeek V4 Flash Vision ExpТекстИзображение

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

от deepseek21 авг. 2026 г.контекст 1,05 млн386 ₩ / 1 158 ₩ · 1M

Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided...

от tencent20 авг. 2026 г.контекст 8 тыс.77 ₩ / 311 ₩ · 1M

Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and...

от tencent20 авг. 2026 г.контекст 8 тыс.130 ₩ / 518 ₩ · 1M

This model always redirects to the latest GLM model from Z.ai.

от ~z-ai19 авг. 2026 г.контекст 1,31 млн2 053 ₩ / 6 950 ₩ · 1M

Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.

от tencent19 авг. 2026 г.контекст 8 тыс.130 ₩ / 518 ₩ · 1M
Z.ai: GLM 5.3Текст
2,24 трлн токенов в неделю

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

от z-ai19 авг. 2026 г.контекст 1,31 млн2 457 ₩ / 7 722 ₩ · 1M
Qwen: Qwen3.8 27BТекстИзображениеВидео

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

от qwen15 авг. 2026 г.контекст 1 млн737 ₩ / 5 265 ₩ · 1M
Dots Studio: Dots3-Note Preview (free)ТекстИзображение

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...

от dots-studio14 авг. 2026 г.контекст 512 тыс.Бесплатно / Бесплатно · 1M
Google: Gemini 3.7 FlashТекстИзображениеВидео
2,63 трлн токенов в неделю

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

от google14 авг. 2026 г.контекст 1,05 млн1 316 ₩ / 6 581 ₩ · 1M
Google: Gemini 3.7 Flash (batch)ТекстИзображениеВидео

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

от google14 авг. 2026 г.контекст 1,05 млн658 ₩ / 3 291 ₩ · 1M
ByteDance Seed: Seed 2.1 TurboТекстИзображениеВидео

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

от bytedance-seed13 авг. 2026 г.контекст 262,14 тыс.878 ₩ / 4 388 ₩ · 1M

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

от qwen13 авг. 2026 г.контекст 1,05 млн3 510 ₩ / 10 530 ₩ · 1M

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

от qwen13 авг. 2026 г.контекст 1,01 млн3 510 ₩ / 10 530 ₩ · 1M
ByteDance Seed: Seed-2.0-CodeТекстИзображениеВидео

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...

от bytedance-seed13 авг. 2026 г.контекст 262,14 тыс.878 ₩ / 5 265 ₩ · 1M
1,08 трлн токенов в неделю

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

от deepseek13 авг. 2026 г.контекст 1,05 млн1 967 ₩ / 5 900 ₩ · 1M

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

от deepseek13 авг. 2026 г.контекст 1,05 млн2 317 ₩ / 6 950 ₩ · 1M
SpaceXAI: Grok 4.6ТекстИзображениеФайл

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

от x-ai13 авг. 2026 г.контекст 500 тыс.3 510 ₩ / 10 530 ₩ · 1M

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

от liquid12 авг. 2026 г.контекст 65,54 тыс.Бесплатно / Бесплатно · 1M

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

от nvidia11 авг. 2026 г.контекст 262,14 тыс.140 ₩ / 351 ₩ · 1M

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

от nvidia11 авг. 2026 г.контекст 1 млнБесплатно / Бесплатно · 1M
Sakana: Sakana NamazuТекстИзображениеФайл

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

от sakana11 авг. 2026 г.контекст 262,14 тыс.1 667 ₩ / 7 020 ₩ · 1M

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity, document-intensive...

от upstage10 авг. 2026 г.контекст 524,29 тыс.53 ₩ / 211 ₩ · 1M
Meta: Muse Glimmer 30BТекстИзображение

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

от meta10 авг. 2026 г.контекст 131,07 тыс.527 ₩ / 1 931 ₩ · 1M
Meta: Muse Glimmer 30B (batch)ТекстИзображение

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

от meta10 авг. 2026 г.контекст 131,07 тыс.614 ₩ / 2 633 ₩ · 1M
Meta: Muse Spark 1.2ТекстИзображениеВидео

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...

от meta6 авг. 2026 г.контекст 1,05 млн2 194 ₩ / 7 459 ₩ · 1M

This model always redirects to the latest model in the DeepSeek V4 Flash family.

от ~deepseek2 авг. 2026 г.контекст 1,31 млн88 ₩ / 175 ₩ · 1M
11,29 трлн токенов в неделю

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

от deepseek31 июл. 2026 г.контекст 1,31 млн114 ₩ / 316 ₩ · 1M

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

от deepseek31 июл. 2026 г.контекст 1,05 млн246 ₩ / 491 ₩ · 1M
Thinking Machines: Inkling SmallТекстИзображениеАудио

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

от thinkingmachines31 июл. 2026 г.контекст 1,05 млн790 ₩ / 2 106 ₩ · 1M
Thinking Machines: Inkling Small (batch)ТекстИзображениеАудио

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

от thinkingmachines31 июл. 2026 г.контекст 524,29 тыс.878 ₩ / 2 106 ₩ · 1M
Thinking Machines: Inkling Small (free)ТекстИзображениеАудио

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

от thinkingmachines31 июл. 2026 г.контекст 1,05 млнБесплатно / Бесплатно · 1M
Qwen: Qwen3.7 FlashТекстИзображениеВидео

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

от qwen28 июл. 2026 г.контекст 1 млн53 ₩ / 228 ₩ · 1M
Claude Opus 5ТекстИзображениеФайл
1,43 трлн токенов в неделю

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

от anthropic25 июл. 2026 г.контекст 1 млн8 775 ₩ / 43 875 ₩ · 1M
Claude Opus 5 (batch)ТекстИзображениеФайл

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

от anthropic25 июл. 2026 г.контекст 1 млн4 388 ₩ / 21 938 ₩ · 1M

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

от inclusionai23 июл. 2026 г.контекст 262,14 тыс.37 ₩ / 111 ₩ · 1M
1,36 трлн токенов в неделю

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

от poolside22 июл. 2026 г.контекст 1,05 млн158 ₩ / 316 ₩ · 1M

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

от poolside22 июл. 2026 г.контекст 262,14 тыс.Бесплатно / Бесплатно · 1M
Google: Gemini 3.6 FlashТекстИзображениеВидео

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

от google22 июл. 2026 г.контекст 1,05 млн1 316 ₩ / 6 581 ₩ · 1M
Google: Gemini 3.6 Flash (batch)ТекстИзображениеВидео

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

от google22 июл. 2026 г.контекст 1,05 млн658 ₩ / 3 291 ₩ · 1M
Google: Gemini 3.5 Flash LiteТекстИзображениеВидео

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

от google22 июл. 2026 г.контекст 1,05 млн527 ₩ / 4 388 ₩ · 1M
Google: Gemini 3.5 Flash Lite (batch)ТекстИзображениеВидео

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

от google22 июл. 2026 г.контекст 1,05 млн263 ₩ / 2 194 ₩ · 1M

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...

от meituan20 июл. 2026 г.контекст 1,05 млн527 ₩ / 2 106 ₩ · 1M
Thinking Machines: InklingТекстИзображениеАудио

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

от thinkingmachines18 июл. 2026 г.контекст 1,05 млн1 755 ₩ / 7 108 ₩ · 1M
Thinking Machines: Inkling (batch)ТекстИзображениеАудио

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

от thinkingmachines18 июл. 2026 г.контекст 524,29 тыс.1 755 ₩ / 7 108 ₩ · 1M
Thinking Machines: Inkling (free)ТекстИзображениеАудио

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

от thinkingmachines18 июл. 2026 г.контекст 1,05 млнБесплатно / Бесплатно · 1M
Auto Router (Beta)ТекстИзображениеАудио

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

от openrouter18 июл. 2026 г.контекст 2 млнБесплатно / Бесплатно · 1M
MoonshotAI: Kimi K3ТекстИзображениеВидео
1,81 трлн токенов в неделю

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

от moonshotai17 июл. 2026 г.контекст 1,05 млн5 265 ₩ / 26 325 ₩ · 1M
MoonshotAI: Kimi K3 (batch)ТекстИзображениеВидео

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

от moonshotai17 июл. 2026 г.контекст 1,05 млн5 265 ₩ / 26 325 ₩ · 1M
Meta: Muse Spark 1.1ТекстИзображениеВидео

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...

от meta17 июл. 2026 г.контекст 1,05 млн2 194 ₩ / 7 459 ₩ · 1M

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

от kwaipilot11 июл. 2026 г.контекст 262,14 тыс.1 299 ₩ / 5 195 ₩ · 1M
OpenAI: GPT-5.6 Luna ProФайлИзображениеТекст

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

от openai9 июл. 2026 г.контекст 1,05 млн351 ₩ / 2 106 ₩ · 1M
OpenAI: GPT-5.6 Luna Pro (batch)ФайлИзображениеТекст

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

от openai9 июл. 2026 г.контекст 1,05 млн176 ₩ / 1 053 ₩ · 1M
OpenAI: GPT-5.6 LunaФайлИзображениеТекст
11,6 трлн токенов в неделю

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

от openai9 июл. 2026 г.контекст 1,05 млн351 ₩ / 2 106 ₩ · 1M
OpenAI: GPT-5.6 Luna (batch)ФайлИзображениеТекст

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

от openai9 июл. 2026 г.контекст 1,05 млн176 ₩ / 1 053 ₩ · 1M
OpenAI: GPT-5.6 Terra ProФайлИзображениеТекст

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

от openai9 июл. 2026 г.контекст 1,05 млн3 510 ₩ / 21 060 ₩ · 1M
OpenAI: GPT-5.6 Terra Pro (batch)ФайлИзображениеТекст

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

от openai9 июл. 2026 г.контекст 1,05 млн1 755 ₩ / 10 530 ₩ · 1M
OpenAI: GPT-5.6 TerraФайлИзображениеТекст

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

от openai9 июл. 2026 г.контекст 1,05 млн3 510 ₩ / 21 060 ₩ · 1M
OpenAI: GPT-5.6 Terra (batch)ФайлИзображениеТекст

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

от openai9 июл. 2026 г.контекст 1,05 млн1 755 ₩ / 10 530 ₩ · 1M
OpenAI: GPT-5.6 Sol ProФайлИзображениеТекст

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

от openai9 июл. 2026 г.контекст 1,05 млн3 510 ₩ / 17 550 ₩ · 1M
OpenAI: GPT-5.6 Sol Pro (batch)ФайлИзображениеТекст

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

от openai9 июл. 2026 г.контекст 1,05 млн1 755 ₩ / 8 775 ₩ · 1M
OpenAI: GPT-5.6 SolФайлИзображениеТекст
1,81 трлн токенов в неделю

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

от openai9 июл. 2026 г.контекст 1,05 млн3 510 ₩ / 17 550 ₩ · 1M
OpenAI: GPT-5.6 Sol (batch)ФайлИзображениеТекст

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

от openai9 июл. 2026 г.контекст 1,05 млн1 755 ₩ / 8 775 ₩ · 1M
SpaceXAI: Grok 4.5ТекстИзображениеФайл

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

от x-ai9 июл. 2026 г.контекст 500 тыс.3 510 ₩ / 10 530 ₩ · 1M
xAI: Grok LatestТекстИзображениеФайл

This model always redirects to the latest Grok model from xAI.

от ~x-ai8 июл. 2026 г.контекст 500 тыс.3 510 ₩ / 10 530 ₩ · 1M

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...

от aion-labs8 июл. 2026 г.контекст 131,07 тыс.1 229 ₩ / 2 457 ₩ · 1M

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...

от aion-labs8 июл. 2026 г.контекст 131,07 тыс.5 265 ₩ / 10 530 ₩ · 1M
Tencent: Hy3Текст
5,25 трлн токенов в неделю

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

от tencent6 июл. 2026 г.контекст 262,14 тыс.232 ₩ / 927 ₩ · 1M

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

от poolside2 июл. 2026 г.контекст 262,14 тыс.105 ₩ / 211 ₩ · 1M

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

от poolside2 июл. 2026 г.контекст 262,14 тыс.Бесплатно / Бесплатно · 1M
Anthropic: Claude Sonnet 5ТекстИзображениеФайл
1,27 трлн токенов в неделю

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

от anthropic1 июл. 2026 г.контекст 1 млн3 510 ₩ / 17 550 ₩ · 1M
Anthropic: Claude Sonnet 5 (batch)ТекстИзображениеФайл

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

от anthropic1 июл. 2026 г.контекст 1 млн1 755 ₩ / 8 775 ₩ · 1M

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation...

от google1 июл. 2026 г.контекст 65,54 тыс.439 ₩ / 2 633 ₩ · 1M
Nex AGI: Nex-N2-MiniТекстИзображение

Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series. It accepts text and image input and is built for coding, tool use,...

от nex-agi24 июн. 2026 г.контекст 262,14 тыс.44 ₩ / 176 ₩ · 1M
Sakana: Fugu UltraТекстИзображение

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

от sakana24 июн. 2026 г.контекст 1 млн8 775 ₩ / 52 650 ₩ · 1M
Google: Nano Banana 2 (Gemini 3.1 Flash Image)ИзображениеТекст

Gemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines advanced...

от google18 июн. 2026 г.контекст 131,07 тыс.878 ₩ / 5 265 ₩ · 1M
Google: Nano Banana Pro (Gemini 3 Pro Image)ИзображениеТекст

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...

от google18 июн. 2026 г.контекст 131,07 тыс.3 510 ₩ / 21 060 ₩ · 1M

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...

от cohere18 июн. 2026 г.контекст 256 тыс.Бесплатно / Бесплатно · 1M
Z.ai: GLM 5.2Текст
2,19 трлн токенов в неделю

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

от z-ai17 июн. 2026 г.контекст 1,05 млн1 695 ₩ / 5 328 ₩ · 1M

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

от z-ai17 июн. 2026 г.контекст 256 тыс.Бесплатно / Бесплатно · 1M

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...

от openrouter14 июн. 2026 г.контекст 1 млнБесплатно / Бесплатно · 1M
MoonshotAI: Kimi K2.7 CodeТекстИзображение

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

от moonshotai12 июн. 2026 г.контекст 262,14 тыс.1 158 ₩ / 5 967 ₩ · 1M
Anthropic: Claude Fable LatestТекстИзображениеФайл

This model always redirects to the latest model in the Claude Fable family.

от ~anthropic10 июн. 2026 г.контекст 1 млн17 550 ₩ / 87 750 ₩ · 1M
Anthropic: Claude Fable 5ТекстИзображениеФайл

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

от anthropic9 июн. 2026 г.контекст 1 млн17 550 ₩ / 87 750 ₩ · 1M
Anthropic: Claude Fable 5 (batch)ТекстИзображениеФайл

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

от anthropic9 июн. 2026 г.контекст 1 млн8 775 ₩ / 43 875 ₩ · 1M
Nex AGI: Nex-N2-ProТекстИзображение

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...

от nex-agi9 июн. 2026 г.контекст 262,14 тыс.439 ₩ / 1 755 ₩ · 1M
NVIDIA: Nemotron 3.5 Content SafetyТекстИзображение

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

от nvidia4 июн. 2026 г.контекст 131,07 тыс.351 ₩ / 351 ₩ · 1M
NVIDIA: Nemotron 3.5 Content Safety (free)ТекстИзображение

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

от nvidia4 июн. 2026 г.контекст 128 тыс.Бесплатно / Бесплатно · 1M
4,1 трлн токенов в неделю

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

от nvidia4 июн. 2026 г.контекст 262,14 тыс.1 097 ₩ / 5 484 ₩ · 1M

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

от nvidia4 июн. 2026 г.контекст 1 млнБесплатно / Бесплатно · 1M
Qwen: Qwen3.7 PlusТекстИзображение

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

от qwen3 июн. 2026 г.контекст 1 млн562 ₩ / 2 246 ₩ · 1M
MiniMax: MiniMax M3ТекстИзображениеВидео
6,6 трлн токенов в неделю

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

от minimax1 июн. 2026 г.контекст 1,05 млн527 ₩ / 2 106 ₩ · 1M
MiniMax: MiniMax M3 (batch)ТекстИзображениеВидео

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

от minimax1 июн. 2026 г.контекст 524,29 тыс.527 ₩ / 2 106 ₩ · 1M
MiniMax: MiniMax M3 (free)ТекстИзображениеВидео

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

от minimax1 июн. 2026 г.контекст 1,05 млнБесплатно / Бесплатно · 1M
StepFun: Step 3.7 FlashТекстИзображениеВидео

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

от stepfun29 мая 2026 г.контекст 262,14 тыс.351 ₩ / 2 018 ₩ · 1M
Anthropic: Claude Opus 4.8ТекстИзображениеФайл

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

от anthropic28 мая 2026 г.контекст 1 млн8 775 ₩ / 43 875 ₩ · 1M
Anthropic: Claude Opus 4.8 (batch)ТекстИзображениеФайл

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

от anthropic28 мая 2026 г.контекст 1 млн4 388 ₩ / 21 938 ₩ · 1M

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

от qwen22 мая 2026 г.контекст 1 млн2 589 ₩ / 7 766 ₩ · 1M
SpaceXAI: Grok Build 0.1ТекстИзображениеФайл

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

от x-ai21 мая 2026 г.контекст 256 тыс.1 755 ₩ / 3 510 ₩ · 1M
Google: Gemini 3.5 FlashТекстИзображениеВидео

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

от google19 мая 2026 г.контекст 1,05 млн2 633 ₩ / 15 795 ₩ · 1M
Google: Gemini 3.5 Flash (batch)ТекстИзображениеВидео

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

от google19 мая 2026 г.контекст 1,05 млн1 316 ₩ / 7 898 ₩ · 1M
Perceptron: Perceptron Mk1ТекстИзображениеВидео

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

от perceptron12 мая 2026 г.контекст 32,77 тыс.263 ₩ / 2 633 ₩ · 1M
Google: Gemini 3.1 Flash LiteТекстИзображениеВидео

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

от google8 мая 2026 г.контекст 1,05 млн439 ₩ / 2 633 ₩ · 1M
Google: Gemini 3.1 Flash Lite (batch)ТекстИзображениеВидео

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

от google8 мая 2026 г.контекст 1,05 млн219 ₩ / 1 316 ₩ · 1M
OpenAI: GPT Chat LatestТекстИзображениеФайл

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

от openai6 мая 2026 г.контекст 400 тыс.8 775 ₩ / 52 650 ₩ · 1M
SpaceXAI: Grok 4.3ТекстИзображениеФайл

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

от x-ai1 мая 2026 г.контекст 1 млн2 194 ₩ / 4 388 ₩ · 1M
SpaceXAI: Grok 4.3 (batch)ТекстИзображениеФайл

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

от x-ai1 мая 2026 г.контекст 1 млн1 755 ₩ / 3 510 ₩ · 1M
Mistral: Mistral Medium 3.5ТекстИзображениеФайл

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

от mistralai1 мая 2026 г.контекст 262,14 тыс.2 633 ₩ / 13 163 ₩ · 1M
Mistral: Mistral Medium 3.5 (batch)ТекстИзображениеФайл

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

от mistralai1 мая 2026 г.контекст 32,77 тыс.1 316 ₩ / 6 581 ₩ · 1M
NVIDIA: Nemotron 3 Nano Omni (free)ТекстАудиоИзображение

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

от nvidia29 апр. 2026 г.контекст 256 тыс.Бесплатно / Бесплатно · 1M
Anthropic Claude Haiku LatestТекстИзображениеФайл

This model always redirects to the latest model in the Anthropic Claude Haiku family.

от ~anthropic28 апр. 2026 г.контекст 200 тыс.1 755 ₩ / 8 775 ₩ · 1M
OpenAI GPT Mini LatestФайлИзображениеТекст

This model always redirects to the latest model in the OpenAI GPT Mini family.

от ~openai28 апр. 2026 г.контекст 400 тыс.1 316 ₩ / 7 898 ₩ · 1M
Google Gemini Pro LatestАудиоФайлИзображение

This model always redirects to the latest model in the Google Gemini Pro family.

от ~google28 апр. 2026 г.контекст 1,05 млн3 510 ₩ / 21 060 ₩ · 1M
MoonshotAI Kimi LatestТекстИзображениеВидео

This model always redirects to the latest model in the MoonshotAI Kimi family.

от ~moonshotai28 апр. 2026 г.контекст 1,05 млн4 475 ₩ / 22 376 ₩ · 1M
Google Gemini Flash LatestТекстИзображениеВидео

This model always redirects to the latest model in the Google Gemini Flash family.

от ~google28 апр. 2026 г.контекст 1,05 млн1 316 ₩ / 6 581 ₩ · 1M
Anthropic Claude Sonnet LatestТекстИзображениеФайл

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

от ~anthropic28 апр. 2026 г.контекст 1 млн3 510 ₩ / 17 550 ₩ · 1M
OpenAI GPT LatestФайлИзображениеТекст

This model always redirects to the latest model in the OpenAI GPT family.

от ~openai28 апр. 2026 г.контекст 1,05 млн3 510 ₩ / 17 550 ₩ · 1M
Qwen: Qwen3.5 Plus 2026-04-20ТекстИзображениеВидео

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

от qwen27 апр. 2026 г.контекст 1 млн527 ₩ / 3 159 ₩ · 1M
Qwen: Qwen3.6 FlashТекстИзображениеВидео

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

от qwen27 апр. 2026 г.контекст 1 млн329 ₩ / 1 974 ₩ · 1M
Qwen: Qwen3.6 35B A3BТекстИзображениеВидео

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

от qwen27 апр. 2026 г.контекст 262,14 тыс.176 ₩ / 1 580 ₩ · 1M

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

от qwen27 апр. 2026 г.контекст 262,14 тыс.1 802 ₩ / 10 814 ₩ · 1M
Qwen: Qwen3.6 27BТекстИзображениеВидео

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

от qwen27 апр. 2026 г.контекст 262,14 тыс.527 ₩ / 3 510 ₩ · 1M
OpenAI: GPT-5.5 ProФайлИзображениеТекст

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

от openai25 апр. 2026 г.контекст 1,05 млн52 650 ₩ / 315 900 ₩ · 1M
OpenAI: GPT-5.5 Pro (batch)ФайлИзображениеТекст

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

от openai25 апр. 2026 г.контекст 1,05 млн26 325 ₩ / 157 950 ₩ · 1M
OpenAI: GPT-5.5ФайлИзображениеТекст

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

от openai25 апр. 2026 г.контекст 1,05 млн8 775 ₩ / 52 650 ₩ · 1M
OpenAI: GPT-5.5 (batch)ФайлИзображениеТекст

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

от openai25 апр. 2026 г.контекст 1,05 млн4 388 ₩ / 26 325 ₩ · 1M
1,49 трлн токенов в неделю

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

от deepseek24 апр. 2026 г.контекст 1,05 млн1 331 ₩ / 2 663 ₩ · 1M
5,18 трлн токенов в неделю

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

от deepseek24 апр. 2026 г.контекст 1,05 млн144 ₩ / 287 ₩ · 1M

Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use. It supports configurable reasoning levels across disabled, low, and high modes, allowing it to...

от tencent23 апр. 2026 г.контекст 262,14 тыс.316 ₩ / 1 053 ₩ · 1M

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

от xiaomi23 апр. 2026 г.контекст 1,05 млн763 ₩ / 1 527 ₩ · 1M
Xiaomi: MiMo-V2.5ТекстАудиоИзображение
4,5 трлн токенов в неделю

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

от xiaomi23 апр. 2026 г.контекст 1,05 млн246 ₩ / 491 ₩ · 1M
OpenAI: GPT-5.4 Image 2ИзображениеТекстФайл

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...

от openai22 апр. 2026 г.контекст 272 тыс.14 040 ₩ / 26 325 ₩ · 1M
Anthropic: Claude Opus LatestТекстИзображениеФайл

This model always redirects to the latest model in the Claude Opus family.

от ~anthropic22 апр. 2026 г.контекст 1 млн8 775 ₩ / 43 875 ₩ · 1M

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...

от openrouter21 апр. 2026 г.контекст 2 млнБесплатно / Бесплатно · 1M
MoonshotAI: Kimi K2.6ТекстИзображение

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

от moonshotai21 апр. 2026 г.контекст 262,14 тыс.1 667 ₩ / 7 020 ₩ · 1M
Anthropic: Claude Opus 4.7ТекстИзображениеФайл

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

от anthropic16 апр. 2026 г.контекст 1 млн8 775 ₩ / 43 875 ₩ · 1M
Anthropic: Claude Opus 4.7 (batch)ТекстИзображениеФайл

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

от anthropic16 апр. 2026 г.контекст 1 млн4 388 ₩ / 21 938 ₩ · 1M
Z.ai: GLM 5.1Текст

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

от z-ai8 апр. 2026 г.контекст 204,8 тыс.1 695 ₩ / 5 328 ₩ · 1M
Google: Gemma 4 26B A4B ИзображениеТекстВидео

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

от google3 апр. 2026 г.контекст 262,14 тыс.123 ₩ / 597 ₩ · 1M
Google: Gemma 4 26B A4B (free)ИзображениеТекстВидео

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

от google3 апр. 2026 г.контекст 262,14 тыс.Бесплатно / Бесплатно · 1M
Google: Gemma 4 31BИзображениеТекстВидео

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

от google3 апр. 2026 г.контекст 262,14 тыс.158 ₩ / 597 ₩ · 1M
Google: Gemma 4 31B (batch)ИзображениеТекстВидео

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

от google3 апр. 2026 г.контекст 262,14 тыс.684 ₩ / 1 702 ₩ · 1M
Google: Gemma 4 31B (free)ИзображениеТекстВидео

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

от google3 апр. 2026 г.контекст 262,14 тыс.Бесплатно / Бесплатно · 1M
Qwen: Qwen3.6 PlusТекстИзображениеВидео

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

от qwen2 апр. 2026 г.контекст 1 млн570 ₩ / 3 422 ₩ · 1M
Z.ai: GLM 5V TurboИзображениеТекстВидео

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

от z-ai2 апр. 2026 г.контекст 202,75 тыс.2 106 ₩ / 7 020 ₩ · 1M

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

от arcee-ai2 апр. 2026 г.контекст 262,14 тыс.439 ₩ / 1 404 ₩ · 1M
SpaceXAI: Grok 4.20 Multi-AgentТекстИзображениеФайл

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

от x-ai1 апр. 2026 г.контекст 2 млн2 194 ₩ / 4 388 ₩ · 1M
SpaceXAI: Grok 4.20ТекстИзображениеФайл

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

от x-ai1 апр. 2026 г.контекст 2 млн2 194 ₩ / 4 388 ₩ · 1M
Google: Lyria 3 Pro PreviewТекстИзображение

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

от google31 мар. 2026 г.контекст 1,05 млнБесплатно / Бесплатно · 1M
Google: Lyria 3 Clip PreviewТекстИзображение

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...

от google31 мар. 2026 г.контекст 1,05 млнБесплатно / Бесплатно · 1M

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...

от kwaipilot28 мар. 2026 г.контекст 262,14 тыс.527 ₩ / 2 106 ₩ · 1M
Reka EdgeИзображениеТекстВидео

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...

от rekaai21 мар. 2026 г.контекст 16,38 тыс.176 ₩ / 176 ₩ · 1M

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

от minimax18 мар. 2026 г.контекст 204,8 тыс.527 ₩ / 2 106 ₩ · 1M

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

от minimax18 мар. 2026 г.контекст 196,61 тыс.Бесплатно / Бесплатно · 1M
OpenAI: GPT-5.4 NanoФайлИзображениеТекст

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

от openai17 мар. 2026 г.контекст 400 тыс.351 ₩ / 2 194 ₩ · 1M
OpenAI: GPT-5.4 Nano (batch)ФайлИзображениеТекст

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

от openai17 мар. 2026 г.контекст 400 тыс.176 ₩ / 1 097 ₩ · 1M
OpenAI: GPT-5.4 MiniФайлИзображениеТекст

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

от openai17 мар. 2026 г.контекст 400 тыс.1 316 ₩ / 7 898 ₩ · 1M
OpenAI: GPT-5.4 Mini (batch)ФайлИзображениеТекст

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

от openai17 мар. 2026 г.контекст 400 тыс.658 ₩ / 3 949 ₩ · 1M
Mistral: Mistral Small 4ТекстИзображение

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

от mistralai17 мар. 2026 г.контекст 262,14 тыс.263 ₩ / 1 053 ₩ · 1M

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

от z-ai15 мар. 2026 г.контекст 202,75 тыс.2 106 ₩ / 7 020 ₩ · 1M

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

от nvidia12 мар. 2026 г.контекст 1 млн149 ₩ / 702 ₩ · 1M

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

от nvidia12 мар. 2026 г.контекст 262,14 тыс.Бесплатно / Бесплатно · 1M
ByteDance Seed: Seed-2.0-LiteТекстИзображениеВидео

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...

от bytedance-seed11 мар. 2026 г.контекст 262,14 тыс.439 ₩ / 3 510 ₩ · 1M
Qwen: Qwen3.5-9BТекстИзображениеВидео

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

от qwen10 мар. 2026 г.контекст 262,14 тыс.176 ₩ / 263 ₩ · 1M
Qwen: Qwen3.5-9B (batch)ТекстИзображениеВидео

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

от qwen10 мар. 2026 г.контекст 262,14 тыс.298 ₩ / 439 ₩ · 1M
OpenAI: GPT-5.4 ProТекстИзображениеФайл

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

от openai6 мар. 2026 г.контекст 1,05 млн52 650 ₩ / 315 900 ₩ · 1M
OpenAI: GPT-5.4 Pro (batch)ТекстИзображениеФайл

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

от openai6 мар. 2026 г.контекст 1,05 млн26 325 ₩ / 157 950 ₩ · 1M
OpenAI: GPT-5.4ТекстИзображениеФайл

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

от openai6 мар. 2026 г.контекст 1,05 млн4 388 ₩ / 26 325 ₩ · 1M
OpenAI: GPT-5.4 (batch)ТекстИзображениеФайл

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

от openai6 мар. 2026 г.контекст 1,05 млн2 194 ₩ / 13 163 ₩ · 1M

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

от inception4 мар. 2026 г.контекст 128 тыс.439 ₩ / 1 316 ₩ · 1M
Google: Gemini 3.1 Flash Lite PreviewТекстИзображениеВидео

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

от google3 мар. 2026 г.контекст 1,05 млн439 ₩ / 2 633 ₩ · 1M
ByteDance Seed: Seed-2.0-MiniТекстИзображениеВидео

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

от bytedance-seed27 февр. 2026 г.контекст 262,14 тыс.176 ₩ / 702 ₩ · 1M

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

от google27 февр. 2026 г.контекст 65,54 тыс.878 ₩ / 5 265 ₩ · 1M
Qwen: Qwen3.5-35B-A3BТекстИзображениеВидео

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

от qwen26 февр. 2026 г.контекст 262,14 тыс.548 ₩ / 2 194 ₩ · 1M
Qwen: Qwen3.5-27BТекстИзображениеВидео

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

от qwen26 февр. 2026 г.контекст 262,14 тыс.342 ₩ / 2 738 ₩ · 1M
Qwen: Qwen3.5-122B-A10BТекстИзображениеВидео

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

от qwen26 февр. 2026 г.контекст 262,14 тыс.509 ₩ / 4 212 ₩ · 1M
Qwen: Qwen3.5-FlashТекстИзображениеВидео

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

от qwen26 февр. 2026 г.контекст 1 млн114 ₩ / 456 ₩ · 1M
Google: Gemini 3.1 Pro Preview Custom ToolsТекстАудиоИзображение

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

от google26 февр. 2026 г.контекст 1,05 млн3 510 ₩ / 21 060 ₩ · 1M
OpenAI: GPT-5.3-CodexТекстИзображениеФайл

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...

от openai25 февр. 2026 г.контекст 400 тыс.3 071 ₩ / 24 570 ₩ · 1M