Modelos

433 modelos

OpenAI: GPT-6 AstraArchivoImagenTexto

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

por openai5 sept 20261,05 M de contexto17.550 KRW / 87.750 KRW · 1M
OpenAI: GPT-6 Astra (batch)ArchivoImagenTexto

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

por openai5 sept 20261,05 M de contexto8775 KRW / 43.875 KRW · 1M
OpenAI: GPT-6 Astra ProArchivoImagenTexto

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

por openai5 sept 20261,05 M de contexto17.550 KRW / 87.750 KRW · 1M

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

por openai5 sept 20261,05 M de contexto8775 KRW / 43.875 KRW · 1M

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

por inclusionai5 sept 2026262,14 mil de contextoGratis / Gratis · 1M
Qwen: Qwen3.8 Max (0902)TextoImagenVideo

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

por qwen4 sept 20261 M de contexto3510 KRW / 10.530 KRW · 1M

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...

por meta3 sept 20261,05 M de contexto176 KRW / 351 KRW · 1M
Meta: Muse Spark 1.3TextoImagenVideo

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...

por meta3 sept 20261,05 M de contexto2194 KRW / 7459 KRW · 1M
Google: Gemini 3.8 FlashTextoImagenVideo

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

por google3 sept 20261,05 M de contexto1316 KRW / 6581 KRW · 1M

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

por google3 sept 20261,05 M de contexto658 KRW / 3291 KRW · 1M
Anthropic: Claude Fable 5.1TextoImagenArchivo

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

por anthropic2 sept 20261 M de contexto17.550 KRW / 87.750 KRW · 1M

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

por anthropic2 sept 20261 M de contexto8775 KRW / 43.875 KRW · 1M

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

por inception1 sept 2026260 mil de contexto70 KRW / 263 KRW · 1M

Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

por ibm-granite1 sept 2026131,07 mil de contexto176 KRW / 263 KRW · 1M
10,97 B tokens semanales

Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...

por tencent28 ago 20261,05 M de contexto1464 KRW / 4389 KRW · 1M

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

por inclusionai28 ago 2026262,14 mil de contexto105 KRW / 316 KRW · 1M

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

por inclusionai28 ago 2026262,14 mil de contextoGratis / Gratis · 1M
Z.ai: GLM Flash LatestTextoImagenVideo

This model always redirects to the latest model in the GLM Flash family.

por ~z-ai27 ago 20261,31 M de contexto132 KRW / 439 KRW · 1M
Qwen: Qwen3.8 FlashTextoImagenVideo

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

por qwen27 ago 20261 M de contexto263 KRW / 825 KRW · 1M
Z.ai: GLM 5.3 FlashTextoImagenVideo
11,95 B tokens semanales

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

por z-ai26 ago 20261,31 M de contexto132 KRW / 439 KRW · 1M

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

por z-ai26 ago 20261,05 M de contexto263 KRW / 878 KRW · 1M

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

por meta22 ago 20261,05 M de contexto176 KRW / 351 KRW · 1M

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

por deepseek21 ago 20261,05 M de contexto386 KRW / 1158 KRW · 1M

Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided...

por tencent20 ago 20268 mil de contexto77 KRW / 311 KRW · 1M

Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and...

por tencent20 ago 20268 mil de contexto130 KRW / 518 KRW · 1M

This model always redirects to the latest GLM model from Z.ai.

por ~z-ai19 ago 20261,31 M de contexto2053 KRW / 6950 KRW · 1M

Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.

por tencent19 ago 20268 mil de contexto130 KRW / 518 KRW · 1M
2,24 B tokens semanales

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

por z-ai19 ago 20261,31 M de contexto2457 KRW / 7722 KRW · 1M
Qwen: Qwen3.8 27BTextoImagenVideo

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

por qwen15 ago 20261 M de contexto737 KRW / 5265 KRW · 1M

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...

por dots-studio14 ago 2026512 mil de contextoGratis / Gratis · 1M
Google: Gemini 3.7 FlashTextoImagenVideo
2,63 B tokens semanales

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

por google14 ago 20261,05 M de contexto1316 KRW / 6581 KRW · 1M

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

por google14 ago 20261,05 M de contexto658 KRW / 3291 KRW · 1M

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

por bytedance-seed13 ago 2026262,14 mil de contexto878 KRW / 4388 KRW · 1M

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

por qwen13 ago 20261,05 M de contexto3510 KRW / 10.530 KRW · 1M

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

por qwen13 ago 20261,01 M de contexto3510 KRW / 10.530 KRW · 1M

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...

por bytedance-seed13 ago 2026262,14 mil de contexto878 KRW / 5265 KRW · 1M
1,08 B tokens semanales

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

por deepseek13 ago 20261,05 M de contexto1967 KRW / 5900 KRW · 1M

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

por deepseek13 ago 20261,05 M de contexto2317 KRW / 6950 KRW · 1M
SpaceXAI: Grok 4.6TextoImagenArchivo

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

por x-ai13 ago 2026500 mil de contexto3510 KRW / 10.530 KRW · 1M

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

por liquid12 ago 202665,54 mil de contextoGratis / Gratis · 1M

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

por nvidia11 ago 2026262,14 mil de contexto140 KRW / 351 KRW · 1M

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

por nvidia11 ago 20261 M de contextoGratis / Gratis · 1M
Sakana: Sakana NamazuTextoImagenArchivo

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

por sakana11 ago 2026262,14 mil de contexto1667 KRW / 7020 KRW · 1M

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity, document-intensive...

por upstage10 ago 2026524,29 mil de contexto53 KRW / 211 KRW · 1M

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

por meta10 ago 2026131,07 mil de contexto527 KRW / 1931 KRW · 1M

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

por meta10 ago 2026131,07 mil de contexto614 KRW / 2633 KRW · 1M
Meta: Muse Spark 1.2TextoImagenVideo

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...

por meta6 ago 20261,05 M de contexto2194 KRW / 7459 KRW · 1M

This model always redirects to the latest model in the DeepSeek V4 Flash family.

por ~deepseek2 ago 20261,31 M de contexto88 KRW / 175 KRW · 1M
11,29 B tokens semanales

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

por deepseek31 jul 20261,31 M de contexto114 KRW / 316 KRW · 1M

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

por deepseek31 jul 20261,05 M de contexto246 KRW / 491 KRW · 1M

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

por thinkingmachines31 jul 20261,05 M de contexto790 KRW / 2106 KRW · 1M

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

por thinkingmachines31 jul 2026524,29 mil de contexto878 KRW / 2106 KRW · 1M

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

por thinkingmachines31 jul 20261,05 M de contextoGratis / Gratis · 1M
Qwen: Qwen3.7 FlashTextoImagenVideo

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

por qwen28 jul 20261 M de contexto53 KRW / 228 KRW · 1M
Claude Opus 5TextoImagenArchivo
1,43 B tokens semanales

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

por anthropic25 jul 20261 M de contexto8775 KRW / 43.875 KRW · 1M
Claude Opus 5 (batch)TextoImagenArchivo

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

por anthropic25 jul 20261 M de contexto4388 KRW / 21.938 KRW · 1M

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

por inclusionai23 jul 2026262,14 mil de contexto37 KRW / 111 KRW · 1M
1,36 B tokens semanales

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

por poolside22 jul 20261,05 M de contexto158 KRW / 316 KRW · 1M

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

por poolside22 jul 2026262,14 mil de contextoGratis / Gratis · 1M
Google: Gemini 3.6 FlashTextoImagenVideo

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

por google22 jul 20261,05 M de contexto1316 KRW / 6581 KRW · 1M

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

por google22 jul 20261,05 M de contexto658 KRW / 3291 KRW · 1M

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

por google22 jul 20261,05 M de contexto527 KRW / 4388 KRW · 1M

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

por google22 jul 20261,05 M de contexto263 KRW / 2194 KRW · 1M

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...

por meituan20 jul 20261,05 M de contexto527 KRW / 2106 KRW · 1M

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

por thinkingmachines18 jul 20261,05 M de contexto1755 KRW / 7108 KRW · 1M

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

por thinkingmachines18 jul 2026524,29 mil de contexto1755 KRW / 7108 KRW · 1M

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

por thinkingmachines18 jul 20261,05 M de contextoGratis / Gratis · 1M
Auto Router (Beta)TextoImagenAudio

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

por openrouter18 jul 20262 M de contextoGratis / Gratis · 1M
MoonshotAI: Kimi K3TextoImagenVideo
1,81 B tokens semanales

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

por moonshotai17 jul 20261,05 M de contexto5265 KRW / 26.325 KRW · 1M

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

por moonshotai17 jul 20261,05 M de contexto5265 KRW / 26.325 KRW · 1M
Meta: Muse Spark 1.1TextoImagenVideo

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...

por meta17 jul 20261,05 M de contexto2194 KRW / 7459 KRW · 1M

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

por kwaipilot11 jul 2026262,14 mil de contexto1299 KRW / 5195 KRW · 1M
OpenAI: GPT-5.6 Luna ProArchivoImagenTexto

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

por openai9 jul 20261,05 M de contexto351 KRW / 2106 KRW · 1M

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

por openai9 jul 20261,05 M de contexto176 KRW / 1053 KRW · 1M
OpenAI: GPT-5.6 LunaArchivoImagenTexto
11,6 B tokens semanales

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

por openai9 jul 20261,05 M de contexto351 KRW / 2106 KRW · 1M

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

por openai9 jul 20261,05 M de contexto176 KRW / 1053 KRW · 1M
OpenAI: GPT-5.6 Terra ProArchivoImagenTexto

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

por openai9 jul 20261,05 M de contexto3510 KRW / 21.060 KRW · 1M

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

por openai9 jul 20261,05 M de contexto1755 KRW / 10.530 KRW · 1M
OpenAI: GPT-5.6 TerraArchivoImagenTexto

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

por openai9 jul 20261,05 M de contexto3510 KRW / 21.060 KRW · 1M

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

por openai9 jul 20261,05 M de contexto1755 KRW / 10.530 KRW · 1M
OpenAI: GPT-5.6 Sol ProArchivoImagenTexto

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

por openai9 jul 20261,05 M de contexto3510 KRW / 17.550 KRW · 1M

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

por openai9 jul 20261,05 M de contexto1755 KRW / 8775 KRW · 1M
OpenAI: GPT-5.6 SolArchivoImagenTexto
1,81 B tokens semanales

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

por openai9 jul 20261,05 M de contexto3510 KRW / 17.550 KRW · 1M
OpenAI: GPT-5.6 Sol (batch)ArchivoImagenTexto

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

por openai9 jul 20261,05 M de contexto1755 KRW / 8775 KRW · 1M
SpaceXAI: Grok 4.5TextoImagenArchivo

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

por x-ai9 jul 2026500 mil de contexto3510 KRW / 10.530 KRW · 1M
xAI: Grok LatestTextoImagenArchivo

This model always redirects to the latest Grok model from xAI.

por ~x-ai8 jul 2026500 mil de contexto3510 KRW / 10.530 KRW · 1M

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...

por aion-labs8 jul 2026131,07 mil de contexto1229 KRW / 2457 KRW · 1M

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...

por aion-labs8 jul 2026131,07 mil de contexto5265 KRW / 10.530 KRW · 1M
5,25 B tokens semanales

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

por tencent6 jul 2026262,14 mil de contexto232 KRW / 927 KRW · 1M

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

por poolside2 jul 2026262,14 mil de contexto105 KRW / 211 KRW · 1M

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

por poolside2 jul 2026262,14 mil de contextoGratis / Gratis · 1M
Anthropic: Claude Sonnet 5TextoImagenArchivo
1,27 B tokens semanales

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

por anthropic1 jul 20261 M de contexto3510 KRW / 17.550 KRW · 1M

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

por anthropic1 jul 20261 M de contexto1755 KRW / 8775 KRW · 1M

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation...

por google1 jul 202665,54 mil de contexto439 KRW / 2633 KRW · 1M

Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series. It accepts text and image input and is built for coding, tool use,...

por nex-agi24 jun 2026262,14 mil de contexto44 KRW / 176 KRW · 1M

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

por sakana24 jun 20261 M de contexto8775 KRW / 52.650 KRW · 1M

Gemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines advanced...

por google18 jun 2026131,07 mil de contexto878 KRW / 5265 KRW · 1M

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...

por google18 jun 2026131,07 mil de contexto3510 KRW / 21.060 KRW · 1M

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...

por cohere18 jun 2026256 mil de contextoGratis / Gratis · 1M
2,19 B tokens semanales

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

por z-ai17 jun 20261,05 M de contexto1695 KRW / 5328 KRW · 1M

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

por z-ai17 jun 2026256 mil de contextoGratis / Gratis · 1M

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...

por openrouter14 jun 20261 M de contextoGratis / Gratis · 1M

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

por moonshotai12 jun 2026262,14 mil de contexto1158 KRW / 5967 KRW · 1M

This model always redirects to the latest model in the Claude Fable family.

por ~anthropic10 jun 20261 M de contexto17.550 KRW / 87.750 KRW · 1M
Anthropic: Claude Fable 5TextoImagenArchivo

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

por anthropic9 jun 20261 M de contexto17.550 KRW / 87.750 KRW · 1M

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

por anthropic9 jun 20261 M de contexto8775 KRW / 43.875 KRW · 1M

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...

por nex-agi9 jun 2026262,14 mil de contexto439 KRW / 1755 KRW · 1M

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

por nvidia4 jun 2026131,07 mil de contexto351 KRW / 351 KRW · 1M

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

por nvidia4 jun 2026128 mil de contextoGratis / Gratis · 1M
4,1 B tokens semanales

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

por nvidia4 jun 2026262,14 mil de contexto1097 KRW / 5484 KRW · 1M

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

por nvidia4 jun 20261 M de contextoGratis / Gratis · 1M

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

por qwen3 jun 20261 M de contexto562 KRW / 2246 KRW · 1M
MiniMax: MiniMax M3TextoImagenVideo
6,6 B tokens semanales

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

por minimax1 jun 20261,05 M de contexto527 KRW / 2106 KRW · 1M

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

por minimax1 jun 2026524,29 mil de contexto527 KRW / 2106 KRW · 1M

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

por minimax1 jun 20261,05 M de contextoGratis / Gratis · 1M
StepFun: Step 3.7 FlashTextoImagenVideo

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

por stepfun29 may 2026262,14 mil de contexto351 KRW / 2018 KRW · 1M
Anthropic: Claude Opus 4.8TextoImagenArchivo

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

por anthropic28 may 20261 M de contexto8775 KRW / 43.875 KRW · 1M

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

por anthropic28 may 20261 M de contexto4388 KRW / 21.938 KRW · 1M

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

por qwen22 may 20261 M de contexto2589 KRW / 7766 KRW · 1M
SpaceXAI: Grok Build 0.1TextoImagenArchivo

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

por x-ai21 may 2026256 mil de contexto1755 KRW / 3510 KRW · 1M
Google: Gemini 3.5 FlashTextoImagenVideo

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

por google19 may 20261,05 M de contexto2633 KRW / 15.795 KRW · 1M

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

por google19 may 20261,05 M de contexto1316 KRW / 7898 KRW · 1M

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

por perceptron12 may 202632,77 mil de contexto263 KRW / 2633 KRW · 1M

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

por google8 may 20261,05 M de contexto439 KRW / 2633 KRW · 1M

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

por google8 may 20261,05 M de contexto219 KRW / 1316 KRW · 1M
OpenAI: GPT Chat LatestTextoImagenArchivo

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

por openai6 may 2026400 mil de contexto8775 KRW / 52.650 KRW · 1M
SpaceXAI: Grok 4.3TextoImagenArchivo

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

por x-ai1 may 20261 M de contexto2194 KRW / 4388 KRW · 1M
SpaceXAI: Grok 4.3 (batch)TextoImagenArchivo

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

por x-ai1 may 20261 M de contexto1755 KRW / 3510 KRW · 1M
Mistral: Mistral Medium 3.5TextoImagenArchivo

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

por mistralai1 may 2026262,14 mil de contexto2633 KRW / 13.163 KRW · 1M

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

por mistralai1 may 202632,77 mil de contexto1316 KRW / 6581 KRW · 1M

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

por nvidia29 abr 2026256 mil de contextoGratis / Gratis · 1M

This model always redirects to the latest model in the Anthropic Claude Haiku family.

por ~anthropic28 abr 2026200 mil de contexto1755 KRW / 8775 KRW · 1M
OpenAI GPT Mini LatestArchivoImagenTexto

This model always redirects to the latest model in the OpenAI GPT Mini family.

por ~openai28 abr 2026400 mil de contexto1316 KRW / 7898 KRW · 1M
Google Gemini Pro LatestAudioArchivoImagen

This model always redirects to the latest model in the Google Gemini Pro family.

por ~google28 abr 20261,05 M de contexto3510 KRW / 21.060 KRW · 1M
MoonshotAI Kimi LatestTextoImagenVideo

This model always redirects to the latest model in the MoonshotAI Kimi family.

por ~moonshotai28 abr 20261,05 M de contexto4475 KRW / 22.376 KRW · 1M

This model always redirects to the latest model in the Google Gemini Flash family.

por ~google28 abr 20261,05 M de contexto1316 KRW / 6581 KRW · 1M

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

por ~anthropic28 abr 20261 M de contexto3510 KRW / 17.550 KRW · 1M
OpenAI GPT LatestArchivoImagenTexto

This model always redirects to the latest model in the OpenAI GPT family.

por ~openai28 abr 20261,05 M de contexto3510 KRW / 17.550 KRW · 1M

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

por qwen27 abr 20261 M de contexto527 KRW / 3159 KRW · 1M
Qwen: Qwen3.6 FlashTextoImagenVideo

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

por qwen27 abr 20261 M de contexto329 KRW / 1974 KRW · 1M
Qwen: Qwen3.6 35B A3BTextoImagenVideo

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

por qwen27 abr 2026262,14 mil de contexto176 KRW / 1580 KRW · 1M

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

por qwen27 abr 2026262,14 mil de contexto1802 KRW / 10.814 KRW · 1M
Qwen: Qwen3.6 27BTextoImagenVideo

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

por qwen27 abr 2026262,14 mil de contexto527 KRW / 3510 KRW · 1M
OpenAI: GPT-5.5 ProArchivoImagenTexto

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

por openai25 abr 20261,05 M de contexto52.650 KRW / 315.900 KRW · 1M
OpenAI: GPT-5.5 Pro (batch)ArchivoImagenTexto

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

por openai25 abr 20261,05 M de contexto26.325 KRW / 157.950 KRW · 1M
OpenAI: GPT-5.5ArchivoImagenTexto

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

por openai25 abr 20261,05 M de contexto8775 KRW / 52.650 KRW · 1M
OpenAI: GPT-5.5 (batch)ArchivoImagenTexto

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

por openai25 abr 20261,05 M de contexto4388 KRW / 26.325 KRW · 1M
1,49 B tokens semanales

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

por deepseek24 abr 20261,05 M de contexto1331 KRW / 2663 KRW · 1M
5,18 B tokens semanales

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

por deepseek24 abr 20261,05 M de contexto144 KRW / 287 KRW · 1M

Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use. It supports configurable reasoning levels across disabled, low, and high modes, allowing it to...

por tencent23 abr 2026262,14 mil de contexto316 KRW / 1053 KRW · 1M

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

por xiaomi23 abr 20261,05 M de contexto763 KRW / 1527 KRW · 1M
Xiaomi: MiMo-V2.5TextoAudioImagen
4,5 B tokens semanales

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

por xiaomi23 abr 20261,05 M de contexto246 KRW / 491 KRW · 1M
OpenAI: GPT-5.4 Image 2ImagenTextoArchivo

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...

por openai22 abr 2026272 mil de contexto14.040 KRW / 26.325 KRW · 1M

This model always redirects to the latest model in the Claude Opus family.

por ~anthropic22 abr 20261 M de contexto8775 KRW / 43.875 KRW · 1M

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...

por openrouter21 abr 20262 M de contextoGratis / Gratis · 1M

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

por moonshotai21 abr 2026262,14 mil de contexto1667 KRW / 7020 KRW · 1M
Anthropic: Claude Opus 4.7TextoImagenArchivo

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

por anthropic16 abr 20261 M de contexto8775 KRW / 43.875 KRW · 1M

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

por anthropic16 abr 20261 M de contexto4388 KRW / 21.938 KRW · 1M

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

por z-ai8 abr 2026204,8 mil de contexto1695 KRW / 5328 KRW · 1M
Google: Gemma 4 26B A4B ImagenTextoVideo

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

por google3 abr 2026262,14 mil de contexto123 KRW / 597 KRW · 1M

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

por google3 abr 2026262,14 mil de contextoGratis / Gratis · 1M
Google: Gemma 4 31BImagenTextoVideo

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

por google3 abr 2026262,14 mil de contexto158 KRW / 597 KRW · 1M

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

por google3 abr 2026262,14 mil de contexto684 KRW / 1702 KRW · 1M

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

por google3 abr 2026262,14 mil de contextoGratis / Gratis · 1M
Qwen: Qwen3.6 PlusTextoImagenVideo

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

por qwen2 abr 20261 M de contexto570 KRW / 3422 KRW · 1M
Z.ai: GLM 5V TurboImagenTextoVideo

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

por z-ai2 abr 2026202,75 mil de contexto2106 KRW / 7020 KRW · 1M

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

por arcee-ai2 abr 2026262,14 mil de contexto439 KRW / 1404 KRW · 1M

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

por x-ai1 abr 20262 M de contexto2194 KRW / 4388 KRW · 1M
SpaceXAI: Grok 4.20TextoImagenArchivo

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

por x-ai1 abr 20262 M de contexto2194 KRW / 4388 KRW · 1M

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

por google31 mar 20261,05 M de contextoGratis / Gratis · 1M

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...

por google31 mar 20261,05 M de contextoGratis / Gratis · 1M

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...

por kwaipilot28 mar 2026262,14 mil de contexto527 KRW / 2106 KRW · 1M
Reka EdgeImagenTextoVideo

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...

por rekaai21 mar 202616,38 mil de contexto176 KRW / 176 KRW · 1M

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

por minimax18 mar 2026204,8 mil de contexto527 KRW / 2106 KRW · 1M

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

por minimax18 mar 2026196,61 mil de contextoGratis / Gratis · 1M
OpenAI: GPT-5.4 NanoArchivoImagenTexto

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

por openai17 mar 2026400 mil de contexto351 KRW / 2194 KRW · 1M

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

por openai17 mar 2026400 mil de contexto176 KRW / 1097 KRW · 1M
OpenAI: GPT-5.4 MiniArchivoImagenTexto

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

por openai17 mar 2026400 mil de contexto1316 KRW / 7898 KRW · 1M

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

por openai17 mar 2026400 mil de contexto658 KRW / 3949 KRW · 1M

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

por mistralai17 mar 2026262,14 mil de contexto263 KRW / 1053 KRW · 1M

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

por z-ai15 mar 2026202,75 mil de contexto2106 KRW / 7020 KRW · 1M

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

por nvidia12 mar 20261 M de contexto149 KRW / 702 KRW · 1M

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

por nvidia12 mar 2026262,14 mil de contextoGratis / Gratis · 1M

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...

por bytedance-seed11 mar 2026262,14 mil de contexto439 KRW / 3510 KRW · 1M
Qwen: Qwen3.5-9BTextoImagenVideo

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

por qwen10 mar 2026262,14 mil de contexto176 KRW / 263 KRW · 1M
Qwen: Qwen3.5-9B (batch)TextoImagenVideo

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

por qwen10 mar 2026262,14 mil de contexto298 KRW / 439 KRW · 1M
OpenAI: GPT-5.4 ProTextoImagenArchivo

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

por openai6 mar 20261,05 M de contexto52.650 KRW / 315.900 KRW · 1M
OpenAI: GPT-5.4 Pro (batch)TextoImagenArchivo

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

por openai6 mar 20261,05 M de contexto26.325 KRW / 157.950 KRW · 1M
OpenAI: GPT-5.4TextoImagenArchivo

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

por openai6 mar 20261,05 M de contexto4388 KRW / 26.325 KRW · 1M
OpenAI: GPT-5.4 (batch)TextoImagenArchivo

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

por openai6 mar 20261,05 M de contexto2194 KRW / 13.163 KRW · 1M

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

por inception4 mar 2026128 mil de contexto439 KRW / 1316 KRW · 1M

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

por google3 mar 20261,05 M de contexto439 KRW / 2633 KRW · 1M

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

por bytedance-seed27 feb 2026262,14 mil de contexto176 KRW / 702 KRW · 1M

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

por google27 feb 202665,54 mil de contexto878 KRW / 5265 KRW · 1M
Qwen: Qwen3.5-35B-A3BTextoImagenVideo

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

por qwen26 feb 2026262,14 mil de contexto548 KRW / 2194 KRW · 1M
Qwen: Qwen3.5-27BTextoImagenVideo

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

por qwen26 feb 2026262,14 mil de contexto342 KRW / 2738 KRW · 1M
Qwen: Qwen3.5-122B-A10BTextoImagenVideo

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

por qwen26 feb 2026262,14 mil de contexto509 KRW / 4212 KRW · 1M
Qwen: Qwen3.5-FlashTextoImagenVideo

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

por qwen26 feb 20261 M de contexto114 KRW / 456 KRW · 1M

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

por google26 feb 20261,05 M de contexto3510 KRW / 21.060 KRW · 1M
OpenAI: GPT-5.3-CodexTextoImagenArchivo

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...

por openai25 feb 2026400 mil de contexto3071 KRW / 24.570 KRW · 1M