Models

433 個 models

OpenAI: GPT-6 AstraFileImageText

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

由 openai 提供2026年9月5日105萬 context₩17,550 / ₩87,750 · 1M

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

由 openai 提供2026年9月5日105萬 context₩8,775 / ₩43,875 · 1M

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

由 openai 提供2026年9月5日105萬 context₩17,550 / ₩87,750 · 1M

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

由 openai 提供2026年9月5日105萬 context₩8,775 / ₩43,875 · 1M

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

由 inclusionai 提供2026年9月5日26.21萬 context免費 / 免費 · 1M

Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

由 qwen 提供2026年9月4日100萬 context₩3,510 / ₩10,530 · 1M

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...

由 meta 提供2026年9月3日104.86萬 context₩176 / ₩351 · 1M
Meta: Muse Spark 1.3TextImageVideo

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...

由 meta 提供2026年9月3日104.86萬 context₩2,194 / ₩7,459 · 1M

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

由 google 提供2026年9月3日104.86萬 context₩1,316 / ₩6,581 · 1M

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

由 google 提供2026年9月3日104.86萬 context₩658 / ₩3,291 · 1M

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

由 anthropic 提供2026年9月2日100萬 context₩17,550 / ₩87,750 · 1M

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

由 anthropic 提供2026年9月2日100萬 context₩8,775 / ₩43,875 · 1M

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

由 inception 提供2026年9月1日26萬 context₩70 / ₩263 · 1M

Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

由 ibm-granite 提供2026年9月1日13.11萬 context₩176 / ₩263 · 1M
每週 10.97兆 tokens

Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...

由 tencent 提供2026年8月28日104.86萬 context₩1,464 / ₩4,389 · 1M

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

由 inclusionai 提供2026年8月28日26.21萬 context₩105 / ₩316 · 1M

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

由 inclusionai 提供2026年8月28日26.21萬 context免費 / 免費 · 1M

This model always redirects to the latest model in the GLM Flash family.

由 ~z-ai 提供2026年8月27日131.07萬 context₩132 / ₩439 · 1M
Qwen: Qwen3.8 FlashTextImageVideo

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

由 qwen 提供2026年8月27日100萬 context₩263 / ₩825 · 1M
Z.ai: GLM 5.3 FlashTextImageVideo
每週 11.95兆 tokens

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

由 z-ai 提供2026年8月26日131.07萬 context₩132 / ₩439 · 1M

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

由 z-ai 提供2026年8月26日104.86萬 context₩263 / ₩878 · 1M

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

由 meta 提供2026年8月22日104.86萬 context₩176 / ₩351 · 1M

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

由 deepseek 提供2026年8月21日104.86萬 context₩386 / ₩1,158 · 1M

Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided...

由 tencent 提供2026年8月20日8192 context₩77 / ₩311 · 1M

Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and...

由 tencent 提供2026年8月20日8192 context₩130 / ₩518 · 1M

This model always redirects to the latest GLM model from Z.ai.

由 ~z-ai 提供2026年8月19日131.07萬 context₩2,053 / ₩6,950 · 1M

Hy-MT2-7B is a 7B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and style-guided translation.

由 tencent 提供2026年8月19日8192 context₩130 / ₩518 · 1M
每週 2.24兆 tokens

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

由 z-ai 提供2026年8月19日131.07萬 context₩2,457 / ₩7,722 · 1M
Qwen: Qwen3.8 27BTextImageVideo

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

由 qwen 提供2026年8月15日100萬 context₩737 / ₩5,265 · 1M

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...

由 dots-studio 提供2026年8月14日51.2萬 context免費 / 免費 · 1M
每週 2.63兆 tokens

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

由 google 提供2026年8月14日104.86萬 context₩1,316 / ₩6,581 · 1M

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

由 google 提供2026年8月14日104.86萬 context₩658 / ₩3,291 · 1M

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

由 bytedance-seed 提供2026年8月13日26.21萬 context₩878 / ₩4,388 · 1M

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

由 qwen 提供2026年8月13日104.86萬 context₩3,510 / ₩10,530 · 1M

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

由 qwen 提供2026年8月13日101萬 context₩3,510 / ₩10,530 · 1M

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...

由 bytedance-seed 提供2026年8月13日26.21萬 context₩878 / ₩5,265 · 1M
每週 1.08兆 tokens

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

由 deepseek 提供2026年8月13日104.86萬 context₩1,967 / ₩5,900 · 1M

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

由 deepseek 提供2026年8月13日104.86萬 context₩2,317 / ₩6,950 · 1M
SpaceXAI: Grok 4.6TextImageFile

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

由 x-ai 提供2026年8月13日50萬 context₩3,510 / ₩10,530 · 1M

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

由 liquid 提供2026年8月12日6.55萬 context免費 / 免費 · 1M

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

由 nvidia 提供2026年8月11日26.21萬 context₩140 / ₩351 · 1M

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

由 nvidia 提供2026年8月11日100萬 context免費 / 免費 · 1M

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

由 sakana 提供2026年8月11日26.21萬 context₩1,667 / ₩7,020 · 1M

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity, document-intensive...

由 upstage 提供2026年8月10日52.43萬 context₩53 / ₩211 · 1M

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

由 meta 提供2026年8月10日13.11萬 context₩527 / ₩1,931 · 1M

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

由 meta 提供2026年8月10日13.11萬 context₩614 / ₩2,633 · 1M
Meta: Muse Spark 1.2TextImageVideo

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...

由 meta 提供2026年8月6日104.86萬 context₩2,194 / ₩7,459 · 1M

This model always redirects to the latest model in the DeepSeek V4 Flash family.

由 ~deepseek 提供2026年8月2日131.07萬 context₩88 / ₩175 · 1M
每週 11.29兆 tokens

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

由 deepseek 提供2026年7月31日131.07萬 context₩114 / ₩316 · 1M

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

由 deepseek 提供2026年7月31日104.86萬 context₩246 / ₩491 · 1M

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

由 thinkingmachines 提供2026年7月31日104.86萬 context₩790 / ₩2,106 · 1M

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

由 thinkingmachines 提供2026年7月31日52.43萬 context₩878 / ₩2,106 · 1M

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

由 thinkingmachines 提供2026年7月31日104.86萬 context免費 / 免費 · 1M
Qwen: Qwen3.7 FlashTextImageVideo

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

由 qwen 提供2026年7月28日100萬 context₩53 / ₩228 · 1M
Claude Opus 5TextImageFile
每週 1.43兆 tokens

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

由 anthropic 提供2026年7月25日100萬 context₩8,775 / ₩43,875 · 1M

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

由 anthropic 提供2026年7月25日100萬 context₩4,388 / ₩21,938 · 1M

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

由 inclusionai 提供2026年7月23日26.21萬 context₩37 / ₩111 · 1M
每週 1.36兆 tokens

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

由 poolside 提供2026年7月22日104.86萬 context₩158 / ₩316 · 1M

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

由 poolside 提供2026年7月22日26.21萬 context免費 / 免費 · 1M

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

由 google 提供2026年7月22日104.86萬 context₩1,316 / ₩6,581 · 1M

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

由 google 提供2026年7月22日104.86萬 context₩658 / ₩3,291 · 1M

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

由 google 提供2026年7月22日104.86萬 context₩527 / ₩4,388 · 1M

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

由 google 提供2026年7月22日104.86萬 context₩263 / ₩2,194 · 1M

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...

由 meituan 提供2026年7月20日104.88萬 context₩527 / ₩2,106 · 1M

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

由 thinkingmachines 提供2026年7月18日104.86萬 context₩1,755 / ₩7,108 · 1M

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

由 thinkingmachines 提供2026年7月18日52.43萬 context₩1,755 / ₩7,108 · 1M

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

由 thinkingmachines 提供2026年7月18日104.86萬 context免費 / 免費 · 1M
Auto Router (Beta)TextImageAudio

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

由 openrouter 提供2026年7月18日200萬 context免費 / 免費 · 1M
MoonshotAI: Kimi K3TextImageVideo
每週 1.81兆 tokens

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

由 moonshotai 提供2026年7月17日104.86萬 context₩5,265 / ₩26,325 · 1M

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

由 moonshotai 提供2026年7月17日104.86萬 context₩5,265 / ₩26,325 · 1M
Meta: Muse Spark 1.1TextImageVideo

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...

由 meta 提供2026年7月17日104.86萬 context₩2,194 / ₩7,459 · 1M

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

由 kwaipilot 提供2026年7月11日26.21萬 context₩1,299 / ₩5,195 · 1M

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

由 openai 提供2026年7月9日105萬 context₩351 / ₩2,106 · 1M

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

由 openai 提供2026年7月9日105萬 context₩176 / ₩1,053 · 1M
每週 11.6兆 tokens

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

由 openai 提供2026年7月9日105萬 context₩351 / ₩2,106 · 1M

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

由 openai 提供2026年7月9日105萬 context₩176 / ₩1,053 · 1M

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

由 openai 提供2026年7月9日105萬 context₩3,510 / ₩21,060 · 1M

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

由 openai 提供2026年7月9日105萬 context₩1,755 / ₩10,530 · 1M

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

由 openai 提供2026年7月9日105萬 context₩3,510 / ₩21,060 · 1M

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

由 openai 提供2026年7月9日105萬 context₩1,755 / ₩10,530 · 1M

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

由 openai 提供2026年7月9日105萬 context₩3,510 / ₩17,550 · 1M

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

由 openai 提供2026年7月9日105萬 context₩1,755 / ₩8,775 · 1M
OpenAI: GPT-5.6 SolFileImageText
每週 1.81兆 tokens

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

由 openai 提供2026年7月9日105萬 context₩3,510 / ₩17,550 · 1M

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

由 openai 提供2026年7月9日105萬 context₩1,755 / ₩8,775 · 1M
SpaceXAI: Grok 4.5TextImageFile

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

由 x-ai 提供2026年7月9日50萬 context₩3,510 / ₩10,530 · 1M
xAI: Grok LatestTextImageFile

This model always redirects to the latest Grok model from xAI.

由 ~x-ai 提供2026年7月8日50萬 context₩3,510 / ₩10,530 · 1M

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...

由 aion-labs 提供2026年7月8日13.11萬 context₩1,229 / ₩2,457 · 1M

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...

由 aion-labs 提供2026年7月8日13.11萬 context₩5,265 / ₩10,530 · 1M
每週 5.25兆 tokens

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

由 tencent 提供2026年7月6日26.21萬 context₩232 / ₩927 · 1M

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

由 poolside 提供2026年7月2日26.21萬 context₩105 / ₩211 · 1M

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

由 poolside 提供2026年7月2日26.21萬 context免費 / 免費 · 1M
每週 1.27兆 tokens

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

由 anthropic 提供2026年7月1日100萬 context₩3,510 / ₩17,550 · 1M

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

由 anthropic 提供2026年7月1日100萬 context₩1,755 / ₩8,775 · 1M

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation...

由 google 提供2026年7月1日6.55萬 context₩439 / ₩2,633 · 1M

Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series. It accepts text and image input and is built for coding, tool use,...

由 nex-agi 提供2026年6月24日26.21萬 context₩44 / ₩176 · 1M

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

由 sakana 提供2026年6月24日100萬 context₩8,775 / ₩52,650 · 1M

Gemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines advanced...

由 google 提供2026年6月18日13.11萬 context₩878 / ₩5,265 · 1M

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...

由 google 提供2026年6月18日13.11萬 context₩3,510 / ₩21,060 · 1M

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...

由 cohere 提供2026年6月18日25.6萬 context免費 / 免費 · 1M
每週 2.19兆 tokens

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

由 z-ai 提供2026年6月17日104.86萬 context₩1,695 / ₩5,328 · 1M

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

由 z-ai 提供2026年6月17日25.6萬 context免費 / 免費 · 1M

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...

由 openrouter 提供2026年6月14日100萬 context免費 / 免費 · 1M

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

由 moonshotai 提供2026年6月12日26.21萬 context₩1,158 / ₩5,967 · 1M

This model always redirects to the latest model in the Claude Fable family.

由 ~anthropic 提供2026年6月10日100萬 context₩17,550 / ₩87,750 · 1M

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

由 anthropic 提供2026年6月9日100萬 context₩17,550 / ₩87,750 · 1M

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

由 anthropic 提供2026年6月9日100萬 context₩8,775 / ₩43,875 · 1M

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...

由 nex-agi 提供2026年6月9日26.21萬 context₩439 / ₩1,755 · 1M

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

由 nvidia 提供2026年6月4日13.11萬 context₩351 / ₩351 · 1M

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

由 nvidia 提供2026年6月4日12.8萬 context免費 / 免費 · 1M
每週 4.1兆 tokens

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

由 nvidia 提供2026年6月4日26.21萬 context₩1,097 / ₩5,484 · 1M

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

由 nvidia 提供2026年6月4日100萬 context免費 / 免費 · 1M

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

由 qwen 提供2026年6月3日100萬 context₩562 / ₩2,246 · 1M
MiniMax: MiniMax M3TextImageVideo
每週 6.6兆 tokens

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

由 minimax 提供2026年6月1日104.86萬 context₩527 / ₩2,106 · 1M

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

由 minimax 提供2026年6月1日52.43萬 context₩527 / ₩2,106 · 1M

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

由 minimax 提供2026年6月1日104.86萬 context免費 / 免費 · 1M

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

由 stepfun 提供2026年5月29日26.21萬 context₩351 / ₩2,018 · 1M

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

由 anthropic 提供2026年5月28日100萬 context₩8,775 / ₩43,875 · 1M

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

由 anthropic 提供2026年5月28日100萬 context₩4,388 / ₩21,938 · 1M

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

由 qwen 提供2026年5月22日100萬 context₩2,589 / ₩7,766 · 1M

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

由 x-ai 提供2026年5月21日25.6萬 context₩1,755 / ₩3,510 · 1M

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

由 google 提供2026年5月19日104.86萬 context₩2,633 / ₩15,795 · 1M

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

由 google 提供2026年5月19日104.86萬 context₩1,316 / ₩7,898 · 1M

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

由 perceptron 提供2026年5月12日3.28萬 context₩263 / ₩2,633 · 1M

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

由 google 提供2026年5月8日104.86萬 context₩439 / ₩2,633 · 1M

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

由 google 提供2026年5月8日104.86萬 context₩219 / ₩1,316 · 1M

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

由 openai 提供2026年5月6日40萬 context₩8,775 / ₩52,650 · 1M
SpaceXAI: Grok 4.3TextImageFile

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

由 x-ai 提供2026年5月1日100萬 context₩2,194 / ₩4,388 · 1M

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

由 x-ai 提供2026年5月1日100萬 context₩1,755 / ₩3,510 · 1M

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

由 mistralai 提供2026年5月1日26.21萬 context₩2,633 / ₩13,163 · 1M

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

由 mistralai 提供2026年5月1日3.28萬 context₩1,316 / ₩6,581 · 1M

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

由 nvidia 提供2026年4月29日25.6萬 context免費 / 免費 · 1M

This model always redirects to the latest model in the Anthropic Claude Haiku family.

由 ~anthropic 提供2026年4月28日20萬 context₩1,755 / ₩8,775 · 1M

This model always redirects to the latest model in the OpenAI GPT Mini family.

由 ~openai 提供2026年4月28日40萬 context₩1,316 / ₩7,898 · 1M

This model always redirects to the latest model in the Google Gemini Pro family.

由 ~google 提供2026年4月28日104.86萬 context₩3,510 / ₩21,060 · 1M

This model always redirects to the latest model in the MoonshotAI Kimi family.

由 ~moonshotai 提供2026年4月28日104.86萬 context₩4,475 / ₩22,376 · 1M

This model always redirects to the latest model in the Google Gemini Flash family.

由 ~google 提供2026年4月28日104.86萬 context₩1,316 / ₩6,581 · 1M

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

由 ~anthropic 提供2026年4月28日100萬 context₩3,510 / ₩17,550 · 1M
OpenAI GPT LatestFileImageText

This model always redirects to the latest model in the OpenAI GPT family.

由 ~openai 提供2026年4月28日105萬 context₩3,510 / ₩17,550 · 1M

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

由 qwen 提供2026年4月27日100萬 context₩527 / ₩3,159 · 1M
Qwen: Qwen3.6 FlashTextImageVideo

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

由 qwen 提供2026年4月27日100萬 context₩329 / ₩1,974 · 1M
Qwen: Qwen3.6 35B A3BTextImageVideo

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

由 qwen 提供2026年4月27日26.21萬 context₩176 / ₩1,580 · 1M

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

由 qwen 提供2026年4月27日26.21萬 context₩1,802 / ₩10,814 · 1M
Qwen: Qwen3.6 27BTextImageVideo

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

由 qwen 提供2026年4月27日26.21萬 context₩527 / ₩3,510 · 1M
OpenAI: GPT-5.5 ProFileImageText

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

由 openai 提供2026年4月25日105萬 context₩52,650 / ₩315,900 · 1M

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

由 openai 提供2026年4月25日105萬 context₩26,325 / ₩157,950 · 1M
OpenAI: GPT-5.5FileImageText

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

由 openai 提供2026年4月25日105萬 context₩8,775 / ₩52,650 · 1M

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

由 openai 提供2026年4月25日105萬 context₩4,388 / ₩26,325 · 1M
每週 1.49兆 tokens

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

由 deepseek 提供2026年4月24日104.86萬 context₩1,331 / ₩2,663 · 1M
每週 5.18兆 tokens

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

由 deepseek 提供2026年4月24日104.86萬 context₩144 / ₩287 · 1M

Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use. It supports configurable reasoning levels across disabled, low, and high modes, allowing it to...

由 tencent 提供2026年4月23日26.21萬 context₩316 / ₩1,053 · 1M

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

由 xiaomi 提供2026年4月23日105萬 context₩763 / ₩1,527 · 1M
Xiaomi: MiMo-V2.5TextAudioImage
每週 4.5兆 tokens

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

由 xiaomi 提供2026年4月23日105萬 context₩246 / ₩491 · 1M

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...

由 openai 提供2026年4月22日27.2萬 context₩14,040 / ₩26,325 · 1M

This model always redirects to the latest model in the Claude Opus family.

由 ~anthropic 提供2026年4月22日100萬 context₩8,775 / ₩43,875 · 1M

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...

由 openrouter 提供2026年4月21日200萬 context免費 / 免費 · 1M

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

由 moonshotai 提供2026年4月21日26.21萬 context₩1,667 / ₩7,020 · 1M

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

由 anthropic 提供2026年4月16日100萬 context₩8,775 / ₩43,875 · 1M

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

由 anthropic 提供2026年4月16日100萬 context₩4,388 / ₩21,938 · 1M

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

由 z-ai 提供2026年4月8日20.48萬 context₩1,695 / ₩5,328 · 1M

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

由 google 提供2026年4月3日26.21萬 context₩123 / ₩597 · 1M

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

由 google 提供2026年4月3日26.21萬 context免費 / 免費 · 1M
Google: Gemma 4 31BImageTextVideo

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

由 google 提供2026年4月3日26.21萬 context₩158 / ₩597 · 1M

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

由 google 提供2026年4月3日26.21萬 context₩684 / ₩1,702 · 1M

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

由 google 提供2026年4月3日26.21萬 context免費 / 免費 · 1M
Qwen: Qwen3.6 PlusTextImageVideo

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...

由 qwen 提供2026年4月2日100萬 context₩570 / ₩3,422 · 1M
Z.ai: GLM 5V TurboImageTextVideo

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

由 z-ai 提供2026年4月2日20.28萬 context₩2,106 / ₩7,020 · 1M

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

由 arcee-ai 提供2026年4月2日26.21萬 context₩439 / ₩1,404 · 1M

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...

由 x-ai 提供2026年4月1日200萬 context₩2,194 / ₩4,388 · 1M
SpaceXAI: Grok 4.20TextImageFile

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

由 x-ai 提供2026年4月1日200萬 context₩2,194 / ₩4,388 · 1M

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

由 google 提供2026年3月31日104.86萬 context免費 / 免費 · 1M

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...

由 google 提供2026年3月31日104.86萬 context免費 / 免費 · 1M

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...

由 kwaipilot 提供2026年3月28日26.21萬 context₩527 / ₩2,106 · 1M
Reka EdgeImageTextVideo

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...

由 rekaai 提供2026年3月21日1.64萬 context₩176 / ₩176 · 1M

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

由 minimax 提供2026年3月18日20.48萬 context₩527 / ₩2,106 · 1M

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

由 minimax 提供2026年3月18日19.66萬 context免費 / 免費 · 1M

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

由 openai 提供2026年3月17日40萬 context₩351 / ₩2,194 · 1M

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...

由 openai 提供2026年3月17日40萬 context₩176 / ₩1,097 · 1M

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

由 openai 提供2026年3月17日40萬 context₩1,316 / ₩7,898 · 1M

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...

由 openai 提供2026年3月17日40萬 context₩658 / ₩3,949 · 1M

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

由 mistralai 提供2026年3月17日26.21萬 context₩263 / ₩1,053 · 1M

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

由 z-ai 提供2026年3月15日20.28萬 context₩2,106 / ₩7,020 · 1M

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

由 nvidia 提供2026年3月12日100萬 context₩149 / ₩702 · 1M

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

由 nvidia 提供2026年3月12日26.21萬 context免費 / 免費 · 1M

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...

由 bytedance-seed 提供2026年3月11日26.21萬 context₩439 / ₩3,510 · 1M
Qwen: Qwen3.5-9BTextImageVideo

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

由 qwen 提供2026年3月10日26.21萬 context₩176 / ₩263 · 1M

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

由 qwen 提供2026年3月10日26.21萬 context₩298 / ₩439 · 1M
OpenAI: GPT-5.4 ProTextImageFile

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

由 openai 提供2026年3月6日105萬 context₩52,650 / ₩315,900 · 1M

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

由 openai 提供2026年3月6日105萬 context₩26,325 / ₩157,950 · 1M
OpenAI: GPT-5.4TextImageFile

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

由 openai 提供2026年3月6日105萬 context₩4,388 / ₩26,325 · 1M

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

由 openai 提供2026年3月6日105萬 context₩2,194 / ₩13,163 · 1M

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

由 inception 提供2026年3月4日12.8萬 context₩439 / ₩1,316 · 1M

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

由 google 提供2026年3月3日104.86萬 context₩439 / ₩2,633 · 1M

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

由 bytedance-seed 提供2026年2月27日26.21萬 context₩176 / ₩702 · 1M

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

由 google 提供2026年2月27日6.55萬 context₩878 / ₩5,265 · 1M
Qwen: Qwen3.5-35B-A3BTextImageVideo

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...

由 qwen 提供2026年2月26日26.21萬 context₩548 / ₩2,194 · 1M
Qwen: Qwen3.5-27BTextImageVideo

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

由 qwen 提供2026年2月26日26.21萬 context₩342 / ₩2,738 · 1M

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

由 qwen 提供2026年2月26日26.21萬 context₩509 / ₩4,212 · 1M
Qwen: Qwen3.5-FlashTextImageVideo

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...

由 qwen 提供2026年2月26日100萬 context₩114 / ₩456 · 1M

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

由 google 提供2026年2月26日104.86萬 context₩3,510 / ₩21,060 · 1M

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...

由 openai 提供2026年2月25日40萬 context₩3,071 / ₩24,570 · 1M