AI Models

400 型号 Free & Paid 更新: 5 hours trước

Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable...

经过 |Aug 2026 |262K 上下文 |自由输入 |自由输出
262K 代币

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, 图片, 视频, audio, and PDF documents, returns text, and offers a 1M-token context...

经过 |Aug 2026 |1M 上下文 |$1.25/米输入 |$4.25/米输出
1M代币

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,...

经过 |Aug 2026 |1M 上下文 |$2.00/米输入 |$6.00/米输出
1M代币

This model always redirects to the latest model in the DeepSeek V4 Flash family.

经过 |Aug 2026 |1M 上下文 |$0.0800/米输入 |$0.2520/米输出
1M代币

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, 推理, and agent workflows.

经过 |Jul 2026 |1M 上下文 |$0.0800/米输入 |$0.1800/米输出
1M代币

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

经过 |Jul 2026 |524K 上下文 |$0.4500/米输入 |$1.20/米输出
524K 代币

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, 空间理解, and real-world...

经过 |Jul 2026 |1M 上下文 |$0.0300/米输入 |$0.1300/米输出
1M代币

Fast-mode variant of [作品 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

经过 |Jul 2026 |1M 上下文 |$10.00/米输入 |$50.00/米输出
1M代币

近距离工作 5 is Anthropic’s flagship model for demanding reasoning, 编码, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

经过 |Jul 2026 |1M 上下文 |$2.50/米输入 |$12.50/米输出
1M代币

近距离工作 5 is Anthropic’s flagship model for demanding reasoning, 编码, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

经过 |Jul 2026 |1M 上下文 |$5.00/米输入 |$25.00/米输出
1M代币

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (教育部) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

经过 |Jul 2026 |262K 上下文 |$0.0210/米输入 |$0.0630/米输出
262K 代币

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

经过 |Jul 2026 |262K 上下文 |自由输入 |自由输出
262K 代币

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

经过 |Jul 2026 |1M 上下文 |$0.0900/米输入 |$0.1800/米输出
1M代币

双子座 3.6 Flash is a high-efficiency model from Google for coding, 代理工作流程, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

经过 |Jul 2026 |1M 上下文 |$1.50/米输入 |$7.50/米输出
1M代币

双子座 3.6 Flash is a high-efficiency model from Google for coding, 代理工作流程, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

经过 |Jul 2026 |1M 上下文 |$0.7500/米输入 |$3.75/米输出
1M代币

双子座 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

经过 |Jul 2026 |1M 上下文 |$0.3000/米输入 |$2.50/米输出
1M代币

双子座 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

经过 |Jul 2026 |1M 上下文 |$0.1500/米输入 |$1.25/米输出
1M代币

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...

经过 |Jul 2026 |1M 上下文 |$0.3000/米输入 |$1.20/米输出
1M代币

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, 编码, agentic and tool-use systems,...

经过 |Jul 2026 |524K 上下文 |$0.5000/米输入 |$2.03/米输出
524K 代币

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, 编码, agentic and tool-use systems,...

经过 |Jul 2026 |1M 上下文 |$0.9500/米输入 |$4.05/米输出
1M代币

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

经过 |Jul 2026 |2M 上下文 |自由输入 |自由输出
2M代币

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

经过 |Jul 2026 |1M 上下文 |$3.00/米输入 |$15.00/米输出
1M代币

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, 图片, 视频, audio, and PDF documents and returns text, with a 1M-token context...

经过 |Jul 2026 |1M 上下文 |$1.25/米输入 |$4.25/米输出
1M代币

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

经过 |Jul 2026 |256K 上下文 |$0.1500/米输入 |$0.6000/米输出
256K 代币

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

经过 |Jul 2026 |256K 上下文 |$0.7400/米输入 |$2.96/米输出
256K 代币

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$0.1000/米输入 |$0.6000/米输出
1.1M代币

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$0.1000/米输入 |$0.6000/米输出
1.1M代币

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

经过 |Jul 2026 |1.1M 上下文 |$0.1000/米输入 |$0.6000/米输出
1.1M代币

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

经过 |Jul 2026 |1.1M 上下文 |$0.1000/米输入 |$0.6000/米输出
1.1M代币

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$1.00/米输入 |$6.00/米输出
1.1M代币

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$1.00/米输入 |$6.00/米输出
1.1M代币

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, 推理, and agentic...

经过 |Jul 2026 |1.1M 上下文 |$1.00/米输入 |$6.00/米输出
1.1M代币

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, 推理, and agentic...

经过 |Jul 2026 |1.1M 上下文 |$1.00/米输入 |$6.00/米输出
1.1M代币

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$5.00/米输入 |$30.00/米输出
1.1M代币

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$2.50/米输入 |$15.00/米输出
1.1M代币

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, 编码, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

经过 |Jul 2026 |1.1M 上下文 |$5.00/米输入 |$30.00/米输出
1.1M代币

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, 编码, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

经过 |Jul 2026 |1.1M 上下文 |$2.50/米输入 |$15.00/米输出
1.1M代币

格罗克 4.5 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

经过 |Jul 2026 |500K 上下文 |$2.00/米输入 |$6.00/米输出
500K 代币

This model always redirects to the latest Grok model from xAI.

经过 |Jul 2026 |500K 上下文 |$2.00/米输入 |$6.00/米输出
500K 代币

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...

经过 |Jul 2026 |131K 上下文 |$0.7000/米输入 |$1.40/米输出
131K 代币

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...

经过 |Jul 2026 |131K 上下文 |$3.00/米输入 |$6.00/米输出
131K 代币

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B 活跃, 192 experts with top-8 routing) built for reasoning, 代理工作流程, and real-world production use. It supports a configurable reasoning effort:...

经过 |Jul 2026 |262K 上下文 |$0.1320/米输入 |$0.5280/米输出
262K 代币

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

经过 |Jul 2026 |262K 上下文 |$0.0600/米输入 |$0.1200/米输出
262K 代币

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

经过 |Jul 2026 |262K 上下文 |自由输入 |自由输出
262K 代币

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

经过 |六月 2026 |1M 上下文 |$1.00/米输入 |$5.00/米输出
1M代币

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

经过 |六月 2026 |1M 上下文 |$2.00/米输入 |$10.00/米输出
1M代币

纳米香蕉 2 精简版 (双子座 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation...

经过 |六月 2026 |66K 上下文 |$0.2500/米输入 |$1.50/米输出
66K 代币

Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series. It accepts text and image input and is built for coding, 工具使用,...

经过 |六月 2026 |262K 上下文 |$0.0250/米输入 |$0.1000/米输出
262K 代币

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

经过 |六月 2026 |1M 上下文 |$5.00/米输入 |$30.00/米输出
1M代币

双子座 3.1 Flash Image, 又名. “纳米香蕉2," 是 Google 最新最先进的图像生成和编辑模型, 以 Flash 速度提供专业级视觉质量. It combines advanced...

经过 |六月 2026 |131K 上下文 |$0.5000/米输入 |$3.00/米输出
131K 代币