AI Models

414 models Free & Paid 更新: 5 hours trước

Qwen3.8 27B is an open-weight dense vision-language model from Qwen. It is suited for coding, professional workflows, research, multimodal interaction, and long-running agent tasks, with flexible thinking that can be...

经过 |八月 2026 |262K 上下文 |$0.4500/米输入 |$3.20/米输出
262K 代币

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is...

经过 |八月 2026 |512K 上下文 |自由输入 |自由输出
512K 代币

双子座 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

经过 |八月 2026 |1M 上下文 |$0.3750/米输入 |$1.88/米输出
1M代币

双子座 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

经过 |八月 2026 |1M 上下文 |$0.1875/米输入 |$0.9375/米输出
1M代币

种子 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

经过 |八月 2026 |262K 上下文 |$0.5000/米输入 |$2.50/米输出
262K 代币

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

经过 |八月 2026 |1M 上下文 |$2.00/米输入 |$6.00/米输出
1M代币

种子 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...

经过 |八月 2026 |262K 上下文 |$0.5000/米输入 |$3.00/米输出
262K 代币

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

经过 |八月 2026 |1M 上下文 |$0.6600/米输入 |$1.98/米输出
1M代币

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

经过 |八月 2026 |500K 上下文 |$2.00/米输入 |$6.00/米输出
500K 代币

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

经过 |八月 2026 |128K 上下文 |自由输入 |自由输出
128K 代币

NVIDIA 神经元 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

经过 |八月 2026 |1M 上下文 |$0.0800/米输入 |$0.2000/米输出
1M代币

NVIDIA 神经元 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

经过 |八月 2026 |1M 上下文 |自由输入 |自由输出
1M代币

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

经过 |八月 2026 |262K 上下文 |$0.9500/米输入 |$4.00/米输出
262K 代币

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity, document-intensive...

经过 |八月 2026 |524K 上下文 |$0.0300/米输入 |$0.1200/米输出
524K 代币

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

经过 |八月 2026 |131K 上下文 |$0.3500/米输入 |$1.50/米输出
131K 代币

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, 图片, 视频, audio, and PDF documents, returns text, and offers a 1M-token context...

经过 |八月 2026 |1M 上下文 |$1.25/米输入 |$4.25/米输出
1M代币

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,...

经过 |八月 2026 |1M 上下文 |$2.00/米输入 |$6.00/米输出
1M代币

This model always redirects to the latest model in the DeepSeek V4 Flash family.

经过 |八月 2026 |1.3M 上下文 |$0.0786/米输入 |$0.1572/米输出
1.3M代币

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

经过 |Jul 2026 |1.3M 上下文 |$0.1400/米输入 |$0.2800/米输出
1.3M代币

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

经过 |Jul 2026 |524K 上下文 |$0.4500/米输入 |$1.20/米输出
524K 代币

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, 空间理解, and real-world...

经过 |Jul 2026 |1M 上下文 |$0.0300/米输入 |$0.1300/米输出
1M代币

Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

经过 |Jul 2026 |1M 上下文 |$10.00/米输入 |$50.00/米输出
1M代币

近距离工作 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

经过 |Jul 2026 |1M 上下文 |$2.50/米输入 |$12.50/米输出
1M代币

近距离工作 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

经过 |Jul 2026 |1M 上下文 |$5.00/米输入 |$25.00/米输出
1M代币

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (教育部) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

经过 |Jul 2026 |262K 上下文 |$0.0210/米输入 |$0.0630/米输出
262K 代币

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

经过 |Jul 2026 |262K 上下文 |自由输入 |自由输出
262K 代币

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

经过 |Jul 2026 |1M 上下文 |$0.0900/米输入 |$0.1800/米输出
1M代币

双子座 3.6 Flash is a high-efficiency model from Google for coding, 代理工作流程, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

经过 |Jul 2026 |1M 上下文 |$0.7500/米输入 |$3.75/米输出
1M代币

双子座 3.6 Flash is a high-efficiency model from Google for coding, 代理工作流程, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

经过 |Jul 2026 |1M 上下文 |$0.3750/米输入 |$1.88/米输出
1M代币

双子座 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

经过 |Jul 2026 |1M 上下文 |$0.3000/米输入 |$2.50/米输出
1M代币

双子座 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

经过 |Jul 2026 |1M 上下文 |$0.1500/米输入 |$1.25/米输出
1M代币

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...

经过 |Jul 2026 |1M 上下文 |$0.3000/米输入 |$1.20/米输出
1M代币

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

经过 |Jul 2026 |524K 上下文 |$1.00/米输入 |$4.05/米输出
524K 代币

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

经过 |Jul 2026 |1M 上下文 |$0.9500/米输入 |$4.05/米输出
1M代币

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

经过 |Jul 2026 |2M 上下文 |自由输入 |自由输出
2M代币

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

经过 |Jul 2026 |1M 上下文 |$3.00/米输入 |$15.00/米输出
1M代币

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, 图片, 视频, audio, and PDF documents and returns text, with a 1M-token context...

经过 |Jul 2026 |1M 上下文 |$1.25/米输入 |$4.25/米输出
1M代币

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

经过 |Jul 2026 |256K 上下文 |$0.1500/米输入 |$0.6000/米输出
256K 代币

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

经过 |Jul 2026 |256K 上下文 |$0.7400/米输入 |$2.96/米输出
256K 代币

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$0.1000/米输入 |$0.6000/米输出
1.1M代币

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$0.2000/米输入 |$1.20/米输出
1.1M代币

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, 分类, and lightweight agentic workflows, providing capable reasoning for...

经过 |Jul 2026 |1.1M 上下文 |$0.1000/米输入 |$0.6000/米输出
1.1M代币

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, 分类, and lightweight agentic workflows, providing capable reasoning for...

经过 |Jul 2026 |1.1M 上下文 |$0.2000/米输入 |$1.20/米输出
1.1M代币

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$1.00/米输入 |$6.00/米输出
1.1M代币

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$2.00/米输入 |$12.00/米输出
1.1M代币

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

经过 |Jul 2026 |1.1M 上下文 |$1.00/米输入 |$6.00/米输出
1.1M代币

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

经过 |Jul 2026 |1.1M 上下文 |$2.00/米输入 |$12.00/米输出
1.1M代币

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$2.50/米输入 |$15.00/米输出
1.1M代币

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

经过 |Jul 2026 |1.1M 上下文 |$1.25/米输入 |$7.50/米输出
1.1M代币

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

经过 |Jul 2026 |1.1M 上下文 |$2.50/米输入 |$15.00/米输出
1.1M代币