人工智能模型

464 型号 Free & Paid 更新: 12 hours trước

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...

by |Oct 2025 |200K context |$1.00/M input |$5.00/M output
200K tokens ⓘ

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, 文件, and temporal sequences. It integrates enhanced multimodal alignment and...

by |Oct 2025 |131K context |$0.1800/M input |$2.10/M output
131K tokens ⓘ

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

by |Oct 2025 |262K context |$0.1170/M input |$0.4550/M output
262K tokens ⓘ

[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while incorporating GPT Image 1's superior instruction following,...

by |Oct 2025 |400K context |$10.00/M input |$10.00/M output
400K tokens ⓘ

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,...

by |Oct 2025 |33K context |$0.3000/M input |$2.50/M output

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...

by |Oct 2025 |262K context |$0.2000/M input |$2.40/M output
262K tokens ⓘ

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

by |Oct 2025 |262K context |$0.1500/M input |$0.6000/M output
262K tokens ⓘ

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. 它针对需要逐步推理的复杂任务进行了优化, instruction following, and...

by |Oct 2025 |400K context |$15.00/M input |$120.00/M output
400K tokens ⓘ

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. 它针对需要逐步推理的复杂任务进行了优化, instruction following, and...

by |Oct 2025 |400K context |$7.50/M input |$60.00/M output
400K tokens ⓘ

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

by |九月 2025 |205K context |$0.4300/M input |$1.75/M output
205K tokens ⓘ

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

by |九月 2025 |1M context |$1.50/M input |$7.50/M output
1M tokens ⓘ

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...

by |九月 2025 |1M context |$3.00/M input |$15.00/M output
1M tokens ⓘ

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. 它引入了 DeepSeek 稀疏注意力 (DSA), a fine-grained sparse attention mechanism...

by |九月 2025 |164K context |$0.2700/M input |$0.4100/M output
164K tokens ⓘ

Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.

by |九月 2025 |131K context |$0.3000/M input |$0.5000/M output
131K tokens ⓘ

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, 克洛德, and others into your files at...

by |九月 2025 |256K context |$0.8500/M input |$1.25/M output
256K tokens ⓘ

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....

by |九月 2025 |131K context |$0.4000/M input |$4.00/M output
131K tokens ⓘ

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...

by |九月 2025 |262K context |$0.2100/M input |$1.90/M output
262K tokens ⓘ

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...

by |九月 2025 |262K context |$0.7800/M input |$3.90/M output
262K tokens ⓘ

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...

by |九月 2025 |1M context |$0.6500/M input |$3.25/M output
1M tokens ⓘ

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...

by |九月 2025 |164K context |$0.3000/M input |$1.00/M output
164K tokens ⓘ

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...

by |九月 2025 |1M context |$0.1950/M input |$0.9750/M output
1M tokens ⓘ

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, 逻辑, and agentic...

by |九月 2025 |262K context |$0.1500/M input |$1.20/M output
262K tokens ⓘ

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

by |九月 2025 |262K context |$0.1000/M input |$1.10/M output
262K tokens ⓘ

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

by |九月 2025 |1M context |$0.2600/M input |$0.7800/M output
1M tokens ⓘ

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...

by |九月 2025 |262K context |$0.6000/M input |$2.50/M output
262K tokens ⓘ

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...

by |八月 2025 |82K context |$0.2000/M input |$2.40/M output
82K tokens ⓘ

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...

by |八月 2025 |131K context |$1.00/M input |$3.00/M output
131K tokens ⓘ

DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...

by |八月 2025 |164K context |$0.2500/M input |$0.9500/M output
164K tokens ⓘ

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

by |八月 2025 |131K context |$0.2000/M input |$1.00/M output
131K tokens ⓘ

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

by |八月 2025 |131K context |$0.4000/M input |$2.00/M output
131K tokens ⓘ

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

by |八月 2025 |66K context |$0.6000/M input |$1.80/M output
66K tokens ⓘ

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. 它针对需要逐步推理的复杂任务进行了优化, instruction following, and accuracy...

by |八月 2025 |400K context |$0.6250/M input |$5.00/M output
400K tokens ⓘ

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. 它针对需要逐步推理的复杂任务进行了优化, instruction following, and accuracy...

by |八月 2025 |400K context |$1.25/M input |$10.00/M output
400K tokens ⓘ

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

by |八月 2025 |400K context |$0.1250/M input |$1.00/M output
400K tokens ⓘ

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

by |八月 2025 |400K context |$0.2500/M input |$2.00/M output
400K tokens ⓘ

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

by |八月 2025 |400K context |$0.0250/M input |$0.2000/M output
400K tokens ⓘ

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

by |八月 2025 |400K context |$0.0500/M input |$0.4000/M output
400K tokens ⓘ

gpt-oss-120b is an open-weight, 117B 参数专家混合 (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

by |八月 2025 |131K context |$0.0296/M input |$0.1360/M output
131K tokens ⓘ

gpt-oss-120b is an open-weight, 117B 参数专家混合 (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

by |八月 2025 |131K context |$0.0370/M input |$0.1700/M output
131K tokens ⓘ

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 执照. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

by |八月 2025 |131K context |$0.0240/M input |$0.1120/M output
131K tokens ⓘ

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 执照. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

by |八月 2025 |131K context |$0.0180/M input |$0.0900/M output
131K tokens ⓘ

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

by |八月 2025 |200K context |$15.00/M input |$75.00/M output
200K tokens ⓘ

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

by |八月 2025 |200K context |$7.50/M input |$37.50/M output
200K tokens ⓘ

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

by |八月 2025 |256K context |$0.3000/M input |$0.9000/M output
256K tokens ⓘ

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

by |八月 2025 |256K context |$0.1500/M input |$0.4500/M output
256K tokens ⓘ

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...

by |Jul 2025 |262K context |$0.0700/M input |$0.2800/M output
262K tokens ⓘ

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...

by |Jul 2025 |262K context |$0.1000/M input |$0.3000/M output
262K tokens ⓘ

GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

by |Jul 2025 |131K context |$0.6000/M input |$2.20/M output
131K tokens ⓘ

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

by |Jul 2025 |131K context |$0.1300/M input |$0.8500/M output
131K tokens ⓘ

Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...

by |Jul 2025 |131K context |$0.2300/M input |$2.30/M output
131K tokens ⓘ