AI Models

414 models Free & Paid Cập nhật: 5 hours trước

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...

による |4月 2026 |131K コンテキスト |$0.0500/M入力 |$0.1000/M出力
131K トークン

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

による |4月 2026 |262K コンテキスト |$1.50/M入力 |$7.50/M出力
262K トークン

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, ビデオ, and...

による |4月 2026 |256K コンテキスト |Miễn phí input |Miễn phí output
256K トークン

This model always redirects to the latest model in the Anthropic Claude Haiku family.

による |4月 2026 |200K コンテキスト |$1.00/M入力 |$5.00/M出力
200K トークン

This model always redirects to the latest model in the OpenAI GPT Mini family.

による |4月 2026 |400K コンテキスト |$0.7500/M入力 |$4.50/M出力
400K トークン

This model always redirects to the latest model in the Google Gemini Pro family.

による |4月 2026 |1M コンテキスト |$2.00/M入力 |$12.00/M出力
1M トークン

This model always redirects to the latest model in the MoonshotAI Kimi family.

による |4月 2026 |1M コンテキスト |$2.60/M入力 |$13.00/M出力
1M トークン

This model always redirects to the latest model in the Google Gemini Flash family.

による |4月 2026 |1M コンテキスト |$0.3750/M入力 |$1.88/M出力
1M トークン

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

による |4月 2026 |1M コンテキスト |$2.00/M入力 |$10.00/M出力
1M トークン

This model always redirects to the latest model in the OpenAI GPT family.

による |4月 2026 |1.1M コンテキスト |$2.50/M入力 |$15.00/M出力
1.1M トークン

Qwen3.5 Plus (4月 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

による |4月 2026 |1M コンテキスト |$0.3000/M入力 |$1.80/M出力
1M トークン

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 シリーズ. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...

による |4月 2026 |1M コンテキスト |$0.1875/M入力 |$1.13/M出力
1M トークン

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

による |4月 2026 |262K コンテキスト |$0.1400/M入力 |$1.00/M出力
262K トークン

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

による |4月 2026 |262K コンテキスト |$1.03/M入力 |$6.16/M出力
262K トークン

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...

による |4月 2026 |262K コンテキスト |$0.6000/M入力 |$3.60/M出力
262K トークン

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

による |4月 2026 |1.1M コンテキスト |$15.00/M入力 |$90.00/M出力
1.1M トークン

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

による |4月 2026 |1.1M コンテキスト |$30.00/M入力 |$180.00/M出力
1.1M トークン

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

による |4月 2026 |1.1M コンテキスト |$2.50/M入力 |$15.00/M出力
1.1M トークン

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

による |4月 2026 |1.1M コンテキスト |$5.00/M入力 |$30.00/M出力
1.1M トークン

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

による |4月 2026 |1M コンテキスト |$1.60/M入力 |$3.20/M出力
1M トークン

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

による |4月 2026 |1M コンテキスト |$0.0886/M入力 |$0.1772/M出力
1M トークン

Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a “fast...

による |4月 2026 |262K コンテキスト |$0.0750/M入力 |$0.6250/M出力
262K トークン

Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use. It supports configurable reasoning levels across disabled, 低い, and high modes, allowing it to...

による |4月 2026 |262K コンテキスト |$0.1800/M入力 |$0.6000/M出力
262K トークン

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

による |4月 2026 |1.1M コンテキスト |$0.4350/M入力 |$0.8700/M出力
1.1M トークン

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

による |4月 2026 |1.1M コンテキスト |$0.1400/M入力 |$0.2800/M出力
1.1M トークン

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...

による |4月 2026 |272K コンテキスト |$8.00/M入力 |$15.00/M出力
272K トークン

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....

による |4月 2026 |262K コンテキスト |$0.0100/M入力 |$0.0300/M出力
262K トークン

This model always redirects to the latest model in the Claude Opus family.

による |4月 2026 |1M コンテキスト |$5.00/M入力 |$25.00/M出力
1M トークン

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://