AIモデル

466 モデル Free & Paid Cập nhật: 49 minutes trước

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

による |ジュン 2026 |131K context |$0.2000/M input |$0.2000/M output
131K tokens ⓘ

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

による |ジュン 2026 |128K context |Miễn phí input |Miễn phí output
128K tokens ⓘ

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

による |ジュン 2026 |262K context |$0.5000/M input |$2.20/M output
262K tokens ⓘ

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

による |ジュン 2026 |1M context |Miễn phí input |Miễn phí output
1M tokens ⓘ

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

による |ジュン 2026 |1M context |$0.3200/M input |$1.28/M output
1M tokens ⓘ

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, 画像, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

による |5月 2026 |1M context |$0.3000/M input |$1.20/M output
1M tokens ⓘ

ステップ 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

による |5月 2026 |262K context |$0.2000/M input |$1.15/M output
262K tokens ⓘ

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, 画像, and file inputs with text output, with reasoning support and a 1M-token...

による |5月 2026 |1M context |$5.00/M input |$25.00/M output
1M tokens ⓘ

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, 画像, and file inputs with text output, with reasoning support and a 1M-token...

による |5月 2026 |1M context |$2.50/M input |$12.50/M output
1M tokens ⓘ

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

による |5月 2026 |1M context |$1.48/M input |$4.43/M output
1M tokens ⓘ

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

による |5月 2026 |256K context |$1.00/M input |$2.00/M output
256K tokens ⓘ

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

による |5月 2026 |1M context |$1.50/M input |$9.00/M output
1M tokens ⓘ

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

による |5月 2026 |1M context |$0.7500/M input |$4.50/M output
1M tokens ⓘ

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

による |5月 2026 |33K context |$0.1500/M input |$1.50/M output

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, 画像, ビデオ, audio, and PDF inputs, and is designed for lightweight agentic...

による |5月 2026 |1M context |$0.1250/M input |$0.7500/M output
1M tokens ⓘ

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, 画像, ビデオ, audio, and PDF inputs, and is designed for lightweight agentic...

による |5月 2026 |1M context |$0.2500/M input |$1.50/M output
1M tokens ⓘ

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

による |5月 2026 |400K context |$5.00/M input |$30.00/M output
400K tokens ⓘ

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

による |4月 2026 |1M context |$1.00/M input |$2.00/M output
1M tokens ⓘ

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

による |4月 2026 |1M context |$1.25/M input |$2.50/M output
1M tokens ⓘ

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

による |4月 2026 |262K context |$1.50/M input |$7.50/M output
262K tokens ⓘ

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

による |4月 2026 |262K context |$0.7500/M input |$3.75/M output
262K tokens ⓘ

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, 画像, ビデオ, and...

による |4月 2026 |256K context |Miễn phí input |Miễn phí output
256K tokens ⓘ

This model always redirects to the latest model in the Claude Haiku family.

による |4月 2026 |200K context |$1.00/M input |$5.00/M output
200K tokens ⓘ

This model always redirects to the latest model in the GPT Mini family.

による |4月 2026 |400K context |$0.7500/M input |$4.50/M output
400K tokens ⓘ

This model always redirects to the latest model in the Gemini Pro family.

による |4月 2026 |1M context |$2.00/M input |$12.00/M output
1M tokens ⓘ

This model always redirects to the latest model in the Kimi family.

による |4月 2026 |1M context |$1.00/M input |$11.36/M output
1M tokens ⓘ

This model always redirects to the latest model in the Gemini Flash family.

による |4月 2026 |1M context |$0.7500/M input |$3.75/M output
1M tokens ⓘ

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, 画像, and video input and produces text output, with a 1M token context window. This...

による |4月 2026 |1M context |$0.3000/M input |$1.80/M output
1M tokens ⓘ

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 シリーズ. It supports text, 画像, and video input with a 1M token context window. Tiered pricing kicks in...

による |4月 2026 |1M context |$0.1875/M input |$1.13/M output
1M tokens ⓘ

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...

による |4月 2026 |262K context |$0.1500/M input |$1.00/M output
262K tokens ⓘ

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...

による |4月 2026 |262K context |$1.03/M input |$6.16/M output
262K tokens ⓘ

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, 画像, and video inputs...

による |4月 2026 |262K context |$0.3200/M input |$3.20/M output
262K tokens ⓘ

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

による |4月 2026 |1.1M context |$30.00/M input |$180.00/M output
1.1M tokens ⓘ

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

による |4月 2026 |1.1M context |$15.00/M input |$90.00/M output
1.1M tokens ⓘ

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

による |4月 2026 |1.1M context |$5.00/M input |$30.00/M output
1.1M tokens ⓘ

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

による |4月 2026 |1.1M context |$2.50/M input |$15.00/M output
1.1M tokens ⓘ

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

による |4月 2026 |1M context |$0.2088/M input |$0.4176/M output
1M tokens ⓘ

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

による |4月 2026 |1M context |$0.0280/M input |$0.0560/M output
1M tokens ⓘ

Hy3 preview is a high-efficiency Mixture-of-Experts model from Tencent designed for agentic workflows and production use. It supports configurable reasoning levels across disabled, low, and high modes, allowing it to...

による |4月 2026 |262K context |$0.1800/M input |$0.6000/M output
262K tokens ⓘ

MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....

による |4月 2026 |1.1M context |$0.4350/M input |$0.8700/M output
1.1M tokens ⓘ

MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...

による |4月 2026 |1.1M context |$0.1400/M input |$0.2800/M output
1.1M tokens ⓘ

[GPT-5.4](https://openrouter.ai/openai/gpt-5.4) Image 2 combines OpenAI's GPT-5.4 model with state-of-the-art image generation capabilities from GPT Image 2. It enables rich multimodal workflows, allowing users to seamlessly move between reasoning, coding, and...

による |4月 2026 |272K context |$8.00/M input |$15.00/M output
272K tokens ⓘ

This model always redirects to the latest model in the Claude Opus family.

による |4月 2026 |1M context |$4.00/M input |$20.00/M output
1M tokens ⓘ

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by [Artificial Analysis](https://artificialanalysis.ai/) coding percentiles. Set min_coding_score between 0 and 1 on the [pareto-router plugin](https://openrouter.ai/docs/guides/routing/routers/pareto-router#the-min_coding_score-parameter) to control how...

による |4月 2026 |2M context |Miễn phí input |Miễn phí output
2M tokens ⓘ

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

による |4月 2026 |262K context |$0.4342/M input |$1.83/M output
262K tokens ⓘ

オーパス 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

による |4月 2026 |1M context |$5.00/M input |$25.00/M output
1M tokens ⓘ

オーパス 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...

による |4月 2026 |1M context |$2.50/M input |$12.50/M output
1M tokens ⓘ

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

による |4月 2026 |205K context |$0.9646/M input |$3.03/M output
205K tokens ⓘ

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

による |4月 2026 |262K context |$0.0675/M input |$0.2250/M output
262K tokens ⓘ