AI Models

412 models Free & Paid Cập nhật: 9 hours trước

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 バージョン. It...

による |9月 2025 |262K コンテキスト |$0.7800/M入力 |$3.90/M出力
262K トークン

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...

による |9月 2025 |1M コンテキスト |$0.6500/M入力 |$3.25/M出力
1M トークン

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. インタラクティブな開発セッションと長時間の開発セッションの両方向けに設計されています。, independent execution of complex engineering tasks....

による |9月 2025 |400K コンテキスト |$0.6250/M入力 |$5.00/M出力
400K トークン

DeepSeek-V3.1 Terminus is an update to [ディープシーク V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...

による |9月 2025 |164K コンテキスト |$0.2700/M入力 |$1.00/M出力
164K トークン

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...

による |9月 2025 |1M コンテキスト |$0.1950/M入力 |$0.9750/M出力
1M トークン

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

による |9月 2025 |262K コンテキスト |$0.1500/M入力 |$1.20/M出力
262K トークン

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

による |9月 2025 |262K コンテキスト |$0.1000/M入力 |$1.10/M出力
262K トークン

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

による |9月 2025 |1M コンテキスト |$0.2600/M入力 |$0.7800/M出力
1M トークン

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

による |9月 2025 |1M コンテキスト |$0.2600/M入力 |$0.7800/M出力
1M トークン

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

による |9月 2025 |128K コンテキスト |Miễn phí input |Miễn phí output
128K トークン

Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...

による |9月 2025 |262K コンテキスト |$0.6000/M入力 |$2.50/M出力
262K トークン

Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...

による |8月 2025 |82K コンテキスト |$0.2000/M入力 |$2.40/M出力
82K トークン

Hermes 4 70B is a hybrid reasoning model from Nous Research, Meta-Llama-3.1-70B に基づいて構築. It introduces the same hybrid mode as the larger 405B release, allowing the model to either...

による |8月 2025 |131K コンテキスト |$0.1300/M入力 |$0.4000/M出力
131K トークン

Hermes 4 Meta-Llama-3.1-405B に基づいて構築され、Nous Research によってリリースされた大規模推論モデルです。. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...

による |8月 2025 |131K コンテキスト |$1.00/M入力 |$3.00/M出力
131K トークン

DeepSeek-V3.1 is a large hybrid reasoning model (671Bパラメータ, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...

による |8月 2025 |164K コンテキスト |$0.2500/M入力 |$0.9500/M出力
164K トークン

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

による |8月 2025 |131K コンテキスト |$0.4000/M入力 |$2.00/M出力
131K トークン

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

による |8月 2025 |66K コンテキスト |$0.6000/M入力 |$1.80/M出力
66K トークン

ジャンバ ラージ 1.7 is the latest model in the Jamba open family, offering improvements in grounding, 指示に従う, 全体的な効率. Built on a hybrid SSM-Transformer architecture with a 256K context...

による |8月 2025 |256K コンテキスト |$2.00/M入力 |$8.00/M出力
256K トークン

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, ユーザーエクスペリエンスと. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

による |8月 2025 |400K コンテキスト |$0.6250/M入力 |$5.00/M出力
400K トークン

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, ユーザーエクスペリエンスと. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

による |8月 2025 |400K コンテキスト |$1.25/M入力 |$10.00/M出力
400K トークン

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

による |8月 2025 |400K コンテキスト |$0.1250/M入力 |$1.00/M出力
400K トークン

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

による |8月 2025 |400K コンテキスト |$0.2500/M入力 |$2.00/M出力
400K トークン

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

による |8月 2025 |400K コンテキスト |$0.0250/M入力 |$0.2000/M出力
400K トークン

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

による |8月 2025 |400K コンテキスト |$0.0500/M入力 |$0.4000/M出力
400K トークン

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

による |8月 2025 |131K コンテキスト |$0.0300/M入力 |$0.1700/M出力
131K トークン

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 ライセンス. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

による |8月 2025 |131K コンテキスト |Miễn phí input |Miễn phí output
131K トークン

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 ライセンス. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

による |8月 2025 |131K コンテキスト |$0.0300/M入力 |$0.1300/M出力
131K トークン

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

による |8月 2025 |200K コンテキスト |$7.50/M入力 |$37.50/M出力
200K トークン

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

による |8月 2025 |200K コンテキスト |$15.00/M入力 |$75.00/M出力
200K トークン

Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)

による |8月 2025 |256K コンテキスト |$0.3000/M入力 |$0.9000/M出力
256K トークン

Qwen3-Coder-30B-A3B-Instruct は 30.5B パラメータの専門家の混合です (MoE) model with 128 experts (8 フォワードパスごとにアクティブになります), 高度なコード生成用に設計, リポジトリ規模の理解, およびエージェントツールの使用. Built on the...

による |7月 2025 |262K コンテキスト |$0.0700/M入力 |$0.2800/M出力
262K トークン

Qwen3-30B-A3B-Instruct-2507 は、Qwen の 30.5B パラメーターの専門家混合言語モデルです。, 推論ごとに 3.3B のアクティブなパラメータを使用. 非思考モードで動作し、次のような質の高い指導ができるように設計されています。, 多言語理解, and...

による |7月 2025 |262K コンテキスト |$0.0482/M入力 |$0.1931/M出力
262K トークン

GLM-4.5は最新のフラッグシップファウンデーションモデルです。, エージェントベースのアプリケーション専用に構築. 専門家の混合を活用します (MoE) アーキテクチャを備えており、最大 128,000 トークンのコンテキスト長をサポートします. GLM-4.5 delivers significantly...

による |7月 2025 |131K コンテキスト |$0.6000/M入力 |$2.20/M出力
131K トークン

GLM-4.5-Air は、当社の最新フラッグシップモデルファミリーの軽量バージョンです。, エージェント中心のアプリケーション向けにも設計されています. Like GLM-4.5, 専門家の混合を採用しています (MoE) architecture but with a more compact parameter...

による |7月 2025 |131K コンテキスト |$0.1300/M入力 |$0.8500/M出力
131K トークン

Qwen3-235B-A22B-Thinking-2507 は高性能です, 無差別級専門家混合 (MoE) 複雑な推論タスク用に最適化された言語モデル. フォワードパスごとに 235B パラメータのうち 22B をアクティブにし、最大でネイティブにサポートします。 262,144...

による |7月 2025 |262K コンテキスト |$0.2300/M入力 |$2.30/M出力
262K トークン

Qwen3-Coder-480B-A35B-Instruct は専門家の混合物です (MoE) Qwen チームによって開発されたコード生成モデル. 関数呼び出しなどのエージェントコーディングタスク用に最適化されています。, tool use, and long-context reasoning over...

による |7月 2025 |262K コンテキスト |$0.3000/M入力 |$1.00/M出力
262K トークン

UI-TARS-1.5 は、GUI ベースの環境に最適化されたマルチモーダル ビジョン言語エージェントです。, デスクトップインターフェースを含む, ウェブブラウザ, モバイルシステム, and games. バイトダンスによって構築されました, it builds upon the UI-TARS framework with reinforcement...

による |7月 2025 |128K コンテキスト |$0.1000/M入力 |$0.2000/M出力
128K トークン

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

による |7月 2025 |1M コンテキスト |$0.0500/M入力 |$0.2000/M出力
1M トークン

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

による |7月 2025 |1M コンテキスト |$0.1000/M入力 |$0.4000/M出力
1M トークン

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

による |7月 2025 |262K コンテキスト |$0.0900/M入力 |$0.5500/M出力
262K トークン

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...

による |7月 2025 |131K コンテキスト |$0.5700/M入力 |$2.30/M出力
131K トークン

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

による |7月 2025 |128K コンテキスト |$0.2000/M入力 |$0.9000/M出力
128K トークン

Hunyuan-A13B は 13B アクティブ パラメータです。 (MoE) テンセントが開発した言語モデル, 合計パラメータ数は 80B で、思考連鎖による推論のサポート. It offers competitive benchmark...

による |7月 2025 |131K コンテキスト |$0.1400/M入力 |$0.5700/M出力
131K トークン

Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: {instruction} {initial_code}...

による |7月 2025 |262K コンテキスト |$0.9000/M入力 |$1.90/M出力
262K トークン

Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: {instruction} {initial_code} {edit_snippet}...

による |7月 2025 |82K コンテキスト |$0.8000/M入力 |$1.20/M出力
82K トークン

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 シリーズ, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...

による |Jun 2025 |123K コンテキスト |$0.4200/M入力 |$1.25/M出力
123K トークン

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, バージョン 3.2 significantly improves accuracy on...

による |Jun 2025 |256K コンテキスト |$0.0938/M入力 |$0.2500/M出力
256K トークン

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

による |Jun 2025 |1M コンテキスト |$0.5500/M入力 |$2.20/M出力
1M トークン

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

による |Jun 2025 |1M コンテキスト |$0.3000/M入力 |$2.50/M出力
1M トークン

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

による |Jun 2025 |1M コンテキスト |$0.1500/M入力 |$1.25/M出力
1M トークン