Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 バージョン. It...
AI Models
Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...
GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. インタラクティブな開発セッションと長時間の開発セッションの両方向けに設計されています。, independent execution of complex engineering tasks....
DeepSeek-V3.1 Terminus is an update to [ディープシーク V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...
Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...
Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal reasoning traces are separated...
Hermes 4 70B is a hybrid reasoning model from Nous Research, Meta-Llama-3.1-70B に基づいて構築. It introduces the same hybrid mode as the larger 405B release, allowing the model to either...
Hermes 4 Meta-Llama-3.1-405B に基づいて構築され、Nous Research によってリリースされた大規模推論モデルです。. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...
DeepSeek-V3.1 is a large hybrid reasoning model (671Bパラメータ, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
ジャンバ ラージ 1.7 is the latest model in the Jamba open family, offering improvements in grounding, 指示に従う, 全体的な効率. Built on a hybrid SSM-Transformer architecture with a 256K context...
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, ユーザーエクスペリエンスと. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, ユーザーエクスペリエンスと. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 ライセンス. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 ライセンス. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
Qwen3-Coder-30B-A3B-Instruct は 30.5B パラメータの専門家の混合です (MoE) model with 128 experts (8 フォワードパスごとにアクティブになります), 高度なコード生成用に設計, リポジトリ規模の理解, およびエージェントツールの使用. Built on the...
Qwen3-30B-A3B-Instruct-2507 は、Qwen の 30.5B パラメーターの専門家混合言語モデルです。, 推論ごとに 3.3B のアクティブなパラメータを使用. 非思考モードで動作し、次のような質の高い指導ができるように設計されています。, 多言語理解, and...
GLM-4.5は最新のフラッグシップファウンデーションモデルです。, エージェントベースのアプリケーション専用に構築. 専門家の混合を活用します (MoE) アーキテクチャを備えており、最大 128,000 トークンのコンテキスト長をサポートします. GLM-4.5 delivers significantly...
GLM-4.5-Air は、当社の最新フラッグシップモデルファミリーの軽量バージョンです。, エージェント中心のアプリケーション向けにも設計されています. Like GLM-4.5, 専門家の混合を採用しています (MoE) architecture but with a more compact parameter...
Qwen3-235B-A22B-Thinking-2507 は高性能です, 無差別級専門家混合 (MoE) 複雑な推論タスク用に最適化された言語モデル. フォワードパスごとに 235B パラメータのうち 22B をアクティブにし、最大でネイティブにサポートします。 262,144...
Qwen3-Coder-480B-A35B-Instruct は専門家の混合物です (MoE) Qwen チームによって開発されたコード生成モデル. 関数呼び出しなどのエージェントコーディングタスク用に最適化されています。, tool use, and long-context reasoning over...
UI-TARS-1.5 は、GUI ベースの環境に最適化されたマルチモーダル ビジョン言語エージェントです。, デスクトップインターフェースを含む, ウェブブラウザ, モバイルシステム, and games. バイトダンスによって構築されました, it builds upon the UI-TARS framework with reinforcement...
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...
Hunyuan-A13B は 13B アクティブ パラメータです。 (MoE) テンセントが開発した言語モデル, 合計パラメータ数は 80B で、思考連鎖による推論のサポート. It offers competitive benchmark...
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: {instruction} {initial_code}...
Morph's fastest apply model for code edits. ~10,500 tokens/sec with 96% accuracy for rapid code transformations. The model requires the prompt to be in the following format: {instruction} {initial_code} {edit_snippet}...
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 シリーズ, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, バージョン 3.2 significantly improves accuracy on...
MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...







