AIモデル

466 モデル Free & Paid Cập nhật: 49 minutes trước

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

による |4月 2026 |262K context |Miễn phí input |Miễn phí output
262K tokens ⓘ

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

による |4月 2026 |262K context |$0.0900/M input |$0.3400/M output
262K tokens ⓘ

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

による |4月 2026 |262K context |Miễn phí input |Miễn phí output
262K tokens ⓘ

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 シリーズ, it delivers...

による |4月 2026 |1M context |$0.3250/M input |$1.95/M output
1M tokens ⓘ

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, ビジョンベースのコーディングとエージェント駆動のタスク向けに構築. It natively handles image, ビデオ, そしてテキスト入力, excels at long-horizon planning, 複雑なコーディング,...

による |4月 2026 |203K context |$1.20/M input |$4.00/M output
203K tokens ⓘ

Trinity Large Thinking は、Arce AI チームによる強力なオープンソース推論モデルです。. It shows strong performance in PinchBench, エージェントのワークロード, そして推論タスク. 起動ビデオ: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...

による |4月 2026 |262K context |$0.2500/M input |$0.8000/M output
262K tokens ⓘ

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, エージェントベースのワークフロー. 複数のエージェントが並行して動作し、詳細な調査を実施します, 座標ツールの使用, and synthesize information...

による |3月 2026 |2M context |$1.25/M input |$2.50/M output
2M tokens ⓘ

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...

による |3月 2026 |2M context |$1.25/M input |$2.50/M output
2M tokens ⓘ

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...

による |3月 2026 |1M context |Miễn phí input |Miễn phí output
1M tokens ⓘ

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate...

による |3月 2026 |1M context |Miễn phí input |Miễn phí output
1M tokens ⓘ

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...

による |3月 2026 |16K context |$0.1000/M input |$0.1000/M output

MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...

による |3月 2026 |205K context |$0.2100/M input |$0.8400/M output
205K tokens ⓘ

GPT-5.4 nano は、GPT-5.4 ファミリの中で最も軽量でコスト効率の高いバージョンです。, スピードが重視される大量のタスク向けに最適化. It supports text and image inputs and is designed for low-latency...

による |3月 2026 |400K context |$0.2000/M input |$1.25/M output
400K tokens ⓘ

GPT-5.4 nano は、GPT-5.4 ファミリの中で最も軽量でコスト効率の高いバージョンです。, スピードが重視される大量のタスク向けに最適化. It supports text and image inputs and is designed for low-latency...

による |3月 2026 |400K context |$0.1000/M input |$0.6250/M output
400K tokens ⓘ

GPT-5.4 mini は、GPT-5.4 のコア機能をさらに高速化します。, 高スループットのワークロード向けに最適化された、より効率的なモデル. 推論全体で強力なパフォーマンスを備えたテキストと画像の入力をサポートします, coding,...

による |3月 2026 |400K context |$0.7500/M input |$4.50/M output
400K tokens ⓘ

GPT-5.4 mini は、GPT-5.4 のコア機能をさらに高速化します。, 高スループットのワークロード向けに最適化された、より効率的なモデル. 推論全体で強力なパフォーマンスを備えたテキストと画像の入力をサポートします, coding,...

による |3月 2026 |400K context |$0.3750/M input |$2.25/M output
400K tokens ⓘ

Mistral Small 4 Mistral Small ファミリーの次のメジャー リリースです, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

による |3月 2026 |262K context |$0.0750/M input |$0.3000/M output
262K tokens ⓘ

Mistral Small 4 Mistral Small ファミリーの次のメジャー リリースです, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

による |3月 2026 |262K context |$0.1500/M input |$0.6000/M output
262K tokens ⓘ

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

による |3月 2026 |203K context |$1.20/M input |$4.00/M output
203K tokens ⓘ

NVIDIA Nemotron 3 Super は 120B パラメータのオープンハイブリッド MoE モデルです, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

による |3月 2026 |262K context |Miễn phí input |Miễn phí output
262K tokens ⓘ

NVIDIA Nemotron 3 Super は 120B パラメータのオープンハイブリッド MoE モデルです, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

による |3月 2026 |262K context |$0.0800/M input |$0.4500/M output
262K tokens ⓘ

Seed-2.0-Lite は多用途です, 強力なマルチモーダル機能とエージェント機能を提供しながら、遅延を大幅に短縮する、コスト効率の高いエンタープライズ主力製品, making it a practical default choice for most production workloads across...

による |3月 2026 |262K context |$0.2500/M input |$2.00/M output
262K tokens ⓘ

Qwen3.5-9B は、Qwen3.5 ファミリのマルチモーダル基礎モデルです。, 強力な推論を提供するように設計されている, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

による |3月 2026 |262K context |$0.1000/M input |$0.1500/M output
262K tokens ⓘ

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

による |3月 2026 |1.1M context |$30.00/M input |$180.00/M output
1.1M tokens ⓘ

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

による |3月 2026 |1.1M context |$15.00/M input |$90.00/M output
1.1M tokens ⓘ

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

による |3月 2026 |1.1M context |$2.50/M input |$15.00/M output
1.1M tokens ⓘ

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

による |3月 2026 |1.1M context |$1.25/M input |$7.50/M output
1.1M tokens ⓘ

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

による |3月 2026 |128K context |$0.2500/M input |$0.7500/M output
128K tokens ⓘ

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

による |3月 2026 |1M context |$0.2500/M input |$1.50/M output
1M tokens ⓘ

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, 256k コンテキストをサポート, 4つの推論努力モード (最小/低/中/高), 多面的な理解,...

による |2月 2026 |262K context |$0.1000/M input |$0.4000/M output
262K tokens ⓘ

Qwen3.5 シリーズ 35B-A3B は、線形注意メカニズムと専門家のまばらな混合モデルを統合するハイブリッド アーキテクチャで設計されたネイティブ ビジョン言語モデルです。, より高い推論効率を実現. Its overall...

による |2月 2026 |262K context |$0.1500/M input |$1.00/M output
262K tokens ⓘ

Qwen3.5 27B ネイティブ ビジョン言語 Dense モデルには、リニア アテンション メカニズムが組み込まれています, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

による |2月 2026 |262K context |$0.1950/M input |$1.56/M output
262K tokens ⓘ

Qwen3.5 122B-A10B ネイティブ ビジョン言語モデルは、線形注意メカニズムと専門家混合モデルを統合したハイブリッド アーキテクチャに基づいて構築されています。, より高い推論効率を実現. In terms of...

による |2月 2026 |262K context |$0.2600/M input |$2.08/M output
262K tokens ⓘ

Qwen3.5 ネイティブ ビジョン言語 Flash モデルは、線形注意メカニズムと専門家のまばらな混合モデルを統合するハイブリッド アーキテクチャに基づいて構築されています。, より高い推論効率を実現. Compared to the...

による |2月 2026 |1M context |$0.0650/M input |$0.2600/M output
1M tokens ⓘ

GPT-5.3-Codex は OpenAI の最も高度なエージェント コーディング モデルです, GPT-5.2-Codex の最先端のソフトウェア エンジニアリング パフォーマンスと、GPT-5.2 のより広範な推論および専門知識機能を組み合わせたもの. It achieves state-of-the-art results...

による |2月 2026 |400K context |$1.75/M input |$14.00/M output
400K tokens ⓘ

Aion-2.0 は、没入型ロールプレイングとストーリーテリング用に最適化された DeepSeek V3.2 のバリアントです。. 緊張感をもたらすのに特に強い, 危機, そして物語への衝突, making narratives feel more engaging....

による |2月 2026 |131K context |$0.8000/M input |$1.60/M output
131K tokens ⓘ

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

による |2月 2026 |1M context |$1.00/M input |$6.00/M output
1M tokens ⓘ

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

による |2月 2026 |1M context |$2.00/M input |$12.00/M output
1M tokens ⓘ

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

による |2月 2026 |1M context |$1.50/M input |$7.50/M output
1M tokens ⓘ

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

による |2月 2026 |1M context |$3.00/M input |$15.00/M output
1M tokens ⓘ

Qwen3.5 ネイティブ ビジョン言語シリーズ Plus モデルは、線形注意メカニズムと専門家のまばらな混合モデルを統合するハイブリッド アーキテクチャに基づいて構築されています。, より高い推論効率を実現. In a variety of...

による |2月 2026 |1M context |$0.2600/M input |$1.56/M output
1M tokens ⓘ

Qwen3.5 シリーズ 397B-A17B ネイティブ ビジョン言語モデルは、線形注意メカニズムと専門家のまばらな混合モデルを統合するハイブリッド アーキテクチャに基づいて構築されています。, より高い推論効率を実現. It delivers...

による |2月 2026 |262K context |$0.5500/M input |$3.50/M output
262K tokens ⓘ

MiniMax-M2.5 は、現実世界の生産性を考慮して設計された SOTA 大規模言語モデルです. 現実世界の多様で複雑なデジタル作業環境でトレーニングを受けています, M2.5 builds upon the coding expertise of M2.1...

による |2月 2026 |205K context |$0.2700/M input |$1.08/M output
205K tokens ⓘ

GLM-5 は、複雑なシステム設計と長期的なエージェント ワークフロー向けに設計された、Z.ai の主力オープンソース基盤モデルです。. 専門開発者向けに構築, 大規模なプログラミング タスクで実稼働グレードのパフォーマンスを実現します。, rivaling leading...

による |2月 2026 |205K context |$0.6000/M input |$1.92/M output
205K tokens ⓘ

Qwen3-Max-Thinking は、Qwen3 シリーズの主力推論モデルです。, 深い知識を必要とする一か八かの認知タスク向けに設計されています, 多段階の推論. モデルの容量と強化学習のコンピューティングを大幅に拡張することにより、, it...

による |2月 2026 |262K context |$0.7800/M input |$3.90/M output
262K tokens ⓘ

オーパス 4.6 Anthropic のコーディングと長期にわたる専門的なタスクのための最強のモデルです. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

による |2月 2026 |1M context |$2.50/M input |$12.50/M output
1M tokens ⓘ

オーパス 4.6 Anthropic のコーディングと長期にわたる専門的なタスクのための最強のモデルです. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...

による |2月 2026 |1M context |$5.00/M input |$25.00/M output
1M tokens ⓘ

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...

による |2月 2026 |262K context |$0.1200/M input |$0.8000/M output
262K tokens ⓘ