AI Models

412 models Free & Paid 更新: 10 hours trước

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

经过 |Jul 2026 |500K 上下文 |$2.00/米输入 |$6.00/米输出
500K 代币

This model always redirects to the latest Grok model from xAI.

经过 |Jul 2026 |500K 上下文 |$2.00/米输入 |$6.00/米输出
500K 代币

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...

经过 |Jul 2026 |131K 上下文 |$0.7000/米输入 |$1.40/米输出
131K 代币

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...

经过 |Jul 2026 |131K 上下文 |$3.00/米输入 |$6.00/米输出
131K 代币

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B 活跃, 192 experts with top-8 routing) built for reasoning, 代理工作流程, and real-world production use. It supports a configurable reasoning effort:...

经过 |Jul 2026 |262K 上下文 |$0.1320/米输入 |$0.5280/米输出
262K 代币

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

经过 |Jul 2026 |262K 上下文 |自由输入 |自由输出
262K 代币

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

经过 |Jul 2026 |262K 上下文 |$0.0600/米输入 |$0.1200/米输出
262K 代币

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

经过 |六月 2026 |1M 上下文 |$2.00/米输入 |$10.00/米输出
1M代币

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

经过 |六月 2026 |1M 上下文 |$1.00/米输入 |$5.00/米输出
1M代币

Nano Banana 2 Lite (双子座 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation...

经过 |六月 2026 |66K 上下文 |$0.2500/米输入 |$1.50/米输出
66K 代币

Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series. It accepts text and image input and is built for coding, 工具使用,...

经过 |六月 2026 |262K 上下文 |$0.0250/米输入 |$0.1000/米输出
262K 代币

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

经过 |六月 2026 |1M 上下文 |$5.00/米输入 |$30.00/米输出
1M代币

双子座 3.1 Flash Image, a.k.a. “纳米香蕉2," 是 Google 最新最先进的图像生成和编辑模型, 以 Flash 速度提供专业级视觉质量. It combines advanced...

经过 |六月 2026 |131K 上下文 |$0.5000/米输入 |$3.00/米输出
131K 代币

Nano Banana Pro 是 Google 最先进的图像生成和编辑模型, 建立在双子座之上 3 Pro. 它扩展了原始 Nano Banana,显着改进了多模态推理, 现实世界的接地, and...

经过 |六月 2026 |131K 上下文 |$2.00/米输入 |$12.00/米输出
131K 代币

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...

经过 |六月 2026 |256K 上下文 |自由输入 |自由输出
256K 代币

广义线性模型 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

经过 |六月 2026 |1M 上下文 |$1.19/米输入 |$3.74/米输出
1M代币

广义线性模型 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

经过 |六月 2026 |512K 上下文 |$0.7000/米输入 |$2.20/米输出
512K 代币

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...

经过 |六月 2026 |1M 上下文 |自由输入 |自由输出
1M代币

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

经过 |六月 2026 |262K 上下文 |$0.7100/米输入 |$3.50/米输出
262K 代币

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

经过 |六月 2026 |262K 上下文 |$0.4750/米输入 |$2.00/米输出
262K 代币

This model always redirects to the latest model in the Claude Fable family.

经过 |六月 2026 |1M 上下文 |$10.00/米输入 |$50.00/米输出
1M代币

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, 图像, and file inputs with text output, with reasoning support and...

经过 |六月 2026 |1M 上下文 |$10.00/米输入 |$50.00/米输出
1M代币

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, 图像, and file inputs with text output, with reasoning support and...

经过 |六月 2026 |1M 上下文 |$5.00/米输入 |$25.00/米输出
1M代币

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...

经过 |六月 2026 |262K 上下文 |$0.2500/米输入 |$1.00/米输出
262K 代币

NVIDIA 神经元 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

经过 |六月 2026 |128K 上下文 |自由输入 |自由输出
128K 代币

NVIDIA 神经元 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (教育部). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

经过 |六月 2026 |1M 上下文 |自由输入 |自由输出
1M代币

NVIDIA 神经元 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (教育部). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

经过 |六月 2026 |512K 上下文 |$0.3000/米输入 |$1.80/米输出
512K 代币

NVIDIA 神经元 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (教育部). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

经过 |六月 2026 |512K 上下文 |$0.6000/米输入 |$3.60/米输出
512K 代币

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...

经过 |六月 2026 |1M 上下文 |$0.3200/米输入 |$1.28/米输出
1M代币

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, 图像, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

经过 |可能 2026 |524K 上下文 |$0.1500/米输入 |$0.6000/米输出
524K 代币

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, 图像, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

经过 |可能 2026 |1M 上下文 |$0.3000/米输入 |$1.20/米输出
1M代币

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

经过 |可能 2026 |262K 上下文 |$0.2000/米输入 |$1.15/米输出
262K 代币

Fast-mode variant of [Opus 4.8](/anthropic/claude-opus-4.8) - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

经过 |可能 2026 |1M 上下文 |$10.00/米输入 |$50.00/米输出
1M代币

近距离工作 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, 图像, and file inputs with text output, with reasoning support and a 1M-token...

经过 |可能 2026 |1M 上下文 |$2.50/米输入 |$12.50/米输出
1M代币

近距离工作 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, 图像, and file inputs with text output, with reasoning support and a 1M-token...

经过 |可能 2026 |1M 上下文 |$5.00/米输入 |$25.00/米输出
1M代币

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...

经过 |可能 2026 |1M 上下文 |$1.48/米输入 |$4.43/米输出
1M代币

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding...

经过 |可能 2026 |256K 上下文 |$1.00/米输入 |$2.00/米输出
256K 代币

双子座 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

经过 |可能 2026 |1M 上下文 |$1.50/米输入 |$9.00/米输出
1M代币

双子座 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

经过 |可能 2026 |1M 上下文 |$0.7500/米输入 |$4.50/米输出
1M代币

Fast-mode variant of [Opus 4.7](/anthropic/claude-opus-4.7) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

经过 |可能 2026 |1M 上下文 |$30.00/米输入 |$150.00/米输出
1M代币

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

经过 |可能 2026 |33K 上下文 |$0.1500/米输入 |$1.50/米输出

Ring-2.6-1T is a 1T-parameter-scale thinking model with 63B active parameters, built for real-world agent workflows that require both strong capability and operational efficiency. It is optimized for coding agents, tool...

经过 |可能 2026 |262K 上下文 |$0.0750/米输入 |$0.6250/米输出
262K 代币

双子座 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, 图像, 视频, audio, and PDF inputs, and is designed for lightweight agentic...

经过 |可能 2026 |1M 上下文 |$0.2500/米输入 |$1.50/米输出
1M代币

双子座 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, 图像, 视频, audio, and PDF inputs, and is designed for lightweight agentic...

经过 |可能 2026 |1M 上下文 |$0.1250/米输入 |$0.7500/米输出
1M代币

GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...

经过 |可能 2026 |400K 上下文 |$5.00/米输入 |$30.00/米输出
400K 代币

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

经过 |4月 2026 |1M 上下文 |$1.25/米输入 |$2.50/米输出
1M代币

Granite 4.1 8B is a dense, decoder-only 8-billion-parameter language model from IBM, part of the Granite 4.1 family. It supports a 131K-token context window and is designed for enterprise tasks...

经过 |4月 2026 |131K 上下文 |$0.0500/米输入 |$0.1000/米输出
131K 代币

米斯特拉尔介质 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...

经过 |4月 2026 |262K 上下文 |$1.50/米输入 |$7.50/米输出
262K 代币

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, 图像, 视频, and...

经过 |4月 2026 |256K 上下文 |自由输入 |自由输出
256K 代币

This model always redirects to the latest model in the Anthropic Claude Haiku family.

经过 |4月 2026 |200K 上下文 |$1.00/米输入 |$5.00/米输出
200K 代币