AI Models

412 models Free & Paid 更新: 7 hours trước

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

经过 |2月 2025 |1M 上下文 |$0.2600/米输入 |$0.7800/米输出
1M代币

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

经过 |扬 2025 |200K 上下文 |$1.10/米输入 |$4.40/米输出
200K 代币

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

经过 |扬 2025 |200K 上下文 |$0.5500/米输入 |$2.20/米输出
200K 代币

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks. Released under the Apache 2.0 license, it features both pre-trained and instruction-tuned versions designed...

经过 |扬 2025 |33K 上下文 |$0.0500/米输入 |$0.0800/米输出

Sonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize sources. It is designed for companies seeking to integrate lightweight question-and-answer features...

经过 |扬 2025 |127K 上下文 |$1.00/米输入 |$1.00/米输出
127K 代币

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...

经过 |扬 2025 |8K 上下文 |$0.8000/米输入 |$0.8000/米输出

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

经过 |扬 2025 |64K 上下文 |$0.7000/米输入 |$2.50/米输出
64K 代币

最小最大-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...

经过 |扬 2025 |1M 上下文 |$0.2000/米输入 |$1.10/米输出
1M代币

[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion...

经过 |扬 2025 |16K 上下文 |$0.0700/米输入 |$0.1400/米输出

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...

经过 |十二月 2024 |164K 上下文 |$0.2574/米输入 |$1.03/米输出
164K 代币

Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.2](/models/sao10k/l3-euryale-70b).

经过 |十二月 2024 |131K 上下文 |$0.6500/米输入 |$0.7500/米输出
131K 代币

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

经过 |十二月 2024 |200K 上下文 |$7.50/米输入 |$30.00/米输出
200K 代币

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

经过 |十二月 2024 |200K 上下文 |$15.00/米输入 |$60.00/米输出
200K 代币

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG, 工具使用, agents, and similar tasks requiring complex reasoning...

经过 |十二月 2024 |128K 上下文 |$0.0375/米输入 |$0.1500/米输出
128K 代币

The Meta Llama 3.3 multilingual large language model (法学硕士) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

经过 |十二月 2024 |131K 上下文 |$0.1000/米输入 |$0.3200/米输出
131K 代币

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, 视频, and text inputs to generate text output. Amazon Nova Lite...

经过 |十二月 2024 |300K 上下文 |$0.0600/米输入 |$0.2400/米输出
300K 代币

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...

经过 |十二月 2024 |128K 上下文 |$0.0350/米输入 |$0.1400/米输出
128K 代币

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks. As of December...

经过 |十二月 2024 |300K 上下文 |$0.8000/米输入 |$3.20/米输出
300K 代币

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...

经过 |Nov 2024 |128K 上下文 |$2.50/米输入 |$10.00/米输出
128K 代币

This is Mistral AI's flagship model, 米斯特拉尔大号 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, 代码, JSON, 聊天, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

经过 |Nov 2024 |131K 上下文 |$2.00/米输入 |$6.00/米输出
131K 代币

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...

经过 |Nov 2024 |33K 上下文 |$0.6600/米输入 |$1.00/米输出

UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.

经过 |Nov 2024 |1M 上下文 |$0.4000/米输入 |$0.4000/米输出
1M代币

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

经过 |Oct 2024 |33K 上下文 |$3.00/米输入 |$5.00/米输出

Qwen2.5 7B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

经过 |Oct 2024 |33K 上下文 |$0.1000/米输入 |$0.2000/米输出

Rocinante 12B is designed for engaging storytelling and rich prose. Early testers have reported: - Expanded vocabulary with unique and expressive word choices - Enhanced creativity for vivid narratives -...

经过 |九月 2024 |66K 上下文 |$0.2500/米输入 |$0.5000/米输出
66K 代币

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

经过 |九月 2024 |131K 上下文 |$0.0500/米输入 |$0.3300/米输出
131K 代币

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

经过 |九月 2024 |60K 上下文 |$0.0270/米输入 |$0.2010/米输出
60K 代币

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

经过 |九月 2024 |33K 上下文 |$0.3600/米输入 |$0.4000/米输出

command-r-plus-08-2024 is an update of the [Command R+](/models/cohere/command-r-plus) with roughly 50% higher throughput and 25% lower latencies as compared to the previous Command R+ version, while keeping the hardware footprint...

经过 |八月 2024 |128K 上下文 |$2.50/米输入 |$10.00/米输出
128K 代币

command-r-08-2024 is an update of the [Command R](/models/cohere/command-r) with improved performance for multilingual retrieval-augmented generation (RAG) and tool use. More broadly, it is better at math, code and reasoning and...

经过 |八月 2024 |128K 上下文 |$0.1500/米输入 |$0.6000/米输出
128K 代币

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).

经过 |八月 2024 |131K 上下文 |$0.8500/米输入 |$0.8500/米输出
131K 代币

Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

经过 |八月 2024 |131K 上下文 |$0.7000/米输入 |$0.7000/米输出
131K 代币

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

经过 |八月 2024 |131K 上下文 |$1.00/米输入 |$1.00/米输出
131K 代币

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3. It's a strategic merge of multiple models, designed to balance creativity with improved logic and general knowledge....

经过 |八月 2024 |8K 上下文 |$0.0400/米输入 |$0.0500/米输出

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...

经过 |八月 2024 |128K 上下文 |$2.50/米输入 |$10.00/米输出
128K 代币

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

经过 |Jul 2024 |131K 上下文 |$0.0500/米输入 |$0.0800/米输出
131K 代币

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...

经过 |Jul 2024 |131K 上下文 |$0.4000/米输入 |$0.4000/米输出
131K 代币

A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...

经过 |Jul 2024 |131K 上下文 |$0.0190/米输入 |$0.0300/米输出
131K 代币

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

经过 |Jul 2024 |128K 上下文 |$0.1500/米输入 |$0.6000/米输出
128K 代币

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

经过 |Jul 2024 |128K 上下文 |$0.0750/米输入 |$0.3000/米输出
128K 代币

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

经过 |Jul 2024 |128K 上下文 |$0.1500/米输入 |$0.6000/米输出
128K 代币

Gemma 2 27B by Google is an open model built from the same research and technology used to create the [Gemini models](/models?q=gemini). Gemma models are well-suited for a variety of...

经过 |Jul 2024 |8K 上下文 |$0.6500/米输入 |$0.6500/米输出

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

经过 |可能 2024 |128K 上下文 |$1.25/米输入 |$5.00/米输出
128K 代币

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

经过 |可能 2024 |128K 上下文 |$5.00/米输入 |$15.00/米输出
128K 代币

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

经过 |可能 2024 |128K 上下文 |$2.50/米输入 |$10.00/米输出
128K 代币

Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...

经过 |4月 2024 |66K 上下文 |$2.00/米输入 |$6.00/米输出
66K 代币

WizardLM-2 8x22B is Microsoft AI's most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state-of-the-art opensource models. It is...

经过 |4月 2024 |66K 上下文 |$0.6200/米输入 |$0.6200/米输出
66K 代币

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

经过 |4月 2024 |128K 上下文 |$5.00/米输入 |$15.00/米输出
128K 代币

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

经过 |4月 2024 |128K 上下文 |$10.00/米输入 |$30.00/米输出
128K 代币

克洛德 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal

经过 |Mar 2024 |200K 上下文 |$0.2500/米输入 |$1.25/米输出
200K 代币