AI Models

412 models Free & Paid 更新: 7 hours trước

双子座 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

经过 |六月 2025 |1M 上下文 |$1.25/米输入 |$10.00/米输出
1M代币

双子座 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

经过 |六月 2025 |1M 上下文 |$0.6250/米输入 |$5.00/米输出
1M代币

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

经过 |六月 2025 |200K 上下文 |$20.00/米输入 |$80.00/米输出
200K 代币

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

经过 |六月 2025 |200K 上下文 |$10.00/米输入 |$40.00/米输出
200K 代币

双子座 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

经过 |六月 2025 |1M 上下文 |$1.25/米输入 |$10.00/米输出
1M代币

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

经过 |可能 2025 |164K 上下文 |$0.5000/米输入 |$2.15/米输出
164K 代币

近距离工作 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

经过 |可能 2025 |200K 上下文 |$15.00/米输入 |$75.00/米输出
200K 代币

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...

经过 |可能 2025 |1M 上下文 |$3.00/米输入 |$15.00/米输出
1M代币

Gemma 3n E4B-it is optimized for efficient execution on mobile and low-resource devices, such as phones, laptops, and tablets. It supports multimodal inputs—including text, visual data, and audio—enabling diverse tasks...

经过 |可能 2025 |33K 上下文 |$0.0600/米输入 |$0.1200/米输出

米斯特拉尔介质 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

经过 |可能 2025 |131K 上下文 |$0.4000/米输入 |$2.00/米输出
131K 代币

双子座 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

经过 |可能 2025 |1M 上下文 |$1.25/米输入 |$10.00/米输出
1M代币

Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning, creative writing and enterprise QA. Unlike many 70 B peers, it retains the 128 k...

经过 |可能 2025 |131K 上下文 |$0.7500/米输入 |$1.20/米输出
131K 代币

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

经过 |4月 2025 |1M 上下文 |$0.1800/米输入 |$0.1800/米输出
1M代币

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (教育部) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...

经过 |4月 2025 |131K 上下文 |$0.1300/米输入 |$0.5200/米输出
131K 代币

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

经过 |4月 2025 |131K 上下文 |$0.1170/米输入 |$0.4550/米输出
131K 代币

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

经过 |4月 2025 |131K 上下文 |$0.1200/米输入 |$0.2400/米输出
131K 代币

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...

经过 |4月 2025 |131K 上下文 |$0.0800/米输入 |$0.2800/米输出
131K 代币

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (教育部) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

经过 |4月 2025 |131K 上下文 |$0.4550/米输入 |$1.82/米输出
131K 代币

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

经过 |4月 2025 |200K 上下文 |$1.10/米输入 |$4.40/米输出
200K 代币

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

经过 |4月 2025 |200K 上下文 |$0.5500/米输入 |$2.20/米输出
200K 代币

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

经过 |4月 2025 |200K 上下文 |$1.00/米输入 |$4.00/米输出
200K 代币

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

经过 |4月 2025 |200K 上下文 |$2.00/米输入 |$8.00/米输出
200K 代币

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

经过 |4月 2025 |200K 上下文 |$0.5500/米输入 |$2.20/米输出
200K 代币

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

经过 |4月 2025 |200K 上下文 |$1.10/米输入 |$4.40/米输出
200K 代币

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

经过 |4月 2025 |1M 上下文 |$1.00/米输入 |$4.00/米输出
1M代币

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

经过 |4月 2025 |1M 上下文 |$2.00/米输入 |$8.00/米输出
1M代币

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

经过 |4月 2025 |1M 上下文 |$0.2000/米输入 |$0.8000/米输出
1M代币

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

经过 |4月 2025 |1M 上下文 |$0.4000/米输入 |$1.60/米输出
1M代币

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

经过 |4月 2025 |1M 上下文 |$0.0500/米输入 |$0.2000/米输出
1M代币

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

经过 |4月 2025 |1M 上下文 |$0.1000/米输入 |$0.4000/米输出
1M代币

Llama 4 Maverick 17B Instruct (128乙) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (教育部) architecture with 128 experts and 17 billion active parameters per forward...

经过 |4月 2025 |1M 上下文 |$0.2000/米输入 |$0.8000/米输出
1M代币

Llama 4 Scout 17B Instruct (16乙) is a mixture-of-experts (教育部) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

经过 |4月 2025 |1.3M 上下文 |$0.1000/米输入 |$0.3000/米输出
1.3M代币

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...

经过 |Mar 2025 |164K 上下文 |$0.2700/米输入 |$1.12/米输出
164K 代币

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...

经过 |Mar 2025 |200K 上下文 |$150.00/米输入 |$600.00/米输出
200K 代币

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...

经过 |Mar 2025 |200K 上下文 |$75.00/米输入 |$300.00/米输出
200K 代币

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities. It provides state-of-the-art performance in text-based reasoning and...

经过 |Mar 2025 |128K 上下文 |$0.3510/米输入 |$0.5550/米输出
128K 代币

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

经过 |Mar 2025 |131K 上下文 |$0.0500/米输入 |$0.1000/米输出
131K 代币

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

经过 |Mar 2025 |131K 上下文 |$0.0500/米输入 |$0.1500/米输出
131K 代币

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases. Compared to other leading proprietary...

经过 |Mar 2025 |256K 上下文 |$2.50/米输入 |$10.00/米输出
256K 代币

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function calling. Featuring a...

经过 |Mar 2025 |66K 上下文 |$0.1000/米输入 |$0.2000/米输出
66K 代币

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

经过 |Mar 2025 |262K 上下文 |$0.0800/米输入 |$0.4500/米输出
262K 代币

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.

经过 |Mar 2025 |33K 上下文 |$0.5500/米输入 |$0.8000/米输出

笔记: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro) Sonar Reasoning Pro is a premier reasoning model powered by DeepSeek R1 with Chain of Thought (CoT). Designed for...

经过 |Mar 2025 |128K 上下文 |$2.00/米输入 |$8.00/米输出
128K 代币

笔记: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro) For enterprises seeking more advanced capabilities, the Sonar Pro API can handle in-depth, multi-step queries with added extensibility, like...

经过 |Mar 2025 |200K 上下文 |$3.00/米输入 |$15.00/米输出
200K 代币

Sonar Deep Research is a research-focused model designed for multi-step retrieval, synthesis, and reasoning across complex topics. It autonomously searches, reads, and evaluates sources, refining its approach as it gathers...

经过 |Mar 2025 |128K 上下文 |$2.00/米输入 |$8.00/米输出
128K 代币

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...

经过 |2月 2025 |33K 上下文 |$0.2000/米输入 |$0.6000/米输出

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

经过 |2月 2025 |200K 上下文 |$1.10/米输入 |$4.40/米输出
200K 代币

OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...

经过 |2月 2025 |200K 上下文 |$0.5500/米输入 |$2.20/米输出
200K 代币

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

经过 |2月 2025 |33K 上下文 |$0.8000/米输入 |$1.60/米输出

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

经过 |2月 2025 |128K 上下文 |$0.8000/米输入 |$1.00/米输出
128K 代币