AI Models

412 models Free & Paid 업데이트: 5 hours trước

엔비디아 네모트론 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

~에 의해 |3 월 2026 |262K context |Miễn phí input |Miễn phí output
262K tokens

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...

~에 의해 |3 월 2026 |262K context |$0.2500/M input |$2.00/M output
262K tokens

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

~에 의해 |3 월 2026 |262K context |$0.1000/M input |$0.1500/M output
262K tokens

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

~에 의해 |3 월 2026 |1.1M context |$30.00/M input |$180.00/M output
1.1M tokens

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...

~에 의해 |3 월 2026 |1.1M context |$15.00/M input |$90.00/M output
1.1M tokens

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

~에 의해 |3 월 2026 |1.1M context |$2.50/M input |$15.00/M output
1.1M tokens

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...

~에 의해 |3 월 2026 |1.1M context |$1.25/M input |$7.50/M output
1.1M tokens

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

~에 의해 |3 월 2026 |128K context |$0.2500/M input |$0.7500/M output
128K tokens

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

~에 의해 |3 월 2026 |1M context |$0.2500/M input |$1.50/M output
1M tokens

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...

~에 의해 |2월 2026 |262K context |$0.1000/M input |$0.4000/M output
262K tokens

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...

~에 의해 |2월 2026 |66K context |$0.5000/M input |$3.00/M output
66K tokens

Qwen3.5 시리즈 35B-A3B는 선형 주의 메커니즘과 희소 전문가 혼합 모델을 통합하는 하이브리드 아키텍처로 설계된 기본 비전 언어 모델입니다., 더 높은 추론 효율성 달성. Its overall...

~에 의해 |2월 2026 |262K context |$0.2250/M input |$1.80/M output
262K tokens

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...

~에 의해 |2월 2026 |262K context |$0.1950/M input |$1.56/M output
262K tokens

Qwen3.5 122B-A10B 기본 비전 언어 모델은 선형 주의 메커니즘과 희박한 전문가 혼합 모델을 통합하는 하이브리드 아키텍처를 기반으로 구축되었습니다., 더 높은 추론 효율성 달성. In terms of...

~에 의해 |2월 2026 |262K context |$0.2900/M input |$2.40/M output
262K tokens

Qwen3.5 기본 비전 언어 플래시 모델은 선형 주의 메커니즘과 희박한 전문가 혼합 모델을 통합하는 하이브리드 아키텍처를 기반으로 구축되었습니다., 더 높은 추론 효율성 달성. Compared to the...

~에 의해 |2월 2026 |1M context |$0.0650/M input |$0.2600/M output
1M tokens

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...

~에 의해 |2월 2026 |1M context |$2.00/M input |$12.00/M output
1M tokens

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...

~에 의해 |2월 2026 |400K context |$1.75/M input |$14.00/M output
400K tokens

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....

~에 의해 |2월 2026 |131K context |$0.8000/M input |$1.60/M output
131K tokens

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

~에 의해 |2월 2026 |1M context |$2.00/M input |$12.00/M output
1M tokens

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...

~에 의해 |2월 2026 |1M context |$1.00/M input |$6.00/M output
1M tokens

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

~에 의해 |2월 2026 |1M context |$3.00/M input |$15.00/M output
1M tokens

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...

~에 의해 |2월 2026 |1M context |$1.50/M input |$7.50/M output
1M tokens