KI-Modelle

466 Modelle Frei & Paid Cập nhật: 2 hours trước

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...

by |Jul 2026 |1M context |$0.0300/M input |$0.1300/M output
1M tokens ⓘ

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

by |Jul 2026 |1M context |$5.00/M input |$25.00/M output
1M tokens ⓘ

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

by |Jul 2026 |1M context |$2.50/M input |$12.50/M output
1M tokens ⓘ

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

by |Jul 2026 |262K context |$0.0210/M input |$0.0630/M output
262K tokens ⓘ

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

by |Jul 2026 |1M context |$0.0900/M input |$0.1800/M output
1M tokens ⓘ

Laguna S 2.1 is the latest coding agent model from [Poolside](). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

by |Jul 2026 |262K context |Miễn phí input |Miễn phí output
262K tokens ⓘ

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

by |Jul 2026 |1M context |$0.3750/M input |$1.88/M output
1M tokens ⓘ

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

by |Jul 2026 |1M context |$0.7500/M input |$3.75/M output
1M tokens ⓘ

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

by |Jul 2026 |1M context |$0.3000/M input |$2.50/M output
1M tokens ⓘ

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

by |Jul 2026 |1M context |$0.1500/M input |$1.25/M output
1M tokens ⓘ

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...

by |Jul 2026 |1M context |$0.3000/M input |$1.20/M output
1M tokens ⓘ

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

by |Jul 2026 |524K context |$0.9500/M input |$4.05/M output
524K tokens ⓘ

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

by |Jul 2026 |1M context |Miễn phí input |Miễn phí output
1M tokens ⓘ

The experimental version of our Auto Router where we test new improvements. Use it to get the latest and greatest version of our general purpose auto router, but expect beta...

by |Jul 2026 |2M context |Miễn phí input |Miễn phí output
2M tokens ⓘ

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

by |Jul 2026 |1M context |$2.70/M input |$13.50/M output
1M tokens ⓘ

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

by |Jul 2026 |1M context |$2.28/M input |$11.40/M output
1M tokens ⓘ

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, Video, and PDF documents and returns text, with a 1M-token context window....

by |Jul 2026 |1M context |$1.25/M input |$4.25/M output
1M tokens ⓘ

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

by |Jul 2026 |262K context |$0.7400/M input |$2.96/M output
262K tokens ⓘ

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. **Cost note:** pro mode spends far more...

by |Jul 2026 |1.1M context |$0.2000/M input |$1.20/M output
1.1M tokens ⓘ

GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. **Cost note:** pro mode spends far more...

by |Jul 2026 |1.1M context |$0.1000/M input |$0.6000/M output
1.1M tokens ⓘ

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

by |Jul 2026 |1.1M context |$0.2000/M input |$1.20/M output
1.1M tokens ⓘ

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

by |Jul 2026 |1.1M context |$0.1000/M input |$0.6000/M output
1.1M tokens ⓘ

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. **Cost note:** pro mode spends far more...

by |Jul 2026 |1.1M context |$2.00/M input |$12.00/M output
1.1M tokens ⓘ

GPT-5.6 Terra Pro is the same underlying model as [GPT-5.6 Terra](https://openrouter.ai/openai/gpt-5.6-terra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. **Cost note:** pro mode spends far more...

by |Jul 2026 |1.1M context |$1.00/M input |$6.00/M output
1.1M tokens ⓘ

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

by |Jul 2026 |1.1M context |$2.00/M input |$12.00/M output
1.1M tokens ⓘ

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

by |Jul 2026 |1.1M context |$1.00/M input |$6.00/M output
1.1M tokens ⓘ

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. **Cost note:** pro mode spends far more...

by |Jul 2026 |1.1M context |$4.00/M input |$20.00/M output
1.1M tokens ⓘ

GPT-5.6 Sol Pro is the same underlying model as [GPT-5.6 Sol](https://openrouter.ai/openai/gpt-5.6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. **Cost note:** pro mode spends far more...

by |Jul 2026 |1.1M context |$1.00/M input |$5.00/M output
1.1M tokens ⓘ

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

by |Jul 2026 |1.1M context |$2.00/M input |$10.00/M output
1.1M tokens ⓘ

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

by |Jul 2026 |1.1M context |$1.00/M input |$5.00/M output
1.1M tokens ⓘ

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

by |Jul 2026 |500K context |$2.00/M input |$6.00/M output
500K tokens ⓘ

This model always redirects to the latest Grok model from xAI.

by |Jul 2026 |500K context |$2.00/M input |$6.00/M output
500K tokens ⓘ

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...

by |Jul 2026 |131K context |$0.7000/M input |$1.40/M output
131K tokens ⓘ

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute...

by |Jul 2026 |131K context |$3.00/M input |$6.00/M output
131K tokens ⓘ

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

by |Jul 2026 |262K context |$0.0825/M input |$0.3300/M output
262K tokens ⓘ

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

by |Jul 2026 |262K context |Miễn phí input |Miễn phí output
262K tokens ⓘ

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

by |Jul 2026 |262K context |$0.0600/M input |$0.1200/M output
262K tokens ⓘ

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

by |Jun 2026 |1M context |$1.00/M input |$5.00/M output
1M tokens ⓘ

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

by |Jun 2026 |1M context |$2.00/M input |$10.00/M output
1M tokens ⓘ

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation...

by |Jun 2026 |66K context |$0.2500/M input |$1.50/M output
66K tokens ⓘ

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

by |Jun 2026 |1M context |$5.00/M input |$30.00/M output
1M tokens ⓘ

Gemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines advanced...

by |Jun 2026 |131K context |$0.5000/M input |$3.00/M output
131K tokens ⓘ

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...

by |Jun 2026 |131K context |$2.00/M input |$12.00/M output
131K tokens ⓘ

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...

by |Jun 2026 |256K context |Miễn phí input |Miễn phí output
256K tokens ⓘ

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

by |Jun 2026 |1M context |$0.4100/M input |$3.99/M output
1M tokens ⓘ

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a...

by |Jun 2026 |1M context |Miễn phí input |Miễn phí output
1M tokens ⓘ

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

by |Jun 2026 |262K context |$0.6712/M input |$3.35/M output
262K tokens ⓘ

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, Bild, and file inputs with text output, with reasoning support and...

by |Jun 2026 |1M context |$5.00/M input |$25.00/M output
1M tokens ⓘ

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, Bild, and file inputs with text output, with reasoning support and...

by |Jun 2026 |1M context |$10.00/M input |$50.00/M output
1M tokens ⓘ