AI Pricing Made Simple
How AI models bill — tokens, input vs. output, hidden cost drivers and three levers to save. With price table and worked examples.
Konkrete KI-Modelle und Modellfamilien im Überblick — wer sie entwickelt, wie sie sich in Reasoning, Coding, Multimodalität und Preis unterscheiden und wofür die einzelnen Linien im Alltag am besten taugen. Gegliedert nach Anbieter-Familien (GPT, Claude, Gemini …) und übergreifenden Einsatzzwecken wie Security, Bild & Video oder Open-Weight.
How AI models bill — tokens, input vs. output, hidden cost drivers and three levers to save. With price table and worked examples.
Claude Fable is Anthropic's top reasoning line — above Opus, pricier, slower. Positioning, where it fits and how to choose it.
Claude Haiku is Anthropic's fast, low-cost line for high volume. Positioning, where it fits and how to choose it.
Claude Opus is Anthropic's pro tier for coding, agents and knowledge work — what the line stands for, where it fits and how to pick it.
Claude Sonnet 5.5: Anthropic's mid tier next to Opus 5.5. Benchmarks, effort-based cost, breaking changes and how to choose vs Opus 5.5 and GPT-6.1 Sol.
Claude Sonnet is Anthropic's balanced line: more capable than Haiku, cheaper than Opus. Positioning, where it fits and how to choose it.
Codestral is Mistral AI's code specialist: autocomplete, function generation, fill-in-the-middle. What the line is for, how it differs, when to pick it.
The major model families in 2026 at a glance. Who builds Claude, GPT, Gemini, Llama, Mistral, DeepSeek, Qwen — and which model to pick when.
Daybreak Blue is not its own AI model, but OpenAI's vetted access to Sol with fewer security refusals for authorized defenders.
DeepSeek V as an open model line: positioning, use profile and how to choose. Flagship DeepSeek V4-Pro — cheap and self-hostable.
DeepSeek V4.1-Flash: rank 1 on coding under 1h, sharp drop on agentic work over 1h. Prices, benchmarks and an honest read in the lexicon.
FLUX.2 by Black Forest Labs (Freiburg): open-weight image model, variants from 4B to 32B, licenses, hardware needs and how it stacks up in the market.
Gemini Flash is Google's fast, low-cost model line: high throughput and native multimodality. Positioning, use profile and a selection guide.
Gemini Pro is Google's reasoning spearhead: the deepest tier of the Gemini family for complex work. Positioning, use profile and a selection guide.
GLM by Zhipu AI as an open line: positioning, use profile and how to choose. Flagship GLM-5.3 with strong coding performance.
GPT Luna is the entry tier of the GPT-5.6 family: fast and cheap for high volume. Positioning, use profile and a selection guide in one lexicon entry.
GPT Sol is OpenAI's flagship line in the GPT-5.6 family: the deepest reasoning and coding. Positioning, use profile and a selection guide in one entry.
GPT Terra is the middle tier of the GPT-5.6 family: balanced across capability, speed and price. Positioning, use profile and a selection guide.
GPT-6 Astra is here: released September 2026. All the info on computer use, the Critical risk tier, reasoning, strengths & limits — in one lexicon entry.
GPT-6.1 Sol: near Astra performance at a fifth of the cost. Benchmarks, pricing, limits and a selection guide against Opus 5.5 and Sonnet 5.5.
Grok by xAI as a model line: positioning, use profile and how to choose. Current flagship Grok 4.6 with real-time access to X.
Kimi by Moonshot AI as an open line: positioning, use profile and how to choose. Flagship Kimi K3 with a very large context window.
Claude, GPT, Gemini, Llama & co. — who builds what, where each family shines, and how to pick the right model for your own use case.
Magistral is Mistral AI's reasoning line: a transparent chain of thought, strong in European languages, partly open weights. When to pick it.
Muse Spark is Metas first frontier model from Superintelligence Labs: closed-weight, thought compression, multimodal. Facts, benchmarks and where it fits.
Google's image model nicknamed Nano Banana — what hides behind the codename, what it can do, and how it differs from FLUX.2 and other image models.
Qwen by Alibaba as a line: open dense/MoE models plus a proprietary Max line. Flagship Qwen 3.8-Max — positioning and how to choose.