knowledge-base/wiki/concepts/llm/sonnet-5-anthropic.md
Hector 8f676acfa7 ingest(youtube): Matthew Berman - 90% less AI costs via Model Routing
- raw: raw/youtube/2026-07-07_berman-model-routing.md
- wiki/tools/model-routing.md (new) — Cost-saving routing patterns
- wiki/architecture/model-routing.md (update) — Berman patterns section
- wiki/index.md — 67. Update with new tool + architecture update
- wiki/log.md — changelog entry
2026-07-07 10:31:37 +02:00

65 lines
No EOL
3 KiB
Markdown
Raw Permalink Blame History

This file contains ambiguous Unicode characters

This file contains Unicode characters that might be confused with other characters. If you think that this is intentional, you can safely ignore this warning. Use the Escape button to reveal them.

---
created: 2026-07-05
updated: 2026-07-05
sources: [youtube/2026-07-05_everlast-ki-news-china-roboter-sonnet5-fable5.md, wiki/tools/anthropic-claude.md]
tags: [concept, llm, anthropic, sonnet-5, claude, cost-inefficiency, default-model, 1m-context]
---
# Claude Sonnet 5 — Anthropic's neues Default-Modell mit Cost-Inefficiency-Problem
> **TL;DR:** Anthropic lanciert Claude Sonnet 5 als neues Default-Modell für alle gratis/Pro User — 1M Kontextfenster, Performance auf Opus 4.8 Niveau. Doch im Cost-to-Run ist Sonnet 5 teurer als Fable 5 und damit "enorm ineffizient" — der ganze Sinn eines leichtgewichtigeren Modells wird widerlegt.
## Spezifikationen
| Eigenschaft | Wert |
|-------------|------|
| **Vendor** | Anthropic |
| **Kontextfenster** | 1 Million Tokens |
| **Performance** | ~Opus 4.8 Niveau ("vielleicht ein bisschen günstiger") |
| **Default-Modell** | Ja — für alle gratis und Pro User |
| **Claude Code** | Verfügbar |
| **Pricing** | Standardpricing |
| **Sicherheit** | Offiziell "sicherer als Sonnet 4.6" |
| **Versionssprung** | 4.6 → 5 (direkt, größer als üblich) |
## Verbesserungen (laut Anthropic)
- Reasoning
- Tool Use
- Coding
- Wissensarbeit
## Cost-to-Run Problem
**Artificial Analysis Cost to Run Index:** Sonnet 5 bei ~$6.000 — **teurer als Claude Fable 5**.
| Modell | Cost to Run | vs Sonnet 5 |
|--------|------------|-------------|
| **Sonnet 5** | ~$6.000 | — |
| **Fable 5** | niedriger | günstiger |
| **GPT 5.5 Extra Hype** | <50% von Sonnet 5 | ~halbe Kosten |
**Kritik (Leo Schmedding):**
- "Man darf nicht auf offizielle Kostenangaben schauen"
- Reasoning-Prozess ist ineffizient viele Output-Tokens reale Kosten viel höher
- "Das ist exakt nicht der Fall bei Sonnet 5" der Sinn eines Sonnet-Modells (leichtgewichtiger, günstiger) wird widerlegt
- "Der ganze Sinn und Zweck eines Modells wird widersinnig"
## Einordnung
Sonnet 5 richtet sich als Default-Modell an den Massenmarkt, scheitert aber an der Cost-Efficiency-Premise. Wer Preis-Leistung sucht, ist mit GPT 5.5 (~halbe Kosten) besser bedient. Wer Frontier-Qualität will, nutzt Fable 5 (das im Cost-to-Run günstiger ist als Sonnet 5 eine Ironie).
**Parallele zu [[fable-5-anthropic.md]]:** Auch Fable 5 hatte ein Pricing-Problem (39× teurer als GLM 5.2 im atomic.chat Benchmark). Sonnet 5 hat das umgekehrte Problem: Es sollte das "günstige" Modell sein, ist aber teurer als das Premium-Modell Fable 5.
## Cross-References
- [[../../tools/anthropic-claude.md]] Anthropic Claude Modellübersicht
- [[fable-5-anthropic.md]] Fable 5 (günstiger im Cost-to-Run als Sonnet 5)
- [[coding-benchmark-price-performance.md]] atomic.chat Benchmark
- [[chinese-model-cost-routing.md]] Cost-Routing-These (chinesische Modelle als Alternative)
- [[../../institutions/anthropic.md]] Anthropic Institutionenseite
## External Sources
- [YouTube: Everlast AI — KI-News vom 05.07.2026](https://www.youtube.com/watch?v=-JCCcR9qtYQ)
- [Anthropic](https://www.anthropic.com)