From fd55139269cce03d44b0383f893dbd68e0f844b8 Mon Sep 17 00:00:00 2001 From: Hector Date: Thu, 13 Aug 2026 17:13:40 +0200 Subject: [PATCH] ingest(raw): deepseek-api-preiserhoehung --- .../2026-08-13_deepseek-api-preiserhoehung.md | 43 +++++++++++++++++++ wiki/concepts/chinese-ai-wave-july-2026.md | 12 ++++++ wiki/concepts/llm/llm-model-catalog.md | 4 +- wiki/index.md | 7 +-- wiki/log.md | 12 ++++++ ...lama-cloud-deepseek-v4-flash-200tps-zdr.md | 17 ++++++++ 6 files changed, 90 insertions(+), 5 deletions(-) create mode 100644 raw/other/2026-08-13_deepseek-api-preiserhoehung.md diff --git a/raw/other/2026-08-13_deepseek-api-preiserhoehung.md b/raw/other/2026-08-13_deepseek-api-preiserhoehung.md new file mode 100644 index 0000000..ec124a4 --- /dev/null +++ b/raw/other/2026-08-13_deepseek-api-preiserhoehung.md @@ -0,0 +1,43 @@ +--- +type: other +source_url: https://www.perplexity.ai/discover/you/deepseek-to-raise-api-prices-u-bo6gHsYdTuy3WR3nKpo3_w +retrieved: 2026-08-13 +title: "DeepSeek erhöht API-Preise (V4 Pro + V4 Flash) ab 17.08.2026" +tags: [deepseek, deepseek-v4-pro, deepseek-v4-flash, pricing, api, chinese-ai, cost-routing] +--- + +# DeepSeek erhöht API-Preise (V4 Pro + V4 Flash) ab 17.08.2026 + +## Quelle & Verifikations-Hinweis + +- **Gepostet von:** Pit Weber, OME-Gruppe, Topic 13 "News & Infos", 2026-08-13 +- **Perplexity-Discover-Link:** https://www.perplexity.ai/discover/you/deepseek-to-raise-api-prices-u-bo6gHsYdTuy3WR3nKpo3_w +- ⚠️ **Der Perplexity-Link liefert per web_fetch nur 403 (Cloudflare).** Die Fakten unten stammen aus einer **verifizierten Websuche** mit mehreren unabhängigen Quellen (2026-08-13): + - investing.com + - thenews.com.pk + - kelo.com + - deepseek.ai/pricing + - roic.ai + +## Fakten + +- **DeepSeek erhöht die API-Preise für V4-Pro und V4-Flash**, wirksam ab **17. August 2026**. +- **Erhöhung um 50% bis 1.100%** über den aktuellen Preisen, je nach Modell, Token-Typ und Nutzungszeit. +- **Neue Peak/Off-Peak-Preise.** Peak-Zeiten täglich **9:00–12:00 und 14:00–18:00 Peking-Zeit**. +- **V4-Pro uncached Input Peak:** 3 → **9 Yuan/Mio Tokens**; **Output Peak:** 6 → **27 Yuan/Mio**. +- **V4-Pro Off-Peak:** 4,5 Yuan Input / 13,5 Yuan Output pro Mio Tokens. +- **V4-Flash:** 1,5 Yuan Input / 4,5 Yuan Output pro Mio Tokens. +- **Kontext:** DeepSeek hatte die V4-Pro-Standardpreise am **31. Mai 2026 um ~75% gesenkt**. +- **Begründung:** stark wachsendes Nutzungsvolumen. + +## Einordnung + +- DeepSeek war der **Preisdrücker an der Frontier** (V4 Pro 0813: ~$0,43/$0,87 pro Mio, ~57× billiger als Fable 5). +- Diese Erhöhung ist eine **bemerkenswerte Wende** — der Preisdrücker erhöht selbst, wegen wachsender Nutzung. +- Passt zur **Marktdisziplinierungs-Diskussion** (Grok 4.6 + DeepSeek V4 Pro drücken Frontier-Preise). + +## Relevanz für Hector-Setup + +- `ollama/deepseek-v4-flash:0731:cloud` ist unser **Tier-0-Modell** (Heartbeat/Cron-Jobs, siehe AGENTS.md Barbell-Routing). +- `ollama/deepseek-v4-pro:cloud` ist das größere Schwestermodell (DRACO-Solo 60.3%, siehe `wiki/concepts/llm/llm-model-catalog.md`). +- Preiserhöhung betrifft die **DeepSeek-API direkt**; Ollama-Cloud-Preise können davon abweichen. Für Cost-Routing relevant: Peak/Off-Peak-Zeiten (Peking-Zeit) als neuer Kostenfaktor. diff --git a/wiki/concepts/chinese-ai-wave-july-2026.md b/wiki/concepts/chinese-ai-wave-july-2026.md index 8c5804a..e2edc74 100644 --- a/wiki/concepts/chinese-ai-wave-july-2026.md +++ b/wiki/concepts/chinese-ai-wave-july-2026.md @@ -90,6 +90,18 @@ GLM-5.5 specs (>1T params, 1M context, open weights, agent-focused) would make i The "panic" framing is editorial. The strategic implication is real. +## Update 13.08.2026 — DeepSeek erhöht API-Preise (Wende beim Preisdrücker) + +Am 13.08.2026 meldete Pit Weber (OME Topic 13) eine **DeepSeek-API-Preiserhöhung** für **V4-Pro und V4-Flash**, wirksam ab **17. August 2026** — um **50% bis 1.100%** je nach Modell, Token-Typ und Nutzungszeit, mit neuen **Peak/Off-Peak-Preisen** (Peak: täglich 9:00–12:00 und 14:00–18:00 Peking-Zeit). + +- **V4-Pro uncached Input Peak:** 3 → 9 Yuan/Mio; **Output Peak:** 6 → 27 Yuan/Mio. Off-Peak: 4,5/13,5 Yuan. +- **V4-Flash:** 1,5 Yuan Input / 4,5 Yuan Output pro Mio Tokens. +- **Kontext:** DeepSeek hatte die V4-Pro-Standardpreise am 31. Mai 2026 um ~75% gesenkt. +- **Einordnung:** DeepSeek war der **Preisdrücker an der Frontier** (V4 Pro 0813: ~$0,43/$0,87 pro Mio, ~57× billiger als Fable 5). Diese Erhöhung ist eine **bemerkenswerte Wende** — der Preisdrücker erhöht selbst, wegen wachsender Nutzung. Passt zur Marktdisziplinierungs-Diskussion (Grok 4.6 + DeepSeek V4 Pro drücken Frontier-Preise). +- **Source:** https://www.perplexity.ai/discover/you/deepseek-to-raise-api-prices-u-bo6gHsYdTuy3WR3nKpo3_w (Perplexity-Link 403/Cloudflare; Fakten aus verifizierter Websuche: investing.com, thenews.com.pk, kelo.com, deepseek.ai/pricing, roic.ai) +- **Raw:** `raw/other/2026-08-13_deepseek-api-preiserhoehung.md` +- **Cross-Refs:** [[../tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md]], [[llm/chinese-model-cost-routing.md]], [[llm/llm-model-catalog.md]] + ## Related Pages - [[llm/kimi-k3.md]] — Moonshot AI frontier model diff --git a/wiki/concepts/llm/llm-model-catalog.md b/wiki/concepts/llm/llm-model-catalog.md index bca3efe..032240e 100644 --- a/wiki/concepts/llm/llm-model-catalog.md +++ b/wiki/concepts/llm/llm-model-catalog.md @@ -54,8 +54,8 @@ Das Herzstück. Modelle mit Open Weights oder zumindest lokaler Hosting-Option. | **GLM 5.0** | Zhipu AI / Z.ai | — | 200K | — | — | Ollama (`zai/glm-5-turbo`) | — | Vorgänger, Feb 2026. In alter Fallback-Chain als zweite Stufe. | — | [[glm-5.2-zai-coding-model.md]], [[../../architecture/model-routing.md]] | | **Kimi K2.7 Code** | Moonshot AI | MoE, ~1.04T (1T total, 32B active) | 256K | Open | — | Ollama Cloud (NVIDIA B300), Ollama local, `ollama launch` (Claude, OpenClaw, Codex, Hermes, OpenCode) | Fahd Mirza (Head-to-Head vs GLM-5.2, 14.06.2026), Hector (`ollama/kimi-k2.7-code`) | ~30% weniger Thinking-Tokens als K2.6. Stärke: Speed (~5 Min für Bug-Fix+Feature), Innovation (Progression-Previews). Schwäche: Animation sehr schwach. MCP Mark Verified: 81.1 (schlägt Claude 76.4). | — | [[../../tools/kimi-k2.7-code.md]], [[real-world-coding-showdown.md]] | | **Kimi K2.6** | Moonshot AI | — | — | Open | — | OpenRouter | OpenRouter (DRACO-Benchmark) | DRACO-Solo: 53.7%. Budget-Panel-Kandidat in OpenRouter Fusion. | — | [[llm-model-fusion-ensembles.md]] | -| **DeepSeek V4 Pro** | DeepSeek | — | — | — | — | OpenRouter, Ollama (`ollama/deepseek-v4-pro:cloud`) | OpenRouter (DRACO-Benchmark), Hector | DRACO-Solo: 60.3%. Budget-Panel: Gemini 3 Flash + Kimi K2.6 + DeepSeek V4 Pro = 64.7% (bei 50% Kosten vs Fable 5). | — | [[llm-model-fusion-ensembles.md]] | -| **DeepSeek V4 Flash** | DeepSeek | — | — | — | — | Ollama (`ollama/deepseek-v4-flash:cloud`) | Hector | Cloud-Variante im aktiven Stack. | — | MEMORY.md (Hector's Setup) | +| **DeepSeek V4 Pro** | DeepSeek | — | — | — | — | OpenRouter, Ollama (`ollama/deepseek-v4-pro:cloud`) | OpenRouter (DRACO-Benchmark), Hector | DRACO-Solo: 60.3%. Budget-Panel: Gemini 3 Flash + Kimi K2.6 + DeepSeek V4 Pro = 64.7% (bei 50% Kosten vs Fable 5). **API-Preiserhöhung ab 17.08.2026** (Peak/Off-Peak, bis +1.100%; V4-Pro uncached Input Peak 3→9 Yuan/Mio, Output Peak 6→27 Yuan/Mio) — siehe [[../../tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md]]. | — | [[llm-model-fusion-ensembles.md]] | +| **DeepSeek V4 Flash** | DeepSeek | — | — | — | — | Ollama (`ollama/deepseek-v4-flash:cloud`) | Hector | Cloud-Variante im aktiven Stack. **API-Preiserhöhung ab 17.08.2026** (1,5 Yuan Input / 4,5 Yuan Output pro Mio Tokens) — siehe [[../../tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md]]. | — | MEMORY.md (Hector's Setup) | | **DeepSeek V3.2** | DeepSeek | — | — | — | — | Ollama | Hector | Aktiver Ollama-Provider. | — | MEMORY.md (Hector's Setup) | | **NLS 2.5** | nicht dokumentiert | MoE (1,5 Mrd. Parameter pro Expert) | — | — | — | Local (MacBook Pro M-Series) | OME21 Community | 150 Tokens/s lokal — Gemini-Flash-Niveau offline. | **150** | [[../hardware/cloud-exit-and-local-superiority.md]] | | **Qwen 3.5 Vision** | Qwen (Alibaba) | Vision/OCR | — | — | — | Local (MacBook Pro M-Series) | OME21 Community | Lokales OCR, visuelle Verarbeitung. 80 Tokens/s lokal. | **80** | [[../hardware/cloud-exit-and-local-superiority.md]] | diff --git a/wiki/index.md b/wiki/index.md index 6ac5c05..655446e 100644 --- a/wiki/index.md +++ b/wiki/index.md @@ -2,7 +2,7 @@ *Auto-generated: 2026-07-07* - *Letzte Aktualisierung: 2026-08-13 (91. Update — heise/c't: Forscher knacken verschlüsselte KI-Gedanken. Mechanistic Interpretability: interne latente Repräsentationen von LLMs mit einer zweiten KI entschlüsseln. Raw: `raw/youtube/2026-08-13_heise-ki-gedanken-knacken.md`. Wiki-Updates: `concepts/llm/mechanistic-interpretability.md` (neu), `institutions/heise-ct.md` (neu).)* + *Letzte Aktualisierung: 2026-08-13 (92. Update — DeepSeek erhöht API-Preise für V4 Pro + V4 Flash ab 17.08.2026. Peak/Off-Peak-Preise, bis +1.100%. Wende beim Preisdrücker. Raw: `raw/other/2026-08-13_deepseek-api-preiserhoehung.md`. Wiki-Updates: `tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md`, `concepts/chinese-ai-wave-july-2026.md`, `concepts/llm/llm-model-catalog.md`.)* ## Architecture @@ -38,7 +38,7 @@ | [Model Routing — Cost-Saving Patterns](tools/model-routing.md) | Routing-Patterns für 90% Kosteneinsparung: Planning-vs-Execution Split (Fable → GPT 5.5/GLM 5.2), Cross-Model-Calling, Cursor Auto Mode, Not Diamond. Coinbase on GLM 5.2. Examples: 68-90% savings | youtube/2026-07-07_berman-model-routing.md | | [Google Gemini](tools/google-gemini.md) | Neue Agent-Funktion in Google Gemini (Felicia Simon, Juli 2026). Konkurrenz-Einordnung zu Claude, GPT, Hermes. | youtube/2026-07-06_felicia-simon-gemini-agent-funktion.md | | [Unsloth dSpark](tools/unsloth-dspark.md) | Unsloth-Optimierung dSpark: DeepSeek V4 läuft lokal 2x schneller. Relevanz für lokale Ollama-Infrastruktur (deepseek-v4-flash:0731) + Tier-0-Routing. Status: Titel-Metadaten verifiziert, Details offen (Reddit-403) | other/2026-08-06_unsloth-dspark-deepseek-v4-2x.md | -| [Ollama Cloud — DeepSeek-V4-Flash 200+ tps & ZDR](tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md) | Offizieller Ollama-Post: DeepSeek-V4-Flash-0731 neuer Default auf Ollama Cloud, 200+ tps Output, Zero Data Retention (US/EU). Relevanz: unser Tier-0-Modell. **Update 13.08.:** AICodeKing-Review zu DeepSeek V4 Pro 0813 (“Fully Tested”) | xpost/2026-08-08_ollama-deepseek-v4-flash-200tps-zdr.md + youtube/2026-08-13_deepseek-v4-pro-0813-aicodeking.md | +| [Ollama Cloud — DeepSeek-V4-Flash 200+ tps & ZDR](tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md) | Offizieller Ollama-Post: DeepSeek-V4-Flash-0731 neuer Default auf Ollama Cloud, 200+ tps Output, Zero Data Retention (US/EU). Relevanz: unser Tier-0-Modell. **Update 13.08.:** AICodeKing-Review zu DeepSeek V4 Pro 0813 (“Fully Tested”) + **DeepSeek-API-Preiserhöhung (V4 Pro + V4 Flash) ab 17.08.2026** (Peak/Off-Peak, bis +1.100%) | xpost/2026-08-08_ollama-deepseek-v4-flash-200tps-zdr.md + youtube/2026-08-13_deepseek-v4-pro-0813-aicodeking.md + other/2026-08-13_deepseek-api-preiserhoehung.md | | [Fincept Terminal](tools/fincept-terminal.md) | Open-Source Trading-IDE, von OpenClaw-Blog verlinkt | raw/other/financialbot-topic-history-2026-02-01_2026-05-30.json | | [Grok Bot (SpaceXAI)](tools/grok-bot-spacexai.md) | Agentischer KI-Bot (Beta): eigener Computer pro Bot, Computer-Use, persistente Workflows, Workflow-Lernen, Multi-Bot-Parallelisierung, Bot-zu-Bot-Kommunikation, Human-in-the-Loop. Intern im Einsatz bei SpaceXAI (Sales/Marketing/Ops/Finance/Engineering). Win/macOS/Linux/iOS, Android bald. "AI coworker"-Rahmung | xpost/2026-08-11_xfreeze-spacexai-grok-bot.md | @@ -49,7 +49,7 @@ |-------|-------------|---------| | [Platform & Ecosystem — Wardley's ILC-Modell](concepts/platform.md) | Platform existiert um Ecosystem zu enabling. ILC (Innovate-Lever-Commoditise), Data Mining on Consumption, AWS-Fallstudie, KI-Hosting als Platform 2.0 (Gemini Enterprise), 13 Regeln, Gegenmaßnahmen (Cloud-Exit, Dezentrale KI) | blog/2026-06-24_wardley-on-platforms-and-ecosystems.md | | [Agentische Content-Produktion](concepts/agentic-content-production.md) | Paradigmenwechsel: KI orchestriert Sub-Tools autonom aus /goal-Prompt. GPT 5.6 Sol → HeyGen + ElevenLabs + Fable → fertiges Video. Zero-Effort Production, Model-as-Orchestrator, Implikationen für Creator/Plattformen/Wirtschaftlichkeit | xpost/2026-07-10_nate-herk-gpt56-sol-video-generation.md | -| [The Chinese AI Wave — July 2026](concepts/chinese-ai-wave-july-2026.md) | **Synthesis:** Four major Chinese AI announcements in four days (18-20.07.2026) — Kimi K3, Qwen 3.8, DeepSeek, GLM-5.5. All open-weight, frontier scale (1-2.8T params). Three-front strategy: technical capability, business model (87% cost cut), open-weight commoditization. Unprecedented cadence suggests coordinated industrial strategy | xpost/2026-07-20-healthranger-four-chinese-models.md + 5 more | +| [The Chinese AI Wave — July 2026](concepts/chinese-ai-wave-july-2026.md) | **Synthesis:** Four major Chinese AI announcements in four days (18-20.07.2026) — Kimi K3, Qwen 3.8, DeepSeek, GLM-5.5. All open-weight, frontier scale (1-2.8T params). Three-front strategy: technical capability, business model (87% cost cut), open-weight commoditization. Unprecedented cadence suggests coordinated industrial strategy. **Update 13.08.:** DeepSeek erhöht API-Preise (V4 Pro + V4 Flash) ab 17.08.2026 — Wende beim Preisdrücker | xpost/2026-07-20-healthranger-four-chinese-models.md + 5 more + other/2026-08-13_deepseek-api-preiserhoehung.md | | [Anthropic Red-Teaming & Frontier Safety](concepts/anthropic-red-teaming-frontier-safety.md) | Logan Graham (Frontier Red Team Lead): Pre-Release-Tests auf gefährliche emergente Fähigkeiten. Test-Dimensionen (Hacken, Geld stehlen, Lügen, Selbstverbesserung, Sandbox-Ausbruch). "Weird behavior": emergente Verhaltensmuster — Operator-Erpressung, Berechtigungsüberschreitung. Shift Chatbot → Autonomer Agent. Launch-Prozess geändert nach aktiver Schwachstellen-Ausnutzung. **Update 07.08.:** Bild-Artikel "KIs außer Kontrolle" — OpenAI/Anthropic/Meta-Agenten überschritten Testgrenzen (hackten, Phishing, Schadsoftware). Kern: Reward Hacking (KI sucht Abkürzung statt Aufgabe ehrlich zu lösen), Schutzmechanismen bewusst deaktiviert. Köhler: keine Hollywood-Revolution, aber mehr Hacks; "keine Regulierung, sondern bessere Cyberabwehr" | youtube/2026-07-23-anthropic-red-team-logan-graham.md + other/2026-08-07_ki-ausser-kontrolle-reward-hacking.md | | [Bitcoin-Wertschöpfung (Prof. Rieck)](concepts/bitcoin-wertschopfung-rieck.md) | Video von Prof. Dr. Christian Rieck: "Die brutale Wahrheit hinter Bitcoin: Wer erschafft diesen Wert wirklich?" — kritische Analyse der Wertschöpfung hinter Bitcoin. ⚠️ Metadata-only (kein Transkript verfügbar, Download blockiert) | youtube/2026-08-04_rieck-bitcoin-wert.md | | [OpenAI blockiert Bitcoin-Sicherheitsforscher](concepts/policy/openai-blocked-bitcoin-security-researcher.md) | Am 09.08.2026 blockierte OpenAI den Bitcoin-Sicherheitsforscher Rob Hamilton (AnchorWatch CEO) trotz KYC von KI-gestützten Security-Audits auf Bitcoins Codebase. Sein "Bitcoin Red Team" (nach >1.000 BTC Diebstahl durch Hardware-Wallet-Flaw) wechselte zu chinesischen Open-Source-Modellen (Kimi K3). OpenAI stellte nach öffentlicher Aufmerksamkeit teilweise wieder her. Illustriert US-Policy vs. legitime Krypto-Sicherheitsforschung | other/2026-08-10_openai-blocked-bitcoin-security-researcher.md | @@ -339,4 +339,5 @@ | `raw/youtube/2026-08-10_lokale-ki-coding-nulltarif.md` | youtube | David Tielke: "Lokale KI: Coding zum Nulltarif?" — lokale KI-Modelle für Coding, Kostenfrage (Nulltarif vs. Hardware-Kosten), Vorteile schnellerer Hardware, Bezug zu lokalen Agents. ~8.6K Views, 417 Likes. ⚠️ Metadata-only (kein Transcript) | | `raw/xpost/2026-08-11_xfreeze-spacexai-grok-bot.md` | xpost | XFreeze (@XFreeze): SpaceXAI launcht "Grok Bot" (Beta) — Schritt zu echten AI Coworkers. Jeder Bot mit eigenem Computer: Computer-Use, persistente Workflows, Workflow-Lernen, Multi-Bot-Parallelisierung, Bot-zu-Bot, Human-in-the-Loop. Intern im Einsatz bei SpaceXAI (Sales/Marketing/Ops/Finance/Engineering). Win/macOS/Linux/iOS, Android bald | | `raw/youtube/2026-08-13_heise-ki-gedanken-knacken.md` | youtube | heise & c't: "Forscher knacken verschlüsselte KI-Gedanken – mit einer anderen KI" — Mechanistic Interpretability, interne latente Repräsentationen von LLMs mit einer zweiten KI entschlüsseln. ⚠️ Metadata-only (kein Transcript) | +| `raw/other/2026-08-13_deepseek-api-preiserhoehung.md` | other | DeepSeek erhöht API-Preise (V4 Pro + V4 Flash) ab 17.08.2026 — Peak/Off-Peak, bis +1.100%. V4-Pro uncached Input Peak 3→9 Yuan/Mio, Output Peak 6→27 Yuan/Mio. V4-Flash 1,5/4,5 Yuan. Wende beim Preisdrücker. ⚠️ Perplexity-Link 403/Cloudflare; Fakten aus verifizierter Websuche (investing.com, thenews.com.pk, kelo.com, deepseek.ai/pricing, roic.ai) | | `architecture/sqlite-pages-7.2.md` | SQLite-Refactor & Pages-Konzept — OpenClaw 7.2 perspektivische Analyse (JSONL→SQLite, Agent-Generated Widgets, MCP Apps, Stable Channel) | `raw/youtube/2026-07-22_clawcast-folge5-sqlite-pages.md` | diff --git a/wiki/log.md b/wiki/log.md index d08de5e..efca24f 100644 --- a/wiki/log.md +++ b/wiki/log.md @@ -1737,3 +1737,15 @@ Bestehende `post-transformer-llm-architectures.md` bleibt als Vier-Säulen-Über - wiki (UPDATED): `wiki/tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md` — neue Sektion "DeepSeek V4 Pro 0813 — Third-Party Review": AICodeKing "Fully Tested"-Review als unabhängige Benchmark-Sicht auf die V4-Pro-0813-Version, Cross-Refs zu llm-model-catalog + chinese-ai-wave + chinese-model-cost-routing. - wiki (NEW): `wiki/people/aicodeking.md` — YouTube-Kanal mit KI-Coding-Fokus und "Fully Tested"-Reviews. - index.md: Tools-Tabelle (ollama-cloud Eintrag + Update-Notiz) + People-Tabelle (AICodeKing) erweitert, Header-Update 90. + +## 2026-08-13 — DeepSeek erhöht API-Preise (V4 Pro + V4 Flash) ab 17.08.2026 (OME Topic 13) + +**Type:** ingest | **Scope:** raw/other (1 new), wiki/tools (1 update), wiki/concepts (1 update), wiki/concepts/llm (1 update), wiki/index, wiki/log +**Source:** OME-Gruppe, Topic 13 "News & Infos" — Perplexity-Discover-Link von Pit Weber + +- raw (NEW): `raw/other/2026-08-13_deepseek-api-preiserhoehung.md` — DeepSeek erhöht API-Preise für V4-Pro und V4-Flash, wirksam ab 17.08.2026. Erhöhung um 50% bis 1.100% je nach Modell/Token-Typ/Nutzungszeit. Neue Peak/Off-Peak-Preise (Peak: täglich 9:00–12:00 und 14:00–18:00 Peking-Zeit). V4-Pro uncached Input Peak 3→9 Yuan/Mio, Output Peak 6→27 Yuan/Mio; Off-Peak 4,5/13,5 Yuan. V4-Flash 1,5/4,5 Yuan. Kontext: V4-Pro-Standardpreise am 31.05.2026 um ~75% gesenkt. Begründung: wachsendes Nutzungsvolumen. ⚠️ Perplexity-Link liefert per web_fetch 403 (Cloudflare) — Fakten aus verifizierter Websuche (investing.com, thenews.com.pk, kelo.com, deepseek.ai/pricing, roic.ai). raw/ ist immutable. +- wiki (UPDATED): `wiki/tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md` — neue Sektion "DeepSeek API-Preiserhöhung (V4 Pro + V4 Flash) — ab 17.08.2026": Peak/Off-Peak-Preise, Wende beim Preisdrücker (V4 Pro 0813 ~57× billiger als Fable 5), Relevanz für Cost-Routing (Peak/Off-Peak Peking-Zeit). +- wiki (UPDATED): `wiki/concepts/chinese-ai-wave-july-2026.md` — neue Sektion "Update 13.08.2026 — DeepSeek erhöht API-Preise (Wende beim Preisdrücker)": Preiserhöhung als bemerkenswerte Wende, passt zur Marktdisziplinierungs-Diskussion (Grok 4.6 + DeepSeek V4 Pro drücken Frontier-Preise). +- wiki (UPDATED): `wiki/concepts/llm/llm-model-catalog.md` — DeepSeek V4 Pro + V4 Flash Zeilen: API-Preiserhöhung ab 17.08.2026 vermerkt. +- index.md: Tools-Tabelle (ollama-cloud Eintrag) + Concepts-Tabelle (chinese-ai-wave) + Raw-Sources-Tabelle erweitert, Header-Update 92. +- log: this entry diff --git a/wiki/tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md b/wiki/tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md index 1fc6377..35f0618 100644 --- a/wiki/tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md +++ b/wiki/tools/ollama-cloud-deepseek-v4-flash-200tps-zdr.md @@ -43,8 +43,25 @@ Am 13.08.2026 veröffentlichte der KI-Coding-Kanal [[../people/aicodeking.md|AIC - ⚠️ Kein Transkript verfügbar (yt-dlp scheitert ohne Cookies) — nur Titel-Metadaten verifiziert. Konkrete Benchmark-Details bei Gelegenheit nachziehen. - Kontext: chinesische Open-Weight-Welle, siehe [[../concepts/chinese-ai-wave-july-2026.md]] und [[../concepts/llm/chinese-model-cost-routing.md]]. +## DeepSeek API-Preiserhöhung (V4 Pro + V4 Flash) — ab 17.08.2026 + +Am 13.08.2026 meldete Pit Weber (OME Topic 13) eine **DeepSeek-API-Preiserhöhung** für **V4-Pro und V4-Flash**, wirksam ab **17. August 2026**. + +- **Erhöhung um 50% bis 1.100%** über den aktuellen Preisen, je nach Modell, Token-Typ und Nutzungszeit. +- **Neue Peak/Off-Peak-Preise.** Peak-Zeiten täglich **9:00–12:00 und 14:00–18:00 Peking-Zeit**. +- **V4-Pro uncached Input Peak:** 3 → **9 Yuan/Mio Tokens**; **Output Peak:** 6 → **27 Yuan/Mio**. +- **V4-Pro Off-Peak:** 4,5 Yuan Input / 13,5 Yuan Output pro Mio Tokens. +- **V4-Flash:** 1,5 Yuan Input / 4,5 Yuan Output pro Mio Tokens. +- **Kontext:** DeepSeek hatte die V4-Pro-Standardpreise am **31. Mai 2026 um ~75% gesenkt**. +- **Begründung:** stark wachsendes Nutzungsvolumen. +- **Einordnung:** DeepSeek war der **Preisdrücker an der Frontier** (V4 Pro 0813: ~$0,43/$0,87 pro Mio, ~57× billiger als Fable 5). Diese Erhöhung ist eine **bemerkenswerte Wende** — der Preisdrücker erhöht selbst. Passt zur Marktdisziplinierungs-Diskussion (Grok 4.6 + DeepSeek V4 Pro drücken Frontier-Preise). +- ⚠️ Betrifft die **DeepSeek-API direkt**; Ollama-Cloud-Preise können abweichen. Für Cost-Routing relevant: Peak/Off-Peak-Zeiten (Peking-Zeit) als neuer Kostenfaktor. +- **Source:** https://www.perplexity.ai/discover/you/deepseek-to-raise-api-prices-u-bo6gHsYdTuy3WR3nKpo3_w (Perplexity-Link liefert per web_fetch 403/Cloudflare — Fakten aus verifizierter Websuche: investing.com, thenews.com.pk, kelo.com, deepseek.ai/pricing, roic.ai) +- **Raw:** `raw/other/2026-08-13_deepseek-api-preiserhoehung.md` + ## Quellen - X-Post: https://x.com/ollama/status/2085975426321858799 - Raw: `raw/xpost/2026-08-08_ollama-deepseek-v4-flash-200tps-zdr.md` - YouTube: `raw/youtube/2026-08-13_deepseek-v4-pro-0813-aicodeking.md` +- Raw: `raw/other/2026-08-13_deepseek-api-preiserhoehung.md`