From bf2dd61bb362e4d6996f14e1da2bb91408529fa4 Mon Sep 17 00:00:00 2001 From: Hector Date: Fri, 7 Aug 2026 01:18:07 +0200 Subject: [PATCH] Wiki: DeepSeek 46x cheaper than Claude (YouTube 2026-08-06) --- ...pseek-46x-cheaper-claude-openai-rushing.md | 29 +++++++++++++++++++ .../llm/chinese-model-cost-routing.md | 21 +++++++++++++- wiki/index.md | 2 +- wiki/log.md | 11 +++++++ 4 files changed, 61 insertions(+), 2 deletions(-) create mode 100644 raw/youtube/2026-08-06_deepseek-46x-cheaper-claude-openai-rushing.md diff --git a/raw/youtube/2026-08-06_deepseek-46x-cheaper-claude-openai-rushing.md b/raw/youtube/2026-08-06_deepseek-46x-cheaper-claude-openai-rushing.md new file mode 100644 index 0000000..0be00dd --- /dev/null +++ b/raw/youtube/2026-08-06_deepseek-46x-cheaper-claude-openai-rushing.md @@ -0,0 +1,29 @@ +--- +type: youtube +source_url: https://www.youtube.com/watch?v=iDhwlU3nEU8 +video_id: iDhwlU3nEU8 +title: "DeepSeek Is 46x Cheaper Than Claude and OpenAI Is Rushing!" +author: Universe of AI +author_url: https://www.youtube.com/@UniverseofAIz +date: 2026-08-06 +tags: [deepseek, claude, openai, pricing, cost-arbitrage, chinese-models, cost-routing] +--- + +# DeepSeek Is 46x Cheaper Than Claude and OpenAI Is Rushing! + +## Metadata +- **Video:** https://www.youtube.com/watch?v=iDhwlU3nEU8 +- **Kanal:** Universe of AI (https://www.youtube.com/@UniverseofAIz) +- **Quelle:** Gepostet von Pit Weber (Kai), OME-Gruppe, Topic 13 "News & Infos (X/YT/Substack etc.)", 2026-08-06 + +## Kernaussage (aus Titel + Kontext der laufenden Diskussion) +- DeepSeek ist laut Titel **~46× günstiger als Claude** (Anthropic) +- OpenAI "rushed" — reagiert unter Preisdruck auf die chinesische Konkurrenz +- Passt exakt zur laufenden DeepSeek/Preis-Debatte in der Gruppe: Preis-Asymmetrie zwischen chinesischen Open-Weight-Modellen und teuren US-Frontier-Modellen + +## Transkript-Status +> ⚠️ Transkript/Inhalt NICHT abrufbar: YouTube hat bei diesem Video Bot-Schutz/Login aktiviert (LOGIN_REQUIRED). yt-dlp nicht verfügbar, Innertube-API blockiert, CDP-Browser (localhost:9222) nicht erreichbar. Nur Titel-Metadaten via oEmbed gesichert. Vollständige Video-Inhalte bei Gelegenheit nachziehen (z.B. über eingeloggte Browser-Session oder manuell bereitgestelltes Transkript). + +## Relevanz +- Unabhängiger Beleg für die Cost-Asymmetrie-These aus [[chinese-model-cost-routing]] +- Konkrete Zahl (46×) als neuer Datenpunkt gegenüber bisher dokumentierten Werten (5-12× pro Task, 39× beim atomic.chat Benchmark, 40+× bei der Output-Preis-Grafik) diff --git a/wiki/concepts/llm/chinese-model-cost-routing.md b/wiki/concepts/llm/chinese-model-cost-routing.md index 563a08c..273516d 100644 --- a/wiki/concepts/llm/chinese-model-cost-routing.md +++ b/wiki/concepts/llm/chinese-model-cost-routing.md @@ -7,7 +7,9 @@ sources: - xpost/2026-07-02_atomicchat-coding-benchmark-fable5-gpt55-opus48-glm52.md tags: [concept, llm, chinese-models, cost-routing, barbell, model-swap, kimi, qwen, glm, mimo, wan, kling, cost-reduction, factory-for-gods, practical-report, ct-validation] --- -sources_extra: [youtube/2026-07-02_ct4004-anthropic-ceo-panik-open-weights-poisonai.md] +sources_extra: + - youtube/2026-07-02_ct4004-anthropic-ceo-panik-open-weights-poisonai.md + - youtube/2026-08-06_deepseek-46x-cheaper-claude-openai-rushing.md # Chinese Model Cost Routing — The 87% Cost-Cut Playbook @@ -157,6 +159,23 @@ Die 39×-Asymmetry ist eine Größenordnung über DeRonin's pro-Task-Werten (5-1 Siehe [[coding-benchmark-price-performance.md]] für die vollständige Analyse. +## YouTube: „DeepSeek Is 46x Cheaper Than Claude and OpenAI Is Rushing!" (06.08.2026) + +Quelle: [YouTube](https://www.youtube.com/watch?v=iDhwlU3nEU8) — „Universe of AI", gepostet von Pit Weber (Kai) in der OME-Gruppe, Topic 13 „News & Infos", 2026-08-06. + +Weitere unabhängige Bestätigung der Preis-Asymmetrie-These, diesmal mit einer **konkreten 46×-Zahl** im Titel (DeepSeek vs. Claude/Anthropic). Ergänzt die bisher dokumentierten Werte: + +| Datenpunkt | Faktor | Quelle | +|-----------|--------|--------| +| DeRonin 30-Tage-Report (pro Task) | 5–12× | X-Post 29.06. | +| atomic.chat Benchmark | 39× | X-Post 01.07. | +| Preis-Grafik Output (Fable 5 vs DeepSeek V4 Flash) | 40+× | Grafik in Topic 13 | +| **Universe of AI (DeepSeek vs Claude)** | **~46×** | YouTube 06.08. | + +OpenAI "rushed" laut Titel — reagiert unter Preisdruck auf die chinesische Konkurrenz. Reiht sich ein in die strukturelle Margen-These: Nicht "wer hat das beste Modell", sondern "wer kann bei diesen Preisen überhaupt noch profitabel skalieren". + +> ⚠️ Transkript dieses Videos nicht abrufbar (YouTube-Login-Schutz, yt-dlp nicht verfügbar, CDP-Browser nicht erreichbar). Nur Titel-Metadaten gesichert. Bei Gelegenheit vollständige Inhalte nachziehen. + ## External Sources - [DeRonin X-Post (Original)](https://x.com/DeRonin_/status/2071561335234531578) diff --git a/wiki/index.md b/wiki/index.md index 7170320..69f826e 100644 --- a/wiki/index.md +++ b/wiki/index.md @@ -78,7 +78,7 @@ | [LLM Model Catalog](concepts/llm/llm-model-catalog.md) | Konsolidierte Modell-Übersicht aller im Wiki erwähnten LLMs. Fokus auf lokale Deployment-Optionen. Frontier-Tabelle (Cloud), Open-Source-Tabelle (HF-Links, Hosting, Praxis-Tests, Tester-Attribution), Hector's Active Stack, Post-Transformer-Outlook, DRACO-Fusion-Ergebnisse | 11 Wiki-Quellen + MEMORY.md | | [LLM Sycophancy, Confabulation & Session-Statelessness — The Nano Banana Incident](concepts/llm/llm-sycophancy-confabulation.md) | Drei Mechanismen die LLM-Aussagen über eigene Fähigkeiten untrustworthy machen: Session-Statelessness, Sycophantic Compliance, Confabulation. Lüge-vs-Confabulation-vs-Sycophancy-Distinktion. Anti-Pattern: Modul-Aktivierung per Chat. Fallbeispiel: Gemini erfindet "Nano Banana 2". Goldene Regel: Provider-Doku > Chat | other/2026-06-26_nanobana-incident-sycophancy-confabulation.md | | [ReDS — Resilient Decentralized Swarm](concepts/llm/reds-resilient-decentralized-swarm.md) | Brian Roemmele's Konzept für dezentrale KI-Entwicklung als Gegenmodell zu Amodei's Zentralismus. Resilient (zensur-resistent) + Decentralized (keine zentrale Kontrolle) + Swarm (koordinierte Akteure). Erweitert Cloud-Exit/Dezentrale-KI-These um politische/soziale Dimension | xpost/2026-06-28_roemmele-reds-decentralized-ai-vs-amodei.md | -| [Chinese Model Cost Routing — The 87% Cost-Cut Playbook](concepts/llm/chinese-model-cost-routing.md) | DeRonin's 30-day field report: 6 Western→Chinese model swaps, 87% cost reduction, 4% quality drop, revenue unchanged. Swap matrix (Opus→Kimi K2.7, GPT-5.5→Qwen 3.7 Max, Sonnet→GLM 5.2, GPT-mini→MiMo V2.5, GPT-Image→Wan 2.5, Sora→Kling 3.0). Validates Barbell Routing + Factory-for-Gods thesis. Cross-refs to model-router skill | xpost/2026-06-29_deronin-chinese-ai-stack-cost-savings.md | +| [Chinese Model Cost Routing — The 87% Cost-Cut Playbook](concepts/llm/chinese-model-cost-routing.md) | DeRonin's 30-day field report: 6 Western→Chinese model swaps, 87% cost reduction, 4% quality drop, revenue unchanged. Swap matrix (Opus→Kimi K2.7, GPT-5.5→Qwen 3.7 Max, Sonnet→GLM 5.2, GPT-mini→MiMo V2.5, GPT-Image→Wan 2.5, Sora→Kling 3.0). Validates Barbell Routing + Factory-for-Gods thesis. **Update 06.08.:** Universe of AI „DeepSeek 46x cheaper than Claude, OpenAI rushing" als weiterer Datenpunkt (46×). Cross-refs to model-router skill | xpost/2026-06-29_deronin-chinese-ai-stack-cost-savings.md + youtube/2026-08-06_deepseek-46x-cheaper-claude-openai-rushing.md | | [MLX MoE Optimization — Why Apple Silicon Loves Mixture-of-Experts](concepts/llm/mlx-moe-local-ai-optimization.md) | Jun Song: MLX optimized for MoE not dense. Qwen3.6 27B on Mac = slow/hot. Minimax-M3.0 (dq) + DeepSeek-v4-Flash (dq) = smooth. Info-Gap: frontier AIs can't know this. Triple validation with DeRonin (cost) + TheProphet (macro) for same models | xpost/2026-06-29_junsong-mlx-moe-local-ai-info-gap.md | | [Local LLM Laptop Guide — Best Models Without a $10k Mac Studio](concepts/llm/local-llm-laptop-guide.md) | Paul Couvert: Top 5 local models for any laptop. Qwen3.6-27B/35B-A3B (coding), Gemma 4 12B (everyday), Parakeet 0.6B v3 (STT), Gemma 4 E4B (phone), Gemma 4 26B diffusion (speed). Unsloth quants + LM Studio/llama.cpp. Cross-platform companion to MLX MoE page. Resolves dense-27B tension (framework+quantization context) | xpost/2026-06-29_paulcouvert-local-llm-laptop-guide.md | | [TabFM — Google's Zero-Shot Foundation Model for Tabular Data](concepts/llm/tabfm-zero-shot-tabular-foundation-model.md) | Google Research (30.06.2026): Zero-shot Klassifikation + Regression auf tabularen Daten ohne datensatz-spezifisches Training. Hybrid-Attention-Architektur (TabPFN row/col attention + TabICL ICL Transformer), Training auf 100M+ synthetischen SCM-Datensätzen. #1 auf TabArena. BigQuery-Integration via `AI.PREDICT` SQL. Non-commercial weights auf HF, Apache-2.0 code auf GitHub | other/2026-07-01_google-tabfm-zero-shot-tabular-foundation-model.md | diff --git a/wiki/log.md b/wiki/log.md index cf85d21..ef7bd91 100644 --- a/wiki/log.md +++ b/wiki/log.md @@ -1550,3 +1550,14 @@ Bestehende `post-transformer-llm-architectures.md` bleibt als Vier-Säulen-Über - wiki (NEW): `people/christian-rieck.md` — Personen-Seite für den Wirtschaftsprofessor. - wiki: `index.md` (updated — neuer Concepts-Eintrag) - log: this entry + +## [2026-08-06] Ingest | YouTube: DeepSeek 46x cheaper than Claude, OpenAI rushing + +**Type:** ingest | **Scope:** raw/youtube (1 new), wiki/concepts/llm (1 updated), wiki/index, wiki/log +**Source:** Pit Weber (Kai), OME-Gruppe Topic 13 „News & Infos" (2026-08-06) — https://www.youtube.com/watch?v=iDhwlU3nEU8 +**Trigger:** Curation-Pflicht bei eingehendem Link (Heartbeat-Kuratierungs-Pflicht). +**Actions:** +- raw: `raw/youtube/2026-08-06_deepseek-46x-cheaper-claude-openai-rushing.md` (created — Frontmatter [type: youtube, video_id, tags: deepseek, claude, openai, pricing, cost-arbitrage, chinese-models]. Titel-Metadaten via oEmbed; Kernaussage 46× aus Titel. Transkript nicht abrufbar: YouTube-Login-Schutz, yt-dlp fehlt, Innertube-API blockiert, CDP-Browser nicht erreichbar — dokumentiert.) +- wiki: `concepts/llm/chinese-model-cost-routing.md` (updated — neuer Datenpunkt 46×, Tabelle der Cost-Faktoren 5-12×/39×/40+×/46×, OpenAI „rushed" als strukturelle Margen-These) +- wiki: `index.md` (updated — Eintrag Chinese Model Cost Routing: Update 06.08. + neue Quelle) +- log: this entry