diff --git a/raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md b/raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md new file mode 100644 index 0000000..0712eb2 --- /dev/null +++ b/raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md @@ -0,0 +1,51 @@ +--- +type: xpost +source_url: https://x.com/bridgemindai/status/2078958257138528373 +retrieved: 2026-07-19 +author: "@bridgemindai" +is_thread: false +view_count: 205000 +like_count: 4700 +reply_count: 133 +tags: [kimi-k3, moonshot-ai, capacity, go-to-market, customer-first, anthropic, competitive-strategy, chinese-ai] +people: [] +institutions: [moonshot-ai, anthropic] +--- + +# BridgeMind AI: Moonshot K3 running 100% faster — capacity strategy vs. Anthropic + +**Source:** [X-Post by @bridgemindai](https://x.com/bridgemindai/status/2078958257138528373) +**Posted:** 2026-07-19 +**Views:** 205K | **Likes:** 4.7K | **Replies:** 133 + +## Post Content + +> Kimi K3 is running 100% FASTER than it was this morning. Why? Moonshot hit capacity and chose to stop selling NEW subscriptions instead of throttling existing ones. Every plan: sold out. On purpose. +> +> When Anthropic hit the same wall in April, they cut existing users' usage 50% during peak hours. +> +> Moonshot cut their revenue. Anthropic cut your usage. +> +> That says everything. + +## Context + +@Kimi_Moonshot announced pausing new subscriptions due to overwhelming demand on Kimi K3, protecting existing users, planning capacity expansions + tiered plans. + +## Key Takeaways + +1. **Customer-first vs. growth-first:** Moonshot chose to protect existing users' experience by stopping new subscriptions — sacrificing new revenue. Anthropic's April approach was the opposite: throttle existing users to keep selling to new customers. +2. **Capacity management as competitive signal:** How a company handles capacity ceilings reveals its priorities. Moonshot's approach builds loyalty; Anthropic's approach breeds frustration. +3. **Market timing:** This event occurs during the Kimi K3 launch wave (open-source release announced for July 27, 2026), amplifying the contrast with US frontier labs. +4. **Go-to-market philosophy as differentiator:** Combined with recent Chinese model signals (Kimi K3 #1 Frontend Code Arena, Qwen 3.8 open-weight), the AI market is shifting not just technically but in go-to-market philosophy. Customer-first vs. growth-first is becoming a competitive differentiator. + +## Analysis (Hector, OME group) + +This is a market signal. Moonshot prioritizes existing customers over new revenue — the opposite of Anthropic's April approach. Combined with recent signals (Kimi K3 #1 Frontend Code Arena, Qwen 3.8, both open-weight Chinese models), the AI market is shifting not just technically but in go-to-market philosophy. Customer-first vs. growth-first is becoming a competitive differentiator. + +## Cross-References + +- [[../../wiki/concepts/llm/kimi-k3.md]] +- [[../../wiki/institutions/moonshot-ai.md]] +- [[../../wiki/institutions/anthropic.md]] +- [[../../wiki/concepts/llm/chinese-model-cost-routing.md]] \ No newline at end of file diff --git a/wiki/concepts/llm/kimi-k3.md b/wiki/concepts/llm/kimi-k3.md index 8ce50e0..fc56313 100644 --- a/wiki/concepts/llm/kimi-k3.md +++ b/wiki/concepts/llm/kimi-k3.md @@ -1,14 +1,15 @@ --- created: 2026-07-18 -updated: 2026-07-18 +updated: 2026-07-19 sources: - raw/xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md + - raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md - raw/xpost/2026-06-29_deronin-chinese-ai-stack-cost-savings.md - raw/other/2026-06-13_kimi-k2.7-code-ollama.md - raw/youtube/2026-06-14_fahd-mirza-kimi-k2.7-vs-glm-5.2.md -tags: [concept, llm, kimi-k3, moonshot-ai, chinese-ai, open-source, pricing, agentic, coding, safety-guardrails] +tags: [concept, llm, kimi-k3, moonshot-ai, chinese-ai, open-source, pricing, agentic, coding, safety-guardrails, capacity, go-to-market, customer-first] people: [mike-adams] -institutions: [moonshot-ai] +institutions: [moonshot-ai, anthropic] --- # Kimi K3 @@ -56,9 +57,28 @@ Kimi K3 is part of a coordinated wave of Chinese model releases that are disrupt | **Qwen 3.7 Max** | Alibaba | General-purpose, multimodal | | **GLM 5.2** | Z.ai (Zhipu) | Coding, MIT license, 1M context | +## Capacity Event: Subscription Pause (July 19, 2026) + +On July 19, 2026, Kimi K3 experienced overwhelming demand, running **100% faster** than earlier in the day because Moonshot stopped selling new subscriptions instead of throttling existing users. Every plan was sold out — on purpose. + +**Moonshot's approach vs. Anthropic's April approach:** + +| Aspect | Moonshot (July 2026) | Anthropic (April 2026) | +|--------|---------------------|----------------------| +| Trigger | Kimi K3 capacity ceiling | Claude capacity ceiling | +| Response | Stopped new subscriptions | Cut existing users' usage 50% during peak | +| Revenue impact | Sacrificed new revenue | Protected new sales | +| User impact | Existing users unaffected | Existing users throttled | +| Signal | Customer-first | Growth-first | + +This event marks a **go-to-market philosophy divergence** between Chinese and US frontier labs. When combined with recent signals (Kimi K3 #1 Frontend Code Arena, Qwen 3.8 open-weight), customer-first vs. growth-first is becoming a competitive differentiator in the AI market. + +**Source:** [BridgeMind AI (@bridgemindai)](https://x.com/bridgemindai/status/2078958257138528373) — 205K views, 4.7K likes. Moonshot announced pausing new subscriptions, protecting existing users, planning capacity expansions + tiered plans. + ## Cross-References - [[../../institutions/moonshot-ai.md]] — Parent company +- [[../../institutions/anthropic.md]] — Competitor comparison (capacity handling) - [[chinese-model-cost-routing.md]] — Broader cost-routing thesis - [[ai-investment-bubble.md]] — AI bubble implications - [[fable-5-anthropic.md]] — Direct competitor comparison diff --git a/wiki/index.md b/wiki/index.md index 09d3f69..d153d70 100644 --- a/wiki/index.md +++ b/wiki/index.md @@ -2,7 +2,7 @@ *Auto-generated: 2026-07-07* - *Letzte Aktualisierung: 2026-07-18 (78. Update — @_The_Prophet__ AI Civilizational Integration wikifiziert. Raw: `raw/xpost/2026-07-18_theprophet-ai-civilizational-integration.md`. Wiki-Neu: `concepts/agi/ai-civilizational-integration.md`. Log: 2026-07-18 ingest: theprophet-ai-civilizational-integration.)* + *Letzte Aktualisierung: 2026-07-19 (79. Update — @bridgemindai Moonshot capacity strategy wikifiziert. Raw: `raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md`. Wiki-Update: `concepts/llm/kimi-k3.md` + `institutions/moonshot-ai.md`. Log: 2026-07-19 ingest: bridgemindai-moonshot-capacity.)* ## Architecture @@ -63,7 +63,7 @@ | [Semantic Similarity Rating (SSR)](concepts/llm/semantic-similarity-rating-ssr.md) | LLM-basierte Kaufintentions-Vorhersage mit 90% Korrelation | xpost/2026-06-11_colgate-llm-purchase-intent-ssr.md | | [GLM 5.2 (Z.ai) — Chinese Frontier Coding Model](concepts/llm/glm-5.2-zai-coding-model.md) | 10x günstiger als Claude, 1M Kontext, MIT-Lizenz, Z.ai Coding Plan, **nativ in OpenClaw v2026.6.8**. Update 22.06.: Arnie-Review mit 4 Tests, Self-Hosting-Pfade (LM Studio, Unsloth, DwarfStar), Kosten-Analyse. **Update 29.06.:** Semgrep IDOR-Benchmark ≈ Opus 4.8 bei Schwachstellen-Suche, Reward Hacking im RL-Training, DSGVO-konforme Security-Nutzung, Geopolitik. **Update 01.07.:** #1 Open-Weights auf Artificial Analysis Intelligence Index v4.1 (Score 51, 4th worldwide), SWE-bench Pro 62.1 beats GPT-5.5, Industry praise from Rauch/Levie/Howard. **Update 02.07.:** atomic.chat One-Shot Benchmark — B+ at $0.08, 39× cheaper than Fable 5, 6th independent validation | youtube/2026-06-15_ichbinfabian-glm-5.2-coding-modell.md + other/2026-06-16_openclaw-releases-v2026.6.8.md + youtube/2026-06-22_ai-mit-arnie-glm-5-2-review.md + blog/2026-06-29_heise-glm52-hacking-cybersecurity.md + blog/2026-07-01_perplexity-glm52-tops-open-weights-intelligence-index.md + xpost/2026-07-02_atomicchat-coding-benchmark-fable5-gpt55-opus48-glm52.md | | [Real-World Coding Showdown](concepts/llm/real-world-coding-showdown.md) | Head-to-Head-Methodik jenseits statischer Benchmarks, Kimi K2.7 vs GLM-5.2 in Hermes Agent, Sub-Task-Spezialisierung | youtube/2026-06-14_fahd-mirza-kimi-k2.7-vs-glm-5.2.md | -| [Kimi K3 — Moonshot AI Frontier Model](concepts/llm/kimi-k3.md) | ~8× günstiger als Claude, ~$15/M Tokens, Open Source angekündigt für 27. Juli 2026. Starke agentic/coding-Performance, extrem lange Kontextfenster. Safeguard-Kontroverse: ungefilterte Antworten vs. Claude-Blockaden. Teil der chinesischen Modell-Offensive (neben DeepSeek, Qwen, GLM) | xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md + xpost/2026-06-29_deronin-chinese-ai-stack-cost-savings.md | +| [Kimi K3 — Moonshot AI Frontier Model](concepts/llm/kimi-k3.md) | ~8× günstiger als Claude, ~$15/M Tokens, Open Source angekündigt für 27. Juli 2026. Starke agentic/coding-Performance, extrem lange Kontextfenster. Safeguard-Kontroverse: ungefilterte Antworten vs. Claude-Blockaden. **Capacity Event (19.07.):** Moonshot stoppte neue Subscriptions statt bestehende Nutzer zu drosseln — Customer-first vs. Anthropic's Growth-first. Teil der chinesischen Modell-Offensive (neben DeepSeek, Qwen, GLM) | xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md + xpost/2026-07-19-bridgemindai-moonshot-capacity.md + xpost/2026-06-29_deronin-chinese-ai-stack-cost-savings.md | | [Hyper-Zusammenfassungen: NotebookLM & Gemini 3.5 Flash](concepts/llm/hyper-summaries-notebooklm.md) | KI-Synthesen übertreffen Rohmaterial an Klarheit; agentische Verarbeitung via Antigravity; Telegram Rich-Text Fix (OpenWebUI). **Update 03.07.:** Short Video Overviews (60s vertikal, Nano Banana 2 Lite, One-Click, Free-Tier) + Pit's Winston+NotebookLM Pipeline | other/2026-06-19_ome21-briefing-ki-fortschritte-lokale-modelle.md + youtube/2026-07-01_futurepedia-notebooklm-short-video-overviews.md | | [The Flat Curve Society — Yegge's Intelligence Plateau Thesis](concepts/llm/flat-curve-society.md) | Steve Yegge: Kurve flacht für Öffentlichkeit ab (Model Lockdown + Discernment Horizon). AI Literacy Cohorts (Netflix). SaaS Revival. 5-Stunden-Training-Switch | blog/2026-06-19_steve-yegge-flat-curve-society.md | | [AI Value Migration — Orchestration + Infrastructure](concepts/llm/ai-value-migration-orchestration.md) | 20VC mit Aravind Srinivas: Wert verschiebt sich von Modellen zu Orchestrierung, Strom, Mindset. Token Value per Watt per User als Schlüsselkennzahl. Yegge-Flat-Curve-Schicht ergänzt | xpost/2026-06-15_harrystebbings-20vc-aravind-srinivas.md + 2 | @@ -307,3 +307,4 @@ | `raw/xpost/2026-07-18_steipete-loops-vs-graphs.md` | xpost | @steipete: "Are we still talking loops or did we shift to graphs yet?" — 248K Views, 402 Quotes. Loop Engineering vs. LangGraph-Debatte | | `raw/xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md` | xpost | @HealthRanger: "Anthropic is panicking over the release of Kimi K3" — 2.8M Views, ~4K Quotes. Kimi K3 vs. Claude Fable 5 Vergleich, Safeguard-Kontroverse, Open Source 27. Juli, ~8× günstiger, AI-Bubble-These | | `raw/xpost/2026-07-18_theprophet-ai-civilizational-integration.md` | xpost | @_The_Prophet__: Strategische Analyse — KI-Wettlauf als zivilisatorische Integration, nicht Benchmark-Race. Schrumpfende Modell-Grenze, ganzer Stack als Schutzwall, UK→US-Analogie, Goldilocks-Governance | +| `raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md` | xpost | @bridgemindai: Moonshot stoppte neue Kimi K3 Subscriptions statt bestehende Nutzer zu drosseln — Customer-first vs. Anthropic's Growth-first. 205K Views, 4.7K Likes. Go-to-market philosophy als competitive differentiator | diff --git a/wiki/institutions/moonshot-ai.md b/wiki/institutions/moonshot-ai.md index c1a772d..5a7c398 100644 --- a/wiki/institutions/moonshot-ai.md +++ b/wiki/institutions/moonshot-ai.md @@ -1,8 +1,8 @@ --- created: 2026-06-24 -updated: 2026-07-18 -sources: [`raw/other/2026-06-13_kimi-k2.7-code-ollama.md`, `raw/youtube/2026-06-14_fahd-mirza-kimi-k2.7-vs-glm-5.2.md`, `raw/xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md`] -tags: [institution, moonshot-ai, kimi-k3] +updated: 2026-07-19 +sources: [`raw/other/2026-06-13_kimi-k2.7-code-ollama.md`, `raw/youtube/2026-06-14_fahd-mirza-kimi-k2.7-vs-glm-5.2.md`, `raw/xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md`, `raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md`] +tags: [institution, moonshot-ai, kimi-k3, capacity, go-to-market] --- # Moonshot AI @@ -22,6 +22,17 @@ tags: [institution, moonshot-ai, kimi-k3] Moonshot AI ist ein chinesisches Unicorn-Startup, das sich durch extrem lange Kontext-Fähigkeiten seiner Kimi-Modelle auszeichnet. Mit dem Kimi K2.7 Code-Modell auf der Ollama Cloud konkurrieren sie direkt mit GLM-5.2 und bieten exzellente agentische Coding-Fähigkeiten für Entwickler-Plattformen. +### Kapazitäts-Strategie: Subscription Pause (19. Juli 2026) + +Am 19. Juli 2026 traf Moonshot eine bemerkenswerte Entscheidung: Bei Erreichen der Kapazitätsgrenze für Kimi K3 stoppte das Unternehmen den Verkauf **neuer** Abonnements — statt bestehende Nutzer zu drosseln. Alle Pläne waren ausverkauft. Dies führte zu einer 100% Performance-Steigerung für bestehende Nutzer. + +Im Kontrast dazu drosselte Anthropic im April 2026 bei einem ähnlichen Kapazitätsproblem bestehende Nutzer um 50% während der Spitzenzeiten — verkaufte aber weiter an neue Kunden. + +**Moonshots Ansatz:** Opferte neuen Umsatz, schützte bestehende Kunden. +**Anthropics Ansatz:** Schützte neuen Umsatz, opferte bestehende Kunden. + +Diese Entscheidung wurde als Markt-Signal wahrgommen: Customer-first vs. Growth-first wird zu einem competitiven Differenzierungsmerkmal im KI-Markt. ([BridgeMind AI, 205K views](https://x.com/bridgemindai/status/2078958257138528373)) + ### Kimi K3 (Juli 2026) Im Juli 2026 veröffentlichte Moonshot AI **Kimi K3**, ein Frontier-Modell das ~8× günstiger ist als Claude-Äquivalente (~$15/M Tokens). Der Open-Source-Release ist für den 27. Juli 2026 angekündigt. Kimi K3 sorgte für Kontroversen wegen seiner ungefilterten Antworten im Vergleich zu Claude's Safety-Guardrails — ein strukturelles Dilemma zwischen US-Safety-First und Chinese-Utility-First. Siehe [[../concepts/llm/kimi-k3.md]] für Details. diff --git a/wiki/log.md b/wiki/log.md index a3deef4..74601af 100644 --- a/wiki/log.md +++ b/wiki/log.md @@ -1430,3 +1430,16 @@ Bestehende `post-transformer-llm-architectures.md` bleibt als Vier-Säulen-Über - log: this entry **Hector-Hauptthese:** Steinbergers Tweet fasst die zentrale Architektur-Debatte im AI-Agent-Bereich 2026 zusammen. Der Konsens "Denk in Loops, implementier als LangGraph" ist pragmatisch: Loops bleiben die konzeptionelle Grundlage (Engineering-Philosophie), Graphen liefern die infrastrukturelle Robustheit (State, Persistenz, Tracing, HITL). Die beiden neuen Wiki-Seiten bilden ein komplementäres Paar — keine Wertung, sondern eine strukturierte Gegenüberstellung. **Subagent-Modell:** ollama/deepseek-v4-flash:cloud + +## [2026-07-19] Ingest | Moonshot Capacity Strategy — BridgeMind AI +**Type:** ingest | **Scope:** raw/xpost, wiki/concepts/llm, wiki/institutions, wiki/index, wiki/log +**Source:** X-Post von @bridgemindai — https://x.com/bridgemindai/status/2078958257138528373 (19.07.2026, 205K Views, 4.7K Likes, 133 Replies) +**Trigger:** Subagent task: Wikify Moonshot capacity/sold-out event + market analysis. +**Actions:** +- raw: `raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md` (created — 2.7 KB; Frontmatter [type: xpost, author: @bridgemindai, view_count: 205000, like_count: 4700, reply_count: 133, tags: kimi-k3, moonshot-ai, capacity, go-to-market, customer-first, anthropic]. Content: Post text, context on @Kimi_Moonshot announcement, key takeaways on customer-first vs. growth-first, Hector's market analysis) +- wiki (UPDATE): `concepts/llm/kimi-k3.md` (updated — neue Section "Capacity Event: Subscription Pause (July 19, 2026)" mit Vergleichstabelle Moonshot vs. Anthropic, updated frontmatter mit neuer source + tags + institutions) +- wiki (UPDATE): `institutions/moonshot-ai.md` (updated — neue Section "Kapazitäts-Strategie: Subscription Pause (19. Juli 2026)" vor Kimi K3 Section, updated frontmatter mit neuer source + tags) +- wiki: `index.md` (updated — Header auf "79. Update", Kimi K3-Eintrag erweitert mit Capacity Event, neuer Raw-Sources-Eintrag) +- log: this entry +**Hector-Hauptthese:** Moonshot's Entscheidung, neue Subscriptions zu stoppen statt bestehende Nutzer zu drosseln, ist mehr als ein Capacity-Workaround — es ist ein go-to-market-philosophy Statement. Im Kontrast zu Anthropic's April-Ansatz (bestehende Nutzer drosseln, weiter an neue verkaufen) wird Customer-first vs. Growth-first zu einem competitiven Differenzierungsmerkmal. Kombiniert mit technischen Signalen (Kimi K3 #1 Frontend Code Arena, Qwen 3.8 open-weight) verschiebt sich der chinesische AI-Markt nicht nur technologisch sondern auch philosophisch. +**Subagent-Modell:** ollama/glm-5.2:cloud diff --git a/wiki/tools/qwen-3.8.md b/wiki/tools/qwen-3.8.md new file mode 100644 index 0000000..c8bbd9d --- /dev/null +++ b/wiki/tools/qwen-3.8.md @@ -0,0 +1,25 @@ +--- +type: model +name: Qwen 3.8 +vendor: Alibaba +date: 2026-07-19 +status: announced +tags: [qwen, alibaba, open-source, moe, china-ai] +--- + +# Qwen 3.8 + +## Spec +- **Parameters:** 2.4T (trillion) +- **Release:** Open-weight (open source) +- **Date:** Announced 2026-07-19 +- **Vendor claim:** "One of the most powerful model available today, second only to Fable 5" +- **Early access:** Alibaba Token Plan, Qoder, QoderWork + +## Context +- Follows Kimi-K3 (2.8T) announced same week — two massive Chinese open-source MoE releases in rapid succession +- Continues trend: Chinese labs releasing open-weight models at scales US labs keep closed +- "Fable 5" referenced as #1 — not a widely established benchmark reference + +## Related +- [[../concepts/llm/kimi-k3.md]] — 2.8T, open-weight, same week \ No newline at end of file