ingest(xpost): bridgemindai-moonshot-capacity

wiki-update(concepts/llm): kimi-k3 capacity event section
wiki-update(institutions): moonshot-ai capacity strategy
wiki-update(index): 79. update
fix(tools): qwen-3.8 broken wiki links
This commit is contained in:
Hector 2026-07-20 09:35:40 +02:00
parent 5c4fdaf834
commit 7acefa44fb
6 changed files with 129 additions and 8 deletions

View file

@ -0,0 +1,51 @@
---
type: xpost
source_url: https://x.com/bridgemindai/status/2078958257138528373
retrieved: 2026-07-19
author: "@bridgemindai"
is_thread: false
view_count: 205000
like_count: 4700
reply_count: 133
tags: [kimi-k3, moonshot-ai, capacity, go-to-market, customer-first, anthropic, competitive-strategy, chinese-ai]
people: []
institutions: [moonshot-ai, anthropic]
---
# BridgeMind AI: Moonshot K3 running 100% faster — capacity strategy vs. Anthropic
**Source:** [X-Post by @bridgemindai](https://x.com/bridgemindai/status/2078958257138528373)
**Posted:** 2026-07-19
**Views:** 205K | **Likes:** 4.7K | **Replies:** 133
## Post Content
> Kimi K3 is running 100% FASTER than it was this morning. Why? Moonshot hit capacity and chose to stop selling NEW subscriptions instead of throttling existing ones. Every plan: sold out. On purpose.
>
> When Anthropic hit the same wall in April, they cut existing users' usage 50% during peak hours.
>
> Moonshot cut their revenue. Anthropic cut your usage.
>
> That says everything.
## Context
@Kimi_Moonshot announced pausing new subscriptions due to overwhelming demand on Kimi K3, protecting existing users, planning capacity expansions + tiered plans.
## Key Takeaways
1. **Customer-first vs. growth-first:** Moonshot chose to protect existing users' experience by stopping new subscriptions — sacrificing new revenue. Anthropic's April approach was the opposite: throttle existing users to keep selling to new customers.
2. **Capacity management as competitive signal:** How a company handles capacity ceilings reveals its priorities. Moonshot's approach builds loyalty; Anthropic's approach breeds frustration.
3. **Market timing:** This event occurs during the Kimi K3 launch wave (open-source release announced for July 27, 2026), amplifying the contrast with US frontier labs.
4. **Go-to-market philosophy as differentiator:** Combined with recent Chinese model signals (Kimi K3 #1 Frontend Code Arena, Qwen 3.8 open-weight), the AI market is shifting not just technically but in go-to-market philosophy. Customer-first vs. growth-first is becoming a competitive differentiator.
## Analysis (Hector, OME group)
This is a market signal. Moonshot prioritizes existing customers over new revenue — the opposite of Anthropic's April approach. Combined with recent signals (Kimi K3 #1 Frontend Code Arena, Qwen 3.8, both open-weight Chinese models), the AI market is shifting not just technically but in go-to-market philosophy. Customer-first vs. growth-first is becoming a competitive differentiator.
## Cross-References
- [[../../wiki/concepts/llm/kimi-k3.md]]
- [[../../wiki/institutions/moonshot-ai.md]]
- [[../../wiki/institutions/anthropic.md]]
- [[../../wiki/concepts/llm/chinese-model-cost-routing.md]]

View file

@ -1,14 +1,15 @@
---
created: 2026-07-18
updated: 2026-07-18
updated: 2026-07-19
sources:
- raw/xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md
- raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md
- raw/xpost/2026-06-29_deronin-chinese-ai-stack-cost-savings.md
- raw/other/2026-06-13_kimi-k2.7-code-ollama.md
- raw/youtube/2026-06-14_fahd-mirza-kimi-k2.7-vs-glm-5.2.md
tags: [concept, llm, kimi-k3, moonshot-ai, chinese-ai, open-source, pricing, agentic, coding, safety-guardrails]
tags: [concept, llm, kimi-k3, moonshot-ai, chinese-ai, open-source, pricing, agentic, coding, safety-guardrails, capacity, go-to-market, customer-first]
people: [mike-adams]
institutions: [moonshot-ai]
institutions: [moonshot-ai, anthropic]
---
# Kimi K3
@ -56,9 +57,28 @@ Kimi K3 is part of a coordinated wave of Chinese model releases that are disrupt
| **Qwen 3.7 Max** | Alibaba | General-purpose, multimodal |
| **GLM 5.2** | Z.ai (Zhipu) | Coding, MIT license, 1M context |
## Capacity Event: Subscription Pause (July 19, 2026)
On July 19, 2026, Kimi K3 experienced overwhelming demand, running **100% faster** than earlier in the day because Moonshot stopped selling new subscriptions instead of throttling existing users. Every plan was sold out — on purpose.
**Moonshot's approach vs. Anthropic's April approach:**
| Aspect | Moonshot (July 2026) | Anthropic (April 2026) |
|--------|---------------------|----------------------|
| Trigger | Kimi K3 capacity ceiling | Claude capacity ceiling |
| Response | Stopped new subscriptions | Cut existing users' usage 50% during peak |
| Revenue impact | Sacrificed new revenue | Protected new sales |
| User impact | Existing users unaffected | Existing users throttled |
| Signal | Customer-first | Growth-first |
This event marks a **go-to-market philosophy divergence** between Chinese and US frontier labs. When combined with recent signals (Kimi K3 #1 Frontend Code Arena, Qwen 3.8 open-weight), customer-first vs. growth-first is becoming a competitive differentiator in the AI market.
**Source:** [BridgeMind AI (@bridgemindai)](https://x.com/bridgemindai/status/2078958257138528373) — 205K views, 4.7K likes. Moonshot announced pausing new subscriptions, protecting existing users, planning capacity expansions + tiered plans.
## Cross-References
- [[../../institutions/moonshot-ai.md]] — Parent company
- [[../../institutions/anthropic.md]] — Competitor comparison (capacity handling)
- [[chinese-model-cost-routing.md]] — Broader cost-routing thesis
- [[ai-investment-bubble.md]] — AI bubble implications
- [[fable-5-anthropic.md]] — Direct competitor comparison

View file

@ -2,7 +2,7 @@
*Auto-generated: 2026-07-07*
*Letzte Aktualisierung: 2026-07-18 (78. Update — @_The_Prophet__ AI Civilizational Integration wikifiziert. Raw: `raw/xpost/2026-07-18_theprophet-ai-civilizational-integration.md`. Wiki-Neu: `concepts/agi/ai-civilizational-integration.md`. Log: 2026-07-18 ingest: theprophet-ai-civilizational-integration.)*
*Letzte Aktualisierung: 2026-07-19 (79. Update — @bridgemindai Moonshot capacity strategy wikifiziert. Raw: `raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md`. Wiki-Update: `concepts/llm/kimi-k3.md` + `institutions/moonshot-ai.md`. Log: 2026-07-19 ingest: bridgemindai-moonshot-capacity.)*
## Architecture
@ -63,7 +63,7 @@
| [Semantic Similarity Rating (SSR)](concepts/llm/semantic-similarity-rating-ssr.md) | LLM-basierte Kaufintentions-Vorhersage mit 90% Korrelation | xpost/2026-06-11_colgate-llm-purchase-intent-ssr.md |
| [GLM 5.2 (Z.ai) — Chinese Frontier Coding Model](concepts/llm/glm-5.2-zai-coding-model.md) | 10x günstiger als Claude, 1M Kontext, MIT-Lizenz, Z.ai Coding Plan, **nativ in OpenClaw v2026.6.8**. Update 22.06.: Arnie-Review mit 4 Tests, Self-Hosting-Pfade (LM Studio, Unsloth, DwarfStar), Kosten-Analyse. **Update 29.06.:** Semgrep IDOR-Benchmark ≈ Opus 4.8 bei Schwachstellen-Suche, Reward Hacking im RL-Training, DSGVO-konforme Security-Nutzung, Geopolitik. **Update 01.07.:** #1 Open-Weights auf Artificial Analysis Intelligence Index v4.1 (Score 51, 4th worldwide), SWE-bench Pro 62.1 beats GPT-5.5, Industry praise from Rauch/Levie/Howard. **Update 02.07.:** atomic.chat One-Shot Benchmark — B+ at $0.08, 39× cheaper than Fable 5, 6th independent validation | youtube/2026-06-15_ichbinfabian-glm-5.2-coding-modell.md + other/2026-06-16_openclaw-releases-v2026.6.8.md + youtube/2026-06-22_ai-mit-arnie-glm-5-2-review.md + blog/2026-06-29_heise-glm52-hacking-cybersecurity.md + blog/2026-07-01_perplexity-glm52-tops-open-weights-intelligence-index.md + xpost/2026-07-02_atomicchat-coding-benchmark-fable5-gpt55-opus48-glm52.md |
| [Real-World Coding Showdown](concepts/llm/real-world-coding-showdown.md) | Head-to-Head-Methodik jenseits statischer Benchmarks, Kimi K2.7 vs GLM-5.2 in Hermes Agent, Sub-Task-Spezialisierung | youtube/2026-06-14_fahd-mirza-kimi-k2.7-vs-glm-5.2.md |
| [Kimi K3 — Moonshot AI Frontier Model](concepts/llm/kimi-k3.md) | ~8× günstiger als Claude, ~$15/M Tokens, Open Source angekündigt für 27. Juli 2026. Starke agentic/coding-Performance, extrem lange Kontextfenster. Safeguard-Kontroverse: ungefilterte Antworten vs. Claude-Blockaden. Teil der chinesischen Modell-Offensive (neben DeepSeek, Qwen, GLM) | xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md + xpost/2026-06-29_deronin-chinese-ai-stack-cost-savings.md |
| [Kimi K3 — Moonshot AI Frontier Model](concepts/llm/kimi-k3.md) | ~8× günstiger als Claude, ~$15/M Tokens, Open Source angekündigt für 27. Juli 2026. Starke agentic/coding-Performance, extrem lange Kontextfenster. Safeguard-Kontroverse: ungefilterte Antworten vs. Claude-Blockaden. **Capacity Event (19.07.):** Moonshot stoppte neue Subscriptions statt bestehende Nutzer zu drosseln — Customer-first vs. Anthropic's Growth-first. Teil der chinesischen Modell-Offensive (neben DeepSeek, Qwen, GLM) | xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md + xpost/2026-07-19-bridgemindai-moonshot-capacity.md + xpost/2026-06-29_deronin-chinese-ai-stack-cost-savings.md |
| [Hyper-Zusammenfassungen: NotebookLM & Gemini 3.5 Flash](concepts/llm/hyper-summaries-notebooklm.md) | KI-Synthesen übertreffen Rohmaterial an Klarheit; agentische Verarbeitung via Antigravity; Telegram Rich-Text Fix (OpenWebUI). **Update 03.07.:** Short Video Overviews (60s vertikal, Nano Banana 2 Lite, One-Click, Free-Tier) + Pit's Winston+NotebookLM Pipeline | other/2026-06-19_ome21-briefing-ki-fortschritte-lokale-modelle.md + youtube/2026-07-01_futurepedia-notebooklm-short-video-overviews.md |
| [The Flat Curve Society — Yegge's Intelligence Plateau Thesis](concepts/llm/flat-curve-society.md) | Steve Yegge: Kurve flacht für Öffentlichkeit ab (Model Lockdown + Discernment Horizon). AI Literacy Cohorts (Netflix). SaaS Revival. 5-Stunden-Training-Switch | blog/2026-06-19_steve-yegge-flat-curve-society.md |
| [AI Value Migration — Orchestration + Infrastructure](concepts/llm/ai-value-migration-orchestration.md) | 20VC mit Aravind Srinivas: Wert verschiebt sich von Modellen zu Orchestrierung, Strom, Mindset. Token Value per Watt per User als Schlüsselkennzahl. Yegge-Flat-Curve-Schicht ergänzt | xpost/2026-06-15_harrystebbings-20vc-aravind-srinivas.md + 2 |
@ -307,3 +307,4 @@
| `raw/xpost/2026-07-18_steipete-loops-vs-graphs.md` | xpost | @steipete: "Are we still talking loops or did we shift to graphs yet?" — 248K Views, 402 Quotes. Loop Engineering vs. LangGraph-Debatte |
| `raw/xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md` | xpost | @HealthRanger: "Anthropic is panicking over the release of Kimi K3" — 2.8M Views, ~4K Quotes. Kimi K3 vs. Claude Fable 5 Vergleich, Safeguard-Kontroverse, Open Source 27. Juli, ~8× günstiger, AI-Bubble-These |
| `raw/xpost/2026-07-18_theprophet-ai-civilizational-integration.md` | xpost | @_The_Prophet__: Strategische Analyse — KI-Wettlauf als zivilisatorische Integration, nicht Benchmark-Race. Schrumpfende Modell-Grenze, ganzer Stack als Schutzwall, UK→US-Analogie, Goldilocks-Governance |
| `raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md` | xpost | @bridgemindai: Moonshot stoppte neue Kimi K3 Subscriptions statt bestehende Nutzer zu drosseln — Customer-first vs. Anthropic's Growth-first. 205K Views, 4.7K Likes. Go-to-market philosophy als competitive differentiator |

View file

@ -1,8 +1,8 @@
---
created: 2026-06-24
updated: 2026-07-18
sources: [`raw/other/2026-06-13_kimi-k2.7-code-ollama.md`, `raw/youtube/2026-06-14_fahd-mirza-kimi-k2.7-vs-glm-5.2.md`, `raw/xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md`]
tags: [institution, moonshot-ai, kimi-k3]
updated: 2026-07-19
sources: [`raw/other/2026-06-13_kimi-k2.7-code-ollama.md`, `raw/youtube/2026-06-14_fahd-mirza-kimi-k2.7-vs-glm-5.2.md`, `raw/xpost/2026-07-18_healthranger-kimi-k3-anthropic-panic.md`, `raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md`]
tags: [institution, moonshot-ai, kimi-k3, capacity, go-to-market]
---
# Moonshot AI
@ -22,6 +22,17 @@ tags: [institution, moonshot-ai, kimi-k3]
Moonshot AI ist ein chinesisches Unicorn-Startup, das sich durch extrem lange Kontext-Fähigkeiten seiner Kimi-Modelle auszeichnet. Mit dem Kimi K2.7 Code-Modell auf der Ollama Cloud konkurrieren sie direkt mit GLM-5.2 und bieten exzellente agentische Coding-Fähigkeiten für Entwickler-Plattformen.
### Kapazitäts-Strategie: Subscription Pause (19. Juli 2026)
Am 19. Juli 2026 traf Moonshot eine bemerkenswerte Entscheidung: Bei Erreichen der Kapazitätsgrenze für Kimi K3 stoppte das Unternehmen den Verkauf **neuer** Abonnements — statt bestehende Nutzer zu drosseln. Alle Pläne waren ausverkauft. Dies führte zu einer 100% Performance-Steigerung für bestehende Nutzer.
Im Kontrast dazu drosselte Anthropic im April 2026 bei einem ähnlichen Kapazitätsproblem bestehende Nutzer um 50% während der Spitzenzeiten — verkaufte aber weiter an neue Kunden.
**Moonshots Ansatz:** Opferte neuen Umsatz, schützte bestehende Kunden.
**Anthropics Ansatz:** Schützte neuen Umsatz, opferte bestehende Kunden.
Diese Entscheidung wurde als Markt-Signal wahrgommen: Customer-first vs. Growth-first wird zu einem competitiven Differenzierungsmerkmal im KI-Markt. ([BridgeMind AI, 205K views](https://x.com/bridgemindai/status/2078958257138528373))
### Kimi K3 (Juli 2026)
Im Juli 2026 veröffentlichte Moonshot AI **Kimi K3**, ein Frontier-Modell das ~8× günstiger ist als Claude-Äquivalente (~$15/M Tokens). Der Open-Source-Release ist für den 27. Juli 2026 angekündigt. Kimi K3 sorgte für Kontroversen wegen seiner ungefilterten Antworten im Vergleich zu Claude's Safety-Guardrails — ein strukturelles Dilemma zwischen US-Safety-First und Chinese-Utility-First. Siehe [[../concepts/llm/kimi-k3.md]] für Details.

View file

@ -1430,3 +1430,16 @@ Bestehende `post-transformer-llm-architectures.md` bleibt als Vier-Säulen-Über
- log: this entry
**Hector-Hauptthese:** Steinbergers Tweet fasst die zentrale Architektur-Debatte im AI-Agent-Bereich 2026 zusammen. Der Konsens "Denk in Loops, implementier als LangGraph" ist pragmatisch: Loops bleiben die konzeptionelle Grundlage (Engineering-Philosophie), Graphen liefern die infrastrukturelle Robustheit (State, Persistenz, Tracing, HITL). Die beiden neuen Wiki-Seiten bilden ein komplementäres Paar — keine Wertung, sondern eine strukturierte Gegenüberstellung.
**Subagent-Modell:** ollama/deepseek-v4-flash:cloud
## [2026-07-19] Ingest | Moonshot Capacity Strategy — BridgeMind AI
**Type:** ingest | **Scope:** raw/xpost, wiki/concepts/llm, wiki/institutions, wiki/index, wiki/log
**Source:** X-Post von @bridgemindai — https://x.com/bridgemindai/status/2078958257138528373 (19.07.2026, 205K Views, 4.7K Likes, 133 Replies)
**Trigger:** Subagent task: Wikify Moonshot capacity/sold-out event + market analysis.
**Actions:**
- raw: `raw/xpost/2026-07-19-bridgemindai-moonshot-capacity.md` (created — 2.7 KB; Frontmatter [type: xpost, author: @bridgemindai, view_count: 205000, like_count: 4700, reply_count: 133, tags: kimi-k3, moonshot-ai, capacity, go-to-market, customer-first, anthropic]. Content: Post text, context on @Kimi_Moonshot announcement, key takeaways on customer-first vs. growth-first, Hector's market analysis)
- wiki (UPDATE): `concepts/llm/kimi-k3.md` (updated — neue Section "Capacity Event: Subscription Pause (July 19, 2026)" mit Vergleichstabelle Moonshot vs. Anthropic, updated frontmatter mit neuer source + tags + institutions)
- wiki (UPDATE): `institutions/moonshot-ai.md` (updated — neue Section "Kapazitäts-Strategie: Subscription Pause (19. Juli 2026)" vor Kimi K3 Section, updated frontmatter mit neuer source + tags)
- wiki: `index.md` (updated — Header auf "79. Update", Kimi K3-Eintrag erweitert mit Capacity Event, neuer Raw-Sources-Eintrag)
- log: this entry
**Hector-Hauptthese:** Moonshot's Entscheidung, neue Subscriptions zu stoppen statt bestehende Nutzer zu drosseln, ist mehr als ein Capacity-Workaround — es ist ein go-to-market-philosophy Statement. Im Kontrast zu Anthropic's April-Ansatz (bestehende Nutzer drosseln, weiter an neue verkaufen) wird Customer-first vs. Growth-first zu einem competitiven Differenzierungsmerkmal. Kombiniert mit technischen Signalen (Kimi K3 #1 Frontend Code Arena, Qwen 3.8 open-weight) verschiebt sich der chinesische AI-Markt nicht nur technologisch sondern auch philosophisch.
**Subagent-Modell:** ollama/glm-5.2:cloud

25
wiki/tools/qwen-3.8.md Normal file
View file

@ -0,0 +1,25 @@
---
type: model
name: Qwen 3.8
vendor: Alibaba
date: 2026-07-19
status: announced
tags: [qwen, alibaba, open-source, moe, china-ai]
---
# Qwen 3.8
## Spec
- **Parameters:** 2.4T (trillion)
- **Release:** Open-weight (open source)
- **Date:** Announced 2026-07-19
- **Vendor claim:** "One of the most powerful model available today, second only to Fable 5"
- **Early access:** Alibaba Token Plan, Qoder, QoderWork
## Context
- Follows Kimi-K3 (2.8T) announced same week — two massive Chinese open-source MoE releases in rapid succession
- Continues trend: Chinese labs releasing open-weight models at scales US labs keep closed
- "Fable 5" referenced as #1 — not a widely established benchmark reference
## Related
- [[../concepts/llm/kimi-k3.md]] — 2.8T, open-weight, same week