diff --git a/raw/xpost/2026-08-23_insiderleak-ox-alpha-anonymous-model.md b/raw/xpost/2026-08-23_insiderleak-ox-alpha-anonymous-model.md new file mode 100644 index 0000000..f814ec2 --- /dev/null +++ b/raw/xpost/2026-08-23_insiderleak-ox-alpha-anonymous-model.md @@ -0,0 +1,43 @@ +--- +type: xpost +source_url: "https://t.me/Insider_leak_of_theday" +retrieved: 2026-08-23 +author: "Insider leak of the day (@Insider_leak_of_theday)" +is_thread: false +forwarded_at: 2026-08-22T15:33:34.000Z +shared_by: "Netbits ⚡️ Stachelbanane (Netbits @NetLightning)" +tags: [model, ox-alpha, openrouter, anonymous-model, chinese-ai, glm, tokenizer] +--- + +# Ox Alpha — Mysteriöses anonymes KI-Modell (Insider-Leak) + +> **Quelle:** Telegram-Kanal "Insider leak of the day", Forward vom 2026-08-22T15:33:34.000Z. Geteilt von Netbits (@NetLightning) in OME-Gruppe "News & Infos (X/YT/Substack etc.)"-Topic am 2026-08-23. + +## Kern-Fakten (neutral aus der Quelle) + +- **Ein mysteriöses neues KI-Modell namens "Ox Alpha"** taucht Berichten zufolge auf und soll beim Coding **Claude Fable 5 und GPT-5.6 Sol schlagen**. +- **Niemand weiß, wer es gebaut hat** (kein offengelegter Owner). +- Das Modell erschien auf **OpenRouter** mit: + - **1-Million-Token-Kontextfenster** + - **multimodalen Fähigkeiten** + - **keinem offengelegten Besitzer/Owner** +- **Spekulation:** Online-Sleuths deuten auf **Google**; andere vermuten ein neues Modell des chinesischen AI-Labs **Z.ai**. +- **Tokenizers-Auffälligkeit:** Ox Alphas Tokenizer soll laut Report **identisch zu GLM** sein. +- **Muster-Hinweis:** Die **vier vorherigen anonymen AI-Modell-Drops** wurden laut Leak **alle letztlich von chinesischen Labs beansprucht**. + +## Kontext aus dem begleitenden Daten-Screenshot (OpenRouter-Databreakdown "AUG 21", Total 9.2T Tokens) + +Relevante Modelle im Nutzungs-Dashboard (quantitativ sortiert): + +- deepseek-v4-flash — 2.5T +- mimo-v2.5 — 1.5T +- deepseek-v4-pro — 272B +- muse-spark-1.2-cont — 1.1T +- nemotron-3-ultra — 411B +- **ox-alpha — 2.6T (hervorgehoben)** +- gpt-5.6-luna — 117B +- hy3 — 365B +- nemotron-3.5-lightning — 99B +- Other — 347B + +> Hinweis: Screenshot-Daten als begleitender Kontext der Quelle; Ox Alpha zeigt im AUG-21-Dashboard ein Nutzungsvolumen von ~2.6T Tokens (hervorgehoben). diff --git a/wiki/concepts/llm/ox-alpha-anonymous-model.md b/wiki/concepts/llm/ox-alpha-anonymous-model.md new file mode 100644 index 0000000..94cb0ab --- /dev/null +++ b/wiki/concepts/llm/ox-alpha-anonymous-model.md @@ -0,0 +1,51 @@ +--- +created: 2026-08-23 +updated: 2026-08-23 +sources: [xpost/2026-08-23_insiderleak-ox-alpha-anonymous-model.md] +tags: [concept, model, ox-alpha, openrouter, anonymous-model, chinese-ai, glm, tokenizer, mystery-model] +--- + +# Ox Alpha — Anonymes KI-Modell auf OpenRouter + +> **Quelle des Konzepts:** Telegram-Leak "Insider leak of the day" (Forward 2026-08-22), geteilt von Netbits (@NetLightning) in OME-Gruppe "News & Infos (X/YT/Substack etc.)"-Topic am 2026-08-23. Siehe `raw/xpost/2026-08-23_insiderleak-ox-alpha-anonymous-model.md`. + +## Kernidee + +**Ox Alpha ist ein mysteriöses, anonym veröffentlichtes KI-Modell auf OpenRouter, das Berichten zufolge beim Coding Claude Fable 5 und GPT-5.6 Sol schlagen soll — ohne offengelegten Owner.** Es gehört zu einer Reihe anonym publizierter Modell-Drops, deren Herkunft erst später geklärt wird. + +## Eigenschaften (laut Leak) + +| Eigenschaft | Wert | +|---|---| +| Name | Ox Alpha | +| Plattform | OpenRouter | +| Kontextfenster | ~1 Million Tokens | +| Modalität | multimodal | +| Owner | nicht offengelegt (anonym) | +| Coding-Leistung | soll Claude Fable 5 + GPT-5.6 Sol schlagen | +| Tokenizer | **identisch zu GLM** (laut Report) | + +## Herkunfts-Spekulationen + +- **Google** — von manchen Online-Sleuths vermutet. +- **Z.ai (Zhipu AI):** von anderen vermutet — gestützt durch den **GLM-identischen Tokenizer**. +- **Muster:** Die vier vorherigen anonymen AI-Modell-Drops wurden laut Leak alle letztlich von **chinesischen Labs** beansprucht — spricht für die China-These. + +> ⚠️ **Status:** Unbestätigte Gerüchte aus einem Leak-Kanal. Kein offizieller Owner, keine offizielle Verifikation. GLM-Tokenizer-Verbindung ist ein Hinweis, kein Beweis. + +## Einordnung + +Ox Alpha reiht sich in die Welle kostengünstiger, teils anonymer Frontier-Coding-Modelle ein — thematisch verwandt mit: + +- [[glm-5.3-z-ai.md]] und [[glm-5.2-zai-coding-model.md]] — die GLM-Familie von [[../../institutions/z-ai.md|Z.ai]] +- [[chinese-model-cost-routing.md]] — chinesische Modelle & Kosten-Routing +- [[open-weights-economics.md]] — Open-Weight-Ökonomie +- [[openrouter-cad-solidworks-benchmark.md]] — OpenRouter-Modellnutzung (dort: DeepSeek/Kimi/MiniMax/Qwen) +- [[../../institutions/openrouter.md|OpenRouter]] — Modell-Aggregator, auf dem Ox Alpha erschienen ist +- [[chinese-models-top-usage-charts.md]] — chinesische Modelle in den Top-Usage-Charts + +**Verwandte Modelle in der GLM-Familie:** siehe [[glm-5.3-z-ai.md]] (Tabellen-Einordnung), [[glm-5.5-z-ai.md]]. + +## Nutzungskontext (begleitender Screenshot, AUG 21) + +Im OpenRouter-Databreakdown "AUG 21" (Total ~9.2T Tokens) zeigt `ox-alpha` ~2.6T Tokens und ist dabei hervorgehoben — vergleichbar mit deepseek-v4-flash (2.5T) und mimo-v2.5 (1.5T). (Kontext aus dem geteilten Screenshot; quantitative Rohdaten im Raw.) diff --git a/wiki/index.md b/wiki/index.md index 8bff796..89da836 100644 --- a/wiki/index.md +++ b/wiki/index.md @@ -2,7 +2,9 @@ *Auto-generated: 2026-07-07* -*Letzte Aktualisierung: 2026-08-22 (122. Update — OpenRouter CAD/SOLIDWORKS-Benchmark (@jack67396128, 2026-08-22). Raw: `raw/xpost/2026-08-22_jack-openrouter-cad-solidworks-benchmark.md`. Wiki-Update: `concepts/llm/openrouter-cad-solidworks-benchmark.md` (neu), Cross-Refs auf `qwen3.8-27b-alibaba.md`, `kimi-k3.md`, `ai-intelligence-commoditization-thesis.md`, `open-weights-economics.md`, `chinese-model-cost-routing.md`. Geteilt von Pit (@PWeber) in OME Topic "News & Infos". Praxis-Benchmark: Open-Weighting-Modelle (Kimi K3, MiniMax M3, QWEN 3.8 Max/27B, DeepSeek V4 Flash Vision Exp) auf mehrstufige SOLIDWORKS-CAD-Aufgabe; DeepSeek V4 Flash Vision Exp ~95 % für ~0,53 $.)* +*Letzte Aktualisierung: 2026-08-23 (123. Update — Ox Alpha: anonymes KI-Modell auf OpenRouter (Insider leak of the day, 2026-08-22). Raw: `raw/xpost/2026-08-23_insiderleak-ox-alpha-anonymous-model.md`. Wiki-Update: `concepts/llm/ox-alpha-anonymous-model.md` (neu), Cross-Refs auf `glm-5.3-z-ai.md`, `glm-5.2-zai-coding-model.md`, `chinese-model-cost-routing.md`, `open-weights-economics.md`, `openrouter-cad-solidworks-benchmark.md`, `chinese-models-top-usage-charts.md`, Institutionen `z-ai.md` + `openrouter.md`. Geteilt von Netbits (@NetLightning) in OME Topic "News & Infos". Gerücht/Leak: mysteriöses anonymes Modell auf OpenRouter (1M-Token-Kontext, multimodal, kein Owner), soll beim Coding Claude Fable 5 + GPT-5.6 Sol schlagen; GLM-identischer Tokenizer; vier vorherige anonyme Drops von chinesischen Labs beansprucht.)* + +*Vorheriges Update: 2026-08-22 (122. Update — OpenRouter CAD/SOLIDWORKS-Benchmark (@jack67396128, 2026-08-22). Raw: `raw/xpost/2026-08-22_jack-openrouter-cad-solidworks-benchmark.md`. Wiki-Update: `concepts/llm/openrouter-cad-solidworks-benchmark.md` (neu), Cross-Refs auf `qwen3.8-27b-alibaba.md`, `kimi-k3.md`, `ai-intelligence-commoditization-thesis.md`, `open-weights-economics.md`, `chinese-model-cost-routing.md`. Geteilt von Pit (@PWeber) in OME Topic "News & Infos". Praxis-Benchmark: Open-Weighting-Modelle (Kimi K3, MiniMax M3, QWEN 3.8 Max/27B, DeepSeek V4 Flash Vision Exp) auf mehrstufige SOLIDWORKS-CAD-Aufgabe; DeepSeek V4 Flash Vision Exp ~95 % für ~0,53 $.)* *Vorheriges Update: 2026-08-21 (121. Update — Mark Klinge: „Wie die Hüter des Geldes auf ganzer Linie versagen“ (The Sovereign Loop, 2026-08-19). Volltext per Web-Fetch ausgelesen (URL: `markklinge.substack.com`; Artikel signiert „The Sovereign Loop“). Raw: `raw/blog/2026-08-21_mark-klinge-huter-des-geldes.md`. Wiki-Updates: `people/mark-klinge.md` (neu), `concepts/finance/sovereign-loop-bitcoin-powerlaw.md` (Update 21.08.: Volltext — Black Wednesday 1992, SNB-Mindestkurs 1.20 CHF/EUR, Japan/Yen-Carry-Trade, Österreichische Schule/Mises/Hayek, Bitcoin-Boden-These ~60k, Halving-Zyklus 2028), Cross-Refs in `concepts/finance/sovereign-loop-antithese.md`, `concepts/finance/gold-zinsdruck-kapitalstruktur.md`, `concepts/finance/schuldenkrise-anleihenmarkt.md` (US-Debt ~40 Bio. Kontext). Identitäts-Hinweis: Artikel signiert „The Sovereign Loop“; bisherige Zuordnung im Wiki: Rüdiger Stelariz/@thesovereignloop (offene Identitätsfrage). Geteilt von Netbits (@NetLightning) in OME Topic „News & Infos“)* @@ -105,6 +107,7 @@ | [GLM 5.2 (Z.ai) — Chinese Frontier Coding Model](concepts/llm/glm-5.2-zai-coding-model.md) | 10x günstiger als Claude, 1M Kontext, MIT-Lizenz, Z.ai Coding Plan, **nativ in OpenClaw v2026.6.8**. Update 22.06.: Arnie-Review mit 4 Tests, Self-Hosting-Pfade (LM Studio, Unsloth, DwarfStar), Kosten-Analyse. **Update 29.06.:** Semgrep IDOR-Benchmark ≈ Opus 4.8 bei Schwachstellen-Suche, Reward Hacking im RL-Training, DSGVO-konforme Security-Nutzung, Geopolitik. **Update 01.07.:** #1 Open-Weights auf Artificial Analysis Intelligence Index v4.1 (Score 51, 4th worldwide), SWE-bench Pro 62.1 beats GPT-5.5, Industry praise from Rauch/Levie/Howard. **Update 02.07.:** atomic.chat One-Shot Benchmark — B+ at $0.08, 39× cheaper than Fable 5, 6th independent validation | youtube/2026-06-15_ichbinfabian-glm-5.2-coding-modell.md + other/2026-06-16_openclaw-releases-v2026.6.8.md + youtube/2026-06-22_ai-mit-arnie-glm-5-2-review.md + blog/2026-06-29_heise-glm52-hacking-cybersecurity.md + blog/2026-07-01_perplexity-glm52-tops-open-weights-intelligence-index.md + xpost/2026-07-02_atomicchat-coding-benchmark-fable5-gpt55-opus48-glm52.md | | [GLM 5.3 (Z.ai) — Neue Generation der GLM-Serie](concepts/llm/glm-5.3-z-ai.md) | AICodeKing Early Access + Bench #1 (2026-08-14). Neue Generation der GLM-Familie (5.0→5.1→5.2→5.3, GLM 5.5 angekündigt). Viert schnellste Frontier-Kadenz der Branche. Relevanz für Model-Routing, da GLM-5.2 Hectors Primary-Modell | youtube/2026-08-14_glm-5.3-aicodeking.md | | [GLM-5.5 (Z.ai) — Trillion-Parameter Announcement](concepts/llm/glm-5.5-z-ai.md) | Successor to GLM 5.2. **>1T parameters**, 1M context, open weights, August 2026 launch. Agent/coding focus. Fourth Chinese AI announcement in four days (20.07.2026). Part of [[concepts/chinese-ai-wave-july-2026.md]]. Comparison table vs GLM 5.2 | xpost/2026-07-20-healthranger-four-chinese-models.md | +| [Ox Alpha — Anonymes KI-Modell auf OpenRouter](concepts/llm/ox-alpha-anonymous-model.md) | Gerücht/Leak (2026-08-22, Insider leak of the day): mysteriöses anonymes Modell auf OpenRouter — 1M-Token-Kontext, multimodal, kein Owner — soll beim Coding Claude Fable 5 + GPT-5.6 Sol schlagen. GLM-identischer Tokenizer; vier vorherige anonyme Drops von chinesischen Labs beansprucht. Offene Herkunfts-Frage (Google vs. Z.ai). ⚠️ Unbestätigt | xpost/2026-08-23_insiderleak-ox-alpha-anonymous-model.md | | [Qwen3.8-27B (Alibaba) — Compact Frontier, "Intelligence Density"](concepts/llm/qwen3.8-27b-alibaba.md) | HuggingFace-Release (Countdown bis 14.08.2026, 4.928 wartend). Kompaktes 27B-Modell der Qwen3.8-Generation mit "unmatched intelligence density". Kontrast zum 2.4T-MoE von Qwen 3.8. Lokal-relevant (27B läuft auf Consumer-HW). Release am selben Tag wie GLM-5.3-Review — chinesischer Release-Zyklus. **Update 18.08.:** jetzt auf Ollama lauffähig (`ollama run qwen3.8:27b`), dichte 27,8B-Architektur, Hybrid-Attention, 262k-Kontext (bis 1M via YaRN), multimodal, MTP-markierte Ollama-Tags für Inferenz-Speedup. **Update 19.08.:** DFlash 2 (Z Lab → Inco AI) erreicht 70 tok/s auf M5 Max MacBook Pro — bis 4,6× schneller als autoregressives Decoding via Speculative Decoding (Jun Song: „biggest breakthrough in local AI this year", nächste Innovation in Prefill/Gewichtskompression). Uncensored-Debatte: gregpr07 („no gates") vs. s1gmoid-Gegenposition (Verhältnismäßigkeit). **Update 20.08.:** Unsloth Dynamic 3.0-GGUFs für Qwen3.8-27B — neue UD-…-Dateien deutlich kleiner (UD-IQ1_S 6,2 GB bis Q6_K 22 GB), Qualität-zu-Größe verbessert, ≠ MTP/Speculative-Decoding | other/2026-08-14_qwen3.8-27b-huggingface-release.md + youtube/2026-08-18_qwen-38-27b-ollama-julian-goldie.md + xpost/2026-08-19_junsong-dflash2-speculative-decoding.md + xpost/2026-08-19_gregpr07-qwen38-uncensored.md + xpost/2026-08-20_teksedge-unsloth-dynamic-3.0-qwen38.md | | [DeepSeek V4-Pro GA + Harness v0.1 (Open Source)](concepts/llm/deepseek-v4-pro-ga-harness-open-source.md) | DeepSeek launcht 13.08.2026 Open Source: V4-Pro GA (App/Web/API, Reasoning-Effort low/high/max, OpenAI-Responses-API + Codex, Peak/Off-Peak-Pricing) + DeepSeek Harness v0.1 (MIT, Open-Source-Agent-Harness, Rivale zu Claude Code). Dritter Baustein der chinesischen Welle in 24h. **Update 21.08.:** Deep-Dive-Video zum Harness (Stephen G. Pope) + Turing-Post-TV-Video "The End Of Coding Agents" | other/2026-08-14_deepseek-v4-pro-ga-harness-open-source.md + youtube/2026-08-13_deepseek-v4-pro-0813-aicodeking.md + youtube/2026-08-21_deepseek-agent-harness-explained.md + youtube/2026-08-21_deepseek-harness-end-of-coding-agents.md | | [Chinesische Modelle räumen global die Usage-Charts ab](concepts/llm/chinese-models-top-usage-charts.md) | OpenRouter: Top-5 der wöchentlichen Token-Nutzung (28.07.–03.08.2026) alle chinesisch, 56,8 Bio. Tokens, 15 Wochen in Folge führend, DeepSeek-V4-Flash Platz 1. Kimi-K3-Schock (2,8T, größtes Open-Weight-Modell, GPU-Kapazität nach 48h erschöpft, PHLX-Semi-Index −20%). Vierter Baustein der chinesischen Welle in 24h — jetzt auf Marktanteils-Ebene | other/2026-08-14_chinese-models-top-usage-charts-kimi-k3-shock.md | diff --git a/wiki/log.md b/wiki/log.md index 3389ae4..f15a72a 100644 --- a/wiki/log.md +++ b/wiki/log.md @@ -2,6 +2,17 @@ *Append-only changelog. Start: 2026-06-05* +## 2026-08-23 — Ox Alpha: Anonymes KI-Modell auf OpenRouter (Insider leak of the day) + +**Type:** ingest | **Scope:** raw/xpost (1 new), wiki/concepts/llm (1 new), wiki/index, wiki/log +**Source:** Telegram-Leak "Insider leak of the day" (Forward 2026-08-22T15:33:34Z), geteilt von Netbits (@NetLightning) in OME Topic "News & Infos" (#11944) am 2026-08-23. +**Inhalt:** Gerücht/Leak über ein mysteriöses anonymes KI-Modell "Ox Alpha" auf OpenRouter (1M-Token-Kontext, multimodal, kein offengelegter Owner), das beim Coding Claude Fable 5 und GPT-5.6 Sol schlagen soll. Spekulation: Google vs. Z.ai; GLM-identischer Tokenizer; vier vorherige anonyme Modell-Drops von chinesischen Labs beansprucht. ⚠️ Unbestätigt (Leak-Quelle). Begleitend OpenRouter-Nutzungs-Databreakdown "AUG 21" (ox-alpha ~2.6T Tokens, hervorgehoben). +**Wiki-Update:** `concepts/llm/ox-alpha-anonymous-model.md` (neu). Cross-Refs auf `glm-5.3-z-ai.md`, `glm-5.2-zai-coding-model.md`, `chinese-model-cost-routing.md`, `open-weights-economics.md`, `openrouter-cad-solidworks-benchmark.md`, `chinese-models-top-usage-charts.md`, Institutionen `z-ai.md` + `openrouter.md`. +- raw (NEW): `raw/xpost/2026-08-23_insiderleak-ox-alpha-anonymous-model.md` +- wiki (NEW): `wiki/concepts/llm/ox-alpha-anonymous-model.md` +- index.md: 123. Update, neue Konzept-Seite + Raw-Katalog (+1) +- log: this entry + ## 2026-08-22 — OpenRouter CAD/SOLIDWORKS-Benchmark (@jack67396128) **Type:** ingest | **Scope:** raw/xpost (1 new), wiki/concepts/llm (1 new), wiki/index, wiki/log