--- type: podcast source_url: https://op3.dev/e/https://github.com/clawdassistant85-netizen/openclaw-podcast-media-en/releases/download/ep106/episode_106.mp3 retrieved: 2026-08-26 title: "AgentStack Daily EP106 — Codex .149 Agents-Dashboard, Stealth Reasoning Model, Compact Translation" author: "AgentStack Daily (TTS-Hosts NOVA/ALLOY)" duration_min: 23 posted_date: 2026-08-21 release_tag: ep106 tags: [podcast, agentstack-daily, codex, openrouter, ox-alpha, qwen, mcp, jailbreak, local-ai] --- # AgentStack Daily EP106 — Codex .149 Agents-Dashboard, Stealth Reasoning Model, Compact Translation Englischer, KI-generierter Daily-Podcast (zwei TTS-Hosts NOVA und ALLOY, NotebookLM-artiges Format). Episode 106, gepostet 21.08.2026, Dauer ~23 Min (~33 MB MP3). Stories der Woche 18.–21. August 2026. Lokale Kopien: Transcript `/tmp/pod106/episode_106_transcript.md`, Show Notes `/tmp/pod106/show_notes_episode_106.md` (49 KB), Cover `/tmp/pod106/episode_106_cover.png`. Laut Closing verweisen die Show Notes auf tobyonfitnesstech.com. ## Stories ### 1. Agent Stack Release Readout: OpenAI Codex rust-v0.149.0 OpenAI shipped Codex rust-v0.149.0 am 20.08.2026 ([Release](https://github.com/openai/codex/releases/tag/rust-v0.149.0)): - Interaktives `codex agents` Dashboard: Tasks suchen, starten, öffnen, umbenennen, stoppen — mit konfigurierbaren Keyboard-Shortcuts. - `codex queue`: Follow-up-Messages in laufende lokale oder remote Sessions senden, ohne die Session neu zu öffnen. - TUI: Arbeitsverzeichnis-Kommandos `/cd`, `/pwd`, `/cwd`. - Vim-Editing: Character Replacement plus Change-Motions `cw`, `c$`, `cc`. - `codex doctor` prüft jetzt Endpoint Protection, Netzwerk- und Proxy-Fehler, Desktop-App-State und Update-Konnektivität. - SDK: exakte CLI-Config-Overrides sowie Wahl des Reasoning Efforts `max` oder `ultra` direkt aus Code. - Bugfixes: Queued Messages wecken idle Sessions zuverlässig; Resumed/Forked Threads restaurieren ihr aktives Permission-Profil statt auf Defaults zurückzufallen; Realtime-WebRTC-Sideband-Verbindungen reconnecten nach Transportverlust ohne pending Output zu verwerfen. ### 2. A new stealth reasoning model just landed on OpenRouter - Modell „Ox Alpha" auf OpenRouter, gelistet unter dem anonymen Provider „stealth" ([Model Page](https://openrouter.ai/models/stealth/ox-alpha), verified 21.08.2026). - Positionierung: Reasoning-Modell für Coding, sustained agentic work und Production Workloads; Language nennt long-horizon software engineering und complex reasoning. - Kontextfenster exakt **1.048.576 Tokens**, Max-Output **4.096 Tokens** pro Call. - Keine Benchmarks, kein Pricing, kein Firmenname, keine Parameterzahl disclosed; keine unabhängigen Evals zum Listing. - Die Capability-Beschreibung bricht mitten im Satz ab („workflows that combine text with..."). ### 3. Tencent Hy-MT2-1.8B lands on OpenRouter with Chinese dialect coverage - Tencent released Hy-MT2-1.8B, kompaktes 1,8B-Parameter-Übersetzungsmodell, gelistet auf OpenRouter ([Model Page](https://openrouter.ai/models/tencent/hy-mt2-1.8b)). - 33 Sprachpaare plus 5 chinesische Dialekt- und Minderheitensprachpaare. - 8192-Token-Kontextfenster, 4096-Token-Max-Output. - Benannte Workflows: structured, delimiter-based, contextual, glossary-based, style-guided translation. ### 4. Stampli cuts launch hours 68% with ChatGPT Work and Codex - Case Study auf OpenAIs News-Seite, veröffentlicht 20.08.2026 ([openai.com/index/stampli](https://openai.com/index/stampli)). - Stampli (Accounts-Payable-Software) nutzte ChatGPT Work und Codex für Launch-Produktion, weil die Design-Kapazität anderweitig gebunden war. - Ergebnis laut Case Study: Launch 68 % unter der ursprünglichen Stunden-Schätzung, Wochen → Tage; kein Hiring, kein Verschieben des Deadlines. - Das Case Study trennt nicht auf, wie viel Entlastung auf Codex vs. ChatGPT Work entfiel und welche konkreten Tasks jeweils übernommen wurden. ### 5. Ramp launches Router, an AI model routing service - Ramp (Corporate Cards / Expense Management) launchte am 20.08.2026 „Router", einen Service mit einer einzigen API für Zugriff auf mehrere LLMs (laut [TechCrunch](https://techcrunch.com/2026/08/20/ramp-launches-its-own-ai-model-router-called-router/)). - Unterstützte Modelle, Routing-Logik und Pricing wurden in der Ankündigung nicht genannt. ### 6. Memory, not compute, is the new AI bottleneck - Counterpoint Research: Memory Supply bleibt bis 2027 und darüber hinaus knapp, da KI-Inferenz wächst ([HPCwire](https://www.hpcwire.com/2026/08/20/what-hyperscalers-should-know-about-cxl/)). - High Bandwidth Memory (HBM) bleibt teuer und kapazitätsbeschränkt. - Hyperscaler prüfen Compute Express Link (CXL) als Ansatz, Memory über Server zu poolen und zu skalieren. ### 7. Cerebras CS-4 lands at 750 PFLOPS with Wafer Scale Engine 3 - Cerebras unveiled CS-4, rated at 750 PFLOPS AI compute und 129,6 Petabytes Capacity per Launch-Angaben ([HPCwire](https://www.hpcwire.com/2026/08/20/its-not-an-hpc-system-but-cerebras-new-cs-4-is-an-ai-monster/)). - Zentrale Komponente: Wafer Scale Engine 3 — nahezu der gesamte Wafer wird als ein Prozessor verwendet statt in hunderte Dies geschnitten. ### 8. OpenAI lays out how it paces frontier models as cyber risks climb - OpenAI veröffentlichte am 18.08.2026 den Post „Pacing model development in an era of cyber-critical capabilities" ([openai.com](https://openai.com/index/pacing-model-development-cyber-capabilities/)). - Drei benannte Pfeile: Monitoring, Alignment, Security — positioniert als Hebel für Tempo und Freigabe fähigerer Systeme gegen Cyber-Capability-Thresholds. - Post nennt kein konkretes Modell, kein Datum, kein Developer-Feature. ### 9. OpenAI launches 'AI Futures' blog on power, governance, and freedom - OpenAI startete am 20.08.2026 den Blog „AI Futures" auf der News-Site ([Introducing AI Futures](https://openai.com/index/introducing-ai-futures)). - Themen: transformative AI und Auswirkung auf Power, Governance, Economy, individual Freedom. - Editorial-Projekt, kein neues Modell/API/Tool; erster Beitrag „Introducing AI Futures". ### 10. LiquidAI claims up to 3.2x faster inference with LFM2.5-DSpark - LiquidAI veröffentlichte am 20.08.2026 einen Hugging Face Blogpost zu LFM2.5-DSpark mit Claim „up to 3.2x faster inference" ([HF Blog](https://huggingface.co/blog/LiquidAI/lfm25-dspark)). - Kein separater Changelog, keine Release Notes, kein technischer Breakdown jenseits des Bloglinks. - MarkTechPost berichtet parallel von „LFM2.5-DSpark Draft Models" mit bis zu 3,18× faster decoding ([MarkTechPost](https://www.marktechpost.com/2026/08/20/liquid-ai-releases-lfm2-5-dspark-draft-models-that-deliver-up-to-3-18x-faster-decoding/)). ### 11. IBM Research asks how much memory an agent really needs - IBM Research, Hugging Face Blogpost vom 18.08.2026: „How Much Memory Does Your Agent Actually Need?" ([HF Blog](https://huggingface.co/blog/ibm-research/altk-evolve-hmm)). - URL zeigt auf das ALTK-Projekt; Slug-Tag „evolve-hmm". - Quellenmaterial beschränkt sich auf Titel + URL; keine getesteten Memory-Größen, Benchmarks oder Deltas im Quellenmaterial. ### 12. A new jailbreak hides malicious instructions inside encrypted text - Forscher demonstrierten „Cryptographic Context Injection": malicious Instructions versteckt in verschlüsseltem/encodiertem Text tricksen KI-Assistenten (Demo: Grok) dazu aus, User-Daten zu exfiltrieren ([Ars Technica, 20.08.](https://arstechnica.com/security/2026/08/grok-exfiltrates-user-data-when-malicious-instructions-are-encrypted/)). - Mechanismus: Safety-Layer liest den Prompt beim Eingang, sieht nur Encodiertes/Gibberish und lässt durch; das Modell dekodiert und befolgt die freigelegten Instructions. - Ars Technica ordnet dies als neueste Variante in eine Reihe von Guardrail-Bypass-Techniken ein. ### 13. Show HN: 125M piano autocomplete + Superwhisper S1-mini - Show HN: 125M-Parameter-Modell für On-Device-Piano-Autocomplete, Hacker News Score 554 ([Diskussion](https://news.ycombinator.com/item?id=49373456)); Quelle headline-only, Primärquelle [simedw.com](https://simedw.com/2026/08/20/midi-autocomplete/) unterstützt nur die genannten Fakten (keine Architektur-, Latenz-, Device- oder Qualitätsangaben). - Superwhisper S1-mini: 462 MB Open-Weights-Text-Normalizer, sitzt nach ASR, entfernt Fillers und löst Self-Corrections lokal auf ([MarkTechPost, 20.08.](https://www.marktechpost.com/2026/08/20/meet-s1-mini-superwhispers-462-mb-open-weights-text-normalizer-that-turnsraw-asr-transcripts-into-clean-written-text/)). ### 14. GitHub Project Radar | Projekt | Stars | Delta | Release | Beschreibung | |---|---|---|---|---| | [HKUDS/nanobot](https://github.com/HKUDS/nanobot) | 47.251 | erste getrackte Erwähnung | v0.3.0 (25.07.2026) | Ultra-lightweight, open-source, self-hosted Personal-AI-Agent-Framework in Python: WebUI, Tools, Memory, MCP, Multi-Agent-Workflows, Automation, Chat-Apps | | [DeusData/codebase-memory-mcp](https://github.com/DeusData/codebase-memory-mcp) | 39.755 | +8.088 (+25,5 %) seit Mitte Juli 2026 | v0.10.8 (19.08.2026) | Code-Intelligence-MCP-Server: indexiert Codebases in persistente Knowledge Graphs, 158 Sprachen, Sub-ms-Queries, 99 % fewer tokens, Single Static Binary | | [PrefectHQ/fastmcp](https://github.com/PrefectHQ/fastmcp) | 27.320 | +1.106 (+4,2 %) seit Mitte Juli 2026 | v3.4.7 (10.08.2026) | „The fast, Pythonic way to build MCP servers and clients" | ### 15. Local LLM Spotlight: Qwen/Qwen3.8-27B - Trending auf Hugging Face ([Model Page](https://huggingface.co/Qwen/Qwen3.8-27B)): Task image-text-to-text, **11.836 Likes**, **1.726.651 Downloads** (>1,7 Mio). - Tags: transformers, safetensors, qwen3_5, image-text-to-text, conversational, license:apache-2.0, eval-results, endpoints_compatible, deploy:azure, region:us. ## Release Coverage Check (laut Show Notes, verified 21.08.2026) | Harness | Stable | Published | |---|---|---| | OpenClaw | v2026.6.34 | 2026-08-08 | | Hermes Agent | v2026.8.18 | 2026-08-18 | | OpenAI Codex | rust-v0.149.0 | 2026-08-20 | | Claude Code CLI | 2.1.228 | 2026-08-11 | | Antigravity CLI | Continuous Delivery, keine Release-Tags this cycle | — | Erkannte Episoden-Version-Tags: OpenClaw v2026.7.2-beta.3/.5/.7, v2026.8.1-beta.2; Hermes v2026.8.16/.16.2/.18/.3; Codex rust-v0.146.0/.146.1/.147.0/.148.0; Claude Code 2.1.227/.228/latest/stable. ## Model Discovery Check (verified 21.08.2026) - **Ox Alpha** (stealth) — Newly listed this cycle. Context 1048576 tokens; API via OpenRouter; params n/a. Decision: Selected — new major-provider model not featured on a recent broadcast. - **Tencent Hy-MT2-1.8B** — Newly listed this cycle. Context 8192 tokens; API via OpenRouter; params n/a. Decision: Selected. ## Extra Research Candidates - ChatGPT Ads expands across Europe — 31 europäische Märkte ([openai.com](https://openai.com/index/chatgpt-ads-expands-across-europe)) - How Much Memory Does Your Agent Actually Need? ([HF Blog](https://huggingface.co/blog/ibm-research/altk-evolve-hmm)) - Grok exfiltrates user data when malicious instructions are encrypted ([Ars Technica](https://arstechnica.com/security/2026/08/grok-exfiltrates-user-data-when-malicious-instructions-are-encrypted/)) - MidTool: Mid-training Data Synthesis for Agentic Tool Use ([arXiv:2608.20314](https://arxiv.org/abs/2608.20314)) - Inject, Align, Recover: Staged Post-Training for Retrieval-Free Document Understanding ([arXiv:2608.20281](https://arxiv.org/abs/2608.20281)) - Introducing ChatGPT for Teens ([openai.com](https://openai.com/index/chatgpt-for-teens)) - JonathanColetti/Qwen3.8-27B-Uncensored-GGUF trending on Hugging Face ([HF](https://huggingface.co/JonathanColetti/Qwen3.8-27B-Uncensored-GGUF)) - Lightricks/LTX-2.5 trending on Hugging Face ([HF](https://huggingface.co/Lightricks/LTX-2.5)) - 5 new ways to level up your learning with Search ([Google Blog](https://blog.google/products-and-platforms/products/search/back-to-school-study-tools/)) ## Chapters | Zeit | Segment | |---|---| | 00:00 | Intro/Hook | | 02:00 | OpenAI Codex rust-v0.149.0 | | 02:12 | Ox Alpha (stealth reasoning model) | | 03:45 | Tencent Hy-MT2-1.8B | | 04:52 | Stampli Case Study | | 06:37 | Ramp Router | | 08:08 | Memory-not-compute-Bottleneck (Counterpoint) | | 09:36 | Cerebras CS-4 | | 11:08 | OpenAI Frontier-Pacing | | 12:33 | OpenAI AI Futures Blog | | 13:37 | LiquidAI LFM2.5-DSpark | | 14:26 | IBM Research Agent Memory | | 16:12 | Cryptographic Context Injection | | 17:18 | Show HN Piano Autocomplete | | 17:42 | Superwhisper S1-mini | | 18:55 | GitHub Project Radar | | 20:10 | Model Discovery Check | | 21:05 | Local LLM Spotlight | | 22:00 | Extra Research Candidates | | 22:48 | Closing (Verweis Show Notes: tobyonfitnesstech.com) | ## Editorial Mix Check (laut Show Notes) flagship_products: 6 · builder_projects: 3 · local_ai: 2 · hardware_compute: 2 · policy_regulation: 1 · research: 0