173 lines
13 KiB
Markdown
173 lines
13 KiB
Markdown
---
|
||
type: podcast
|
||
source_url: https://op3.dev/e/https://github.com/clawdassistant85-netizen/openclaw-podcast-media-en/releases/download/ep106/episode_106.mp3
|
||
retrieved: 2026-08-26
|
||
title: "AgentStack Daily EP106 — Codex .149 Agents-Dashboard, Stealth Reasoning Model, Compact Translation"
|
||
author: "AgentStack Daily (TTS-Hosts NOVA/ALLOY)"
|
||
duration_min: 23
|
||
posted_date: 2026-08-21
|
||
release_tag: ep106
|
||
tags: [podcast, agentstack-daily, codex, openrouter, ox-alpha, qwen, mcp, jailbreak, local-ai]
|
||
---
|
||
|
||
# AgentStack Daily EP106 — Codex .149 Agents-Dashboard, Stealth Reasoning Model, Compact Translation
|
||
|
||
Englischer, KI-generierter Daily-Podcast (zwei TTS-Hosts NOVA und ALLOY, NotebookLM-artiges Format). Episode 106, gepostet 21.08.2026, Dauer ~23 Min (~33 MB MP3). Stories der Woche 18.–21. August 2026. Lokale Kopien: Transcript `/tmp/pod106/episode_106_transcript.md`, Show Notes `/tmp/pod106/show_notes_episode_106.md` (49 KB), Cover `/tmp/pod106/episode_106_cover.png`. Laut Closing verweisen die Show Notes auf tobyonfitnesstech.com.
|
||
|
||
## Stories
|
||
|
||
### 1. Agent Stack Release Readout: OpenAI Codex rust-v0.149.0
|
||
|
||
OpenAI shipped Codex rust-v0.149.0 am 20.08.2026 ([Release](https://github.com/openai/codex/releases/tag/rust-v0.149.0)):
|
||
|
||
- Interaktives `codex agents` Dashboard: Tasks suchen, starten, öffnen, umbenennen, stoppen — mit konfigurierbaren Keyboard-Shortcuts.
|
||
- `codex queue`: Follow-up-Messages in laufende lokale oder remote Sessions senden, ohne die Session neu zu öffnen.
|
||
- TUI: Arbeitsverzeichnis-Kommandos `/cd`, `/pwd`, `/cwd`.
|
||
- Vim-Editing: Character Replacement plus Change-Motions `cw`, `c$`, `cc`.
|
||
- `codex doctor` prüft jetzt Endpoint Protection, Netzwerk- und Proxy-Fehler, Desktop-App-State und Update-Konnektivität.
|
||
- SDK: exakte CLI-Config-Overrides sowie Wahl des Reasoning Efforts `max` oder `ultra` direkt aus Code.
|
||
- Bugfixes: Queued Messages wecken idle Sessions zuverlässig; Resumed/Forked Threads restaurieren ihr aktives Permission-Profil statt auf Defaults zurückzufallen; Realtime-WebRTC-Sideband-Verbindungen reconnecten nach Transportverlust ohne pending Output zu verwerfen.
|
||
|
||
### 2. A new stealth reasoning model just landed on OpenRouter
|
||
|
||
- Modell „Ox Alpha" auf OpenRouter, gelistet unter dem anonymen Provider „stealth" ([Model Page](https://openrouter.ai/models/stealth/ox-alpha), verified 21.08.2026).
|
||
- Positionierung: Reasoning-Modell für Coding, sustained agentic work und Production Workloads; Language nennt long-horizon software engineering und complex reasoning.
|
||
- Kontextfenster exakt **1.048.576 Tokens**, Max-Output **4.096 Tokens** pro Call.
|
||
- Keine Benchmarks, kein Pricing, kein Firmenname, keine Parameterzahl disclosed; keine unabhängigen Evals zum Listing.
|
||
- Die Capability-Beschreibung bricht mitten im Satz ab („workflows that combine text with...").
|
||
|
||
### 3. Tencent Hy-MT2-1.8B lands on OpenRouter with Chinese dialect coverage
|
||
|
||
- Tencent released Hy-MT2-1.8B, kompaktes 1,8B-Parameter-Übersetzungsmodell, gelistet auf OpenRouter ([Model Page](https://openrouter.ai/models/tencent/hy-mt2-1.8b)).
|
||
- 33 Sprachpaare plus 5 chinesische Dialekt- und Minderheitensprachpaare.
|
||
- 8192-Token-Kontextfenster, 4096-Token-Max-Output.
|
||
- Benannte Workflows: structured, delimiter-based, contextual, glossary-based, style-guided translation.
|
||
|
||
### 4. Stampli cuts launch hours 68% with ChatGPT Work and Codex
|
||
|
||
- Case Study auf OpenAIs News-Seite, veröffentlicht 20.08.2026 ([openai.com/index/stampli](https://openai.com/index/stampli)).
|
||
- Stampli (Accounts-Payable-Software) nutzte ChatGPT Work und Codex für Launch-Produktion, weil die Design-Kapazität anderweitig gebunden war.
|
||
- Ergebnis laut Case Study: Launch 68 % unter der ursprünglichen Stunden-Schätzung, Wochen → Tage; kein Hiring, kein Verschieben des Deadlines.
|
||
- Das Case Study trennt nicht auf, wie viel Entlastung auf Codex vs. ChatGPT Work entfiel und welche konkreten Tasks jeweils übernommen wurden.
|
||
|
||
### 5. Ramp launches Router, an AI model routing service
|
||
|
||
- Ramp (Corporate Cards / Expense Management) launchte am 20.08.2026 „Router", einen Service mit einer einzigen API für Zugriff auf mehrere LLMs (laut [TechCrunch](https://techcrunch.com/2026/08/20/ramp-launches-its-own-ai-model-router-called-router/)).
|
||
- Unterstützte Modelle, Routing-Logik und Pricing wurden in der Ankündigung nicht genannt.
|
||
|
||
### 6. Memory, not compute, is the new AI bottleneck
|
||
|
||
- Counterpoint Research: Memory Supply bleibt bis 2027 und darüber hinaus knapp, da KI-Inferenz wächst ([HPCwire](https://www.hpcwire.com/2026/08/20/what-hyperscalers-should-know-about-cxl/)).
|
||
- High Bandwidth Memory (HBM) bleibt teuer und kapazitätsbeschränkt.
|
||
- Hyperscaler prüfen Compute Express Link (CXL) als Ansatz, Memory über Server zu poolen und zu skalieren.
|
||
|
||
### 7. Cerebras CS-4 lands at 750 PFLOPS with Wafer Scale Engine 3
|
||
|
||
- Cerebras unveiled CS-4, rated at 750 PFLOPS AI compute und 129,6 Petabytes Capacity per Launch-Angaben ([HPCwire](https://www.hpcwire.com/2026/08/20/its-not-an-hpc-system-but-cerebras-new-cs-4-is-an-ai-monster/)).
|
||
- Zentrale Komponente: Wafer Scale Engine 3 — nahezu der gesamte Wafer wird als ein Prozessor verwendet statt in hunderte Dies geschnitten.
|
||
|
||
### 8. OpenAI lays out how it paces frontier models as cyber risks climb
|
||
|
||
- OpenAI veröffentlichte am 18.08.2026 den Post „Pacing model development in an era of cyber-critical capabilities" ([openai.com](https://openai.com/index/pacing-model-development-cyber-capabilities/)).
|
||
- Drei benannte Pfeile: Monitoring, Alignment, Security — positioniert als Hebel für Tempo und Freigabe fähigerer Systeme gegen Cyber-Capability-Thresholds.
|
||
- Post nennt kein konkretes Modell, kein Datum, kein Developer-Feature.
|
||
|
||
### 9. OpenAI launches 'AI Futures' blog on power, governance, and freedom
|
||
|
||
- OpenAI startete am 20.08.2026 den Blog „AI Futures" auf der News-Site ([Introducing AI Futures](https://openai.com/index/introducing-ai-futures)).
|
||
- Themen: transformative AI und Auswirkung auf Power, Governance, Economy, individual Freedom.
|
||
- Editorial-Projekt, kein neues Modell/API/Tool; erster Beitrag „Introducing AI Futures".
|
||
|
||
### 10. LiquidAI claims up to 3.2x faster inference with LFM2.5-DSpark
|
||
|
||
- LiquidAI veröffentlichte am 20.08.2026 einen Hugging Face Blogpost zu LFM2.5-DSpark mit Claim „up to 3.2x faster inference" ([HF Blog](https://huggingface.co/blog/LiquidAI/lfm25-dspark)).
|
||
- Kein separater Changelog, keine Release Notes, kein technischer Breakdown jenseits des Bloglinks.
|
||
- MarkTechPost berichtet parallel von „LFM2.5-DSpark Draft Models" mit bis zu 3,18× faster decoding ([MarkTechPost](https://www.marktechpost.com/2026/08/20/liquid-ai-releases-lfm2-5-dspark-draft-models-that-deliver-up-to-3-18x-faster-decoding/)).
|
||
|
||
### 11. IBM Research asks how much memory an agent really needs
|
||
|
||
- IBM Research, Hugging Face Blogpost vom 18.08.2026: „How Much Memory Does Your Agent Actually Need?" ([HF Blog](https://huggingface.co/blog/ibm-research/altk-evolve-hmm)).
|
||
- URL zeigt auf das ALTK-Projekt; Slug-Tag „evolve-hmm".
|
||
- Quellenmaterial beschränkt sich auf Titel + URL; keine getesteten Memory-Größen, Benchmarks oder Deltas im Quellenmaterial.
|
||
|
||
### 12. A new jailbreak hides malicious instructions inside encrypted text
|
||
|
||
- Forscher demonstrierten „Cryptographic Context Injection": malicious Instructions versteckt in verschlüsseltem/encodiertem Text tricksen KI-Assistenten (Demo: Grok) dazu aus, User-Daten zu exfiltrieren ([Ars Technica, 20.08.](https://arstechnica.com/security/2026/08/grok-exfiltrates-user-data-when-malicious-instructions-are-encrypted/)).
|
||
- Mechanismus: Safety-Layer liest den Prompt beim Eingang, sieht nur Encodiertes/Gibberish und lässt durch; das Modell dekodiert und befolgt die freigelegten Instructions.
|
||
- Ars Technica ordnet dies als neueste Variante in eine Reihe von Guardrail-Bypass-Techniken ein.
|
||
|
||
### 13. Show HN: 125M piano autocomplete + Superwhisper S1-mini
|
||
|
||
- Show HN: 125M-Parameter-Modell für On-Device-Piano-Autocomplete, Hacker News Score 554 ([Diskussion](https://news.ycombinator.com/item?id=49373456)); Quelle headline-only, Primärquelle [simedw.com](https://simedw.com/2026/08/20/midi-autocomplete/) unterstützt nur die genannten Fakten (keine Architektur-, Latenz-, Device- oder Qualitätsangaben).
|
||
- Superwhisper S1-mini: 462 MB Open-Weights-Text-Normalizer, sitzt nach ASR, entfernt Fillers und löst Self-Corrections lokal auf ([MarkTechPost, 20.08.](https://www.marktechpost.com/2026/08/20/meet-s1-mini-superwhispers-462-mb-open-weights-text-normalizer-that-turnsraw-asr-transcripts-into-clean-written-text/)).
|
||
|
||
### 14. GitHub Project Radar
|
||
|
||
| Projekt | Stars | Delta | Release | Beschreibung |
|
||
|---|---|---|---|---|
|
||
| [HKUDS/nanobot](https://github.com/HKUDS/nanobot) | 47.251 | erste getrackte Erwähnung | v0.3.0 (25.07.2026) | Ultra-lightweight, open-source, self-hosted Personal-AI-Agent-Framework in Python: WebUI, Tools, Memory, MCP, Multi-Agent-Workflows, Automation, Chat-Apps |
|
||
| [DeusData/codebase-memory-mcp](https://github.com/DeusData/codebase-memory-mcp) | 39.755 | +8.088 (+25,5 %) seit Mitte Juli 2026 | v0.10.8 (19.08.2026) | Code-Intelligence-MCP-Server: indexiert Codebases in persistente Knowledge Graphs, 158 Sprachen, Sub-ms-Queries, 99 % fewer tokens, Single Static Binary |
|
||
| [PrefectHQ/fastmcp](https://github.com/PrefectHQ/fastmcp) | 27.320 | +1.106 (+4,2 %) seit Mitte Juli 2026 | v3.4.7 (10.08.2026) | „The fast, Pythonic way to build MCP servers and clients" |
|
||
|
||
### 15. Local LLM Spotlight: Qwen/Qwen3.8-27B
|
||
|
||
- Trending auf Hugging Face ([Model Page](https://huggingface.co/Qwen/Qwen3.8-27B)): Task image-text-to-text, **11.836 Likes**, **1.726.651 Downloads** (>1,7 Mio).
|
||
- Tags: transformers, safetensors, qwen3_5, image-text-to-text, conversational, license:apache-2.0, eval-results, endpoints_compatible, deploy:azure, region:us.
|
||
|
||
## Release Coverage Check (laut Show Notes, verified 21.08.2026)
|
||
|
||
| Harness | Stable | Published |
|
||
|---|---|---|
|
||
| OpenClaw | v2026.6.34 | 2026-08-08 |
|
||
| Hermes Agent | v2026.8.18 | 2026-08-18 |
|
||
| OpenAI Codex | rust-v0.149.0 | 2026-08-20 |
|
||
| Claude Code CLI | 2.1.228 | 2026-08-11 |
|
||
| Antigravity CLI | Continuous Delivery, keine Release-Tags this cycle | — |
|
||
|
||
Erkannte Episoden-Version-Tags: OpenClaw v2026.7.2-beta.3/.5/.7, v2026.8.1-beta.2; Hermes v2026.8.16/.16.2/.18/.3; Codex rust-v0.146.0/.146.1/.147.0/.148.0; Claude Code 2.1.227/.228/latest/stable.
|
||
|
||
## Model Discovery Check (verified 21.08.2026)
|
||
|
||
- **Ox Alpha** (stealth) — Newly listed this cycle. Context 1048576 tokens; API via OpenRouter; params n/a. Decision: Selected — new major-provider model not featured on a recent broadcast.
|
||
- **Tencent Hy-MT2-1.8B** — Newly listed this cycle. Context 8192 tokens; API via OpenRouter; params n/a. Decision: Selected.
|
||
|
||
## Extra Research Candidates
|
||
|
||
- ChatGPT Ads expands across Europe — 31 europäische Märkte ([openai.com](https://openai.com/index/chatgpt-ads-expands-across-europe))
|
||
- How Much Memory Does Your Agent Actually Need? ([HF Blog](https://huggingface.co/blog/ibm-research/altk-evolve-hmm))
|
||
- Grok exfiltrates user data when malicious instructions are encrypted ([Ars Technica](https://arstechnica.com/security/2026/08/grok-exfiltrates-user-data-when-malicious-instructions-are-encrypted/))
|
||
- MidTool: Mid-training Data Synthesis for Agentic Tool Use ([arXiv:2608.20314](https://arxiv.org/abs/2608.20314))
|
||
- Inject, Align, Recover: Staged Post-Training for Retrieval-Free Document Understanding ([arXiv:2608.20281](https://arxiv.org/abs/2608.20281))
|
||
- Introducing ChatGPT for Teens ([openai.com](https://openai.com/index/chatgpt-for-teens))
|
||
- JonathanColetti/Qwen3.8-27B-Uncensored-GGUF trending on Hugging Face ([HF](https://huggingface.co/JonathanColetti/Qwen3.8-27B-Uncensored-GGUF))
|
||
- Lightricks/LTX-2.5 trending on Hugging Face ([HF](https://huggingface.co/Lightricks/LTX-2.5))
|
||
- 5 new ways to level up your learning with Search ([Google Blog](https://blog.google/products-and-platforms/products/search/back-to-school-study-tools/))
|
||
|
||
## Chapters
|
||
|
||
| Zeit | Segment |
|
||
|---|---|
|
||
| 00:00 | Intro/Hook |
|
||
| 02:00 | OpenAI Codex rust-v0.149.0 |
|
||
| 02:12 | Ox Alpha (stealth reasoning model) |
|
||
| 03:45 | Tencent Hy-MT2-1.8B |
|
||
| 04:52 | Stampli Case Study |
|
||
| 06:37 | Ramp Router |
|
||
| 08:08 | Memory-not-compute-Bottleneck (Counterpoint) |
|
||
| 09:36 | Cerebras CS-4 |
|
||
| 11:08 | OpenAI Frontier-Pacing |
|
||
| 12:33 | OpenAI AI Futures Blog |
|
||
| 13:37 | LiquidAI LFM2.5-DSpark |
|
||
| 14:26 | IBM Research Agent Memory |
|
||
| 16:12 | Cryptographic Context Injection |
|
||
| 17:18 | Show HN Piano Autocomplete |
|
||
| 17:42 | Superwhisper S1-mini |
|
||
| 18:55 | GitHub Project Radar |
|
||
| 20:10 | Model Discovery Check |
|
||
| 21:05 | Local LLM Spotlight |
|
||
| 22:00 | Extra Research Candidates |
|
||
| 22:48 | Closing (Verweis Show Notes: tobyonfitnesstech.com) |
|
||
|
||
## Editorial Mix Check (laut Show Notes)
|
||
|
||
flagship_products: 6 · builder_projects: 3 · local_ai: 2 · hardware_compute: 2 · policy_regulation: 1 · research: 0
|