knowledge-base/raw/xpost/2026-09-08_promiscfx-robocurve-gpt6-astra-yam.md

29 lines
1.6 KiB
Markdown
Raw Permalink Normal View History

---
type: xpost
source_url: https://x.com/promiscfx/status/2097080991475257839
retrieved: 2026-09-08
author: "@promiscfx / Cognitive f(x)"
is_thread: false
quote_count: 1
tags: [robotics, physical-ai, robocurve, gpt-6-astra, claude-fable-5.1, yam, benchmarking]
---
# Cognitive f(x): GPT-6 Astra vs. Claude Fable 5.1 auf YAM-Robotarmen
**Autor:** [@promiscfx / Cognitive f(x)](https://x.com/promiscfx)
**Quelle:** https://x.com/promiscfx/status/2097080991475257839
**Datum des Posts:** 07.09.2026, 21:54 UTC
**Kontext:** Von Pit Weber in OME Topic „aGi / eMergence“ geteilt, 08.09.2026.
## Vollständiger Post-Text
> Independent benchmarking on Robocurve recently tested OpenAI's GPT-6 Astra against Claude Fable 5.1 controlling YAM robot arms. The shift in performance is massive:
>
> 🔹 Task Success (Block-into-bowl): Jumped from 40% (Fable 5.1) to 95% with Astra.
> 🔹 Token Efficiency: Astra used ~2.1k output tokens per run vs 12.9k for Fable 5.1 (a ~6x reduction in token overhead).
> 🔹 Speed & Cost: Execution time dropped from 6.8 minutes down to 2.5 minutes, nearly cutting operational costs in half ($0.94 vs $2.12/run).
## Einordnung der Quelle
Der Post bezeichnet Robocurve als unabhängigen Benchmarking-Kontext und enthält ein 14,3-sekündiges Video. Die Website [Robocurve](https://robocurve.org/) beschreibt die Organisation als Public Benefit Corporation und unabhängigen Evaluator für Physical AI; sie entwickelt offene Werkzeuge und Benchmarks. Der konkrete Astra/Fable-5.1-Datensatz, die YAM-Versuchsparameter und die Zahlen oben wurden in diesem Ingest nicht direkt auf der Robocurve-Website verifiziert.