28 lines
1.6 KiB
Markdown
28 lines
1.6 KiB
Markdown
---
|
|
type: xpost
|
|
source_url: https://x.com/promiscfx/status/2097080991475257839
|
|
retrieved: 2026-09-08
|
|
author: "@promiscfx / Cognitive f(x)"
|
|
is_thread: false
|
|
quote_count: 1
|
|
tags: [robotics, physical-ai, robocurve, gpt-6-astra, claude-fable-5.1, yam, benchmarking]
|
|
---
|
|
|
|
# Cognitive f(x): GPT-6 Astra vs. Claude Fable 5.1 auf YAM-Robotarmen
|
|
|
|
**Autor:** [@promiscfx / Cognitive f(x)](https://x.com/promiscfx)
|
|
**Quelle:** https://x.com/promiscfx/status/2097080991475257839
|
|
**Datum des Posts:** 07.09.2026, 21:54 UTC
|
|
**Kontext:** Von Pit Weber in OME Topic „aGi / eMergence“ geteilt, 08.09.2026.
|
|
|
|
## Vollständiger Post-Text
|
|
|
|
> Independent benchmarking on Robocurve recently tested OpenAI's GPT-6 Astra against Claude Fable 5.1 controlling YAM robot arms. The shift in performance is massive:
|
|
>
|
|
> 🔹 Task Success (Block-into-bowl): Jumped from 40% (Fable 5.1) to 95% with Astra.
|
|
> 🔹 Token Efficiency: Astra used ~2.1k output tokens per run vs 12.9k for Fable 5.1 (a ~6x reduction in token overhead).
|
|
> 🔹 Speed & Cost: Execution time dropped from 6.8 minutes down to 2.5 minutes, nearly cutting operational costs in half ($0.94 vs $2.12/run).
|
|
|
|
## Einordnung der Quelle
|
|
|
|
Der Post bezeichnet Robocurve als unabhängigen Benchmarking-Kontext und enthält ein 14,3-sekündiges Video. Die Website [Robocurve](https://robocurve.org/) beschreibt die Organisation als Public Benefit Corporation und unabhängigen Evaluator für Physical AI; sie entwickelt offene Werkzeuge und Benchmarks. Der konkrete Astra/Fable-5.1-Datensatz, die YAM-Versuchsparameter und die Zahlen oben wurden in diesem Ingest nicht direkt auf der Robocurve-Website verifiziert.
|