knowledge-base/raw/xpost/2026-09-08_promiscfx-robocurve-gpt6-astra-yam.md

1.6 KiB

type source_url retrieved author is_thread quote_count tags
xpost https://x.com/promiscfx/status/2097080991475257839 2026-09-08 @promiscfx / Cognitive f(x) false 1
robotics
physical-ai
robocurve
gpt-6-astra
claude-fable-5.1
yam
benchmarking

Cognitive f(x): GPT-6 Astra vs. Claude Fable 5.1 auf YAM-Robotarmen

Autor: @promiscfx / Cognitive f(x) Quelle: https://x.com/promiscfx/status/2097080991475257839 Datum des Posts: 07.09.2026, 21:54 UTC Kontext: Von Pit Weber in OME Topic „aGi / eMergence“ geteilt, 08.09.2026.

Vollständiger Post-Text

Independent benchmarking on Robocurve recently tested OpenAI's GPT-6 Astra against Claude Fable 5.1 controlling YAM robot arms. The shift in performance is massive:

🔹 Task Success (Block-into-bowl): Jumped from 40% (Fable 5.1) to 95% with Astra. 🔹 Token Efficiency: Astra used ~2.1k output tokens per run vs 12.9k for Fable 5.1 (a ~6x reduction in token overhead). 🔹 Speed & Cost: Execution time dropped from 6.8 minutes down to 2.5 minutes, nearly cutting operational costs in half ($0.94 vs $2.12/run).

Einordnung der Quelle

Der Post bezeichnet Robocurve als unabhängigen Benchmarking-Kontext und enthält ein 14,3-sekündiges Video. Die Website Robocurve beschreibt die Organisation als Public Benefit Corporation und unabhängigen Evaluator für Physical AI; sie entwickelt offene Werkzeuge und Benchmarks. Der konkrete Astra/Fable-5.1-Datensatz, die YAM-Versuchsparameter und die Zahlen oben wurden in diesem Ingest nicht direkt auf der Robocurve-Website verifiziert.