GPT-6 Astra appears to show a "step change" in spatial reasoning based on early benchmarks
The Decoder··作者 Matthias Bastian
资讯摘要
GPT-6 Astra appears to show a "step change" in spatial reasoning based on early benchmarks GPT-6 Astra appears to be a big leap forward for spatial reasoning. A new robotics benchmark called StationeryBench pits OpenAI's GPT-6 Astra against Ai2's MolmoAct2 across five desk-object tasks like uncapping a marker, pouring out paper clips, or passing a ruler between two robot arms. Both models controlled the same dual-arm YAM robots across 200 trials. Astra fully completed 7 out of 100 tasks; MolmoAct2 completed zero. Astra's median progress score hit 46 out of 100, MolmoAct2 managed 12.

来源与参考
收录于 2026-09-13