GPT-6 Astra appears to show a "step change" in spatial reasoning based on early benchmarks

The Decoder··作者 Matthias Bastian

资讯摘要

GPT-6 Astra appears to show a "step change" in spatial reasoning based on early benchmarks GPT-6 Astra appears to be a big leap forward for spatial reasoning. A new robotics benchmark called StationeryBench pits OpenAI's GPT-6 Astra against Ai2's MolmoAct2 across five desk-object tasks like uncapping a marker, pouring out paper clips, or passing a ruler between two robot arms. Both models controlled the same dual-arm YAM robots across 200 trials. Astra fully completed 7 out of 100 tasks; MolmoAct2 completed zero. Astra's median progress score hit 46 out of 100, MolmoAct2 managed 12.

GPT-6 Astra appears to show a "step change" in spatial reasoning based on early benchmarks

来源与参考

  1. 原始链接
  2. GPT-6 Astra appears to show a "step change" in spatial reasoning based on early benchmarks

收录于 2026-09-13