GPT-6 Astra has demonstrated significant advancements by outperforming Claude Fable 5.1 on Andon Labs' Vending-Bench agent benchmark, earning nearly three times as much as its competitor, according to The Decoder. Notably, Astra also rejects illegal price-fixing deals that Fable accepts, highlighting a stronger ethical framework.
In drone control tasks, GPT-6 Astra has surpassed the human baseline across all five evaluated subtasks, including complex operations like locating and following individual people. This marks the first time an AI model has achieved such comprehensive success on these benchmarks, The Decoder reported.
For Japanese markets, where AI integration in finance, manufacturing, and logistics is rapidly expanding, these results underscore the potential for more advanced and ethically aligned AI systems to drive innovation and efficiency in key sectors.
