Moonshot AI has introduced PerceptionBench, a new benchmark designed to assess the visual perception capabilities of multimodal AI models. This data release aims to provide clearer insights into how well these advanced systems interpret and reason about visual information.

According to The Decoder, no leading frontier model currently achieves 60 percent accuracy on PerceptionBench, highlighting the challenge of visual perception tasks. Among competitors, GPT-5.6 Sol holds a slight lead, though many errors attributed to reasoning are actually linked to difficulties in the image-reading stage.

For Japanese markets, where AI-driven technologies are increasingly integrated into sectors like finance and manufacturing, benchmarks such as PerceptionBench offer valuable metrics to guide investment and development in AI models with enhanced multimodal understanding.