Lotu Radar About · RSS

New benchmark confirms AI models still perform poorly at visual perception

The Decoder AI Score 9/10

Summary

Moonshot AI's PerceptionBench tests how well multimodal AI models can actually "see," separate from logical reasoning. No frontier model reaches 60 percent accuracy, and GPT-5.6 Sol leads by a narrow margin. Many supposed reasoning errors actually happen as early as the image-reading stage. The article New benchmark confirms AI models still perform poorly at visual perception appeared first on The Decoder .

AIResearch

Lotu Radar provides attributed news summaries and links to the original publisher. Full reporting and copyright remain with the source.