AIThe Decoder2h ago

New benchmark confirms AI models still perform poorly at visual

New benchmark confirms AI models still perform poorly at visual perception

New benchmark confirms AI models still perform poorly at visual

Moonshot AI's PerceptionBench tests how well multimodal AI models can actually "see," separate from logical reasoning. No frontier model reaches 60 percent accuracy, and GPT-5.6 Sol leads by a narrow margin. Many supposed reasoning errors actually happen as early as the…

Read full article

Source: The Decoder · Opens in new tab