AIThe Decoder2h ago
New benchmark confirms AI models still perform poorly at visual
New benchmark confirms AI models still perform poorly at visual perception

Moonshot AI's PerceptionBench tests how well multimodal AI models can actually "see," separate from logical reasoning. No frontier model reaches 60 percent accuracy, and GPT-5.6 Sol leads by a narrow margin. Many supposed reasoning errors actually happen as early as the…
Read full articleSource: The Decoder · Opens in new tab