Multimodal AI Models Struggle to Surpass 50% Accuracy in Visual Entity Recognition
A new benchmark reveals that even the top multimodal AI models fail to reliably identify specific visual entities, highlighting challenges in AI’s understanding of detailed visual content.
