Models Claim Images Without Seeing Them
The finding shows that widely used benchmarks miss this hallucination, raising global concerns over AI reliability, medical decision support, and the need for more robust evaluation standards across the industry.