AI Models Fail Miserably at This One Easy Task: Telling Time
View original at spectrum.ieee.orgAI Models Fail Miserably at This One Easy Task: Telling Time <img src="https://spectrum.ieee.org/media-library/a-digitally-structured-tree-with-a-melting-clock-hanging-off-one-of-its-branches-the-concept-resembles-salvador-dali-s-persist.jpg?id=62053134&width=1200&height=800&coordinates=0%2C133%2C0%2C134" /><br /><br /…
What we drew from this source
The claims Via News extracted from this document. We point to the source; we don't replace it.
If a MLLM struggles with one facet of image analysis, this can cause a cascading effect that impacts other aspects of its image analysis
80% confidenceWe cannot take model performance for granted and extensive training and testing with varied inputs is necessary to ensure models remain robust against diverse real-world scenarios
80% confidenceIf the MLLMs made an error in recognizing the clock hands, this in turn resulted in greater spatial errors
80% confidenceReading the time is not as simple a task as it may seem, since the model must identify the clock hands, determine their orientations, and combine these observations to infer the correct time
80% confidenceWhile such variations pose little difficulty for humans, models often fail at this task
80% confidence
Cited in these Via News reports
- AI Governance Body AAIF Adds 97 Members as Visual Reasoning Gaps Challenge Global Deployment →
- Big Tech AI Models Force Specialized Language Startups to Close as Investors Flee →
- Big Tech AI Releases Trigger Investor Flight from Small Language Model Startups Globally →
- DeepSeek's efficiency challenge to Big Tech AI sparks global debate as funding shifts threaten 55-country startup ecosystem →
- Meta's 200-Language AI Model Killed African Language Startups, Safety Researcher Says →
- Specialized AI Models Challenge Big Tech's Data-Intensive Development Doctrine →
