[AI]■ STORY TIMELINE
AI X-RAY READERS OVERCONFIDENT WHEN WRONG
AI chatbots analyzing X-rays frequently deliver incorrect diagnoses with unwarranted confidence, according to the RadLE 2.0 benchmark. The test reveals that many models fail to recognize the limits of their capabilities, a critical flaw for medical applications.
The Decoder+0m
The RadLE 2.0 benchmark tests whether AI models in radiology can tell when they should leave a diagnosis to a human. Man…