:
[AI]■ STORY TIMELINE

AI X-RAY READERS OVERCONFIDENT WHEN WRONG

AI chatbots analyzing X-rays frequently deliver incorrect diagnoses with unwarranted confidence, according to the RadLE 2.0 benchmark. The test reveals that many models fail to recognize the limits of their capabilities, a critical flaw for medical applications.

1 SOURCEFIRST SEEN JUL 19, 07:35 AM► READ THE ARTICLE
The Decoder+0m

The RadLE 2.0 benchmark tests whether AI models in radiology can tell when they should leave a diagnosis to a human. Man…