← Back to Model Beat
Products·Jul 19·all news from July 19, 2026

AI chatbots reading X-rays can be dangerously confident even when they're wrong

A new benchmark called RadLE 2.0 reveals that many medical AI models provide incorrect radiology findings with high levels of confidence. The study shows that these systems often fail to identify when they should defer to human expertise, highlighting a significant safety gap in current automated diagnostic technology. Human radiologists continue to outperform these models, suggesting that AI is not yet ready to function autonomously in clinical settings.

Covered by 1 source

Related stories

ProductsIntroducing the ChatGPT for small business programJul 21 · 9 sourcesProductsLaunching Health in ChatGPTJul 23 · 3 sourcesProductsEnvironment-free Synthetic Data Generation for API-Calling AgentsJul 21 · 2 sourcesProductsThe Download: AI hiring biases, and weather data sabotageJul 20 · 4 sources