Biomedical Machine Learning (ML) is increasingly evaluated through Independent and Identically Distributed (IID) benchmarks, even when deployment involves irregular observation, missingness, distribution shift, and high failure costs. This position p...
Nurse-led skin cancer screening extends specialist reach but depends on the nurse selecting which lesions to forward for diagnosis, and performance varies with experience. We evaluated whether real-time decision support using artificial intelligence ...
Purpose: Digital-twin frameworks can potentially offer a pathway toward individualized treatment but remain largely untested in radiation oncology. We present a causal digital-twin model for head-and-neck (H&N) cancer that estimates patient-specific ...
Background: Spontaneous intracerebral hemorrhage (ICH) has high disability and mortality. Accurate early prognostication remains challenging in routine practice. Conventional CT assessment relies mainly on hematoma volume and rough anatomical locatio...
Objective To evaluate the information quality, reliability, readability, and content deficits of ChatGPT, DeepSeek, and Doubao in answering frequently asked questions about bee stings. Methods Twenty-five high-frequency patient questions were selecte...
Background. Large language models (LLMs) are increasingly used to support systematic review and meta-analysis, but complex evidence flow, trial-family identification, study-level data extraction, and statistical execution remain vulnerable to omissio...
Youth are increasingly engaging with generative AI beyond its use as a productivity tool, with a growing share turning to AI for social interaction and emotional support. This leads to unknown risks for their well-being and development. Online commun...
Existing prediction models for obstetric anal sphincter injuries (OASIS) often rely on traditional regression, lack external validation, and, when fitted on multicenter data, fail to account for hospital-level clustering. We aimed to develop and temp...
Digital photography remains the gold standard for dermatological assessment, but reliance on standardized stationary image-capture systems limits trial scalability, whereas smartphones combined with computer vision artificial intelligence (AI) could ...
Large language models (LLMs) screen titles and abstracts without review-specific training, but generating screening decisions as text takes processing time and incurs API charges. We evaluated Jev, a non-generative model returning classification prob...
Join thousands of healthcare professionals staying informed about the latest AI breakthroughs in medicine. Get curated insights delivered to your inbox.