AIMC Journal:
medRxiv

Showing 61 to 70 of 4073 articles

Safety-Relevant Biomedical Machine Learning Should Adopt Stability-First Reporting

medRxiv
Biomedical Machine Learning (ML) is increasingly evaluated through Independent and Identically Distributed (IID) benchmarks, even when deployment involves irregular observation, missingness, distribution shift, and high failure costs. This position p...

AI-assisted nurse-led skin cancer screening in a teledermoscopy framework: a multi-site evaluation with one million lesions

medRxiv
Nurse-led skin cancer screening extends specialist reach but depends on the nurse selecting which lesions to forward for diagnosis, and performance varies with experience. We evaluated whether real-time decision support using artificial intelligence ...

HPV-Stratified Causal Digital Twins for Personalized Survival Benefit Estimation in Head and Neck Cancer

medRxiv
Purpose: Digital-twin frameworks can potentially offer a pathway toward individualized treatment but remain largely untested in radiation oncology. We present a causal digital-twin model for head-and-neck (H&N) cancer that estimates patient-specific ...

An automated, explainable, NCCT-based clinical decision-support system for spontaneous intracerebral hemorrhage.

medRxiv
Background: Spontaneous intracerebral hemorrhage (ICH) has high disability and mortality. Accurate early prognostication remains challenging in routine practice. Conventional CT assessment relies mainly on hematoma volume and rough anatomical locatio...

Quality, Readability, and Content Deficiencies of Three Large Language Models in Bee Sting Health Information: A Comparative Evaluation with Patient Perspectives

medRxiv
Objective To evaluate the information quality, reliability, readability, and content deficits of ChatGPT, DeepSeek, and Doubao in answering frequently asked questions about bee stings. Methods Twenty-five high-frequency patient questions were selecte...

Benchmarking modular skills and fixed orchestration across six large language models in a meta-analysis task

medRxiv
Background. Large language models (LLMs) are increasingly used to support systematic review and meta-analysis, but complex evidence flow, trial-family identification, study-level data extraction, and statistical execution remain vulnerable to omissio...

'I love my AI girlfriend': a content analysis of youth social AI risk in online communities

medRxiv
Youth are increasingly engaging with generative AI beyond its use as a productivity tool, with a growing share turning to AI for social interaction and emotional support. This leads to unknown risks for their well-being and development. Online commun...

Mixed-Effects Machine Learning improved prediction of Obstetric Anal Sphincter Injury: a nationwide multicenter cohort study.

medRxiv
Existing prediction models for obstetric anal sphincter injuries (OASIS) often rely on traditional regression, lack external validation, and, when fitted on multicenter data, fail to account for hospital-level clustering. We aimed to develop and temp...

Comparison of Facial Feature Grading by an AI-Based System Versus Human Expert Grading on Images Under Standardized Clinical and At-Home Settings

medRxiv
Digital photography remains the gold standard for dermatological assessment, but reliance on standardized stationary image-capture systems limits trial scalability, whereas smartphones combined with computer vision artificial intelligence (AI) could ...

Title and abstract screening for systematic reviews with Jev, a System One model: comparison with generative large language models

medRxiv
Large language models (LLMs) screen titles and abstracts without review-specific training, but generating screening decisions as text takes processing time and incurs API charges. We evaluated Jev, a non-generative model returning classification prob...