AIMC Journal:
medRxiv

Showing 241 to 250 of 4073 articles

An Interpretable Cost-Aware Framework for Mitigating Bias in Skin Lesion Classification Across Diverse Skin Tones

medRxiv
Despite advances in dermatological AI, skin lesion predictions continue to exhibit significant bias, consistently exhibiting underperformance in brown and darker tones. This disparity stems largely from the lack of representation in commonly used dat...

Certified large language model-based diagnostic decision support in rheumatology: the ALLIANCE multicentre randomised controlled trial

medRxiv
Objectives To evaluate whether access to a certified large language model (LLM)-based clinical decision support system improves physician diagnostic performance in rheumatology compared with conventional diagnostic resources alone. Methods In this mu...

ECG-based longitudinal risk prediction across diseases and organ systems

medRxiv
Artificial intelligence applied to routine electrocardiograms (ECGs) has largely focused on detecting existing disease or predicting individual cardiovascular outcomes. Whether ECGs can support prediction of multiple future diseases across organ syst...

Are Frontier Large Language Models Safer Than Government-Backed Symptom Checkers for Clinical Self-Triage? A Standardised Vignette Evaluation

medRxiv
Background: Consumer use of AI chatbots for health advice is rising, yet triage safety relative to established services remains unclear. Australia's Healthdirect, a government-backed symptom checker with 2.4 million uses in FY2024-25, remains unevalu...

Benchmarking ten frontier large language models on 1,477 board style multiple choice questions in hematology

medRxiv
Large Language Models (LLMs) are increasingly used by clinicians and patients for medical queries, yet their accuracy and safety at the specialist level in hematology remain insufficiently characterised. We benchmarked ten frontier proprietary and op...

When medical credentials conflict with stated accuracy: A factorial study of source credibility and answer revision in medical LLM interactions

medRxiv
Large language models perform well on medical examinations, but users routinely challenge their answers and invoke professional roles, and it is unclear what a system does when a medical credential and a stated task-specific accuracy point in opposit...

PCGS: biomarker and risk group identification for Pediatric Cancers via explainable Graph neural networks with Shapley values

medRxiv
Improvements in data availability, sharing, and integration, together with the development of explainable artificial intelligence (XAI) techniques, are advancing precision medicine for pediatric cancer by facilitating diagnosis, biomarker discovery, ...

A Pragmatic Randomized Trial of an EHR-Integrated Generative AI Chart Summarization Tool for Ambulatory Clinicians

medRxiv
BACKGROUND Generative AI (genAI) chart summarization tools embedded in electronic health records (EHRs) are being rapidly deployed across U.S. health systems. Although these tools represent a promising solution to alleviate cognitive burdens, their e...

Machine learning analysis of Autism phenotype data supports a four-dimensional continuum with three overlapping subtypes

medRxiv
Autism Spectrum Disorder (ASD) is a heterogeneous neurodevelopmental condition defined by differences in social communication and restricted, repetitive behaviours. As diagnostic criteria have broadened, ASD is now recognised across a wider range of ...

Image transmission through a multimode fibre in reflection mode with physics-guided deep learning towards ultrathin endoscopy

medRxiv
Ultrathin endoscopy is highly attractive for real-time tissue imaging in narrow and hard-to-reach regions of the body. A single multimode fibre (MMF) is an attractive probe because of its small diameter, flexibility, and diffraction-limited spatial r...