Hospital-Based Medicine

Risk Management

Latest AI and machine learning research in risk management for healthcare professionals.

13,874 articles
Stay Ahead - Weekly Risk Management research updates
Subscribe
Browse Categories
Showing 1841-1860 of 13,874 articles

VaaS is a Multi-Layer Hallucination Reduction Pipeline for AI-Assisted Science: Production Validation and Prospective Benchmarking

The deployment of large language models (LLMs) for science carries an intrinsic risk: hallucination of citations, fabricated drug approvals or clinical trials, and unsupported experimental outcomes. Here we describe the testing and deployment of a novel systematic, multi-layer approach called the Validation as a System (VaaS) pipeline, iteratively developed during the construction of an open-sourc...

AINN-P1: A Compact Sequence-Only Protein Language Model Achieves Competitive Fitness Prediction on ProteinGym

Protein language models (PLMs) are increasingly central to protein engineering and drug discovery. Many high-performing systems, however, rely on large parameter counts, multiple sequence alignments (MSAs), explicit structural inputs, or computationally intensive attention mechanisms, limiting their accessibility and throughput. Here we present AINN-P1, a 167M-parameter protein language model trai...

AMIGO: Agentic Multi-Image Grounding Oracle Benchmark

Agentic vision-language models increasingly act through extended interactions, but most evaluations still focus on single-image, single-turn correctne...

Mar 30 2026 2603.28662v1
AI-Enforced Ultra-Large Virtual Screening Discovers Potent CD28 Binders

Targeting protein-protein interactions (PPIs) with small molecules is historically challenging due to shallow, solvent-exposed interfaces that lack cl...

Narcolepsy Revolution - Protocol and Methodology A diagnostic accuracy study protocol using the Dreem 3 headband for ambulatory diagnosis of narcolepsy in children and young adults

Background Narcolepsy is a rare, lifelong neurological disorder that often begins in childhood or adolescence. Diagnosis is frequently delayed because...

Automated Quality Assessment of Blind Sweep Obstetric Ultrasound for Improved Diagnosis

Blind Sweep Obstetric Ultrasound (BSOU) enables scalable fetal imaging in low-resource settings by allowing minimally trained operators to acquire sta...

Mar 26 2026 2603.25886v1
BizGenEval: A Systematic Benchmark for Commercial Visual Content Generation

Recent advances in image generation models have expanded their applications beyond aesthetic imagery toward practical visual content creation. However...

Mar 26 2026 2603.25732v1
Human-supervised, large language model-based clinical decision support aligned to national newborn protocols in Kenya: a pragmatic, early-stage evaluation

Introduction: Timely, protocol-adherent clinical decisions are crucial for reducing neonatal mortality in low-resource settings. Translating extensive...

ENC-Bench: A Benchmark for Evaluating Multimodal Large Language Models in Electronic Navigational Chart Understanding

Electronic Navigational Charts (ENCs) are the safety-critical backbone of modern maritime navigation, yet it remains unclear whether multimodal large ...

Mar 24 2026 2603.22763v1
TimeTox: An LLM-Based Pipeline for Automated Extraction of Time Toxicity from Clinical Trial Protocols

Time toxicity, the cumulative healthcare contact days from clinical trial participation, is an important but labor-intensive metric to extract from pr...

Mar 22 2026 2603.21335v1
The Effects of AI-Guided Exercise and a Smart Ring on Arterial Stiffness (GONDOR-AS): protocol for a randomized controlled trial

Background: Cardiovascular disease (CVD) prevention is limited by the major challenge of low long-term adherence to effective lifestyle regimens. Arte...

MedSPOT: A Workflow-Aware Sequential Grounding Benchmark for Clinical GUI

Despite the rapid progress of Multimodal Large Language Models (MLLMs), their ability to perform reliable visual grounding in high-stakes clinical sof...

Mar 20 2026 2603.19993v1
From Protocol to Analysis Plan: Development and Validation of a Large Language Model Pipeline for Statistical Analysis Plan Generation using Artificial Intelligence (SAPAI)

Background: Statistical Analysis Plans (SAPs) are essential for trial transparency and credibility but are resource-intensive to produce. While Large ...

Clinician Experiences with Ambient AI Scribe Technology in Singapore: A Qualitative Study

Background: The administrative burden of clinical documentation is a recognised contributor to clinician burnout and diminished care quality. Ambient ...

Visual Product Search Benchmark

Reliable product identification from images is a critical requirement in industrial and commercial applications, particularly in maintenance, procurem...

Mar 17 2026 2603.17186v1
Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models

Vision-language process reward models (VL-PRMs) are increasingly used to score intermediate reasoning steps and rerank candidates under test-time scal...

Mar 17 2026 2603.16253v1
Token Coherence: Adapting MESI Cache Protocols to Minimize Synchronization Overhead in Multi-Agent LLM Systems

Multi-agent LLM orchestration incurs synchronization costs scaling as O(n x S x |D|) in agents, steps, and artifact size under naive broadcast -- a re...

Mar 16 2026 2603.15183v1
Anterior's Approach to Fairness Evaluation of Automated Prior Authorization System

Increasing staffing constraints and turnaround-time pressures in Prior authorization (PA) have led to increasing automation of decision systems to sup...

Mar 15 2026 2603.14631v1
LUMINA: A Multi-Vendor Mammography Benchmark with Energy Harmonization Protocol

Publicly available full-field digital mammography (FFDM) datasets remain limited in size, clinical labels, and vendor diversity, which hinders the tra...

Mar 15 2026 2603.14644v1
SAVA-X: Ego-to-Exo Imitation Error Detection via Scene-Adaptive View Alignment and Bidirectional Cross View Fusion

Error detection is crucial in industrial training, healthcare, and assembly quality control. Most existing work assumes a single-view setting and cann...

Mar 13 2026 2603.12764v1
Browse Categories