Surgery

Surveys

Latest AI and machine learning research in surveys for healthcare professionals.

5,349 articles
Stay Ahead - Weekly Surveys research updates
Subscribe
Browse Categories
Showing 2801-2820 of 5,349 articles

Hard to Halt: Automation Bias in Agent-Driven Sequencing Prior Authorization Workflows

Purpose: Prior authorization (PA) for exome or genome sequencing is a time-consuming process that impedes timely rare disease diagnosis. Large language model-based browser agents offer potential for automating these workflows, but their clinical reliability remain uncharacterized. Methods: We developed a sandbox compromising a simulated ES/GS PA submission payer portal and a synthetic EHR containi...

Artificial Intelligence-informed mobile behavioural interventions to support adolescents mental health in schools: protocol for a randomised controlled trial using the MindCraft app

Background: Children and young people (CYP) are particularly affected by mental health problems. Mobile apps provide a scalable and accessible approach to adolescent mental health support, and schools are well-positioned to address multiple risk factors and deliver large-scale interventions. By combining active (self-reported) and passive (sensor-derived) data, mobile apps can model mental states ...

ARTEMIS: Agent-guided Reliability-aware Temporal Mask Evolution for Imperfectly Supervised Video Polyp Segmentation

Imperfectly supervised video polyp segmentation (VPS) aims to learn dense, temporally consistent masks from inexpensive supervision, including weak an...

Jun 18 2026 2606.20161v1
BAFIS: Dataset + Framework to assess occupational Bias and Human Preference in modern Text-to-image Models

Generative artificial intelligence has the potential to improve productivity and transform the production of creative content. However, existing resea...

Jun 18 2026 2606.20241v1
Judging to Improve: A De-biased VLM-as-3D-Judge Protocol for Single-Image 3D Generation

A companion study established a de-biased, cross-model VLM-as-3D-judge that reliably ranks single-image-to-3D mesh quality where cheap geometry and CL...

Jun 18 2026 2606.20364v1
StylisticBias: A Few Human Visual Cues Drive Most Social Biases in MLLMs

Multimodal large language models (MLLMs) are increasingly deployed in personally and societally consequential settings, yet the visual cues that shape...

Jun 18 2026 2606.20527v1
The Unreliable Judges: Assessing Reproducibility and Self-Preference Bias of LLMs as Free-Text Evaluators

Large Language Models (LLMs) are transforming clinical practice and research, but their adoption requires rigorous evaluation. While human assessment ...

A Cross-Model VLM-Judge Protocol for Single-Image 3D Mesh Quality (and Why Cheap Proxies Fall Short)

Single-image-to-3D generators are improving quickly, but there is no agreed, human-free way to tell whether one generated mesh is better than another....

Jun 16 2026 2606.18451v1
The Illusion of Improvement: Reject Inference Strategies in Credit Scoring

Reject inference methods are widely used to mitigate survival bias in credit scoring, yet their effectiveness remains poorly understood. We systematic...

Jun 16 2026 2606.18479v1
A Transformer-derived transcriptomic score associates with ex-vivo drug response in AML

Background Drug-tolerant persister (DTP) cell states have been implicated in relapse across multiple cancers, including acute myeloid leukaemia (AML) ...

Toward Controllable Catalyst Inverse Design via Large-Scale Autoregressive Pretraining

Inverse design of heterogeneous catalysts remains challenging because catalyst surfaces exhibit substantial structural complexity with coupled surface...

Jun 16 2026 2606.17445v1
Domain-Validity-Gated Metamorphic Testing of Scientific ML Surrogates

Scientific machine-learning (SciML) surrogates approximate expensive simulations, but exact expected outputs for arbitrary inputs are unavailable (the...

Jun 16 2026 2606.17529v1
ARES: A Platform for Adaptive Role-Based Evaluation of Social Engineering Risks in Human--AI Games

This work introduces ARES, a platform and open pilot dataset for auditing adaptive social engineering risks in LLM-mediated social decision-making thr...

Jun 16 2026 2606.17793v1
AURA: Active-Response Attribution under Treatment Ambiguity in Bacterial Cytological Profiling

When a bacterial sample is exposed to several antibiotics, not every applied drug necessarily acts: if the organism is resistant to one of them, that ...

Jun 15 2026 2606.16477v1
MIRAGE: Auditing Anti-Muslim Bias in Frontier LLMs Across Reasoning, Agentic, and Time-Coupled Conditions

Five years after the discovery of persistent anti-Muslim bias in large language models, most evaluations remain confined to single-turn prompt complet...

Jun 15 2026 2606.16562v1
RaLMPH: Reliability-aware Learning for Multi-Pathologist Harmonization in Whole-Slide Image Classification

Multiple Instance Learning (MIL) is a standard paradigm for Whole-Slide Image (WSI) analysis and has achieved strong results in computational patholog...

Jun 14 2026 2606.15554v1
Positive Interpretation of Emotional Ambiguity Across Development: LC-dlPFC Circuitry

Emotionally ambiguous facial expressions provide a tractable model for studying how uncertain affective information is resolved into categorical judgm...

GermRL: Alleviating The Germline Bias In Autoregressive Antibody Language Models Through Reinforcement Learning

Antibodies are powerful therapeutics whose antigen specificity arises from sequence diversity shaped during development. Recently, language models tra...

Multimodal Graph Negative Learning

Multimodal attributed graphs (MAGs) integrate graph topology with heterogeneous modality attributes, such as text and images, thereby enabling richer ...

Jun 11 2026 2606.12863v1
Reinforcement Learning for Neural Model Editing

Editing pretrained neural networks requires specialized algorithms tailored to specific objectives. Designing such algorithms is often time-consuming ...

Jun 11 2026 2606.13461v1
Browse Categories