Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 28,611 to 28,620 of 219,260 articles

Trustworthy Endoscopic Super-Resolution

arXiv
Super-resolution (SR) models are attracting growing interest for enhancing minimally invasive surgery and diagnostic videos under hardware constraints. However, valid concerns remain regarding the introduction of hallucinated structures and amplified... read more 

CFSR: Geometry-Conditioned Shadow Removal via Physical Disentanglement

arXiv
Traditional shadow removal networks often treat image restoration as an unconstrained mapping, lacking the physical interpretability required to balance localized texture recovery with global illumination consistency. To address this, we propose CFSR... read more 

HABIT: Chrono-Synergia Robust Progressive Learning Framework for Composed Image Retrieval

arXiv
Composed Image Retrieval (CIR) is a flexible image retrieval paradigm that enables users to accurately locate the target image through a multimodal query composed of a reference image and modification text. Although this task has demonstrated promisi... read more 

INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval

arXiv
Composed Image Retrieval (CIR) is a challenging image retrieval paradigm that enables to retrieve target images based on multimodal queries consisting of reference images and modification texts. Although substantial progress has been made in recent y... read more 

Sonata: A Hybrid World Model for Inertial Kinematics under Clinical Data Scarcity

arXiv
We introduce Sonata, a compact latent world model for six-axis trunk IMU representation learning under clinical data scarcity. Clinical cohorts typically comprise tens to hundreds of patients, making web-scale masked-reconstruction objectives poorly ... read more 

Class-specific diffusion models improve military object detection in a low-data domain

arXiv
Diffusion-based image synthesis has emerged as a promising source of synthetic training data for AI-based object detection and classification. In this work, we investigate whether images generated with diffusion can improve military vehicle detection... read more 

Autonomous Unmanned Aircraft Systems for Enhanced Search and Rescue of Drowning Swimmers: Image-Based Localization and Mission Simulation

arXiv
Drowning is an omnipresent risk associated with any activity on or in the water, and rescuing a drowning person is particularly challenging because of the time pressure, making a short response time important. Further complicating water rescue are un... read more 

Culture-Aware Humorous Captioning: Multimodal Humor Generation across Cultural Contexts

arXiv
Recent multimodal large language models have shown promising ability in generating humorous captions for images, yet they still lack stable control over explicit cultural context, making it difficult to jointly maintain image relevance, contextual ap... read more 

Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?

arXiv
Recent advancements in self-supervised learning have led to powerful surgical vision encoders capable of spatiotemporal understanding. However, extending these visual foundations to multi-modal reasoning tasks is severely bottlenecked by the prohibit... read more 

Soft Label Pruning and Quantization for Large-Scale Dataset Distillation

arXiv
Large-scale dataset distillation requires storing auxiliary soft labels that can be 30-40x larger on ImageNet-1K and 200x larger on ImageNet-21K than the condensed images, undermining the goal of dataset compression. We identify two fundamental issue... read more