Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 37,661 to 37,670 of 223,469 articles

SPROUT: A Scalable Diffusion Foundation Model for Agricultural Vision

arXiv
Vision Foundation Models (VFM) pre-trained on large-scale unlabeled data have achieved remarkable success on general computer vision tasks, yet typically suffer from significant domain gaps when applied to agriculture. In this context, we introduce $... read more 

Hidden Ads: Behavior Triggered Semantic Backdoors for Advertisement Injection in Vision Language Models

arXiv
Vision-Language Models (VLMs) are increasingly deployed in consumer applications where users seek recommendations about products, dining, and services. We introduce Hidden Ads, a new class of backdoor attacks that exploit this recommendation-seeking ... read more 

MV-RoMa: From Pairwise Matching into Multi-View Track Reconstruction

arXiv
Establishing consistent correspondences across images is essential for 3D vision tasks such as structure-from-motion (SfM), yet most existing matchers operate in a pairwise manner, often producing fragmented and geometrically inconsistent tracks when... read more 

BLOSSOM: Block-wise Federated Learning Over Shared and Sparse Observed Modalities

arXiv
Multimodal federated learning (FL) is essential for real-world applications such as autonomous systems and healthcare, where data is distributed across heterogeneous clients with varying and often missing modalities. However, most existing FL approac... read more 

PANDORA: Pixel-wise Attention Dissolution and Latent Guidance for Zero-Shot Object Removal

arXiv
Removing objects from natural images is challenging due to difficulty of synthesizing semantically coherent content while preserving background integrity. Existing methods often rely on fine-tuning, prompt engineering, or inference-time optimization,... read more 

Structured Observation Language for Efficient and Generalizable Vision-Language Navigation

arXiv
Vision-Language Navigation (VLN) requires an embodied agent to navigate complex environments by following natural language instructions, which typically demands tight fusion of visual and language modalities. Existing VLN methods often convert raw im... read more 

A Robust Low-Rank Prior Model for Structured Cartoon-Texture Image Decomposition with Heavy-Tailed Noise

arXiv
Cartoon-texture image decomposition is a fundamental yet challenging problem in image processing. A significant hurdle in achieving accurate decomposition is the pervasive presence of noise in the observed images, which severely impedes robust result... read more 

An Energy-Efficient Spiking Neural Network Architecture for Predictive Insulin Delivery

arXiv
Diabetes mellitus affects over 537 million adults worldwide. Insulin-dependent patients require continuous glucose monitoring and precise dose calculation while operating under strict power budgets on wearable devices. This paper presents PDDS - an i... read more 

You Only Erase Once: Erasing Anything without Bringing Unexpected Content

arXiv
We present YOEO, an approach for object erasure. Unlike recent diffusion-based methods which struggle to erase target objects without generating unexpected content within the masked regions due to lack of sufficient paired training data and explicit ... read more 

Clore: Interactive Pathology Image Segmentation with Click-based Local Refinement

arXiv
Recent advancements in deep learning-based interactive segmentation methods have significantly improved pathology image segmentation. Most existing approaches utilize user-provided positive and negative clicks to guide the segmentation process. Howev... read more