Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 54,431 to 54,440 of 226,183 articles

RISE: Interactive Visual Diagnosis of Fairness in Machine Learning Models

arXiv
Evaluating fairness under domain shift is challenging because scalar metrics often obscure exactly where and how disparities arise. We introduce \textit{RISE} (Residual Inspection through Sorted Evaluation), an interactive visualization tool that con... read more 

Explicit Uncertainty Modeling for Active CLIP Adaptation with Dual Prompt Tuning

arXiv
Pre-trained vision-language models such as CLIP exhibit strong transferability, yet adapting them to downstream image classification tasks under limited annotation budgets remains challenging. In active learning settings, the model must select the mo... read more 

Finding NeMO: A Geometry-Aware Representation of Template Views for Few-Shot Perception

arXiv
We present Neural Memory Object (NeMO), a novel object-centric representation that can be used to detect, segment and estimate the 6DoF pose of objects unseen during training using RGB images. Our method consists of an encoder that requires only a fe... read more 

VecSet-Edit: Unleashing Pre-trained LRM for Mesh Editing from Single Image

arXiv
3D editing has emerged as a critical research area to provide users with flexible control over 3D assets. While current editing approaches predominantly focus on 3D Gaussian Splatting or multi-view images, the direct editing of 3D meshes remains unde... read more 

When and Where to Attack? Stage-wise Attention-Guided Adversarial Attack on Large Vision Language Models

arXiv
Adversarial attacks against Large Vision-Language Models (LVLMs) are crucial for exposing safety vulnerabilities in modern multimodal systems. Recent attacks based on input transformations, such as random cropping, suggest that spatially localized pe... read more 

SparVAR: Exploring Sparsity in Visual AutoRegressive Modeling for Training-Free Acceleration

arXiv
Visual AutoRegressive (VAR) modeling has garnered significant attention for its innovative next-scale prediction paradigm. However, mainstream VAR paradigms attend to all tokens across historical scales at each autoregressive step. As the next scale ... read more 

Reducing the labeling burden in time-series mapping using Common Ground: a semi-automated approach to tracking changes in land cover and species over time

arXiv
Reliable classification of Earth Observation data depends on consistent, up-to-date reference labels. However, collecting new labelled data at each time step remains expensive and logistically difficult, especially in dynamic or remote ecological sys... read more 

Enabling Real-Time Colonoscopic Polyp Segmentation on Commodity CPUs via Ultra-Lightweight Architecture

arXiv
Early detection of colorectal cancer hinges on real-time, accurate polyp identification and resection. Yet current high-precision segmentation models rely on GPUs, making them impractical to deploy in primary hospitals, mobile endoscopy units, or cap... read more 

Quantile Transfer for Reliable Operating Point Selection in Visual Place Recognition

arXiv
Visual Place Recognition (VPR) is a key component for localisation in GNSS-denied environments, but its performance critically depends on selecting an image matching threshold (operating point) that balances precision and recall. Thresholds are typic... read more 

Interactive Spatial-Frequency Fusion Mamba for Multi-Modal Image Fusion

arXiv
Multi-Modal Image Fusion (MMIF) aims to combine images from different modalities to produce fused images, retaining texture details and preserving significant information. Recently, some MMIF methods incorporate frequency domain information to enhanc... read more