Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 53,621 to 53,630 of 225,548 articles

VecSet-Edit: Unleashing Pre-trained LRM for Mesh Editing from Single Image

arXiv
3D editing has emerged as a critical research area to provide users with flexible control over 3D assets. While current editing approaches predominantly focus on 3D Gaussian Splatting or multi-view images, the direct editing of 3D meshes remains unde... read more 

Finding NeMO: A Geometry-Aware Representation of Template Views for Few-Shot Perception

arXiv
We present Neural Memory Object (NeMO), a novel object-centric representation that can be used to detect, segment and estimate the 6DoF pose of objects unseen during training using RGB images. Our method consists of an encoder that requires only a fe... read more 

Explicit Uncertainty Modeling for Active CLIP Adaptation with Dual Prompt Tuning

arXiv
Pre-trained vision-language models such as CLIP exhibit strong transferability, yet adapting them to downstream image classification tasks under limited annotation budgets remains challenging. In active learning settings, the model must select the mo... read more 

RISE: Interactive Visual Diagnosis of Fairness in Machine Learning Models

arXiv
Evaluating fairness under domain shift is challenging because scalar metrics often obscure exactly where and how disparities arise. We introduce \textit{RISE} (Residual Inspection through Sorted Evaluation), an interactive visualization tool that con... read more 

GeneralVLA: Generalizable Vision-Language-Action Models with Knowledge-Guided Trajectory Planning

arXiv
Large foundation models have shown strong open-world generalization to complex problems in vision and language, but similar levels of generalization have yet to be achieved in robotics. One fundamental challenge is that the models exhibit limited zer... read more 

Beyond Static Cropping: Layer-Adaptive Visual Localization and Decoding Enhancement

arXiv
Large Vision-Language Models (LVLMs) have advanced rapidly by aligning visual patches with the text embedding space, but a fixed visual-token budget forces images to be resized to a uniform pretraining resolution, often erasing fine-grained details a... read more 

SkeletonGaussian: Editable 4D Generation through Gaussian Skeletonization

arXiv
4D generation has made remarkable progress in synthesizing dynamic 3D objects from input text, images, or videos. However, existing methods often represent motion as an implicit deformation field, which limits direct control and editability. To addre... read more 

Aortic Valve Disease Detection from PPG via Physiology-Informed Self-Supervised Learning

arXiv
Traditional diagnosis of aortic valve disease relies on echocardiography, but its cost and required expertise limit its use in large-scale early screening. Photoplethysmography (PPG) has emerged as a promising screening modality due to its widespread... read more 

ACIL: Active Class Incremental Learning for Image Classification

arXiv
Continual learning (or class incremental learning) is a realistic learning scenario for computer vision systems, where deep neural networks are trained on episodic data, and the data from previous episodes are generally inaccessible to the model. Exi... read more 

An Improved Boosted DC Algorithm for Nonsmooth Functions with Applications in Image Recovery

arXiv
We propose a new approach to perform the boosted difference of convex functions algorithm (BDCA) on non-smooth and non-convex problems involving the difference of convex (DC) functions. The recently proposed BDCA uses an extrapolation step from the p... read more