Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 58,991 to 59,000 of 227,876 articles

DSFedMed: Dual-Scale Federated Medical Image Segmentation via Mutual Distillation Between Foundation and Lightweight Models

arXiv
Foundation Models (FMs) have demonstrated strong generalization across diverse vision tasks. However, their deployment in federated settings is hindered by high computational demands, substantial communication overhead, and significant inference cost... read more 

Masked Modeling for Human Motion Recovery Under Occlusions

arXiv
Human motion reconstruction from monocular videos is a fundamental challenge in computer vision, with broad applications in AR/VR, robotics, and digital content creation, but remains challenging under frequent occlusions in real-world settings.Existi... read more 

Masked Modeling for Human Motion Recovery Under Occlusions

arXiv
Human motion reconstruction from monocular videos is a fundamental challenge in computer vision, with broad applications in AR/VR, robotics, and digital content creation, but remains challenging under frequent occlusions in real-world settings. Exist... read more 

Clustering-Guided Spatial-Spectral Mamba for Hyperspectral Image Classification

arXiv
Although Mamba models greatly improve Hyperspectral Image (HSI) classification, they have critical challenges in terms defining efficient and adaptive token sequences for improve performance. This paper therefore presents CSSMamba (Clustering-guided ... read more 

Rethinking Composed Image Retrieval Evaluation: A Fine-Grained Benchmark from Image Editing

arXiv
Composed Image Retrieval (CIR) is a pivotal and complex task in multimodal understanding. Current CIR benchmarks typically feature limited query categories and fail to capture the diverse requirements of real-world scenarios. To bridge this evaluatio... read more 

Learning to Watermark in the Latent Space of Generative Models

arXiv
Existing approaches for watermarking AI-generated images often rely on post-hoc methods applied in pixel space, introducing computational overhead and potential visual artifacts. In this work, we explore latent space watermarking and introduce DistSe... read more 

360Anything: Geometry-Free Lifting of Images and Videos to 360°

arXiv
Lifting perspective images and videos to 360° panoramas enables immersive 3D world generation. Existing approaches often rely on explicit geometric alignment between the perspective and the equirectangular projection (ERP) space. Yet, this requires k... read more 

Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders

arXiv
Representation Autoencoders (RAEs) have shown distinct advantages in diffusion modeling on ImageNet by training in high-dimensional semantic latent spaces. In this work, we investigate whether this framework can scale to large-scale, freeform text-to... read more 

GR3EN: Generative Relighting for 3D Environments

arXiv
We present a method for relighting 3D reconstructions of large room-scale environments. Existing solutions for 3D scene relighting often require solving under-determined or ill-conditioned inverse rendering problems, and are as such unable to produce... read more 

FeTTL: Federated Template and Task Learning for Multi-Institutional Medical Imaging

arXiv
Federated learning enables collaborative model training across geographically distributed medical centers while preserving data privacy. However, domain shifts and heterogeneity in data often lead to a degradation in model performance. Medical imagin... read more