Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 33,701 to 33,710 of 221,422 articles

GenLCA: 3D Diffusion for Full-Body Avatars from In-the-Wild Videos

arXiv
We present GenLCA, a diffusion-based generative model for generating and editing photorealistic full-body avatars from text and image inputs. The generated avatars are faithful to the inputs, while supporting high-fidelity facial and full-body animat... read more 

GenLCA: 3D Diffusion for Full-Body Avatars from In-the-Wild Videos

arXiv
We present GenLCA, a diffusion-based generative model for generating and editing photorealistic full-body avatars from text and image inputs. The generated avatars are faithful to the inputs, while supporting high-fidelity facial and full-body animat... read more 

A Systematic Study of Retrieval Pipeline Design for Retrieval-Augmented Medical Question Answering

arXiv
Large language models (LLMs) have demonstrated strong capabilities in medical question answering; however, purely parametric models often suffer from knowledge gaps and limited factual grounding. Retrieval-augmented generation (RAG) addresses this li... read more 

Are Face Embeddings Compatible Across Deep Neural Network Models?

arXiv
Automated face recognition has made rapid strides over the past decade due to the unprecedented rise of deep neural network (DNN) models that can be trained for domain-specific tasks. At the same time, foundation models that are pretrained on broad v... read more 

Region-Graph Optimal Transport Routing for Mixture-of-Experts Whole-Slide Image Classification

arXiv
Multiple Instance Learning (MIL) is the dominant framework for gigapixel whole-slide image (WSI) classification in computational pathology. However, current MIL aggregators route all instances through a shared pathway, constraining their capacity to ... read more 

Distilling Photon-Counting CT into Routine Chest CT through Clinically Validated Degradation Modeling

arXiv
Photon-counting CT (PCCT) provides superior image quality with higher spatial resolution and lower noise compared to conventional energy-integrating CT (EICT), but its limited clinical availability restricts large-scale research and clinical deployme... read more 

Appear2Meaning: A Cross-Cultural Benchmark for Structured Cultural Metadata Inference from Images

arXiv
Recent advances in vision-language models (VLMs) have improved image captioning for cultural heritage. However, inferring structured cultural metadata (e.g., creator, origin, period) from visual input remains underexplored. We introduce a multi-categ... read more 

TC-AE: Unlocking Token Capacity for Deep Compression Autoencoders

arXiv
We propose TC-AE, a ViT-based architecture for deep compression autoencoders. Existing methods commonly increase the channel number of latent representations to maintain reconstruction quality under high compression ratios. However, this strategy oft... read more 

Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health

arXiv
Maternal and child health is a critical concern around the world. In many global health programs disseminating preventive care and health information, limited healthcare worker resources prevent continuous, personalised engagement with vulnerable ben... read more 

Data Warmup: Complexity-Aware Curricula for Efficient Diffusion Training

arXiv
A key inefficiency in diffusion training occurs when a randomly initialized network, lacking visual priors, encounters gradients from the full complexity spectrum--most of which it lacks the capacity to resolve. We propose Data Warmup, a curriculum s... read more