Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 33,851 to 33,860 of 221,510 articles

TC-AE: Unlocking Token Capacity for Deep Compression Autoencoders

arXiv
We propose TC-AE, a ViT-based architecture for deep compression autoencoders. Existing methods commonly increase the channel number of latent representations to maintain reconstruction quality under high compression ratios. However, this strategy oft... read more 

Decisions and Deployment: The Five-Year SAHELI Project (2020-2025) on Restless Multi-Armed Bandits for Improving Maternal and Child Health

arXiv
Maternal and child health is a critical concern around the world. In many global health programs disseminating preventive care and health information, limited healthcare worker resources prevent continuous, personalised engagement with vulnerable ben... read more 

Data Warmup: Complexity-Aware Curricula for Efficient Diffusion Training

arXiv
A key inefficiency in diffusion training occurs when a randomly initialized network, lacking visual priors, encounters gradients from the full complexity spectrum--most of which it lacks the capacity to resolve. We propose Data Warmup, a curriculum s... read more 

Accelerating Training of Autoregressive Video Generation Models via Local Optimization with Representation Continuity

arXiv
Autoregressive models have shown superior performance and efficiency in image generation, but remain constrained by high computational costs and prolonged training times in video generation. In this study, we explore methods to accelerate training fo... read more 

GAN-based Domain Adaptation for Image-aware Layout Generation in Advertising Poster Design

arXiv
Layout plays a crucial role in graphic design and poster generation. Recently, the application of deep learning models for layout generation has gained significant attention. This paper focuses on using a GAN-based model conditioned on images to gene... read more 

FORGE:Fine-grained Multimodal Evaluation for Manufacturing Scenarios

arXiv
The manufacturing sector is increasingly adopting Multimodal Large Language Models (MLLMs) to transition from simple perception to autonomous execution, yet current evaluations fail to reflect the rigorous demands of real-world manufacturing environm... read more 

Multimodal Large Language Models for Multi-Subject In-Context Image Generation

arXiv
Recent advances in text-to-image (T2I) generation have enabled visually coherent image synthesis from descriptions, but generating images containing multiple given subjects remains challenging. As the number of reference identities increases, existin... read more 

An Analysis of Artificial Intelligence Adoption in NIH-Funded Research

arXiv
Understanding the landscape of artificial intelligence (AI) and machine learning (ML) adoption across the National Institutes of Health (NIH) portfolio is critical for research funding strategy, institutional planning, and health policy. The advent o... read more 

Personalizing Text-to-Image Generation to Individual Taste

arXiv
Modern text-to-image (T2I) models generate high-fidelity visuals but remain indifferent to individual user preferences. While existing reward models optimize for "average" human appeal, they fail to capture the inherent subjectivity of aesthetic judg... read more 

SMFD-UNet: Semantic Face Mask Is The Only Thing You Need To Deblur Faces

arXiv
For applications including facial identification, forensic analysis, photographic improvement, and medical imaging diagnostics, facial image deblurring is an essential chore in computer vision allowing the restoration of high-quality images from blur... read more