State Required CME

Cultural Competence

Latest AI and machine learning research in cultural competence for healthcare professionals.

4,455 articles
Stay Ahead - Weekly Cultural Competence research updates
Subscribe
Browse Categories
Showing 1881-1900 of 4,455 articles

DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models

Large Multimodal Models (LMMs) have emerged as powerful models capable of understanding various data modalities, including text, images, and videos. LMMs encode both text and visual data into tokens that are then combined and processed by an integrated Large Language Model (LLM). Including visual tokens substantially increases the total token count, often by thousands. The increased input length...

Group Relative Policy Optimization for Image Captioning

Image captioning tasks usually use two-stage training to complete model optimization. The first stage uses cross-entropy as the loss function for optimization, and the second stage uses self-critical sequence training (SCST) for reinforcement learning optimization. However, the SCST algorithm has certain defects. SCST relies only on a single greedy decoding result as a baseline. If the model its...

Parameter Expanded Stochastic Gradient Markov Chain Monte Carlo

Bayesian Neural Networks (BNNs) provide a promising framework for modeling predictive uncertainty and enhancing out-of-distribution robustness (OOD)...

Image-based food groups and portion prediction by using deep learning.

Chronic diseases such as obesity and hypertension due to malnutrition can be prevented by following the appropriate diet, correct diet intake with cor...

Mar 1 2025 40052549
Transparency and Representation in Clinical Research Utilizing Artificial Intelligence in Oncology: A Scoping Review.

INTRODUCTION: Artificial intelligence (AI) has significant potential to improve health outcomes in oncology. However, as AI utility increases, it is i...

Mar 1 2025 40059400
Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generation

Autoregressive (AR) modeling, known for its next-token prediction paradigm, underpins state-of-the-art language and visual generative models. Tradit...

Learning to Generalize without Bias for Open-Vocabulary Action Recognition

Leveraging the effective visual-text alignment and static generalizability from CLIP, recent video learners adopt CLIP initialization with further r...

UIFace: Unleashing Inherent Model Capabilities to Enhance Intra-Class Diversity in Synthetic Face Recognition

Face recognition (FR) stands as one of the most crucial applications in computer vision. The accuracy of FR models has significantly improved in rec...

The erasure of intensive livestock farming in text-to-image generative AI

Generative AI (e.g., ChatGPT) is increasingly integrated into people's daily lives. While it is known that AI perpetuates biases against marginalize...

Revealing Treatment Non-Adherence Bias in Clinical Machine Learning Using Large Language Models

Machine learning systems trained on electronic health records (EHRs) increasingly guide treatment decisions, but their reliability depends on the cr...

Effect of Gender Fair Job Description on Generative AI Images

STEM fields are traditionally male-dominated, with gender biases shaping perceptions of job accessibility. This study analyzed gender representation...

FairGen: Controlling Sensitive Attributes for Fair Generations in Diffusion Models via Adaptive Latent Guidance

Text-to-image diffusion models often exhibit biases toward specific demographic groups, such as generating more males than females when prompted to ...

Defining bias in AI-systems: Biased models are fair models

The debate around bias in AI systems is central to discussions on algorithmic fairness. However, the term bias often lacks a clear definition, despi...

Assessing Large Language Models in Agentic Multilingual National Bias

Large Language Models have garnered significant attention for their capabilities in multilingual natural language processing, while studies on risks...

CLEP-GAN: An Innovative Approach to Subject-Independent ECG Reconstruction from PPG Signals

This study addresses the challenge of reconstructing unseen ECG signals from PPG signals, a critical task for non-invasive cardiac monitoring. While...

HybridLinker: Topology-Guided Posterior Sampling for Enhanced Diversity and Validity in 3D Molecular Linker Generation

Linker generation is critical in drug discovery applications such as lead optimization and PROTAC design, where molecular fragments are assembled in...

Measuring Data Diversity for Instruction Tuning: A Systematic Analysis and A Reliable Metric

Data diversity is crucial for the instruction tuning of large language models. Existing studies have explored various diversity-aware data selection...

Fair Foundation Models for Medical Image Analysis: Challenges and Perspectives

Ensuring equitable Artificial Intelligence (AI) in healthcare demands systems that make unbiased decisions across all demographic groups, bridging t...

FedBM: Stealing Knowledge from Pre-trained Language Models for Heterogeneous Federated Learning

Federated learning (FL) has shown great potential in medical image computing since it provides a decentralized learning paradigm that allows multipl...

Visual Reasoning Evaluation of Grok, Deepseek Janus, Gemini, Qwen, Mistral, and ChatGPT

Traditional evaluations of multimodal large language models (LLMs) have been limited by their focus on single-image reasoning, failing to assess cru...

Browse Categories