Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 33,831 to 33,840 of 221,510 articles

USCNet: Transformer-Based Multimodal Fusion with Segmentation Guidance for Urolithiasis Classification

arXiv
Kidney stone disease ranks among the most prevalent conditions in urology, and understanding the composition of these stones is essential for creating personalized treatment plans and preventing recurrence. Current methods for analyzing kidney stones... read more 

Learning to Search: A Decision-Based Agent for Knowledge-Based Visual Question Answering

arXiv
Knowledge-based visual question answering (KB-VQA) requires vision-language models to understand images and use external knowledge, especially for rare entities and long-tail facts. Most existing retrieval-augmented generation (RAG) methods adopt a f... read more 

Learning to Search: A Decision-Based Agent for Knowledge-Based Visual Question Answering

arXiv
Knowledge-based visual question answering (KB-VQA) requires vision-language models to understand images and use external knowledge, especially for rare entities and long-tail facts. Most existing retrieval-augmented generation (RAG) methods adopt a f... read more 

Bridging MRI and PET physiology: Untangling complementarity through orthogonal representations

arXiv
Multimodal imaging analysis often relies on joint latent representations, yet these approaches rarely define what information is shared versus modality-specific. Clarifying this distinction is clinically relevant, as it delineates the irreducible con... read more 

DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification

arXiv
Although visual foundation models like DINOv2 provide state-of-the-art performance as feature extractors, their complex, high-dimensional representations create substantial hurdles for interpretability. This work proposes DINO-QPM, which converts the... read more 

Multiple Domain Generalization Using Category Information Independent of Domain Differences

arXiv
Domain generalization is a technique aimed at enabling models to maintain high accuracy when applied to new environments or datasets (unseen domains) that differ from the datasets used in training. Generally, the accuracy of models trained on a speci... read more 

Energy-based Tissue Manifolds for Longitudinal Multiparametric MRI Analysis

arXiv
We propose a geometric framework for longitudinal multi-parametric MRI analysis based on patient-specific energy modelling in sequence space. Rather than operating on images with spatial networks, each voxel is represented by its multi-sequence inten... read more 

BRIDGE: Multimodal-to-Text Retrieval via Reinforcement-Learned Query Alignment

arXiv
Multimodal retrieval systems struggle to resolve image-text queries against text-only corpora: the best vision-language encoder achieves only 27.6 nDCG@10 on MM-BRIGHT, underperforming strong text-only retrievers. We argue the bottleneck is not the r... read more 

VersaVogue: Visual Expert Orchestration and Preference Alignment for Unified Fashion Synthesis

arXiv
Diffusion models have driven remarkable advancements in fashion image generation, yet prior works usually treat garment generation and virtual dressing as separate problems, limiting their flexibility in real-world fashion workflows. Moreover, fashio... read more 

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing

arXiv
Achieving physically accurate object manipulation in image editing is essential for its potential applications in interactive world models. However, existing visual generative models often fail at precise spatial manipulation, resulting in incorrect ... read more