Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 49,641 to 49,650 of 224,513 articles

CT-Bench: A Benchmark for Multimodal Lesion Understanding in Computed Tomography

arXiv
Artificial intelligence (AI) can automatically delineate lesions on computed tomography (CT) and generate radiology report content, yet progress is limited by the scarcity of publicly available CT datasets with lesion-level annotations. To bridge thi... read more 

Web-Scale Multimodal Summarization using CLIP-Based Semantic Alignment

arXiv
We introduce Web-Scale Multimodal Summarization, a lightweight framework for generating summaries by combining retrieved text and image data from web sources. Given a user-defined topic, the system performs parallel web, news, and image searches. Ret... read more 

Picking the Right Specialist: Attentive Neural Process-based Selection of Task-Specialized Models as Tools for Agentic Healthcare Systems

arXiv
Task-specialized models form the backbone of agentic healthcare systems, enabling the agents to answer clinical queries across tasks such as disease diagnosis, localization, and report generation. Yet, for a given task, a single "best" model rarely e... read more 

Wrivinder: Towards Spatial Intelligence for Geo-locating Ground Images onto Satellite Imagery

arXiv
Aligning ground-level imagery with geo-registered satellite maps is crucial for mapping, navigation, and situational awareness, yet remains challenging under large viewpoint gaps or when GPS is unreliable. We introduce Wrivinder, a zero-shot, geometr... read more 

Activation-Space Uncertainty Quantification for Pretrained Networks

arXiv
Reliable uncertainty estimates are crucial for deploying pretrained models; yet, many strong methods for quantifying uncertainty require retraining, Monte Carlo sampling, or expensive second-order computations and may alter a frozen backbone's predic... read more 

PAct: Part-Decomposed Single-View Articulated Object Generation

arXiv
Articulated objects are central to interactive 3D applications, including embodied AI, robotics, and VR/AR, where functional part decomposition and kinematic motion are essential. Yet producing high-fidelity articulated assets remains difficult to sc... read more 

ThermEval: A Structured Benchmark for Evaluation of Vision-Language Models on Thermal Imagery

arXiv
Vision language models (VLMs) achieve strong performance on RGB imagery, but they do not generalize to thermal images. Thermal sensing plays a critical role in settings where visible light fails, including nighttime surveillance, search and rescue, a... read more 

Efficient Sampling with Discrete Diffusion Models: Sharp and Adaptive Guarantees

arXiv
Diffusion models over discrete spaces have recently shown striking empirical success, yet their theoretical foundations remain incomplete. In this paper, we study the sampling efficiency of score-based discrete diffusion models under a continuous-tim... read more 

Cold-Start Personalization via Training-Free Priors from Structured World Models

arXiv
Cold-start personalization requires inferring user preferences through interaction when no user-specific historical data is available. The core challenge is a routing problem: each task admits dozens of preference dimensions, yet individual users car... read more 

Image Generation with a Sphere Encoder

arXiv
We introduce the Sphere Encoder, an efficient generative framework capable of producing images in a single forward pass and competing with many-step diffusion models using fewer than five steps. Our approach works by learning an encoder that maps nat... read more