Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 54,461 to 54,470 of 226,183 articles

VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text?

arXiv
Vision-Language Models (VLMs) have achieved impressive performance in cross-modal understanding across textual and visual inputs, yet existing benchmarks predominantly focus on pure-text queries. In real-world scenarios, language also frequently appe... read more 

X2HDR: HDR Image Generation in a Perceptually Uniform Space

arXiv
High-dynamic-range (HDR) formats and displays are becoming increasingly prevalent, yet state-of-the-art image generators (e.g., Stable Diffusion and FLUX) typically remain limited to low-dynamic-range (LDR) output due to the lack of large-scale HDR t... read more 

XtraLight-MedMamba for Classification of Neoplastic Tubular Adenomas

arXiv
Accurate risk stratification of precancerous polyps during routine colonoscopy screenings is essential for lowering the risk of developing colorectal cancer (CRC). However, assessment of low-grade dysplasia remains limited by subjective histopatholog... read more 

Toward Reliable and Explainable Nail Disease Classification: Leveraging Adversarial Training and Grad-CAM Visualization

arXiv
Human nail diseases are gradually observed over all age groups, especially among older individuals, often going ignored until they become severe. Early detection and accurate diagnosis of such conditions are important because they sometimes reveal ou... read more 

When LLaVA Meets Objects: Token Composition for Vision-Language-Models

arXiv
Current autoregressive Vision Language Models (VLMs) usually rely on a large number of visual tokens to represent images, resulting in a need for more compute especially at inference time. To address this problem, we propose Mask-LLaVA, a framework t... read more 

Laminating Representation Autoencoders for Efficient Diffusion

arXiv
Recent work has shown that diffusion models can generate high-quality images by operating directly on SSL patch features rather than pixel-space latents. However, the dense patch grids from encoders like DINOv2 contain significant redundancy, making ... read more 

PerpetualWonder: Long-Horizon Action-Conditioned 4D Scene Generation

arXiv
We introduce PerpetualWonder, a hybrid generative simulator that enables long-horizon, action-conditioned 4D scene generation from a single image. Current works fail at this task because their physical state is decoupled from their visual representat... read more 

Reinforced Attention Learning

arXiv
Post-training with Reinforcement Learning (RL) has substantially improved reasoning in Large Language Models (LLMs) via test-time scaling. However, extending this paradigm to Multimodal LLMs (MLLMs) through verbose rationales yields limited gains for... read more 

Prioritizing nurse-patient relationships in the digital transformation of nursing: A discussion paper.

International journal of nursing studies
Healthcare is undergoing rapid digital transformation, with artificial intelligence (AI), automated documentation, and predictive analytics now integrated into clinical workflows. These technologies promise to reduce administrative burden and expand ... read more 

Learning with less: A survey of deep learning in medical imaging under varying supervision levels.

Artificial intelligence in medicine
The rapid evolution of deep learning has significantly advanced the field of medical image analysis. However, despite these achievements, further progress is hindered by the scarcity of large, well-annotated datasets. To overcome this limitation, rec... read more