Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 44,331 to 44,340 of 224,055 articles

PatchCue: Enhancing Vision-Language Model Reasoning with Patch-Based Visual Cues

arXiv
Vision-Language Models (VLMs) have achieved remarkable progress on a wide range of challenging multimodal understanding and reasoning tasks. However, existing reasoning paradigms, such as the classical Chain-of-Thought (CoT), rely solely on textual i... read more 

Shifting Adaptation from Weight Space to Memory Space: A Memory-Augmented Agent for Medical Image Segmentation

arXiv
Medical image segmentation is fundamental to clinical workflows, yet models trained on a single dataset often fail to generalize across institutions, scanners, or patient populations. While vision foundation models have shown great promise in address... read more 

Systematic Evaluation of Novel View Synthesis for Video Place Recognition

arXiv
The generation of synthetic novel views has the potential to positively impact robot navigation in several ways. In image-based navigation, a novel overhead view generated from a scene taken by a ground robot could be used to guide an aerial robot to... read more 

PixARMesh: Autoregressive Mesh-Native Single-View Scene Reconstruction

arXiv
We introduce PixARMesh, a method to autoregressively reconstruct complete 3D indoor scene meshes directly from a single RGB image. Unlike prior methods that rely on implicit signed distance fields and post-hoc layout optimization, PixARMesh jointly p... read more 

Measuring Perceptions of Fairness in AI Systems: The Effects of Infra-marginality

arXiv
Differences in data distributions between demographic groups, known as the problem of infra-marginality, complicate how people evaluate fairness in machine learning models. We present a user study with 85 participants in a hypothetical medical decisi... read more 

InnoAds-Composer: Efficient Condition Composition for E-Commerce Poster Generation

arXiv
E-commerce product poster generation aims to automatically synthesize a single image that effectively conveys product information by presenting a subject, text, and a designed style. Recent diffusion models with fine-grained and efficient controllabi... read more 

Mitigating Bias in Concept Bottleneck Models for Fair and Interpretable Image Classification

arXiv
Ensuring fairness in image classification prevents models from perpetuating and amplifying bias. Concept bottleneck models (CBMs) map images to high-level, human-interpretable concepts before making predictions via a sparse, one-layer classifier. Thi... read more 

CollabOD: Collaborative Multi-Backbone with Cross-scale Vision for UAV Small Object Detection

arXiv
Small object detection in unmanned aerial vehicle (UAV) imagery is challenging, mainly due to scale variation, structural detail degradation, and limited computational resources. In high-altitude scenarios, fine-grained features are further weakened ... read more 

Beyond Geometry: Artistic Disparity Synthesis for Immersive 2D-to-3D

arXiv
Current 2D-to-3D conversion methods achieve geometric accuracy but are artistically deficient, failing to replicate the immersive and emotionally resonant experience of professional 3D cinema. This is because geometric reconstruction paradigms mistak... read more 

Pano3DComposer: Feed-Forward Compositional 3D Scene Generation from Single Panoramic Image

arXiv
Current compositional image-to-3D scene generation approaches construct 3D scenes by time-consuming iterative layout optimization or inflexible joint object-layout generation. Moreover, most methods rely on limited field-of-view perspective images, h... read more