Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 47,291 to 47,300 of 224,199 articles

Learning to Fuse and Reconstruct Multi-View Graphs for Diabetic Retinopathy Grading

arXiv
Diabetic retinopathy (DR) is one of the leading causes of vision loss worldwide, making early and accurate DR grading critical for timely intervention. Recent clinical practices leverage multi-view fundus images for DR detection with a wide coverage ... read more 

Bayesian Generative Adversarial Networks via Gaussian Approximation for Tabular Data Synthesis

arXiv
Generative Adversarial Networks (GAN) have been used in many studies to synthesise mixed tabular data. Conditional tabular GAN (CTGAN) have been the most popular variant but struggle to effectively navigate the risk-utility trade-off. Bayesian GAN ha... read more 

MindDriver: Introducing Progressive Multimodal Reasoning for Autonomous Driving

arXiv
Vision-Language Models (VLM) exhibit strong reasoning capabilities, showing promise for end-to-end autonomous driving systems. Chain-of-Thought (CoT), as VLM's widely used reasoning strategy, is facing critical challenges. Existing textual CoT has a ... read more 

Global-Local Dual Perception for MLLMs in High-Resolution Text-Rich Image Translation

arXiv
Text Image Machine Translation (TIMT) aims to translate text embedded in images in the source-language into target-language, requiring synergistic integration of visual perception and linguistic understanding. Existing TIMT methods, whether cascaded ... read more 

Robustness in sparse artificial neural networks trained with adaptive topology

arXiv
We investigate the robustness of sparse artificial neural networks trained with adaptive topology. We focus on a simple yet effective architecture consisting of three sparse layers with 99% sparsity followed by a dense layer, applied to image classif... read more 

Global-Aware Edge Prioritization for Pose Graph Initialization

arXiv
The pose graph is a core component of Structure-from-Motion (SfM), where images act as nodes and edges encode relative poses. Since geometric verification is expensive, SfM pipelines restrict the pose graph to a sparse set of candidate edges, making ... read more 

Dream-SLAM: Dreaming the Unseen for Active SLAM in Dynamic Environments

arXiv
In addition to the core tasks of simultaneous localization and mapping (SLAM), active SLAM additionally in- volves generating robot actions that enable effective and efficient exploration of unknown environments. However, existing active SLAM pipelin... read more 

When LoRA Betrays: Backdooring Text-to-Image Models by Masquerading as Benign Adapters

arXiv
Low-Rank Adaptation (LoRA) has emerged as a leading technique for efficiently fine-tuning text-to-image diffusion models, and its widespread adoption on open-source platforms has fostered a vibrant culture of model sharing and customization. However,... read more 

PatchDenoiser: Parameter-efficient multi-scale patch learning and fusion denoiser for medical images

arXiv
Medical images are essential for diagnosis, treatment planning, and research, but their quality is often degraded by noise from low-dose acquisition, patient motion, or scanner limitations, affecting both clinical interpretation and downstream analys... read more 

PanoEnv: Exploring 3D Spatial Intelligence in Panoramic Environments with Reinforcement Learning

arXiv
360 panoramic images are increasingly used in virtual reality, autonomous driving, and robotics for holistic scene understanding. However, current Vision-Language Models (VLMs) struggle with 3D spatial reasoning on Equirectangular Projection (ERP) im... read more