Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 44,341 to 44,350 of 224,055 articles

CORE-Seg: Reasoning-Driven Segmentation for Complex Lesions via Reinforcement Learning

arXiv
Medical image segmentation is undergoing a paradigm shift from conventional visual pattern matching to cognitive reasoning analysis. Although Multimodal Large Language Models (MLLMs) have shown promise in integrating linguistic and visual knowledge, ... read more 

BlackMirror: Black-Box Backdoor Detection for Text-to-Image Models via Instruction-Response Deviation

arXiv
This paper investigates the challenging task of detecting backdoored text-to-image models under black-box settings and introduces a novel detection framework BlackMirror. Existing approaches typically rely on analyzing image-level similarity, under t... read more 

Beyond Static Frames: Temporal Aggregate-and-Restore Vision Transformer for Human Pose Estimation

arXiv
Vision Transformers (ViTs) have recently achieved state-of-the-art performance in 2D human pose estimation due to their strong global modeling capability. However, existing ViT-based pose estimators are designed for static images and process each fra... read more 

FTSplat: Feed-forward Triangle Splatting Network

arXiv
High-fidelity three-dimensional (3D) reconstruction is essential for robotics and simulation. While Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS) achieve impressive rendering quality, their reliance on time-consuming per-scene optimi... read more 

OD-RASE: Ontology-Driven Risk Assessment and Safety Enhancement for Autonomous Driving

arXiv
Although autonomous driving systems demonstrate high perception performance, they still face limitations when handling rare situations or complex road structures. Such road infrastructures are designed for human drivers, safety improvements are typic... read more 

SLER-IR: Spherical Layer-wise Expert Routing for All-in-One Image Restoration

arXiv
Image restoration under diverse degradations remains challenging for unified all-in-one frameworks due to feature interference and insufficient expert specialization. We propose SLER-IR, a spherical layer-wise expert routing framework that dynamicall... read more 

Adaptive Radial Projection on Fourier Magnitude Spectrum for Document Image Skew Estimation

arXiv
Skew estimation is one of the vital tasks in document processing systems, especially for scanned document images, because its performance impacts subsequent steps directly. Over the years, an enormous number of researches focus on this challenging pr... read more 

LucidNFT: LR-Anchored Multi-Reward Preference Optimization for Generative Real-World Super-Resolution

arXiv
Generative real-world image super-resolution (Real-ISR) can synthesize visually convincing details from severely degraded low-resolution (LR) inputs, yet its stochastic sampling makes a critical failure mode hard to avoid: outputs may look sharp but ... read more 

Energy-Driven Adaptive Visual Token Pruning for Efficient Vision-Language Models

arXiv
Visual token reduction is critical for accelerating Vision-Language Models (VLMs), yet most existing approaches rely on a fixed budget shared across all inputs, overlooking the substantial variation in image information density. We propose E-AdaPrune... read more 

Exploring Open-Vocabulary Object Recognition in Images using CLIP

arXiv
To address the limitations of existing open-vocabulary object recognition methods, specifically high system complexity, substantial training costs, and limited generalization, this paper proposes a novel Open-Vocabulary Object Recognition (OVOR) fram... read more