Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 53,641 to 53,650 of 225,930 articles

LocateEdit-Bench: A Benchmark for Instruction-Based Editing Localization

arXiv
Recent advancements in image editing have enabled highly controllable and semantically-aware alteration of visual content, posing unprecedented challenges to manipulation localization. However, existing AI-generated forgery localization methods prima... read more 

LoGoSeg: Integrating Local and Global Features for Open-Vocabulary Semantic Segmentation

arXiv
Open-vocabulary semantic segmentation (OVSS) extends traditional closed-set segmentation by enabling pixel-wise annotation for both seen and unseen categories using arbitrary textual descriptions. While existing methods leverage vision-language model... read more 

Geometric Observability Index: An Operator-Theoretic Framework for Per-Feature Sensitivity, Weak Observability, and Dynamic Effects in SE(3) Pose Estimation

arXiv
We present a unified operator-theoretic framework for analyzing per-feature sensitivity in camera pose estimation on the Lie group SE(3). Classical sensitivity tools - conditioning analyses, Euclidean perturbation arguments, and Fisher information bo... read more 

A Mixed Reality System for Robust Manikin Localization in Childbirth Training

arXiv
Opportunities for medical students to gain practical experience in vaginal births are increasingly constrained by shortened clinical rotations, patient reluctance, and the unpredictable nature of labour. To alleviate clinicians' instructional burden ... read more 

CAViT -- Channel-Aware Vision Transformer for Dynamic Feature Fusion

arXiv
Vision Transformers (ViTs) have demonstrated strong performance across a range of computer vision tasks by modeling long-range spatial interactions via self-attention. However, channel-wise mixing in ViTs remains static, relying on fixed multilayer p... read more 

Empowering Time Series Analysis with Large-Scale Multimodal Pretraining

arXiv
While existing time series foundation models primarily rely on large-scale unimodal pretraining, they lack complementary modalities to enhance time series understanding. Building multimodal foundation models is a natural next step, but it faces key c... read more 

ShapeUP: Scalable Image-Conditioned 3D Editing

arXiv
Recent advancements in 3D foundation models have enabled the generation of high-fidelity assets, yet precise 3D manipulation remains a significant challenge. Existing 3D editing frameworks often face a difficult trade-off between visual controllabili... read more 

Perception-Based Beliefs for POMDPs with Visual Observations

arXiv
Partially observable Markov decision processes (POMDPs) are a principled planning model for sequential decision-making under uncertainty. Yet, real-world problems with high-dimensional observations, such as camera images, remain intractable for tradi... read more 

Poster: Camera Tampering Detection for Outdoor IoT Systems

arXiv
Recently, the use of smart cameras in outdoor settings has grown to improve surveillance and security. Nonetheless, these systems are susceptible to tampering, whether from deliberate vandalism or harsh environmental conditions, which can undermine t... read more 

Disc-Centric Contrastive Learning for Lumbar Spine Severity Grading

arXiv
This work examines a disc-centric approach for automated severity grading of lumbar spinal stenosis from sagittal T2-weighted MRI. The method combines contrastive pretraining with disc-level fine-tuning, using a single anatomically localized region o... read more