Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 36,041 to 36,050 of 223,137 articles

STRNet: Visual Navigation with Spatio-Temporal Representation through Dynamic Graph Aggregation

arXiv
Visual navigation requires the robot to reach a specified goal such as an image, based on a sequence of first-person visual observations. While recent learning-based approaches have made significant progress, they often focus on improving policy head... read more 

NavCrafter: Exploring 3D Scenes from a Single Image

arXiv
Creating flexible 3D scenes from a single image is vital when direct 3D data acquisition is costly or impractical. We introduce NavCrafter, a novel framework that explores 3D scenes from a single image by synthesizing novel-view video sequences with ... read more 

PaveBench: A Versatile Benchmark for Pavement Distress Perception and Interactive Vision-Language Analysis

arXiv
Pavement condition assessment is essential for road safety and maintenance. Existing research has made significant progress. However, most studies focus on conventional computer vision tasks such as classification, detection, and segmentation. In rea... read more 

EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors

arXiv
Vision-Language Models (VLMs) excel at multimodal tasks, but they remain vulnerable to hallucinations that are factually incorrect or ungrounded in the input image. Recent work suggests that hallucination detection using internal representations is m... read more 

A Unified Perspective on Adversarial Membership Manipulation in Vision Models

arXiv
Membership inference attacks (MIAs) aim to determine whether a specific data point was part of a model's training set, serving as effective tools for evaluating privacy leakage of vision models. However, existing MIAs implicitly assume honest query i... read more 

InverseDraping: Recovering Sewing Patterns from 3D Garment Surfaces via BoxMesh Bridging

arXiv
Recovering sewing patterns from draped 3D garments is a challenging problem in human digitization research. In contrast to the well-studied forward process of draping designed sewing patterns using mature physical simulation engines, the inverse proc... read more 

Differentiable Stroke Planning with Dual Parameterization for Efficient and High-Fidelity Painting Creation

arXiv
In stroke-based rendering, search methods often get trapped in local minima due to discrete stroke placement, while differentiable optimizers lack structural awareness and produce unstructured layouts. To bridge this gap, we propose a dual representa... read more 

Visual Instruction-Finetuned Language Model for Versatile Brain MR Image Tasks

arXiv
LLMs have demonstrated remarkable capabilities in linguistic reasoning and are increasingly adept at vision-language tasks. The integration of image tokens into transformers has enabled direct visual input and output, advancing research from image-to... read more 

Task-Guided Prompting for Unified Remote Sensing Image Restoration

arXiv
Remote sensing image restoration (RSIR) is essential for recovering high-fidelity imagery from degraded observations, enabling accurate downstream analysis. However, most existing methods focus on single degradation types within homogeneous data, res... read more 

ExploreVLA: Dense World Modeling and Exploration for End-to-End Autonomous Driving

arXiv
End-to-end autonomous driving models based on Vision-Language-Action (VLA) architectures have shown promising results by learning driving policies through behavior cloning on expert demonstrations. However, imitation learning inherently limits the mo... read more