Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 34,711 to 34,720 of 221,633 articles

Token Warping Helps MLLMs Look from Nearby Viewpoints

arXiv
Can warping tokens, rather than pixels, help multimodal large language models (MLLMs) understand how a scene appears from a nearby viewpoint? While MLLMs perform well on visual reasoning, they remain fragile to viewpoint changes, as pixel-wise warpin... read more 

SPG: Sparse-Projected Guides with Sparse Autoencoders for Zero-Shot Anomaly Detection

arXiv
We study zero-shot anomaly detection and segmentation using frozen foundation model features, where all learnable parameters are trained only on a labeled auxiliary dataset and deployed to unseen target categories without any target-domain adaptation... read more 

InstructTable: Improving Table Structure Recognition Through Instructions

arXiv
Table structure recognition (TSR) holds widespread practical importance by parsing tabular images into structured representations, yet encounters significant challenges when processing complex layouts involving merged or empty cells. Traditional visu... read more 

High-dimensional Many-to-many-to-many Mediation Analysis

arXiv
We study high-dimensional mediation analysis in which exposures, mediators, and outcomes are all multivariate, and both exposures and mediators may be high-dimensional. We formalize this as a many (exposures)-to-many (mediators)-to-many (outcomes) (M... read more 

Toward an Artificial General Teacher: Procedural Geometry Data Generation and Visual Grounding with Vision-Language Models

arXiv
We study visual explanation in geometry education as a Referring Image Segmentation (RIS) problem: given a diagram and a natural language description, the task is to produce a pixel-level mask for the referred geometric element. However, existing RIS... read more 

EvaNet: Towards More Efficient and Consistent Infrared and Visible Image Fusion Assessment

arXiv
Evaluation is essential in image fusion research, yet most existing metrics are directly borrowed from other vision tasks without proper adaptation. These traditional metrics, often based on complex image transformations, not only fail to capture the... read more 

BioUNER: A Benchmark Dataset for Clinical Urdu Named Entity Recognition

arXiv
In this article, we present a gold-standard benchmark dataset for Biomedical Urdu Named Entity Recognition (BioUNER), developed by crawling health-related articles from online Urdu news portals, medical prescriptions, and hospital health blogs and we... read more 

Learning from Synthetic Data via Provenance-Based Input Gradient Guidance

arXiv
Learning methods using synthetic data have attracted attention as an effective approach for increasing the diversity of training data while reducing collection costs, thereby improving the robustness of model discrimination. However, many existing me... read more 

Visual Prototype Conditioned Focal Region Generation for UAV-Based Object Detection

arXiv
Unmanned aerial vehicle (UAV) based object detection is a critical but challenging task, when applied in dynamically changing scenarios with limited annotated training data. Layout-to-image generation approaches have proved effective in promoting det... read more 

Extending deep learning U-Net architecture for predicting unsteady fluid flows in textured microchannels

arXiv
In this study, we have explored an application of deep learning architecture of the U-Net model, originally designed for biomedical image segmentation, in a regression analysis aimed at predicting fluid flows through textured microchannels. The data ... read more