Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 36,031 to 36,040 of 223,137 articles

EvaNet: Towards More Efficient and Consistent Infrared and Visible Image Fusion Assessment

arXiv
Evaluation is essential in image fusion research, yet most existing metrics are directly borrowed from other vision tasks without proper adaptation. These traditional metrics, often based on complex image transformations, not only fail to capture the... read more 

Toward an Artificial General Teacher: Procedural Geometry Data Generation and Visual Grounding with Vision-Language Models

arXiv
We study visual explanation in geometry education as a Referring Image Segmentation (RIS) problem: given a diagram and a natural language description, the task is to produce a pixel-level mask for the referred geometric element. However, existing RIS... read more 

High-dimensional Many-to-many-to-many Mediation Analysis

arXiv
We study high-dimensional mediation analysis in which exposures, mediators, and outcomes are all multivariate, and both exposures and mediators may be high-dimensional. We formalize this as a many (exposures)-to-many (mediators)-to-many (outcomes) (M... read more 

InstructTable: Improving Table Structure Recognition Through Instructions

arXiv
Table structure recognition (TSR) holds widespread practical importance by parsing tabular images into structured representations, yet encounters significant challenges when processing complex layouts involving merged or empty cells. Traditional visu... read more 

SPG: Sparse-Projected Guides with Sparse Autoencoders for Zero-Shot Anomaly Detection

arXiv
We study zero-shot anomaly detection and segmentation using frozen foundation model features, where all learnable parameters are trained only on a labeled auxiliary dataset and deployed to unseen target categories without any target-domain adaptation... read more 

Token Warping Helps MLLMs Look from Nearby Viewpoints

arXiv
Can warping tokens, rather than pixels, help multimodal large language models (MLLMs) understand how a scene appears from a nearby viewpoint? While MLLMs perform well on visual reasoning, they remain fragile to viewpoint changes, as pixel-wise warpin... read more 

Few-Shot Distribution-Aligned Flow Matching for Data Synthesis in Medical Image Segmentation

arXiv
Data heterogeneity hinders clinical deployment of medical image analysis models, and generative data augmentation helps mitigate this issue. However, recent diffusion-based methods that synthesize image-mask pairs often ignore distribution shifts bet... read more 

HairOrbit: Multi-view Aware 3D Hair Modeling from Single Portraits

arXiv
Reconstructing strand-level 3D hair from a single-view image is highly challenging, especially when preserving consistent and realistic attributes in unseen regions. Existing methods rely on limited frontal-view cues and small-scale/style-restricted ... read more 

Adaptive Local Frequency Filtering for Fourier-Encoded Implicit Neural Representations

arXiv
Fourier-encoded implicit neural representations (INRs) have shown strong capability in modeling continuous signals from discrete samples. However, conventional Fourier feature mappings use a fixed set of frequencies over the entire spatial domain, ma... read more 

ViraHinter: a dual-modal artificial intelligence framework for predicting virus-host interactions

arXiv
Protein-protein interactions (PPIs) between a virus and its host govern infection, replication, and pathogenesis. While high-throughput mapping has identified thousands of virus-host associations, much of the virus-host interactome remains uncharacte... read more