Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 46,681 to 46,690 of 224,199 articles

TabDLM: Free-Form Tabular Data Generation via Joint Numerical-Language Diffusion

arXiv
Synthetic tabular data generation has attracted growing attention due to its importance for data augmentation, foundation models, and privacy. However, real-world tabular datasets increasingly contain free-form text fields (e.g., reviews or clinical ... read more 

LoR-LUT: Learning Compact 3D Lookup Tables via Low-Rank Residuals

arXiv
We present LoR-LUT, a unified low-rank formulation for compact and interpretable 3D lookup table (LUT) generation. Unlike conventional 3D-LUT-based techniques that rely on fusion of basis LUTs, which are usually dense tensors, our unified approach ex... read more 

Coded-E2LF: Coded Aperture Light Field Imaging from Events

arXiv
We propose Coded-E2LF (coded event to light field), a computational imaging method for acquiring a 4-D light field using a coded aperture and a stationary event-only camera. In a previous work, an imaging system similar to ours was adopted, but both ... read more 

CGSA: Class-Guided Slot-Aware Adaptation for Source-Free Object Detection

arXiv
Source-Free Domain Adaptive Object Detection (SF-DAOD) aims to adapt a detector trained on a labeled source domain to an unlabeled target domain without retaining any source data. Despite recent progress, most popular approaches focus on tuning pseud... read more 

Instruction-based Image Editing with Planning, Reasoning, and Generation

arXiv
Editing images via instruction provides a natural way to generate interactive content, but it is a big challenge due to the higher requirement of scene understanding and generation. Prior work utilizes a chain of large language models, object segment... read more 

DiffBMP: Differentiable Rendering with Bitmap Primitives

arXiv
We introduce DiffBMP, a scalable and efficient differentiable rendering engine for a collection of bitmap images. Our work addresses a limitation that traditional differentiable renderers are constrained to vector graphics, given that most images in ... read more 

Plug, Play, and Fortify: A Low-Cost Module for Robust Multimodal Image Understanding Models

arXiv
Missing modalities present a fundamental challenge in multimodal models, often causing catastrophic performance degradation. Our observations suggest that this fragility stems from an imbalanced learning process, where the model develops an implicit ... read more 

Interactive Medical-SAM2 GUI: A Napari-based semi-automatic annotation tool for medical images

arXiv
Interactive Medical-SAM2 GUI is an open-source desktop application for semi-automatic annotation of 2D and 3D medical images. Built on the Napari multi-dimensional viewer, box/point prompting is integrated with SAM2-style propagation by treating a 3D... read more 

Denoising as Path Planning: Training-Free Acceleration of Diffusion Models with DPCache

arXiv
Diffusion models have demonstrated remarkable success in image and video generation, yet their practical deployment remains hindered by the substantial computational overhead of multi-step iterative sampling. Among acceleration strategies, caching-ba... read more 

ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport

arXiv
Image-text retrieval has become a fundamental component in intelligent multimedia systems; however, most existing vision-language models are optimized for highresource languages and remain suboptimal for low-resource settings such as Vietnamese. This... read more