Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 44,301 to 44,310 of 224,055 articles

CLoPA: Continual Low Parameter Adaptation of Interactive Segmentation for Medical Image Annotation

arXiv
Interactive segmentation enables clinicians to guide annotation, but existing zero-shot models like nnInteractive fail to consistently reach expert-level performance across diverse medical imaging tasks. Because annotation campaigns produce a growing... read more 

CaTok: Taming Mean Flows for One-Dimensional Causal Image Tokenization

arXiv
Autoregressive (AR) language models rely on causal tokenization, but extending this paradigm to vision remains non-trivial. Current visual tokenizers either flatten 2D patches into non-causal sequences or enforce heuristic orderings that misalign wit... read more 

Pinterest Canvas: Large-Scale Image Generation at Pinterest

arXiv
While recent image generation models demonstrate a remarkable ability to handle a wide variety of image generation tasks, this flexibility makes them hard to control via prompting or simple inference adaptation alone, rendering them unsuitable for us... read more 

Training Flow Matching: The Role of Weighting and Parameterization

arXiv
We study the training objectives of denoising-based generative models, with a particular focus on loss weighting and output parameterization, including noise-, clean image-, and velocity-based formulations. Through a systematic numerical study, we an... read more 

Do Foundation Models Know Geometry? Probing Frozen Features for Continuous Physical Measurement

arXiv
Vision-language models encode continuous geometry that their text pathway fails to express: a 6,000-parameter linear probe extracts hand joint angles at 6.1 degrees MAE from frozen features, while the best text output achieves only 20.0 degrees -- a ... read more 

GreenRFM: Toward a resource-efficient radiology foundation model

arXiv
The development of radiology foundation models (RFMs) is hindered by a reliance on brute-force scaling. Existing approaches often directly translate methods for natural images, which prioritize scale over precision and hence lead to brittle and expen... read more 

Match4Annotate: Propagating Sparse Video Annotations via Implicit Neural Feature Matching

arXiv
Acquiring per-frame video annotations remains a primary bottleneck for deploying computer vision in specialized domains such as medical imaging, where expert labeling is slow and costly. Label propagation offers a natural solution, yet existing appro... read more 

PONTE: Personalized Orchestration for Natural Language Trustworthy Explanations

arXiv
Explainable Artificial Intelligence (XAI) seeks to enhance the transparency and accountability of machine learning systems, yet most methods follow a one-size-fits-all paradigm that neglects user differences in expertise, goals, and cognitive needs. ... read more 

Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis

arXiv
Strong semantic representations improve the convergence and generation quality of diffusion and flow models. Existing approaches largely rely on external models, which require separate training, operate on misaligned objectives, and exhibit unexpecte... read more 

When One Modality Rules Them All: Backdoor Modality Collapse in Multimodal Diffusion Models

arXiv
While diffusion models have revolutionized visual content generation, their rapid adoption has underscored the critical need to investigate vulnerabilities, e.g., to backdoor attacks. In multimodal diffusion models, it is natural to expect that attac... read more