Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 55,821 to 55,830 of 226,731 articles

Interacted Planes Reveal 3D Line Mapping

arXiv
3D line mapping from multi-view RGB images provides a compact and structured visual representation of scenes. We study the problem from a physical and topological perspective: a 3D line most naturally emerges as the edge of a finite 3D planar patch. ... read more 

Interaction-Consistent Object Removal via MLLM-Based Reasoning

arXiv
Image-based object removal often erases only the named target, leaving behind interaction evidence that renders the result semantically inconsistent. We formalize this problem as Interaction-Consistent Object Removal (ICOR), which requires removing n... read more 

ReDiStory: Region-Disentangled Diffusion for Consistent Visual Story Generation

arXiv
Generating coherent visual stories requires maintaining subject identity across multiple images while preserving frame-specific semantics. Recent training-free methods concatenate identity and frame prompts into a unified representation, but this oft... read more 

StoryState: Agent-Based State Control for Consistent and Editable Storybooks

arXiv
Large multimodal models have enabled one-click storybook generation, where users provide a short description and receive a multi-page illustrated story. However, the underlying story state, such as characters, world settings, and page-level objects, ... read more 

DeCorStory: Gram-Schmidt Prompt Embedding Decorrelation for Consistent Storytelling

arXiv
Maintaining visual and semantic consistency across frames is a key challenge in text-to-image storytelling. Existing training-free methods, such as One-Prompt-One-Story, concatenate all prompts into a single sequence, which often induces strong embed... read more 

FlowCast: Trajectory Forecasting for Scalable Zero-Cost Speculative Flow Matching

arXiv
Flow Matching (FM) has recently emerged as a powerful approach for high-quality visual generation. However, their prohibitively slow inference due to a large number of denoising steps limits their potential use in real-time or interactive application... read more 

Beyond Pixels: Visual Metaphor Transfer via Schema-Driven Agentic Reasoning

arXiv
A visual metaphor constitutes a high-order form of human creativity, employing cross-domain semantic fusion to transform abstract concepts into impactful visual rhetoric. Despite the remarkable progress of generative AI, existing models remain largel... read more 

Deep Variational Contrastive Learning for Joint Risk Stratification and Time-to-Event Estimation

arXiv
Survival analysis is essential for clinical decision-making, as it allows practitioners to estimate time-to-event outcomes, stratify patient risk profiles, and guide treatment planning. Deep learning has revolutionized this field with unprecedented p... read more 

PromptRL: Prompt Matters in RL for Flow-Based Image Generation

arXiv
Flow matching models (FMs) have revolutionized text-to-image (T2I) generation, with reinforcement learning (RL) serving as a critical post-training strategy for alignment with reward objectives. In this research, we show that current RL pipelines for... read more 

Stronger Semantic Encoders Can Harm Relighting Performance: Probing Visual Priors via Augmented Latent Intrinsics

arXiv
Image-to-image relighting requires representations that disentangle scene properties from illumination. Recent methods rely on latent intrinsic representations but remain under-constrained and often fail on challenging materials such as metal and gla... read more