Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 47,731 to 47,740 of 224,199 articles

Training-Free Generative Modeling via Kernelized Stochastic Interpolants

arXiv
We develop a kernel method for generative modeling within the stochastic interpolant framework, replacing neural network training with linear systems. The drift of the generative SDE is $\hat b_t(x) = \nablaφ(x)^\topη_t$, where $η_t\in\R^P$ solves a ... read more 

SemanticNVS: Improving Semantic Scene Understanding in Generative Novel View Synthesis

arXiv
We present SemanticNVS, a camera-conditioned multi-view diffusion model for novel view synthesis (NVS), which improves generation quality and consistency by integrating pre-trained semantic feature extractors. Existing NVS methods perform well for vi... read more 

StructXLIP: Enhancing Vision-language Models with Multimodal Structural Cues

arXiv
Edge-based representations are fundamental cues for visual understanding, a principle rooted in early vision research and still central today. We extend this principle to vision-language alignment, showing that isolating and aligning structural cues ... read more 

Transcending the Annotation Bottleneck: AI-Powered Discovery in Biology and Medicine

arXiv
The dependence on expert annotation has long constituted the primary rate-limiting step in the application of artificial intelligence to biomedicine. While supervised learning drove the initial wave of clinical algorithms, a paradigm shift towards un... read more 

Conformal Risk Control for Non-Monotonic Losses

arXiv
Conformal risk control is an extension of conformal prediction for controlling risk functions beyond miscoverage. The original algorithm controls the expected value of a loss that is monotonic in a one-dimensional parameter. Here, we present risk con... read more 

Flow3r: Factored Flow Prediction for Scalable Visual Geometry Learning

arXiv
Current feed-forward 3D/4D reconstruction systems rely on dense geometry and pose supervision -- expensive to obtain at scale and particularly scarce for dynamic real-world scenes. We present Flow3r, a framework that augments visual geometry learning... read more 

A Very Big Video Reasoning Suite

arXiv
Rapid progress in video models has largely focused on visual quality, leaving their reasoning capabilities underexplored. Video reasoning grounds intelligence in spatiotemporally consistent visual environments that go beyond what text can naturally c... read more 

A Very Big Video Reasoning Suite

arXiv
Rapid progress in video models has largely focused on visual quality, leaving their reasoning capabilities underexplored. Video reasoning grounds intelligence in spatiotemporally consistent visual environments that go beyond what text can naturally c... read more 

tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction

arXiv
We propose tttLRM, a novel large 3D reconstruction model that leverages a Test-Time Training (TTT) layer to enable long-context, autoregressive 3D reconstruction with linear computational complexity, further scaling the model's capability. Our framew... read more 

Mobile-O: Unified Multimodal Understanding and Generation on Mobile Device

arXiv
Unified multimodal models can both understand and generate visual content within a single architecture. Existing models, however, remain data-hungry and too heavy for deployment on edge devices. We present Mobile-O, a compact vision-language-diffusio... read more