Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 64,151 to 64,160 of 231,309 articles

Enhancing the quality of gauge images captured in smoke and haze scenes through deep learning

arXiv
Images captured in hazy and smoky environments suffer from reduced visibility, posing a challenge when monitoring infrastructures and hindering emergency services during critical situations. The proposed work investigates the use of the deep learning... read more 

Inference-time Physics Alignment of Video Generative Models with Latent World Models

arXiv
State-of-the-art video generative models produce promising visual content yet often violate basic physics principles, limiting their utility. While some attribute this deficiency to insufficient physics understanding from pre-training, we find that t... read more 

DeepUrban: Interaction-Aware Trajectory Prediction and Planning for Automated Driving by Aerial Imagery

arXiv
The efficacy of autonomous driving systems hinges critically on robust prediction and planning capabilities. However, current benchmarks are impeded by a notable scarcity of scenarios featuring dense traffic, which is essential for understanding and ... read more 

Representation-Aware Unlearning via Activation Signatures: From Suppression to Knowledge-Signature Erasure

arXiv
Selective knowledge erasure from LLMs is critical for GDPR compliance and model safety, yet current unlearning methods conflate behavioral suppression with true knowledge removal, allowing latent capabilities to persist beneath surface-level refusals... read more 

Jordan-Segmentable Masks: A Topology-Aware definition for characterizing Binary Image Segmentation

arXiv
Image segmentation plays a central role in computer vision. However, widely used evaluation metrics, whether pixel-wise, region-based, or boundary-focused, often struggle to capture the structural and topological coherence of a segmentation. In many ... read more 

RSATalker: Realistic Socially-Aware Talking Head Generation for Multi-Turn Conversation

arXiv
Talking head generation is increasingly important in virtual reality (VR), especially for social scenarios involving multi-turn conversation. Existing approaches face notable limitations: mesh-based 3D methods can model dual-person dialogue but lack ... read more 

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding

arXiv
Today's strongest video-language models (VLMs) remain proprietary. The strongest open-weight models either rely on synthetic data from proprietary VLMs, effectively distilling from them, or do not disclose their training data or recipe. As a result, ... read more 

Alterbute: Editing Intrinsic Attributes of Objects in Images

arXiv
We introduce Alterbute, a diffusion-based method for editing an object's intrinsic attributes in an image. We allow changing color, texture, material, and even the shape of an object, while preserving its perceived identity and scene context. Existin... read more 

DInf-Grid: A Neural Differential Equation Solver with Differentiable Feature Grids

arXiv
We present a novel differentiable grid-based representation for efficiently solving differential equations (DEs). Widely used architectures for neural solvers, such as sinusoidal neural networks, are coordinate-based MLPs that are both computationall... read more 

Future Optical Flow Prediction Improves Robot Control & Video Generation

arXiv
Future motion representations, such as optical flow, offer immense value for control and generative tasks. However, forecasting generalizable spatially dense motion representations remains a key challenge, and learning such forecasting from noisy, re... read more