Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 44,661 to 44,670 of 224,055 articles

Uncertainty-aware Blood Glucose Prediction from Continuous Glucose Monitoring Data

arXiv
In this work, we investigate uncertainty-aware neural network models for blood glucose prediction and adverse glycemic event identification in Type 1 diabetes. We consider three families of sequence models based on LSTM, GRU, and Transformer architec... read more 

VisionPangu: A Compact and Fine-Grained Multimodal Assistant with 1.7B Parameters

arXiv
Large Multimodal Models (LMMs) have achieved strong performance in vision-language understanding, yet many existing approaches rely on large-scale architectures and coarse supervision, which limits their ability to generate detailed image captions. I... read more 

Revisiting an Old Perspective Projection for Monocular 3D Morphable Models Regression

arXiv
We introduce a novel camera model for monocular 3D Morphable Model (3DMM) regression methods that effectively captures the perspective distortion effect commonly seen in close-up facial images. Fitting 3D morphable models to video is a key techniqu... read more 

BiEvLight: Bi-level Learning of Task-Aware Event Refinement for Low-Light Image Enhancement

arXiv
Event cameras, with their high dynamic range, show great promise for Low-light Image Enhancement (LLIE). Existing works primarily focus on designing effective modal fusion strategies. However, a key challenge is the dual degradation from intrinsic ba... read more 

A Simple Baseline for Unifying Understanding, Generation, and Editing via Vanilla Next-token Prediction

arXiv
In this work, we introduce Wallaroo, a simple autoregressive baseline that leverages next-token prediction to unify multi-modal understanding, image generation, and editing at the same time. Moreover, Wallaroo supports multi-resolution image input an... read more 

MultiGO++: Monocular 3D Clothed Human Reconstruction via Geometry-Texture Collaboration

arXiv
Monocular 3D clothed human reconstruction aims to generate a complete and realistic textured 3D avatar from a single image. Existing methods are commonly trained under multi-view supervision with annotated geometric priors, and during inference, thes... read more 

Physics-consistent deep learning for blind aberration recovery in mobile optics

arXiv
Mobile photography is often limited by complex, lens-specific optical aberrations. While recent deep learning methods approach this as an end-to-end deblurring task, these "black-box" models lack explicit optical modeling and can hallucinate details.... read more 

How far have we gone in Generative Image Restoration? A study on its capability, limitations and evaluation practices

arXiv
Generative Image Restoration (GIR) has achieved impressive perceptual realism, but how far have its practical capabilities truly advanced compared with previous methods? To answer this, we present a large-scale study grounded in a new multi-dimension... read more 

Tell2Adapt: A Unified Framework for Source Free Unsupervised Domain Adaptation via Vision Foundation Model

arXiv
Source Free Unsupervised Domain Adaptation (SFUDA) is critical for deploying deep learning models across diverse clinical settings. However, existing methods are typically designed for low-gap, specific domain shifts and cannot generalize into a unif... read more 

Direct Contact-Tolerant Motion Planning With Vision Language Models

arXiv
Navigation in cluttered environments often requires robots to tolerate contact with movable or deformable objects to maintain efficiency. Existing contact-tolerant motion planning (CTMP) methods rely on indirect spatial representations (e.g., prebuil... read more