Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 40,521 to 40,530 of 223,737 articles

A Creative Agent is Worth a 64-Token Template

arXiv
Text-to-image (T2I) models have substantially improved image fidelity and prompt adherence, yet their creativity remains constrained by reliance on discrete natural language prompts. When presented with fuzzy prompts such as ``a creative vinyl record... read more 

SpiderCam: Low-Power Snapshot Depth from Differential Defocus

arXiv
We introduce SpiderCam, an FPGA-based snapshot depth-from-defocus camera which produces 480x400 sparse depth maps in real-time at 32.5 FPS over a working range of 52 cm while consuming 624 mW of power in total. SpiderCam comprises a custom camera tha... read more 

SegFly: A 2D-3D-2D Paradigm for Aerial RGB-Thermal Semantic Segmentation at Scale

arXiv
Semantic segmentation for uncrewed aerial vehicles (UAVs) is fundamental for aerial scene understanding, yet existing RGB and RGB-T datasets remain limited in scale, diversity, and annotation efficiency due to the high cost of manual labeling and the... read more 

A practical artificial intelligence framework for legal age estimation using clavicle computed tomography scans

arXiv
Legal age estimation plays a critical role in forensic and medico-legal contexts, where decisions must be supported by accurate, robust, and reproducible methods with explicit uncertainty quantification. While prior artificial intelligence (AI)-based... read more 

TransText: Transparency Aware Image-to-Video Typography Animation

arXiv
We introduce the first method, to the best of our knowledge, for adapting image-to-video models to layer-aware text (glyph) animation, a capability critical for practical dynamic visual design. Existing approaches predominantly handle the transparenc... read more 

TransText: Alpha-as-RGB Representation for Transparent Text Animation

arXiv
We introduce the first method, to the best of our knowledge, for adapting image-to-video models to layer-aware text (glyph) animation, a capability critical for practical dynamic visual design. Existing approaches predominantly handle the transparenc... read more 

LaDe: Unified Multi-Layered Graphic Media Generation and Decomposition

arXiv
Media design layer generation enables the creation of fully editable, layered design documents such as posters, flyers, and logos using only natural language prompts. Existing methods either restrict outputs to a fixed number of layers or require eac... read more 

Robust-ComBat: Mitigating Outlier Effects in Diffusion MRI Data Harmonization

arXiv
Harmonization methods such as ComBat and its variants are widely used to mitigate diffusion MRI (dMRI) site-specific biases. However, ComBat assumes that subject distributions exhibit a Gaussian profile. In practice, patients with neurological disord... read more 

AdaRadar: Rate Adaptive Spectral Compression for Radar-based Perception

arXiv
Radar is a critical perception modality in autonomous driving systems due to its all-weather characteristics and ability to measure range and Doppler velocity. However, the sheer volume of high-dimensional raw radar data saturates the communication l... read more 

The Unreasonable Effectiveness of Text Embedding Interpolation for Continuous Image Steering

arXiv
We present a training-free framework for continuous and controllable image editing at test time for text-conditioned generative models. In contrast to prior approaches that rely on additional training or manual user intervention, we find that a simpl... read more