Artificial Intelligence Medical Compendium

Explore the latest research on artificial intelligence and machine learning in medicine.

Showing 16,171 to 16,180 of 213,568 articles

Virtual 3D H&E Staining from Phase-contrast Back-illumination Interference Tomography

arXiv
Three-dimensional (3D) histopathology of unprocessed tissues has the potential to transform disease management by enabling volumetric characterization of tissue microarchitecture and in-vivo assessment. Back-illumination Interference Tomography (BIT)... read more 

ConvNeXt-FD: A Fractal-Based Deep Model for Robust Biomedical Image Segmentation

arXiv
Biomedical image segmentation is a critical task in medical diagnosis and treatment planning, enabling precise delineation of anatomical structures and pathological regions. Despite significant advancements, challenges persist due to the inherent var... read more 

Rethinking Token Reduction for Diffusion Models via Output-Similarity-Awareness

arXiv
Diffusion Transformers (DiTs) achieve superior image generation quality but suffer from quadratic computational complexity relative to token count. While various token reduction (TR) methods have been proposed to mitigate this cost, they overlook the... read more 

PointLLM-R: Enhancing 3D Point Cloud Reasoning via Chain-of-Thought

arXiv
Understanding 3D point clouds through language remains a fundamental challenge in computer graphics and visual computing, due to the irregular structure of point cloud data and the lack of explicit reasoning in existing 3D multimodal models. While Ch... read more 

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration

arXiv
Diffusion Transformer (DiT) has emerged as a powerful model architecture for generating high-quality images and videos. In the case of video DiT, 3D Spatio-Temporal Attention increases token length in proportion to the number of frames, sharply incre... read more 

FRED: A Multi-Modal Autonomous Driving Dataset for Flooded Road Environments

arXiv
The Flooded Road Environments Dataset (FRED) is, to our knowledge, the first multi-modal autonomous driving dataset specifically targeting the collection of data from scenarios involving water hazards on the road. The dataset contains images from a 2... read more 

AgroVG: A Large-Scale Multi-Source Benchmark for Agricultural Visual Grounding

arXiv
Visual grounding, the task of localizing objects described by natural-language expressions, is a foundational capability for agricultural AI systems, enabling applications such as selective weeding, disease monitoring, and targeted harvesting. Reliab... read more 

Broken Memories: Detecting and Mitigating Memorization in Diffusion Models with Degraded Generations

arXiv
While diffusion models excel at generating high-quality images, their tendency to memorize training data poses significant privacy and copyright risks. In this work, we for the first time identify that memorization induces internal numerical instabil... read more 

Distributed Image Compression with Multimodal Side Information at Extremely Low Bitrates

arXiv
Distributed Image Compression (DIC) is crucial for multi-view transmission, especially when operating at extremely low bitrates (< 0.1 bpp). Its core challenge is effectively utilizing side information to achieve high-quality reconstruction under str... read more 

Echo4DIR: 4D Implicit Heart Reconstruction from 2D Echocardiography Videos

arXiv
Reconstructing 4D (3D+t) cardiac geometry from sparse 2D echocardiography is highly desirable yet fundamentally challenged by geometric ambiguity and temporal discontinuity. To tackle these issues, we propose Echo4DIR, a novel test-time 4D implicit r... read more