Latest AI and machine learning research in identifying and reporting dependent adult abuse for healthcare professionals.
Recent advances in open-vocabulary object detection focus primarily on two aspects: scaling up datasets and leveraging contrastive learning to align language and vision modalities. However, these approaches often neglect internal consistency within a single modality, particularly when background or environmental changes occur. This lack of consistency leads to a performance drop because the model ...
Composed Image Retrieval (CIR) is a challenging image retrieval paradigm. It aims to retrieve target images from large-scale image databases that are consistent with the modification semantics, based on a multimodal query composed of a reference image and modification text. Although existing methods have made significant progress in cross-modal alignment and feature fusion, a key flaw remains: the...
The Brain Tumor Reporting and Data System (BT-RADS) standardizes post-treatment MRI response assessment in patients with diffuse gliomas but requires ...
Cardiac ultrasound diagnosis is critical for cardiovascular disease assessment, but acquiring standard views remains highly operator-dependent. Existi...
In clinical practice, crossmodal information including medical images and tabular data is essential for disease diagnosis. There exists a significant ...
The rapid advancement of Multimodal Large Language Models (MLLMs) has enabled browsing agents to acquire and reason over multimodal information in the...
Background: Multiple stakeholders need to locate results of registered clinical trials but frequently struggle to find them. Summary results of clinic...
Human Activity Recognition using wearable inertial sensors is foundational to healthcare monitoring, fitness analytics, and context-aware computing, y...
The rapid advancement of Multimodal Large Language Models (MLLMs) has enabled browsing agents to acquire and reason over multimodal information in the...
Structured radiology reporting promises faster, more consistent communication than free text, but automation remains difficult as models must make man...
Mechanical ventilation (MV) is a life-saving intervention for patients with acute respiratory failure (ARF) in the ICU. However, inappropriate ventila...
Generative AI has advanced rapidly in medical report generation; however, its application to oral and maxillofacial CBCT reporting remains limited, la...
Large pretrained diffusion models have significantly enhanced the quality of generated videos, and yet their use in real-time streaming remains limite...
Large pretrained diffusion models have significantly enhanced the quality of generated videos, and yet their use in real-time streaming remains limite...
Predicting hospital outcomes for patients with severe acute respiratory infections is critical for risk stratification and resource planning, yet hete...
Existing concept customization methods have achieved remarkable outcomes in high-fidelity and multi-concept customization. However, they often neglect...
Accurate estimation of enzyme kinetic parameters is essential for enzyme engineering and industrial biocatalysis, yet their experimental measurement r...
Infrared image super-resolution (IISR) under real-world conditions is a practically significant yet rarely addressed task. Pioneering works are often ...
Safe visual navigation is critical for indoor mobile robots operating in cluttered environments. Existing benchmarks, however, often neglect collision...
Rapid progress in vision-language modeling has enabled pathology report generation from gigapixel whole-slide images, but most approaches assume stati...