Latest AI and machine learning research in prescriptions for healthcare professionals.
Single-modal object detection tasks often experience performance degradation when encountering diverse scenarios. In contrast, multimodal object detection tasks can offer more comprehensive information about object features by integrating data from various modalities. Current multimodal object detection methods generally use various fusion techniques, including conventional neural networks and t...
The personalization model has gained significant attention in image generation yet remains underexplored for large vision-language models (LVLMs). Beyond generic ones, with personalization, LVLMs handle interactive dialogues using referential concepts (e.g., ``Mike and Susan are talking.'') instead of the generic form (e.g., ``a boy and a girl are talking.''), making the conversation more custom...
With the advancement of deep learning, object detectors (ODs) with various architectures have achieved significant success in complex scenarios like...
This study introduces an adaptive user interface generation technology, emphasizing the role of Human-Computer Interaction (HCI) in optimizing user ...
Large language models (LLMs) excel at clinical information extraction but their computational demands limit practical deployment. Knowledge distilla...
For efficient human-agent interaction, an agent should proactively recognize their target user and prepare for upcoming interactions. We formulate t...
While cancer has traditionally been considered a genetic disease, mounting evidence indicates an important role for non-genetic (epigenetic) mechani...
The intuitive nature of drag-based interaction has led to its growing adoption for controlling object trajectories in image-to-video synthesis. Stil...
Many inverse problems are ill-posed and need to be complemented by prior information that restricts the class of admissible models. Bayesian approac...
This study presents a novel approach for intelligent user interaction interface generation and optimization, grounded in the variational autoencoder...
Considering the difficulty of interpreting generative model output, there is significant current research focused on determining meaningful evaluati...
Anatomical abnormality detection and report generation of chest X-ray (CXR) are two essential tasks in clinical practice. The former aims at localiz...
There is a growing trend toward AI systems interacting with humans to revolutionize a range of application domains such as healthcare and transporta...
The emergence of diffusion models has significantly advanced image synthesis. The recent studies of model interaction and self-corrective reasoning ...
Generic sentences express generalisations about the world without explicit quantification. Although generics are central to everyday communication, ...
The drug development process is a critical challenge in the pharmaceutical industry due to its time-consuming nature and the need to discover new dr...
Section identification is an important task for library science, especially knowledge management. Identifying the sections of a paper would help fil...
Analysts in Security Operations Centers (SOCs) are often occupied with time-consuming investigations of alerts from Network Intrusion Detection Syst...
To break through the limitations of pre-training models on fixed categories, Open-Set Object Detection (OSOD) and Open-Set Segmentation (OSS) have a...
Image editing has advanced significantly with the development of diffusion models using both inversion-based and instruction-based methods. However,...