Latest AI and machine learning research in schizophrenia for healthcare professionals.
Radiological diagnosis is a perceptual process in which careful visual inspection and language reasoning are repeatedly interleaved. Most medical large vision language models (LVLMs) perform visual inspection only once and then rely on text-only chain-of-thought (CoT) reasoning, which operates purely in the linguistic space and is prone to hallucination. Recent methods attempt to mitigate this iss...
Neurological health score (NHS), indicating the health of brain and nervous system, helps in identifying high risk individuals, and in recommending lifestyle modifications. In the present study, we developed NHS based on genetic, lifestyle and biochemical variables associated with eight neurological disorders - dementia, stroke, Parkinsons disease, amyotrophic lateral sclerosis, schizophrenia, bip...
Large Vision-Language Models (VLMs) have achieved remarkable success across diverse multimodal tasks but remain vulnerable to hallucinations rooted in...
Multimodal large language models (MLLMs) are increasingly adopted in remote sensing (RS) and have shown strong performance on tasks such as RS visual ...
The growing volume of video-based news content has heightened the need for transparent and reliable methods to extract on-screen information. Yet the ...
The robustness of Vision Language Models (VLMs) is commonly assessed through output-level invariance, implicitly assuming that stable predictions refl...
Open-vocabulary semantic segmentation (OVSS) extends traditional closed-set segmentation by enabling pixel-wise annotation for both seen and unseen ca...
In pharmacovigilance, analyzing drug safety cases is often time consuming due to the abundance of laboratory data, complex medical histories, and intr...
Identifying robust neuroimaging markers associated with schizophrenia is essential for advancing research and informing clinical understanding. Howeve...
We introduce a framework that automates the transformation of static anime illustrations into manipulatable 2.5D models. Current professional workflow...
Large multimodal reasoning models solve challenging visual problems via explicit long-chain inference: they gather visual clues from images and decode...
Human social interactions rely on the ability to reflect on one's own and others' internal states and traits--a process known as mentalizing. Impaired...
Dopamine (DA) has been implicated in exploration-exploitation behaviour, i.e., exploring novel, potentiallybetter options vs. exploiting known, previo...
Despite progress in Large Vision Language Models (LVLMs), object hallucination remains a critical issue in image captioning task, where models generat...
Multimodal Large Language Models (MLLMs) have shown remarkable capability in assisting disease diagnosis in medical visual question answering (VQA). H...
The reliability of Large Language Models (LLMs) in high-stakes domains such as healthcare, law, and scientific discovery is often compromised by hallu...
Recent advances in fMRI-based image reconstruction have achieved remarkable photo-realistic fidelity. Yet, a persistent limitation remains: while reco...
Real-world text image super-resolution aims to restore overall visual quality and text legibility in images suffering from diverse degradations and te...
High-precision facial landmark detection (FLD) relies on high-resolution deep feature representations. However, low-resolution face images or the comp...
We introduce ELITE, an Efficient Gaussian head avatar synthesis from a monocular video via Learned Initialization and TEst-time generative adaptation....