Psychiatry

Schizophrenia

Latest AI and machine learning research in schizophrenia for healthcare professionals.

3,500 articles
Stay Ahead - Weekly Schizophrenia research updates
Subscribe
Browse Categories
Showing 821-840 of 3,500 articles

Improving clinical interpretability of linear neuroimaging models through feature whitening

Linear models are widely used in computational neuroimaging to identify biomarkers associated with brain pathologies. However, interpreting the learned weights remains challenging, as they do not always yield clinically meaningful insights. This difficulty arises in part from the inherent correlation between brain regions, which causes linear weights to reflect shared rather than region-specific c...

Apr 22 2026 2604.20675v1

R-CoV: Region-Aware Chain-of-Verification for Alleviating Object Hallucinations in LVLMs

Large vision-language models (LVLMs) have demonstrated impressive performance in various multimodal understanding and reasoning tasks. However, they still struggle with object hallucinations, i.e., the claim of nonexistent objects in the visual input. To address this challenge, we propose Region-aware Chain-of-Verification (R-CoV), a visual chain-of-verification method to alleviate object hallucin...

Apr 22 2026 2604.20696v1
Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps

Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on gold-standard out...

Apr 21 2026 2604.19565v1
Lucky High Dynamic Range Smartphone Imaging

While the human eye can perceive an impressive twenty stops of dynamic range, smartphone camera sensors remain limited to about twelve stops despite d...

Apr 21 2026 2604.19976v1
Optimally Bridging Semantics and Data: Generative Semantic Communication via Schrödinger Bridge

Generative Semantic Communication (GSC) is a promising solution for image transmission over narrow-band and high-noise channels. However, existing GSC...

Apr 20 2026 2604.17802v1
Region-Grounded Report Generation for 3D Medical Imaging: A Fine-Grained Dataset and Graph-Enhanced Framework

Automated medical report generation for 3D PET/CT imaging is fundamentally challenged by the high-dimensional nature of volumetric data and a critical...

Apr 20 2026 2604.18145v1
When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models

Recent advances in Vision-Language Models (VLMs) have substantially enhanced their ability across multimodal video understanding benchmarks spanning t...

Apr 19 2026 2604.17375v1
Aligning What Vision-Language Models See and Perceive with Adaptive Information Flow

Vision-Language Models (VLMs) have demonstrated strong capability in a wide range of tasks such as visual recognition, document parsing, and visual gr...

Apr 17 2026 2604.15809v1
Aakhyan: An AI-Powered Vernacular Patient Communication Platform for Oncology in Resource-Limited Settings - System Architecture and Pilot Randomised Trial Protocol

Inadequate discharge communication is a well-documented contributor to medication non-adherence, missed follow-ups, and preventable readmissions acros...

MetaMuse: A Multi-Agent AI System for Biomedical Metadata Curation and Harmonization

Inconsistent and unstructured metadata in public biomedical repositories, such as the Gene Expression Omnibus (GEO), severely limits data discoverabil...

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Lar...

Apr 14 2026 2604.13328v1
Domain-Specific Latent Representations Improve the Fidelity of Diffusion-Based Medical Image Super-Resolution

Latent diffusion models for medical image super-resolution universally inherit variational autoencoders designed for natural photographs. We show that...

Apr 14 2026 2604.12152v1
Decoding by Perturbation: Mitigating MLLM Hallucinations via Dynamic Textual Perturbation

Multimodal Large Language Models frequently suffer from inference hallucinations, partially stemming from language priors dominating visual evidence. ...

Apr 14 2026 2604.12424v1
T2I-BiasBench: A Multi-Metric Framework for Auditing Demographic and Cultural Bias in Text-to-Image Models

Text-to-image (T2I) generative models achieve impressive visual fidelity but inherit and amplify demographic imbalances and cultural biases embedded i...

Apr 14 2026 2604.12481v1
SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis

Large Language Models (LLMs) and Vision-Language Models (VLMs) increasingly generate indoor scenes through intermediate structures such as layouts and...

Apr 14 2026 2604.13035v1
LogitDynamics: Reliable ViT Error Detection from Layerwise Logit Trajectories

Reliable confidence estimation is critical when deploying vision models. We study error prediction: determining whether an image classifier's output i...

Apr 12 2026 2604.10643v1
MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-...

Apr 9 2026 2604.08203v1
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions

Accurately detecting and localizing hallucinations is a critical task for ensuring high reliability of image captions. In the era of Multimodal Large ...

Apr 7 2026 2604.05623v1
Leveraging Image Editing Foundation Models for Data-Efficient CT Metal Artifact Reduction

Metal artifacts from high-attenuation implants severely degrade CT image quality, obscuring critical anatomical structures and posing a challenge for ...

Apr 7 2026 2604.05934v1
HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models

Large vision-language models can produce object hallucinations in image descriptions, highlighting the need for effective detection and mitigation str...

Apr 7 2026 2604.06165v1
Browse Categories