Psychiatry

Schizophrenia

Latest AI and machine learning research in schizophrenia for healthcare professionals.

3,224 articles
Stay Ahead - Weekly Schizophrenia research updates
Subscribe
Browse Categories
Showing 721-740 of 3,224 articles

Optimally Bridging Semantics and Data: Generative Semantic Communication via Schrödinger Bridge

Generative Semantic Communication (GSC) is a promising solution for image transmission over narrow-band and high-noise channels. However, existing GSC methods rely on long, indirect transport trajectories from a Gaussian to an image distribution guided by semantics, causing severe hallucination and high computational cost. To address this, we propose a general framework named Schrödinger Bridge-ba...

Apr 20 2026 2604.17802v1

Region-Grounded Report Generation for 3D Medical Imaging: A Fine-Grained Dataset and Graph-Enhanced Framework

Automated medical report generation for 3D PET/CT imaging is fundamentally challenged by the high-dimensional nature of volumetric data and a critical scarcity of annotated datasets, particularly for low-resource languages. Current black-box methods map whole volumes to reports, ignoring the clinical workflow of analyzing localized Regions of Interest (RoIs) to derive diagnostic conclusions. In th...

Apr 20 2026 2604.18145v1
When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models

Recent advances in Vision-Language Models (VLMs) have substantially enhanced their ability across multimodal video understanding benchmarks spanning t...

Apr 19 2026 2604.17375v1
Aligning What Vision-Language Models See and Perceive with Adaptive Information Flow

Vision-Language Models (VLMs) have demonstrated strong capability in a wide range of tasks such as visual recognition, document parsing, and visual gr...

Apr 17 2026 2604.15809v1
Aakhyan: An AI-Powered Vernacular Patient Communication Platform for Oncology in Resource-Limited Settings - System Architecture and Pilot Randomised Trial Protocol

Inadequate discharge communication is a well-documented contributor to medication non-adherence, missed follow-ups, and preventable readmissions acros...

MetaMuse: A Multi-Agent AI System for Biomedical Metadata Curation and Harmonization

Inconsistent and unstructured metadata in public biomedical repositories, such as the Gene Expression Omnibus (GEO), severely limits data discoverabil...

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Lar...

Apr 14 2026 2604.13328v1
Domain-Specific Latent Representations Improve the Fidelity of Diffusion-Based Medical Image Super-Resolution

Latent diffusion models for medical image super-resolution universally inherit variational autoencoders designed for natural photographs. We show that...

Apr 14 2026 2604.12152v1
Decoding by Perturbation: Mitigating MLLM Hallucinations via Dynamic Textual Perturbation

Multimodal Large Language Models frequently suffer from inference hallucinations, partially stemming from language priors dominating visual evidence. ...

Apr 14 2026 2604.12424v1
T2I-BiasBench: A Multi-Metric Framework for Auditing Demographic and Cultural Bias in Text-to-Image Models

Text-to-image (T2I) generative models achieve impressive visual fidelity but inherit and amplify demographic imbalances and cultural biases embedded i...

Apr 14 2026 2604.12481v1
SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis

Large Language Models (LLMs) and Vision-Language Models (VLMs) increasingly generate indoor scenes through intermediate structures such as layouts and...

Apr 14 2026 2604.13035v1
LogitDynamics: Reliable ViT Error Detection from Layerwise Logit Trajectories

Reliable confidence estimation is critical when deploying vision models. We study error prediction: determining whether an image classifier's output i...

Apr 12 2026 2604.10643v1
MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-...

Apr 9 2026 2604.08203v1
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions

Accurately detecting and localizing hallucinations is a critical task for ensuring high reliability of image captions. In the era of Multimodal Large ...

Apr 7 2026 2604.05623v1
Leveraging Image Editing Foundation Models for Data-Efficient CT Metal Artifact Reduction

Metal artifacts from high-attenuation implants severely degrade CT image quality, obscuring critical anatomical structures and posing a challenge for ...

Apr 7 2026 2604.05934v1
HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models

Large vision-language models can produce object hallucinations in image descriptions, highlighting the need for effective detection and mitigation str...

Apr 7 2026 2604.06165v1
EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors

Vision-Language Models (VLMs) excel at multimodal tasks, but they remain vulnerable to hallucinations that are factually incorrect or ungrounded in th...

Apr 3 2026 2604.02784v1
Overconfidence and Calibration in Medical VQA: Empirical Findings and Hallucination-Aware Mitigation

As vision-language models (VLMs) are increasingly deployed in clinical decision support, more than accuracy is required: knowing when to trust their p...

Apr 2 2026 2604.02543v1
How and why does deep ensemble coupled with transfer learning increase performance in bipolar disorder and schizophrenia classification?

Transfer learning (TL) and deep ensemble learning (DE) have recently been shown to outperform simple machine learning in classifying psychiatric disor...

Apr 2 2026 2604.02002v1
Evaluating the Large Language Model-Based Quality Assurance Tool for Auto-Contouring

Purpose: Manual verification of AI-based auto-contouring is labor-intensive and prone to fatigue-related errors. This study developed the large langua...

Browse Categories