Psychiatry

Schizophrenia

Latest AI and machine learning research in schizophrenia for healthcare professionals.

3,500 articles
Stay Ahead - Weekly Schizophrenia research updates
Subscribe
Browse Categories
Showing 881-900 of 3,500 articles

Dehallu3D: Hallucination-Mitigated 3D Generation from Single Image via Cyclic View Consistency Refinement

Large 3D reconstruction models have revolutionized the 3D content generation field, enabling broad applications in virtual reality and gaming. Just like other large models, large 3D reconstruction models suffer from hallucinations as well, introducing structural outliers (e.g., odd holes or protrusions) that deviate from the input data. However, unlike other large models, hallucinations in large 3...

Mar 2 2026 2603.01601v1

CARE: Towards Clinical Accountability in Multi-Modal Medical Reasoning with an Evidence-Grounded Agentic Framework

Large visual language models (VLMs) have shown strong multi-modal medical reasoning ability, but most operate as end-to-end black boxes, diverging from clinicians' evidence-based, staged workflows and hindering clinical accountability. Complementarily, expert visual grounding models can accurately localize regions of interest (ROIs), providing explicit, reliable evidence that improves both reasoni...

Mar 2 2026 2603.01607v1
FireRed-OCR Technical Report

We present FireRed-OCR, a systematic framework to specialize general VLMs into high-performance OCR models. Large Vision-Language Models (VLMs) have d...

Mar 2 2026 2603.01840v1
Semantic Similarity is a Spurious Measure of Comic Understanding: Lessons Learned from Hallucinations in a Benchmarking Experiment

A system that enables blind or visually impaired users to access comics/manga would introduce a new medium of storytelling to this community. However,...

Mar 2 2026 2603.01950v1
3D Field of Junctions: A Noise-Robust, Training-Free Structural Prior for Volumetric Inverse Problems

Volume denoising is a foundational problem in computational imaging, as many 3D imaging inverse problems face high levels of measurement noise. Inspir...

Mar 2 2026 2603.02149v1
AgilePruner: An Empirical Study of Attention and Diversity for Adaptive Visual Token Pruning in Large Vision-Language Models

Large Vision-Language Models (LVLMs) have adopted visual token pruning strategies to mitigate substantial computational overhead incurred by extensive...

Mar 1 2026 2603.01236v1
Suppressing Prior-Comparison Hallucinations in Radiology Report Generation via Semantically Decoupled Latent Steering

Automated radiology report generation using vision-language models (VLMs) is limited by the risk of prior-comparison hallucination, where the model ge...

Feb 27 2026 2602.23676v1
Toward Guarantees for Clinical Reasoning in Vision Language Models via Formal Verification

Vision-language models (VLMs) show promise in drafting radiology reports, yet they frequently suffer from logical inconsistencies, generating diagnost...

Feb 27 2026 2602.24111v1
Onco-Shikshak: An AI-Native Adaptive Learning Ecosystem for Medical Oncology Education

Medical oncology education faces a dual crisis: knowledge velocity that outpaces static curricula and large language model (LLM) risks hallucination a...

Disentangling Symptom Heterogeneity in Large-Scale Psychiatric Text: Domain-Adapted vs. Instruction-Tuned Transformers

Psychiatric disorders are fundamentally challenged by symptom heterogeneity, high comorbidity, and the absence of objective biomarkers, which together...

Plug-and-Play Diffusion Meets ADMM: Dual-Variable Coupling for Robust Medical Image Reconstruction

Plug-and-Play diffusion prior (PnPDP) frameworks have emerged as a powerful paradigm for solving imaging inverse problems by treating pretrained gener...

Feb 26 2026 2602.23214v1
Beyond Dominant Patches: Spatial Credit Redistribution For Grounded Vision-Language Models

Vision-language models (VLMs) frequently hallucinate objects absent from the input image. We trace this failure to spatial credit collapse: activation...

Feb 25 2026 2602.22469v1
Patient-centric radiology: Utilising large language models (LLMs) to improve patient communication and education

Purpose: To evaluate whether large language models (LLMs) can enhance clinician-patient communication by simplifying radiology reports to improve pati...

See It, Say It, Sorted: An Iterative Training-Free Framework for Visually-Grounded Multimodal Reasoning in LVLMs

Recent large vision-language models (LVLMs) have demonstrated impressive reasoning ability by generating long chain-of-thought (CoT) responses. Howeve...

Feb 25 2026 2602.21497v1
NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language Priors

Object hallucination is a critical issue in Large Vision-Language Models (LVLMs), where outputs include objects that do not appear in the input image....

Feb 25 2026 2602.22144v1
Causal Decoding for Hallucination-Resistant Multimodal Large Language Models

Multimodal Large Language Models (MLLMs) deliver detailed responses on vision-language tasks, yet remain susceptible to object hallucination (introduc...

Feb 24 2026 2602.21441v1
Continual-NExT: A Unified Comprehension And Generation Continual Learning Framework

Dual-to-Dual MLLMs refer to Multimodal Large Language Models, which can enable unified multimodal comprehension and generation through text and image ...

Feb 20 2026 2602.18055v1
LGD-Net: Latent-Guided Dual-Stream Network for HER2 Scoring with Task-Specific Domain Knowledge

It is a critical task to evalaute HER2 expression level accurately for breast cancer evaluation and targeted treatment therapy selection. However, the...

Feb 19 2026 2602.17793v1
A Large-Scale Computer-Vision Mapping of the Geometric Structures of Stroboscopically-Induced Visual Hallucinations

Visual hallucinations (VHs) occur across psychedelic states and diverse psychiatric and neurological conditions, yet their phenomenology remains diffi...

Bridging Day and Night: Target-Class Hallucination Suppression in Unpaired Image Translation

Day-to-night unpaired image translation is important to downstream tasks but remains challenging due to large appearance shifts and the lack of direct...

Feb 17 2026 2602.15383v1
Browse Categories