Psychiatry

Schizophrenia

Latest AI and machine learning research in schizophrenia for healthcare professionals.

3,231 articles
Stay Ahead - Weekly Schizophrenia research updates
Subscribe
Browse Categories
Showing 1001-1020 of 3,231 articles

SHAPE : Self-Improved Visual Preference Alignment by Iteratively Generating Holistic Winner

Large Visual Language Models (LVLMs) increasingly rely on preference alignment to ensure reliability, which steers the model behavior via preference fine-tuning on preference data structured as ``image - winner text - loser text'' triplets. However, existing approaches often suffer from limited diversity and high costs associated with human-annotated preference data, hindering LVLMs from fully a...

TPC: Cross-Temporal Prediction Connection for Vision-Language Model Hallucination Reduction

Vision-language models (VLMs) have achieved remarkable advancements, capitalizing on the impressive capabilities of large language models (LLMs) across diverse tasks. Despite this, a critical challenge known as hallucination occurs when models overconfidently describe objects or attributes absent from the image, a problem exacerbated by the tendency of VLMs to rely on linguistic priors. This lim...

Towards Understanding Text Hallucination of Diffusion Models via Local Generation Bias

Score-based diffusion models have achieved incredible performance in generating realistic images, audio, and video data. While these models produce ...

See What You Are Told: Visual Attention Sink in Large Multimodal Models

Large multimodal models (LMMs) "see" images by leveraging the attention mechanism between text and visual tokens in the transformer decoder. Ideally...

WMNav: Integrating Vision-Language Models into World Models for Object Goal Navigation

Object Goal Navigation-requiring an agent to locate a specific object in an unseen environment-remains a core challenge in embodied AI. Although rec...

MedHEval: Benchmarking Hallucinations and Mitigation Strategies in Medical Large Vision-Language Models

Large Vision Language Models (LVLMs) are becoming increasingly important in the medical domain, yet Medical LVLMs (Med-LVLMs) frequently generate ha...

The order in speech disorder: a scoping review of state of the art machine learning methods for clinical speech classification

Background:Speech patterns have emerged as potential diagnostic markers for conditions with varying etiologies. Machine learning (ML) presents an op...

Explainable Depression Detection in Clinical Interviews with Personalized Retrieval-Augmented Generation

Depression is a widespread mental health disorder, and clinical interviews are the gold standard for assessment. However, their reliance on scarce p...

Tackling Hallucination from Conditional Models for Medical Image Reconstruction with DynamicDPS

Hallucinations are spurious structures not present in the ground truth, posing a critical challenge in medical image reconstruction, especially for ...

HalCECE: A Framework for Explainable Hallucination Detection through Conceptual Counterfactuals in Image Captioning

In the dynamic landscape of artificial intelligence, the exploration of hallucinations within vision-language (VL) models emerges as a critical fron...

Hybrid Retrieval for Hallucination Mitigation in Large Language Models: A Comparative Analysis

Large Language Models (LLMs) excel in language comprehension and generation but are prone to hallucinations, producing factually incorrect or unsupp...

MedHallTune: An Instruction-Tuning Benchmark for Mitigating Medical Hallucination in Vision-Language Models

The increasing use of vision-language models (VLMs) in healthcare applications presents great challenges related to hallucinations, in which the mod...

Mitigating Hallucinations in Large Vision-Language Models by Adaptively Constraining Information Flow

Large vision-language models show tremendous potential in understanding visual information through human languages. However, they are prone to suffe...

Towards Statistical Factuality Guarantee for Large Vision-Language Models

Advancements in Large Vision-Language Models (LVLMs) have demonstrated promising performance in a variety of vision-language tasks involving image-c...

One-for-More: Continual Diffusion Model for Anomaly Detection

With the rise of generative models, there is a growing interest in unifying all tasks within a generative framework. Anomaly detection methods also ...

ProAPO: Progressively Automatic Prompt Optimization for Visual Classification

Vision-language models (VLMs) have made significant progress in image classification by training with large-scale paired image-text data. Their perf...

Human-Centered AI in Multidisciplinary Medical Discussions: Evaluating the Feasibility of a Chat-Based Approach to Case Assessment

In this study, we investigate the feasibility of using a human-centered artificial intelligence (AI) chat platform where medical specialists collabo...

Medical Hallucinations in Foundation Models and Their Impact on Healthcare

Foundation Models that are capable of processing and generating multi-modal data have transformed AI's role in medicine. However, a key limitation o...

On the Importance of Text Preprocessing for Multimodal Representation Learning and Pathology Report Generation

Vision-language models in pathology enable multimodal case retrieval and automated report generation. Many of the models developed so far, however, ...

Stealthy Backdoor Attack in Self-Supervised Learning Vision Encoders for Large Vision Language Models

Self-supervised learning (SSL) vision encoders learn high-quality image representations and thus have become a vital part of developing vision modal...

Browse Categories