Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 5881-5900 of 9,853 articles

Vision-Enhanced Time Series Forecasting via Latent Diffusion Models

Diffusion models have recently emerged as powerful frameworks for generating high-quality images. While recent studies have explored their application to time series forecasting, these approaches face significant challenges in cross-modal modeling and transforming visual information effectively to capture temporal patterns. In this paper, we propose LDM4TS, a novel framework that leverages the p...

MC-BEVRO: Multi-Camera Bird Eye View Road Occupancy Detection for Traffic Monitoring

Single camera 3D perception for traffic monitoring faces significant challenges due to occlusion and limited field of view. Moreover, fusing information from multiple cameras at the image feature level is difficult because of different view angles. Further, the necessity for practical implementation and compatibility with existing traffic infrastructure compounds these challenges. To address the...

AnyRefill: A Unified, Data-Efficient Framework for Left-Prompt-Guided Vision Tasks

In this paper, we present a novel Left-Prompt-Guided (LPG) paradigm to address a diverse range of reference-based vision tasks. Inspired by the huma...

Knowledge Graph-Driven Retrieval-Augmented Generation: Integrating Deepseek-R1 with Weaviate for Advanced Chatbot Applications

Large language models (LLMs) have significantly advanced the field of natural language generation. However, they frequently generate unverified outp...

Can LVLMs and Automatic Metrics Capture Underlying Preferences of Blind and Low-Vision Individuals for Navigational Aid?

Vision is a primary means of how humans perceive the environment, but Blind and Low-Vision (BLV) people need assistance understanding their surround...

MET-Bench: Multimodal Entity Tracking for Evaluating the Limitations of Vision-Language and Reasoning Models

Entity tracking is a fundamental challenge in natural language understanding, requiring models to maintain coherent representations of entities. Pre...

Disentangle Nighttime Lens Flares: Self-supervised Generation-based Lens Flare Removal

Lens flares arise from light reflection and refraction within sensor arrays, whose diverse types include glow, veiling glare, reflective flare and s...

ProMRVL-CAD: Proactive Dialogue System with Multi-Round Vision-Language Interactions for Computer-Aided Diagnosis

Recent advancements in large language models (LLMs) have demonstrated extraordinary comprehension capabilities with remarkable breakthroughs on vari...

Ocular Disease Classification Using CNN with Deep Convolutional Generative Adversarial Network

The Convolutional Neural Network (CNN) has shown impressive performance in image classification because of its strong learning capabilities. However...

A Roadmap to Address Burnout in the Cybersecurity Profession: Outcomes from a Multifaceted Workshop

This paper addresses the critical issue of burnout among cybersecurity professionals, a growing concern that threatens the effectiveness of digital ...

VisCon-100K: Leveraging Contextual Web Data for Fine-tuning Vision Language Models

Vision-language models (VLMs) excel in various visual benchmarks but are often constrained by the lack of high-quality visual fine-tuning data. To a...

Compress image to patches for Vision Transformer

The Vision Transformer (ViT) has made significant strides in the field of computer vision. However, as the depth of the model and the resolution of ...

A novel approach to data generation in generative model

Variational Autoencoders (VAEs) and other generative models are widely employed in artificial intelligence to synthesize new data. However, current ...

Granite Vision: a lightweight, open-source multimodal model for enterprise Intelligence

We introduce Granite Vision, a lightweight large language model with vision capabilities, specifically designed to excel in enterprise use cases, pa...

Insect-Foundation: A Foundation Model and Large Multimodal Dataset for Vision-Language Insect Understanding

Multimodal conversational generative AI has shown impressive capabilities in various vision and language understanding through learning massive text...

HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation

We present HealthGPT, a powerful Medical Large Vision-Language Model (Med-LVLM) that integrates medical visual comprehension and generation capabili...

[Vigorously advancing the application of AI in the diagnosis and treatment of ocular surface and tear diseases].

Ocular surface and tear diseases are among the most common and significant ocular conditions affecting eye health. In recent years, research and clini...

Feb 14 2025 39939001
[Advancements of artificial intelligence in dry eye].

With the continuous evolution of computer technology and the surging advent of the big data era, artificial intelligence (AI) has already manifested e...

Feb 14 2025 39939010
Development of a pressure ulcer stage determination system for community healthcare providers using a vision transformer deep learning model.

This study reports the first steps toward establishing a computer vision system to help caregivers of bedridden patients detect pressure ulcers (PUs) ...

Feb 14 2025 39960905
GAIA: A Global, Multi-modal, Multi-scale Vision-Language Dataset for Remote Sensing Image Analysis

The continuous operation of Earth-orbiting satellites generates vast and ever-growing archives of Remote Sensing (RS) images. Natural language prese...

Browse Categories