Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 6281-6300 of 9,853 articles

[Artificial intelligence in assessment of individual risks of age-related macular degeneration progression].

Age-related macular degeneration (AMD) is a progressive degenerative retinal disease and a leading cause of blindness in older adults worldwide. According to numerous studies, the number of affected individuals reached 196 million in 2020, with projections estimating an increase to 288 million by 2040, including 18.6 million cases of advanced AMD. The advent of optical coherence tomography (OCT) h...

Jan 1 2025 40353550

Probing Visual Language Priors in VLMs

Despite recent advances in Vision-Language Models (VLMs), they may over-rely on visual language priors existing in their training data rather than true visual reasoning. To investigate this, we introduce ViLP, a benchmark featuring deliberately out-of-distribution images synthesized via image generation models and out-of-distribution Q&A pairs. Each question in ViLP is coupled with three potenti...

Dual Diffusion for Unified Image Generation and Understanding

Diffusion models have gained tremendous success in text-to-image generation, yet still lag behind with visual understanding tasks, an area dominated...

Readability and Appropriateness of Responses Generated by ChatGPT 3.5, ChatGPT 4.0, Gemini, and Microsoft Copilot for FAQs in Refractive Surgery.

OBJECTIVES: To assess the appropriateness and readability of large language model (LLM) chatbots' answers to frequently asked questions about refracti...

Dec 31 2024 39743925
Real-Time Computational Visual Aberration Correcting Display Through High-Contrast Inverse Blurring

This paper presents a framework for developing a live vision-correcting display (VCD) to address refractive visual aberrations without the need for ...

Minimalist Vision with Freeform Pixels

A minimalist vision system uses the smallest number of pixels needed to solve a vision task. While traditional cameras use a large grid of square pi...

ReFlow6D: Refraction-Guided Transparent Object 6D Pose Estimation via Intermediate Representation Learning

Transparent objects are ubiquitous in daily life, making their perception and robotics manipulation important. However, they present a major challen...

Are Vision-Language Models Truly Understanding Multi-vision Sensor?

Large-scale Vision-Language Models (VLMs) have advanced by aligning vision inputs with text, significantly improving performance in computer vision ...

UniRS: Unifying Multi-temporal Remote Sensing Tasks through Vision Language Models

The domain gap between remote sensing imagery and natural images has recently received widespread attention and Vision-Language Models (VLMs) have d...

Slow Perception: Let's Perceive Geometric Figures Step-by-step

Recently, "visual o1" began to enter people's vision, with expectations that this slow-thinking design can solve visual reasoning tasks, especially ...

Polarimetric BSSRDF Acquisition of Dynamic Faces

Acquisition and modeling of polarized light reflection and scattering help reveal the shape, structure, and physical characteristics of an object, w...

Towards a Systematic Evaluation of Hallucinations in Large-Vision Language Models

Large Vision-Language Models (LVLMs) have demonstrated remarkable performance in complex multimodal tasks. However, these models still suffer from h...

Enhancing autonomous vehicle safety in rain: a data-centric approach for clear vision

Autonomous vehicles face significant challenges in navigating adverse weather, particularly rain, due to the visual impairment of camera-based syste...

PTQ4VM: Post-Training Quantization for Visual Mamba

Visual Mamba is an approach that extends the selective space state model, Mamba, to vision tasks. It processes image tokens sequentially in a fixed ...

Impact of Data Distribution on Fairness Guarantees in Equitable Deep Learning

We present a comprehensive theoretical framework analyzing the relationship between data distributions and fairness guarantees in equitable deep lea...

Machine Learning-Enabled Multidimensional Data Utilization Through Multi-Resonance Architecture: A Pathway to Enhanced Accuracy in Biosensing

A novel framework is proposed that combines multi-resonance biosensors with machine learning (ML) to significantly enhance the accuracy of parameter...

Multi-Modality Driven LoRA for Adverse Condition Depth Estimation

The autonomous driving community is increasingly focused on addressing corner case problems, particularly those related to ensuring driving safety u...

ErgoChat: a Visual Query System for the Ergonomic Risk Assessment of Construction Workers

In the construction sector, workers often endure prolonged periods of high-intensity physical work and prolonged use of tools, resulting in injuries...

Toward Modality Gap: Vision Prototype Learning for Weakly-supervised Semantic Segmentation with CLIP

The application of Contrastive Language-Image Pre-training (CLIP) in Weakly Supervised Semantic Segmentation (WSSS) research powerful cross-modal se...

Enhancing Vision-Language Tracking by Effectively Converting Textual Cues into Visual Cues

Vision-Language Tracking (VLT) aims to localize a target in video sequences using a visual template and language description. While textual cues enh...

Browse Categories