Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 5181-5200 of 9,853 articles

Detailed Evaluation of Modern Machine Learning Approaches for Optic Plastics Sorting

According to the EPA, only 25% of waste is recycled, and just 60% of U.S. municipalities offer curbside recycling. Plastics fare worse, with a recycling rate of only 8%; an additional 16% is incinerated, while the remaining 76% ends up in landfills. The low plastic recycling rate stems from contamination, poor economic incentives, and technical difficulties, making efficient recycling a challeng...

AnchorFormer: Differentiable Anchor Attention for Efficient Vision Transformer

Recently, vision transformers (ViTs) have achieved excellent performance on vision tasks by measuring the global self-attention among the image patches. Given $n$ patches, they will have quadratic complexity such as $\mathcal{O}(n^2)$ and the time cost is high when splitting the input image with a small granularity. Meanwhile, the pivotal information is often randomly gathered in a few regions o...

Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression

Despite their remarkable progress in multimodal understanding tasks, large vision language models (LVLMs) often suffer from "hallucinations", genera...

Dynamic Caustics by Ultrasonically Modulated Liquid Surface

This paper presents a method for generating dynamic caustic patterns by utilising dual-optimised holographic fields with Phased Array Transducer (PA...

MAGE: A Multi-task Architecture for Gaze Estimation with an Efficient Calibration Module

Eye gaze can provide rich information on human psychological activities, and has garnered significant attention in the field of Human-Robot Interact...

FPQVAR: Floating Point Quantization for Visual Autoregressive Model with FPGA Hardware Co-design

Visual autoregressive (VAR) modeling has marked a paradigm shift in image generation from next-token prediction to next-scale prediction. VAR predic...

Deep learning-based automatic differentiation of acute angle closure with or without zonulopathy using ultrasound biomicroscopy: a comparison of diagnostic performance with ophthalmologists.

OBJECTIVE: This study aims to develop ultrasound biomicroscopy (UBM)-based artificial intelligence (AI) models for preoperative differentiation of acu...

May 22 2025 40409764
Self-supervised model-informed deep learning for low-SNR SS-OCT domain transformation.

This article introduces a novel deep-learning based framework, Super-resolution/Denoising network (SDNet), for simultaneous denoising and super-resolu...

May 22 2025 40404743
Pixels Versus Priors: Controlling Knowledge Priors in Vision-Language Models through Visual Counterfacts

Multimodal Large Language Models (MLLMs) perform well on tasks such as visual question answering, but it remains unclear whether their reasoning rel...

Learning Interpretable Representations Leads to Semantically Faithful EEG-to-Text Generation

Pretrained generative models have opened new frontiers in brain decoding by enabling the synthesis of realistic texts and images from non-invasive b...

A Paradigm for Creative Ownership

As generative AI tools become more integrated into creative workflows, questions of ownership in co-creative contexts have become increasingly urgen...

Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning

Chain-of-thought reasoning has significantly improved the performance of Large Language Models (LLMs) across various domains. However, this reasonin...

OViP: Online Vision-Language Preference Learning

Large vision-language models (LVLMs) remain vulnerable to hallucination, often generating content misaligned with visual inputs. While recent approa...

VideoGameQA-Bench: Evaluating Vision-Language Models for Video Game Quality Assurance

With video games now generating the highest revenues in the entertainment industry, optimizing game development workflows has become essential for t...

A Deep Learning Framework for Two-Dimensional, Multi-Frequency Propagation Factor Estimation

Accurately estimating the refractive environment over multiple frequencies within the marine atmospheric boundary layer is crucial for the effective...

Exploring The Visual Feature Space for Multimodal Neural Decoding

The intrication of brain signals drives research that leverages multimodal AI to align brain modalities with visual and textual data for explainable...

FragFake: A Dataset for Fine-Grained Detection of Edited Images with Vision Language Models

Fine-grained edited image detection of localized edits in images is crucial for assessing content authenticity, especially given that modern diffusi...

Bayesian Ensembling: Insights from Online Optimization and Empirical Bayes

We revisit the classical problem of Bayesian ensembles and address the challenge of learning optimal combinations of Bayesian models in an online, c...

SNAP: A Benchmark for Testing the Effects of Capture Conditions on Fundamental Vision Tasks

Generalization of deep-learning-based (DL) computer vision algorithms to various image perturbations is hard to establish and remains an active area...

LENS: Multi-level Evaluation of Multimodal Reasoning with Large Language Models

Multimodal Large Language Models (MLLMs) have achieved significant advances in integrating visual and linguistic information, yet their ability to r...

Browse Categories