Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 4361-4380 of 9,853 articles

GazeVaLM: A Multi-Observer Eye-Tracking Benchmark for Evaluating Clinical Realism in AI-Generated X-Rays

We introduce GazeVaLM, a public eye-tracking dataset for studying clinical perception during chest radiograph authenticity assessment. The dataset comprises 960 gaze recordings from 16 expert radiologists interpreting 30 real and 30 synthetic chest X-rays (generated by diffusion based generative AI) under two conditions: diagnostic assessment and real-fake classification (Visual Turing test). For ...

Apr 13 2026 2604.11653v1

LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

While the shortage of explicit action data limits Vision-Language-Action (VLA) models, human action videos offer a scalable yet unlabeled data source. A critical challenge in utilizing large-scale human video datasets lies in transforming visual signals into ontology-independent representations, known as latent actions. However, the capacity of latent action representation to derive robust control...

Apr 13 2026 2604.11689v1
fMRI-Based Prediction of Eye Gaze During Naturalistic Movie Viewing Reveals Eye-Movement-Related Brain Activity

Background: Eye gaze provides crucial insights into perceptual and cognitive processes during naturalistic movie viewing, yet concurrent eye tracking ...

Data-Efficient Surgical Phase Segmentation in Small-Incision Cataract Surgery: A Controlled Study of Vision Foundation Models

Surgical phase segmentation is central to computer-assisted surgery, yet robust models remain difficult to develop when labeled surgical videos are sc...

Apr 12 2026 2604.10514v1
Retinal Cyst Detection from Optical Coherence Tomography Images

Retinal Cysts are formed by leakage and accumulation of fluid in the retina due to the incompetence of retinal vasculature. These cystic spaces have s...

Apr 12 2026 2604.10843v1
Nested Radially Monotone Polar Occupancy Estimation: Clinically-Grounded Optic Disc and Cup Segmentation for Glaucoma Screening

Valid segmentation of the optic disc (OD) and optic cup (OC) from fundus photographs is essential for glaucoma screening. Unfortunately, existing deep...

Apr 10 2026 2604.09062v1
Off-the-shelf Vision Models Benefit Image Manipulation Localization

Image manipulation localization (IML) and general vision tasks are typically treated as two separate research directions due to the fundamental differ...

Apr 10 2026 2604.09096v1
FIRE-CIR: Fine-grained Reasoning for Composed Fashion Image Retrieval

Composed image retrieval (CIR) aims to retrieve a target image that depicts a reference image modified by a textual description. While recent vision-l...

Apr 10 2026 2604.09114v1
Arbitration Failure, Not Perceptual Blindness: How Vision-Language Models Resolve Visual-Linguistic Conflicts

When a Vision-Language Model (VLM) sees a blue banana and answers "yellow", is the problem of perception or arbitration? We explore the question in te...

Apr 10 2026 2604.09364v1
Do Vision Language Models Need to Process Image Tokens?

Vision Language Models (VLMs) have achieved remarkable success by integrating visual encoders with large language models (LLMs). While VLMs process de...

Apr 10 2026 2604.09425v1
VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu...

Apr 10 2026 2604.09508v1
VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning

Large Vision Language Models (LVLMs) achieve strong multimodal reasoning but frequently exhibit hallucinations and incorrect responses with high certa...

Apr 10 2026 2604.09529v1
VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images

Vision-language models (VLMs) still struggle with visual perception tasks such as spatial understanding and viewpoint recognition. One plausible contr...

Apr 10 2026 2604.09531v1
Seeing is Believing: Robust Vision-Guided Cross-Modal Prompt Learning under Label Noise

Prompt learning is a parameter-efficient approach for vision-language models, yet its robustness under label noise is less investigated. Visual conten...

Apr 10 2026 2604.09532v1
T-Gated Adapter: A Lightweight Temporal Adapter for Vision-Language Medical Segmentation

Medical image segmentation traditionally relies on fully supervised 3D architectures that demand a large amount of dense, voxel-level annotations from...

Apr 9 2026 2604.08167v1
MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-...

Apr 9 2026 2604.08203v1
Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment

General-purpose vision-language models demonstrate strong performance in everyday domains but struggle with specialized technical fields requiring pre...

Apr 9 2026 2604.08212v1
Fundus-R1: Training a Fundus-Reading MLLM with Knowledge-Aware Reasoning on Public Data

Fundus imaging such as CFP, OCT and UWF is crucial for the early detection of retinal anomalies and diseases. Fundus image understanding, due to its k...

Apr 9 2026 2604.08322v1
What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric

Scanpath similarity metrics are central to eye-movement research, yet existing methods predominantly evaluate spatial and temporal alignment while neg...

Apr 9 2026 2604.08494v1
Meta-learning In-Context Enables Training-Free Cross Subject Brain Decoding

Visual decoding from brain signals is a key challenge at the intersection of computer vision and neuroscience, requiring methods that bridge neural re...

Apr 9 2026 2604.08537v1
Browse Categories