Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 3761-3780 of 9,853 articles

Dynamic SpectraFormer for Ultra-High-Definition Underwater Image Enhancement

Underwater images suffer from color distortion, haze, and poor visibility due to light refraction and absorption in water. These challenges significantly impact the utilization of Autonomous Underwater Vehicles (AUVs) or marine robots. Typically, color and brightness distortions manifest at lower frequencies, while edge and texture distortions are prevalent at higher frequencies. Traditional metho...

Aug 19 2026 2608.18662v1

CL4D: Contrastive Language-4D Pretraining for Vision-Language Reasoning in Dynamic Scenes

4D understanding and reasoning is a fundamental capability for embodied AI agents operating in dynamic physical environments. However, existing vision encoders are largely limited to static 2D images or 3D point clouds without temporal modeling, or to 2D videos that lack accurate geometric depth reasoning. Consequently, current approaches fail to jointly capture spatial structure and motion evolut...

Aug 19 2026 2608.18734v1
Breaking the weakest link to evade vision language models

Vision Language Models (VLMs) have recently emerged as a critical component of multimodal AI systems, enabling joint reasoning over visual and textual...

Aug 19 2026 2608.18938v1
Uncertainty-Aware Art-Historical Dating with Vision-Language Models

Museum and archival datasets do not mirror historical artistic production, but materialize the contingent histories of collecting, preservation, catal...

Aug 19 2026 2608.18984v1
ReWEIGH the Evidence: Calibrating Token-Level Ordinal Visual Evidence to Mitigate Hallucinations in Large Vision-Language Models

Large vision-language models (LVLMs) often hallucinate, generating content that the input image does not support. Preventing such content during decod...

Aug 19 2026 2608.19075v1
3D Gaussian Accelerated Ray Tracing: Fast training through particle-based backward propagation

3D Gaussian Splatting has made Gaussian primitives a highly efficient representation for real-time novel view synthesis, but its rasterisation-based f...

Aug 18 2026 2608.17298v1
Primitive-Driven Compositional Forensic Visual Prompting for Open-World Face Anti-Spoofing

Open-world face anti-spoofing must address both covariate and semantic shifts: source and target domains differ in imaging conditions, while target do...

Aug 18 2026 2608.17351v1
MoE-ViE: Mixture of Experts Vision Encoder for Efficient Image and Video Understanding

Vision encoders are a critical component of vision-language models, and scaling their capacity effectively improves performance. However, dense scalin...

Aug 18 2026 2608.17402v1
RetiWave-Mamba: A Dual-Stream Network for Retinal Disease Detection based on Multi-scale Context and Frequency-Adaptive Mamba Projection

Retinal diseases are a leading cause of irreversible vision impairment, making early and accurate diagnosis essential for effective treatment. Optical...

Aug 18 2026 2608.17623v1
Evaluation of AI-based Visual Crack Detection in Steel Bridges Using Probability of Detection

Bridge structures are regularly inspected for structural damage such as cracks and corrosion in order to ensure public safety and reduce maintenance c...

Aug 18 2026 2608.17726v1
Optic Disc Segmentation in Fundus Images: From Classical Image Processing and Deformable Models to Modern AI

Accurate localization and segmentation of the optic disc (OD) are important for retinal image analysis and glaucoma assessment, yet remain challenging...

Aug 18 2026 2608.18367v1
AlignJEPA: Predictive Vision-Language Alignment for Remote Sensing Foundation Models

Remote sensing (RS) foundation models provide transferable Earth observation representations across sensors, resolutions, and geographies, yet most re...

Aug 16 2026 2608.15456v1
Population Structure Analysis of an Inbred Population using Quantitative Shape Phenotyping from Stereo Retinal Photographs

The population structure of an inbred population of 781 people on Norfolk Island in the Pacific, 318 of which are descendants of the original Mutineer...

Aug 16 2026 2608.15471v1
Attention Capture Is Not Detection: A Two-Stage Account of How Humans Miss Localized AI Image Edits

As AI-generated image edits proliferate, the platforms meant to curb the resulting disinformation treat detectability as a single, undifferentiated pr...

Aug 14 2026 2608.13865v1
CSG-Mamba: A Convolutional Scoring Gating Vision State Space Network for Endoscopic Polyp Segmentation

Accurate polyp segmentation is critical for computer-aided colonoscopy, yet endoscopic images often contain low-contrast boundaries, mucosal texture i...

Aug 14 2026 2608.14146v1
Seeing Red, Thinking Bad: Color Bias in Vision Language Models

Vision language models (VLMs) are increasingly used in industrial decision-making systems, such as recruitment support and recommendation. This motiva...

Aug 14 2026 2608.14286v1
TRIAGE: Risk-Controlled Pseudo-Label Admission for Annotation-Efficient Semi-Supervised Retinal OCT Classification

The advanced retinal disease diagnosing imaging modality, optical coherence tomography (OCT), encounters a lack of automation because of the high expe...

Aug 14 2026 2608.14321v1
A Vision-Language Model for Coronary Angiography Interpretation and Clinical Decision Support

BACKGROUND: Coronary angiography remains the reference standard for diagnosing coronary artery disease and guiding revascularization, yet its interpre...

Learning from human and chemical languages to predict biological function

Understanding how molecular structure encodes biological function remains a grand challenge in drug discovery. Here, we present PubCheF-1, a deep lear...

MedPlex: Deep Vision-Language Co-Adaptation for Clinically Grounded Medical Segmentation

Medical image segmentation is still largely treated as a vision-only problem, although clinical interpretation often relies on textual knowledge of an...

Aug 13 2026 2608.13690v1
Browse Categories