Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 4601-4620 of 9,853 articles

Crop-OCT: a Fully Integrated Imageomics Pipeline to Identify Regional and Focal Retinopathy in Murine Models

Imageomics uses machine learning to accelerate our understanding of biological traits and human disease processes. Some of the earliest imageomics applications used deep learning to assess human diseases. For example, retinal fundus images were analyzed to diagnose diabetic retinopathy. The imaging modality optical coherence tomography (OCT) is widely used to diagnose and monitor the progression o...

ATA: Bridging Implicit Reasoning with Attention-Guided and Action-Guided Inference for Vision-Language Action Models

Vision-Language-Action (VLA) models rely on current observations, including images, language instructions, and robot states, to predict actions and complete tasks. While accurate visual perception is crucial for precise action prediction and execution, recent work has attempted to further improve performance by introducing explicit reasoning during inference. However, such approaches face signific...

Mar 2 2026 2603.01490v1
Rate-Distortion Signatures of Generalization and Information Trade-offs

Generalization to novel visual conditions remains a central challenge for both human and machine vision, yet standard robustness metrics offer limited...

Mar 2 2026 2603.01568v1
Shape-Interpretable Visual Self-Modeling Enables Geometry-Aware Continuum Robot Control

Continuum robots possess high flexibility and redundancy, making them well suited for safe interaction in complex environments, yet their continuous d...

Mar 2 2026 2603.01751v1
Learning to Read Where to Look: Disease-Aware Vision-Language Pretraining for 3D CT

Recent 3D CT vision-language models align volumes with reports via contrastive pretraining, but typically rely on limited public data and provide only...

Mar 2 2026 2603.02026v1
Leveraging Model Soups to Classify Intangible Cultural Heritage Images from the Mekong Delta

The classification of Intangible Cultural Heritage (ICH) images in the Mekong Delta poses unique challenges due to limited annotated data, high visual...

Mar 2 2026 2603.02181v1
Vision-Language Feature Alignment for Road Anomaly Segmentation

Safe autonomous systems in complex environments require robust road anomaly segmentation to identify unknown obstacles. However, existing approaches o...

Mar 1 2026 2603.01029v1
Unified Vision-Language Modeling via Concept Space Alignment

We introduce V-SONAR, a vision-language embedding space extended from the text-only embedding space SONAR (Omnilingual Embeddings Team et al., 2026), ...

Mar 1 2026 2603.01096v1
GroundedSurg: A Multi-Procedure Benchmark for Language-Conditioned Surgical Tool Segmentation

Clinically reliable perception of surgical scenes is essential for advancing intelligent, context-aware intraoperative assistance such as instrument h...

Mar 1 2026 2603.01108v1
GuiDINO: Rethinking Vision Foundation Model in Medical Image Segmentation

Foundation vision models are increasingly adopted in medical image analysis. Due to domain shift, these pretrained models misalign with medical image ...

Mar 1 2026 2603.01115v1
ClinCoT: Clinical-Aware Visual Chain-of-Thought for Medical Vision Language Models

Medical Vision-Language Models have shown promising potential in clinical decision support, yet they remain prone to factual hallucinations due to ins...

Mar 1 2026 2603.01124v1
AgilePruner: An Empirical Study of Attention and Diversity for Adaptive Visual Token Pruning in Large Vision-Language Models

Large Vision-Language Models (LVLMs) have adopted visual token pruning strategies to mitigate substantial computational overhead incurred by extensive...

Mar 1 2026 2603.01236v1
When Does RL Help Medical VLMs? Disentangling Vision, SFT, and RL Gains

Reinforcement learning (RL) is increasingly used to post-train medical Vision-Language Models (VLMs), yet it remains unclear whether RL improves medic...

Mar 1 2026 2603.01301v1
CycleBEV: Regularizing View Transformation Networks via View Cycle Consistency for Bird's-Eye-View Semantic Segmentation

Transforming image features from perspective view (PV) space to bird's-eye-view (BEV) space remains challenging in autonomous driving due to depth amb...

Feb 27 2026 2602.23575v1
Pseudo Contrastive Learning for Diagram Comprehension in Multimodal Models

Recent multimodal models such as Contrastive Language-Image Pre-training (CLIP) have shown remarkable ability to align visual and linguistic represent...

Feb 27 2026 2602.23589v1
Annotation-Free Visual Reasoning for High-Resolution Large Multimodal Models via Reinforcement Learning

Current Large Multimodal Models (LMMs) struggle with high-resolution visual inputs during the reasoning process, as the number of image tokens increas...

Feb 27 2026 2602.23615v1
3D Modality-Aware Pre-training for Vision-Language Model in MRI Multi-organ Abnormality Detection

Vision-language models (VLMs) show strong potential for complex diagnostic tasks in medical imaging. However, applying VLMs to multi-organ medical ima...

Feb 27 2026 2602.23652v1
Vision-Language Semantic Grounding for Multi-Domain Crop-Weed Segmentation

Fine-grained crop-weed segmentation is essential for enabling targeted herbicide application in precision agriculture. However, existing deep learning...

Feb 27 2026 2602.23677v1
StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation

Vision-language-action (VLA) models integrate visual observations and language instructions to predict robot actions, demonstrating promising generali...

Feb 27 2026 2602.23721v1
OPTIAGENT: A Physics-Driven Agentic Framework for Automated Optical Design

Optical design is the process of configuring optical elements to precisely manipulate light for high-fidelity imaging. It is inherently a highly non-c...

Feb 27 2026 2602.23761v1
Browse Categories