Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 4781-4800 of 9,853 articles

Camera-Agnostic Autonomous Diagnosis of Glaucomatous Optic Neuropathy using Macular Fundus Imaging and Machine Learning

Abstract Purpose: Glaucoma, a leading cause of irreversible vision loss, often remains undiagnosed due to its asymptomatic progression and the limitations of existing screening methods. This study aimed to validate an artificial intelligence machine learning algorithm for the camera-agnostic detection of glaucomatous optic neuropathy using macula-centered fundus images. Methods: Data were collecte...

Understanding the Transfer Limits of Vision Foundation Models

Foundation models leverage large-scale pretraining to capture extensive knowledge, demonstrating generalization in a wide range of language tasks. By comparison, vision foundation models (VFMs) often exhibit uneven improvements across downstream tasks, despite substantial computational investment. We postulate that this limitation arises from a mismatch between pretraining objectives and the deman...

Jan 22 2026 2601.15888v1
DTP: A Simple yet Effective Distracting Token Pruning Framework for Vision-Language Action Models

Vision-Language Action (VLA) models have shown remarkable progress in robotic manipulation by leveraging the powerful perception abilities of Vision-L...

Jan 22 2026 2601.16065v1
DeepMoLM: Leveraging Visual and Geometric Structural Information for Molecule-Text Modeling

AI models for drug discovery and chemical literature mining must interpret molecular images and generate outputs consistent with 3D geometry and stere...

Jan 21 2026 2601.14732v1
Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning

Chain-of-Thought (CoT) prompting has achieved remarkable success in unlocking the reasoning capabilities of Large Language Models (LLMs). Although CoT...

Jan 21 2026 2601.14750v1
BayesianVLA: Bayesian Decomposition of Vision Language Action Models via Latent Action Queries

Vision-Language-Action (VLA) models have shown promise in robot manipulation but often struggle to generalize to new instructions or complex multi-tas...

Jan 21 2026 2601.15197v1
VisTIRA: Closing the Image-Text Modality Gap in Visual Math Reasoning via Structured Tool Integration

Vision-language models (VLMs) lag behind text-only language models on mathematical reasoning when the same problems are presented as images rather tha...

Jan 20 2026 2601.14440v1
CARPE: Context-Aware Image Representation Prioritization via Ensemble for Large Vision-Language Models

Recent advancements in Large Vision-Language Models (LVLMs) have pushed them closer to becoming general-purpose assistants. Despite their strong perfo...

Jan 20 2026 2601.13622v1
DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes

Social media imagery provides a low-latency source of situational information during natural and human-induced disasters, enabling rapid damage assess...

Jan 20 2026 2601.13839v1
Left-Right Symmetry Breaking in CLIP-style Vision-Language Models Trained on Synthetic Spatial-Relation Data

Spatial understanding remains a key challenge in vision-language models. Yet it is still unclear whether such understanding is truly acquired, and if ...

Jan 19 2026 2601.12809v1
CLIP-Guided Adaptable Self-Supervised Learning for Human-Centric Visual Tasks

Human-centric visual analysis plays a pivotal role in diverse applications, including surveillance, healthcare, and human-computer interaction. With t...

Jan 19 2026 2601.13133v1
Reasoning with Pixel-level Precision: QVLM Architecture and SQuID Dataset for Quantitative Geospatial Analytics

Current Vision-Language Models (VLMs) fail at quantitative spatial reasoning because their architectures destroy pixel-level information required for ...

Jan 19 2026 2601.13401v1
SGW-GAN: Sliced Gromov-Wasserstein Guided GANs for Retinal Fundus Image Enhancement

Retinal fundus photography is indispensable for ophthalmic screening and diagnosis, yet image quality is often degraded by noise, artifacts, and uneve...

Jan 19 2026 2601.13417v1
Effects of Gabor Filters on Classification Performance of CNNs Trained on a Limited Number of Conditions

In this study, we propose a technique to improve the accuracy and reduce the size of convolutional neural networks (CNNs) running on edge devices for ...

Jan 17 2026 2601.11918v1
A Unified Masked Jigsaw Puzzle Framework for Vision and Language Models

In federated learning, Transformer, as a popular architecture, faces critical challenges in defending against gradient attacks and improving model per...

Jan 17 2026 2601.12051v1
Vision-as-Inverse-Graphics Agent via Interleaved Multimodal Reasoning

Vision-as-inverse-graphics, the concept of reconstructing an image as an editable graphics program is a long-standing goal of computer vision. Yet eve...

Jan 16 2026 2601.11109v1
Sociotechnical Challenges of Machine Learning in Healthcare and Social Welfare

Sociotechnical challenges of machine learning in healthcare and social welfare are mismatches between how a machine learning tool functions and the st...

Jan 16 2026 2601.11417v1
Karhunen-Loève Expansion-Based Residual Anomaly Map for Resource-Efficient Glioma MRI Segmentation

Accurate segmentation of brain tumors is essential for clinical diagnosis and treatment planning. Deep learning is currently the state-of-the-art for ...

Jan 16 2026 2601.11833v1
Vision-Language Models vs Autonomous AI Agents for Anterior Capsular Radial Folds: A Diagnostic Study

ImportanceVision-language models (VLMs) enable generalist multimodal reasoning, but their ability to resolve brief, low-contrast cues in surgical vide...

DanQing: An Up-to-Date Large-Scale Chinese Vision-Language Pre-training Dataset

Vision-Language Pre-training (VLP) models demonstrate strong performance across various downstream tasks by learning from large-scale image-text pairs...

Jan 15 2026 2601.10305v1
Browse Categories