Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 3921-3940 of 9,853 articles

GDTR: Layer-wise Settling Depth Reveals Biological Grammar in Genomic Foundation Models

Genomic foundation models capture sequence regularities, yet existing interpretability tools rarely ask where in the layer stack a biological grammar becomes stable. We introduce GDTR, the Genomic Deep-Thinking Ratio, a training-free residual-stream lens that assigns each nucleotide token a settling depth : the first layer at which its representation stabilizes against the post-final-norm referenc...

From Hodgkin-Huxley to Pretrained Neural Inference AI

High-density probes record from thousands of neurons simultaneously, yet resolving single-neuron identity remains an ill-posed inverse problem. While detailed simulations precisely characterize the biophysical forward process, their utility for interpreting brain signal remains unclear. Here we show that biophysical simulations of population neuronal electrical signals serve as an effective bridge...

From Pixel to Prognosis: Convolutional and GLCM Feature Fusion for Automated Four-Class Cataract Severity Classification

Objective: To develop a low-cost automated cataract severity classification system operating on standard consumer-grade colour photographs of the eye,...

Jul 20 2026 2607.18349v1
PC-Seg: Progressive Cross-View Consistency for 3D OCT Segmentation from Sparse 2D Annotations

Volumetric segmentation of optical coherence tomography (OCT) images is essential for diagnosing ocular diseases but requires labor-intensive voxel-wi...

Jul 20 2026 2607.17718v2
CANDOR: Chance-Calibrated Discordance in Frozen Foundation Encoders

Frozen encoders are chosen by how well a lightweight head reads a finding from their features, not whether the geometry separates it. Nearest-neighbor...

Jul 20 2026 2607.18451v1
Luminosity-Adaptive Contrast Enhancement Using CLAHE for Retinal Fundus Images with Quantitative Validation and Comparative Analysis

Background: Retinal fundus imaging is central to the early diagnosis of sight-threatening conditions including diabetic retinopathy, glaucoma, and ret...

Jul 20 2026 2607.17691v1
PC-Seg: Progressive Cross-View Consistency for 3D OCT Segmentation from Sparse 2D Annotations

Volumetric segmentation of optical coherence tomography (OCT) images is essential for diagnosing ocular diseases but requires labor-intensive voxel-wi...

Jul 20 2026 2607.17718v1
Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation

End-to-end vision-language navigation (VLN) with causal vision-language models can map instructions and egocentric observations directly to actions, b...

Jul 20 2026 2607.18042v1
VGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy Prediction

Vision-only occupancy prediction requires recovering a semantic 3D occupancy field from calibrated surround-view images, where each view provides obse...

Jul 20 2026 2607.18078v1
FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry

In line with the prevailing direction of vision research, we explore the integration of both generation and editing capabilities for video and image m...

Jul 20 2026 2607.18227v1
Searching for Task-Specific Vision Paths: Evolutionary Block Pruning Across Vision-Language Models

Vision-language models normally execute the same complete vision encoder for every question, even when OCR, counting, object, attribute, and spatial q...

Jul 19 2026 2607.17052v1
A Preoperative Electroencephalography Signature for Predicting Treatment Response to Deep Brain Stimulation in Obsessive-Compulsive Disorder

Deep brain stimulation (DBS) is effective for treatment-refractory obsessive-compulsive disorder (OCD), but outcomes are heterogeneous and non-respond...

Physics-aware Masked Diffusion-based Flood Simulation for Urban Fisheye Disaster Detection

Physical simulations that predict the behavior of urban disasters, such as climate-related flooding, play a crucial role in disaster prevention and th...

Jul 17 2026 2607.15527v1
Ask Twice, Look Twice: Prompt Echoing Resolves the Question-First Paradox in Vision-Language Models

Where should the question go in a vision-language model (VLM) prompt: before the image or after it? Intuition says before: knowing what is asked shoul...

Jul 17 2026 2607.15565v1
Attention-Guided Saliency Maps for Interpreting Visualization Literacy in VLMs

Understanding how vision-language models (VLMs) interpret data visualizations remains an open problem, and is increasingly important as these models a...

Jul 17 2026 2607.16105v1
Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos

We present Audio-Visual Flamingo (AV-Flamingo), a fully open state-of-the-art audio-visual large language model (AV-LLM) for joint understanding and r...

Jul 17 2026 2607.16107v1
Clean-Reference Streaming Detection of Lens Occlusion and Photometric Transitions for Camera Tamper Monitoring

A surveillance camera is an image sensor whose silent physical degradation invalidates every downstream consumer of its data. In-situ integrity alarms...

Jul 16 2026 2607.14760v1
Rotational Motion-Induced Error Compensation for Phase-Shifting Profilometry-Based Eye Reconstruction

With the proliferation of immersive Head-Mounted Displays (HMDs) for Virtual and Augmented Reality (VR/AR), reliable and high-precision eye tracking h...

Jul 16 2026 2607.14876v1
A vision foundation model for single-cell biology via spatial gene cartography

Most single-cell foundation models are adapted from language models, representing each cell as a sequence of gene tokens. This discards the relationsh...

Jul 15 2026 2607.14163v1
SeeSE3: Emergence of 3D Space in Vision Features

In this paper, we ask whether vision foundation models construct representations that reflect the intrinsic properties of 3D Euclidean space. Unlike p...

Jul 15 2026 2607.14228v1
Browse Categories