Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 4021-4040 of 9,853 articles

AC3S: Adaptive Conditioning for 3D-Aware Synthetic Data Generation

Synthetic data generation has emerged as a powerful tool for improving data scalability in computer vision. Recent diffusion-based pipelines have demonstrated strong photorealism. However, how to enforce precise 3D structure and pose consistency in generated images remains challenging. Existing methods leverage visual prompts such as edge maps to guide diffusion models, but often suffer from over-...

Jun 30 2026 2606.31204v1

Decodable Is Not Grounded: A Vision-Ablation Arbiter for VLM Spatial Reasoning

The standard way to read latent knowledge out of a model, a linear probe confirmed by a steering recovery, can systematically overstate what a vision-language model (VLM) actually grounds in the image. We show this on spatial reasoning, where the error is invisible to both probing and steering yet exposed by a one-line causal control: replacing the image with a gray blank. Probes decode the within...

Jun 30 2026 2606.31257v1
Visual Semantic Entropy: Do Vision Language Models Recognize Visual Ambiguity?

Vision-language models can produce confident answers on visually ambiguous inputs, resulting in biased predictions. Common entropy-based methods, such...

Jun 30 2026 2606.31407v1
Fully Automated High-Precision Segmentation of Retinal Atrophy and Ellipsoid Zone Thickness in OCT: A Reliable Tool for Real-World GA Monitoring

Geographic atrophy (GA) secondary to age-related macular degeneration (AMD) requires precise monitoring of relevant structural biomarkers to assess di...

Jun 30 2026 2606.31502v1
Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement Learning

Vision-language models (VLMs) combining reinforcement learning (RL) ignite remarkable progress in multimodal reasoning, yet still struggle with medica...

Jun 30 2026 2606.31599v1
InstanceControl: Controllable Complex Image Generation without Instance Labeling

Controllable image generation methods, such as ControlNet, have demonstrated a remarkable capacity to introduce visual conditions(e.g., depth maps) to...

Jun 30 2026 2606.31924v1
Dynamic Prediction of Alternating Recurrent Events via Neural Network

Alternating recurrent events -- event-times of a specific nature that trigger a secondary refractory period -- occur in a wide-range of fields, includ...

Jun 29 2026 2606.30889v1
Computer-Vision Procedural Telemetry for Airway Guidance: A Public 30-Run Manikin Evidence-Package Audit

Background: Computer vision-enabled airway workflows can turn airway video into timestamped model-observation fields, but later blinded review and tra...

FalconTrack: Photorealistic Auto-Labeled Perception and Physics-Aware Vision-Based Aerial Tracking

Vision-based aerial tracking is critical in GPS-denied environments. Reliable perception for tracking depends on large-scale labeled data, yet most ph...

Jun 29 2026 2606.29783v1
Cross-Modal Iteration Distillation for Robust IHD Screening: The IDNet Framework and A New Benchmark

Color Fundus Photography (CFP) offers a low-cost and non-invasive route for ischemic heart disease (IHD) screening, but current studies are limited by...

Jun 29 2026 2606.30027v1
CogSENet: Blind Image Deblurring with Blur-Conditioned Semantic Routing and Explicit Frequency Fusion

Blind image deblurring demands the recovery of high-fidelity details and coherent structures from complex, unknown degradations. Current blind image d...

Jun 29 2026 2606.30030v1
Latent Noise Mask for Reducing Visual Redundancy in Multimodal Large Language Models

Multimodal large language models (MLLMs) often fail in fine-grained visual reasoning, as question-relevant visual cues are diluted by dense and redund...

Jun 29 2026 2606.30168v1
SHOVIR: A Benchmark for Evaluating Vision Shortcut Learning in Radiology Report Generation

Current evaluation protocols for Vision-Language Models (VLMs) in Radiology Report Generation (RRG) rely on report-level metrics that measure lexical ...

Jun 29 2026 2606.30201v1
VisReflect: Latent Visual Reflection for Fine-Grained Perception in Long Visual Context

Large Vision Language Models (LVLMs) have achieved remarkable success on vision-language tasks, yet fine-grained perception over high-resolution image...

Jun 29 2026 2606.30288v1
Beyond Point Estimates for Glaucoma Visual Field Forecasting with Diffusion Models

Forecasting visual fields (VFs) is critical for personalized monitoring and treatment planning in glaucoma. This is inherently uncertain due to hetero...

Jun 29 2026 2606.30417v1
GLACIER: Rethinking Mass Spectrum Prediction as an Object Detection Problem

Predicting tandem mass spectra (MS/MS) from molecular structures represents a central task in analytical chemistry with direct relevance to clinical m...

Jun 28 2026 2606.29161v1
Fast Enough to Act: Spatio-Temporal Visual Token Merging for Low-Latency Robotic VLMs and VLAs

Vision-language models and vision-language action models endow the robot with unprecedented capabilities. However, the input of video and high-resolut...

Jun 28 2026 2606.29350v1
Can Machines Really See Objects in Images? A Study Based on Syntactic Distance and Visual Self-Referential Instances

Can a vision model truly see an object, or does it only fit surface-level visual cues? Following Wittgenstein's view that the limits of language are t...

Jun 28 2026 2606.29416v1
Bit-ViP: Leveraging Bit-planes to Preserve Visual Privacy in Images through Obfuscation

The unprecedented growth of computer vision applications, such as surveillance systems and social media, raises security and visual privacy concerns, ...

Jun 28 2026 2606.29417v1
MIRROR: Aligning Semantic Relations from Language to Image via Gromov--Wasserstein

Multimodal Large Language Models (MLLMs) inherit rich relational priors from their language backbones, yet often fail when asked to apply these relati...

Jun 28 2026 2606.29462v1
Browse Categories