Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 5921-5940 of 9,853 articles

A Survey on Mamba Architecture for Vision Applications

Transformers have become foundational for visual tasks such as object detection, semantic segmentation, and video understanding, but their quadratic complexity in attention mechanisms presents scalability challenges. To address these limitations, the Mamba architecture utilizes state-space models (SSMs) for linear scalability, efficient processing, and improved contextual awareness. This paper i...

Choroidal image analysis for OCT image sequences with applications in systemic health

The choroid, a highly vascular layer behind the retina, is an extension of the central nervous system and has parallels with the renal cortex, with blood flow far exceeding that of the brain and kidney. Thus, there has been growing interest of choroidal blood flow reflecting physiological status of systemic disease. Optical coherence tomography (OCT) enables high-resolution imaging of the choroi...

Geometry-aware RL for Manipulation of Varying Shapes and Deformable Objects

Manipulating objects with varying geometries and deformable objects is a major challenge in robotics. Tasks such as insertion with different objects...

Universal Vessel Segmentation for Multi-Modality Retinal Images

We identify two major limitations in the existing studies on retinal vessel segmentation: (1) Most existing works are restricted to one modality, i....

Visual Agentic AI for Spatial Reasoning with a Dynamic API

Visual reasoning -- the ability to interpret the visual world -- is crucial for embodied agents that operate within three-dimensional scenes. Progre...

Sparse Autoencoders for Scientifically Rigorous Interpretation of Vision Models

To truly understand vision models, we must not only interpret their learned features but also validate these interpretations through controlled expe...

Efficient Spatial Estimation of Perceptual Thresholds for Retinal Implants via Gaussian Process Regression

Retinal prostheses restore vision by electrically stimulating surviving neurons, but calibrating perceptual thresholds (i.e., the minimum stimulus i...

SparseFocus: Learning-based One-shot Autofocus for Microscopy with Sparse Content

Autofocus is necessary for high-throughput and real-time scanning in microscopic imaging. Traditional methods rely on complex hardware or iterative ...

KMT2B-related disorders: expansion of the phenotypic spectrum and long-term efficacy of deep brain stimulation

Heterozygous mutations in KMT2B are associated with an early-onset, progressive, and often complex dystonia (DYT28). Key characteristics of typical ...

Is an Ultra Large Natural Image-Based Foundation Model Superior to a Retina-Specific Model for Detecting Ocular and Systemic Diseases?

The advent of foundation models (FMs) is transforming medical domain. In ophthalmology, RETFound, a retina-specific FM pre-trained sequentially on 1...

Fully Exploiting Vision Foundation Model's Profound Prior Knowledge for Generalizable RGB-Depth Driving Scene Parsing

Recent vision foundation models (VFMs), typically based on Vision Transformer (ViT), have significantly advanced numerous computer vision tasks. Des...

Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models

While recent Large Vision-Language Models (LVLMs) have shown remarkable performance in multi-modal tasks, they are prone to generating hallucinatory...

Event Vision Sensor: A Review

By monitoring temporal contrast, event-based vision sensors can provide high temporal resolution and low latency while maintaining low power consump...

Redefining Robot Generalization Through Interactive Intelligence

Recent advances in large-scale machine learning have produced high-capacity foundation models capable of adapting to a broad array of downstream tas...

A Generative Framework for Bidirectional Image-Report Understanding in Chest Radiography

The rapid advancements in large language models (LLMs) have unlocked their potential for multimodal tasks, where text and visual data are processed ...

Can Generative Agent-Based Modeling Replicate the Friendship Paradox in Social Media Simulations?

Generative Agent-Based Modeling (GABM) is an emerging simulation paradigm that combines the reasoning abilities of Large Language Models with tradit...

Effective Black-Box Multi-Faceted Attacks Breach Vision Large Language Model Guardrails

Vision Large Language Models (VLLMs) integrate visual data processing, expanding their real-world applications, but also increasing the risk of gene...

Exploring Visual Embedding Spaces Induced by Vision Transformers for Online Auto Parts Marketplaces

This study examines the capabilities of the Vision Transformer (ViT) model in generating visual embeddings for images of auto parts sourced from onl...

Segmentation-free integration of nuclei morphology and spatial transcriptomics for retinal images

This study introduces SEFI (SEgmentation-Free Integration), a novel method for integrating morphological features of cell nuclei with spatial transc...

Beyond Vision: How Large Language Models Interpret Facial Expressions from Valence-Arousal Values

Large Language Models primarily operate through text-based inputs and outputs, yet human emotion is communicated through both verbal and non-verbal ...

Browse Categories