Ophthalmology

Latest AI and machine learning research in ophthalmology for healthcare professionals.

9,853 articles
Stay Ahead - Weekly Ophthalmology research updates
Subscribe
Browse Categories
Showing 6101-6120 of 9,853 articles

VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Recent Multimodal Large Language Models (MLLMs) have typically focused on integrating visual and textual modalities, with less emphasis placed on the role of speech in enhancing interaction. However, speech plays a crucial role in multimodal dialogue systems, and implementing high-performance in both vision and speech tasks remains a significant challenge due to the fundamental modality differen...

MoColl: Agent-Based Specific and General Model Collaboration for Image Captioning

Image captioning is a critical task at the intersection of computer vision and natural language processing, with wide-ranging applications across various domains. For complex tasks such as diagnostic report generation, deep learning models require not only domain-specific image-caption datasets but also the incorporation of relevant general knowledge to provide contextual accuracy. Existing appr...

GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models

In recent years, 2D Vision-Language Models (VLMs) have made significant strides in image-text understanding tasks. However, their performance in 3D ...

Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models

Latent diffusion models with Transformer architectures excel at generating high-fidelity images. However, recent studies reveal an optimization dile...

Training Medical Large Vision-Language Models with Abnormal-Aware Feedback

Existing Medical Large Vision-Language Models (Med-LVLMs), which encapsulate extensive medical knowledge, demonstrate excellent capabilities in unde...

Optimized Relay Lens Design For High-Resolution Image Transmission In Military Target Detection Systems

The design and performance analysis of relay lenses that provide high-performance image transmission for target acquisition and tracking in military...

Real-time Cross-modal Cybersickness Prediction in Virtual Reality

Cybersickness remains a significant barrier to the widespread adoption of immersive virtual reality (VR) experiences, as it can greatly disrupt user...

TexAVi: Generating Stereoscopic VR Video Clips from Text Descriptions

While generative models such as text-to-image, large language models and text-to-video have seen significant progress, the extension to text-to-virt...

Generalized Task-Driven Medical Image Quality Enhancement with Gradient Promotion

Thanks to the recent achievements in task-driven image quality enhancement (IQE) models like ESTR, the image enhancement model and the visual recogn...

Deep Learning-Based SD-OCT Layer Segmentation Quantifies Outer Retina Changes in Patients With Biallelic RPE65 Mutations Undergoing Gene Therapy.

PURPOSE: To quantify outer retina structural changes and define novel biomarkers of inherited retinal degeneration associated with biallelic mutations...

Jan 2 2025 39745677
The Associations Between Myopia and Fundus Tessellation in School Children: A Comparative Analysis of Macular and Peripapillary Regions Using Deep Learning.

PURPOSE: To evaluate the refractive differences among school-aged children with macular or peripapillary fundus tessellation (FT) distribution pattern...

Jan 2 2025 39775798
New Directions for Ophthalmic OCT - Handhelds, Surgery, and Robotics.

The introduction of optical coherence tomography (OCT) in the 1990s revolutionized diagnostic ophthalmic imaging. Initially, OCT's role was primarily ...

Jan 2 2025 39808124
Artificial Intelligence for Optical Coherence Tomography in Glaucoma.

PURPOSE: The integration of artificial intelligence (AI), particularly deep learning (DL), with optical coherence tomography (OCT) offers significant ...

Jan 2 2025 39854198
Artificial Intelligence in Predicting Ocular Hypertension After Descemet Membrane Endothelial Keratoplasty.

PURPOSE: Descemet membrane endothelial keratoplasty (DMEK) has emerged as a novel approach in corneal transplantation over the past two decades. This ...

Jan 2 2025 39869086
Visual-like Template Diffusion: Boosting Single-Sequence Protein Structure Prediction by Adapting Image Diffusion Models

Single-sequence protein structure prediction has drawn increasing attention due to the high computational costs associated with obtaining homologous i...

Intrinsic plasticity underlies malleability of neural network heterogeneity

Diversity exists throughout biology, playing an important role in maintaining robustness and stability. The same is true of the brain, as has become i...

Post-operative tissue fragment puzzling using histopathological vision transformer alignment HiViTAlign

In pathology, reconstructing adjacent tissue parts enables an overview of the macro environment of objects like tumors. Especially, malignoma are of i...

Holotomography-driven learning for in-silico staining of single cells in flow cytometry avoiding co-registration

Virtual staining is the current state-of-the-art computational technique to cleverly enhance intracellular specificity in unstained biological samples...

Clinical and molecular characterisation of primary refractoriness to atezolizumab plus bevacizumab in patients with unresectable hepatocellular carcinoma

Despite improved outcomes with atezolizumab plus bevacizumab (A+B) in hepatocellular carcinoma (HCC), primary refractoriness (PRef), characterised by ...

Anticipatory Eye Gaze as a Marker of Memory

Human memory is typically studied by direct questioning, and the recollection of events is investigated through verbal reports. Thus, current research...

Browse Categories