AIMC Topic: Data Curation

Clear Filters Showing 31 to 40 of 147 articles

Self-Supervised Robust Feature Matching Pipeline for Teach and Repeat Navigation.

Sensors (Basel, Switzerland)
The performance of deep neural networks and the low costs of computational hardware has made computer vision a popular choice in many robotic systems. An attractive feature of deep-learned methods is their ability to cope with appearance changes caus...

Active label cleaning for improved dataset quality under resource constraints.

Nature communications
Imperfections in data annotation, known as label noise, are detrimental to the training of machine learning models and have a confounding effect on the assessment of model performance. Nevertheless, employing experts to remove label noise by fully re...

The Challenge of Data Annotation in Deep Learning-A Case Study on Whole Plant Corn Silage.

Sensors (Basel, Switzerland)
Recent advances in computer vision are primarily driven by the usage of deep learning, which is known to require large amounts of data, and creating datasets for this purpose is not a trivial task. Larger benchmark datasets often have detailed proces...

The BMS-LM ontology for biomedical data reporting throughout the lifecycle of a research study: From data model to ontology.

Journal of biomedical informatics
Biomedical research data reuse and sharing is essential for fostering research progress. To this aim, data producers need to master data management and reporting through standard and rich metadata, as encouraged by open data initiatives such as the F...

Tracking cell lineages in 3D by incremental deep learning.

eLife
Deep learning is emerging as a powerful approach for bioimage analysis. Its use in cell tracking is limited by the scarcity of annotated data for the training of deep-learning models. Moreover, annotation, training, prediction, and proofreading curre...

Generalisation Gap of Keyword Spotters in a Cross-Speaker Low-Resource Scenario.

Sensors (Basel, Switzerland)
Models for keyword spotting in continuous recordings can significantly improve the experience of navigating vast libraries of audio recordings. In this paper, we describe the development of such a keyword spotting system detecting regions of interest...

A localization strategy combined with transfer learning for image annotation.

PloS one
This study aims to solve the overfitting problem caused by insufficient labeled images in the automatic image annotation field. We propose a transfer learning model called CNN-2L that incorporates the label localization strategy described in this stu...

Whole-cell segmentation of tissue images with human-level performance using large-scale data annotation and deep learning.

Nature biotechnology
A principal challenge in the analysis of tissue imaging data is cell segmentation-the task of identifying the precise boundary of every cell in an image. To address this problem we constructed TissueNet, a dataset for training segmentation models tha...

The value of human data annotation for machine learning based anomaly detection in environmental systems.

Water research
Anomaly detection is the process of identifying unexpected data samples in datasets. Automated anomaly detection is either performed using supervised machine learning models, which require a labelled dataset for their calibration, or unsupervised mod...

Harnessing clinical annotations to improve deep learning performance in prostate segmentation.

PloS one
PURPOSE: Developing large-scale datasets with research-quality annotations is challenging due to the high cost of refining clinically generated markup into high precision annotations. We evaluated the direct use of a large dataset with only clinicall...