Quantity beats quality for semantic segmentation of corrosion in images. - Citegraph

Paper Info

Title
Quantity beats quality for semantic segmentation of corrosion in images.

Abstract
Dataset creation is typically one of the first steps when applying Artificial Intelligence methods to a new task; and the real world performance of models hinges on the quality and quantity of data available. Producing an image dataset for semantic segmentation is resource intensive, particularly for specialist subjects where class segmentation is not able to be effectively farmed out. The benefit of producing a large, but poorly labelled, dataset versus a small, expertly segmented dataset for semantic segmentation is an open question. Here we show that a large, noisy dataset outperforms a small, expertly segmented dataset for training a Fully Convolutional Network model for semantic segmentation of corrosion in images. A large dataset of 250 images with segmentations labelled by undergraduates and a second dataset of just 10 images, with segmentations labelled by subject matter experts were produced. The mean Intersection over Union and micro F-score metrics were compared after training for 50,000 epochs. This work is illustrative for researchers setting out to develop deep learning models for detection and location of specialist features.

Year	Venue	Field
2018	arXiv: Computer Vision and Pattern Recognition	Pattern recognition,Segmentation,Subject-matter expert,Computer science,Artificial intelligence,Deep learning,Network model,Machine learning
DocType	Volume	Citations
Journal	abs/1807.03138	0
PageRank	References	Authors
0.34	0	3

Authors (3 rows)

Cited by (0 rows)

References (0 rows)

Name	Order	Citations	PageRank
Will Nash	1	0	0.34
Tom Drummond	2	2676	159.45
Nick Birbilis	3	0	0.34

1