Conceptual data sampling for breast cancer histology image classification
View/ Open
Publisher version (Check access options)
Check access options
Date
2017Author
Rezk, EmanAwan, Zainab
Islam, Fahad
Jaoua, Ali
Al Maadeed, Somaya
Zhang, Nan
Das, Gautam
Rajpoot, Nasir
...show more authors ...show less authors
Metadata
Show full item recordAbstract
Data analytics have become increasingly complicated as the amount of data has increased. One technique that is used to enable data analytics in large datasets is data sampling, in which a portion of the data is selected to preserve the data characteristics for use in data analytics. In this paper, we introduce a novel data sampling technique that is rooted in formal concept analysis theory. This technique is used to create samples reliant on the data distribution across a set of binary patterns. The proposed sampling technique is applied in classifying the regions of breast cancer histology images as malignant or benign. The performance of our method is compared to other classical sampling methods. The results indicate that our method is efficient and generates an illustrative sample of small size. It is also competing with other sampling methods in terms of sample size and sample quality represented in classification accuracy and F1 measure. 1 2017 Elsevier Ltd
Collections
- Computer Science & Engineering [2402 items ]