Local rough set: A solution to rough data analysis in big data

Yuhua Qian,Xinyan Liang,Qi Wang,Jiye Liang,Bing Liu,Andrzej Skowron,Yiyu Yao,Jianmin Ma,Chuangyin Dang

doi:10.1016/j.ijar.2018.01.008

Abstract

As a supervised learning method, classical rough set theory often requires a large amount of labeled data, in which concept approximation and attribute reduction are two key issues. With the advent of the age of big data however, labeling data is an expensive and laborious task and sometimes even infeasible, while unlabeled data are cheap and easy to collect. Hence, techniques for rough data analysis in big data using a semi-supervised approach, with limited labeled data, are desirable. Although many concept approximation and attribute reduction algorithms have been proposed in the classical rough set theory, quite often, these methods are unable to work well in the context of limited labeled big data. The challenges to classical rough set theory can be summarized with three issues: limited labeled property of big data, computational inefficiency and over-fitting in attribute reduction. To address these three challenges, we introduce a theoretic framework called local rough set, and develop a series of corresponding concept approximation and attribute reduction algorithms with linear time complexity, which can efficiently and effectively work in limited labeled big data. Theoretical analysis and experimental results show that each of the algorithms in the local rough set significantly outperforms its original counterpart in classical rough set theory. It is worth noting that the performances of the algorithms in the local rough set become more significant when dealing with larger data sets.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Local rough set: A solution to rough data analysis in big data

Abstract

Talk to us

Similar Papers

More From: International Journal of Approximate Reasoning

Lead the way for us

Journal: International Journal of Approximate Reasoning	Publication Date: Apr 5, 2018
Citations: 127

Similar Papers

Uncertainty Estimation Using Probabilistic Dependency: An Extended Rough Set-Based Approach
Indrajit Ghosh
-
Indrajit GhoshIndrajit Ghosh
01 Jan 2020
01 Jan 2020

Rough set analysis of relational structures
Tuan-Fang Fan
Information Sciences | VOL. 221
Tuan-Fang FanTuan-Fang Fan
05 Oct 2012
Information Sciences | VOL. 221

Some Concepts of Incomplete Multigranulation Based on Rough Intuitionistic Fuzzy Sets
B K Tripathy ... Arnab Mitra
-
B K Tripathy, et. al.B K Tripathy ... Arnab Mitra
01 Jan 2012
01 Jan 2012

Data-driven Valued Tolerance Relation Based on the Extended Rough Set
Guoyin Wang ... Lihe Guan
Fundamenta Informaticae | VOL. 132
Guoyin Wang, et. al.Guoyin Wang ... Lihe Guan
01 Jan 2014
Fundamenta Informaticae | VOL. 132

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Local rough set: A solution to rough data analysis in big data

Abstract

Talk to us

Similar Papers

More From: International Journal of Approximate Reasoning