The role of domain knowledge in data mining

Sarabjot S Anand,John G Hughes,David A Bell

doi:10.1145/221270.221321

Abstract

The ideal situation for a Data Mining or Knowledge Discovery system would be for the user to be able to pose a query of the form “Give me something interesting that could be useful” and for the system to discover some useful knowledge for the user. But such a system would be unrealistic as databases in the real world are very large and so it would be too inefficient to be workable. So the role of the human within the discovery process is essential. Moreover, the measure of what is meant by “interesting to the user” is dependent on the user as well as the domain within which the Data Mining system is being used. In this paper we discuss the use of domain knowledge within Data Mining. We define three classes of domain knowledge: Hierarchical Generalization Trees ( HG-Trees), Attribute Relationship Rules (AR-rules) and EnvironmentBased Constraints (EBC). We discuss how each one of these types of domain knowledge is incorporated into the discovery process within the EDM (Evidential Data Mining) framework for Data Mining proposed earlier by the authors [ANAN94], and in particular within the STRIP (Strong Rule Induction in Parallel) algorithm [ANAN95] implemented within the EDM framework. We highlight the advantages of using domain knowledge within the discovery process by providing results from the application of the STRIP algorithm in the actuarial domain.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

The role of domain knowledge in data mining

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Research of data mining system
Ruiying Xu
-
Ruiying XuRuiying Xu
01 Jan 2015
01 Jan 2015

Data mining and knowledge discovery: The third generation
Gregory Piatetsky-Shapiro
-
Gregory Piatetsky-ShapiroGregory Piatetsky-Shapiro
01 Jan 1997
01 Jan 1997

Research on the Data Mining System Based on B/S Framework and Algrotithm
Hui Ding ... Yao Zong Liu
Applied Mechanics and Materials | VOL. 380-384
Hui Ding, et. al.Hui Ding ... Yao Zong Liu
30 Aug 2013
Applied Mechanics and Materials | VOL. 380-384

Research on the data mining system based on B/S framework and algrotithm
Hui Ding ... Yaozong Liu
-
Hui Ding, et. al. Hui Ding ... Yaozong Liu
01 Nov 2012
01 Nov 2012

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

The role of domain knowledge in data mining

Abstract

Talk to us

Similar Papers