ACM SIGMOD Anthology ACM SIGMOD dblp.uni-trier.de

The Role of Domain Knowledge in Data Mining.

Sarabjot S. Anand, David A. Bell, John G. Hughes: The Role of Domain Knowledge in Data Mining. CIKM 1995: 37-43
@inproceedings{DBLP:conf/cikm/AnandBH95,
  author    = {Sarabjot S. Anand and
               David A. Bell and
               John G. Hughes},
  title     = {The Role of Domain Knowledge in Data Mining},
  booktitle = {CIKM '95, Proceedings of the 1995 International Conference on
               Information and Knowledge Management, November 28 - December
               2, 1995, Baltimore, Maryland, USA},
  publisher = {ACM},
  year      = {1995},
  pages     = {37-43},
  ee        = {db/conf/cikm/AnandBH95.html, http://doi.acm.org/10.1145/221270.221321},
  crossref  = {DBLP:conf/cikm/95},
  bibsource = {DBLP, http://dblp.uni-trier.de}
}
BibTeX

Abstract

The ideal situation for a Data Mining or Knowledge Discovery system would be for the user to be able to pose a query of the form "Give me something interesting that could be useful" and for the system to discover some useful knowledge for the user. But such a system would be unrealistic as databases in the real world are very large and so it would be too inefficient to be workable. So the role of the human within the discovery process is essential. Moreover, the measure of what is meant by "interesting to the user" is dependent on the user as well as the domain within which the Data Mining system is being used.

In this paper we discuss the use of domain knowledge within Data Mining. We define three classes of domain knowledge: Hierarchical Generalization Trees ( HG-Trees), Attribute Relationship Rules (AR-rules) and Environment-Based Constraints (EBC). We discuss how each one of these types of domain knowledge is incorporated into the discovery process within the EDM (Evidential Data Mining) framework for Data Mining proposed earlier by the authors [ANAN94], and in particular within the STRIP (Strong Rule Induction in Parallel) algorithm [ANAN95] implemented within the EDM framework. We highlight the advantages of using domain knowledge within the discovery process by providing results from the application of the STRIP algorithm in the actuarial domain.

Copyright © 1995 by the ACM, Inc., used by permission. Permission to make digital or hard copies is granted provided that copies are not made or distributed for profit or direct commercial advantage, and that copies show this notice on the first page or initial screen of a display along with the full citation.


ACM SIGMOD Anthology

CDROM Version: Load the CDROM "Volume 2 Issue 4, CIKM, DOLAP, GIS, SIGFIDET, ..." and ... DVD Version: Load ACM SIGMOD Anthology DVD 1" and ... BibTeX

Printed Edition

CIKM '95, Proceedings of the 1995 International Conference on Information and Knowledge Management, November 28 - December 2, 1995, Baltimore, Maryland, USA. ACM 1995
Contents BibTeX

Online Edition

Citation Page BibTeX

References

[AGRA93]
Rakesh Agrawal, Tomasz Imielinski, Arun N. Swami: Database Mining: A Performance Perspective. IEEE Trans. Knowl. Data Eng. 5(6): 914-925(1993) BibTeX
[ANAN94]
...
[ANAN95]
...
[BELL93]
David A. Bell: From Data Properties to Evidence. IEEE Trans. Knowl. Data Eng. 5(6): 965-969(1993) BibTeX
[BELL94]
...
[FASM94]
...
[FRAW91]
William J. Frawley, Gregory Piatetsky-Shapiro, Christopher J. Matheus: Knowledge Discovery in Databases: An Overview. Knowledge Discovery in Databases 1991: 1-30 BibTeX
[GUAN91]
...
[GUAN92]
...
[GUAN93]
...
[HAN94]
Jiawei Han: Towards Efficient Induction Mechanisms in Database Systems. Theor. Comput. Sci. 133(2): 361-385(1994) BibTeX
[HAN94a]
Jiawei Han, Yongjian Fu: Dynamic Generation and Refinement of Concept Hierarchies for Knowledge Discovery in Databases. KDD Workshop 1994: 157-168 BibTeX
[MAGI95]
...
[MALL94]
...
[PIAT91]
Gregory Piatetsky-Shapiro, William J. Frawley (Eds.): Knowledge Discovery in Databases. AAAI/MIT Press 1991, ISBN 0-262-62080-4
Contents BibTeX
[PIAT91a]
Gregory Piatetsky-Shapiro: Discovery, Analysis, and Presentation of Strong Rules. Knowledge Discovery in Databases 1991: 229-248 BibTeX
[PIAT93]
...

Referenced by

  1. Alex G. Büchner, Maurice D. Mulvenna: Discovering Internet Marketing Intelligence through Online Analytical Web Usage Mining. SIGMOD Record 27(4): 54-61(1998)
BibTeX
ACM SIGMOD Anthology - DBLP: [Home | Search: Author, Title | Conferences | Journals]
CIKM 1995 Proceedings, ACM SIGMOD Anthology: Copyright © by ACM (info@acm.org), Corrections: anthology@acm.org
DBLP: Copyright © by Michael Ley (ley@uni-trier.de), last change: Sat May 16 23:01:47 2009