Knowledge Discovery in Textual Databases (KDT)

Ronen Feldman, Ido Dagan

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

352 Scopus citations

Abstract

The information age is characterized by a rapid growth in the amount of information available in electronic media. Traditional data handling methods are not adequate to cope with this information flood. Knowledge Discovery in Databases (KDD) is a new paradigm that focuses on computerized exploration of large amounts of data and on discovery of relevant and interesting patterns within them. While most work on KDD is concerned with structured databases, it is clear that this paradigm is required for handling the huge amount of information that is available only in unstructured textual form. To apply traditional KDD on texts it is necessary to impose some structure on the data that would be rich enough to allow for interesting KDD operations. On the other hand, we have to consider the severe limitations of current text processing technology and define rather simple structures that can be extracted from texts fairly automatically and in a reasonable cost. We propose using a text categorization paradigm to annotate text articles with meaningful concepts that are organized in hierarchical structure. We suggest that this relatively simple annotation is rich enough to provide the basis for a KDD framework, enabling data summarization, exploration of interesting patterns, and trend analysis. This research combines the KDD and text categorization paradigms and suggests advances to the state of the art in both areas.

Original languageEnglish
Title of host publicationKDD 1995 - Proceedings of the 1st International Conference on Knowledge Discovery and Data Mining
EditorsUsama M. Fayyad, Ramasamy Uthurusamy
PublisherAAAI Press
Pages112-117
Number of pages6
ISBN (Electronic)0929280822, 9780929280820
StatePublished - 1995
Externally publishedYes
Event1st International Conference on Knowledge Discovery and Data Mining, KDD 1995 - Montreal, Canada
Duration: 20 Aug 199521 Aug 1995

Publication series

NameKDD 1995 - Proceedings of the 1st International Conference on Knowledge Discovery and Data Mining

Conference

Conference1st International Conference on Knowledge Discovery and Data Mining, KDD 1995
Country/TerritoryCanada
CityMontreal
Period20/08/9521/08/95

Bibliographical note

Publisher Copyright:
© 1995, AAAI (www.aaai.org). All rights reserved.

Fingerprint

Dive into the research topics of 'Knowledge Discovery in Textual Databases (KDT)'. Together they form a unique fingerprint.

Cite this