Instance Pruning by Filtering Uninformative Words: An Information Extraction Case Study

In this paper we present a novel instance pruning technique for Information Extraction (IE). In particular, our technique filters out uninformative words from texts on the basis of the assumption that very frequent words in the language do not provide any specific information about the text in which they appear, therefore their expectation of being (part of… CONTINUE READING