Skip to search form
Skip to main content
Skip to account menu
Semantic Scholar
Semantic Scholar's Logo
Search 238,226,770 papers from all fields of science
Search
Sign In
Create Free Account
Wrapper (data mining)
Known as:
Wrapper
, Wrapper induction
Wrapper in data mining is a program that extracts content of a particular information source and translates it into a relational form. Many web pages…
Expand
Wikipedia
(opens in a new tab)
Create Alert
Alert
Related topics
Related topics
5 relations
Data mining
Information extraction
Relational model
Web scraping
Broader (1)
World Wide Web
Papers overview
Semantic Scholar uses AI to extract papers important to this topic.
2014
2014
"Linked data as background knowledge for information extraction on the web" by Ziqi Zhang, Anna Lisa Gentile and Isabelle Augenstein with Martin Vesely as coordinator
Ziqi Zhang
,
Anna Lisa Gentile
,
Isabelle Augenstein
SIGWEB Newsl.
2014
Corpus ID: 8278075
Information Extraction (IE) is the technique for transforming textual data into structured representation that can be understood…
Expand
Review
2013
Review
2013
An Overview of Web Data Extraction Techniques
K. Devika
,
S. Surendran
2013
Corpus ID: 16942651
Web pages are usually generated for visualization not for data exchange. Each page may contain several groups of structured data…
Expand
2011
2011
Extract knowledge from semi-structured websites for search task simplification
Yingqin Gu
,
Jun Yan
,
+4 authors
Zheng Chen
International Conference on Information and…
2011
Corpus ID: 16402232
Simplifying the key tasks of search engine users by directly retrieving to them structured knowledge according to their queries…
Expand
2005
2005
Parameterless Information Extraction Using (k,l)-Contextual Tree Languages
S. Raeymaekers
,
M. Bruynooghe
2005
Corpus ID: 14324054
Recently, several wrapper induction algorithms for structured documents have been introduced. They are based on contextual tree…
Expand
2005
2005
Automatically maintaining wrappers for Web sources
Juan Raposo
,
Alberto Pan
,
Manuel Álvarez
,
Justo Hidalgo
International Database Engineering and…
2005
Corpus ID: 1530278
A substantial subset of the Web data follows some kind of underlying structure. Nevertheless, HTML does not contain any schema or…
Expand
2004
2004
Tree pattern inference and matching for wrapper induction on the World Wide Web
A. Hogue
2004
Corpus ID: 108788945
We develop a method for learning patterns from a set of positive examples to retrieve semantic content from tree-structured data…
Expand
2003
2003
( LP ) 2 : Rule Induction for Information Extraction Using Linguistic Constraints
F. Ciravegna
,
F. Ciravegna
2003
Corpus ID: 703505
Machine learning has been widely used in information extraction from texts in the last years. Two directions of research can be…
Expand
2003
2003
Roma , Italy On Automatic Information Extraction from Large Web Sites
V. Crescenzi
2003
Corpus ID: 31648470
Information extraction from Web sites is nowadays a relevant problem, usually performed by software modules called wrappers. A…
Expand
Review
2002
Review
2002
Knowledge Discovery from Semistructured Texts
Hiroshi Sakamoto
,
Hiroki Arimura
,
S. Arikawa
Progress in Discovery Science
2002
Corpus ID: 45345546
This paper surveys our recent results on the knowledge discovery from semistructured texts, which contain heterogeneous…
Expand
2000
2000
Automatic Extraction of Information Blocks Using PAT Trees
Chia-Hui Chang
,
Chun-Nan Hsu
2000
Corpus ID: 14724522
Information extraction from semi-structured Web documents is a critical issue for software agents on the Internet. Previous work…
Expand