|
IGI Global
Main Office
701 E. Chocolate Avenue
Hershey, PA 17033, USA
Tel: 717-533-8845 x100
Toll Free: 1-866-342-6657
Fax: 717-533-8661
or 717-533-7115
|
|
|
Acquiring Semantic Sibling Associations from Web Documents:
| Our Price: |
$30.00 US |
| Article #: |
ITJ3906 |
| Number of pages: |
83-98 pages |
| Source: |
International Journal of Data Warehousing and Mining, Vol. 3, Issue 4 |
| Author(s): |
Brunzel, Marko; Spiliopoulou, Myra |
| Affiliation(s): |
German Research Center for Artificial Intelligence (DFKI GmbH), Germany; Otto-von-Guericke-Universitat Magdeburg, Germany |
Order Now!
This document will be delivered electronically. Terms of Delivery |
|
Description
The automated discovery of relationships among terms contributes to the automation of the ontology engineering process and allows for sophisticated query expansion in information retrieval. While there are many findings on the identification of direct hierarchical relations among concepts, less attention has been paid on the discovery sibling terms. These are terms that share a common, a priori unknown parent such as co-hyponyms and co-meronyms. In this study, we present our results on the discovery of pairs or groups of sibling terms with XTREEM-SA (Xhtml TREE mining for sibling associations), an algorithm that extracts semantics from Web documents. While conventional methods process an appropriately prepared corpus, XTREEM-SA takes as input an arbitrary collection of Web documents on a given topic and finds sibling relations between terms in this corpus. It is thus independent of domain and language, does not require linguistic preprocessing, and does not rely on syntactic or other rules on text formation. We describe XTREEM-SA and evaluate it toward two reference ontologies. In this context, we also elaborate on the challenges of evaluating semantics extracted from the Web against handcrafted ontologies of high quality but possibly low coverage. |