Tech Reports
Tech Report kmi-05-14 Abstract
Extracting Domain Ontologies with CORDER
Techreport ID: kmi-05-14
Date: 2005
Author(s): Camilo Thorne, Jianhan Zhu, Victoria Uren
The CORDER web mining engine developed at the Knowledge Media Institute computes a lexical coocurrence network out of websites - a binary relation R. A natural extension of CORDER would be that of learning an ontology. However, our work shows that coocurrence proves insufficient to discover concepts and conceptual taxonomies (i.e. very simple ontologies) out of this network. To tackle this problem two unsupervised learning methods were studied based, on the one hand, on set similarity (and thus on a set-based representation of the data) and, on the other hand, on cosine similarity (and thus on a vector-space representation of the data). The underlying idea being that of taking into account, for the clustering, as features, their related coocurring entities (and thus the indirect links among the entities), as suggested, for instance, by O. Ferret. For the purposes of this study, we restricted ourselves to (solely) research areas. The most promising results in our experiments were given by the vector-space representation. To validate the results we used the ACM classification of computer science research areas as our gold standard.
Future Internet
KnowledgeManagementMultimedia &
Information SystemsNarrative
HypermediaNew Media SystemsSemantic Web &
Knowledge ServicesSocial Software
Future Internet is...

To succeed the Future Internet will need to address a number of cross-cutting challenges including:
- Scalability in the face of peer-to-peer traffic, decentralisation, and increased openness
- Trust when government, medical, financial, personal data are increasingly trusted to the cloud, and middleware will increasingly use dynamic service selection
- Interoperability of semantic data and metadata, and of services which will be dynamically orchestrated
- Pervasive usability for users of mobile devices, different languages, cultures and physical abilities
- Mobility for users who expect a seamless experience across spaces, devices, and velocities
Future Internet from KMi.
Check out these Hot Future Internet Projects:
List all Future Internet Projects
Check out these Hot Future Internet Technologies:
List all Future Internet Technologies
List all Future Internet Projects
Check out these Hot Future Internet Technologies:
List all Future Internet Technologies

