KMi Publications

Tech Reports

Tech Report kmi-04-19 Abstract


Ontosophie: A Semi-Automatic System for Ontology Population from Text
Techreport ID: kmi-04-19
Date: 2004
Author(s): David Celjuska, Maria Vargas-Vera
Download PDF

This paper describes a system for semi-automatic population of ontologies with instances from unstructured text. It is based on supervised learning, learns extraction rules from annotated text and then applies those rules on new articles for ontology population. Therefore, the system classifies stories and populates a hand-crafted ontology with new instances of classes defined in it. It is based on three components: Marmot - a natural language processor; Crystal - a dictionary induction tool; and Badger - an information extraction tool. A part of the entire cycle is a user who accepts, rejects or modifies newly extracted and suggested instances to be populated. A description of experiments performed with text corpus consisting of 91 articles is given in turn. The results cover the paper and support a presented hypothesis of assigning a rule confidence value to each extraction rule for improving the performance.
 
KMi Publications
 

New Media Systems is...


Our New Media Systems research theme aims to show how new media devices, standards, architectures and concepts can change the nature of learning.

Our work involves the development of short life-cycle working prototypes of innovative technologies or concepts that we believe will influence the future of open learning within a 3-5 year timescale. Each new media concept is built into a working prototype of how the innovation may change a target community. The working prototypes are all available (in some form) from this website.

Our prototypes themselves are not designed solely for traditional Open Learning, but include a remit to show how that innovation can and will change learning at all levels and in all forms; in education, at work and play.