Building ontology based-on heterogeneous data
Keywords:Domain ontology, information extraction, natural language processing.
Ontologies play an important role in the distinct areas, such as information retrieval, information extraction, question and answer. They help us in capturing and storing knowledge in a particular domain and can be used for distinct applications. In recent years, research relevant to ontology development has produced tangible results concerning semantic web, information extraction, etc. In this paper, a domain specific ontology called Information Technology Ontology (ITO) is proposed. This ontology is built basing on three distinct sources of Wikipedia, WordNet and ACM Digital Library. An information extraction system focusing on computing domain based on this ontology in the future will be built. In order to have an ontology with highest quality and performance as expected, the authors combine some algorithms between machine learning and natural language processing (NLP) for building ontology. Results generated by such experiments show that these algorithms outperform others, especially in semantic relations among entities of ontology.
How to Cite
License1. We hereby assign copyright of our article (the Work) in all forms of media, whether now known or hereafter developed, to the Journal of Computer Science and Cybernetics. We understand that the Journal of Computer Science and Cybernetics will act on my/our behalf to publish, reproduce, distribute and transmit the Work.
2. This assignment of copyright to the Journal of Computer Science and Cybernetics is done so on the understanding that permission from the Journal of Computer Science and Cybernetics is not required for me/us to reproduce, republish or distribute copies of the Work in whole or in part. We will ensure that all such copies carry a notice of copyright ownership and reference to the original journal publication.
3. We warrant that the Work is our results and has not been published before in its current or a substantially similar form and is not under consideration for another publication, does not contain any unlawful statements and does not infringe any existing copyright.
4. We also warrant that We have obtained the necessary permission from the copyright holder/s to reproduce in the article any materials including tables, diagrams or photographs not owned by me/us.