“All you can eat” ontology-building: Feeding Wikipedia to Cyc

Abstract
In order to achieve genuine web intelligence, building some kind of large general machine-readable conceptual scheme (i.e. ontology) seems inescapable. Yet the past 20 years have shown that manual ontology-building is not practicable. The recent explosion of free user-supplied knowledge on the Web has led to great strides in automatic ontology building, but quality-control is still a major issue. Ideally one should automatically build onto an already intelligent base. We suggest that the long-running Cyc project is able to assist here. We describe methods used to add 35K new concepts mined from Wikipedia to collections in ResearchCyc entirely automatically. Evaluation with 22 human subjects shows high precision both for the new concepts’ categorization, and their assignment as individuals or collections. Most importantly we show how Cyc itself can be leveraged for ontological quality control by ‘feeding’ it assertions one by one, enabling it to reject those that contradict its other knowledge.
Keywords ontology  Cyc  Wikipedia  concept mapping  knowledge mining
Categories (categorize this paper)
Options
 Save to my reading list
Follow the author(s)
My bibliography
Export citation
Find it on Scholar
Edit this record
Mark as duplicate
Revision history Request removal from index Translate to english
 
Download options
PhilPapers Archive


Upload a copy of this paper     Check publisher's policy on self-archival     Papers currently archived: 11,399
External links
Setup an account with your affiliations in order to access resources via your University's proxy server
Configure custom proxy (use this if your affiliation does not provide a proxy)
Through your library
References found in this work BETA

No references found.

Citations of this work BETA

No citations found.

Similar books and articles
Olena Medelyan & Catherine Legg (2008). Integrating Cyc and Wikipedia: Folksonomy Meets Rigorously Defined Common-Sense. Proceedings of Wikipedia and AI Workshop at the AAAI-08 Conference. Chicago, US, July 12 2008.
Catherine Legg & Samuel Sarjant (2012). Bill Gates is Not a Parking Meter: Philosophical Quality Control in Automated Ontology Building. Proceedings of the Symposium on Computational Philosophy, AISB/IACAP World Congress 2012 (Birmingham, England, July 2-6).
David Milne, Catherine Legg, Medelyan Olena & Witten Ian (2009). Mining Meaning From Wikipedia. International Journal of Human-Computer Interactions 67 (9):716-754.
P. D. Magnus (2006). Epistemology and the Wikipedia. North American Computing and Philosophy Conference.
Analytics

Monthly downloads

Added to index

2012-10-12

Total downloads

15 ( #109,932 of 1,102,971 )

Recent downloads (6 months)

2 ( #183,254 of 1,102,971 )

How can I increase my downloads?

My notes
Sign in to use this feature


Discussion
Start a new thread
Order:
There  are no threads in this forum
Nothing in this forum yet.