DocumentCode
3229680
Title
Geographic Named Entity Disambiguation with Automatic Profile Generation
Author
Peng, Yefei ; He, Daqing ; Mao, Ming
Author_Institution
Sch. of Inf. Sci., Pittsburgh Univ., PA
fYear
2006
fDate
18-22 Dec. 2006
Firstpage
522
Lastpage
525
Abstract
Knowledge rich approach of processing documents has been viewed as a method to improve over simple bag-of-word representation. Extracting location information from documents and link them to some ontology such as world gazetteer through a disambiguation process becomes an interesting and important topic. Lacking of training data is a problem in disambiguation method. In this paper we described a method to automatically extract training data from large collection of documents based on local context disambiguation, and then sense profiles are generated automatically for disambiguation use. Another topic of this paper is to describe a linear combination method to combine different types of evidences of disambiguation. We explored three different evidences including location sense context in training documents, local neighbor context, and the popularity of individual location sense. Our results show that combining the three evidences generates reasonable results
Keywords
information retrieval; ontologies (artificial intelligence); automatic profile generation; entity disambiguation; extract training data; geographic distribution; linear combination method; location information; Cities and towns; Data mining; Helium; Humans; Information analysis; Information processing; Information science; Ontologies; Training data; Visualization;
fLanguage
English
Publisher
ieee
Conference_Titel
Web Intelligence, 2006. WI 2006. IEEE/WIC/ACM International Conference on
Conference_Location
Hong Kong
Print_ISBN
0-7695-2747-7
Type
conf
DOI
10.1109/WI.2006.87
Filename
4061424
Link To Document