مرکز منطقه ای اطلاع رساني علوم و فناوري - Aggregation of Information Resources on the Invisible Web

DocumentCode :

483311

Title :

Aggregation of Information Resources on the Invisible Web

Author :

Li, Gang ; Kou, Gang

Author_Institution :

Sch. of Inf. Manage., Wuhan Univ., Wuhan

fYear :

2009

fDate :

23-25 Jan. 2009

Firstpage :

773

Lastpage :

776

Abstract :

There are huge numbers of valuable information resources resided on Invisible Web. However, it is hard to use for us. In this paper we propose a system called NewsReaper that is capable of making Invisible Web to be visible, especially the huge number of real-time information, which update frequently and are time-sensitive. NewsReaper makes use of information extraction, text classification, full text index, RSS technologies to aggregate Invisible Web information resource. In first, this paper analyzes the reasons why it is invisible and four types of Invisible Web. At the same time it summarizes the characteristics of Invisible Web. On this basis, the system architecture of NewsReaper will to be introduced. In order to verify this system, there is a test for campus recruitment information compared with general-purpose search engines. Finally, the weaknesses of this system and some further works are discussed.

Keywords :

classification; information resources; information retrieval; NewsReaper; RSS technologies; campus recruitment information; full text index; general-purpose search engines; information extraction; information resource aggregation; invisible Web; real-time information; text classification; Costs; Data mining; Databases; Information management; Information resources; Real time systems; Recruitment; Search engines; Software systems; Web pages; Keywords-Invisible Web; information resources aggregation;

fLanguage :

English

Publisher :

ieee

Conference_Titel :

Knowledge Discovery and Data Mining, 2009. WKDD 2009. Second International Workshop on

Conference_Location :

Moscow

Print_ISBN :

978-0-7695-3543-2

Type :

conf

DOI :

10.1109/WKDD.2009.165

Filename :

4772050

Link To Document :

https://search.ricest.ac.ir/dl/search/defaultta.aspx?DTC=49&DC=483311