Title of article :
Determining the context of text using augmented latent semantic indexing
Author/Authors :
Tom Rishel، نويسنده , , Louise A. Perkins، نويسنده , ,
Sumanth Yenduri، نويسنده , ,
Farnaz Zand، نويسنده ,
Issue Information :
ماهنامه با شماره پیاپی سال 2007
Abstract :
Latent semantic analysis has been used for several years to improve the performance of document library searches. We show that latent semantic analysis, augmented with a Part–of–Speech Tagger, may be an effective algorithm for classifying a textual document as well. Using Brilleʹs Part–of–Speech Tagger, we truncate the singular value decomposition used in latent semantic analysis to reduce the size of the word–frequency matrix. This method is then tested on a toy problem, and has shown to increase search accuracy. We then relate these results to natural language processing and show that latent semantic analysis can be combined with context free grammars to infer semantic meaning from natural language. English is the natural language currently being used.
Journal title :
Journal of the American Society for Information Science and Technology
Journal title :
Journal of the American Society for Information Science and Technology