DocumentCode :
2881999
Title :
Evaluating text visualization: An experiment in authorship analysis
Author :
Benjamin, Victor ; Wingyan Chung ; Abbasi, Ali ; Chuang, Jen-Hui ; Larson, Catherine A. ; Hsinchun Chen
Author_Institution :
MIS Dept., Univ. of Arizona, Tucson, AZ, USA
fYear :
2013
fDate :
4-7 June 2013
Firstpage :
16
Lastpage :
20
Abstract :
Analyzing authorship of online texts is an important analysis task in security-related areas such as cybercrime investigation and counter-terrorism, and in any field of endeavor in which authorship may be uncertain or obfuscated. This paper presents an automated approach for authorship analysis using machine learning methods, a robust stylometric feature set, and a series of visualizations designed to facilitate analysis at the feature, author, and message levels. A testbed consisting of 506,554 forum messages, in English and Arabic, from 14,901 authors was first constructed. A prototype portal system was then developed to support feasibility analysis of the approach. A preliminary evaluation to assess the efficacy of the text visualizations was conducted. The evaluation showed that task performance with the visualization functions was more accurate and more efficient than task performance without the visualizations.
Keywords :
Internet; authorisation; data visualisation; feature extraction; learning (artificial intelligence); natural language processing; portals; text analysis; Arabic; English; author level; automated authorship analysis; counter-terrorism; cybercrime; feasibility analysis; feature level; machine learning method; message level; online text visualization function; prototype portal system; robust stylometric feature set; task performance; Accuracy; Feature extraction; HTML; Heating; Portals; Visualization; Writing; authorship analysis; online forum; terrorism; text visualization;
fLanguage :
English
Publisher :
ieee
Conference_Titel :
Intelligence and Security Informatics (ISI), 2013 IEEE International Conference on
Conference_Location :
Seattle, WA
Print_ISBN :
978-1-4673-6214-6
Type :
conf
DOI :
10.1109/ISI.2013.6578778
Filename :
6578778
Link To Document :
بازگشت