DocumentCode
2881999
Title
Evaluating text visualization: An experiment in authorship analysis
Author
Benjamin, Victor ; Wingyan Chung ; Abbasi, Ali ; Chuang, Jen-Hui ; Larson, Catherine A. ; Hsinchun Chen
Author_Institution
MIS Dept., Univ. of Arizona, Tucson, AZ, USA
fYear
2013
fDate
4-7 June 2013
Firstpage
16
Lastpage
20
Abstract
Analyzing authorship of online texts is an important analysis task in security-related areas such as cybercrime investigation and counter-terrorism, and in any field of endeavor in which authorship may be uncertain or obfuscated. This paper presents an automated approach for authorship analysis using machine learning methods, a robust stylometric feature set, and a series of visualizations designed to facilitate analysis at the feature, author, and message levels. A testbed consisting of 506,554 forum messages, in English and Arabic, from 14,901 authors was first constructed. A prototype portal system was then developed to support feasibility analysis of the approach. A preliminary evaluation to assess the efficacy of the text visualizations was conducted. The evaluation showed that task performance with the visualization functions was more accurate and more efficient than task performance without the visualizations.
Keywords
Internet; authorisation; data visualisation; feature extraction; learning (artificial intelligence); natural language processing; portals; text analysis; Arabic; English; author level; automated authorship analysis; counter-terrorism; cybercrime; feasibility analysis; feature level; machine learning method; message level; online text visualization function; prototype portal system; robust stylometric feature set; task performance; Accuracy; Feature extraction; HTML; Heating; Portals; Visualization; Writing; authorship analysis; online forum; terrorism; text visualization;
fLanguage
English
Publisher
ieee
Conference_Titel
Intelligence and Security Informatics (ISI), 2013 IEEE International Conference on
Conference_Location
Seattle, WA
Print_ISBN
978-1-4673-6214-6
Type
conf
DOI
10.1109/ISI.2013.6578778
Filename
6578778
Link To Document