Applying authorship analysis to Arabic web content

43Citations
Citations of this article
52Readers
Mendeley users who have this article in their library.
Get full text

Abstract

The advent and rapid proliferation of internet communication has allowed the realization of numerous security issues. The anonymous nature of online mediums such as email, web sites, and forums provides an attractive communication method for criminal activity. Increased globalization and the boundless nature of the internet have further amplified these concerns due to the addition of a multilingual dimension. The world's social and political climate has caused Arabic to draw a great deal of attention. In this study we apply authorship identification techniques to Arabic web forum messages. Our research uses lexical, syntactic, structural, and content-specific writing style features for authorship identification. We address some of the problematic characteristics of Arabic in route to the development of an Arabic language model that provides a respectable level of classification accuracy for authorship discrimination. We also run experiments to evaluate the effectiveness of different feature types and classification techniques on our dataset. © Springer-Verlag Berlin Heidelberg 2005.

Cite

CITATION STYLE

APA

Abbasi, A., & Chen, H. (2005). Applying authorship analysis to Arabic web content. In Lecture Notes in Computer Science (Vol. 3495, pp. 183–197). Springer Verlag. https://doi.org/10.1007/11427995_15

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free