Skip to main navigation Skip to search Skip to main content

Authorship attribution of Arabic tweets

  • Jordan University of Science and Technology
  • Zayed University

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

25 Scopus citations

Abstract

In tweet authentication, we are concerned with correctly attributing a tweet to its true author based on its textual content. The more general problem of authenticating long documents has been studied before and the most common approach relies on the intuitive idea that each author has a unique style that can be captured using stylometric features (SF). Inspired by the success of modern automatic document classification problem, some researchers followed the Bag-Of-Words (BOW) approach for authenticating long documents. In this work, we consider both approaches and their application on authenticating tweets, which represent additional challenges due to the limitation in their sizes. We focus on the Arabic language due to its importance and the scarcity of works related on it. We create different sets of features from both approaches and compare the performance of different classifiers using them. To the best of our knowledge, this is the first study of its kind to combine these different sets of features for authorship analysis of Arabic tweets. The results show that combining all the feature sets we compute yields the best results.

Original languageEnglish
Title of host publication2016 IEEE/ACS 13th International Conference of Computer Systems and Applications, AICCSA 2016 - Proceedings
PublisherIEEE Computer Society
ISBN (Electronic)9781509043200
DOIs
StatePublished - 2 Jul 2016
Externally publishedYes
Event13th IEEE/ACS International Conference of Computer Systems and Applications, AICCSA 2016 - Agadir, Morocco
Duration: 29 Nov 20162 Dec 2016

Publication series

NameProceedings of IEEE/ACS International Conference on Computer Systems and Applications, AICCSA
Volume0
ISSN (Print)2161-5322
ISSN (Electronic)2161-5330

Conference

Conference13th IEEE/ACS International Conference of Computer Systems and Applications, AICCSA 2016
Country/TerritoryMorocco
CityAgadir
Period29/11/162/12/16

Keywords

  • Authorship Authentication
  • Bag-Of-Words
  • Online Social Networks
  • Stylometric Features

Fingerprint

Dive into the research topics of 'Authorship attribution of Arabic tweets'. Together they form a unique fingerprint.

Cite this