Authorship Authentication of Short Messages from Social Networks Using Recurrent Artificial Neural Networks

Nesibe Merve Demir

Abstract


Dataset consists of 17000 tweets collected from Twitter, as 500 tweets for each of 34 authors that meet certain criteria. Raw data is collected by using the software Nvivo. The collected raw data is preprocessed to extract frequencies of 200 features. In the data analysis 128 of features are eliminated since they are rare in tweets. As a progressive presentation, five – ten – fifteen – twenty - thirty and thirty four of these 34 authors are selected each time. Since recurrent artificial neural networks are more stable and iterations converge more quickly, in this work this architecture is preferred. In general, ANNs are more successful in distinguishing two classes, therefore for N authors, N×N neural networks are trained for pair wise classification. These N×N experts then organized as N special teams (CANNT) to aggregate decisions of these N×N experts. Number of authors is seen not so effective on the accuracy of the authentication, and around 80% accuracy is achieved for any number of authors.

Keywords


Authorship Authentication; short massages; committee machines; recurrent neural network

Full Text:

PDF


DOI: http://dx.doi.org/10.21533/scjournal.v7i2.163

Refbacks

  • There are currently no refbacks.


Copyright (c) 2018 Nesibe Merve Demir

ISSN 2233 -1859

Digital Object Identifier DOI: 10.21533/scjournal

Creative Commons License
This work is licensed under a Creative Commons Attribution 4.0 International License