全部 标题 作者
关键词 摘要

OALib Journal期刊
ISSN: 2333-9721
费用:99美元

查看量下载量

相关文章

更多...

Authorship Attribution Using Principal Component Analysis and Nearest Neighbor Rule for Neural Networks

Keywords: principal components , stylometry , stylistic features , syntactic characteristics , multilayer preceptor , competitive learning , artificial neural network.

Full-Text   Cite this paper   Add to My Lib

Abstract:

Feature extraction is a common problem in statistical pattern recognition. It refers to a process whereby a data space is transformed into a feature space that, in theory, has exactly the same dimension as the original data space. However, the transformation is designed in such a way that the data set may be represented by a reduced number of "effective" features and yet retain most of the intrinsic information content of the data; in other words, the data set undergoes a dimensionality reduction. Principal component analysis is one of these processes. In this paper the data collected by counting selected syntactic characteristics in around a thousand paragraphs of each of the sample books underwent a principal component analysis. To make a comparison, the original data is also processed. Authors of texts identified with higher success by the competitive neural networks, which use principal components. The process repeated on another group of authors, and similar results are obtained.

Full-Text

Contact Us

service@oalib.com

QQ:3279437679

WhatsApp +8615387084133