Anda belum login :: 02 May 2025 15:03 WIB
Home
|
Logon
Hidden
»
Administration
»
Collection Detail
Detail
Assessing the Effectiveness of Feature Groups in Author Recognition Tasks With the SOM Model
Oleh:
Tambouratzis, George
Jenis:
Article from Journal - ilmiah internasional
Dalam koleksi:
IEEE Transactions on Systems, Man, and Cybernetics: Part C Applications and Reviews vol. 36 no. 2 (Mar. 2006)
,
page 249-259.
Topik:
Author Identification
;
Self-Organizing Map (SOM)
;
Stylometry
Ketersediaan
Perpustakaan Pusat (Semanggi)
Nomor Panggil:
II69.2
Non-tandon:
1 (dapat dipinjam: 0)
Tandon:
tidak ada
Lihat Detail Induk
Isi artikel
The present paper focuses on studying the effectiveness of the self-organizing map (SOM) when applied to the task of categorizing a corpus of texts according to the style of their authors. This task is of particular importance for information retrieval applications using very large databases of documents. The emphasis of this article is to determine the extent to which the SOM possesses the ability to analyze such data, successfully uncovering the stylistic differences among authors in an unsupervised manner. To that end, a variety of feature vectors are studied, each of which either 1) comprises a single category of linguistic features or 2) spans several different categories of linguistic features, in order to determine the effectiveness of each feature category. It is shown that the highest accuracy is achieved when using a vector covering multiple linguistic categories. A comparison of the results obtained to the results of statistical methods indicates the ability of the SOM network to reveal the clustering potential of isolated parameter groups and its effectiveness in handling efficiently high-dimensional data vectors. Potential extensions to related text-organization techniques, such as the WEBSOM, thus become evident.
Opini Anda
Klik untuk menuliskan opini Anda tentang koleksi ini!
Kembali
Process time: 0 second(s)