Pembobotan Berdasarkan Tingkat Kesamaan Semantik pada Metode Fuzzy Semi-Supervised Co-Clustering untuk Pengelompokkan Dokumen Teks

  • Galang Amanda Dwi P. Institut Teknologi Sepuluh November (ITS)
  • Gregorius Edwadr Institut Teknologi Sepuluh November (ITS)
  • Agus Zainal Arifin Institut Teknologi Sepuluh November (ITS)

Abstract

Nowadays, a large number of information can not be reached by the reader because of the misclassification of text-based documents. The misclassified data can also make the readers obtain the wrong information. The method which is proposed by this paper is aiming to classify the documents into the correct group.  Each document will have a membership value in several different classes. The method will be used to find the degree of similarity between the two documents is the semantic similarity. In fact, there is no document that doesn’t have a relationship with the other but their relationship might be close to 0. This method calculates the similarity between two documents by taking into account the level of similarity of words and their synonyms. After all inter-document similarity values obtained, a matrix will be created. The matrix is then used as a semi-supervised factor. The output of this method is the value of the membership of each document, which must be one of the greatest membership value for each document which indicates where the documents are grouped. Classification result computed by the method shows a good value which is 90 %.

Index Terms - Fuzzy co-clustering, Heuristic, Semantica Similiarity, Semi-supervised learning.

Downloads

Download data is not yet available.
Published
2014-12-01
How to Cite
Dwi P., G., Edwadr, G., & Arifin, A. (2014). Pembobotan Berdasarkan Tingkat Kesamaan Semantik pada Metode Fuzzy Semi-Supervised Co-Clustering untuk Pengelompokkan Dokumen Teks. Ultimatics : Jurnal Teknik Informatika, 6(2), 46-51. https://doi.org/https://doi.org/10.31937/ti.v6i2.333