Garuda - Garba Rujukan Digital

p-Index From 2021 - 2026

P-Index

This Author published in this journals

All Journal Knowledge Engineering and Data Science

Andrew Nafalski

UniSA Education Futures, School of Engineering, University of South Australia SCT2-39 Mawson Lakes Campus, Adelaide, South Australia 5095, Australia

Author-ID : 2635044

Computer Science & IT Engineering

Published : 1 Documents Claim Missing Document

Claim Missing Document

Articles

Generating Javanese Stopwords List using K-means Clustering Algorithm Aji Prasetya Wibawa; Hidayah Kariima Fithri; Ilham Ari Elbaith Zaeni; Andrew Nafalski
Knowledge Engineering and Data Science Vol 3, No 2 (2020)
Publisher : Universitas Negeri Malang

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.17977/um018v3i22020p106-111

Stopword removal necessary in Information Retrieval. It can remove frequently appeared and general words to reduce memory storage. The algorithm eliminates each word that is precisely the same as the word in the stopword list. However, generating the list could be time-consuming. The words in a specific language and domain must be collected and validated by specialists. This research aims to develop a new way to generate a stop word list using the K-means Clustering method. The proposed approach groups words based on their frequency. The confusion matrix calculates the difference between the findings with a valid stopword list created by a Javanese linguist. The accuracy of the proposed method is 78.28% (K=7). The result shows that the generation of Javanese stopword lists using a clustering method is reliable.

Co-Authors Aji Prasetya Wibawa Hidayah Kariima Fithri Zaeni, Ilham Ari Elbaith

Title

Found 1 Documents
Search

Abstract

Title Search

Found 1 Documents Search

Abstract

Title

Found 1 Documents
Search