STEMMING BAHASA JAWA MENGGUNAKAN DAMERAU LEVENSHTEIN DISTANCE (DLD)
Autor: | Muhammad Nu’man Hakim, Aji Prasetya Wibawa |
---|---|
Rok vydání: | 2021 |
Předmět: | |
Zdroj: | JURNAL TEKNIK INFORMATIKA. 14:22-27 |
ISSN: | 2549-7901 1979-9160 |
DOI: | 10.15408/jti.v14i1.15010 |
Popis: | Stemming is one of the essential stages of text mining. This process removes prefixes and suffixes to produce root words in a text. This study uses a string matching algorithm, namely Damerau Levenshtein Distance (DLD), to find the basic word forms of Javanese. Test data of 300 words that have a prefix, insertion, suffix, a combination of prefix and suffix, and word repetition. The results of this study indicate that the Damerau Levenshtein Distance (DLD) algorithm can be used for Stemming Javanese text with an accuracy value of 49.6%. |
Databáze: | OpenAIRE |
Externí odkaz: |