SimCleaner -- Sistema de Padroniza\c{c}\~ao de Bases de Dados utilizando Fun\c{c}\~oes de Similaridade

Autor: Damasceno, Carlos Diego Nascimento, Lobato, Fabio Manoel França, Moutinho, Elton Rocha, de França, Arilene Santos, de Oliveira, Ivan Ikikame, de Santana, Ádamo Lima
Jazyk: portugalština
Rok vydání: 2021
Předmět:
Druh dokumentu: Working Paper
Popis: The Knowledge Discovery in Database (KDD) process permits the detection of pattern in databases, where this analysis may be compromised if database is not consistent, making necessary the use of data cleaning techniques. This paper presents a tool based in similarity functions to help the preprocessing of databases and it behaved efficiently in the standardization of a System of Public Security of the State of Par\'a database and may be reused with other databases and other data mining projects.
Comment: 6 pages, 5 figures, 1 table, Published in Portuguese in the Anais da XIV Semana de Inform\'atica (SEMINF) e Escola Regional de Inform\'atica Norte (ERIN), 2011
Databáze: arXiv