SimCleaner -- Sistema de Padronizac{c}~ao de Bases de Dados utilizando Func{c}~oes de Similaridade


Abstract in English

The Knowledge Discovery in Database (KDD) process permits the detection of pattern in databases, where this analysis may be compromised if database is not consistent, making necessary the use of data cleaning techniques. This paper presents a tool based in similarity functions to help the preprocessing of databases and it behaved efficiently in the standardization of a System of Public Security of the State of Para database and may be reused with other databases and other data mining projects.

Download