Utilização de agrupamento como método de pré-processamento em problemas de regressão linear

Abstract

For real-world data to be explored, pre-processing is needed in order to ease machine learning applications. Nowadays, various pre-processing techniques are available, and this paper aims to show the impacts of using clustering as one of them, more specifically in linear regression problems. Two experiments were carried out, using two different databases: the first one describing data from a series of properties from Ames, a city from Iowa, in the United States of America, and the second one containing information about COVID-19. It was observed that in both cases, clustering before applying the regression model improves regressor performance, based on database nature. Besides the improvement, it is recommended to use clustering alongside other pre-processing techniques.

Description

Citation

MOTTA, Gabriel Gonçalves. Utilização de agrupamento como método de pré-processamento em problemas de regressão linear. 2021. Trabalho de Conclusão de Curso (Graduação em Engenharia de Computação) – Universidade Federal de São Carlos, São Carlos, 2021. Disponível em: https://repositorio.ufscar.br/handle/20.500.14289/15156.

Collections

Endorsement

Review

Supplemented By

Referenced By

Creative Commons license

Except where otherwise noted, this item's license is described as Attribution-NonCommercial-NoDerivs 3.0 Brazil