Busca avançada
Ano de início
Entree


Time Series Prediction via Similarity Search: Exploring Invariances, Distance Measures and Ensemble Functions

Texto completo
Autor(es):
Parmezan, Antonio R. S. ; Souza, Vinicius M. A. ; Batista, Gustavo E. A. P. A.
Número total de Autores: 3
Tipo de documento: Artigo Científico
Fonte: IEEE ACCESS; v. 10, p. 22-pg., 2022-01-01.
Resumo

The rapid advance of scientific research in data mining has led to the adaptation of conventional pattern extraction methods to the context of time series analysis. The forecasting (or prediction) task has been supported mainly by regression algorithms based on artificial neural networks, support vector machines, and k-Nearest Neighbors (kNN). However, some studies provided empirical evidence that similarity-based methods, i.e. variations of kNN, constitute a promising approach compared with more complex predictive models from both machine learning and statistics. Although the scientific community has made great strides in increasing the visibility of these easy-to-fit and impressively accurate algorithms, previous work has failed to recognize the right invariances needed for this task. We propose a novel extension of kNN, namely kNN - Time Series Prediction with Invariances (kNN-TSPI), that differs from the literature by combining techniques to obtain amplitude and offset invariance, complexity invariance, and treatment of trivial matches. Our predictor enables more meaningful matches between reference queries and data subsequences. From a comprehensive evaluation with real-world datasets, we demonstrate that kNN-TSPI is a competitive algorithm against two conventional similarity-based approaches and, most importantly, against 11 popular predictors. To assist future research and provide a better understanding of similarity-based method behaviors, we also explore different settings of kNN-TSPI regarding invariances to distortions in time series, distance measures, complexity-invariant distances, and ensemble functions. Results show that kNN-TSPI stands out for its robustness and stability both concerning the parameter k and the accuracy of the projection horizon trends. (AU)

Processo FAPESP: 13/10978-8 - Predição de Séries Temporais por Similaridade
Beneficiário:Antonio Rafael Sabino Parmezan
Modalidade de apoio: Bolsas no Brasil - Mestrado
Processo FAPESP: 16/04986-6 - Armadilhas e sensores inteligentes: uma abordagem inovadora para controle de insetos peste e vetores de doenças
Beneficiário:Gustavo Enrique de Almeida Prado Alves Batista
Modalidade de apoio: Auxílio à Pesquisa - Programa eScience e Data Science - Regular
Processo FAPESP: 18/05859-3 - Armadilhas e sensores inteligentes: uma abordagem inovadora para controle de insetos peste e vetores de doenças
Beneficiário:Vinícius Mourão Alves de Souza
Modalidade de apoio: Bolsas no Brasil - Pós-Doutorado