Logo image
Tree-based methods for clustering time series using domain-relevant attributes
期刊文章

Tree-based methods for clustering time series using domain-relevant attributes

Mahsa Ashouri, Galit ShmueliChor-Yiu Sin
Journal of Business Analytics, 卷.2(1), 頁碼.1-23
01/2019

摘要

ARIMA clustering forecasting linear regression model-based partitioning tree Time series Industrial and Manufacturing Engineering Information Systems Management Information Systems
We propose two methods for time-series clustering that capture temporal information (trend, seasonality, autocorrelation) and domain-relevant cross-sectional attributes. The methods are based on model-based partitioning (MOB) trees and can be used as automated yet transparent tools for clustering large collections of time series. We address the challenge of using common time-series models in MOB by instead utilising least squares regression. We propose two methods. The single-step method clusters series using trend, seasonality, lags and domain-relevant cross-sectional attributes. The two-step method first clusters by trend, seasonality and cross-sectional attributes, and then clusters the residuals by autocorrelation and domain-relevant attributes. Both methods produce clusters interpretable by domain experts. We illustrate our approach by considering one-step-ahead forecasting and compare to autoregressive integrated moving average (ARIMA) models for forecasting many Wikipedia pageviews time series. The tree-based approach produces forecasts on par with ARIMA, yet is significantly faster and more efficient, thereby suitable for large collections of time-series. The simple parametric forecasting models allow for interpretable time-series clusters.

相關連結

指標

1 檢視次數

詳細資料

Logo image