On the use of cross-validation for time series predictor evaluation

被引：603

作者：

Bergmeir, Christoph ^{[1
]}

Benitez, Jose M. ^{[1
]}

机构：

[1] Univ Granada, Dept Comp Sci & Artificial Intelligence, ETS Ingn Informat & Telecomunicac, CITIC UGR, E-18071 Granada, Spain

来源：

INFORMATION SCIENCES | 2012年 / 191卷

关键词：

Cross-validation; Time series; Predictor evaluation; Error measures; Machine learning; Regression; ARTIFICIAL NEURAL-NETWORKS; MODEL-SELECTION; SIGNIFICANCE TESTS;

D O I：

10.1016/j.ins.2011.12.028

中图分类号：

TP [自动化技术、计算机技术];

学科分类号：

0812 ;

摘要：

In time series predictor evaluation, we observe that with respect to the model selection procedure there is a gap between evaluation of traditional forecasting procedures, on the one hand, and evaluation of machine learning techniques on the other hand. In traditional forecasting, it is common practice to reserve a part from the end of each time series for testing, and to use the rest of the series for training. Thus it is not made full use of the data, but theoretical problems with respect to temporal evolutionary effects and dependencies within the data as well as practical problems regarding missing values are eliminated. On the other hand, when evaluating machine learning and other regression methods used for time series forecasting, often cross-validation is used for evaluation, paying little attention to the fact that those theoretical problems invalidate the fundamental assumptions of cross-validation. To close this gap and examine the consequences of different model selection procedures in practice, we have developed a rigorous and extensive empirical study. Six different model selection procedures, based on (i) cross-validation and (ii) evaluation using the series' last part, are used to assess the performance of four machine learning and other regression techniques on synthetic and real-world time series. No practical consequences of the theoretical flaws were found during our study, but the use of cross-validation techniques led to a more robust model selection. To make use of the "best of both worlds", we suggest that the use of a blocked form of cross-validation for time series evaluation became the standard procedure, thus using all available information and circumventing the theoretical problems. (C) 2012 Elsevier Inc. All rights reserved.

引用

页码：192 / 213

页数：22

共 54 条

[1] Application of a new hybrid neuro-evolutionary system for day-ahead price forecasting of electricity markets [J].