Authors :
Tharakesvulu Vangalapat; Priyank Raj Sharma; Somnath Banerjee
Volume/Issue :
Volume 11 - 2026, Issue 9 - September
Google Scholar :
https://tinyurl.com/3w48fxsc
DOI :
https://doi.org/10.38124/ijisrt/26sep096
Note : A published paper may take 4-5
working days from the publication date to appear in PlumX Metrics, Semantic Scholar, and
ResearchGate.
Abstract :
Background:
Practitioners receive contradictory guidance on forecasting method selection because existing comparisons use fixed
train-test splits without examining how training data size affects rankings.
Methods:
We evaluate 11 methods across 11 datasets (2,515–48,204 observations) using two protocols (60/20/20 and 70/15/15
splits) with five metrics (MAPE, SMAPE, RMSE, MAE, MASE).
Results:
With 60% training data, Exponential Smoothing ranks best (4.80). With 70% training data, Prophet dominates (4.47),
followed by LightGBM (4.75) and ETS (5.02). ETS improves 1.75 positions and LightGBM improves 1.56 positions, whereas
Exponential Smoothing declines 1.24 positions. ElasticNet ranks last under both protocols (7.69 and 8.31).
Keywords :
Time Series Forecasting, Method Selection, Training Data Requirements, Empirical Evaluation, Comparative Study.
References :
- O. B. Sezer, M. U. Gudelek, and A. M. Ozbayoglu, “Financial time series forecasting with deep learning: A systematic literature review: 2005–2019,” Appl. Soft Comput., vol. 90, p. 106181, 2020.
- V. K. R. Chimmula and L. Zhang, “Time series forecasting of COVID-19 transmission in Canada using LSTM networks,” Chaos Solitons Fractals, vol. 135, p. 109864, 2020.
- T. Ahmad, H. Zhang, and B. Yan, “A review on renewable energy and electricity requirement forecasting models for smart grid and buildings,” Sustain. Cities Soc., vol. 55, p. 102052, 2020.
- Y. Lv, Y. Duan, W. Kang, Z. Li, and F. Wang, “Traffic flow prediction with big data: A deep learning approach,” IEEE Trans. Intell. Transp. Syst., vol. 16, no. 2, pp. 865–873, 2015.
- S. Makridakis, E. Spiliotis, and V. Assimakopoulos, “Statistical and machine learning forecasting methods: Concerns and ways forward,” PLoS One, vol. 13, no. 3, p. e0194889, 2018.
- H. Zhou et al., “Informer: Beyond efficient transformer for long sequence time-series forecasting,” in Proc. AAAI Conf. Artif. Intell., 2021, pp. 11106–11115.
- G. E. P. Box, G. M. Jenkins, G. C. Reinsel, and G. M. Ljung, Time Series Analysis: Forecasting and Control, 5th ed. Hoboken, NJ: Wiley, 2015.
- R. J. Hyndman, A. B. Koehler, J. K. Ord, and R. D. Snyder, Forecasting with Exponential Smoothing: The State Space Approach. Berlin: Springer, 2008.
- S. J. Taylor and B. Letham, “Forecasting at scale,” Amer. Statist., vol. 72, no. 1, pp. 37–45, 2018.
- O. Triebe et al., “NeuralProphet: Explainable forecasting at scale,” arXiv preprint arXiv:2111.15397, 2021.
- H. Zou and T. Hastie, “Regularization and variable selection via the elastic net,” J. R. Stat. Soc. B, vol. 67, no. 2, pp. 301–320, 2005.
- T. Hastie, R. Tibshirani, and J. Friedman, The Elements of Statistical Learning, 2nd ed. New York: Springer, 2009.
- L. Breiman, “Random forests,” Mach. Learn., vol. 45, no. 1, pp. 5–32, 2001.
- T. Chen and C. Guestrin, “XGBoost: A scalable tree boosting system,” in Proc. 22nd ACM SIGKDD, 2016, pp. 785–794.
- G. Ke et al., “LightGBM: A highly efficient gradient boosting decision tree,” in Adv. Neural Inf. Process. Syst., vol. 30, 2017, pp. 3146–3154.
- L. Prokhorenkova, G. Gusev, A. Vorobev, A. V. Dorogush, and A. Gulin, “CatBoost: Unbiased boosting with categorical features,” in Adv. Neural Inf. Process. Syst., vol. 31, 2018, pp. 6638–6648.
- B. N. Oreshkin, D. Carpov, N. Chapados, and Y. Bengio, “N-BEATS: Neural basis expansion analysis for interpretable time series forecasting,” in Int. Conf. Learn. Represent., 2020.
- T. Zhou et al., “FEDformer: Frequency enhanced decomposed transformer for long-term series forecasting,” in Proc. 39th ICML, 2022, pp. 27268–27286.
- Y. Nie, N. H. Nguyen, P. Sinthong, and J. Kalagnanam, “A time series is worth 64 words: Long-term forecasting with transformers,” in Int. Conf. Learn. Represent., 2023.
- S. Makridakis, E. Spiliotis, and V. Assimakopoulos, “The M4 Competition: 100,000 time series and 61 forecasting methods,” Int. J. Forecast., vol. 36, no. 1, pp. 54–74, 2020.
- S. Makridakis, E. Spiliotis, and V. Assimakopoulos, “M5 accuracy competition: Results, findings, and conclusions,” Int. J. Forecast., vol. 38, no. 4, pp. 1346–1364, 2022.
- V. Cerqueira, L. Torgo, and I. Mozetic, “Evaluating time series forecast-ˇ ing models: An empirical study on performance estimation methods,” Mach. Learn., vol. 109, no. 11, pp. 1997–2028, 2020.
- R. J. Hyndman and G. Athanasopoulos, Forecasting: Principles and Practice, 2nd ed. Melbourne: OTexts, 2018.
Background:
Practitioners receive contradictory guidance on forecasting method selection because existing comparisons use fixed
train-test splits without examining how training data size affects rankings.
Methods:
We evaluate 11 methods across 11 datasets (2,515–48,204 observations) using two protocols (60/20/20 and 70/15/15
splits) with five metrics (MAPE, SMAPE, RMSE, MAE, MASE).
Results:
With 60% training data, Exponential Smoothing ranks best (4.80). With 70% training data, Prophet dominates (4.47),
followed by LightGBM (4.75) and ETS (5.02). ETS improves 1.75 positions and LightGBM improves 1.56 positions, whereas
Exponential Smoothing declines 1.24 positions. ElasticNet ranks last under both protocols (7.69 and 8.31).
Keywords :
Time Series Forecasting, Method Selection, Training Data Requirements, Empirical Evaluation, Comparative Study.