Submitted:
13 July 2024
Posted:
16 July 2024
You are already at the latest version
Abstract
Keywords:
1. Introduction
2. Related Work
2.1. sLSTM
2.2. mLSTM
3. Proposed Method
4. Results
| Datasets | Features | Timesteps | Granularity |
|---|---|---|---|
| Weather | 21 | 52,696 | 10 min |
| Traffic | 862 | 17,544 | 1 h |
| ILI | 7 | 966 | 1 week |
| ETTh1/ ETTh2 | 7 | 17,420 | 1 h |
| ETTm1/ETTm2 | 7 | 69,680 | 5 min |
| PEMS03 | 358 | 26,209 | 5 min |
| PEMS04 | 307 | 16,992 | 5 min |
| PEMS07 | 883 | 28,224 | 5 min |
| PEMS08 | 170 | 17,856 | 5 min |
| Electricity | 321 | 26,304 | 1 h |
5. Discussion
6. Conclusions
Author Contributions
Funding
Data Availability Statement
Conflicts of Interest
References
- Box, G.E.; Jenkins, G.M.; Reinsel, G.C.; Ljung, G.M. Time Series Analysis: Forecasting and Control; John Wiley & Sons: Hoboken, NJ, USA, 2015. [Google Scholar]
- Dubey, Ashutosh Kumar, Abhishek Kumar, Vicente García-Díaz, Arpit Kumar Sharma, and Kishan Kanhaiya. “Study and analysis of SARIMA and LSTM in forecasting time series data.” Sustainable Energy Technologies and Assessments 47 (2021): 101474. [CrossRef]
- Zhang, Guoqiang Peter. “An investigation of neural networks for linear time-series forecasting.” Computers & Operations Research 28, no. 12 (2001): 1183-1202. [CrossRef]
- De Livera, A.M.; Hyndman, R.J.; Snyder, R.D. Forecasting time series with complex seasonal patterns using exponential smoothing. J. Am. Stat. Assoc. 2011, 106, 1513–1527. [Google Scholar] [CrossRef]
- Ristanoski, Goce, Wei Liu, and James Bailey. “Time series forecasting using distribution enhanced linear regression.” In Advances in Knowledge Discovery and Data Mining: 17th Pacific-Asia Conference, PAKDD 2013, Gold Coast, Australia, April 14–17, 2013, Proceedings, Part I 17, pp. 484-495. Springer Berlin Heidelberg, 2013. [CrossRef]
- Chen, Tianqi, and Carlos Guestrin. “Xgboost: A scalable tree boosting system.” In Proceedings of the 22nd acm sigkdd international conference on knowledge discovery and data mining, pp. 785-794. 2016. [CrossRef]
- Hewamalage, Hansika, Christoph Bergmeir, and Kasun Bandara. “Recurrent neural networks for time series forecasting: Current status and future directions.” International Journal of Forecasting 37, no. 1 (2021): 388-427. [CrossRef]
- Petneházi, G. Recurrent neural networks for time series forecasting. arXiv 2019, arXiv:1901.00069. [Google Scholar]
- Zhao, Bendong, Huanzhang Lu, Shangfeng Chen, Junliang Liu, and Dongya Wu. “Convolutional neural networks for time series classification.” Journal of Systems Engineering and Electronics 28, no. 1 (2017): 162-169. [CrossRef]
- Borovykh, Anastasia, Sander Bohte, and Cornelis W. Oosterlee. “Conditional time series forecasting with convolutional neural networks.” arXiv 2017, arXiv:1703.04691. arXiv:1703.04691.
- Koprinska, Irena, Dengsong Wu, and Zheng Wang. “Convolutional neural networks for energy time series forecasting.” In 2018 international joint conference on neural networks (IJCNN), pp. 1-8. IEEE, 2018. [CrossRef]
- Zhou, H.; Zhang, S.; Peng, J.; Zhang, S.; Li, J.; Xiong, H.; Zhang, W. Informer: Beyond efficient transformer for long sequence time-series forecasting. Proc. AAAI Conf. Artif. Intell. 2021, 35, 11106–11115. [Google Scholar] [CrossRef]
- Wu, H.; Xu, J.; Wang, J.; Long, M. Autoformer: Decomposition transformers with auto-correlation for long-term series forecasting. In Advances in Neural Information Processing Systems 34 (NeurIPS 2021); Neural Information Processing Systems Foundation, Inc. (NeurIPS): San Diego, CA, USA, 2021; pp. 22419–22430. [Google Scholar]
- Zhou, T.; Ma, Z.; Wen, Q.; Wang, X.; Sun, L.; Jin, R. Fedformer: Frequency enhanced decomposed transformer for long-term series forecasting. In Proceedings of the 39th International Conference on Machine Learning PMLR 2022, Baltimore, MD, USA, 17–23 July 2022; pp. 27268–27286. [Google Scholar]
- Li, S.; Jin, X.; Xuan, Y.; Zhou, X.; Chen, W.; Wang, Y.X.; Yan, X. Enhancing the locality and breaking the memory bottleneck of transformer on time series forecasting. In Advances in Neural Information Processing Systems 32 (NeurIPS 2019); Neural Information Processing Systems Foundation, Inc. (NeurIPS): San Diego, CA, USA, 2019. [Google Scholar]
- Nie, Y.; Nguyen, N.H.; Sinthong, P.; Kalagnanam, J. A Time Series is worth 64 words: Long-term forecasting with Transformers. arXiv arXiv:2211.14730, 2022.
- Liu, S.; Yu, H.; Liao, C.; Li, J.; Lin, W.; Liu, A.X.; Dustdar, S. Pyraformer: Low-complexity pyramidal attention for long-range time series modeling and forecasting. In Proceedings of the International Conference on Learning Representations 2022, Online, 25–29 April 2022. [Google Scholar]
- Liu, Yong, Tengge Hu, Haoran Zhang, Haixu Wu, Shiyu Wang, Lintao Ma, and Mingsheng Long. “itransformer: Inverted transformers are effective for time series forecasting.” arXiv preprint arXiv:2310.06625 (2023).
- Zeng, Ailing, Muxi Chen, Lei Zhang, and Qiang Xu. “Are transformers effective for time series forecasting?.” In Proceedings of the AAAI conference on artificial intelligence, vol. 37, no. 9, pp. 11121-11128. 2023. [CrossRef]
- Alharthi, Musleh, and Ausif Mahmood. “Enhanced Linear and Vision Transformer-Based Architectures for Time Series Forecasting.” Big Data and Cognitive Computing 8, no. 5 (2024): 48. [CrossRef]
- Wu, Haixu, Tengge Hu, Yong Liu, Hang Zhou, Jianmin Wang, and Mingsheng Long. “Timesnet: Temporal 2d-variation modeling for general time series analysis.” arXiv preprint arXiv:2210.02186 (2022).
- Gu, Albert, Karan Goel, and Christopher Ré. “Efficiently modeling long sequences with structured state spaces.” arXiv preprint arXiv:2111.00396 (2021).
- Zhang, Michael, Khaled K. Saab, Michael Poli, Tri Dao, Karan Goel, and Christopher Ré. “Effectively modeling time series with simple discrete state spaces.” arXiv preprint arXiv:2303.09489 (2023).
- Beck, Maximilian, Korbinian Pöppel, Markus Spanring, Andreas Auer, Oleksandra Prudnikova, Michael Kopp, Günter Klambauer, Johannes Brandstetter, and Sepp Hochreiter. “xLSTM: Extended Long Short-Term Memory.” arXiv preprint arXiv:2405.04517 (2024).
- Vaswani, Ashish, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, and Illia Polosukhin. “Attention is all you need.” Advances in neural information processing systems 30 (2017).
- Ioffe, Sergey, and Christian Szegedy. “Batch normalization: Accelerating deep network training by reducing internal covariate shift.” In International conference on machine learning, pp. 448-456. pmlr, 2015.
- Kim, Taesung, Jinhee Kim, Yunwon Tae, Cheonbok Park, Jang-Ho Choi, and Jaegul Choo. “Reversible instance normalization for accurate time-series forecasting against distribution shift.” In International Conference on Learning Representations. 2021.






| Models | xLSTMTime | PatchTST | DLinear | FEDformer | Autoformer | Informer | Pyraformer | |
| Metric | MSE MAE | MSE MAE | MSE MAE | MSE MAE | MSE MAE | MSE MAE | MSE MAE | |
| Weather | 96 | 0.144 0.187 | 0.149 0.198 | 0.176 0.237 | 0.238 0.314 | 0.249 0.329 | 0.354 0.405 | 0.896 0.556 |
| 192 | 0.192 0.236 | 0.194 0.241 | 0.220 0.282 | 0.275 0.329 | 0.325 0.370 | 0.419 0.434 | 0.622 0.624 | |
| 336 | 0.237 0.272 | 0.245 0.282 | 0.265 0.319 | 0.339 0.377 | 0.351 0.391 | 0.583 0.543 | 0.739 0.753 | |
| 720 | 0.313 0.326 | 0.314 0.334 | 0.323 0.362 | 0.389 0.409 | 0.415 0.426 | 0.916 0.705 | 1.004 0.934 | |
| Traffic | 96 | 0.358 0.242 | 0.360 0.249 | 0.410 0.282 | 0.576 0.359 | 0.597 0.371 | 0.733 0.410 | 2.085 0.468 |
| 192 | 0.378 0.253 | 0.379 0.256 | 0.423 0.287 | 0.610 0.380 | 0.607 0.382 | 0.777 0.435 | 0.867 0.467 | |
| 336 | 0,392 0,261 | 0.392 0.264 | 0.436 0.296 | 0.608 0.375 | 0.623 0.387 | 0.776 0.434 | 0.869 0.469 | |
| 720 | 0,434 0,287 | 0.432 0.286 | 0.466 0.315 | 0.621 0.375 | 0.639 0.395 | 0.827 0.466 | 0.881 0.473 | |
| Electricity | 96 | 0.128 0.221 | 0.129 0.222 | 0.14 0.237 | 0.186 0.302 | 0.196 0.313 | 0.304 0.393 | 0.386 0.449 |
| 192 | 0.150 0.243 | 0.147 0.240 | 0.153 0.249 | 0.197 0.311 | 0.211 0.324 | 0.327 0.417 | 0.386 0.443 | |
| 336 | 0.166 0.259 | 0.163 0.259 | 0.169 0.267 | 0.213 0.328 | 0.214 0.327 | 0.333 0.422 | 0.378 0.443 | |
| 720 | 0.185 0.276 | 0.197 0.290 | 0.203 0.301 | 0.233 0.344 | 0.236 0.342 | 0.351 0.427 | 0.376 0.445 | |
| Illness | 24 | 1.514 0.694 | 1.319 0.754 | 2.215 1.081 | 2.624 1.095 | 2.906 1.182 | 4.657 1.449 | 1.420 2.012 |
| 36 | 1.519 0.722 | 1.579 0.870 | 1.963 0.963 | 2.516 1.021 | 2.585 1.038 | 4.650 1.463 | 7.394 2.031 | |
| 48 | 1.500 0.725 | 1.553 0.815 | 2.130 1.024 | 2.505 1.041 | 3.024 1.145 | 5.004 1.542 | 7.551 2.057 | |
| 60 | 1.418 0.715 | 1.470 0.788 | 2.368 1.096 | 2.742 1.122 | 2.761 1.114 | 5.071 1.543 | 7.662 2.100 | |
| ETTh1 | 96 | 0.368 0.395 | 0.370 0.400 | 0.375 0.399 | 0.376 0.415 | 0.435 0.446 | 0.941 0.769 | 0.664 0.612 |
| 192 | 0.401 0.416 | 0.413 0.429 | 0.405 0.416 | 0.423 0.446 | 0.456 0.457 | 1.007 0.786 | 0.790 0.681 | |
| 336 | 0.422 0.437 | 0.422 0.440 | 0.439 0.443 | 0.444 0.462 | 0.486 0.487 | 1.038 0.784 | 0.891 0.738 | |
| 720 | 0.441 0.465 | 0.447 0.468 | 0.472 0.490 | 0.469 0.492 | 0.515 0.517 | 1.144 0.857 | 0.963 0.782 | |
| ETTh2 | 96 | 0.273 0.333 | 0.274 0.337 | 0.289 0.353 | 0.332 0.374 | 0.332 0.368 | 1.549 0.952 | 0.645 0.597 |
| 192 | 0.340 0.378 | 0.341 0.382 | 0.383 0.418 | 0.407 0.446 | 0.426 0.434 | 3.792 1.542 | 0.788 0.683 | |
| 336 | 0.373 0.403 | 0.329 0.384 | 0.448 0.465 | 0.400 0.447 | 0.477 0.479 | 4.215 1.642 | 0.907 0.747 | |
| 720 | 0.398 0.430 | 0.379 0.422 | 0.605 0.551 | 0.412 0.469 | 0.453 0.490 | 3.656 1.619 | 0.963 0.783 | |
| ETTm1 | 96 | 0.286 0.335 | 0.293 0.346 | 0.299 0.343 | 0.326 0.390 | 0.510 0.492 | 0.626 0.560 | 0.543 0.510 |
| 192 | 0.329 0.361 | 0.333 0.370 | 0.335 0.365 | 0.365 0.415 | 0.514 0.495 | 0.725 0.619 | 0.557 0.537 | |
| 336 | 0.358 0.379 | 0.369 0.392 | 0.369 0.386 | 0.392 0.425 | 0.510 0.492 | 1.005 0.741 | 0.754 0.655 | |
| 720 | 0.416 0.411 | 0.416 0.420 | 0.425 0.421 | 0.446 0.458 | 0.527 0.493 | 1.133 0.845 | 0.908 0.724 | |
| ETTm2 | 96 | 0.164 0.250 | 0.166 0.256 | 0.167 0.260 | 0.180 0.271 | 0.205 0.293 | 0.355 0.462 | 0.435 0.507 |
| 192 | 0.218 0.288 | 0.223 0.296 | 0.224 0.303 | 0.252 0.318 | 0.278 0.336 | 0.595 0.586 | 0.730 0.673 | |
| 336 | 0.271 0.322 | 0.274 0.329 | 0.281 0.342 | 0.324 0.364 | 0.343 0.379 | 1.270 0.871 | 1.201 0.845 | |
| 720 | 0.361 0.380 | 0.362 0.385 | 0.397 0.421 | 0.410 0.420 | 0.414 0.419 | 3.001 1.267 | 3.625 1.451 | |
| Models | Our xlstmTime | iTransformer | RLinear | PatchTST | Crossformer | DLinear | SCINet | |
|---|---|---|---|---|---|---|---|---|
| Metric | MSE MAE | MSE MAE | MSE MAE | MSE MAE | MSE MAE | MSE MAE | MSE MAE | |
| PEMS03 | 12 | 0.065 0.166 | 0.071 0.174 | 0.126 0.236 | 0.099 0.216 | 0.090 0.203 | 0.122 0.243 | 0.066 0.172 |
| 24 | 0.087 0.194 | 0.093 0.201 | 0.246 0.334 | 0.142 0.259 | 0.121 0.240 | 0.201 0.317 | 0.085 0.198 | |
| 48 | 0.125 0.232 | 0.125 0.236 | 0.551 0.529 | 0.211 0.319 | 0.202 0.317 | 0.333 0.425 | 0.127 0.238 | |
| 96 | 0.192 0.291 | 0.071 0.174 | 0.126 0.236 | 0.099 0.216 | 0.090 0.203 | 0.457 0.515 | 0.178 0.287 | |
| PEMS0 | 12 | 0.074 0.175 | 0.078 0.183 | 0.138 0.252 | 0.105 0.224 | 0.098 0.218 | 0.148 0.272 | 0.073 0.177 |
| 24 | 0.090 0.195 | 0.095 0.205 | 0.258 0.348 | 0.1530.275 | 0.131 0.256 | 0.224 0.340 | 0.084 0.193 | |
| 48 | 0.123 0.230 | 0.120 0.233 | 0.572 0.544 | 0.229 0.339 | 0.205 0.326 | 0.355 0.437 | 0.099 0.211 | |
| 96 | 0.174 0.280 | 0.150 0.262 | 1.137 0.820 | 0.291 0.389 | 0.402 0.457 | 0.452 0.504 | 0.114 0.227 | |
| PEMS0 | 12 | 0.059 0.151 | 0.067 0.165 | 0.118 0.235 | 0.095 0.207 | 0.094 0.200 | 0.115 0.242 | 0.068 0.171 |
| 24 | 0.077 0.170 | 0.088 0.190 | 0.242 0.341 | 0.150 0.262 | 0.139 0.247 | 0.210 0.329 | 0.119 0.225 | |
| 48 | 0.105 0.204 | 0.110 0.215 | 0.562 0.541 | 0.253 0.340 | 0.311 0.369 | 0.398 0.458 | 0.149 0.237 | |
| 96 | 0.148 0.247 | 0.139 0.245 | 1.096 0.795 | 0.346 0.404 | 0.396 0.442 | 0.594 0.553 | 0.141 0.234 | |
| PEMS08 | 12 | 0.072 0.169 | 0.079 0.182 | 0.133 0.247 | 0.168 0.232 | 0.165 0.214 | 0.154 0.276 | 0.087 0.184 |
| 24 | 0.101 0.199 | 0.115 0.219 | 0.249 0.343 | 0.224 0.281 | 0.215 0.260 | 0.248 0.353 | 0.122 0.221 | |
| 48 | 0.149 0.238 | 0.186 0.235 | 0.569 0.544 | 0.321 0.354 | 0.315 0.355 | 0.440 0.470 | 0.189 0.270 | |
| 96 | 0.224 0.289 | 0.221 0.267 | 1.166 0.814 | 0.408 0.417 | 0.377 0.397 | 0.674 0.565 | 0.236 0.300 | |
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content. |
© 2024 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (http://creativecommons.org/licenses/by/4.0/).