About the journal
Browse
Collections
Multimedia collections
Authors & reviewers
A survey of deep time series forecasting backbone architectures: progress, pitfalls, and a systematic comparison
Xiang LI , Yanping ZHENG , Zhewei WEI
Front. Comput. Sci. ›› 2026, Vol. 20 ›› Issue (8) : 2008371
Deep learning-based time series forecasting has largely converged to standardized training protocols and a narrow set of public benchmarks. While such standardization improves comparability, evaluations based on averaged metrics over fixed windows obscure variable-level differences, mask long-horizon degradation, and diverge from real-world rolling-forecasting scenarios. This study revisits these limitations by surveying seven major backbone architectures and systematically evaluating 21 representative models across diverse datasets. A fine-grained variable-level analysis shows that, in several settings, extending the input window contributes more to forecasting accuracy than architectural innovations, yet such gains diminish rapidly and may even become detrimental as the window grows excessively. Furthermore, models with strong average performance often behave inconsistently across variables and differ markedly in their ability to capture short- and long-term temporal dynamics. These findings highlight inherent constraints on predictability and call for new research directions, including domain-specific end-to-end forecasting pipelines, forecasting with exogenous drivers, and the development of large time series models.
time series forecasting / backbone architectures / variable-level analysis
| [1] |
|
| [2] |
|
| [3] |
|
| [4] |
|
| [5] |
|
| [6] |
|
| [7] |
|
| [8] |
|
| [9] |
|
| [10] |
Abdelmalak I, Madhusudhanan K, Kloetergens C, Yalavarit V K, Stubbemann M, Schmidt-Thieme L. Channel dependence, limited lookback windows, and the simplicity of datasets: how biased is time series forecasting? 2025, arXiv preprint arXiv: 2502.09683 |
| [11] |
|
| [12] |
|
| [13] |
|
| [14] |
|
| [15] |
|
| [16] |
|
| [17] |
|
| [18] |
|
| [19] |
|
| [20] |
|
| [21] |
|
| [22] |
|
| [23] |
|
| [24] |
|
| [25] |
|
| [26] |
|
| [27] |
|
| [28] |
|
| [29] |
|
| [30] |
|
| [31] |
|
| [32] |
|
| [33] |
|
| [34] |
|
| [35] |
|
| [36] |
|
| [37] |
|
| [38] |
|
| [39] |
|
| [40] |
|
| [41] |
|
| [42] |
|
| [43] |
|
| [44] |
|
| [45] |
|
| [46] |
|
| [47] |
Nie X, Zhou X, Li Z, Wang L, Lin X, Tong T. LogTrans: providing efficient local-global fusion with transformer and CNN parallel network for biomedical image segmentation. In: Proceedings of the 24th IEEE International Conference on High Performance Computing & Communications; 8th International Conference on Data Science & Systems; 20th International Conference on Smart City; 8th International Conference on Dependability in Sensor, Cloud & Big Data Systems & Application. 2022, 769−776 |
| [48] |
|
| [49] |
|
| [50] |
|
| [51] |
|
| [52] |
|
| [53] |
|
| [54] |
|
| [55] |
|
| [56] |
|
| [57] |
|
| [58] |
|
| [59] |
|
| [60] |
|
| [61] |
|
| [62] |
|
| [63] |
|
| [64] |
|
| [65] |
|
| [66] |
|
| [67] |
|
| [68] |
|
| [69] |
|
| [70] |
|
| [71] |
|
| [72] |
Wang Z, Kong F, Feng S, Wang M, Yang X, Zhao H, Wang D, Zhang Y. Is Mamba effective for time series forecasting? Neurocomputing, 2025, 619: 129178 |
| [73] |
|
| [74] |
|
| [75] |
|
| [76] |
|
| [77] |
|
| [78] |
|
| [79] |
|
| [80] |
|
| [81] |
|
| [82] |
|
| [83] |
|
| [84] |
|
| [85] |
|
| [86] |
|
| [87] |
|
| [88] |
|
| [89] |
|
| [90] |
|
| [91] |
|
| [92] |
|
| [93] |
|
| [94] |
|
| [95] |
|
| [96] |
|
| [97] |
|
| [98] |
|
| [99] |
|
| [100] |
|
| [101] |
|
| [102] |
|
| [103] |
|
| [104] |
|
| [105] |
|
| [106] |
|
| [107] |
|
| [108] |
|
| [109] |
|
| [110] |
|
| [111] |
|
| [112] |
|
| [113] |
|
| [114] |
|
| [115] |
|
| [116] |
|
| [117] |
|
| [118] |
|
| [119] |
|
| [120] |
|
| [121] |
|
| [122] |
|
| [123] |
|
| [124] |
|
| [125] |
|
| [126] |
|
| [127] |
|
| [128] |
|
| [129] |
|
| [130] |
|
| [131] |
|
| [132] |
|
| [133] |
|
| [134] |
|
| [135] |
|
| [136] |
|
| [137] |
|
| [138] |
|
Higher Education Press
/
| 〈 |
|
〉 |