xlstm-0d6d068b·2 events·first seen Aliases: xLSTM
TiRex-2 is a recurrent xLSTM-based time series foundation model that extends the univariate TiRex to multivariate forecasting with past and future covariates, while supporting streaming inference at constant per-patch cost. The model uses a bidirectional time mixer and asymmetric grouped-attention variate mixer to handle future-known covariates without violating causality over target variables. A synthetic coupling pipeline enables scalable multivariate pretraining from univariate corpora. TiRex-2 claims state-of-the-art zero-shot performance on GIFT-Eval and fev-bench benchmarks with 38.4M–82.5M parameters depending on mode.
A new arXiv paper compares three subquadratic sequence modeling architectures — xLSTM, Mamba-2, and Gated DeltaNet — across code model pre-training, LLM distillation, and time-series foundation model pre-training. xLSTM consistently delivers the strongest performance, which the authors attribute to more flexible and stable memory correction via its gating scheme. The paper provides a unified formulation and analysis of state tracking and memory dynamics across the three architectures, with corroborating results on synthetic length-generalization tasks.