Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series Forecasting

Last update: Jan 08, 2023

Related tags

Overview

Autoformer (NeurIPS 2021)

Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series Forecasting

Time series forecasting is a critical demand for real applications. Enlighted by the classic time series analysis and stochastic process theory, we propose the Autoformer as a general series forecasting model [paper]. Autoformer goes beyond the Transformer family and achieves the series-wise connection for the first time.

In long-term forecasting, Autoformer achieves SOTA, with a 38% relative improvement on six benchmarks, covering five practical applications: energy, traffic, economics, weather and disease.

Autoformer vs. Transformers

1. Deep decomposition architecture

We renovate the Transformer as a deep decomposition architecture, which can progressively decompose the trend and seasonal components during the forecasting process.

Figure 1. Overall architecture of Autoformer.

2. Series-wise Auto-Correlation mechanism

Inspired by the stochastic process theory, we design the Auto-Correlation mechanism, which can discover period-based dependencies and aggregate the information at the series level. This empowers the model with inherent log-linear complexity. This series-wise connection contrasts clearly from the previous self-attention family.

Figure 2. Auto-Correlation mechansim.

Get Started

Install Python 3.6, PyTorch 1.9.0.
Download data. You can obtain all the six benchmarks from Tsinghua Cloud or Google Drive. All the datasets are well pre-processed and can be used easily.
Train the model. We provide the experiment scripts of all benchmarks under the folder ./scripts. You can reproduce the experiment results by:

bash ./scripts/ETT_script/Autoformer_ETTm1.sh
bash ./scripts/ECL_script/Autoformer.sh
bash ./scripts/Exchange_script/Autoformer.sh
bash ./scripts/Traffic_script/Autoformer.sh
bash ./scripts/Weather_script/Autoformer.sh
bash ./scripts/ILI_script/Autoformer.sh

Sepcial-designed implementation

Speedup Auto-Correlation: We built the Auto-Correlation mechanism as a batch-normalization-style block to make it more memory-access friendly. See the paper for details.
Without the position embedding: Since the series-wise connection will inherently keep the sequential information, Autoformer does not need the position embedding, which is different from Transformers.

Main Results

We experiment on six benchmarks, covering five main-stream applications. We compare our model with ten baselines, including Informer, N-BEATS, etc. Generally, for the long-term forecasting setting, Autoformer achieves SOTA, with a 38% relative improvement over previous baselines.

Citation

If you find this repo useful, please cite our paper.

@inproceedings{wu2021autoformer,
  title={Autoformer: Decomposition Transformers with {Auto-Correlation} for Long-Term Series Forecasting},
  author={Haixu Wu and Jiehui Xu and Jianmin Wang and Mingsheng Long},
  booktitle={Advances in Neural Information Processing Systems},
  year={2021}
}

Contact

If you have any question or want to use the code, please contact [email protected] .

Acknowledgement

We appreciate the following github repos a lot for their valuable code base or datasets:

https://github.com/zhouhaoyi/Informer2020

https://github.com/zhouhaoyi/ETDataset

https://github.com/laiguokun/multivariate-time-series-data

Autoformer: Decomposition Transformers with Auto-Correlation for Long-Term Series Forecasting

Related tags

Overview

Autoformer (NeurIPS 2021)

Autoformer vs. Transformers

Get Started

Main Results

Citation

Contact

Acknowledgement

Owner

THUML @ Tsinghua University

Pretrained language model and its related optimization techniques developed by Huawei Noah's Ark Lab.

Convolutional Neural Networks on Graphs with Fast Localized Spectral Filtering

Video-based open-world segmentation

GPOEO is a micro-intrusive GPU online energy optimization framework for iterative applications

Reinforcement learning framework and algorithms implemented in PyTorch.

Code for "Offline Meta-Reinforcement Learning with Advantage Weighting" [ICML 2021]

Simple torch.nn.module implementation of Alias-Free-GAN style filter and resample

Python version of the amazing Reaction Mechanism Generator (RMG).

An Efficient Implementation of Analytic Mesh Algorithm for 3D Iso-surface Extraction from Neural Networks

An Official Repo of CVPR '20 "MSeg: A Composite Dataset for Multi-Domain Segmentation"

[NeurIPS2021] Code Release of Learning Transferable Perturbations

LoL Runes Recommender With Python

Python scripts for performing lane detection using the LSTR model in ONNX

Official implementation of "Articulation Aware Canonical Surface Mapping"

Sparse Physics-based and Interpretable Neural Networks

PuppetGAN - Cross-Domain Feature Disentanglement and Manipulation just got way better! 🚀

Charsiu: A transformer-based phonetic aligner

Notes, programming assignments and quizzes from all courses within the Coursera Deep Learning specialization offered by deeplearning.ai

Implementation of "The Power of Scale for Parameter-Efficient Prompt Tuning"

Source codes for Improved Few-Shot Visual Classification (CVPR 2020), Enhancing Few-Shot Image Classification with Unlabelled Examples