PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 20, 20250 citationsOpen Access

WDformer: A Wavelet-based Differential Transformer Model for Time Series Forecasting

View Full Paper
XWXiaojian WangCZChaoli ZhangZZZhonglong Zheng

Key Points

  • WDformer improves accuracy in time series forecasting tasks through advanced modeling techniques.
  • Achieving state-of-the-art results on multiple datasets, it leverages both time-frequency domain information and attention mechanisms.
  • By focusing on key information and reducing noise, the model enhances forecasting performance in various applications.
  • The innovative differential attention mechanism ensures effective extraction and representation of relevant data features.

Abstract

Time series forecasting has various applications, such as meteorological rainfall prediction, traffic flow analysis, financial forecasting, and operational load monitoring for various systems. Due to the sparsity of time series data, relying solely on time-domain or frequency-domain modeling limits the model's ability to fully leverage multi-domain information. Moreover, when applied to time series forecasting tasks, traditional attention mechanisms tend to over-focus on irrelevant historical information, which may introduce noise into the prediction process, leading to biased results. We proposed WDformer, a wavelet-based differential Transformer model. This study employs the wavelet transform to conduct a multi-resolution analysis of time series data. By leveraging the advantages of joint representation in the time-frequency domain, it accurately extracts the key information components that reflect the essential characteristics of the data. Furthermore, we apply attention mechanisms on inverted dimensions, allowing the attention mechanism to capture relationships between multiple variables. When performing attention calculations, we introduced the differential attention mechanism, which computes the attention score by taking the difference between two separate softmax attention matrices. This approach enables the model to focus more on important information and reduce noise. WDformer has achieved state-of-the-art (SOTA) results on multiple challenging real-world datasets, demonstrating its accuracy and effectiveness. Code is available at https://github.com/xiaowangbc/WDformer.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Wang et al. (2025) studied this question.

synapsesocial.com/papers/68f5fcd68d54a28a75cf1ea9https://doi.org/10.48550/arxiv.2509.25231
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1SDformer: Transformer with Spectral Filter and Dynamic Attention for Multivariate Time Series Long-term Forecasting2024 · 8 citations
  2. 2Beyond flat patching: wavelet tokens with cross-scale and decoupled attention for electricity load forecasting2026
  3. 3Wavelet-Enhanced Transformer for Adaptive Multi-Period Time Series Forecasting2025 · 5 citations
  4. 4Wavelet-Enhanced Transformer for Adaptive Multi-Period Time Series Forecasting2025
  5. 5A Multiscale Transformer Model for Long Time Series Forecasting Based on Discrete Wavelet Transform and Residual Learning Modules2025 · 2 citations