Pith. sign in

REVIEW 1 cited by

UnitNorm: Rethinking Normalization for Transformers in Time Series

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.15903 v1 pith:XM7LYBKP submitted 2024-05-24 cs.LG

classification cs.LG
keywords normalizationattentionseriestimeunitnormperformanceanalysisclassification
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Normalization techniques are crucial for enhancing Transformer models' performance and stability in time series analysis tasks, yet traditional methods like batch and layer normalization often lead to issues such as token shift, attention shift, and sparse attention. We propose UnitNorm, a novel approach that scales input vectors by their norms and modulates attention patterns, effectively circumventing these challenges. Grounded in existing normalization frameworks, UnitNorm's effectiveness is demonstrated across diverse time series analysis tasks, including forecasting, classification, and anomaly detection, via a rigorous evaluation on 6 state-of-the-art models and 10 datasets. Notably, UnitNorm shows superior performance, especially in scenarios requiring robust attention mechanisms and contextual comprehension, evidenced by significant improvements by up to a 1.46 decrease in MSE for forecasting, and a 4.89% increase in accuracy for classification. This work not only calls for a reevaluation of normalization strategies in time series Transformers but also sets a new direction for enhancing model performance and stability. The source code is available at https://anonymous.4open.science/r/UnitNorm-5B84.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. TOKON: TOKenization-Optimized Normalization for time series analysis with a large language model

    cs.LG 2025-02 reject novelty 6.0 of 10

    TOKON rounds normalized time series values into integer tokens and adds a 'forecast with care' prompt, reporting RMSE improvements of 7 to 28 percent on two datasets with GPT-4o-mini.

Pith tools