Pith. sign in

REVIEW 7 cited by

Deep Transformer Models for Time Series Forecasting: The Influenza Prevalence Case

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2001.08317 v1 pith:QHFERS3H submitted 2020-01-23 cs.LG stat.ML

classification cs.LGstat.ML
keywords seriestimedataforecastingapproachcaselearningmachine
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In this paper, we present a new approach to time series forecasting. Time series data are prevalent in many scientific and engineering disciplines. Time series forecasting is a crucial task in modeling time series data, and is an important area of machine learning. In this work we developed a novel method that employs Transformer-based machine learning models to forecast time series data. This approach works by leveraging self-attention mechanisms to learn complex patterns and dynamics from time series data. Moreover, it is a generic framework and can be applied to univariate and multivariate time series data, as well as time series embeddings. Using influenza-like illness (ILI) forecasting as a case study, we show that the forecasting results produced by our approach are favorably comparable to the state-of-the-art.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. A Benchmark for Electrical Load Forecasting Across Grid Levels: Time-Series Transformers Outperform Established Methods

    cs.LG 2026-07 conditional novelty 6.0 of 10

    Transformers—especially a standard encoder-decoder—yield the lowest hourly load forecast errors across TSO, low-voltage feeder, and client-level datasets, with 6.6–10.7% error reduction over the best non-Transformer baseline.

  2. NEST: Tackling Dataset-Level Distribution Shifts via Regime-Oriented Mixture-of-Experts

    cs.LG 2026-07 conditional novelty 6.0 of 10

    NEST improves long-term multivariate forecasting under dataset-level distribution shifts by clustering regimes in moment-entropy space and recomposing specialized variate-attention experts via a content-plus-geometry router.

  3. Decomposing the Time Series Forecasting Pipeline: A Modular Approach for Time Series Representation, Information Extraction, and Projection

    cs.AI 2025-07 conditional novelty 6.0 of 10

    REP-Net, a modular pipeline of representation, memory, and projection modules, achieves competitive forecasting accuracy on seven multivariate benchmarks with lower computational cost.

  4. A corrective agentic hybrid RAG and an operations-grounded evaluation for a scientific facility

    physics.acc-ph 2026-07 conditional novelty 5.5 of 10

    A deployed hybrid RAG for APS operations improves vital-nugget recall over BM25 mainly via cross-encoder reranking; graph and corrective loops help only marginally on a 50-question facility benchmark.

  5. Evaluating the Performance of Deep Learning Models in Whole-body Dynamic 3D Posture Prediction During Load-reaching Activities

    cs.CV 2025-11 conditional novelty 5.0 of 10

    Transformer networks with a segment-length-constraining loss predict whole-body 3D posture during load-reaching with ~41 mm RMSE, beating BLSTM on long-horizon recursive prediction.

  6. Diffusion Models for Time Series Forecasting: A Survey

    stat.ML 2025-07 conditional novelty 4.0 of 10

    A survey classifies diffusion-based time series forecasting models into a two-axis taxonomy by conditioning source and integration method.

  7. Self-supervised Learning Method Using Transformer for Multi-dimensional Sensor Data Processing

    cs.LG 2025-05 conditional novelty 3.0 of 10

    An n-dimensional numerical Transformer with linear embedding, bin-based discretization, and parallel output heads improves human activity recognition accuracy by 10-15% over a tokenized vanilla Transformer.

Pith tools