Pith. sign in

REVIEW 21 cited by

N-BEATS: Neural basis expansion analysis for interpretable time series forecasting

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1905.10437 v4 pith:472JKLW2 submitted 2019-05-24 cs.LG stat.ML

classification cs.LGstat.ML
keywords architecturedatasetsdeepseriesforecastinginterpretableneuraltime
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We focus on solving the univariate times series point forecasting problem using deep learning. We propose a deep neural architecture based on backward and forward residual links and a very deep stack of fully-connected layers. The architecture has a number of desirable properties, being interpretable, applicable without modification to a wide array of target domains, and fast to train. We test the proposed architecture on several well-known datasets, including M3, M4 and TOURISM competition datasets containing time series from diverse domains. We demonstrate state-of-the-art performance for two configurations of N-BEATS for all the datasets, improving forecast accuracy by 11% over a statistical benchmark and by 3% over last year's winner of the M4 competition, a domain-adjusted hand-crafted hybrid between neural network and statistical time series models. The first configuration of our model does not employ any time-series-specific components and its performance on heterogeneous datasets strongly suggests that, contrarily to received wisdom, deep learning primitives such as residual blocks are by themselves sufficient to solve a wide range of forecasting problems. Finally, we demonstrate how the proposed architecture can be augmented to provide outputs that are interpretable without considerable loss in accuracy.

Discussion (0). Sign in to comment.

Forward citations

Cited by 21 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 160 citations worldwide. Full citation record

  1. Multi-Source Dynamic Graph Learning for Compound-Flood Forecasting in Managed Coastal Systems

    cs.LG 2026-08 conditional novelty 6.0 of 10

    An anchored forecaster that adds bounded, regime-gated corrections from a dynamic multi-station graph to a local temporal forecast improves sustained high-water plateau prediction in South Florida without harming rout...

  2. Dual-Prototype Disentanglement: A Context-Aware Enhancement Framework for Time Series Forecasting

    cs.LG 2026-01 conditional novelty 6.0 of 10

    A model-agnostic module that retrieves common and rare prototype patterns improves forecasting error on many standard benchmarks, but not on all reported cases.

  3. Fremer: Lightweight and Effective Frequency Transformer for Workload Forecasting in Cloud Services

    cs.LG 2025-07 conditional novelty 6.0 of 10

    Fremer forecasts cloud workloads by aligning frequency spectra via linear padding, filtering noise, and attending over frequency combinations.

  4. Leveraging External Factors in Household-Level Electrical Consumption Forecasting using Hypernetworks

    cs.LG 2025-06 conditional novelty 6.0 of 10

    A hypernetwork that generates per-household linear forecast weights is the only global model in the test that improves its error when weather, holiday, and football-event data are added, beating other global models on...

  5. GatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series Forecasting

    cs.LG 2026-07 conditional novelty 5.5 of 10

    Adaptive soft routing among three complementary linear bases via a channel-horizon-phase gate yields competitive multivariate forecasting accuracy with a small, interpretable model.

  6. Enhancing Irregular Time Series Forecasting with Continuous-Time Modeling Framework

    cs.LG 2026-07 conditional novelty 5.0 of 10

    WrapFlow combines continuous-time event/gap tokenization with simulation-free residual flow matching on a Transformer to improve irregular multivariate time-series forecasting.

  7. Hierarchical Spatio-Temporal Transformer for Coherent Emergency Department Forecasting

    cs.LG 2026-07 conditional novelty 5.0 of 10

    A top-down hierarchical Transformer with a coherence loss jointly forecasts Portuguese ED demand at 81 hospitals, 5 regions, and national level, cutting mean WAPE ~32% versus the best non-hierarchical deep baseline wh...

  8. A Predict-then-Correct Loop Based on Few-Shot Continuous Contextual Bandit for Demand Forecasting

    cs.LG 2026-07 conditional novelty 5.0 of 10

    A contextual-bandit correction layer with few-shot masked updates improves ML demand forecasts by 3.7–14.9% and cuts inventory costs in two retail datasets.

  9. OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning

    cs.CV 2025-07 conditional novelty 5.0 of 10

    A shared-backbone transformer with pairwise modality training reports top results across 25 datasets spanning 12 modalities.

  10. Teaching Time Series to See and Speak: Forecasting with Aligned Visual and Textual Perspectives

    cs.LG 2025-06 reject novelty 5.0 of 10

    TimesCLIP aligns image-based and text-based views of the same time series via contrastive learning to improve forecasting accuracy on several benchmarks, but the full multimodal model is not used on two of the six lon...

  11. Probabilistic Forecasting for Building Energy Systems using Time-Series Foundation Models

    cs.LG 2025-05 conditional novelty 5.0 of 10

    Fine-tuned time-series foundation models, especially Chronos with LoRA, outperform trained-from-scratch deep forecasters on multi-signal building energy forecasting with limited data.

  12. Forecasting Multivariate Urban Data via Decomposition and Spatio-Temporal Graph Analysis

    cs.LG 2025-05 conditional novelty 5.0 of 10

    A graph neural network that learns one dependency graph per decomposed time-series component improves long-term urban forecasts by a few percent over baselines.

  13. Contextual Deconvolution for Variance-Stable Demand Sensing: Kernel-Modulated Operators in Promotional Retail

    cs.LG 2026-07 conditional novelty 4.0 of 10

    A smooth-baseline-plus-sparse-shock decomposition lowers forecast variance and safety stock but increases stockout costs, reducing total inventory cost only when holding costs exceed ~20% of stockout costs.

  14. Towards Reliable Zero-Shot Crowd Forecasting: Evaluating Time Series Foundation Models for Special Event Pedestrian Forecasting

    cs.LG 2026-07 conditional novelty 4.0 of 10

    Zero-shot time-series foundation models, especially Chronos-2 with increasing context, can produce probabilistically reliable crowd-flow forecasts up to about 30–45 minutes ahead for a five-day special event, though t...

  15. Cross-device Zero-shot Label Transfer via Alignment of Time Series Foundation Model Embeddings

    eess.SP 2025-08 reject novelty 4.0 of 10

    A framework using adversarial alignment of frozen time-series foundation model embeddings transfers labels to a simulated target domain, but the target is synthetic and no real consumer device data is tested.

  16. N-BEATS-MOE: N-BEATS with a Mixture-of-Experts Layer for Heterogeneous Time Series Forecasting

    cs.LG 2025-08 conditional novelty 4.0 of 10

    Adding a gating network on top of N-BEATS block outputs gives modest SMAPE improvements on some heterogeneous benchmark series, but the gains are small and not statistically validated.

  17. Towards Measuring and Modeling Geometric Structures in Time Series Forecasting via Image Modality

    cs.CV 2025-07 conditional novelty 4.0 of 10

    A new image-based similarity metric (TGSI) and a three-part training loss (SATL) that together aim to improve the geometric fidelity of time series forecasts.

  18. MamNet: A Novel Hybrid Model for Time-Series Forecasting and Frequency Pattern Analysis in Network Traffic

    cs.LG 2025-06 reject novelty 4.0 of 10

    MamNet combines Mamba time-domain modeling with Fourier frequency features and claims 2-4% gains over five baselines on UNSW-NB15 and CAIDA.

  19. FAF: A Feature-Adaptive Framework for Few-Shot Time Series Forecasting

    cs.LG 2025-06 reject novelty 4.0 of 10

    A feature-adaptive meta-learning framework for few-shot time series forecasting reports large gains, but its evaluation uses one to nine test tasks per dataset, lacks error bars, and contains numerical and preprocessi...

  20. Wavelet-Enhanced Neural ODE and Graph Attention for Interpretable Energy Forecasting

    cs.LG 2025-07 reject novelty 3.0 of 10

    The proposed hybrid forecaster claims consistent superiority on ETT and EIA energy datasets, but the benchmark setup and reported numbers do not support that claim.

  21. Entanglement for Pattern Learning in Temporal Data with Logarithmic Complexity: Benchmarking on IBM Quantum Hardware

    quant-ph 2025-05 reject novelty 3.0 of 10

    A fixed, untrained 10-qubit circuit with CNOT entanglement forecasts Z500 weather data with MSE near classical AR models, but the claimed logarithmic training complexity is an artifact of defining the window size as l...

Pith tools