Benchmarking time-series foundation models (Chronos-Bolt, zero-shot) vs. supervised (PatchTST) and classical (seasonal-naive, Croston) baselines on the M5 Walmart dataset, scored with MASE and WQL. No single model dominates: foundation/deep models win on dense SKUs, classical methods win on the intermittent tail.
-
Updated
Jul 19, 2026 - Python