Forecasting Benchmarks

Latest papers 147

All topics
CardsList
  1. Are We Really Benchmarking Forecasting Models? The Impact of Preprocessing on Time Series Performance

    Oct 6, 2026Guilherme Afonso Galindo Padilha, Paulo Salgado Gomes de Mattos Neto, Rafael Menelau Oliveira e CruzTime Series ForecastingForecasting Benchmarks

  2. Machine Learning for German Redispatch Forecasting under Data Delays and Temporal Distribution Shift

    Oct 6, 2026Faraz Shamim, Faris ShamimDistribution Shift RobustnessTime Series Forecasting

  3. HouseholdBench: Evaluating Large Language Models as Predictors of Household Economic Behavior

    Oct 6, 2026Jin Huang, Diego Ferreras Garrucho, Yutong Xie +4LLM EvaluationLLM Fine-Tuning

  4. Towards Explainable Benchmarking for Data-driven Post-Wildfire Debris Flow Prediction

    Oct 5, 2026Zhisheng Qi, Li Zhu, Utkarsh Sahu +3Perturbation-Based Feature AttributionForecasting Benchmarks

  5. Benchmarking Time Series Foundation Models for Load Forecasting Under Covariate Uncertainty

    Oct 5, 2026Tomas Kaljevic, Ivan Arzola, Yu ZhangZero-Shot Time Series ForecastingTime Series Forecasting

  6. Do Your Own Research: Learning to Forecast by Learning to Search

    Oct 1, 2026Yusuf Afifi, Artur Kiulian, Anton Polishko +3Agentic RLForecasting Benchmarks

  7. On the Divergence of Accuracy and Mechanism Consistency in Time Series World Models

    Oct 1, 2026Haochen Zhang, Jiaheng Guo, Zhen Xu +4World ModelsTime Series Forecasting

  8. The Nixtlaverse: An Open-Source Ecosystem for Forecasting

    Sep 30, 2026Olivier Sprangers, Max Mergenthaler Canseco, Marco Peixeiro +8Time Series ForecastingForecasting Benchmarks

  9. M2^2Weather: A Benchmark for Joint Multi-Station and Multi-Variable Weather Forecasting

    Sep 30, 2026Rongwen Li, Xiao Wang, Mingyang Wang +4Multivariate Time Series ForecastingForecasting Benchmarks

  10. PDE-OBS: Controlled Evaluation Across Observation Patterns

    Sep 29, 2026Ruichen Xu, Siyao Wang, Fang Wan +7Benchmark DesignPDE Surrogate Modeling

  11. BITS: Rethinking Fair and Comprehensive Evaluation for Irregular Time Series Forecasting

    Sep 27, 2026Kangjia Yan, Linfeng Wang, Tianen Shen +7Irregular Time-Series ModelingForecasting Benchmarks

  12. Forecast-Dojo: Replayable Environments for Benchmarking and Training LLM Forecasting Agents

    Sep 24, 2026Liqin Ye, Haorui Wang, Fardin Ahmed +8LLM Agent EvaluationForecasting Benchmarks

  13. fable.intermittent: benchmarking probabilistic forecasting methods for intermittent time series

    Sep 23, 2026Stefano Damato, Lorenzo Zambon, Giorgio Corani +1Time Series ForecastingForecasting Benchmarks

  14. Overlay_dx - Automating forecasting evaluation

    Sep 21, 2026Long Ngo, Mohammed Amine Chamli, Jonathan Rivalan +1Time Series ForecastingForecasting Benchmarks

  15. WPBench: A Comprehensive Benchmark for Wind Power Forecasting

    Sep 21, 2026Yuhan Zhu, Jilin Hu, Xinying Cai +7Multivariate Time Series ForecastingBenchmark Design

  16. How Good Are Time-Series Foundation Models for Pedestrian Crowd Count Forecasting? A Cross-Dataset Comparative Study

    Sep 14, 2026Theivaprakasham Hari, Ziteng Li, Yanan Xin +2Crowd CountingTime Series Forecasting

  17. MoveBench: A Benchmark for Global-Scale Wildlife Movement Forecasting

    Sep 14, 2026Justin Kay, Shir Bar, Ellen O. Aikens +28Forecasting BenchmarksProbabilistic Forecasting

  18. Vishing-Tactics-Bench: Forecasting Exploitation Trajectories in Voice Phishing Calls

    Sep 7, 2026Jeongmin Lee, Dongmyung Sul, Seung Yun +1Forecasting Benchmarks

  19. Can Large Language Models Forecast What Researchers Study Next?

    Sep 1, 2026Fenghai Li, Zihan Tang, Haofei Yu +2LLM EvaluationForecasting Benchmarks

  20. Can LLMs Take the Pulse of the Economy? A Real-Time Evaluation of LLM Nowcasts on Macroeconomic Indicators

    Aug 31, 2026Xinyue Zhao, Ruiyi Zhang, Liqin Ye +3Time Series ForecastingLLM Agent Evaluation

  21. A Critical Audit of Spatiotemporal Forecasting Benchmark Datasets and Models

    Aug 21, 2026Kenneth Martin, Simon Heilig, Asja Fischer +3Graph Neural NetworksTime Series Forecasting

  22. Long-Horizon Forecasting of Complete Financial Statements with Forma

    Aug 11, 2026Travis L. Johnson, Jiannan Jiang, Soumyabrata Chaudhuri +3Long-Term Time Series ForecastingFinancial Forecasting

  23. Benchmarking Time Series Generation Methods for Privacy-Preserving Forecasting

    Aug 11, 2026Luis Amorim, Vitor Cerqueira, Moises Santos +2Synthetic Data GenerationTime Series Forecasting

  24. Evaluating Generative Time-Series Models on Data with Point Masses

    Aug 10, 2026Jian XuTime Series ForecastingTime Series Generation

  25. Machine Learning and ARIMA Model Averaging for Adaptive Public Health Forecasting: Comparative Evaluation and an Ontario COVID-19 Case Study

    Aug 7, 2026Yushu Zou, Ye Li, Johra Moosa +3Epidemic ForecastingGradient-Boosted Decision Trees

  26. Enhancing Anomaly Resilience in Research Networks: A Large-Scale Forecasting Benchmark for Dynamic Security Baselining

    Aug 6, 2026Mohammad Arafath Uddin Shariff, Byrav RamamurthyNetwork Intrusion DetectionForecasting Benchmarks

  27. TIDE: A Physically Diverse 3D Turbulence Benchmark Dataset for Advancing Scientific Machine Learning

    Aug 4, 2026Yilong Dai, Yiming Sun, Yiheng Chen +4Forecasting BenchmarksScientific ML

  28. FinVerse: Financial Time-Series Benchmark

    Aug 4, 2026Jaehoon Lee, Jun Seo, Seunghan Lee +9Benchmark DesignFinancial Forecasting

  29. Benchmarking ConvLSTM for One-Day-Ahead IMDAA Rainfall-Field Prediction across Four Indian Cities

    Jul 29, 2026Tanmay Ghosh, Shaurabh Anand, Rakesh Gomaji Nannewar +1Convolutional RNNsPrecipitation Forecasting

  30. LLM-SoccerArena: Benchmarking LLMs on Real-World Predictions in Sports

    Jul 27, 2026Jonas Schröder, Jonas Schweisthal, Oliver Müller +2LLM EvaluationForecasting Benchmarks

  31. Variational Quantum Conditional Boltzmann Machines for Time-Series Forecasting: Architectures, Symmetric Hyperparameter Evaluation, and a Nonlinear Benchmark

    Jul 27, 2026Gerhard Hellstern, Danyal Maheshwari, Martin Zaefferer +2Time Series ForecastingForecasting Benchmarks

  32. GlucoTune: A Unified Framework for Blood Glucose Preprocessing, Forecasting, and Benchmarking in Diabetes

    Jul 23, 2026Davide Marelli, Giorgia Rigamonti, Mirko Paolo Barbato +1Type 1 DiabetesTime Series Forecasting

  33. WorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football Forecasting

    Jul 20, 2026Zhaokai Wang, Tianlin Gui, Jiayuan Rao +3LLM EvaluationForecasting Benchmarks

  34. FIFA World Cup 2026 as a Contamination-Free Benchmark for LLM Forecasting Agents: Four Models, a Bookmaker, and 104 Matches

    Jul 20, 2026Jiacheng Ding, Cong Guo, Jason XuLLM Agent EvaluationForecasting Benchmarks

  35. A Benchmark for Electrical Load Forecasting Across Grid Levels: Time-Series Transformers Outperform Established Methods

    Jul 17, 2026Matthias Hertel, Sebastian Pütz, Jonathan Kolar +3Time Series ForecastingForecasting Benchmarks

  36. Hindcast: Replaying Prediction Markets to Evaluate LLM Forecasters

    Jul 15, 2026Xiao Ye, Jacob Dineen, Evan Zhu +3LLM EvaluationForecasting Benchmarks

  37. Institutional Equity Holdings Prediction Using Node Affinities of Dynamic Graphs

    Jul 13, 2026Emad Izadifar, Zahed RahmatiTemporal GNNsLink Prediction

  38. Evaluating the Generalizability of Foundation Models for Extreme Environmental Events: Case Study of California Wildfire PM2.5

    Jul 8, 2026Yongcan Huang, Li Jiang, Ze Yu LiuWildfire ForecastingOOD Generalization

  39. Rethinking Multimodal Time-Series Forecasting Evaluation

    Jul 8, 2026Haoxin Liu, Yichen Zhou, Rajat Sen +2Multivariate Time Series ForecastingZero-Shot Time Series Forecasting

  40. When Do Foundation Models Pay Off? A Break-Even Analysis of Pretrained Time Series Forecasters

    Jul 6, 2026Nicholas Tan Jerome, Frank SimonZero-Shot Time Series ForecastingTime Series Forecasting

  41. The Granularity Paradox: How Temporal Disaggregation Inflates In-Sample Fit and Compounds Out-of-Sample Error

    Jul 5, 2026Hugo MoreiraTime Series ForecastingForecasting Benchmarks

  42. Evaluating Time Series Foundation Models for Electricity Price Forecasting: Contamination Risk, Distributional Shifts, and Covariate Dependence

    Jul 2, 2026Zhenghua Pan, Ahmed Aziz EzzatZero-Shot Time Series ForecastingElectricity Price Forecasting

  43. Timesynth: A Temporal Fidelity Framework for Health Signal Digital Twins

    Jul 1, 2026Md Rakibul Haque, Shireen Elhabian, Warren Woodrich PettineTime Series ForecastingForecasting Benchmarks

  44. Diversity is the Strength of the AI Crowd

    Jun 29, 2026Matthew Aitchison, Scott Jeen, Toby Shevlane +1Forecasting BenchmarksLanguage Model Ensembles