CORTEXA
← Browse
arxivecon.GNstat.MEstat.ML2026-07-14

Forecasting Inflation with Microdata: An Adaptive Machine Learning Approach

Catherine Chen, Chen Gao, Jonathon Hazell, Lihua Lei, Chen Lian

Does microeconomic heterogeneity help to forecast aggregate inflation in a non-stationary environment? We develop a scan test for whether one forecast outperforms another, over an interval with unknown starting point and duration. To exploit any occasional forecasting power that the scan test detects, we design an adaptive machine learning pipeline. We encode the distribution of price changes into a high-dimensional vector, which we combine with a gradient boosted trees algorithm. We then combine this micro forecast with other benchmark forecasts, using an adaptive algorithm that makes use of the micro forecast only when it performs well. We apply the pipeline to UK microdata, with four main results. First, the micro forecast outperforms a univariate benchmark, but only in the volatile period after 2020. Second, the scan test detects periods of micro outperformance, so the micro forecast enters the combined forecast. Third, the combined forecast performs comparably to the univariate benchmark before 2020 and better at every horizon after 2020. Fourth, the value of microdata for the combined forecast materializes after 2020. We conclude that microdata are valuable for forecasting aggregate inflation, but only after large shocks.

View free PDFSource page

Related papers

arxivstat.MEmath.STstat.ML2026-07-03

Outcome-adapted Automatic Debiased Machine Learning

Asger Waagepetersen, Asbjørn Risom, Niels Richard Hansen, Anton Rask Lundborg

Parameters of interest in causal inference, such as treatment or policy effects, can often be expressed as linear functionals of an outcome regression function. Automatic debiased machine learning (AutoDML) is a unified framework for obtaining asymptotically normal estimators of…

View free PDFSource page
arxivstat.MEstat.ML2026-06-27

Doubly cross-fit debiased machine learning of heterogeneous treatment effects under principal stratification

Jiaqi Tong, Fan Li

Principal stratification provides a foundational framework for causal inference with intermediate outcomes by defining causal effects within subpopulations, yet existing work has largely focused on average effects across strata rather than treatment effect heterogeneity within st…

View free PDFSource page
arxivecon.EMmath.STstat.MEstat.ML2026-07-07

Factor-Augmented Machine Learning Panel Regressions

Andrii Babii, Luca Barbaglia, Eric Ghysels, Jonas Striaukas

This paper develops the asymptotic theory for high-dimensional panel data regressions in settings with cross-sectionally dependent errors driven by common shocks. We consider a factor-augmented sparse-group LASSO estimator that combines MIDAS aggregation with latent factors. The…

View free PDFSource page
arxivstat.MLcs.LGstat.ME2026-07-02

Autorelevance function and other feature relevance measures for univariate time series

Julian Cardenas, Jamie Arjona, Pedro Delicado

We propose a model agnostic methodology to measure lag relevance in machine learning forecasting models applied to univariate time series. Particularly, we are working in the context of time series using the frameworks of Ghost variables and Shapley values, together with additive…

View free PDFSource page
arxivstat.MEcs.LGstat.ML2026-07-23

Longitudinal Random Forests for Sparse and Irregular Response Trajectories

Yangsheng Wang, Xiaotian Dai, Haoda Fu, Guifang Fu

Longitudinal studies often collect data at sparse, irregular, and unequally spaced time points. Such heterogeneity is often driven by subject-specific covariates, yet existing methods have been restricted to a scalar endpoint value, completely neglecting the underlying response t…

View free PDFSource page
arxivstat.MEstat.ML2026-06-28

Multi-Source Transfer Learning of Sparse Single-Index Models

Ye Tian

Transfer learning leverages knowledge from related source domains to improve learning in a target domain. Recent theoretical advances cover a broad range of regression settings within (generalized) linear models. Despite their diversity, these methods share two common constraints…

View free PDFSource page