SearcharxivSearch

arXiv · 2609.20554

Does Training on Future Data Pay? Look-Ahead Bias in Forecasting with Pretrained Models

Abstract

We examine whether post-origin training information inflates the measured accuracy and economic value of financial forecasts. We evaluate five sets of financial time-series foundation models, each comprising independently trained annual vintages under U.S., global, and factor-augmented training environments, across 14 equity markets and four forecast horizons. Rolling comparisons vary the annual vintage for a fixed forecast; fixed-vintage comparisons hold the vintage fixed as target windows move across its training cutoff. Each alternative forecast is paired with an origin-aligned point-in-time (PIT) benchmark using identical numerical histories and inference protocols. In the U.S.-trained reference environment, post-origin vintages materially revise informative PIT forecasts but generally reduce accuracy in both designs. Pooled rolling comparisons yield higher mean squared forecast errors in 18 of 20 U.S. model-set-horizon combinations. The origin-crossing update also performs worse on average than an equally long pre-origin update. Under a common constrained allocation rule using one-month forecasts, median exposed-minus-PIT differences in annualized certainty-equivalent returns are -1.77 percentage points in the United States and -2.14 points internationally. Global and factor-augmented training produce more mixed predictive effects. An exact squared-error decomposition shows that revisions improve accuracy when their error-correcting benefit exceeds their mean squared magnitude; under U.S. training, alignment with PIT errors generally falls short of this requirement. Temporal exposure therefore establishes an information-set violation, not sufficient evidence of inflated predictive accuracy or investor value.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Haiqiang Chen, Li Chen, Yunlong Chen, Difang Huang, Bo Zhang. 2026-09-17. Does Training on Future Data Pay? Look-Ahead Bias in Forecasting with Pretrained Models. https://arxiv.org/abs/2609.20554

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Computing Endogenous Transformations in Processing Networks: A Dynamic Calibration Approach

Understanding how supply chains endogenously transform requires a parametric model of processing networks with non-neutral substitution elasticities. While the Cascaded CES production function provides a rigorous framework, dynamically calibrating its structural parameters from time-series data constitutes a highly non-convex inverse optimization problem. Since enforcing strict microeconomic concavity renders standard monolithic approaches computationally intractable, we propose a novel structure-exploiting algorithm to bypass this limitation. By leveraging the physical upstreamness topology of the network, our hybrid heuristic effectively breaks the curse of dimensionality inherent in economywide processing networks. Applying this framework to U.S. time-series data, we provide a scalable computational engine to fully endogenize complex supply-chain transformations, ultimately uncovering the elastic origins of asymmetric macroeconomic tail risks.

econ.GN

Access to Live AI Advice and Behavior Under Risk: An Incentivized Experiment

Generative AI has become an everyday advisor, and the systems people consult are live and interactive, not pre-scripted. We ask whether access to such a system changes behavior under risk. In an incentivized experiment (N = 158), participants made lottery choices with an optional decision aid presented as a conventional pre-written tool, a live one-shot AI, or a live interactive AI they could query, with information format held equivalent across conditions. Risk preferences are elicited via DOSE. We find no evidence that access to a live AI advisor changes risk aversion.

econ.GN

Screening Out the Needy: The Effects of SNAP Work Requirements

We examine the effectiveness of work requirements as a screening device in the Supplemental Nutrition Assistance Program (SNAP). Work requirements for "able-bodied adults without dependents" were suspended after the Great Recession and gradually reinstated across counties and states in the 2010s. Using linked administrative SNAP and employment data from five states and a triple-differences design, we find that work requirements reduce SNAP participation by seven percent without increasing labor supply and disproportionately screen out low-income individuals. We develop a welfare framework to interpret these results and find that the social costs of work requirements exceed budget savings.

econ.GN