arXiv ScienceSearch

arXiv · 2609.20554

Does Training on Future Data Pay? Look-Ahead Bias in Forecasting with Pretrained Models

Abstract

We examine whether post-origin training information inflates the measured accuracy and economic value of financial forecasts. We evaluate five sets of financial time-series foundation models, each comprising independently trained annual vintages under U.S., global, and factor-augmented training environments, across 14 equity markets and four forecast horizons. Rolling comparisons vary the annual vintage for a fixed forecast; fixed-vintage comparisons hold the vintage fixed as target windows move across its training cutoff. Each alternative forecast is paired with an origin-aligned point-in-time (PIT) benchmark using identical numerical histories and inference protocols. In the U.S.-trained reference environment, post-origin vintages materially revise informative PIT forecasts but generally reduce accuracy in both designs. Pooled rolling comparisons yield higher mean squared forecast errors in 18 of 20 U.S. model-set-horizon combinations. The origin-crossing update also performs worse on average than an equally long pre-origin update. Under a common constrained allocation rule using one-month forecasts, median exposed-minus-PIT differences in annualized certainty-equivalent returns are -1.77 percentage points in the United States and -2.14 points internationally. Global and factor-augmented training produce more mixed predictive effects. An exact squared-error decomposition shows that revisions improve accuracy when their error-correcting benefit exceeds their mean squared magnitude; under U.S. training, alignment with PIT errors generally falls short of this requirement. Temporal exposure therefore establishes an information-set violation, not sufficient evidence of inflated predictive accuracy or investor value.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Haiqiang Chen, Li Chen, Yunlong Chen, Difang Huang, Bo Zhang. 2026-09-17. Does Training on Future Data Pay? Look-Ahead Bias in Forecasting with Pretrained Models. https://arxiv.org/abs/2609.20554

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

The time interpretation of expected utility theory

Ergodicity economics is a new branch of economic theory that notes the conceptual difference between time averages and expectation values, which coincide only for ergodic observables. It postulates that individual agents maximise the time average growth rate of wealth, known widely as growth optimality. This contrasts with the dominant behavioural model in economics, expected utility theory, in which agents maximise expectation values of changes in psychologically transformed wealth. Historically, growth optimality was explored for additive and multiplicative gambles. Here we apply it to a general class of wealth dynamics, extending the range of economic situations where it may be used. Moreover, we show a correspondence between growth optimality and expected utility theory, in which the ergodicity transformation in the former is identified as the utility function in the latter. This correspondence offers a theoretical basis for choosing utility functions and predicts that wealth dynamics are strong determinants of risk preferences.

econ.GN

Monetary Regimes and Trade before the Classical Gold Standard: Evidence from the Latin Monetary Union

This paper reexamines the trade effects of the Latin Monetary Union (LMU), a 19th century agreement to standardize gold and silver coinage among several European countries. The LMU provides a useful setting for studying whether monetary arrangements fostered trade before the classical gold standard, when gold, silver, bimetallic, and paper regimes coexisted. Because some countries already shared other monetary standards, treating all non-member pairs as a single control group mixes pairs with and without alternative forms of monetary coordination. I classify pairs by standard and estimate the LMU effect relative to pairs without a common standard, bringing the comparison closer to those used in the literature on the gold standard and contemporary currency unions. The results suggest that the LMU increased trade between its members by approximately 30\% during its early years, when bimetallism was still credible. These effects subsequently faded, converging to zero by the end of the 1870s. More broadly, these findings also highlight the importance of accounting for the existing monetary regimes when estimating the trade effects of other international policies.

econ.GN

Access to Live AI Advice and Behavior Under Risk: An Incentivized Experiment

Generative AI has become an everyday advisor, and the systems people consult are live and interactive, not pre-scripted. We ask whether access to such a system changes behavior under risk. In an incentivized experiment (N = 158), participants made lottery choices with an optional decision aid presented as a conventional pre-written tool, a live one-shot AI, or a live interactive AI they could query, with information format held equivalent across conditions. Risk preferences are elicited via DOSE. We find no evidence that access to a live AI advisor changes risk aversion.

econ.GN