Field note · December 15, 2024
Forward-chaining split, in short
Validate the way you will deploy — train on the past, predict the future, and leave the gap your real prediction lag imposes.
Added a note to Forward-chaining split today, which is a good excuse to say the short version here.
Validate the way you will deploy — train on the past, predict the future, and leave the gap your real prediction lag imposes.
What makes it a pattern rather than a tip is that the wrong version is the one you write naturally. It reads correctly, it runs, and it returns something. The failure is in the result, not in the execution — which means the only defence is recognising the shape before you are in it.
Two datasets on this site have the shape built in: grid-energy-load and sensor-telemetry. Both are small enough to run the broken version, see the number, then run the corrected one and see it change.
It lives under modelling because the model is not where the mistake is. The mistake is upstream, in what the training rows knew.
The long-form treatment is in the course (validation-that-matches-deployment); the pattern page is the version to read at 4pm with a query open.
The pattern page has the version that holds and the version that looks right and is not, side by side. Reading them together is the point; either one alone is just code.