The study

How early can the splits reveal how the race will finish?

In unseen 2024–2026 race editions, the trend model’s median finish-time error was 6.6 minutes at 12.43 mi and 4.0 minutes at 18.64 mi.

This page shows the current checkpoint model’s held-out validation. Use the interactive checkpoint comparison for a selected course and elapsed time; 12.43 mi is before halfway.

861,691 eligible finishes · validated forecast

How we measured it. At each checkpoint, start with elapsed time × 42.195 / distance. The elapsed-only model multiplies this by the training median actual/projected ratio in 30-second-per-mi elapsed-pace bands. The trend model adds the most recent section’s change from the previous section (faster than −2%, within ±2%, or slower than +2%). At 3.11 mi no trend exists.

Full methodology & sources

Hold out the latest three observed calendar years (2024–2026). Fit every factor and the 10th–90th percentile ratio interval on earlier years only. Training cells need at least 100 records; missing trend cells fall back to the pace-band model, then the pooled training model. No future splits, finishing-time groups, identities or supplied ability fields enter a prediction.

Report median absolute error for all three methods, plus observed coverage and median width of the nominal 80% prediction interval and the 90th percentile absolute error. Model selection is fixed before examining these results. A runner may occur in training and test in different years; identities are not used. Results apply to complete eligible finishers and do not predict withdrawals.

Use complete, strictly increasing elapsed checkpoints at 3.11, 6.21, 9.32, 12.43, 15.53, 18.64, 21.75, 24.85 and 26.22 mi. Clock strings must parse as H:MM:SS or M:SS. No missing splits are interpolated.

Remove exact duplicate race records, ignoring database IDs, ingestion timestamps and source URLs. Retain finishes from 90 minutes to 12 hours with every section between 3.22 and 32.19 minutes/mi. These quality filters can exclude genuine unusual performances; the analysis describes this eligible cohort, not every entrant.

Full source records are available in public GitHub Releases. These chart tables require at least 100 eligible observations per cell for estimate reliability. Counts refer to finishes, linked pairs or event observations as specified in that answer.

These are observational results. Fitness changes, intentions, training, selection into the dataset and unmeasured conditions can explain differences. Outcome percentiles describe variation among performances, not confidence intervals or advice about the best strategy.

Apply the reviewed source-quality edition exclusions for this exact export after the timing checks. Known invalid split grids, incomplete ingestion, unreconciled HOLD editions and a selected top-finisher field do not contribute to the analyses or prior benchmarks. Report source exclusions separately; an already invalid timing row is not counted twice. Missing age or recorded gender alone does not exclude an otherwise eligible finish from the overall cohort. Other sparse editions are not declared incomplete merely from their size.

Marathons, years and recorded data. Chart samples may be smaller than the eligible analysis cohort.

Median finish-time error in later race years

Minutes

Even-pace extrapolationLearned from elapsed paceElapsed pace + recent trend
View exact values and sample sizes
Checkpoint (mi)Even-pace extrapolationLearned from elapsed paceElapsed pace + recent trendEven-pace extrapolation: ObservationsLearned from elapsed pace: ObservationsElapsed pace + recent trend: Observations
3.1 mi13.2 min10.5 min10.5 min861,691861,691861,691
6.2 mi12.5 min9.3 min9 min861,691861,691861,691
9.3 mi11.8 min8.4 min7.9 min861,691861,691861,691
12.4 mi10.8 min7.4 min6.6 min861,691861,691861,691
15.5 mi9 min6 min5.2 min861,691861,691861,691
18.6 mi6.6 min4.5 min4 min861,691861,691861,691
21.7 mi3.6 min2.6 min2.2 min861,691861,691861,691
24.9 mi0.7 min0.7 min0.7 min861,691861,691861,691

Observed coverage of the 80% prediction interval

Percent (%)

580.2%
1080.6%
1579.9%
2080.3%
2580.4%
3080.4%
3580%
4080.7%
View exact values and sample sizes
Checkpoint (mi)Finish inside intervalObservations
580.2%861,691
1080.6%861,691
1579.9%861,691
2080.3%861,691
2580.4%861,691
3080.4%861,691
3580%861,691
4080.7%861,691

Prediction width and larger errors

Minutes

Median interval width90th percentile absolute error
5
Median interval width49.1 min
90th percentile absolute error35.6 min
10
Median interval width42.2 min
90th percentile absolute error28.8 min
15
Median interval width35.8 min
90th percentile absolute error24.6 min
20
Median interval width29.9 min
90th percentile absolute error20.9 min
25
Median interval width24.1 min
90th percentile absolute error16.6 min
30
Median interval width18.8 min
90th percentile absolute error11.9 min
35
Median interval width10.3 min
90th percentile absolute error7 min
40
Median interval width2.9 min
90th percentile absolute error2.1 min
View exact values and sample sizes
Checkpoint (mi)Median interval width90th percentile absolute errorMedian interval width: Observations90th percentile absolute error: Observations
549.1 min35.6 min861,691861,691
1042.2 min28.8 min861,691861,691
1535.8 min24.6 min861,691861,691
2029.9 min20.9 min861,691861,691
2524.1 min16.6 min861,691861,691
3018.8 min11.9 min861,691861,691
3510.3 min7 min861,691861,691
402.9 min2.1 min861,691861,691

All research questions