The study

How early can the splits reveal how the race will finish?

In unseen 2024–2026 race editions, the trend model’s median finish-time error was 6.5 minutes at 12.43 mi and 4.0 minutes at 18.64 mi.

This page shows the current checkpoint model’s held-out validation. Use the interactive checkpoint comparison for a selected course and elapsed time; 12.43 mi is before halfway.

869,761 eligible finishes · validated forecast

How we measured it. At each checkpoint, start with elapsed time × 42.195 / distance. The elapsed-only model multiplies this by the training median actual/projected ratio in 30-second-per-mi elapsed-pace bands. The trend model adds the most recent section’s change from the previous section (faster than −2%, within ±2%, or slower than +2%). At 3.11 mi no trend exists.

Full methodology & sources

Hold out the latest three observed calendar years (2024–2026). Fit every factor and the 10th–90th percentile ratio interval on earlier years only. Training cells need at least 100 records; missing trend cells fall back to the pace-band model, then the pooled training model. No future splits, finishing-time groups, identities or supplied ability fields enter a prediction.

Report median absolute error for all three methods, plus observed coverage and median width of the nominal 80% prediction interval and the 90th percentile absolute error. Model selection is fixed before examining these results. A runner may occur in training and test in different years; identities are not used. Results apply to complete eligible finishers and do not predict withdrawals.

Use complete, strictly increasing elapsed checkpoints at 3.11, 6.21, 9.32, 12.43, 15.53, 18.64, 21.75, 24.85 and 26.22 mi. Clock strings must parse as H:MM:SS or M:SS. No missing splits are interpolated.

Remove exact duplicate race records, ignoring database IDs, ingestion timestamps and source URLs. Retain finishes from 90 minutes to 12 hours with every section between 3.22 and 32.19 minutes/mi. These quality filters can exclude genuine unusual performances; the analysis describes this eligible cohort, not every entrant.

Full source records are available in public GitHub Releases. These chart tables require at least 100 eligible observations per cell for estimate reliability. Counts refer to finishes, linked pairs or event observations as specified in that answer.

These are observational results. Fitness changes, intentions, training, selection into the dataset and unmeasured conditions can explain differences. Outcome percentiles describe variation among performances, not confidence intervals or advice about the best strategy.

Apply the reviewed source-quality edition exclusions for this exact export after the timing checks. Known invalid split grids, incomplete ingestion, unreconciled HOLD editions and a selected top-finisher field do not contribute to the analyses or prior benchmarks. Report source exclusions separately; an already invalid timing row is not counted twice. Missing age or recorded gender alone does not exclude an otherwise eligible finish from the overall cohort. Other sparse editions are not declared incomplete merely from their size.

Marathons, years and recorded data. Chart samples may be smaller than the eligible analysis cohort.

Median finish-time error in later race years

Minutes

Even-pace extrapolationLearned from elapsed paceElapsed pace + recent trend
View exact values and sample sizes
Checkpoint (mi)Even-pace extrapolationLearned from elapsed paceElapsed pace + recent trendEven-pace extrapolation: ObservationsLearned from elapsed pace: ObservationsElapsed pace + recent trend: Observations
3.1 mi13.1 min10.5 min10.5 min869,761869,761869,761
6.2 mi12.5 min9.2 min9 min869,761869,761869,761
9.3 mi11.8 min8.3 min7.9 min869,761869,761869,761
12.4 mi10.8 min7.3 min6.5 min869,761869,761869,761
15.5 mi9 min6 min5.2 min869,761869,761869,761
18.6 mi6.6 min4.4 min4 min869,761869,761869,761
21.7 mi3.6 min2.6 min2.2 min869,761869,761869,761
24.9 mi0.7 min0.7 min0.7 min869,761869,761869,761

Observed coverage of the 80% prediction interval

Percent (%)

580.4%
1080.7%
1580.2%
2080.7%
2580.8%
3080.8%
3580.6%
4081.7%
View exact values and sample sizes
Checkpoint (mi)Finish inside intervalObservations
580.4%869,761
1080.7%869,761
1580.2%869,761
2080.7%869,761
2580.8%869,761
3080.8%869,761
3580.6%869,761
4081.7%869,761

Prediction width and larger errors

Minutes

Median interval width90th percentile absolute error
5
Median interval width48.8 min
90th percentile absolute error36.3 min
10
Median interval width41.8 min
90th percentile absolute error28.9 min
15
Median interval width35.6 min
90th percentile absolute error24.6 min
20
Median interval width29.7 min
90th percentile absolute error20.9 min
25
Median interval width24 min
90th percentile absolute error16.6 min
30
Median interval width18.8 min
90th percentile absolute error11.9 min
35
Median interval width10.3 min
90th percentile absolute error7 min
40
Median interval width2.9 min
90th percentile absolute error2.1 min
View exact values and sample sizes
Checkpoint (mi)Median interval width90th percentile absolute errorMedian interval width: Observations90th percentile absolute error: Observations
548.8 min36.3 min869,761869,761
1041.8 min28.9 min869,761869,761
1535.6 min24.6 min869,761869,761
2029.7 min20.9 min869,761869,761
2524 min16.6 min869,761869,761
3018.8 min11.9 min869,761869,761
3510.3 min7 min869,761869,761
402.9 min2.1 min869,761869,761

All research questions