- Home
- Case studies
- How a replay works
How a replay works: quarter by quarter, with no hindsight.
A replay asks how early a capital committee could have seen a project's overrun coming, using only what was public at each date. This page sets out the method and walks one invented project through it, so you can see what each step produces. The real megaproject replay, and its results, are shown in a walkthrough after sign-in.
The question
When could a committee have seen the overrun?
Large projects tend to overrun for structural reasons: a single-point estimate at sanction, threats treated as independent when they move together, and in-flight reviews that confirm an overrun once it is hard to act on.
A replay turns that into a question you can test. Suppose a committee had used Capibud at sanction and every quarter after, with nothing but the information public at the time. When would it have seen the overrun coming, how far ahead of the public disclosure, and what could it have done about it?
The method
Six steps, with no look-ahead
- Assemble a point-in-time record. Filings, regulator documents, press reports and market prices, each stored with the date it was published. Nothing enters the replay before its publication date.
- Set a range at sanction. Each project gets a P50 and P80 for cost and finish on its sanction date, from its reference class and the shared factors, before any construction evidence exists.
- Step forward one quarter at a time. For each quarter the replay shows the public estimate, Capibud's P50 and P80, the rules that fired and the documents behind them.
- Test for look-ahead. An automated test checks that no quarter's forecast uses any information dated after that quarter. If it finds one, the replay fails.
- Measure the warning and price the responses. Time in hand runs from the first in-flight warning to the date half of the eventual overrun was public. At the first warning, responses are priced as they would be for a live project, and any of them can be adopted, routed for approval and recorded.
- Score against outcomes. For completed projects, the range set at sanction and each quarter's range are compared with the final cost and finish, and with the sanctioned budget. Misses are reported alongside hits.
An invented example
One illustrative LNG export project, step by step
The invented project is a two-train LNG export plant sanctioned with a budget of $8.0bn. Quarters are counted from sanction.
| Quarter | Public estimate | Capibud P50 | Capibud P80 | What the replay shows |
|---|---|---|---|---|
| Q0 · sanction | 8.0 | 9.1 | 10.6 | Range from the reference class for two-train LNG builds |
| Q4 | 8.0 | 9.2 | 10.7 | Module yard reports a labour shortage |
| Q6 | 8.0 | 9.6 | 11.3 | First warning Module deliveries slip while contractor claims rise |
| Q9 | 8.4 | 9.8 | 11.4 | Owner discloses a first revision |
| Q13 | 9.1 | 10.0 | 11.5 | Half of the eventual overrun is now public |
| Q18 | 10.2 | 10.1 | 11.4 | Latest public estimate |
Illustrative example — figures invented for illustration; not a client or a Capibud demo. P50 is the middle outcome; P80 is the cost the project stays below in four scenarios out of five.
Reading the table
In this invented case the overrun is $2.2bn on an $8.0bn budget, so half of it is public once the estimate reaches $9.1bn, in quarter 13. The first in-flight warning fires in quarter 6. The time in hand is the gap between the two: seven quarters, or about 21 months.
Pricing the responses at the first warning
At quarter 6 the replay prices the responses open to the committee. Three invented ones show the form:
- Bring train 1 forward and let train 2 follow: about $0.5bn of expected value in the example.
- Put the marine works on a fixed-price contract: about $0.2bn, with a narrower range.
- Hedge the foreign-currency share of spend: narrows the range but adds little expected value, about $0.05bn.
In the real product, each response carries its range and can be adopted, routed for approval and recorded, exactly as for a live project.
What the example does not show
An illustration of the method, not evidence for it
- No accuracy claim. We invented both the forecasts and the outcome, so the example cannot show whether the method works. It only shows what each step produces.
- Public information is coarse. A replay sees only what was published. A live deployment reads monthly actuals and contractor data, which a replay cannot see.
- A replay cannot show what people would have done. It prices the options a warning opens; it cannot say whether an owner's team would have taken them.
- The real results are behind sign-in. The megaproject replay, its projects, its hits and its misses are shown in a guided walkthrough. Our published forecast tests are on the case studies page.
What this means for you
Ask for the same test on your own history
Ranges at sanction
Your own completed projects can tell you whether a reference-class range at sanction would have caught your overruns before construction began.
Warnings in flight
Public signals are thin. With monthly actuals and contractor dates, Execute has far more to work with.
Priced responses
A warning is only useful if it arrives with options priced in money and routed to someone who can approve them.
Related
Forecast tests, use cases and industries
Our published forecast tests are on the case studies page. The industry packs below hold the kind of reference classes a replay starts from.
Tests 2 and 3: synthetic backtest and market data
77% and 84% P80 coverage under our generator, and the four generators built to break it.
See the testsA data-centre developer: keeping phase 1 on track
Execute on monthly actuals rather than public filings.
Read the caseEvaluate: sanction and stage-gates
Ranges at sanction, the value of waiting and the shift that would flip the case.
Read the guideRequest the megaproject replay walkthrough.
The real replay and its results are shown after sign-in. We walk you through the method, the projects and the misses.