Forecast workbench · Snapshot 06 May 2026

UK Renewable Delivery Outlook

Explore 6,189 active projects, compare aggregate capacity outlooks, rank delivery evidence, map the pipeline and stress-test external conditions. The public two-year signal is a validated relative ranking—not a literal project probability.

Loading 6,189 forecast records…

Validated challenger

Forecast Model Evidence

The richer model is used only where it proved useful: resolving ties inside the stable empirical ordering.

Empirical ranking AUC0.767Stage, technology and annual delivery hazard
AI tie-break ranking AUC0.791+0.025 on the untouched latest cohort
Average precision0.073Up from 0.043 for the empirical ranker
Primary orderingEmpirical survival score
+
AI tie-break evidenceStage age · milestones · capacity revisions · developer history · regional track record · CfD · portfolio pressure
Public outputRelative percentile only

The standalone CatBoost challenger was rejected as a replacement because its overall ranking was weaker. Its within-bucket ordering improved both AUC and average precision, passed the pre-declared ranking gate and is the only AI component promoted to Forecast v2.2.

Product governance

Horizon Release Matrix

Each horizon is released independently. Ranking quality cannot justify publishing an uncalibrated percentage.

HorizonTest sampleRanking AUCObservedRaw meanProbability errorPublic output
2 years8 rolling cohorts 4,01963 observed events 0.791 1.6% 8.9% +33.9% worse Ranking signal
3 years5 rolling cohorts 4,03577 observed events 0.693 1.9% 10.6% +40.2% worse Research only
5 years2 rolling cohorts 1,22289 observed events 0.519 7.3% 9.8% +3.5% worse Withheld

The two-year score is exposed only as a percentile rank across the current forecast universe. Three-year values remain methodology research. Five-year values are removed from the public decision surface.

Held-out evidence

Ranking & Reliability Evidence

Higher ROC-AUC means better ordering. The rate comparison explains why the same scores are not released as probabilities.

Ranking Discrimination

ROC-AUC by horizon · dark marker at the 0.5 random baseline

2-year horizon8 rolling cohorts
0.791
3-year horizon5 rolling cohorts
0.693
5-year horizon2 rolling cohorts
0.519

Predicted and Observed Delivery Rates

Raw mean score versus realised outcome rate on held-out rows

2 years4,019 held-out rows
Raw score mean 8.9% Observed 1.6%
3 years4,035 held-out rows
Raw score mean 10.6% Observed 1.9%
5 years1,222 held-out rows
Raw score mean 9.8% Observed 7.3%

Decision workflow

Delivery-Signal Usage

The signal narrows the research queue; the underlying evidence remains the decision input.

01

Rank

Sort comparable projects by the two-year delivery signal. A score of 80 means stronger model evidence than roughly 80% of the current forecast universe.

02

Verify

Open the project record and inspect its planning stage, dated milestones, authority, CfD evidence and recorded constraints.

03

Stress

Use the Scenario Lab to test how rates, construction costs, grid delay and policy assumptions change portfolio delivery pressure.

Leakage controls

Temporal Validation Controls

Fully Observed Outcomes

An origin is evaluated only when the entire forecast horizon has elapsed. Unresolved recent projects are censored rather than labelled as failures.

Project-Purged Temporal Splits

Test projects are removed from survival-training rows. Features such as developer track record count only information dated before the forecast origin.

Calibration Governance

Log-odds, Platt and isotonic calibration are fitted on earlier out-of-time cohorts and tested on the latest complete cohort. They are not promoted using in-sample fit.

Release Abstention

The public product now abstains: a failed horizon is shown as research-only or withheld instead of publishing the least-bad candidate as a probability.

Feature roadmap

External Driver Register

Project evidence enters the trained challenger now. Market, political and news variables enter only after they have dated histories, source QA and a measurable out-of-time lift.

VariableSourceCurrent treatmentStatus
Prior portfolio size and delivery track recordDeveloper Derived from dated DESNZ REPD snapshots ↗ Trained feature Live in AI challenger
Connection agreement, queue position and target connection dateGrid National Energy System Operator connections registers ↗ Project-match feature after entity-resolution QA Next ingestion
Planning stage, stage age and milestone historyProject delivery DESNZ Renewable Energy Planning Database ↗ Trained feature Live in AI challenger
Capacity revisions, data completeness and observation countProject delivery Derived from dated DESNZ REPD snapshots ↗ Trained feature Live in AI challenger
CfD award, strike price, delivery year and allocation roundSupport DESNZ allocation results and Low Carbon Contracts Company register ↗ Project-match feature and scenario input Award flag live; contract detail next
Infrastructure construction pricesCosts Office for National Statistics Construction Output Price Indices ↗ Manual stress input Live context
Copper, aluminium, iron ore and energy pricesCosts World Bank Commodity Markets Pink Sheet ↗ Technology-weighted scenario input Next ingestion
Bank Rate, gilt yields and credit conditionsFinance Bank of England statistical database ↗ Live context and manual stress input Bank Rate live; yield curve next
Dated policy events, planning reform and technology-specific supportPolicy and news GOV.UK Content API and UK Parliament APIs ↗ Source-linked event features; no free-form LLM probability adjustment Research pipeline
Wholesale electricity price level and volatilityRevenue Elexon Insights market-index price API ↗ Manual stress input until dated history is archived Scenario-ready
Accounts recency, insolvency flags and filing statusDeveloper finance Companies House API ↗ Entity-linked feature for UK companies Research pipeline
Curtailment, negative-price hours and regional constraint exposureSystem value Elexon Insights and NESO data ↗ Regional scenario and portfolio-risk feature Research pipeline
Developer share return, volatility and drawdownMarkets Licensed market-data provider ↗ Publicly listed developers only; never imputed to private firms Research only

Every external feature must be dated, source-linked and available at the forecast origin. Revisions and publication lags are retained so future information cannot leak into historical tests.

Politics, prices and external conditions

Macroeconomic & Policy Scenarios

Bank Rate, wholesale electricity prices, construction and material costs, grid delay, developer-financing stress, policy support and CfD assumptions can materially change delivery conditions. However, 15 independent REPD dates are not enough to estimate credible macroeconomic coefficients. The Scenario Lab therefore exposes them as bounded sensitivities, separate from the trained project ranking. This avoids pretending that thousands of projects sharing one date are thousands of independent observations of politics or inflation.

Public release decision. Promote the empirical ranking with the validated AI tie-breaker; continue withholding literal probabilities. The saved evidence is available in forecast-challenger.json, with the original calibration audit in forecast-audit.json.