Technical documentation

Methodology

How the British Resilience Index is constructed, what it measures, and what it does not.

What this index measures

The British Resilience Index is a non-partisan data project measuring stress across key systems of national life. It does not assign blame to political parties. It tracks whether core systems are improving, deteriorating or entering dangerous levels of strain.

The index combines 11 domains, each measuring a different national system, weighted by their estimated contribution to systemic resilience. Each domain score is the median of its component metric stress scores. The national score is a weighted sum of domain medians.

Current release: August 2026 · National score: 52/100 · Band: Fragile

What this index does NOT measure

The index does not:

  • Attribute blame to any political party, government, or named individual.
  • Predict policy outcomes or future conditions.
  • Model cascading system failures or tipping points.
  • Provide individual-level data on any person.
  • Measure cultural, moral, or ideological conditions.
  • Claim to know at what score a society becomes ungovernable.

It tracks whether measured systems are under more or less stress than in previous periods. The interpretation of why conditions have changed is left to the reader.

What the score means

The index produces a single 0 to 100 reading. Higher scores indicate greater measured stress across national systems. The score is grouped into five bands:

  • Stable (0–24): measured stress is low and within the normal historical range.
  • Strained (25–49): under noticeable pressure, but not yet at dangerous levels.
  • Fragile (50–74): significant strain; several systems are weak and vulnerable to shocks.
  • Critical (75–89): severe strain; the system is close to failing to function normally.
  • Crisis (90–100): extreme strain, at or near the point of breakdown.

The score is not a prediction of societal collapse; it tracks whether measured systems are under more or less stress than in previous periods.

Domain model and weights

The index currently comprises 11 domains. Weights sum to 100. Each domain is scored as the median of its component metric stress scores. Weights represent the estimated share of systemic national resilience carried by each domain. These weights are currently set by editorial judgement and will be reviewed formally.

Metric scoring (0.6 / 0.3 / 0.1 formula)

Each metric produces a stress score 0–100, composed of three sub-scores:

metric_stress = clamp(0.6 × level_score + 0.3 × trend_score + 0.1 × volatility_score, 0, 100)
  • Level score (60%): where the current value sits relative to the metric’s own historical distribution. Metrics with ≥ 8 verified historical points use historical-percentile normalisation: a value at the 90th percentile of its own history scores ~90; the median scores ~50. Direction of badness is automatically accounted for. Shorter series fall back to provisional best/worst bounds, flagged in the metadata as the weaker method. Currently 41 of 73 metrics are scored on historical-percentile rank and 32 on provisional bounds.
  • Trend score (30%): direction and velocity of recent change. trend = clamp(50 + 500 × worse_frac, 0, 100) where worse_frac is the signed fractional change (positive = worsening). A 10% deterioration or improvement spans the full 0–100 range.
  • Volatility score (10%): standard deviation of period-over-period fractional changes (if ≥ 4 history points; else 50). volatility = clamp(sd × 800, 0, 100). Marked as provisional.

Trend scoring and direction

trendDirection is derived from the trend score: >55 = deteriorating; <45 = improving; else stable.

The retrospective national trend line shown on the dashboard uses a level-only method (no trend or volatility component) applied consistently back to 2018. This ensures the historical trajectory uses a single method end-to-end. The current headline composite includes all three components and may differ from the final plotted point; this difference is disclosed beneath the chart.

Confidence scoring and thresholds

Each metric carries a confidenceScore (0–1) reflecting source quality, update frequency, and known weaknesses. The index dataConfidence is the mean of all metric confidence scores.

ThresholdLevel
≥ 0.85High
≥ 0.70Medium
< 0.70Low

Current index data confidence: 0.83. Every colour-coded confidence signal includes a text label; colour is never the sole indicator.

Weighting model

The national score is a weighted arithmetic mean of domain medians, normalised over the domains present in the view. Register weights sum to 100, so for the UK and the four nations (all eleven domains present) this is simply the weighted sum divided by 100; for views carrying a subset of domains (the nine English regions) the denominator is the present weight, so an absent domain is treated as no information rather than a score of zero.

national_score = Σ(weight_d × domain_median_d) / Σ(weight_d present)

Correction, July 2026 release: earlier releases divided by the full register weight in every view, which understated the nine English regional scores. Two further changes in this release: provisional seed placeholders are excluded from all scoring, and the headline movement is now the change since the previous published release.

Higher-weighted domains have a proportionally larger influence on the headline figure. The Stress Contribution chart on the dashboard visualises this: each bar equals score × weight.

These eleven domain weights are an editorial interim. A future version will derive them through a documented budget-allocation expert elicitation (Step 6 of the OECD/JRC Handbook on Constructing Composite Indicators), with the weights remaining the single editable source in the index data. Crucially, the headline does not hinge on the exact figures: the robustness analysis below re-scores the index under randomised weights (and other methodological choices) and the UK national score stays within 4552 (median 49, headline 52), so it is robust to the weighting choice.

Robustness & sensitivity analysis

A single set of methodological choices, these domain weights, this normalisation method, the median domain rule, the weighted-arithmetic national aggregation, and the 0.6/0.3/0.1 blend, produces the headline figure. To test how much the headline depends on those choices rather than on the underlying data, the index is re-scored under 2,000 plausible methodological combinations (a Monte Carlo analysis, seeded for reproducibility).

The headline national score is 52. Across 2,000 plausible methodological combinations, varying the domain weights, normalisation method, aggregation rule, and the level/trend/volatility blend, it ranges 4552 (median 49). The published headline can sit a point or two from the simulation median: the headline uses the production method exactly, while the median averages over many alternative methods, several of which score systematically lower. The score is therefore robust to methodological choice; the largest single driver is the domain aggregation rule (median vs mean vs geometric).

DomainHeadlineMedianRobust range (p5–p95)
Living Standards40352643
Work and Productivity27422649
Public Service Capacity61565063
Population Health84716486
Housing and Household Formation62625964
Childhood and Social Floor54554458
Rule of Law and Safety56555058
Fiscal Resilience34393244
Environment and Climate33332438
Infrastructure47464148
Trust and Wellbeing47473849

This follows Step 8 (“Robustness and sensitivity analysis”) of the OECD/JRC Handbook on Constructing Composite Indicators, which treats a published range under alternative methods as a core requirement for a defensible composite rather than an optional extra. The headline figure remains the weighted mean of domain medians described above; the range quantifies the uncertainty around it.

Systemic risk penalty

When three or more domains simultaneously reach the Critical band (score ≥ 75), a systemic penalty is added to the national score to reflect the increased probability that concurrent failures become self-reinforcing:

  • 3–4 domains in Critical: +5 points
  • 5 or more domains in Critical or Crisis: +10 points

Penalties are not cumulative. The national score is capped at 100. The applied penalty is recorded as systemicPenaltyApplied (0, 5, or 10). Current release: +0 pts.

Source hierarchy and tiers

All metrics must be sourced. Sources are classified into five tiers reflecting data quality and methodological rigour:

Official Statistics

Published by national statistical authorities (ONS, DWP, DfE, NHS England, etc.) under the Code of Practice for Statistics. Highest confidence.

Official (In Development)

Officially published but methodology or coverage is still maturing (e.g. some NHS Digital series). Medium confidence.

Administrative

Operational records collected for administrative rather than statistical purposes (e.g. Environment Agency event monitoring). Medium confidence; may reflect reporting changes.

Survey

Probability or omnibus surveys (e.g. Crime Survey for England and Wales, Opinions and Lifestyle Survey). Confidence depends on response rates and survey design.

Independent Body

Published by independent public bodies with statutory remits (OBR, CCC, Bank of England). High authority; methodology may differ from ONS conventions.

Full source list: data sources page.

Release and revision policy

Each monthly release is published in the first week of the month, with a data cutoff of the 25th of the preceding month: source publications up to the cutoff are incorporated, and anything published later appears in the next release. The index publishes every month regardless of how many indicators received new data, and each Monthly Review states how many did. The cutoff date is recorded on the release itself.

All releases are versioned and preserved. When source data is revised (as is common with ONS administrative series), the metric value is updated but the original release reading is retained in the metric history, labelled with its release period.

Methodology changes are disclosed in the Monthly Review section 10 of the relevant release. Significant methodology changes trigger a re-calibration notice and a comparison table showing how scores would have differed under the old method.

Political neutrality rules

The following ten rules govern the construction of every metric, domain, and narrative in this index. They are reproduced verbatim from the project contracts.

  1. The index does not assign blame to parties.
  2. The index tracks conditions and trends.
  3. Every metric must have a source.
  4. Every metric must have a stated weakness.
  5. Weightings are transparent.
  6. Revisions are preserved.
  7. Survey/perception data is separated from hard administrative data.
  8. Culture-war claims are excluded from the core score.
  9. Immigration is only measured through neutral capacity-pressure indicators, not moral framing.
  10. The project distinguishes between deterioration, low baseline, and poor data quality.

Approved vocabulary: under pressure, elevated stress, fragile, critical, deteriorating, improving, material movement, data confidence, requires monitoring, stable, resilient.

Banned: “collapse”, “collapsing”, “broken”, any UK political party name, any named politician, “the government has failed”. “crisis” is allowed only inside a registered source’s own dataset name.

Known weaknesses

The following limitations are documented honestly. They do not invalidate the index but should be borne in mind when interpreting any specific reading.

  • England-only bias: several public services and housing metrics use England-only administrative data, reflecting devolved structures. Scottish, Welsh and Northern Irish equivalents are registered for future addition.
  • Survey lag: welfare and perception indicators (HBAI, CSEW, community surveys) carry a 12–18 month lag. Current readings may understate or overstate current conditions.
  • Short-history metrics: some metrics use provisional best/worst bounds rather than historical-percentile normalisation; each metric carries a "normalisation" field for transparency, and bounds-based metrics migrate to percentile as their verified series lengthen (which can shift the headline slightly).
  • Band boundaries are thresholds on a continuous score, not sharp categories: a national or domain reading within a point or two of a threshold should be read as sitting between bands rather than as a sharp category change.
  • Percentile interpretation: a metric that has been chronically stressed will score around its own historical median (~50) even if the absolute level is high, because percentile is relative to its own track record. The trend component (30% weight) separately captures whether stress is worsening or improving.
  • Domain weights: the eleven domain weights are an editorial interim, set by judgement rather than empirical optimisation. A future version will derive them via a documented budget-allocation expert elicitation (OECD/JRC Handbook Step 6). Crucially, the sensitivity analysis already demonstrates that the headline is robust to weight variation, re-scoring under randomised weights leaves the national score within the published robust range, so the interim weights are not load-bearing for the headline verdict.
  • Volatility score: the volatility component is marked provisional. With fewer than four history points per metric it defaults to 50, adding noise to the composite.
  • Metric independence: some metrics within a domain (e.g. housing affordability and social-housing waiting lists) are positively correlated; using the median partially mitigates double-counting but does not eliminate it.

International benchmarks & faster signals (context only)

Each domain page includes an International context card showing how the UK compares to the OECD or G7 on a single headline indicator for that domain. These figures are sourced separately from the index metrics, primarily from the OECD, World Bank, and IMF, and are shown purely as context. They are un-scored: international benchmarks do not affect any domain score, the national score, or risk bands. They are not part of the British Resilience Index score. Each benchmark carries a confidence level (high / med) and a source link, and is labelled “context only, not part of the score.”

Where a domain has a reliable higher-frequency series, its page also shows a Faster signal · a monthly or quarterly indicator (for example CPIH inflation, job vacancies, public-sector borrowing, or mortgage approvals) showing which way the domain has moved since the last annual scored release. These are sourced from the ONS, Bank of England and others and, like the benchmarks, are un-scored context that never affects the index score.

Future improvements

The following improvements are planned. They are listed in rough priority order.

  • Migrate the remaining bounds-based metrics to historical-percentile normalisation as verified historical series accumulate.
  • Regional and sub-national breakdowns for all eleven domains.
  • Live data ingestion pipeline with automated source refresh.
  • Formal domain-weight elicitation exercise.
  • Independent methodology review by a panel of statisticians.
  • Confidence intervals on domain and national scores, not just point estimates.
  • Machine-readable API for the full index and metric time series.

Further technical documentation is available in the repository’s methodology/ directory. Source ingestion specifications are in CONTRACTS.md.

Risk band reference

Score rangeLabel
0–24Stable
25–49Strained
50–74Fragile
75–89Critical
90–100Crisis