Methodology

How a metric gets chosen, how its peer set is picked, where the figures come from, and what a verdict does and does not claim.

Every claim on this site is a comparison, and a comparison is only as good as the choices behind it — which countries, which measure, which year, which source. This page sets out those choices, so that a reader can check them and disagree with them specifically rather than in general.

For why the project exists and how the five pillars fit together, see about. What follows is the working method.

How a metric is selected

Not everything worth caring about can be benchmarked, and not everything that can be benchmarked is worth a page. A candidate has to clear five tests before it is built.

Related measures are paired, not collapsed. Where two measures of the same subject move differently — a level and its rate of change, the capital raised and where it goes, the assessment stage and the licensing stage — they are built as separate pages that link to each other. Averaging them into one figure would conceal precisely the disagreement that makes the pair informative.

Choosing the peer set

The peer set is chosen on the merits of the comparison and before the result is known, and it is stated on the page. The default is the G7: countries of comparable income and institutional type, against which Canada already measures itself in ordinary public debate.

Two departures from that default are used, both declared where they apply.

Like for like. A cross-country comparison must hold constant everything except the country: the same profession, the same stage of the same process, the same type of institution, the same units. A six-week turnaround at one country's nursing regulator set beside a fifteen-week one at another country's engineering body is not a country comparison — it is a profession comparison wearing a flag. Where the like-for-like series does not exist, the page compares within a single profession or process and states what it therefore cannot conclude.

How to read a verdict

Every metric carries one of three verdicts, describing where Canada sits within that metric's peer set, on that measure, at the vintage stated on the page.

Strong
Canada sits in the upper part of its peer set — a position other comparable countries would want.
Watch
Canada sits in the middle — no crisis, but no advantage either, and often a direction worth watching.
Weak
Canada sits at or near the bottom of its peer set.

The verdict follows Canada's position in the peer set, but it is a judgement rather than the output of a formula, and the page shows the working that produced it. Where a measure's level and its direction disagree — a strong position eroding quickly, a weak one improving steadily — the page says which of the two the verdict reflects, and why. A verdict is a one-word summary of a page, not a substitute for reading it.

Three things a verdict is not. It is not a grade for any government: most of these measures move over decades and across administrations. It is not a claim about cause; establishing that Canada trails its peers is a different exercise from establishing why. And it is not comparable between metrics — each metric has its own peer set, so Weak on one page and Weak on another are not necessarily the same distance from the front.

Pillar-level verdicts on the metrics overview are computed from the metrics currently in view, so filtering the dashboard recomputes them rather than leaving a stale headline in place.

An important exception. For some measures no comparable international series exists — no other country publishes passport processing times or veterans' disability claim backlogs on a basis that can be set beside Canada's. Where that is the case, the benchmark is Canada's own published service standard, and those pages say so explicitly. We would rather compare Canada to its own commitments than manufacture a false international ranking out of mismatched definitions.

Where the figures come from

Sources are used in order of preference, and the page says which level it is drawing on.

  1. The primary series, from the agency that produces it. Statistics Canada, the Bank of Canada, the IMF, the OECD, the World Bank, departmental performance reports and departmental plans, and the equivalent national agencies in peer countries.
  2. The underlying records, assembled here. Where no published series exists but the case-level or administrative records do, the series is constructed from those records and the construction is described on the page, so that someone else could rebuild it.
  3. A named institutional source, cited and flagged. Where neither of the above is obtainable — a research institute, an industry association, a private data holder — the holder is named, the page says the figure depends on them, and the limitation travels with the number.

Not used: figures taken from news reports, values read off someone else's chart, aggregator sites and encyclopaedias, and any number that cannot be traced back to a body prepared to stand behind it. Every chart on every metric page carries its source, the specific table or series identifier, and the vintage of the data.

How figures are checked

Each metric page has a companion workbook holding the figures behind its charts. Before a page is published, and again before any change to it ships, it passes four checks.

  1. Page to workbook Every number and every chart point on the page traces to a cell in the workbook behind it. Nothing appears on a page that does not exist in its data, and no value is ever invented to close a gap or illustrate a point.
  2. Workbook to source Every figure in the workbook matches the cited source at the stated vintage. Where a source has revised its series since the figure was taken, the revision is adopted rather than the original quietly retained.
  3. Comparability Where countries define a measure differently, the difference is disambiguated on the page rather than averaged away. Where a ratio can be computed on more than one basis, the page states which basis it uses.
  4. Internal consistency No claim contradicts another: a series described as a record high has to exceed the maximum of the series shown. Where the same figure appears on more than one metric page, it agrees across them, or the difference is explained on both.

Charts

Vintage, updates, and revisions

Figures carry the vintage of the data, not the date of the page. A metric published this month may show a figure from two years ago because that is the most recent complete period its source has released; where that is the case the page says so, rather than substituting something more current and less comparable.

Metrics are updated when their sources publish, not on a fixed calendar. A page is not refreshed with a partial period presented as a complete one. Where a source revises history, the revised series is adopted; where a revision changes a finding, the change is noted on the page rather than absorbed silently.

Corrections

If you believe a figure here is wrong, please say so. Corrections are published rather than quietly absorbed.

A challenge to a specific figure will be answered specifically: the source it came from, the judgement calls made in constructing it, and any revisions since publication. That is the standard this project holds itself to. A number that cannot be defended in that detail should not be on the site in the first place.

How metrics are tagged

Besides its pillar, each metric carries two sets of tags, used by the filters on the metrics overview.

What this method cannot tell you

Think something here is wrong? Questions about a figure, corrections, and suggestions for metrics worth adding are all welcome — see contact.