How It Works

The Narrative Gap monitors news coverage from a curated set of sources across multiple regions and languages. Our system identifies what is being claimed, who is saying it, and how consistently those claims hold up across independent reporting.

Multi-source analysis

Rather than relying on any single outlet, we look at how a story is reported across different media ecosystems. Claims that appear independently in multiple regions carry more weight than those from a single source or media tradition. When coverage is heavily concentrated in one region, we flag that so readers can judge accordingly.

Competing explanations

For contested questions — where reasonable people disagree about what is happening or why — we track multiple explanations and assess how well each one holds up against the available evidence. This helps readers see the full landscape of possibilities rather than being funnelled toward a single narrative.

Confidence levels

Our confidence assessments reflect how strongly the available evidence supports a claim — how strong and consistent that evidence is, weighed for its quality and independence. Confidence is not a count of how many outlets carried a story: a single strong, uncontested source can warrant high confidence, while many copies of the same report add little. How widely a claim was reported is a separate signal, shown alongside the claim. We use plain language rather than misleading percentages:

Two ways we describe likelihood

Event pages describe competing explanations two ways at once, because they answer two different questions:

A field of weak explanations can have a "Leading" one that's still unlikely in absolute terms — both readings are honest at once. We never call an explanation "Most likely" unless it also clears a minimum bar on the absolute scale; a weak leader is labelled "Leading" instead.

How likely an explanation is, on its own

The seven bands, from most to least likely:

For events we treat as graver — where getting it wrong matters more — we hold competing explanations to a stricter standard before calling one "likely" or better.

How we phrase a question's verdict

When a question has enough signal to say something, we phrase it one of four ways:

How clearly a source states something

Separately from confidence and likelihood, we also note how directly the underlying source material states a claim — "very clearly stated", "clearly stated", "less clearly stated", or "ambiguously stated" by the source. This is about how cleanly the claim was extracted from the reporting, not about how believable the claim itself is.

Intent and motive assessments

When we assess what an actor "likely" intends or what their motives may be, these are analytical inferences drawn from patterns in reported evidence — not allegations or statements of personal knowledge. Evidence may be incomplete or misleading, and our inferences may be wrong. They represent the best reading of available reporting, not established fact.

What we don't do

How we measure confidence

Our confidence in a claim reflects how strongly the available evidence supports it. We weigh the quality of that evidence, whether it is independent and mutually corroborating, and whether any of it points the other way — and for higher-stakes claims we set a higher bar. Confidence is not a headcount of outlets: a claim carried by a single strong, uncontested source can reach high confidence, and a story echoed by many outlets that all trace back to one report is not thereby more certain. How widely a claim has been reported is shown separately, alongside the claim, so you can weigh reach and confidence as two distinct things.

Checking our predictions

When a claim we track makes a specific, checkable prediction, we follow up once it's due and record what actually happened, using three labels:

Our track record page shows the current state of that checking — including if we've had to withdraw judged outcomes. Not every prediction can be checked before it expires, so the judged set is a subset of everything we track, and it skews toward the predictions we were most confident about, since that's the set we prioritise checking. An expired, unchecked prediction is never counted as wrong.

Limitations

All analysis is advisory. Confidence assessments reflect the balance of available evidence at the time of analysis and may change as new information emerges.