Example

You will be asked to stand behind
the sentence at the top of the report.

Northline’s draft says the job-readiness program increased lasting employment by 30% for the people they serve. The intake log, the follow-up tracker, and the field definitions are already on the shared drive. This page follows that sentence through those files, while the wording can still be changed.

Synthetic organization · illustrative example

Northline Employment Partnership

Northline runs a job-readiness program at two sites. 214 people completed last year. An Ashgrove Foundation renewal is approaching, and the board has asked for evidence of sustained employment outcomes.

2 sites · 214 completions · $3.1M annual budget

Ahead of the renewalWhat they already had

The waitlist is already in the intake log.

The draft annual report says the program increased lasting employment by 30%. The comparison behind that number is people who enquired and never enrolled. The same folder has an intake log that offers places in referral-date order until the 60 are filled.

The decision
Can this sentence go in front of the board?
What is being checked
Can we say this?
Read it in
90 seconds

Our job-readiness program increased lasting employment by 30% for the people we serve.

A board will remember a sentence like this. The comparison behind it is people who enquired and never enrolled.

In the files already

Places are capped at 60. Later referrals wait.

Places are offered in the order people were referred, until the 60 are filled. The 87 who wait look like the 60 who got in. What separates them is that the places ran out. That comparison is fairer than people who enquired and never enrolled.

Cohort capacity: 60 places per intake.

Applicants are ordered by referral date. Places are offered in that order until capacity is reached.

Applicants above the cut are enrolled. Applicants below the cut are held to the next intake.

The consideration this raised

Counterfactual Construction

The methods by which a credible comparison condition is constructed when randomization is not feasible.

What the record supports

The intended causal sentence fits within a quasi-experimental ceiling, with visible limits.

The sentence type fits, but its population, measure, source, or analytic status needs to travel with it.

What this comparison can support

Quasi-experimentalfrom a cutoff, waitlist, capacity rule, boundary, or phased rollout

The files

Report

Draft Annual Report 2025–26

page 4 · circulated to the board 2 April

Program results

Our job-readiness program increased lasting employment by 30% for the people we serve.

Across two sites, 214 people completed the program in the reporting year.

We are grateful to the Ashgrove Foundation for three years of support.

What the check returned

What if they change one thing

The comparison gets stronger, and that finding holds. How many people were reached, and what employed means, still have to sit with the sentence. Those limits are the next part of the check.

What the record can carry

Within the selected operational comparison, employment was 30% higher than among people who did not enrol among program completers. This can support a causal interpretation of the job-readiness program only if the design's identifying assumptions hold. The selected measure is a proxy for the outcome named in the claim. Outcomes were visible for 42% of the relevant group.

The smallest useful next step

Report the observed and missing groups separately, then test how good and bad outcomes among the missing group would change the conclusion.

1of 149 considerations in the method applied here
  • Counterfactual Construction

Applied because the files and the decision made them relevant.

The same filesThe 90 people

90 of 214 people answered the six-month call.

Follow-up reached 90 of 214 people. The field records any paid work in the 30 days before the call. A funder is being asked to rely on that number. A narrower sentence is one the board can be shown.

The decision
What can the sentence say?
What is being checked
Can we say this?
Read it in
90 seconds

What the record supports

The intended causal sentence fits within a quasi-experimental ceiling, with visible limits.

The sentence type fits, but its population, measure, source, or analytic status needs to travel with it.

What this comparison can support

Quasi-experimentalfrom a cutoff, waitlist, capacity rule, boundary, or phased rollout

The files

Data export

Six-month follow-up tracker

program database · exported 28 March

Completed program: 214

Reached at six months: 90

Follow-up rate: 42%

Not reached: 124 — no working number (71), no response after three attempts (38), declined (15)

Outreach calls were made Tuesday and Thursday, 9am–4pm.

What the check returned

What if they change one thing

The comparison holds. How many people were reached, what employed means, and who holds the numbers still sit with the sentence.

Work that can be dropped

The new survey can be cancelled.

Northline had budgeted a fresh six-month outcome survey to strengthen the renewal case. It would ask the same 30-day question of the same reachable group. A larger sample would leave the sentence where it is, three months later.

The consideration this raised

Proportionality of Rigor

The matching of design rigor to decision stakes and value of information, so that the intensity of investigation is calibrated to what resolving the uncertainty is worth rather than defaulting to the most rigorous design the budget allows.

Where this leaves them

There is a sentence the board can be shown.

The waitlist already in the intake log is a fairer comparison than the one in the draft. The sentence still has to name the 90 people reached, the 30-day measure, and the remaining limits. The expansion question needs a different check.

What the record can carry

Within the selected operational comparison, employment was 30% higher than among people who did not enrol among program completers. This can support a causal interpretation of the job-readiness program only if the design's identifying assumptions hold. The selected measure is a proxy for the outcome named in the claim. Outcomes were visible for 42% of the relevant group.

The smallest useful next step

Report the observed and missing groups separately, then test how good and bad outcomes among the missing group would change the conclusion.

3of 149 considerations in the method applied here
  • Partial Identification and Bounds
  • Instrument Design and Validation
  • Measurement Independence Governance

Applied because the files and the decision made them relevant.

Six weeks laterThe expansion question

The board is being asked to double the program.

The same files now have to answer a larger question. Staging the rollout and opening one new site leave the answer where it was. Filling the coach posts moves it.

The decision
Should we extend the program to two more sites next year?
What is being checked
Would this still work at two more sites?
Read it in
4 minutes

What the record supports

Delivery will not hold at the next size.

Workforce, training, inputs, quality measurement, or founder-led delivery will break before the proposed jump. The pilot result is not a delivery system.

The files

Proposal

Expansion proposal to the board

FY27 · tabled 6 April

Recommendation

Extend the job-readiness program to Eastvale and Rill Street in FY27.

Target: 2× current annual enrolment.

New sites will report the same monthly indicators as existing sites.

What the check returned

What if they change one thing

The coach posts, the shared employers, and how they would know if it worked still carry cautions.

What to do before the vote

Secure the constrained input, or redesign delivery around what the ordinary supply can hold.

  • Do not treat the founding team or extra coaching as part of the model unless they will be present at scale.
  • Install a quality signal you can see while delivery is happening.

This result is a structured decision aid based on the answers you selected. It does not verify the underlying facts, estimate a treatment effect at scale, certify readiness, or replace a professionally owned expansion judgment. No result here means the program is ready for full scale.

5of 149 considerations in the method applied here
  • Institutional Capacity Assessment
  • Moderator Theory and Transportability Priors
  • General Equilibrium Sensitivity
  • Scaling and Adoption Pathway
  • Counterfactual Construction

Applied because the files and the decision made them relevant.

Your own decision

You can put the same questions to your own report.

The Decision Brief asks what is being decided, who needs the answer, by when, what information already exists, and where the uncertainty sits. It runs in your browser, and it reaches Prova only if you choose to send it.

Prepare a decision brief

Northline is a synthetic organization, and this example is illustrative. It carries no conclusion about any real program or client.