DawaniعStart a project

KnowledgeEvidence: CUpdated:

How does a nonprofit measure a children's program's impact? From theory of change to the donor report

From a theory of change to a mandatory guardrail indicator to a donor report speaking of outcomes, not only activities: how a nonprofit honestly measures a childhood program.

In this article

How does a nonprofit honestly measure its program's impact on children?

Start from a clear theory of change linking the program's activity to a specific outcome, measure a proximal indicator that moves during the program itself, make at least one guardrail indicator mandatory rather than optional, and write the donor report in the language of outcomes, not only activities. These four steps together are the difference between a report describing what happened and one proving what changed.

Many childhood programs measure what is easy to measure, such as how many children attended, rather than what is actually worth measuring, such as what changed in them. The difference between the two is not only a matter of effort required; it is a difference in the question asked from the start: are we describing our activity, or measuring our outcome?

What is a theory of change, and how is one built for a childhood program?

A theory of change is a sentence or a short chain of sentences linking what a program does to what is assumed to change in a child because of it, with clear intermediate steps between the two. A program with no written theory of change resembles a trip with no map: you may arrive somewhere, but you do not know if it is where you meant to go.

Building a good theory of change for a childhood program needs the children themselves involved in deciding what matters to them, not adults alone assuming what they need. Roger Hart's ladder of children's participation (Hart, 1992), a well-known reference in the participation literature, ranks degrees of participation from a token one with no real influence to a genuine one where children lead part of the decision. The Saudi Child Protection Law, in its Article 16, requires all bodies to consider the child's interest in every measure taken concerning them, and to account for their mental, psychological, physical, developmental and educational needs in line with their age and health. These two references, one academic and one statutory, form a reasonable foundation for any theory of change that takes a child's voice and best interest seriously from the start, rather than treating it as an extra detail added once the design is finished.

What proximal indicators suit a childhood program?

A proximal indicator moves during the program itself and can be attributed to it with reasonable confidence: a skill a child demonstrates in a specific task, an intended behaviour that recurs in an observable way, or a self-report carefully designed in language close to a child's age rather than borrowed from a questionnaire built for a different age group. These indicators, explained in further detail in the article on measuring adolescent impact, differ from distal indicators such as overall academic attainment, which need a rigorous research design rarely available to a small community program.

The share who keep participating voluntarily, with no required attendance, is an honestly proximal indicator on its own: it measures real desire rather than imposed obligation. A program that conflates the two might declare success based on high attendance, while attendance itself is required by some other condition unrelated to a participant's actual desire for the program.

What is the mandatory guardrail indicator, and why can it not be skipped?

Every childhood program carries a specific possible harm that must be watched by name, not left to chance or discovered too late. In a community program, this harm might be attendance driven by fear of losing a benefit tied to the program rather than real desire to take part, or pressure on a child to show improvement in front of an evaluation team regardless of their actual state. An impact report with no guardrail indicator clearly stated is itself a warning sign, whatever the rest of its numbers look like.

Choosing the right guardrail indicator starts with a plain question the program team asks itself: what is the worst unintended harm this specific program could cause? An honest answer to that question is where the indicator comes from, not a general list copied from an entirely different program in a different context.

How does a program scale without losing delivery quality?

Rapid scaling threatens delivery quality when a program is copied to new sites with no adequate training for the delivering team at each one. The fix is not freezing growth, but separating what must stay fixed in every version of the program, such as its guardrail indicator and its principle of child participation, from what can adapt to each local context, such as session timing and the language used with families.

A periodic review of delivery quality at every site, not only the first one the program started from, reveals the difference between scaling that preserves a program's substance and scaling that preserves only its name.

How is protection designed for volunteers and children inside the program?

Volunteers who deal directly with children need a clear path for supervision and reporting, not good intentions assumed without verification. A sound protection path includes training before any direct contact with children, a written code of conduct every volunteer signs, and a way to report any concern that any child or fellow volunteer can use without fear of consequences. The absence of this path does not necessarily mean real harm exists, but it does mean any harm will be discovered by chance rather than by a design that protects against it from the start.

How does a donor report shift from activities to outcomes?

The difference between the two forms shows in one example. Activity language (illustrative example): "The program ran a number of workshops this term, attended by a number of children." Outcome language (illustrative example): "Of the children who attended the workshops, a notable share returned to practise the same skill with no reminder from the program team, while the voluntary follow-up participation rate remains an indicator we watch, not a final result we announce." The second version names a specific proximal indicator and honestly keeps the limits of what it knows, instead of a total figure that does not distinguish attendance from outcome.

Writing a report in this form needs proximal indicator data collected from the program's first day, not in its last week in preparation for handing in the report. A report written retroactively from a team's memory rarely carries the accuracy of one whose data was collected during the program itself.

Child Social Impact Canvas: a list of its fields

One working canvas gathers everything above onto a single page the team reviews with every program cycle, as an internal working tool rather than a standard approved by any outside body:

  • The precisely targeted age band, not a wide general range.
  • The theory of change in one short sentence someone who did not write it can understand.
  • The specific proximal indicator to be measured, and how it will be measured.
  • The mandatory guardrail indicator, and the specific worst harm it watches.
  • The data collection method, who collects it, and when within the program cycle.
  • Which children themselves took part in designing the program or its indicators.
  • The limits of how far this cycle's results can generalize to other cycles or sites.
  • The date of the next review of the theory of change itself, not only its results.

Checklist: from activities to outcomes in your next report

  • Does every results sentence state its correct type: observed, measured, or interpretation?
  • Is at least one guardrail indicator named explicitly, rather than assumed?
  • Is the voluntary continuation rate measured, not only total attendance?
  • Did the children themselves take part in deciding what matters to measure?
  • Does the report state the limits of what cannot be claimed with the same confidence as what was actually measured?

What are the limits of this framework?

The Child Social Impact Canvas is an internal working tool, not a peer-reviewed research instrument or a measurement standard approved by a regulatory body. Hart's ladder is a well-known academic reference for the principle of child participation, and the Child Protection Law article is a real Saudi statutory text, but neither is a standard that dictates how a specific program's results are measured numerically. Any program needs to adapt this canvas to its own context, not apply it literally as a final list.

Where should you start?

If you run a childhood program and need a clearer theory of change, start with one sentence describing what changes in a child because of your program, and write it on a page separate from any attendance number. If your program serves the youngest ages specifically, the article on designing beyond screens for early childhood offers related design principles from a different angle. You can read about our pathway for nonprofits and foundations, try the Outcome to Experience tool to turn your program's outcome into a situation a child actually lives through, or tell us about your program through the Start a project form.

References

  • Hart, R. A. (1992). Children's Participation: From Tokenism to Citizenship. UNICEF Innocenti Essays, No. 4. A standard academic reference, cited as (Hart, 1992), with no link.
  • Saudi Child Protection Law, issued by Royal Decree M/14 of 3/2/1436H (25 November 2014), Article 16. Per the Council of Ministers' Bureau of Experts, retrieved 6 September 2026.

Quick questions

Is a theory of change a document written once and never revisited?

No. A theory of change is a working hypothesis that gets tested and revisited every program cycle, and adjusted when reality shows a step in it does not lead to what it assumed. A document never revisited is a sign it is being read, not used.

Is it enough for a report to mention beneficiary satisfaction as an impact indicator?

Satisfaction is a useful indicator but not sufficient alone, because it measures a momentary feeling rather than an actual change in behaviour or skill. A good impact report names satisfaction alongside a proximal indicator of real change, not instead of one.

Who decides what the mandatory guardrail indicator is for a specific program?

The program team itself, after asking one question: what is the worst unintended harm this program could cause? An honest answer to that question is where a guardrail indicator comes from, not a general list copied from another program.

Have an idea for children to try?

Tell us the goal in your own words; someone on the team will read it.

Start a project