Library / Glossary / The context glossary
Programme evaluation glossary
30 terms from C22, Programme evaluation — each defined in the primer's own words. Every term the primer teaches links to the slide that teaches it.
C22 · Evidence and change
Programme evaluation
“Evaluation is research done on a programme.”
Every term below is defined in the words of programme evaluation, primer C22 of understanding the context of education, and opens the primer at the slide where it is taught. 3 terms are also defined by another primer in the series; where the two differ, both wordings are given. The whole context glossary holds all of them together.
| Term | Definition | Referred to in | Read further |
|---|---|---|---|
| A | |||
| Attribution against contribution | Attribution claims a share of the observed change for the programme and needs a counterfactual. Contribution claims the programme was one cause among several, none sufficient alone, and is established by testing the theory and the rival explanations (Mayne, 2012). | ||
| C | |||
| CIPP model | Stufflebeam's model of four kinds of evaluation, each answering one question: context, "What needs to be done?"; input, "How should it be done?"; process, "Is it being done?"; product, "Did it succeed?" (Stufflebeam, 2007). |
| |
| Context–mechanism–outcome configurations | The unit of a realist evaluation: a statement of how a programme's resources trigger a mechanism, among whom and in which conditions, to produce a given outcome. Summarised as "context + mechanism = outcome" (Pawson & Tilley, 2004). |
| |
| Counterfactual | What would have happened to the same people without the programme. It cannot be observed, so it is estimated from a comparison or control group. Primer C21 covers the designs used to do this. C21 What would have happened to the same students without the programme. That world cannot be observed, so every causal design is a way of building a stand-in for it. |
| |
| D | |||
| Developmental evaluation | Evaluation of an innovation that is still changing, where the model is not settled and may never be. It supplies continuous feedback rather than a verdict (Patton, 1994; Gamble, 2008). |
| |
| E | |||
| Economic evaluation and cost-effectiveness | Analysis of what a programme consumes as well as what it achieves. Cost-effectiveness compares outcomes per unit of cost against alternatives; cost-benefit analysis puts a monetary value on the benefits. Costs are built up by the ingredients method, which lists every resource a programme needs (Hollands et al., 2020). |
| |
| Empowerment evaluation | A participatory approach in which the evaluator supports programme staff and participants to evaluate their own work, with capacity and ownership as intended outcomes (Fetterman, 1994). |
| |
| Evaluability assessment | A short, structured check on whether a programme can usefully be evaluated: are its goals clear, is its theory plausible, are the data available, and would anyone act on the answer (Wholey, 1987; Craig & Campbell, 2015)? |
| |
| Evaluation against research | Research sets its own questions and aims at knowledge that generalises; an evaluation answers questions set by a programme and its users, about that programme, and ends in a judgement and usually a recommendation (Levin-Rozalis, 2003). |
| |
| Evaluation of professional development | Evaluation of training for teachers and staff. Guskey (2002) sets out five levels: reactions, learning, organisation support and change, use of new knowledge and skills, and student learning outcomes. |
| |
| Evaluative rubrics and reasoning | The explicit path from evidence to a value conclusion: criteria saying what matters, and standards saying what weak, adequate and strong performance would look like on each, agreed before the data arrive (Davidson, 2005). |
| |
| F | |||
| Formative and summative evaluation | Formative evaluation is done to improve something while it is still being shaped; summative evaluation is done to reach a verdict about it. Scriven (1996) argues that the same activity can be either, depending on the context it is used in. |
| |
| G | |||
| Goal-free evaluation | Evaluation in which the evaluator is deliberately kept from the programme's stated goals, so as to record what it is actually doing and to catch effects nobody intended, including harms (Scriven, 1991b, quoted in Youker, 2024). | ||
| K | |||
| Kirkpatrick's four levels | A level model used for courses, training and edtech that runs from reaction to learning to behaviour to results: a useful ladder and a dangerous shortcut (Kirkpatrick & Kirkpatrick, 2006). |
| |
| L | |||
| Logic model | A diagram of a programme as a chain: resources, activities, outputs, immediate outcomes, intermediate outcomes and ultimate outcomes (Treasury Board of Canada Secretariat, 2021). It lists the boxes but does not explain the arrows. |
| |
| M | |||
| Merit and worth | Merit is how good something is in itself; worth is how valuable it is to a particular set of people in meeting their needs. Stufflebeam (2007) adds probity, meaning integrity and honesty, and significance, meaning importance beyond its own setting. |
| |
| Meta-evaluation | Evaluation of an evaluation, against professional standards. The Program Evaluation Standards devote a whole category to it, and Stufflebeam (2001) argued that it is an obligation rather than an extra. |
| |
| N | |||
| Needs assessment | A study of what the people a programme serves actually need, which supplies the criteria against which worth is later judged. In the CIPP model this is context evaluation, asking "What needs to be done?" (Stufflebeam, 2007). |
| |
| O | |||
| OECD evaluation criteria | Six criteria for development evaluation: relevance, coherence, effectiveness, efficiency, impact and sustainability, which together "provide a normative framework used to determine the merit or worth" of an intervention (OECD, n.d.). |
| |
| Outcome and impact evaluation | Outcome evaluation asks what changed among those the programme reached. Impact evaluation asks how much of that change would not have happened otherwise, which requires a credible comparison. |
| |
| P | |||
| Pilot evaluation | The evaluation of a small first run, normally to test whether the programme can be delivered and whether an outcome study is feasible. A pilot with volunteers, extra support and no comparison cannot tell you whether the programme works. |
| |
| Process and implementation evaluation | Evaluation of whether a programme was delivered as intended and how it was experienced, covering fidelity, quality, reach and the contextual factors behind variation in outcomes (Moore et al., 2015). |
| |
| Program Evaluation Standards | Thirty standards in five categories: utility, feasibility, propriety, accuracy and accountability. They are the field's nearest thing to a shared code, and utility comes first (Joint Committee on Standards for Educational Evaluation, 2018). |
| |
| Programme evaluation | The systematic assessment of a programme's merit, worth, probity and significance, for people who have to decide something about it (Stufflebeam, 2007). It is a field with its own models, standards and professional bodies. C21 Judging the merit, worth or value of a programme for the people who must decide about it. It has its own discipline, standards and models, covered in Programme evaluation (C22). |
| |
| Programme theory | The account of how and why a programme is supposed to produce its effects, including the mechanisms it relies on and the conditions it needs (Weiss, 1997b; Coryn et al., 2011). |
| |
| R | |||
| Realist evaluation | Evaluation that asks "What works for whom in what circumstances and in what respects, and how?", on the claim that an outcome follows from a mechanism firing in a context (Pawson & Tilley, 2004). C21 An approach that asks what works for whom in what circumstances, treating context as part of the finding rather than as noise (Pawson & Tilley, 1997). |
| |
| S | |||
| Stakeholder and participatory evaluation | Approaches that involve the people with a stake in a programme in framing questions, interpreting findings or running the evaluation. Involvement is the factor most consistently linked to whether evaluations get used (Johnson et al., 2009). |
| |
| T | |||
| Theory of change | The explanation of why each step in the chain should follow from the one before, with its assumptions, risks and context made explicit (Treasury Board of Canada Secretariat, 2021). |
| |
| U | |||
| Use and influence of evaluation | What an evaluation actually changes: a decision, a practice, or how people understand a problem. Utilisation-focused evaluation treats use by named intended users as the goal of the whole exercise (Patton, 2013). |
| |
| Utilisation-focused evaluation | An approach that starts from the user: identify the primary intended users, design every step around their use, and report in a form they can act on (Patton, 2013). |
| |
No term matches. Try fewer letters.
Nearby