Group Designs for Outcome Evaluation of Training
The case hands you a design opportunity: seven centres starting at different times. Recognising that as a stepped-wedge opportunity is the strongest move available.
Editorial process
Last reviewed · August 16, 2026
Choose a design, then generate criteria you could actually measure
Read the case for its design constraints before choosing anything, because it contains a gift. Seven regional centres are willing to participate and they are starting the training at different times — which is a naturally staggered rollout, and a staggered rollout supports a stepped-wedge or multiple-baseline design in which every centre eventually receives the training and each acts as its own control beforehand. That design answers the ethical objection to withholding training and gives you comparison at the same time, which is why recognising it is worth more than defaulting to a randomised or a one-group-pretest-posttest design. Set out the alternatives so the choice is visible: a one-group pretest-posttest is feasible and confounded by everything that happens over time; a non-equivalent comparison group is stronger but risks selection bias; randomised assignment is strongest and often refused by agencies. Say which threat to validity each alternative leaves open, and which one your choice closes.
The criteria are the second half and they are where evaluations quietly fail. Foster parent training has outcomes at several distances from the training itself, and confusing them is the error to avoid: satisfaction with the training is not learning, learning is not behaviour change, and behaviour change is not child outcome. Name criteria at more than one level — knowledge measured before and after, self-efficacy on a validated scale, observed or reported parenting behaviour, placement stability, and child wellbeing measures — and be honest that the outcomes furthest from the training are the ones that matter most and are hardest to attribute to it. Then say who supplies each measure and when, because a criterion with no data source is a wish. Note attrition too: foster carers leave, placements end, and a design that assumes a complete follow-up sample will not survive contact with the agency's records.
Likely learning objectives
Inferred from the brief — check these against your own rubric.
- 01Match a group evaluation design to the constraints a real agency imposes.
- 02Recognise a staggered rollout as a design opportunity.
- 03Distinguish outcome levels from satisfaction through to child wellbeing.
- 04Attach a data source and a timing to each criterion.
Read the full question
Review every instruction before using the planning guidance that follows.
Turn the brief into deliverables
- 01The design alternatives considered.
- 02The design selected, with its justification from the case.
- 03The comparison the design provides.
- 04Criteria at more than one outcome level.
- 05A data source and timing for each criterion.
- 06Attrition and attribution addressed.
The programme, the design, the comparison, then the criteria
What the case constrains
Identify participation, timing and ethical constraints from the case.
Design alternatives
Set out the candidate designs with their strengths and threats.
The design selected
Choose and justify from the case constraints.
Criteria: proximal outcomes
Specify knowledge and self-efficacy measures.
Criteria: distal outcomes
Specify behaviour, placement stability and child wellbeing measures.
Data sources, timing and attrition
Say who supplies each measure, when, and what happens to dropouts.
Evaluation design sources with real trade-offs
Recommended databases
- PubMed Central
- Campbell Collaboration
- The assigned case study
- Program evaluation texts
Search sequence
- 1.Read the case for participation and timing details before choosing a design.
- 2.Look up stepped-wedge and multiple-baseline designs and their assumptions.
- 3.Search for validated foster carer self-efficacy or knowledge measures.
- 4.Find evidence on placement stability as an outcome and how it is measured.
Reference shortlist
These are authoritative starting points, not a ready-made bibliography. A qualified reviewer must confirm that each source fits the assignment and supports the claim beside which it is cited.
Nothing here is cleared for citation until you have read it.
- 01
Logic analysis: testing program theory to better evaluate complex interventions
Canadian Journal of Program Evaluation · 2011
Logic analysis for testing programme theory — how criteria are derived rather than invented.
- 02
Quantitative Methods
University of Southern California Libraries · 2025
Group design options with their validity threats set out.
- 03
Process evaluation of a person-centred outcome measures-based quality improvement program in a hospital setting
PMC / National Library of Medicine · 2025
A worked process evaluation of an outcome-measure-based programme.
- 04
Continuous Quality Improvement
StatPearls, NCBI Bookshelf · 2023
Continuous improvement, for the difference between evaluating and improving.
- 05
The impact of substance use disorders on families and children: from theory to practice
Social Work in Public Health · 2013
Family-level outcomes and how they are measured in practice.
Review before submission
Common mistakes
- Defaulting to a one-group pretest-posttest without considering alternatives.
- Missing that the staggered start supports a stronger design.
- Treating satisfaction as an outcome measure.
- Listing criteria with no data source or measurement point.
Submission checklist
- Did you consider more than one design before choosing?
- Is the choice justified from the case's own constraints?
- Are criteria at more than one level of outcome?
- Does each criterion have a source and a timing?
Use this guide to plan and review your own work. Follow your institution's rules and read our academic-integrity policy.

Written by
Aaron Bishop
MA, Education
assignment interpretation and research-methods coaching across disciplines
Aaron leads the EssayCrackers editorial desk. He works on how assignment briefs are read — what a rubric is actually asking for, and where students most often answer a different question than the one set.

Reviewed by
Dr. Nathan Cole
PhD, Rhetoric & Composition
Argumentation and thesis development
Nathan teaches first-year composition and directs a university writing center. He reviews EssayCrackers guides for argumentative soundness and citation accuracy.