HSCI 312 · Lesson 11

Evaluation and Careers
in Health Promotion

Health Promotion: Individuals and Communities

Learning objectives for this lesson:

  • Explain what evaluation is in plain terms, why it is needed, and how the accountability movement and evidence-based practice made it a standard part of the program environment.
  • Describe how theory and evaluation are linked: a theory-based program specifies what ought to change, and therefore what an evaluation has to measure.
  • Distinguish process, outcome, and impact evaluation by the question each answers, the data each needs, and the time frame each requires, and use related terms such as fidelity, dosage, baseline, and formative evaluation correctly.
  • Build a basic logic model that links inputs, the health problem, outputs, short-term outcomes, and long-term impacts, and connect each component to the type of evaluation that tests it.
  • Compare evaluation designs from simple record keeping to the classic experiment, and identify the confounds that weaken a claim that a program caused a change.
  • Describe the career settings in which social and behavioural theory is applied, both as a general typology and in Canadian public health, and explain how theory is used in each.

This course was developed by Dr. Kiffer G. Card, Faculty of Health Sciences, Simon Fraser University, to accompany Edberg, M. Essentials of Health Behavior: Social and Behavioral Theory in Public Health. Jones & Bartlett Learning. This lesson follows Chapters 14 and 16 of the text.

Reference

Glossary: Key Terms, People & Concepts

📚 Reference page, available throughout the lesson

This glossary collects the key concepts, people, and ideas you will meet in this lesson. Use it as a reference while you work through the material, or as a review before assessments. Type in the search box to filter entries.

Core concepts
Evaluation In plain terms, the process of making sure you did what you proposed to do, determining whether the program had an effect based on what you were trying to achieve, and assessing whether the program model and the theory used to design it were useful. It may also cover cost-effectiveness and whether community goals were met.
Accountability The requirement to show that public or donor resources invested in a program had some effect. In the United States the Government Performance and Results Act of 1993 and the Program Assessment Rating Tool made performance monitoring a condition of funding; Canadian federal and provincial bodies apply comparable results policies.
Evidence-based practice The move within public health to build a body of interventions supported by evaluation evidence, in the same way that medicine developed evidence-based standards of care. Evaluation data are the evidence; registries compile and rate the programs that have it.
Model program A program that has demonstrated good evidence of effectiveness, especially one implemented in several places with different populations and found effective in each. Funders increasingly require grant recipients to use model or evidence-based programs, or to justify why none fits the community.
Fidelity The degree to which a program is implemented according to its design, with all components delivered as planned. Replicating a program with fidelity is what allows the replication to serve as an additional test of the model; process evaluation supplies fidelity data.
Dosage In a health promotion or other nonmedical intervention, the amount of the intervention (number of education sessions, number of activities, and so on) that results in a particular outcome or effect.
Baseline Data collected on the things a program expects to change before the program begins, so that the same information collected at one or more follow-up points can measure change. Also called pretest data; the later collection is posttest or follow-up data.
Inputs The resources, staff, program components, funds, and the like invested in or used for an intervention. Some of this information comes from the administrative and policy assessment of PRECEDE-PROCEED.
Outputs The activities a program delivers, such as educational materials, community events, health screenings, or trainings, together with the factors each activity is intended to change. The choice of activities is guided by the factors identified in assessment and by the theory judged appropriate.
Indicators and measures The criteria and tools for measuring outcome and impact: health data, survey or qualitative data, measures of skill, and the like. They are the last component of a logic model and they tie theory to observable evidence.
Confound An unplanned occurrence or other factor that may affect an intervention and complicate the claim that only the intervention caused a change. Seven common confounds are history, maturation, testing, regression to the mean, selection bias, mortality or attrition, and diffusion of treatment.
Independent variable A characteristic of a person or situation that the intervention is not trying to change; something that just is, such as being a woman or a recent immigrant. Such variables tell you a great deal about who and what the program can change.
Dependent variable A characteristic of a person or situation that the intervention is trying to change, such as diet or exercise behaviour. It is what an evaluation of outcome or impact measures.
Quantitative data Information that can be expressed numerically and analyzed statistically, such as fixed-choice survey answers or official demographic and epidemiological records. Because it is collected the same way from many people, it supports comparisons and generalized conclusions about a population.
Qualitative data Narrative, descriptive, and subjective information from extended interviews, focus groups, and observation, usually gathered from fewer people. It reveals the details and meaning of events, how people interpret situations, and the context in which behaviour occurs.
Program officer A staff member of a government or public funding agency who helps decide the program framework behind a funding announcement and then manages or oversees the organizations selected to carry out the work, providing technical support and monitoring evaluation, without delivering the program directly.
Advocacy organization A nonprofit whose focus is increasing public engagement about an issue and affecting public policy on it, typically through media advocacy, social marketing, other communications efforts, and direct policy work such as position papers and draft legislation.
Types, models, and methods
Process evaluation Evaluation of how implementation fared. Its basic question is whether the components of the intervention were implemented as planned, answered by keeping records of materials, sessions, partners, and participation and comparing what was done with what was planned. It also supplies fidelity and dosage data.
Outcome evaluation Evaluation of the short-term or immediate effect of an intervention, measured by collecting baseline and follow-up data on the things theory and assessment predicted would change, such as knowledge, service use, or policy. In the convention this course uses, outcome refers to the short term; Green and Kreuter's PRECEDE-PROCEED framework uses the label the other way around.
Impact evaluation Evaluation of whether the intervention affected the health problem that was its ultimate target, for example cancer morbidity and mortality, measured with baseline and follow-up data over an extended period or by following a cohort of participants. In the convention this course uses, impact refers to the long term; Green and Kreuter's PRECEDE-PROCEED framework uses the label the other way around.
Formative evaluation Evaluation used during program development to shape the program before or as it is implemented, for example by pretesting materials or gathering feedback. This lesson names it, with cost-effectiveness and quality assurance evaluation, as a type it does not cover in detail.
Logic model A diagram or structure that links what a program plans to do with its expected outcomes and impacts. It connects the health problem, inputs, outputs, projected short-term outcomes and long-term impacts, and the indicators and measures used to evaluate them.
Quasi-experimental design An outcome and impact evaluation design that identifies a community or population similar to the intervention sample but not receiving the intervention, and collects pre- and posttest or time series data from both. Its problems are comparability, contamination of the comparison group, and other events in the intervention community.
Classic experimental design The most rigorous evaluation design: individuals in the target population are randomly assigned to the intervention or to serve as controls, and pre- and posttest or time series data are collected from both groups. Its problems include accidental exposure of controls and the ethics of withholding an intervention from a community that needs it.
Empowerment evaluation A collaborative approach, drawing on Paulo Freire's community theories and participatory programming, in which the community identifies its goals and the best ways to measure them, often instead of standard validated instruments. Its key purpose is to help a community improve its program on its own terms.
Guide to Community Preventive Services (Community Guide) A compendium from the United States Centers for Disease Control and Prevention that summarizes what is known about the effectiveness, economic efficiency, and feasibility of community interventions, based on systematic reviews, with recommendations from the Task Force on Community Preventive Services.
People
Paulo Freire Brazilian educator whose Pedagogy of the Oppressed (1970) argued that people should analyze and act on the conditions of their own lives. Empowerment evaluation traces its roots to his community theories and to participatory programming more broadly.
David Fetterman Evaluator whose Foundations of Empowerment Evaluation (2000) set out the approach described in this lesson: a collaborative effort with a community to define goals and measures, not an experimental design meant to produce rigorous comparative data.
Lawrence Green and Marshall Kreuter Developers of the PRECEDE-PROCEED planning framework, whose sequence of social, epidemiological, behavioural and environmental, educational and ecological, and administrative and policy assessments forms the logical chain that this lesson turns into a logic model.
No matching entries. Try a different search term.
Section 1

Evaluation: What It Is and Why It Matters

⏱ Estimated reading time: 15 minutes

Section 1 of 4

Evaluation: What It Is and Why It Matters

Did the whole process, from assessment to theory to program, make a difference?

Without the jargon

Evaluation is a short list of questions

  • Did we do what we proposed to do?
  • Did the program have an effect, judged against what we were trying to achieve?
  • Were the program model and the theory behind it useful?
  • Sometimes: what did it cost, and were community goals met?
The program environment

The accountability movement

1980s

Consensus that better evaluation was needed to improve planning and performance.

1993 onward

The Government Performance and Results Act, then the Program Assessment Rating Tool: monitor, report, account.

Planning documents

Stated goals become the measure of success. Healthy People in the United States; Treasury Board results policies in Canada.

Evidence-based practice

Evaluation data are the evidence

Evaluate

Without evaluation there is no evidence to enter into the record.

Registries rate

The Community Guide, NREPP, Health Evidence. Replication across settings counts most.

Funders require

Use a model program, or justify why none fits your community.

Why evaluate?

Four reasons

Accountability

Show that resources had an effect; meet reporting requirements.

Learning

See what is working and change course as you go.

Theory

Test a theoretical linkage or a model program in a new setting.

Efficiency

What did it cost to achieve the goals?

Theory and evaluation

Theory says what ought to change

Assessment identifies the factors. Theory predicts how changing them changes behaviour and health. Evaluation checks whether the predicted changes occurred.

The result is evidence about the program and, at the same time, evidence about the theory. A shortfall can mean the theory did not fit, or that the program was not delivered as designed.

Carry forward

What to take into the next section

  • Evaluation: did we do it, did it work, was the theory useful, what did it cost?
  • Accountability and evidence-based practice make evaluation a condition of funding.
  • Theory names the changes an evaluation must measure.
  • Next: process, outcome, and impact evaluation, and the questions each answers.

Evaluation: what it is and why it matters

Learning objectives for this section

  • State, in plain terms, what evaluation is and what it asks about a program.
  • Explain how the accountability movement and the move to evidence-based practice made evaluation a standard part of the program environment, in the United States and in Canada.
  • Name four reasons for evaluating and sort real situations into them.
  • Describe how evaluation carries the use of theory through to a judgment about whether the whole process made a difference.

Every lesson in this course has followed the same path: assess a situation, identify the factors that matter, choose a theory that explains them, and design a program. This lesson asks the question that path has been leading toward. Did it work? This is not a course on evaluation, and the lesson will not go into technical detail. Its aim is familiarity, because evaluation is so much a part of the program environment that anyone who plans programs will eventually be doing it. This section covers what evaluation is, why the funding environment now demands it, and how it connects to the theories you have studied.

Evaluation without the jargon

Evaluation sounds formal and technical. Stripped down, it is a short list of questions. Click each card for the question and an example.

Did we do what
we proposed?
Click to learn more
Did it have
an effect?
Click to learn more
Was the theory
useful?
Click to learn more
What else
matters?
Click to learn more

The point can be summarized in one sentence: evaluation is relevant to the entire scope of this course, because it is the way you carry through your use of social and behavioural theory to the point where you can determine whether the whole process, of assessing a situation, identifying factors to address, and using the assessment with relevant theory to plan and implement an intervention, made a difference.

The current program environment

There is also a pragmatic reason to evaluate. In the current and future environment of program funding there is a strong emphasis on evidence and accountability. To keep funding, a program needs evidence that it is doing something about the problem it was meant to address. This imperative is relatively recent. Serious evaluation of health promotion and behavioural interventions did not become the norm until a consensus formed among researchers and practitioners in the 1980s that better evaluation was needed to improve program planning and performance.

In the United States the turning point was legislative. The Government Performance and Results Act (GPRA) of 1993, followed by the Program Assessment Rating Tool (PART) of the White House Office of Management and Budget, required agencies to develop performance monitoring and accountability procedures to ensure that public funds were well spent. The influence of this accountability movement has since been integrated into standard practice at every level, including private and global health promotion efforts. One consequence is the spread of strategic planning documents that set goals as a guide to action and as a method for evaluating progress. If the plan says "we will reduce the number of youth who smoke by 25%," then documented change in youth smoking, compared with that goal, becomes the measure of success. The Healthy People documents of the United States Department of Health and Human Services (Healthy People 2010, Healthy People 2020, and their successors) are large, comprehensive planning documents of this kind.

The Canadian version of the accountability movement

Canada arrived at the same place by a different route. Federal departments and agencies, including the Public Health Agency of Canada, work under Treasury Board results policies that require them to set expected results, report performance against them, and evaluate their programs on a regular cycle. Provincial bodies apply similar expectations to the organizations they fund: Ontario's public health standards, for example, expect boards of health to plan and evaluate programs using evidence, and regional health authorities in British Columbia report against provincially set performance measures. The Canadian Evaluation Society offers a Credentialed Evaluator designation, a sign of how much of a profession program evaluation has become. The practical effect for a health promotion program is the same on both sides of the border: a planning document states a goal, and the evaluation measures progress against it.

Evidence-based practice and model programs

An even more profound change than accountability, arguably, is the move within public health to develop a body of evidence-based practice, in the same way that medicine developed evidence-based standards of care. For health promotion and prevention this means three things. First, evaluating the program becomes very important, because evaluation data are the evidence. Second, many agencies have developed directories of interventions that have evidence of effectiveness, and these directories often rate programs by the quality of their evidence. A program implemented in several places with different populations and found effective in all of them is rated more highly than a program tried once, even if it worked that one time. Third, programs with good evidence are called model programs or something similar, and public and even private funders increasingly require grant recipients to use a model program, or to justify why no appropriate program of that type exists for the community, population, or situation to be addressed.

Registry or guideWho maintains itWhat it does
Guide to Community Preventive Services (the Community Guide)United States Centers for Disease Control and Prevention, with the Task Force on Community Preventive ServicesActs as a filter for a scientific literature that can be large, inconsistent, uneven in quality, and inaccessible. Systematic reviews summarize the effectiveness, economic efficiency, and feasibility of community interventions, and the Task Force issues recommendations.
National Registry of Evidence-Based Programs and Practices (NREPP)United States Substance Abuse and Mental Health Services AdministrationProvided descriptive information and peer-reviewed ratings of outcome-specific evidence for interventions that prevent or treat mental and substance use disorders. The registry in its original form was closed in 2018 and replaced by a smaller evidence resource.
Health EvidenceMcMaster University, Hamilton, OntarioA searchable, quality-rated collection of systematic reviews on the effectiveness of public health interventions, built for Canadian practitioners and decision makers.
National Collaborating Centre for Methods and ToolsOne of six National Collaborating Centres for Public Health funded by the Public Health Agency of Canada, hosted at McMaster UniversitySupports evidence-informed decision making in Canadian public health with methods, tools, and training for finding and appraising evidence.

The consequence for research is that a great deal of effort now goes into identifying model programs, which means building or locating the evidence base, and along with that, compiling evidence for the usefulness or inapplicability of the theoretical approaches used to structure those programs. Every well-evaluated Health Belief Model program, whether it succeeds or fails, is also a data point about the Health Belief Model.

Common confusion

"Evidence-based" does not mean a program was tested with a randomized trial. A registry rates the quality of evidence on a scale, and replication across settings and populations counts heavily. As Section 2 will show, a process evaluation that documents faithful implementation is part of the evidence too, because without it no one can tell whether a disappointing result reflects the theory or the delivery.

Why evaluate? Four reasons

Four reasons to evaluate cover most real situations. Accountability responds to the public need to show that the cost and resources invested in a program had some effect, and, since GPRA and PART, to legal requirements for monitoring systems. Learning and improvement means that as a program runs, evaluation data show what is working and what is not, feedback that lets you make changes as you go. Theory means testing the validity of a theoretical linkage, or testing a model program or approach that has been shown to work somewhere else. Efficiency and other issues covers what it costs for the program to achieve its goals, and other kinds of assessment. Use the sorter to practise telling them apart.

Interactive: why is this evaluation happening? Click a situation, then click the reason from the list of four that best fits it. The feedback explains the reasoning.
Accountability
Learning and improvement
Theory
Efficiency and other issues
Select a situation to begin.
0 of 8 placed

How evaluation relates to theory

How does evaluation relate to theory? The answer runs through the whole course. A theory-based program is a set of predictions. The Health Belief Model predicts that if people come to see themselves as susceptible to a serious condition, and see a course of action whose benefits outweigh its barriers, they will act. Social Cognitive Theory predicts that observing a credible model and building self-efficacy will change behaviour. Each prediction names something that ought to change, and each of those things is something an evaluation can measure. Theory, in other words, tells the evaluator what to look for. The diagram shows the loop.

Assessment what is the problem, and why Theory what ought to change, and how Intervention activities aimed at those factors Evaluation did the predicted changes occur? evidence about the theory and about the program records, surveys, health data

The loop also explains why evaluation can be called the way to test theory. When a program falls short, evaluation can separate two very different explanations: the theory did not fit this population, or the program was not delivered as designed. Section 2 returns to this distinction under the heading of process evaluation and fidelity. For now, hold on to the idea that a program without a theory has nothing specific to evaluate except whether it happened, and a theory without an evaluation is a prediction nobody checked.

Case study: The funder asks for evidence

A community organization in Winnipeg has run a peer-led program for three years in which trained older teens deliver sessions on vaping and nicotine to grade 7 and 8 classes. The sessions are built on Social Cognitive Theory: the peers model refusal skills, students rehearse them, and the program aims to raise students' confidence that they can turn down a vape without losing face. The provincial funder's renewal letter asks two questions. Is the program evidence-based? And what evidence does the organization have that it is working? The director has attendance sheets, a stack of thank-you notes from teachers, and a vague memory of a similar program in Ontario that was rated well in a registry. She has no baseline data and no follow-up survey.

Which of the four reasons to evaluate does the funder's letter reflect, and which reason would the director invoke if she argued the case on her own terms? What would she need to have collected, starting three years ago, to answer whether the Social Cognitive Theory prediction (higher refusal self-efficacy, then less vaping) held?

The director's problem is common, and it points to the rest of the lesson. Attendance sheets answer the first of the evaluation questions (did we do what we proposed?) but not the second or third. To answer those, she would have needed to decide at the start what the theory said should change, measure it before the program, and measure it again afterward. The next section names those different kinds of evaluation and the questions each one answers.

Reflection

A nonprofit in your community runs a sexual health program for newcomer youth, built on Social Cognitive Theory (peer role models, skills rehearsal, self-efficacy). Its board has asked the coordinator to "do an evaluation" before the next funding cycle, but nobody has said why. Using the four reasons to evaluate, write a short memo to the board explaining what each reason would require the program to measure, and then explain in two or three sentences how the program's theory determines what the evaluation should look for.

Model answerA strong answer works through all four reasons and ties each to a measurement. Accountability: the funder will want evidence that the money produced an effect, so the program needs to show, against the goals in its proposal, that the planned sessions happened and that the intended change (for example, condom use or testing among participants) moved. Learning and improvement: session feedback, attendance patterns, and facilitator notes collected while the program runs show what to adjust now. Theory: Social Cognitive Theory predicts that observing peer models and rehearsing skills will raise self-efficacy, and that self-efficacy will change behaviour, so the evaluation should measure self-efficacy before and after the program and check whether behaviour followed. Efficiency: the board should know what each participant and each behaviour change cost, so that it can compare the program with alternatives. The final point is the one this section stresses: the theory names the links in the chain (modelling, rehearsal, self-efficacy, behaviour), so it tells the coordinator exactly what to collect at baseline and follow-up. Without that theory the program could only report that it happened.

Minimum 20 characters required.

✓ Reflection saved

Key Takeaways

  • In plain terms, evaluation means making sure you did what you proposed, determining whether the program had an effect on what it was trying to achieve, and assessing whether the program model and theory were useful; sometimes it also covers cost-effectiveness and community goals.
  • Serious evaluation of health promotion became the norm only after a 1980s consensus that better evaluation was needed. The Government Performance and Results Act of 1993 and the Program Assessment Rating Tool made accountability a condition of public funding in the United States; Canadian federal results policies and provincial standards play the same role here.
  • The move to evidence-based practice means evaluation data are the evidence. Registries such as the Community Guide and Health Evidence compile and rate interventions, replication across settings counts heavily, and funders increasingly require model programs or a justification for not using one.
  • The four reasons to evaluate are accountability, learning and improvement, theory, and efficiency and other issues.
  • Theory tells the evaluator what to look for: a theory-based program predicts specific changes, and evaluation checks whether they occurred, producing evidence about the program and the theory at the same time.
Knowledge Check: this section

1. A regional health authority tells a community organization that its harm reduction outreach program will be renewed only if it can document that the funds produced a measurable effect. Which of the four reasons to evaluate does this request reflect?

Accountability is the need to show that the cost and resources invested in a program had some effect, and it is tied to the reporting requirements that followed the accountability movement. Learning and improvement would be the organization using data to adjust the program as it runs; theory would be testing a predicted linkage; efficiency would be asking what each result cost.

2. Why is a program that has been implemented in several places with different populations, and found effective in each, rated more highly in an evidence registry than a program tried once with success?

Directories rate programs on the quality of their evidence, and a program that shows effectiveness across several settings and populations earns a higher rating than one that worked a single time. Randomized trials are not a requirement for listing, and cost and implementation quality are separate questions.

3. A youth mental health program built on Social Cognitive Theory shows no change in help-seeking. Which statement best captures how theory and evaluation are linked in interpreting this result?

Evaluation is the way to carry theory through to a judgment about whether the whole process made a difference. A theory-based program predicts a chain of changes, and measuring each link lets the evaluator separate a theory that did not fit this population from a program that was not implemented faithfully. A single disappointing result neither refutes the theory nor makes the program a model program.

4. The accountability movement accelerated the use of strategic planning documents. In evaluation terms, what role does such a document play?

Consider a plan that says youth smoking will fall by 25%: documented change in smoking, compared with that goal, is then the measure of success. Planning documents set goals as a guide to action and a method for evaluating progress; they do not supply baseline data, list model programs, or design the evaluation.

✦ Pass the knowledge check with 100% and complete the reflection to continue

Section 2

Process, Outcome, and Impact Evaluation

⏱ Estimated reading time: 15 minutes

Section 2 of 4

Process, Outcome, and Impact Evaluation

Three questions, three kinds of data, three time frames.

Three questions

Process, outcome, impact

Process

Were the components of the program implemented as planned?

Outcome

What short-term or immediate effect did the program have?

Impact

How did the program affect the health problem that was the ultimate target?

Process

Did we do what we planned?

Fidelity

All components delivered according to the design. Without it, a replication tests nothing.

Dosage

The amount of the intervention (sessions, activities) that results in an effect.

Records of sessions, materials, partners, and participation separate a failed theory from a failed delivery.

Outcome

Short-term effects, measured before and after

Baseline, then one or more follow-ups, on the things theory said would change first.

  • Knowledge, attitudes, skills
  • Health service utilization
  • Policy or regulatory change
  • The reach of a three- or four-year program
Impact

The health problem itself

Awareness

Susceptibility, benefits, barriers: the short-term outcome.

Quitting

The behaviour change the theory predicts will follow.

Cancer

Morbidity and mortality, years later: the long-term impact.

A note on labels

Which word means long term?

This course

Outcome = short term. Impact = long term.

PRECEDE-PROCEED

Impact = immediate effects on factors and behaviour. Outcome = health and quality of life.

Same idea, swapped labels. Check definitions before comparing results.

Carry forward

What to take into the next section

  • Process: implemented as planned? Fidelity and dosage.
  • Outcome: short-term change in targeted factors, baseline to follow-up.
  • Impact: the health problem itself, over years or a cohort.
  • Next: the logic model that gives each type its place, then methods and confounds.

Process, outcome, and impact evaluation

Learning objectives for this section

  • State the question that process, outcome, and impact evaluation each answer, and the kind of data each needs.
  • Explain why process evaluation matters for testing theory, using the ideas of fidelity and dosage.
  • Describe the pre- and posttest (baseline and follow-up) logic of outcome and impact evaluation and the time frames each requires.
  • Recognize that different sources label short- and long-term effects differently, and read a study's definitions before comparing results.

Once you look closely at evaluation it becomes more complex, because there are many things one could evaluate in any program and several goals an evaluation can serve. The way to make this manageable is to start with three basic categories: process, outcome, and impact. A project may use one, two, or all three at once, because each has a different purpose. Other types exist that this lesson does not cover in detail: cost-effectiveness evaluation, formative evaluation used during program development, and quality assurance evaluation. This section works through the three basic types, the vocabulary that comes with them, and a Canadian program that needs all three.

Three questions

The simplest way to hold the three types apart is by the question each one asks. The tabs give the question, the way it is answered, and where the answer comes from.

Process evaluation

Question: Were the components of the intervention implemented as planned?

How: keep records. How many materials were developed and distributed, how many education sessions were held, how many community coalition partners took part, how many people showed up. Then compare what you actually did with what you planned to do.

When: throughout implementation. The historical recordkeeping approach, which tracks project activities and client participation, is the basic process evaluation method.

Outcome evaluation

Question: What short-term or immediate effect did the intervention have?

How: collect data on the things you expect to change in the short term before the project begins, then collect the same information at one or, preferably, more points after the intervention is in place. This is pre- and posttest data, or baseline and follow-up data.

What: depending on program goals, knowledge change, changes in health service utilization, policy changes, or other educational, ecological, or policy data. It is the kind of change a program funded for three or four years could expect.

Impact evaluation

Question: Did the intervention affect the health problem or issue that was the ultimate target?

How: the same baseline and follow-up logic, but the data concern the long-term impacts you expect, such as cancer morbidity and mortality, and follow-up continues for years.

When: only if the project is in place for an extended time (more than three or four years) or if you can follow the people who took part, a cohort, over an extended period.

Types this lesson names but does not cover

Formative evaluation is used in program development, before or as a program is rolled out, to shape it: pretesting messages, piloting a session, gathering feedback from the intended audience. Cost-effectiveness evaluation asks what it costs to achieve a unit of result. Quality assurance evaluation checks that services meet defined standards. Qualitative approaches can be part of formative work as well as of process, outcome, and impact evaluation.

Process evaluation: did we do what we planned?

Process evaluation is more than bookkeeping. It matters for testing theory. Suppose an intervention is based on the stages of change model from the Transtheoretical Model. Its effectiveness may depend in part on periodically assessing each participant to identify their stage in changing, say, their diet, because the program's components are tied to a stage for each individual. If the intervention turns out less effective than expected, it would be very helpful to know, through process evaluation, whether the stage assessments were done properly. If they were not, then at least you know that the problem may lie in the details of implementation rather than in the theory or the intervention design. Two terms follow from this. Click each card.

FidelityClick to learn more
DosageClick to learn more
Record
keeping
Click to learn more
Baseline and
follow-up
Click to learn more

Outcome evaluation: short-term effects

With outcome evaluation you start looking at the effects of a program. How you answer the question depends on your program goals. What did you expect to happen in the short term, and was that expectation guided by theory and by your assessment? The expected outcome has everything to do with theory, assessment, and how these were incorporated into the program design, for example through the PRECEDE-PROCEED process. The outcome data might be pre- and posttest scores on knowledge, records of health service use, evidence of a policy change, or other educational, ecological, or policy data. The diagram shows the timeline that outcome and impact evaluation share, and where the two part company.

Baseline pretest data Intervention process evaluation: records Follow-up 1 Follow-up 2 Outcome evaluation knowledge, attitudes, skills, service use, policy (a three- or four-year program) Years later Impact evaluation behaviour change sustained, morbidity and mortality

Impact evaluation: the health problem itself

Impact evaluation reaches the real health impact of a program, and with it the theoretical links you may have drawn between short-term outcome and long-term impact. Consider a smoking intervention with several activities designed to raise awareness of risks and of prevention options. The hypothesis, drawn from the Health Belief Model, is that if a group of people come to see their own susceptibility to cancer from smoking, and see quitting options that do not present an insurmountable barrier, they are likely to quit. The program is therefore linking awareness change to behaviour change, and behaviour change to a reduction in cancer. Awareness is the short-term outcome. Cancer is the long-term impact. Measuring the impact requires a program that stays in place for an extended period, or the ability to follow participants as a cohort for years, and it means collecting follow-up data on cancer morbidity and mortality rather than on awareness.

Common confusion: which word means long term?

The terms outcome evaluation and impact evaluation are attached to different time frames by different sources in the field, even though the underlying idea of short- and long-term evaluation is the same everywhere. You will meet both conventions. In the convention this course uses, which follows much current usage, outcome is short term and impact is long term. In Green and Kreuter's PRECEDE-PROCEED framework the labels run the other way: impact evaluation covers the immediate effects on predisposing, enabling, and reinforcing factors and on behaviour, while outcome evaluation covers health and quality of life. Some public health evaluation frameworks avoid the problem by speaking of short-, intermediate-, and long-term outcomes. Whenever you read an evaluation report, check its definitions before you compare its results with anyone else's.

TypeThe questionTypical dataTime frameWhat theory and the logic model say
ProcessWere the components of the program implemented as planned?Counts of materials, sessions, partners, participants; fidelity and dosage recordsDuring implementationWhether the planned outputs were produced from the inputs
OutcomeWhat short-term or immediate effect did the program have?Pre- and posttest knowledge, attitudes, skills; service utilization; policy changeWithin a program of three or four yearsWhether the predisposing, enabling, and reinforcing factors the theory targeted moved
ImpactHow did the program affect the health problem that was the ultimate target?Sustained behaviour change; morbidity and mortality; health statusMore than three or four years, or a followed cohortWhether the chain from factors to risk to health problem held

Practise: process, outcome, or impact?

Each type reduces to one question, and for outcome and impact the answer is what your logic model and theory say should have happened. Use the classifier below to test yourself on evaluation questions drawn from Canadian programs.

Interactive: which type of evaluation asks this? Read the question, choose a type, then move to the next one. Ten questions.
Were all eight community kitchen sessions delivered in the Saskatoon neighbourhood, and how many people attended each one?
Six months after enrolling in the provincial cessation program, what proportion of participants report that they have quit smoking?
Fifteen years on, has lung cancer incidence fallen among the cohort of people who used the cessation program?
Did the youth workers complete the stage-of-change assessment with each participant, as the program design required?
Did knowledge of the signs of stroke increase between the baseline survey and the follow-up survey in the campaign region?
Has the rate of new type 2 diabetes diagnoses in the county changed over the ten years since the program began?
How many pharmacies dispensed nicotine replacement products under the program, and how many residents picked them up?
Did the municipal council adopt the smoke-free parks bylaw that the coalition was working toward?
Has the proportion of grade 9 students who say they would seek help for depression increased since the mental health literacy curriculum began?
Did emergency department visits for self-harm among youth in the district fall over the following decade?
Question 1 of 10
Choose the type of evaluation that asks this question.
0 answered

Case study: Evaluating British Columbia's Smoking Cessation Program

Since 2011 British Columbia's Smoking Cessation Program has supplied nicotine replacement products such as patches and gum at no cost through community pharmacies, and has covered prescription cessation medications through the provincial PharmaCare plan. The program rests on a familiar logic: a barrier (cost) is removed, a cue to action arrives when a pharmacist offers the product, and the province's QuitNow service provides support for the attempt. Imagine you have been asked to evaluate it. A process evaluation would count participating pharmacies, products dispensed, and referrals to counselling, and would check whether pharmacists delivered the brief advice the program design calls for. An outcome evaluation would measure quit attempts and self-reported abstinence among participants at follow-up, compared with their smoking at enrolment. An impact evaluation would ask whether smoking prevalence in the province, and eventually tobacco-related disease, changed over the years that followed, a question that needs population data over a long period and cannot be answered by the program's own records.

The ministry wants "results" within the program's first eighteen months. Which type of evaluation can honestly deliver results in that time, which cannot, and how would you explain the difference using the distinction between short-term outcomes and long-term impacts?

Case study: A replication that disappointed

A school district in Alberta adopted a mental health literacy curriculum that had shown good results in another province. At follow-up, students' knowledge and attitudes had barely moved. The superintendent concluded that the curriculum did not work for Alberta students. The district's process evaluation, however, showed that half the teachers had delivered only two of the six sessions, that the teacher training day had been cut to a lunch-hour briefing, and that the referral pathway was never set up.

Using the ideas of fidelity and dosage, what can and cannot be concluded from this replication about the curriculum and the theory behind it?

The second case is the stages-of-change example from earlier in this section in different clothes. The process evaluation showed that the replication lacked fidelity and delivered a low dose, so the disappointing outcome says little about the curriculum and nothing about the theory. It says a great deal about implementation. The next section puts all three types of evaluation into a single structure, the logic model, and introduces the methods and pitfalls of making a case that a program caused a change.

Reflection

A public health unit in Ontario is launching a three-year program in which public health nurses visit new parents at home to promote safe infant sleep, using modelling and skill building from Social Cognitive Theory. Write one process question, one outcome question, and one impact question for its evaluation. For each, say what data would answer it, when the data would be collected, and whether the answer could be available before the program's funding ends.

Model answerA strong answer keeps the three questions distinct and honest about time. Process: Were the planned home visits made to every eligible family, and did nurses deliver the modelling and practice components as designed? Data: visit logs, checklists of components delivered, and counts of families reached, collected throughout; fully available during the program. Outcome: Did parents' confidence in setting up a safe sleep space (self-efficacy) and their reported sleep practices change between the first visit (baseline) and a follow-up visit or survey some months later? Data: the same questions asked before and after, plus perhaps observation of the sleep space; available within the three years. Impact: Did sleep-related infant deaths in the unit's area fall compared with the years before the program? Data: coroner and mortality records over many years; because such deaths are rare and the time frame is long, this answer would not be available before funding ends, and the outcome-impact distinction explains why. Outcome is the short-term change in the factors the theory targeted; impact is the health problem itself, which needs an extended program or a followed cohort. A good answer also notes that the process data protect the interpretation: if outcomes disappoint, they show whether the program was delivered with fidelity.

Minimum 20 characters required.

✓ Reflection saved

Key Takeaways

  • Process evaluation asks whether the components of the program were implemented as planned, answered by keeping records and comparing what was done with what was planned. It supplies fidelity and dosage data, and it lets you tell a failed theory from a failed delivery.
  • Outcome evaluation asks what short-term or immediate effect the program had. It collects baseline (pretest) data on what theory and assessment said would change, then the same data at one or more follow-up points. Knowledge, service use, and policy changes are typical outcomes for a program of three or four years.
  • Impact evaluation asks whether the program affected the health problem that was its ultimate target, such as cancer morbidity and mortality. It needs a program in place for an extended period or a cohort followed for years.
  • Different sources in the field attach the labels outcome and impact to different time frames. This course uses outcome for the short term and impact for the long term, following much current usage; Green and Kreuter's PRECEDE-PROCEED uses them the other way around. Always check a report's definitions before comparing results.
  • Formative, cost-effectiveness, and quality assurance evaluation exist alongside the three basic types; a project may use one, two, or all three of the basic types at once.
Knowledge Check: this section

1. A coordinator compares the number of cooking classes actually delivered in each community with the number promised in the proposal, and records how many sessions each participant attended. Which type of evaluation is this, and what related concept do the attendance records support?

Comparing what was done with what was planned is the basic process evaluation question, answered by record keeping. Attendance records also give dosage, the amount of the intervention each person received, which an outcome analysis can then take into account.

2. A stage-based diet program shows weaker results than expected. Process data reveal that staff skipped the periodic stage assessments that the program design required. What does this tell the evaluator?

This is the classic example of process evaluation testing fidelity. Because the program components are tied to each participant's stage, skipping the assessments means the program was not delivered as designed. The process evaluation therefore points to implementation, and the disappointing result cannot be read as evidence against the theory.

3. A provincial cessation program has been running for two years. Which of the following is an impact evaluation question in the sense this course uses?

Impact evaluation asks how the program affected the health problem that was the ultimate target, measured over an extended period. Pharmacy counts are process data; six-month quit rates and belief change are short-term outcomes measured from baseline to follow-up.

4. Two reports evaluate similar school programs. One, following the convention used in this course, calls the change in students' attitudes an "outcome"; the other, following PRECEDE-PROCEED, calls the same change an "impact." What does the difference in terminology imply for a reader?

Different sources in the field attach the two labels to different time frames, while the underlying idea of short- and long-term evaluation is the same everywhere. The convention used in this course follows much current usage; Green and Kreuter's PRECEDE-PROCEED attaches the labels the other way around. Neither report is wrong; the reader must translate before comparing.

✦ Pass the knowledge check with 100% and complete the reflection to continue

Section 3

Logic Models, Methods, and Confounds

⏱ Estimated reading time: 18 minutes

Section 3 of 4

Logic Models, Methods, and Confounds

Setting up an evaluation, and making a case that will hold.

The logical chain

Back up the chain

Factors

Predisposing, enabling, reinforcing: what the intervention targets.

Risks

Behavioural and environmental risk factors.

Health problem

Morbidity and mortality in the affected group.

Context

Social conditions; administrative and policy resources.

The parts

Five linked components

Inputs: resources and staffProblem / goal: the health problemOutputs: activitiesShort-term outcomesLong-term impactsIndicators and measures

Each column is also a type of evaluation: process for outputs, outcome for the short term, impact for the long term.

Worked example

Harfield County

Outputs

Outreach; diet classes and cooking groups through churches and veterans centres; diet modelling by leaders.

Short-term outcomes

Knowledge of food risk and improved diet; skills to choose healthier diets.

Long-term impact

Reduced incidence of diabetes and its consequences.

Methods

A ladder of designs

  • Historical record keeping: are inputs producing the planned outputs?
  • Qualitative approaches: what changed, from whose viewpoint, and why?
  • Periodic inventory / time series and benchmarking: is progress being made, and how does it compare?
  • Quasi-experimental: a similar community that did not get the program.
  • Classic experimental: random assignment to intervention or control.
Confounds

Seven ways a case gets weaker

HistoryMaturationTestingRegression to the meanSelection biasMortality / attritionDiffusion of treatment

Independent variables just are (being a recent immigrant); dependent variables are what the program tries to change (diet, exercise).

Fit

No cookie-cutter standard

Quantitative

Numbers collected the same way from many people: comparison and generalization.

Qualitative

Interviews, focus groups, observation: meaning, context, and the dynamics of change.

Process when getting started is the achievement; outcome when time is short; impact when the program is long enough to show it.

Carry forward

What to take into the next section

  • A logic model gives each type of evaluation a column, and indicators make the columns measurable.
  • Designs trade rigour against cost, feasibility, and ethics; confounds are the alternatives a design must rule out.
  • Fit the evaluation to the program; qualitative and quantitative data answer different questions.
  • Next: the career settings where theory and evaluation are put to work, in general and in Canada.

Logic models, methods, and confounds

Learning objectives for this section

  • Explain how the PRECEDE assessments form a logical chain that an intervention and its evaluation travel back up.
  • Name the components of a logic model and place program elements into inputs, problem, outputs, short-term outcomes, and long-term impacts.
  • Compare evaluation designs from record keeping to the classic experiment, and state what each can claim.
  • Identify the confounds that weaken a causal claim, and distinguish independent from dependent variables.
  • Explain why there is no cookie-cutter standard for evaluation.

The last section separated three types of evaluation. This one puts them into a single structure, the logic model, which we reach by way of PRECEDE-PROCEED. It then turns to method: how do you make the case that a program caused a change? The answer is a ladder of designs, a list of confounds, and a reminder that the right evaluation is the one that fits the program.

From PRECEDE to a logical chain

Recall the PRECEDE assessments: social (the community context and quality-of-life issues), epidemiological (the health problems and affected group, from morbidity and mortality data), behavioural and environmental (the risk factors behind those problems), educational and ecological (norms, attitudes, awareness, and policies understood as predisposing, enabling, and reinforcing factors), and administrative and policy (resources, community politics, and structures that help or hinder implementation). Can you see the logical chain? Health problems result from risks, which result from predisposing, enabling, and reinforcing factors, within a community and policy context. Real situations rarely line up so neatly, but for designing an intervention and its evaluation the chain is useful, because the intervention goes back up it: you target several factors, which should affect risks, which should affect the health problem. The diagram shows the PROCEED side of that logic.

Intervention Educational / ecological change predisposing, enabling, reinforcing Behavioural / environmental change risk factors Change in epidemiology Change in social conditions Administrative / policy change and resources people, funds, community politics, structures that help or hinder implementation Process evaluation covers the intervention box; outcome evaluation the educational / ecological box; impact evaluation the rest.

The parts of a logic model

A logic model is a diagram or structure that links what you plan to do with its expected outcomes and impacts. It has five linked components. The health problem comes from the epidemiological assessment (HIV, cancer, malaria, diabetes, and who it affects) and may name the contributing factors you intend to address. Outputs are the activities you plan, such as materials, events, screenings, or trainings, and the factors each is meant to change; the choice is guided by the factors identified and the theory you believe appropriate. Inputs are the resources, staff, components, and funds invested, some identified in the administrative and policy assessment. Projected outcomes and impacts are the short-term effects hypothesized for the outputs, typically enabling and predisposing factors such as knowledge, and the longer-term effects on risk behaviour and health status. Indicators and measures are the criteria and tools for measuring them: health data, survey or qualitative data, measures of skill.

Take a worked example: Harfield County, a fictional rural county where type 2 diabetes is rising among adults and young adults. Large supermarkets are scarce, the newspaper rarely covers health, and the county's strength is its churches, service organizations, and a veterans group, full of people who are leaders or at least "movers." From that assessment the planners choose a social-cognitive approach with social network components, and the logic model follows.

InputsProblem / goalOutputsShort-term outcomesLong-term impacts
Church and veterans organization staffType 2 diabetes caused by dietCommunity outreach; diet classes and cooking groups through churches and veterans centres; diet modelling by community leadersKnowledge and awareness of food risk for diabetes; knowledge of improved diet practices; skills to choose healthier dietsReduced incidence of diabetes and its consequences

Read the table against the theory. Modelling by leaders is Social Cognitive Theory's observational learning; delivery through churches and veterans centres is a social network strategy; knowledge and skills are the outcomes the theory predicts; diabetes incidence is the impact. Each column is also a type of evaluation: process for outputs, outcome for the short-term column, impact for the long-term column. Now build one.

Interactive: build the logic model. Click a component, then click its column. When all fifteen are placed, the model reveals its indicators and the type of evaluation that tests each column.
The case. A Canadian school district finds rising anxiety and depression symptoms among grade 9 students and very low help-seeking. Guided by Social Cognitive Theory and the school as a setting, it adopts a mental health literacy curriculum delivered by trained teachers, in partnership with the local Canadian Mental Health Association branch.
Inputs
resources and staff
Problem / goal
the health problem
Outputs
activities
Short-term outcomes
of the activities
Long-term impacts
health status
Select a component to begin.
0 of 15 placed
Indicators and measures, and the evaluation that tests each column
ColumnIndicator or measureEvaluation
OutputsSessions delivered per class; teachers trained; parents attending; referral pathway in useProcess (fidelity, dosage)
Short-term outcomesKnowledge, stigma, and help-seeking intention scales at baseline and after the curriculum; counsellor referralsOutcome
Long-term impactsEmergency department records for self-harm; district survey data on untreated depression; graduation records, over yearsImpact

Evaluation methods: making a case

This lesson does not present methods in detail, but their shared rationale is worth stating. Evaluation methods are tools for making a case for the effectiveness, or lack of effectiveness, of an intervention; it is a little like being a lawyer. The more rigorous the method, the more you can argue that the intervention, and nothing else, was responsible for the change. As designs get more rigorous they add control groups (people who do not get the intervention) and, at the top, random assignment. The accordion runs from simple to rigorous.

Historical recordkeeping approach (process)▼

What you do: track project activities and client participation. Data: sessions conducted, services delivered, calls taken, advertisements posted, client utilization. Question: are program inputs creating the planned outputs? This is the floor beneath every other design.

Qualitative evaluation approaches (any type)▼

What you do: assess what participants and staff have experienced in terms of change. Data: interviews, focus groups, and observations using semistructured guides. Questions: what is the nature of the change that has or has not occurred, from the viewpoint of implementers and participants? Was it the intended change? What factors were involved? Especially useful when randomized designs are not possible and risk behaviour must be understood in context.

Periodic inventory / time series (process and outcome)▼

What you do: at set intervals, beginning with a baseline, tabulate program data and complete surveys to check progress toward goals, such as the number who quit smoking or changed their attitude toward quitting. No comparison group. Questions: are short-term goals being attained? Is progress being made?

Benchmarking (process and outcome)▼

What you do: assess program and participant data against a comparable benchmark, such as national or population-specific data. A Canadian program might compare its quit rate with Canadian Tobacco and Nicotine Survey figures, or its screening uptake with the provincial average. Question: are program results comparable to results documented elsewhere?

Quasi-experimental design (outcome and impact)▼

What you do: identify a similar community or population sample that does not receive the intervention, and collect pre- and posttest or time series data from both. Problems: are the two groups really comparable? Did the comparison group get any of the intervention (contamination)? Did anything else happen in the intervention community? Question: are changes the result of the intervention, or would they have occurred anyway?

Classic experimental design (outcome and impact)▼

What you do: randomly assign individuals to the intervention or to serve as controls, for example every other name on a list, then collect pre- and posttest or time series data from both. Problems: did controls receive some of the intervention by accident? Is it ethical to provide an intervention only to some people in a community that needs it? A program that includes a full-scale experimental design can be viewed as an evaluation project.

Empowerment evaluation (a different purpose)▼

Drawing on the community theories of Paulo Freire and other participatory approaches, empowerment evaluation is not an experimental design meant to produce rigorous data. It is a collaborative effort with the community to identify goals and decide how to measure them, often instead of standard validated instruments, so that the community improves its program on its own terms rather than by outside models. In Canada it sits naturally with the First Nations principles of ownership, control, access, and possession (OCAP) for community information, and with Indigenous-led evaluation.

Think about it: rigour and its assumptions

These designs connect to the history of positivist social science from earlier in the course. What Enlightenment assumptions underlie the methods now regarded as rigorous? When does a quasi-experimental or classic experimental design truly isolate the effect of the program? How else could you know whether a program made an impact? Qualitative and empowerment approaches trade the ability to isolate a cause for a richer account of what changed and why.

Independent and dependent variables

An independent variable is a characteristic of a person or situation that the intervention is not trying to change. It just is, yet it tells you a great deal about who and what your program can change. A dependent variable is what the intervention is trying to change, and what outcome or impact evaluation measures. An example: an intervention to change diet and exercise among women who are recent immigrants. Independent: being a woman, being a recent immigrant. Dependent: diet and exercise behaviour.

Confounds: what weakens the case

When you argue that a program caused a change, you must also address what could weaken the case. Some weaknesses are methodological. Others are confounds: unplanned occurrences or other issues that complicate the claim that only the intervention caused the change. Seven are listed here; designs often build in ways to offset them.

ConfoundWhat it isA Canadian illustration
HistoryAn unconnected event during the program causes change; the standard example is a famous film star dying of AIDS.A federal tobacco tax increase arrives midway through a cessation program.
MaturationParticipants change simply because they gain life experience during the program.Grade 9 students in a two-year program become grade 11 students; some change in risk-taking is growing up.
TestingChange from baseline to follow-up reflects familiarity with the survey or its questions.Parents answer a safe-sleep questionnaire at three home visits and learn the "right" answers.
Regression to the meanExtreme starting values move toward the average regardless of the program; with extreme poverty or very high-risk behaviour, scores may improve only because they cannot get worse.A campus program recruits the heaviest drinkers; their drinking falls whether or not the program worked.
Selection biasParticipants are not representative, for example because answering is voluntary, so responses come only from the motivated.An online follow-up survey is completed mainly by participants who liked the program.
Mortality or attritionParticipants leave over time, so follow-up data come only from those who stayed.Half of a cessation cohort is lost to follow-up; those who relapsed are least likely to answer.
Diffusion of treatmentThe control group receives some or all of the intervention, for example when radio advertisements reach the control community, so the comparison shows little difference.A campaign in one Lower Mainland municipality is seen by residents of the neighbouring comparison municipality who commute through it.

What kinds of outcome or impact?

Different evaluations give different results, and results short of a classic experiment are not without merit. Programs run in the real world, and success has many levels. When getting a program in place and operating is itself an achievement, a process evaluation may be all you need. When the time period is short, or barriers are many and only a limited effect can be expected, an outcome evaluation may be the best you can do. When longer-term impacts are realistic, because the program runs for ten years or includes a long-term follow-up, look for them. A clinical intervention calls for clinical outcomes.

The kind of data matters too. Quantitative data can be expressed numerically and analyzed statistically: fixed-choice survey answers, official demographic or epidemiological records. Collected the same way from many people, they support comparison and generalization. Qualitative data are narrative, descriptive, and subjective: extended interviews, focus groups, and observation, where people have time to explain their actions, beliefs, and interpretations. Gathered from fewer people, they reveal meaning, context, and the dynamics of interactions such as those between patient and doctor. The conclusion: there is no cookie-cutter standard. Evaluation is tailored to the situation, the program, and the evidence needed.

Case study: Designing the evaluation of a social prescribing pilot

A network of community health centres in Ontario pilots social prescribing for one year. Primary care providers refer patients who are isolated, lonely, or coping with chronic conditions to a link worker, who connects them with walking groups, arts activities, and volunteering. The logic model runs from referrals (outputs) to reduced loneliness and greater confidence in managing health (short-term outcomes) to fewer primary care and emergency visits (long-term impacts). There is no money for a comparison group, the centres serve very different neighbourhoods, and patients join whenever their provider refers them.

Which design from the ladder above is realistic here, and what could it claim? Name two confounds that threaten the claim that social prescribing caused any improvement, and say what qualitative data would add that surveys cannot.

A realistic answer is a periodic inventory with baseline and follow-up surveys, benchmarked where possible against provincial data, plus interviews with patients and link workers. It could claim that participants improved, not that the program alone caused it. Regression to the mean threatens the claim, since providers refer the most isolated patients; so does attrition, since those who did not benefit stop attending. Interviews would reveal what "connection" meant to participants and why some referrals never became attendance. The final section asks where, in a career, all of this gets used.

Reflection

A First Nation in the BC Interior, working with the First Nations Health Authority, launches a two-year program in which community Elders and youth grow, harvest, and cook traditional foods together, aiming at both diabetes risk and cultural connection. Sketch a logic model for the program in five columns (inputs, problem, outputs, short-term outcomes, long-term impacts). Then choose an evaluation approach from the ladder of designs, explain why a classic experiment would be inappropriate here, and name one confound you would still need to address. Finally, say what an empowerment evaluation approach would change about who defines success.

Model answerA strong answer produces a plausible model: inputs (FNHA and community funding, Elders' knowledge, garden land, a coordinator); problem (high diabetes risk linked to diet, and weakened connection to traditional food practices); outputs (weekly garden and harvest sessions, cooking gatherings, youth-Elder pairings); short-term outcomes (knowledge of traditional foods and their preparation, skills, confidence, stronger youth-Elder relationships, more traditional food eaten); long-term impacts (lower diabetes incidence, sustained cultural practice). For design, a periodic inventory with baseline and follow-up measures, combined with qualitative interviews and perhaps benchmarking against regional data, fits a two-year community program. A classic experiment would require randomly withholding the program from some community members, which is ethically problematic in a community that needs the intervention, and which would be at odds with a program whose value is partly collective. Diffusion of treatment would also be unavoidable in a small community. Confounds still matter: maturation among youth and history (a new grocery store, a regional health campaign) could explain change, and attrition could bias follow-up. An empowerment evaluation approach would make the community the one that identifies goals and decides how to measure them, which here might mean indicators of cultural connection and food sovereignty that no standard validated instrument captures, consistent with OCAP principles on ownership of community information.

Minimum 20 characters required.

✓ Reflection saved

Key Takeaways

  • The PRECEDE assessments form a logical chain: health problems result from risks, which result from predisposing, enabling, and reinforcing factors within a community and policy context. An intervention and its evaluation travel back up that chain.
  • A logic model links the health problem, inputs, outputs, projected short-term outcomes and long-term impacts, and the indicators and measures for each. Process evaluation tests the outputs, outcome evaluation the short-term column, impact evaluation the long-term column.
  • Evaluation methods are tools for making a case, like a lawyer's. Designs climb from historical record keeping, qualitative approaches, periodic inventories, and benchmarking to quasi-experimental and classic experimental designs; each rung adds comparison or randomization and brings its own problems of comparability, contamination, and ethics.
  • Confounds (history, maturation, testing, regression to the mean, selection bias, mortality or attrition, and diffusion of treatment) are alternative explanations a design must anticipate. Independent variables just are; dependent variables are what the program tries to change and what outcome or impact evaluation measures.
  • There is no cookie-cutter standard. Success has many levels, quantitative and qualitative data answer different questions, and empowerment evaluation lets a community define and measure success on its own terms.
Knowledge Check: this section

1. In a logic model for a school mental health program, where does "six classroom sessions delivered by trained teachers" belong, and which type of evaluation tests that column?

Sessions are activities the program delivers, which a logic model calls outputs. Process evaluation asks whether the planned outputs were produced from the inputs, so it is the evaluation that tests this column. The trained teachers themselves would be an input; knowledge change would be an outcome.

2. An evaluator compares a cessation program's community with a similar neighbouring community that did not receive the program, collecting baseline and follow-up data in both, but without random assignment. Which design is this, and which problem does it raise?

A similar but non-randomized comparison community with pre- and posttest data is the quasi-experimental design. Its characteristic problems are whether the communities are truly comparable, whether the comparison group received some of the intervention (contamination), and whether something else happened in the intervention community. Random assignment would make it a classic experiment.

3. During a youth drug-use program, a widely reported overdose death of a well-known musician shifts attitudes among young people in both the intervention and comparison communities. Which confound is this?

History is an event unconnected to the program that occurs during it and produces change; the standard example is a famous film star dying of AIDS. Maturation is change from gaining life experience, testing is familiarity with the survey, and regression to the mean concerns extreme starting values.

4. An intervention aims to change diet and exercise among women who are recent immigrants. Which is a dependent variable?

A dependent variable is a characteristic the intervention is trying to change and is what an outcome or impact evaluation measures; here that is diet and exercise behaviour. Being a woman and being a recent immigrant are independent variables, characteristics the program is not trying to change but that tell you who and what the program can change.

5. What is the key purpose of empowerment evaluation?

Empowerment evaluation, drawing on Paulo Freire's community theories and participatory programming, is a collaborative effort to identify goals and decide how to measure them, often instead of standard validated instruments. It is explicitly not an experimental design meant to produce rigorous comparative data; its purpose is improvement on the community's own terms.

✦ Pass the knowledge check with 100% and complete the reflection to continue

Section 4

Careers: Putting Theory to Work

⏱ Estimated reading time: 15 minutes

Section 4 of 4

Careers: Putting Theory to Work

Where social and behavioural theory and evaluation are used, and by whom.

The argument

Theory and evaluation: where the money is

Coherence

Theory connected to real circumstances is what makes a program make sense.

Evaluability

Theory says what ought to change, and therefore what to evaluate.

Together they produce the evidence that funders, registries, and communities now expect.

Six paths

Where theory is used

Government or public agencyNonprofit, community-based, schoolsPrivate consultingPrivate industryHealthcare providersAcademic settings

In Canada: PHAC and ministries, health authorities and public health units, charities and community health centres, consultancies, employers and unions, universities.

Two ends

Agency and community organization

Public agency

Adopts a framework, issues the funding announcement, convenes the review panel, oversees through a program officer.

Community organization

Knows the community, writes the application, delivers the program, evaluates or partners to evaluate.

A grant application: problem, theory-justified plan, evaluation plan, staff and resources, experience, budget and timeline.

Three more

Consulting, industry, health care

Consulting

Contracts, project teams, hands-on work with the funder; often the ones writing the evaluation plan.

Industry

Workplace wellness, screening, exercise; associations and unions on policy.

Health care

Information, interpersonal communication and role modelling, feedback from practice.

Academic, and a Canadian addition

Universities and Indigenous health organizations

Academic settings

Generate and test theory; evaluate; train; move research into practice through knowledge translation.

Indigenous health organizations

First Nations Health Authority, national and local Indigenous organizations; community-defined success, OCAP principles.

Carry forward

What to take into the final review

  • Theory lends coherence to a program and tells the evaluator what to measure.
  • Six general settings, plus Indigenous health organizations in Canada; each uses theory in design, implementation, or evaluation.
  • Agencies fund and oversee; community organizations design, deliver, and evaluate; consultants and academics work on both sides.
  • Next: the final assessment, integrating evaluation, logic models, methods, and careers.

Careers: putting theory to work

Learning objectives for this section

  • Explain why a working knowledge of social and behavioural theory is an increasingly useful qualification in public health, and why theory and evaluation together are "where the money is."
  • Describe six career settings for public health work and the role theory plays in each.
  • Map those settings onto Canadian public health, including Indigenous health organizations.
  • Outline the parts of a grant application and explain where theory and evaluation appear in it.

This section is short and practical. Its argument is that a good working knowledge of social and behavioural theory and its real-world application has become an increasingly important qualification for public health work, because so much weight now falls on programs that make sense, are effective, and are evidence-based. Theory, properly connected to real circumstances and people, is what lends coherence to an intervention. Theory is also what allows a meaningful evaluation, because the theoretical base tells you what ought to change and therefore what to evaluate. Theory plus evaluation produces the evidence that improves the knowledge base about what works. That, to put it bluntly, is where the money is.

Six career paths

Public health careers where theory matters can be grouped into six settings, each with a characteristic role. The table adds the Canadian organizations where you would find them.

SettingRole for theoryWhere in Canada
Government or public agency, domestic and globalProgram design and managementPublic Health Agency of Canada; Health Canada; Indigenous Services Canada; provincial ministries and agencies such as the BC Centre for Disease Control, Public Health Ontario, and the Institut national de santé publique du Québec; regional health authorities; Ontario's local public health units
Nonprofit and community-based organizations, and schoolsProgram design and implementation (including evaluation); advocacyCanadian Cancer Society, Heart and Stroke Foundation, Canadian Mental Health Association, community health centres, friendship centres, AIDS service organizations, school boards
Private sector consulting organizationsProgram design and evaluation; technical assistanceEvaluation and research consultancies contracted by governments, health authorities, and foundations
Private industryProgram design and implementation (including evaluation) for workplace healthEmployers with extended health benefits and wellness programs; industry associations; labour unions and their health and safety committees
Healthcare providersProgram design and implementation (including evaluation)Hospitals and health authorities, primary care networks and community health centres, pharmacies, public health nursing
Academic settingsBehavioural research, program design, evaluation, research collaboration, training, capacity buildingFaculties of health sciences and public health; research funded by the Canadian Institutes of Health Research and provincial funders such as Michael Smith Health Research BC
Interactive: career explorer. Click a setting to see what people do there, how theory enters the work, and where it happens in Canada. Eight settings: the six from the table, with advocacy organizations shown separately, plus a Canadian addition.

Select a setting on the left.

0 of 8 settings viewed
Government or public agency

What you do Disseminate and manage funds for a health problem: write funding announcements, support review panels, then oversee the selected organizations as a program officer without delivering the program yourself. Policy offices develop guidelines and regulations and build consensus.

Theory Announcements are increasingly built on a program framework or theoretical approach the agency has adopted. Policy work draws on the ecological model, communications strategies, and organizational mobilization.

In Canada Program consultants and policy analysts at the Public Health Agency of Canada and Health Canada; provincial ministry staff; health promotion leads at regional health authorities; program managers in Ontario public health units.

Nonprofit or community-based organization

What you do Work at the other end of the funding process, in direct contact with the community. Respond to solicitations, or develop a program that meets a need and seek funds for it. Deliver it, then evaluate it or partner with an organization that can. Write grant applications.

Theory The application justifies the program with theory and an analysis of the problem, a little like a PRECEDE-PROCEED analysis, and the evaluation plan follows from it.

In Canada Health promotion coordinators at community health centres, AIDS service organizations, immigrant-serving agencies, friendship centres, and national charities such as the Canadian Mental Health Association.

Advocacy organization

What you do Increase public engagement about an issue and affect policy on it: media advocacy, social marketing and other communications, position papers, help drafting legislation, and the research behind them.

Theory Communications theory and social marketing from Lesson 8, with community and organizational change theory; goals are set in terms of behaviour and policy change.

In Canada Policy and communications roles at the Canadian Cancer Society, the Heart and Stroke Foundation, Physicians for a Smoke-Free Canada, and the Canadian Public Health Association.

Private consulting

What you do Carry out work agencies contract to firms with specific expertise: evaluation, program design, organizational development, cultural competency, communications. You work in a project team, the firm competes through proposals, and contracts mean closer contact with the funder's project officer than grants do.

Theory Knowledge of theory and its relationship to design, implementation, and evaluation is very important here, because consultants often write the logic model and evaluation plan.

In Canada Evaluation and research consultancies working for governments, health authorities, and foundations; many hold the Canadian Evaluation Society's Credentialed Evaluator designation.

Private industry and workplaces

What you do Design, manage, and evaluate workplace health promotion: wellness screening and education, exercise programs, facility memberships. Industry associations advocate workplace health policy, and labour unions often run wellness programming and lobby.

Theory The worksite theories from Lesson 7 (stage-based tailoring, social support, organizational change); evaluation shows the employer or union what the program returns.

In Canada Wellness and occupational health roles at large employers and benefits providers; union health and safety committees; workplace mental health initiatives run with the Canadian Mental Health Association.

Healthcare providers

What you do Beyond care: disseminate health information, provide interpersonal communication and role modelling, and give real-world feedback on which public health approaches work. Providers are also a setting for research on health beliefs, program approaches, and barriers to treatment.

Theory Role modelling and interpersonal communication are Social Cognitive Theory in practice; a pharmacist offering nicotine replacement is a cue to action from the Health Belief Model.

In Canada Health promotion positions in hospitals and health authorities, primary care networks, community health centres, public health nursing, and community pharmacy.

Academic settings

What you do Generate theory through research, evaluate theory-driven interventions, identify useful evaluation methods, and compete for peer-reviewed grants, often with community partners. Share results in journals, briefs, policy recommendations, and with partner communities. Teach and train; consult for agencies.

Theory This is where theory is built, tested, and taught.

In Canada Faculties such as SFU Health Sciences; research funded by the Canadian Institutes of Health Research, which uses the term knowledge translation for moving findings into practice.

Indigenous health organizations (a Canadian addition)

What you do First Nations, Inuit, and Métis organizations design, deliver, and evaluate health promotion in their own communities and at regional and national levels. The First Nations Health Authority in British Columbia took over federal First Nations health programs in the province in 2013; Inuit Tapiriit Kanatami, the National Collaborating Centre for Indigenous Health, and local friendship centres also employ health promotion and evaluation staff.

Theory Community and cultural theories from Lessons 5 and 10 sit alongside Indigenous knowledge; evaluation often follows the empowerment model, with the community defining success and controlling its data under OCAP principles.

In Canada Wellness and evaluation roles with the First Nations Health Authority, Métis and Inuit organizations, tribal councils, and friendship centres.

Two ends of the funding process

The first two settings are two ends of one process. A government agency, federal, provincial or state, local, or international, disseminates and manages funds allocated to a health problem. Its grant announcements describe the purpose, the activities requested, the funds available, and the criteria for success, and are increasingly based on a program framework or theoretical approach the agency has adopted. Program heads and program officers make those decisions but do not carry out the program; the organizations selected do, overseen by a program officer. A panel of experts and agency staff rates applications on the best plan at the best price with adequate organizational capability. Agencies also do policy work, issuing guidelines and regulations (air and water pollution, secondhand smoke) and building consensus on recommended policies.

Public agency adopts a program framework issues a funding announcement review panel rates applications program officer oversees Community organization knows the community writes the application designs and delivers the program evaluates, or partners to evaluate funds, framework, success criteria reports, evaluation data, evidence Consultants and academics work on either side Private foundations act like agencies at the funding end

Community-based organizations are at the other end, in direct contact with people and with intimate knowledge of the situations surrounding health behaviour. Here theory is used in program design and evaluation, in response to a solicitation or when the organization has developed a program it believes meets a need and must seek funding. The exception is the private foundation, a nonprofit that acts like a public agency by disseminating funds; the Ford Foundation and the Kaiser Family Foundation are United States examples, and Canada's community foundations, such as the Vancouver Foundation, play a similar role. A task common to all these organizations is writing grant applications, which typically contain six parts.

1. A description of the health problem to be addressed▼

The epidemiological and social assessment in miniature: who is affected, how badly, and in what context. In logic model terms, the problem column.

2. A plan to address the problem, justified by theory and an analysis of the problem▼

This is a little like a PRECEDE-PROCEED analysis: the factors identified, the theory that explains them, and the activities chosen to change them. This is where a reviewer decides whether the program makes sense.

3. A plan to evaluate the effectiveness of the program▼

Process, outcome, and (if realistic) impact questions, the design, the indicators, and the baseline plan. Because part 2 says what ought to change, part 3 follows from it.

4. Staff and agency resources, with their qualifications and capabilities▼

The inputs column of the logic model, and the reviewer's basis for judging organizational capability.

5. The organization's experience doing this kind of work▼

Track record: previous programs, previous evaluations, and relationships with the community. The Winnipeg director in Section 1 needed stronger material here.

6. A budget and a timeline for implementation▼

What it will cost and when each output will be delivered. Reviewers rate applications partly on price, and the timeline becomes the yardstick for process evaluation.

Consulting, industry, and health care

More and more of the work of government agencies is carried out by private consulting organizations with specific expertise: evaluation, program design, organizational development, cultural competency, communications campaigns, materials, information technology. You work in a project team, and the firm competes through proposals much like grant applications. Consultants usually work under contracts rather than grants, which means more hands-on interaction with the funding agency's project officer. Knowledge of theory and its relationship to design, implementation, and evaluation is very important here.

Private businesses that provide health insurance may run workplace health promotion to manage costs: wellness screening and education, exercise programs, health facility memberships, sometimes in-house staff. Industry associations advocate workplace health policy, and labour unions often handle wellness programming and lobby as well. Healthcare providers, public or private, large or small, have three key roles beyond care: disseminating health information, providing interpersonal communication and role modelling, and giving real-world feedback on which public health approaches work. They are also a setting for research on health beliefs and barriers to treatment.

Academic settings and knowledge transfer

Universities generate theory through research, evaluate theory-driven interventions, and identify useful evaluation methods, all of which means competing for research grants through peer-reviewed applications, often with community partners. Results are shared in journals and at conferences, and also in research briefs, policy recommendations, and with the partnering communities, which is how research gets transferred to practice. Canadian funders call it knowledge translation. Academics also teach students and practitioners and advise agencies. The possibilities are many and can only increase.

Four people, four settings

Four hypothetical profiles, none of them real cases, illustrate the settings. Three are set in the United States, where this typology of settings originated; the fourth is set in a Canadian setting the typology does not cover. Click each card.

Government:
program officer
Click to learn more
Advocacy:
communications
Click to learn more
Consulting:
research director
Click to learn more
Indigenous health:
evaluation specialist
Click to learn more

Why the fourth profile

The six-setting typology was written for the United States, where private insurance drives workplace programs and federal grants shape the nonprofit sector. In Canada, ministries, health authorities, and public health units deliver much of what a United States nonprofit might do under a grant, and Indigenous self-determination in health has created organizations with no counterpart in that typology. The roles for theory and evaluation are the same. The organizations that hold them are not.

Case study: Three postings

A graduating SFU health sciences student is weighing three job postings: a health promotion coordinator at a regional health authority in the BC Interior, developing and evaluating community programs on physical activity and healthy eating; an evaluation and quality improvement analyst with an Indigenous health organization, supporting communities that design their own wellness programs; and a research coordinator on a university trial, funded by the Canadian Institutes of Health Research, testing a theory-based intervention to increase HPV vaccination among young adults.

For each posting, which of the six settings is it closest to, how would theory show up in a typical week, and which type of evaluation would the person spend most time on? Which posting would let the student test a theory, and which would let a community test its own?

Case study: A wellness program in a unionized workplace

A large Canadian employer with a unionized workforce wants to reduce sick days and improve mental health. Management proposes an app-based wellness challenge. The union's health and safety committee argues that shift scheduling and workload are the real problem and wants organizational change. A health promotion graduate has been hired to design something both sides will accept and to evaluate it.

Drawing on the account of private industry and unions above, and the worksite theories from Lesson 7, how would you use theory to reconcile the two proposals, and what would the process and outcome evaluation each need to show for the employer and for the union?

One closing observation applies to all four profiles and both cases. Every setting rewards the same combination: theory connected to real circumstances, a program built from it, and an evaluation that checks whether the predicted changes occurred. That combination is what this course has been assembling since its first lesson, and it is the qualification public health increasingly asks for.

Reflection

Choose two of the six career settings (or one of them plus an Indigenous health organization in Canada). For each, describe a single week of work in a health promotion role and point to at least two moments in that week where a specific theory from this course, and at least one moment where evaluation, would shape what you do. Then explain why the combination of theory and evaluation is "where the money is."

Model answerA strong answer is concrete about the week and names the theories. For a community-based organization, a coordinator might spend Monday drafting the theory section of a grant application (justifying a peer-education program with Social Cognitive Theory and diffusion of innovations), Wednesday running a session and logging attendance for the process evaluation, and Friday reviewing baseline survey results to see whether the self-efficacy measure is capturing what the theory predicts. For a government agency, a program consultant might spend the week writing the framework section of a funding call (an ecological approach combining education and environmental change), reviewing grantees' logic models to check that their outcomes follow from their stated theory, and meeting a policy team using communications and organizational mobilization strategies to build consensus on a new regulation. For an Indigenous health organization, an evaluation specialist might help a community define indicators for a land-based program, drawing on community and cultural theory and the empowerment evaluation model, and then translate those indicators into the process and outcome reporting a funder expects. The closing explanation should follow the argument of this section: theory connected to real circumstances gives a program coherence and tells you what ought to change; evaluation checks whether it changed; together they produce the evidence that funders, registries, and communities now require, which is why the combination is what public health employers increasingly pay for.

Minimum 20 characters required.

✓ Reflection saved

Key Takeaways

  • A working knowledge of social and behavioural theory is an increasingly important qualification in public health because theory, connected to real circumstances, gives a program coherence and tells the evaluator what ought to change. Theory plus evaluation produces the evidence funders now require: put bluntly, that is where the money is.
  • The six career settings are government or public agencies (program design and management), nonprofit and community-based organizations and schools (design, implementation, evaluation, advocacy), private consulting (design, evaluation, technical assistance), private industry (workplace programs), healthcare providers (information, interpersonal communication and role modelling, feedback), and academic settings (research, evaluation, training, capacity building).
  • Agencies adopt a framework, issue funding announcements, convene review panels, and oversee grantees through program officers; community organizations write the applications, deliver the programs, and evaluate them. A grant application describes the problem, a theory-justified plan, an evaluation plan, staff and resources, experience, and a budget and timeline.
  • In Canada these settings correspond to the Public Health Agency of Canada and provincial ministries and agencies, regional health authorities and local public health units, charities and community health centres, evaluation consultancies, employers and unions, health authorities and primary care, and universities funded by the Canadian Institutes of Health Research, with Indigenous health organizations such as the First Nations Health Authority as a setting the six-part typology does not name.
  • Research is transferred to practice through publication, research briefs, policy recommendations, and sharing with partner communities, which Canadian funders call knowledge translation.
Knowledge Check: this section

1. A program officer at a federal agency manages a set of grants for youth obesity prevention. Which of the following does she NOT do?

Program heads and officers make decisions about the framework and oversee the organizations selected, but do not actually carry out the program or the evaluation. Delivery is the work of the community-based or other organizations that won the grant.

2. Which career setting usually works on a contract basis, with more hands-on interaction with the funding agency's project officer than a grant involves?

Private consulting organizations, which increasingly carry out work that agencies contract to firms with specific expertise, typically operate under contracts rather than grants, and contracts involve closer day-to-day contact with the project officer. Community organizations and academics usually work under grants.

3. A nonprofit spends most of its effort on media advocacy, social marketing, drafting position papers, and helping to draft legislation on sugary drink taxation. What kind of organization is it, and which theories from this course would its staff draw on most?

Advocacy organizations are nonprofits focused on increasing public engagement about an issue and affecting public policy, through media advocacy, social marketing, other communications efforts, and direct policy work. Those activities draw on the communications and community and organizational change theories covered earlier in the course.

4. Sharing research results through research briefs, policy recommendations, and with partnering communities, and not only in journals, is important for which reason?

Results shared in briefs, policy recommendations, and with partner communities are how research is transferred to practice; Canadian funders call this knowledge translation. Peer review still governs research grants and journal publication, and evaluation remains necessary.

5. Beyond providing care, which set of roles do healthcare providers play in health promotion?

Healthcare providers have three key roles beyond care: disseminating health information, interpersonal communication and role modelling, and feedback on public health approaches that work and do not. Regulations and grant programs belong to public agencies, and lobbying on workplace policy to industry associations and unions.

✦ Pass the knowledge check with 100% and complete the reflection to continue

Section 5

Final Review & Assessment

⏱ Estimated time: 25 minutes

Bringing It All Together

This lesson closed the loop the course has been drawing since its first pages. Evaluation, in plain terms, asks whether you did what you proposed, whether the program had an effect on what it set out to achieve, and whether the program model and theory were useful, with cost and community goals sometimes added. The funding environment now demands it: the accountability movement that followed the Government Performance and Results Act and the Program Assessment Rating Tool, and its Canadian counterparts in Treasury Board results policies and provincial standards, made stated goals the measure of success, and the move to evidence-based practice made evaluation data the evidence that registries compile and funders require. The four reasons to evaluate, accountability, learning and improvement, theory, and efficiency, each shape what gets collected. And theory is what makes an evaluation specific, because a theory-based program predicts particular changes that an evaluation can check.

The three basic types of evaluation ask three questions. Process evaluation asks whether the components were implemented as planned, and supplies fidelity and dosage data that separate a failed theory from a failed delivery. Outcome evaluation asks what short-term effect the program had, measured from baseline to follow-up on the factors theory targeted. Impact evaluation asks whether the health problem itself changed, which needs an extended program or a followed cohort. The logic model gives each its column: inputs, the problem, outputs, short-term outcomes, long-term impacts, and the indicators for each, all descending from the logical chain of PRECEDE assessments. Making a case that the program caused the change is like being a lawyer: designs climb from record keeping through qualitative approaches, periodic inventories, and benchmarking to quasi-experimental and classic experimental designs, and each must anticipate confounds such as history, maturation, testing, regression to the mean, selection bias, attrition, and diffusion of treatment. There is no cookie-cutter standard; success has many levels, quantitative and qualitative data answer different questions, and empowerment evaluation lets a community define success on its own terms.

The final section asked where all of this is practised. The answer was six settings, government and public agencies, nonprofit and community-based organizations, private consulting, private industry, healthcare providers, and academic settings, each using theory in program design, implementation, or evaluation, and each mapped in this lesson onto Canadian institutions from the Public Health Agency of Canada to regional health authorities, community health centres, unions, universities, and Indigenous health organizations such as the First Nations Health Authority. The closing argument of that section is the course's argument: theory connected to real circumstances gives a program coherence and tells the evaluator what to measure, and theory plus evaluation produces the evidence public health now runs on.

Key Takeaways from this lesson

  • Evaluation asks whether you did what you proposed, whether the program had an effect, and whether the theory was useful; the accountability movement and evidence-based practice made it a condition of funding in the United States and in Canada.
  • The four reasons to evaluate are accountability, learning and improvement, theory, and efficiency; a theory-based program predicts specific changes, so theory tells the evaluator what to measure.
  • Process evaluation checks implementation (fidelity, dosage); outcome evaluation measures short-term change from baseline to follow-up; impact evaluation measures the health problem itself over years. The convention this course uses for the outcome and impact labels is reversed relative to Green and Kreuter's PRECEDE-PROCEED, so check definitions.
  • A logic model links inputs, the problem, outputs, short-term outcomes, long-term impacts, and indicators, and gives each type of evaluation a column; it descends from the logical chain of PRECEDE assessments.
  • Evaluation designs climb from record keeping to the classic experiment, buying stronger causal claims at a cost in feasibility and ethics; confounds are the alternative explanations a design must rule out, and there is no cookie-cutter standard.
  • Theory and evaluation are used across six general career settings, and in Canada across federal and provincial agencies, health authorities and public health units, nonprofits, consultancies, workplaces, health care, universities, and Indigenous health organizations.

Reflection

A provincial health authority in Canada has funded a community organization for three years to deliver a program that aims to increase physical activity among newcomer families in a mid-sized city, using Social Cognitive Theory (family role models, group walks, skill building) and social network strategies (recruitment through settlement agencies and faith communities). You are the evaluator the organization has hired. Integrating this lesson: (1) name the two reasons to evaluate that matter most to the health authority and the two that matter most to the organization; (2) write one process, one outcome, and one impact question, and say which can be answered within three years; (3) sketch the logic model in five columns with one indicator per column; (4) choose an evaluation design from the ladder of designs, justify it, and name two confounds you must address; and (5) describe the roles the health authority's program officer, the organization, you as evaluator, and a university partner would each play in producing evidence that could enter a registry.

Model answerA strong answer moves through all five parts. (1) The health authority's reasons are accountability (showing that public funds produced a change against the stated goal) and efficiency (what each result cost); the organization's are learning and improvement (adjusting recruitment and sessions as it goes) and theory (did role modelling and group support raise self-efficacy and activity as Social Cognitive Theory predicts?). (2) Process: were the planned group walks and skill sessions delivered in each community, with what attendance (answerable now)? Outcome: did families' self-efficacy for activity and their weekly minutes of activity change from baseline to follow-up (answerable within three years)? Impact: did rates of physical inactivity, and eventually cardiovascular and metabolic disease, fall among newcomer families in the city (not answerable within three years; needs long follow-up or a cohort). (3) Inputs: funding, trained peer leaders, settlement agency partners (indicator: leaders trained). Problem: low physical activity among newcomer families (indicator: baseline activity). Outputs: weekly walks, skill sessions, family role-model events (indicator: sessions delivered, dosage per family). Short-term outcomes: self-efficacy, knowledge of local facilities, activity minutes (indicator: pre and post survey). Long-term impacts: sustained activity, lower chronic disease risk (indicator: health status data over years). (4) A periodic inventory with baseline and repeated follow-ups, benchmarked against provincial activity data, is realistic; a quasi-experimental comparison with a similar city would strengthen the causal claim if a partner could be found; a classic experiment would mean withholding the program from families who need it. Confounds: selection bias (families who join are already motivated), attrition (those who stop attending do not answer follow-ups), history (a new municipal recreation subsidy), and maturation. (5) The program officer sets the framework and success criteria, monitors, and supports the evaluation; the organization delivers with fidelity and keeps the process records; the evaluator designs the measures and analysis and reports in the funder's language; a university partner could add rigour, run the analysis, and publish the result, which is how the program's evidence reaches a registry such as Health Evidence and becomes usable elsewhere.

Minimum 20 characters required.

✓ Reflection saved

Final Knowledge Assessment

This assessment covers all material from this lesson. You must score 100% to complete the lesson. Review the feedback for any incorrect answers and try again.

Final Assessment: Evaluation and Careers in Health Promotion (15 Questions)

1. A community organization keeps attendance sheets and teacher thank-you notes but has no baseline or follow-up data. Which of the basic evaluation questions can it answer, and which can it not?

Attendance records show that the planned activities happened, the first of the evaluation questions. Answering whether the program had an effect, or whether the theory was useful, requires measuring what the theory said would change before and after the program, which needs baseline and follow-up data.

2. Which development can be described as an even more profound change than the accountability movement?

The move toward evidence-based practice, modelled on evidence-based standards of care in medicine, is arguably an even more profound change than accountability requirements. It makes evaluation data the evidence, leads to rated registries of interventions, and drives funders to require model programs.

3. A health unit collects session feedback each week and changes the next session in response. Which reason to evaluate is this?

Learning and improvement is the use of evaluation data while a program runs to see what is working and what is not, feedback that lets you make changes as you go. Accountability is about showing funders an effect; theory is about testing a predicted linkage; efficiency is about cost.

4. A funder requires grantees to use a model program or justify why none fits. What does "model program" mean?

Model programs are those with good evidence of effectiveness, and registries rate that evidence more highly when a program has been implemented in several places with different populations and found effective in each. Neither the funder's identity, the theory used, nor the program's age defines the term.

5. A replication of a mental health curriculum shows no effect, but process data show teachers delivered only two of six sessions. What conclusion does the idea of fidelity support?

For a replication to count as an additional test of a model, it must be implemented with fidelity, with all components delivered according to the design. Process data showing a low dose and missing components mean the disappointing result reflects implementation, and tells you little about the curriculum and nothing about the theory.

6. Which of the following is an example of the kind of data an outcome evaluation collects?

Outcome evaluation looks at short-term effects, and typical examples include pre- and posttest knowledge change, changes in health service utilization, and policy changes. Cancer morbidity and mortality are impact data; counts of coalition partners are process data.

7. Why can long-term impact be measured only if the project is in place for more than three or four years or you can follow a cohort of participants?

Impact evaluation asks whether the program affected the health problem that was the ultimate target, and changes in disease morbidity and mortality emerge slowly. Measuring them requires collecting follow-up data for years, either within a long-running program or by following participants as a cohort.

8. In a logic model, which component is described as the criteria and tools for measuring outcome and impact, such as health data, survey or qualitative data, and measures of skill?

Indicators and measures are the last of the five linked components and they turn projected outcomes and impacts into things that can actually be observed. Inputs are resources, outputs are activities, and the health problem comes from the epidemiological assessment.

9. In the Harfield County example, why did the planners choose a social-cognitive approach with social network components?

The assessment found limited supermarket access, a newspaper that did not cover health, and a prevalence of social organizations with established leaders. The theory choice followed the assessment: leaders could model diet change (Social Cognitive Theory) and the organizations offered channels for classes and cooking groups (social network strategy).

10. An evaluator randomly assigns every other person on a waiting list to receive a program now, with the rest as controls, and collects pre- and posttest data from both groups. Which design is this, and which problem does it raise?

Random assignment of individuals to intervention or control, with pre- and posttest data from both, is the classic experimental design; selecting every other name is a standard way of doing it. Its characteristic problems are accidental exposure of controls and the ethics of providing an intervention only to some people in a community that needs it.

11. A campus program recruits the heaviest drinkers. Their drinking falls at follow-up. Which confound most directly threatens the claim that the program caused the drop?

Regression to the mean is the tendency of participants whose behaviour is extreme at the start to change simply because it was so extreme; with very high-risk behaviour, scores may improve only because they cannot get worse. Diffusion, testing, and history are different threats.

12. Which statement best describes the value of results that fall short of a classic experimental design?

Just because results are not rigorous evidence from a classic experiment does not mean they are without merit. Sometimes getting a program operating is the achievement and process evaluation is enough; sometimes outcome evaluation is the best that can be done; there is no cookie-cutter standard.

13. What does theory contribute to a program that makes it evaluable?

Using theory to design and implement a program allows a meaningful evaluation because the theoretical base tells you what ought to change and, therefore, what to evaluate. Theory also lends coherence to the intervention, but it does not replace design choices or the budget.

14. Which parts does a grant application typically contain?

A grant application has six parts: the health problem, a plan justified by theory and analysis (a little like PRECEDE-PROCEED), an evaluation plan, staff and resources with qualifications, the organization's experience, and a budget and timeline. Theory appears in the plan and evaluation appears as its own section.

15. A graduate takes a position with the First Nations Health Authority helping communities define their own indicators of program success and control their own data. Which evaluation approach does this most resemble, and how does it differ from a classic experiment?

Empowerment evaluation, drawing on Paulo Freire and participatory programming, is a collaborative effort with a community to identify goals and the best ways to measure them, often instead of standard validated instruments, so that the community improves its program on its own terms. It stands in explicit contrast to experimental designs intended to produce rigorous comparative data.

✦ Complete the final reflection above before submitting