Definition

An education research concept defining methods used to evaluate interventions, programs, and policy impacts. It governs study design, measurement, and interpretation practices used to estimate effects and assess implementation quality. It does not establish causation without appropriate design choices and careful handling of bias and uncertainty. It supports improvement by identifying what works and under what implementation conditions. The concept is generally stable, though methods and reporting standards evolve over time.

Principle

Principle
Clarify why the evaluation is being done and how its methods and data will produce credible, usable answers to stakeholder questions so findings are valid for decision-making and improvement.

Demonstration

Demonstration
A plan for a literacy intervention that includes: purpose (formative improvement and summative reporting), evaluation questions about fidelity and student progress, a logic model linking activities to short- and medium-term outcomes, mixed-methods design (pre/post assessments, classroom observations, parent surveys), sampling frame, analysis approach, schedule, responsible staff, and estimated costs.

Misapplication

Misapplication
Writing vague goals with no clear evaluation questions, selecting methods that cannot answer those questions (e.g., only surveys for causal claims), or omitting timelines and budget so the evaluation cannot be delivered.

Consequence

Consequence
A clear plan produces focused data collection and analysis, produces trustworthy findings to inform program refinement or scale decisions, helps allocate resources, and clarifies stakeholder expectations.

Reversal

Reversal
No plan or an ad hoc approach results in irrelevant or uninterpretable data, wasted resources, limited learning, and poor decisions about continuation or scaling.

Boundary

Boundary
Covers program-level evaluation planning for implementation and outcomes. It does not replace detailed institutional review board protocols or the operational monitoring dashboard used for day-to-day management, though it should reference those tools as appropriate.

Semantic Tension

Semantic Tension
Tension exists between evaluations designed for accountability (summative) and those designed for improvement (formative); the plan must make the intended use explicit because the design choices differ.

Synthesis

Synthesis
A Program Evaluation Plan links purpose and questions to a coherent design—methods, measures, timeline, roles and resources—so stakeholders can generate valid, usable evidence about a program's functioning and effects.