09 · Norms, standards and quality
The Program Evaluation Standards and the AEA Guiding Principles
North American evaluation rests on two distinct instruments that are routinely conflated. The JCSEE Program Evaluation Standards (3rd edition, 2010) judge evaluations: thirty standards in five attribute groups — utility, feasibility, propriety, accuracy and evaluation accountability. The AEA Guiding Principles for Evaluators, first adopted in 1994 and most recently updated in 2025, guide evaluators: systematic inquiry, competence, integrity, respect for people, and common good and equity. One assesses the work; the other governs the worker — and rigorous practice uses both.
Last updated · Reviewed against 3 cited sources
Two instruments, two jobs
The most useful thing to know about North America’s evaluation standards is that there are two of them and they do different work. The Program Evaluation Standards, maintained by the Joint Committee on Standards for Educational Evaluation (JCSEE), are criteria for judging evaluations: is this study useful, feasible, proper, accurate and accountable? The Guiding Principles for Evaluators, maintained by the American Evaluation Association (AEA), are professional ethics for evaluators: how should the person doing the work behave, whatever the study looks like? [1] [2]
The distinction matters in practice. A technically excellent evaluation can be conducted by someone behaving badly — misrepresenting their competence, concealing conflicts of interest — and an evaluator of impeccable ethics can produce a study that fails every utility standard. Quality frameworks that cite “the standards” without saying which instrument they mean tend to inherit this confusion. When you write a terms-of-reference quality clause or a meta-evaluation protocol, name the instrument and the layer: product standards from JCSEE, conduct principles from AEA.
A note on names: this site uses British spelling throughout, but the instruments keep their published titles — the Program Evaluation Standards and the AEA’s principles as the association itself words them.
The JCSEE standards: five attributes, thirty standards
The Joint Committee — a coalition of North American professional associations concerned with evaluation — first published program evaluation standards in 1981 and revised them in 1994 and 2010. The third edition (Yarbrough, Shulha, Hopson and Caruthers, 2010) is the current text, and it organises thirty individual standards into five attribute groups [1]:
| Attribute group | Standards | Stated purpose |
|---|---|---|
| Utility | U1–U8 | Increase the extent to which program stakeholders find evaluation processes and products valuable in meeting their needs |
| Feasibility | F1–F4 | Increase evaluation effectiveness and efficiency |
| Propriety | P1–P7 | Support what is proper, fair, legal, right and just in evaluations |
| Accuracy | A1–A8 | Increase the dependability and truthfulness of evaluation representations, propositions, and findings |
| Evaluation accountability | E1–E3 | Encourage adequate documentation of evaluations and a metaevaluative perspective |
Three features distinguish the instrument from a generic quality checklist.
Utility leads. The ordering is deliberate: an evaluation nobody can use fails first, before its methods are even examined. This is the same judgement that animates utilization-focused evaluation, and it is why the JCSEE standards and Patton’s approach cite each other’s territory so comfortably.
Propriety is a full group, not a footnote. Seven standards address fairness, legality, human rights and conflicts of interest — the recognition that evaluations exercise power over the people and programmes they examine.
Accountability closes the loop. The evaluation accountability group (E1–E3), added as a named group in the third edition, expects evaluations to be documented well enough to be assessed themselves, and expects that assessment — meta-evaluation — to actually happen [1]. An evaluation that cannot survive scrutiny against the other twenty-seven standards was, on this view, never finished.
The AEA Guiding Principles: ethics for the evaluator
The AEA’s Guiding Principles were first adopted in 1994 and have been periodically revised since; the current text was last updated in 2025, with input from the association’s Foundational Documents Task Force, leadership and members [2]. Five principles:
- Systematic inquiry — evaluators conduct data-based inquiry that is methodical and defensible, and communicate its methods, strengths and limitations honestly.
- Competence — evaluators provide skilled professional services, work within the limits of their capability, and maintain and grow it.
- Integrity — honesty and transparency throughout: about design, findings, funding, conflicts of interest and changes to agreed plans.
- Respect for people — evaluators honour the dignity, well-being and self-worth of everyone they interact with, participants and clients alike.
- Common good and equity — evaluators contribute to the advancement of an equitable and just society, attending to the public interest beyond the client’s.
The principles are aspirational and self-enforced: the AEA is a membership association, not a licensing body, and no one loses a credential for violating them. Their force is reputational and contractual — which is precisely why commissioners increasingly write them into contracts, converting professional aspiration into an enforceable term.
Using the instruments: meta-evaluation and trade-offs
The working combination looks like this. At commissioning, the terms of reference cite the JCSEE attribute groups as acceptance criteria and the AEA principles as conduct requirements. During the evaluation, disputes about scope or method are argued in the standards’ vocabulary rather than as clashes of preference. At review, the report is assessed against the standards — the Western Michigan University Evaluation Center maintains freely available checklists, including meta-evaluation checklists, that turn the thirty standards into a workable review instrument rather than a reading exercise [3]. Our page on evaluation quality assessment treats that review practice, and its UN-system counterparts, in depth.
Honest use also means admitting that the standards conflict — the third edition treats this as a feature of real evaluations, not a defect of the instrument. The recurring tensions:
- Utility vs. accuracy. The decision-maker needs an answer in March; defensible measurement needs data through June. Neither standard automatically wins; the trade-off must be made explicitly and recorded.
- Propriety vs. feasibility. Full consent and confidentiality protections cost time and coverage. Cutting the protections to hit the budget is a propriety failure; pretending the budget can absorb them is a feasibility failure.
- Accuracy vs. utility, again, at reporting. Every caveat added for accuracy makes the report harder to use; every simplification for use shaves accuracy. The standards demand the tension be adjudicated on the page, not resolved silently in favour of whichever pressure is closest.
A quality clause that says “the evaluation shall meet all applicable standards” has missed the point. A better one says which attribute groups take precedence for this study, and requires the evaluator to document departures and trade-offs as they occur.
Reach and limits outside North America
Both instruments were written by North American bodies for institutional contexts — school systems, federal programmes, a professionalised consulting market — that do not exist everywhere. What transfers well: the five-attribute vocabulary (utility through accountability names real properties of any evaluation, anywhere), the meta-evaluation discipline, and the insistence on documented trade-offs. What transfers poorly: the assumption of a litigation-conscious institutional environment behind the propriety standards, and a professional-association enforcement model that presumes a dense market of credentialled evaluators.
Outside North America the instruments are therefore best used in combination rather than adopted wholesale: the UNEG Norms and Standards supply the institutional and independence layer the JCSEE instrument largely assumes, and the African Evaluation Principles supply the contextual grounding — whose knowledge counts, who validates quality — that neither North American instrument addresses. Treating the JCSEE standards as a universal default is itself a standards violation of a kind: it fails the context analysis the third edition’s own utility and propriety standards require.
Sources
- The Program Evaluation Standards, 3rd edition (summary of record) — Joint Committee on Standards for Educational Evaluation (JCSEE), 2010.Yarbrough, Shulha, Hopson & Caruthers; published by Corwin. The JCSEE page lists all thirty standards and the five attribute-group purpose statements quoted on this page.
- Guiding Principles for Evaluators — American Evaluation Association (AEA), 2025.First adopted 1994 and periodically revised since; the AEA page states the principles were last updated in 2025 with input from the Foundational Documents Task Force, AEA leadership and members.
- Evaluation Checklists — The Evaluation Center, Western Michigan University, updated continuously.The standing repository of practical checklists — including meta-evaluation checklists — that turn the standards into review instruments.