Master bibliography
Every work cited anywhere in this reference — 117sources across 50pages — aggregated automatically from page frontmatter, so it can never drift out of date.
- 'Improving ratings': audit in the British university system — European Review, 5(3), 305–321, 1997.Strathern's formulation of Goodhart's law — 'When a measure becomes a target, it ceases to be a good measure' — the standing warning over every consequential metric regime.Cited by: OKRs and KPI cascades in institutions
- A Guide to the Project Management Body of Knowledge (PMBOK® Guide) — Seventh Edition and The Standard for Project Management — Project Management Institute, 2021.The profession's standard of record. Paywalled; cited here for its existence and structure — including its shift to principles and performance domains, among them measurement — not for paraphrased content.Cited by: The project management–M&E interface, Earned value management, Risk registers and risk monitoring
- A Practical Introduction to Regression Discontinuity Designs: Foundations — Cambridge Elements: Quantitative and Computational Methods for the Social Sciences (arXiv 1911.09511), 2019.Cattaneo, Idrobo & Titiunik — the modern step-by-step estimation and falsification workflow.Cited by: Regression discontinuity designs
- A Synthesis Paper on the Made in Africa Evaluation Concept — Commissioned by AfrEA, 2015.Bagele Chilisa's synthesis — the foundational analysis of what Made in Africa Evaluation means and the positions within it.Cited by: The African Evaluation Principles and Made in Africa Evaluation
- About DHIS2 — HISP Centre, University of Oslo, continuously updated.The platform's own description of record: open-source, global public good, ministry-anchored national deployments.Cited by: DHIS2 in monitoring and evaluation
- Closing the Loop: Effective Feedback in Humanitarian Contexts — Practitioner Guidance — ALNAP / CDA Collaborative Learning Projects, 2014.Bonino, Jean & Knox Clarke. The practitioner spine on what makes feedback mechanisms effective, built from field research.Cited by: Feedback and complaints mechanisms
- Collaborating, Learning and Adapting (CLA) framework and toolkit — USAID Learning Lab, n.d..The CLA framework: collaborating, learning and adapting practices plus enabling conditions. Continuously maintained resource; USAID web properties were reorganised in 2025 and the link state is re-verified at publication.Cited by: Learning agendas, adaptive management and after-action reviews
- Contribution Analysis (approach page) — BetterEvaluation (Global Evaluation Initiative), updated continuously.Practitioner overview, applications and resources.Cited by: Contribution analysis
- Contribution Analysis: An Approach to Exploring Cause and Effect — ILAC Brief 16, Institutional Learning and Change Initiative, 2008.Mayne — the contribution logic OM's causal claims rest on when its monitoring record is used as evaluative evidence.Cited by: Outcome Mapping, Contribution analysis
- Contribution analysis: Coming of age? — Evaluation, 18(3), 2012.Mayne — refines the approach, including the treatment of contributory causes and causal packages.Cited by: Contribution analysis
- Core Humanitarian Standard on Quality and Accountability, 2024 edition — CHS Alliance, Groupe URD & Sphere, 2024.The nine commitments in their revised, people-centred formulation; the verifiable standard organisations assess themselves against.Cited by: Accountability to Affected Populations (AAP), Feedback and complaints mechanisms
- Cost Estimating and Assessment Guide: Best Practices for Developing and Managing Program Costs (GAO-20-195G) — U.S. Government Accountability Office, 2020.The open reference for cost baselines and earned value as an oversight instrument, used to audit major U.S. federal programmes.Cited by: Earned value management, Risk registers and risk monitoring
- Data Quality Assurance. Module 1: Framework and metrics — World Health Organization, 2017.The DQR framework: dimensions, standard metrics, and the desk-review / verification split.Cited by: Data quality assessment (DQA)
- Data Quality Audit Tool: Guidelines for Implementation — MEASURE Evaluation (USAID), 2008.The external-audit variant: protocols for audit teams verifying reported programme results.Cited by: Data quality assessment (DQA)
- Data Quality Review (DQR): a toolkit for facility data quality assessment. Module 1: Framework and metrics — World Health Organization, 2017.The DQR dimension set — completeness, timeliness, internal consistency, external consistency — with standard metrics for each.Cited by: Data quality dimensions in M&E, Dashboards and data visualisation for M&E, M&E information systems: data flow, interoperability and architecture, DHIS2 in monitoring and evaluation
- Data Quality Tools (DQA and RDQA tool family) — MEASURE Evaluation (USAID), 2017.Index of the audit (DQA) and routine (RDQA) instruments, guidelines and workbooks.Cited by: Data quality assessment (DQA)
- Data Visualization Checklist — Stephanie Evergreen & Ann K. Emery, 2016.The evaluation field's standard chart-grading instrument: text, arrangement, colour and lines, including proportional-accuracy items.Cited by: Dashboards and data visualisation for M&E
- Designing Household Survey Questionnaires for Developing Countries: Lessons from 15 Years of the Living Standards Measurement Study — The World Bank / Oxford University Press, 2000.Grosh & Glewwe (eds.). The instrument side of the survey craft; sampling and questionnaire quality fail together.Cited by: Sampling for M&E: methods and sample size, Survey and questionnaire design
- Designing Household Survey Samples: Practical Guidelines (Studies in Methods, Series F No. 98) — United Nations Statistics Division, 2005.The standard open reference for probability designs, design effects, weighting and frame problems in developing-country surveys.Cited by: Sampling for M&E: methods and sample size, Survey and questionnaire design, Baselines, targets and disaggregation
- Determining Sample Size for Research Activities — Educational and Psychological Measurement, 30(3), 1970.Krejcie & Morgan — the source of the ubiquitous sample-size table, with the assumptions the table's users rarely restate.Cited by: Sampling for M&E: methods and sample size
- Developmental evaluation (approach page) — BetterEvaluation (Global Evaluation Initiative), n.d..Living overview of DE's niche — innovation, adaptation in dynamic environments — and the evaluator's embedded team position. Accessed 18 August 2026.Cited by: Developmental Evaluation
- Developmental Evaluation: Applying Complexity Concepts to Enhance Innovation and Use — Guilford Press, 2010.Patton's account of where UFE's use-driven logic leads when the programme itself is still developing — the sibling approach.Cited by: Utilization-Focused Evaluation, Developmental Evaluation
- DIME Wiki — survey design, piloting and questionnaire resources — World Bank Development Impact (DIME), n.d..Continuously updated practitioner reference (accessed August 2026) for questionnaire design, translation, piloting and enumerator training in impact evaluations.Cited by: Survey and questionnaire design, Indicator types and levels, Performance indicator reference sheets (PIRS), Proxy indicators and composite indices, Baselines, targets and disaggregation, M&E information systems: data flow, interoperability and architecture
- Earned Value Management (EVM) resource site — NASA Office of the Chief Financial Officer, maintained continuously.Freely available tutorials, handbooks and implementation guidance from one of the longest-running institutional EVM practices.Cited by: Earned value management
- Empowerment Evaluation — Evaluation Practice 15(1), 1994.Fetterman's founding statement, from his 1993 American Evaluation Association presidential address.Cited by: Participatory and Empowerment Evaluation
- Empowerment Evaluation (approach page) — BetterEvaluation (Global Evaluation Initiative), n.d..Living overview of the approach, its principles and step sequence as practised. Accessed 18 August 2026.Cited by: Participatory and Empowerment Evaluation, Most Significant Change (MSC)
- Empowerment Evaluation as Evaluation Ideology — American Journal of Evaluation, 2007.Nick L. Smith's critique — reading empowerment evaluation as ideology and social movement rather than a bounded evaluation model.Cited by: Participatory and Empowerment Evaluation
- Empowerment Evaluation: Yesterday, Today, and Tomorrow — American Journal of Evaluation, 2007.Fetterman & Wandersman's reply to the critics — the defence side of the debate.Cited by: Participatory and Empowerment Evaluation
- eNIMES — National Integrated Monitoring and Evaluation System — Republic of Kenya, live platform.The electronic platform through which NIMES indicator reporting and dashboards are delivered.Cited by: National M&E systems
- Ethical Guidelines for Evaluation — United Nations Evaluation Group (UNEG), 2020.The UN system's ethical obligations for evaluators — the frame for safety and dignity in gender data collection. 2020 revision.Cited by: Feminist and Gender-Responsive Evaluation, Research ethics, do no harm and informed consent in M&E, Data protection for M&E: Kenya's DPA 2019 and the GDPR, The UNEG Norms and Standards for Evaluation
- Evaluation Checklists — The Evaluation Center, Western Michigan University, n.d..The checklist repository in which the UFE checklist tradition sits, alongside meta-evaluation and design checklists. Accessed 18 August 2026.Cited by: Utilization-Focused Evaluation, The Program Evaluation Standards and the AEA Guiding Principles
- Evaluation of Humanitarian Action Guide — ALNAP/ODI, 2016.Buchanan-Smith & Cosgrave. Field-realistic guidance on interviews, group methods and triangulation under operational constraints.Cited by: Qualitative data collection: interviews, focus groups, observation, Accountability to Affected Populations (AAP), Feedback and complaints mechanisms, Writing and judging evaluation reports
- Framing Participatory Evaluation — New Directions for Evaluation, no. 80, Jossey-Bass / Wiley, 1998.Cousins & Whitmore — the canonical map of the participatory family: two streams, three process dimensions.Cited by: Participatory and Empowerment Evaluation
- Glossary of Key Terms in Evaluation and Results-Based Management for Sustainable Development (Second Edition) — OECD Publishing, 2023.The definition of record for indicator, output, outcome and impact in development evaluation.Cited by: Indicator types and levels, Indicator quality criteria: SMART, CREAM, SPICED, Proxy indicators and composite indices, Results-based management (RBM)
- Guidance for After Action Review (AAR) (WHO/WHE/CPI/2019.4) — World Health Organization, 2019.A tested AAR methodology: the review questions and four formats — debrief, working group, key informant, mixed.Cited by: Learning agendas, adaptive management and after-action reviews
- Guidance for Conducting Terminal Evaluations of UNDP-Supported, GEF-Financed Projects — UNDP Independent Evaluation Office, current edition (undated on locator).Operational guidance in which the mid-term review is part of the project's evaluative record feeding the terminal evaluation.Cited by: Mid-term reviews
- Guide: Set Goals with OKRs — Google re:Work, 2016.Google's public account of its OKR practice: ambitious objectives, measurable key results, 0–1.0 grading with ~0.6–0.7 as the sweet spot for stretch goals.Cited by: OKRs and KPI cascades in institutions
- Guidelines for Conducting Terminal Evaluations of Full-Size Projects — GEF Independent Evaluation Office, 2024.The terminal end of the sequence the MTR opens: updated guidelines effective 1 January 2024.Cited by: Mid-term reviews
- Guiding Principles for Evaluators — American Evaluation Association, 2025.The profession's principles, including its explicit commitment to the common good and equity. Most recently updated 2025.Cited by: Feminist and Gender-Responsive Evaluation, The Program Evaluation Standards and the AEA Guiding Principles
- Handbook on Constructing Composite Indicators: Methodology and User Guide — OECD Publishing & European Commission Joint Research Centre, 2008.The standard of record for composite-index construction: the step sequence, the normalisation and weighting menus, and the sensitivity-analysis obligation.Cited by: Proxy indicators and composite indices
- How Much Should We Trust Differences-in-Differences Estimates? — The Quarterly Journal of Economics, 119(1), 249–275, 2004.Bertrand, Duflo & Mullainathan on why serial correlation wrecks naive DiD standard errors.Cited by: Difference-in-differences
- How to Build M&E Systems to Support Better Government — World Bank Independent Evaluation Group, 2007.Mackay. Why utilisation is the measure of an M&E system's success, and how governments institutionalise results management.Cited by: Results-based management (RBM), National M&E systems
- IASC Revised Commitments on Accountability to Affected People and Protection from Sexual Exploitation and Abuse — Inter-Agency Standing Committee, 2017.The system-wide commitments, endorsed November 2017, revising the original 2011 commitments and integrating PSEA.Cited by: Accountability to Affected Populations (AAP), Feedback and complaints mechanisms
- Impact Evaluation in Practice, Second Edition — World Bank / Inter-American Development Bank, 2016.Gertler, Martinez, Premand, Rawlings & Vermeersch. Part on sampling and power places sample-size decisions inside evaluation design.Cited by: Sampling for M&E: methods and sample size, Mixed methods and secondary data in evaluation, Difference-in-differences
- Impact Evaluation in Practice, Second Edition — World Bank / Inter-American Development Bank, 2016.Gertler, Martinez, Premand, Rawlings & Vermeersch. Chapters 3–4 cover the counterfactual problem and randomised assignment.Cited by: Randomised controlled trials, Regression discontinuity designs, Matching and propensity score methods, Interrupted time series analysis, Synthetic control methods
- Indigenous Made in Africa Evaluation Frameworks: Addressing Epistemic Violence and Contributing to Social Transformation — American Journal of Evaluation, 2021.Chilisa & Mertens — the case for African-rooted evaluation frameworks and the critique of imported paradigms.Cited by: Feminist and Gender-Responsive Evaluation, The African Evaluation Principles and Made in Africa Evaluation
- International Ethical Guidelines for Health-Related Research Involving Humans, 4th edition — Council for International Organizations of Medical Sciences (CIOMS), in association with WHO, 2016.The reference standard for health-adjacent studies: community engagement, vulnerable participants, and research in low-resource settings.Cited by: Research ethics, do no harm and informed consent in M&E
- Interrupted time series regression for the evaluation of public health interventions: a tutorial — International Journal of Epidemiology, 46(1), 348–355, 2017.Lopez Bernal, Cummins & Gasparrini — the standard applied tutorial: impact models, segmented regression, seasonality and autocorrelation.Cited by: Interrupted time series analysis
- Introduction to Mixed Methods in Impact Evaluation (Impact Evaluation Notes No. 3) — InterAction / The Rockefeller Foundation, 2012.Bamberger. The standard practitioner treatment of mixed-methods designs, priority notation and integration stages in evaluation.Cited by: Mixed methods and secondary data in evaluation
- ISO 31000:2018 — Risk Management: Guidelines — International Organization for Standardization, 2018.The international reference standard for risk management principles and process. Paywalled; cited for its existence and high-level structure only.Cited by: Risk registers and risk monitoring
- J-PAL Research Resources — Abdul Latif Jameel Poverty Action Lab, MIT, updated continuously.Open handbooks and tools on power calculations, randomisation and measurement.Cited by: Randomised controlled trials
- Kenya National Monitoring and Evaluation Policy — Republic of Kenya, The National Treasury and Planning, 2022.The policy instrument for Kenya's whole-of-government M&E; hosted by the Twende Mbele African M&E partnership.Cited by: National M&E systems, Performance contracting in Kenya's public service
- Matching Methods for Causal Inference: A Review and a Look Forward — Statistical Science, 25(1), 2010.Stuart — the standard review of matching estimators, diagnostics and practical guidance.Cited by: Matching and propensity score methods
- Minimum Wages and Employment: A Case Study of the Fast Food Industry in New Jersey and Pennsylvania — NBER Working Paper 4509 (published in American Economic Review 84(4), 772–793), 1994.Card & Krueger — the canonical two-group, two-period DiD application.Cited by: Difference-in-differences
- Monitoring and Evaluation Directorate — Policies and Guidelines — Republic of Kenya, State Department for Economic Planning, continuously updated.The custodian directorate's repository of NIMES and CIMES guidance and the national M&E toolkit.Cited by: National M&E systems
- Most Significant Change (approach page) — BetterEvaluation (Global Evaluation Initiative), n.d..Living overview of the technique and its typical uses. Accessed 18 August 2026.Cited by: Most Significant Change (MSC)
- Norms and Standards for Evaluation — United Nations Evaluation Group (UNEG), 2016.Places ethics among the norms every credible evaluation must satisfy, alongside independence and impartiality.Cited by: Research ethics, do no harm and informed consent in M&E, Evaluation use and the management response, The UNEG Norms and Standards for Evaluation
- Office of the Data Protection Commissioner (ODPC), Kenya — official site and Act repository — Office of the Data Protection Commissioner, Kenya, 2019.The supervisory authority established under the Act: registration, guidance, complaints and enforcement. Year given is the Act's; the site is continuously updated.Cited by: Data protection for M&E: Kenya's DPA 2019 and the GDPR
- OpenHIE Framework and Architecture Specification — OpenHIE Community, living specification.The reference architecture for health-information exchange: registries as shared sources of truth, an interoperability layer between point systems.Cited by: M&E information systems: data flow, interoperability and architecture, DHIS2 in monitoring and evaluation
- Outcome Mapping (approach page) — BetterEvaluation (Global Evaluation Initiative), n.d..Living overview of the approach and its typical applications. Accessed 18 August 2026.Cited by: Outcome Mapping
- Outcome Mapping: Building Learning and Reflection into Development Programs — International Development Research Centre (IDRC), Ottawa, 2001.Earl, Carden & Smutylo — the founding manual: definitions, the three stages and twelve steps, journals and progress markers.Cited by: Outcome Mapping
- Pathways to Change: Evaluating Development Interventions with Qualitative Comparative Analysis (QCA) — Expert Group for Aid Studies (EBA), Stockholm, 2016.Befani — the standard treatment of QCA for development evaluation, including step-by-step guidance and quality-assurance checks for commissioners.Cited by: Qualitative comparative analysis (QCA)
- Performance Contracting Guidelines for FY 2025/26 (22nd Cycle) — Republic of Kenya, Executive Office of the President, 2025.The statutory cousin: Kenya's negotiated, scored public-service performance contracts — a target regime with consequences, and therefore a different instrument from OKRs.Cited by: OKRs and KPI cascades in institutions, Performance contracting in Kenya's public service
- Process Tracing (method page) — BetterEvaluation (Global Evaluation Initiative), updated continuously.Practitioner overview of process tracing in evaluation.Cited by: Process tracing for evaluation
- Process-Tracing Methods: Foundations and Guidelines, 2nd edition — University of Michigan Press, 2019.Beach & Pedersen — the book-length treatment, including the theory-testing, theory-building and explaining-outcome variants.Cited by: Process tracing for evaluation
- Public Service Commission (Performance Management) Regulations, 2021 — Legal Notice 114 of 2021 — Republic of Kenya, Kenya Law, 2021.The statutory instrument underpinning performance management in the public service, including performance contracting and staff appraisal.Cited by: Performance contracting in Kenya's public service
- Public Service Performance Management Unit (PSPMU) — Republic of Kenya, official pages.The administering unit for performance contracting within the Government's performance and delivery machinery.Cited by: Performance contracting in Kenya's public service
- Qualitative Comparative Analysis (approach page) — BetterEvaluation (Global Evaluation Initiative), updated continuously.Practitioner overview, origins with Ragin, and applications in evaluation.Cited by: Qualitative comparative analysis (QCA)
- Qualitative Research Methods: A Data Collector's Field Guide — Family Health International (FHI), 2005.Mack, Woodsong, MacQueen, Guest & Namey. Interviewing technique and sensitive-topic practice that survey pretesting borrows from.Cited by: Survey and questionnaire design, Qualitative data collection: interviews, focus groups, observation, Mixed methods and secondary data in evaluation
- Quality Standards for Development Evaluation — OECD DAC Network on Development Evaluation, 2010.The DAC's process standards, including what a completed evaluation report must contain and disclose.Cited by: Writing and judging evaluation reports, The UNEG Norms and Standards for Evaluation, Evaluation Quality Assessment and Meta-Evaluation
- Quasi-Experimental Design and Methods. Methodological Briefs: Impact Evaluation No. 8 — UNICEF Office of Research, Florence, 2014.White & Sabarwal — a compact practitioner orientation to RDD among the quasi-experimental options.Cited by: Regression discontinuity designs, Matching and propensity score methods, Interrupted time series analysis
- Randomized Controlled Trials (RCTs). Methodological Briefs: Impact Evaluation No. 7 — UNICEF Office of Research, Florence, 2014.White, Sabarwal & de Hoop — a compact practitioner brief on RCT logic and threats.Cited by: Randomised controlled trials
- Realist Evaluation — Paper prepared for the British Cabinet Office, 2004.Pawson & Tilley — the originators' own summary of the approach they first set out in Realistic Evaluation (SAGE, 1997).Cited by: Realist evaluation
- Realist Evaluation (approach page) — BetterEvaluation (Global Evaluation Initiative), updated continuously.Practitioner overview and resource collection.Cited by: Realist evaluation
- Recommended Performance Indicator Reference Sheet (PIRS): Guidance & Template — USAID, 2017.The reference-sheet template that requires known data limitations to be documented against USAID's five data quality standards.Cited by: Data quality dimensions in M&E, Performance indicator reference sheets (PIRS)
- Regression Discontinuity Designs in Economics — Journal of Economic Literature, 48(2), 281–355, 2010.Lee & Lemieux — the standard survey: identification logic, sharp and fuzzy designs, and the validity-check toolkit.Cited by: Regression discontinuity designs
- Regulation (EU) 2016/679 — General Data Protection Regulation — European Parliament & Council (EUR-Lex official text), 2016.The GDPR: adopted 2016, applicable from 25 May 2018. Principles, lawful bases, special categories, transfers, DPIAs and breach duties.Cited by: Data protection for M&E: Kenya's DPA 2019 and the GDPR
- Results-Based Management Handbook: Harmonizing RBM Concepts and Approaches for Improved Development Results at Country Level — United Nations Development Group (UNDG), 2011.The UN system's harmonised treatment of results levels and the grammar of results statements.Cited by: Indicator types and levels, Indicator quality criteria: SMART, CREAM, SPICED, Baselines, targets and disaggregation, The Balanced Scorecard and strategy maps, OKRs and KPI cascades in institutions, Results-based management (RBM)
- Routine Data Quality Assessment (RDQA) Tool: User Manual — MEASURE Evaluation (USAID), 2017.Operationalises dimension checks as verification and systems-assessment questions.Cited by: Data quality dimensions in M&E, Performance indicator reference sheets (PIRS), DHIS2 in monitoring and evaluation
- Routine Data Quality Assessment Tool: User Manual — MEASURE Evaluation (USAID), 2017.Operational manual for the RDQA: verification steps, system checklist, dashboards and action planning.Cited by: Data quality assessment (DQA)
- Running Randomized Evaluations: A Practical Guide — Princeton University Press, 2014.Glennerster & Takavarasha — field-level guidance on randomisation mechanics, ethics and implementation.Cited by: Randomised controlled trials
- Sampling for Household Surveys — UNHCR Assessment and Monitoring Resource Centre, 2024.Guidance written for displacement settings, where sampling frames are weakest; publication date per the resource-centre record (October 2024).Cited by: Sampling for M&E: methods and sample size
- Schedule Assessment Guide: Best Practices for Project Schedules (GAO-16-89G) — U.S. Government Accountability Office, 2015.An open, auditable statement of what a healthy schedule looks like: logic-linked activities, a credible critical path, reasonable float, a maintained baseline.Cited by: The project management–M&E interface, Earned value management
- Ten Steps to a Results-Based Monitoring and Evaluation System — The World Bank, 2004.Kusek & Rist. Chapter on data collection places quality choices at system-design time.Cited by: Data quality dimensions in M&E, Indicator types and levels, Indicator quality criteria: SMART, CREAM, SPICED, Performance indicator reference sheets (PIRS), Proxy indicators and composite indices, Baselines, targets and disaggregation, The project management–M&E interface, Dashboards and data visualisation for M&E, M&E information systems: data flow, interoperability and architecture, Results-based management (RBM)
- The 'Most Significant Change' (MSC) Technique: A Guide to Its Use — Rick Davies & Jess Dart (self-published guide, funded by CARE International, Oxfam and other agencies), 2005.Version 1.00, April 2005 — the canon: origins, the ten steps, and the technique's own account of its limits.Cited by: Most Significant Change (MSC), Qualitative data collection: interviews, focus groups, observation
- The African Evaluation Principles — African Evaluation Association (AfrEA), 2021.The instrument itself, launched at the 10th AfrEA Conference, Addis Ababa, November 2021; successor to the African Evaluation Guidelines. Full text also mirrored by ALNAP.Cited by: The African Evaluation Principles and Made in Africa Evaluation
- The balance on the balanced scorecard — a critical analysis of some of its assumptions — Management Accounting Research, 11(1), 65–88, 2000.Nørreklit's standard critique: the cause-and-effect chain between perspectives is assumed rather than demonstrated, and the scorecard's claims as a strategic control model are overdrawn.Cited by: The Balanced Scorecard and strategy maps
- The Balanced Scorecard — Measures that Drive Performance — Harvard Business Review, January–February 1992, 1992.Kaplan & Norton's founding article: the four perspectives and the case against steering by financial measures alone.Cited by: The Balanced Scorecard and strategy maps
- The Belmont Report: Ethical Principles and Guidelines for the Protection of Human Subjects of Research — US National Commission for the Protection of Human Subjects of Biomedical and Behavioral Research (hosted by HHS/OHRP), 1979.The source of the respect-beneficence-justice trio and of consent's three elements: information, comprehension, voluntariness.Cited by: Research ethics, do no harm and informed consent in M&E
- The central role of the propensity score in observational studies for causal effects — Biometrika, 70(1), 41–55, 1983.Rosenbaum & Rubin — the founding result: the propensity score is a balancing score.Cited by: Matching and propensity score methods
- The Data Protection Act, No. 24 of 2019 (consolidated text) — Republic of Kenya / Kenya Law, 2019.The official consolidated text: definitions, principles, data-subject rights, controller and processor obligations, transfers and enforcement.Cited by: Data protection for M&E: Kenya's DPA 2019 and the GDPR
- The DCED Standard for Results Measurement — Donor Committee for Enterprise Development (DCED), n.d..The results-measurement standard whose audit discipline includes defensible baselines and projections. Accessed 18 August 2026.Cited by: Baselines, targets and disaggregation
- The GEF Evaluation Policy — Global Environment Facility, Independent Evaluation Office, 2019.The policy frame of one of the most fully institutionalised review-and-evaluation systems in development finance.Cited by: Mid-term reviews
- The M&E Universe — practitioner papers on monitoring and evaluation — INTRAC, n.d..Continuously maintained library of short practitioner papers, including the indicator papers that document the SPICED approach for participatory measurement.Cited by: Indicator quality criteria: SMART, CREAM, SPICED
- The Magenta Book: Central Government Guidance on Evaluation — HM Treasury, United Kingdom, 2020.Distinguishes process and impact questions and ties evaluation planning to the delivery cycle.Cited by: The project management–M&E interface, Mid-term reviews, Randomised controlled trials, Realist evaluation
- The Orange Book: Management of Risk — Principles and Concepts — HM Treasury and Government Finance Function, United Kingdom, 2023.The open, freely available public-sector risk framework this page leans on for its working detail: principles, appetite, and governance.Cited by: Risk registers and risk monitoring
- The Principles for Digital Development — Principles for Digital Development community, living guidance.Nine design guardrails for digital tools in development practice, including designing with the user, reusing existing systems and using open standards.Cited by: M&E information systems: data flow, interoperability and architecture
- The Program Evaluation Standards, 3rd edition (summary of record) — Joint Committee on Standards for Educational Evaluation (JCSEE), 2010.Yarbrough, Shulha, Hopson & Caruthers; published by Corwin. The JCSEE page lists all thirty standards and the five attribute-group purpose statements quoted on this page.Cited by: The Program Evaluation Standards and the AEA Guiding Principles, Evaluation Quality Assessment and Meta-Evaluation
- The RAMESES Projects: quality standards, reporting standards and training materials for realist evaluation and realist synthesis — RAMESES Projects, updated continuously.Home of the RAMESES II reporting and quality standards for realist evaluation.Cited by: Realist evaluation
- The Sphere Handbook: Humanitarian Charter and Minimum Standards in Humanitarian Response, 4th edition — Sphere Association, 2018.Embeds the CHS and the rights-based framing under which information and participation are owed to affected people.Cited by: Accountability to Affected Populations (AAP)
- Theory-Based Impact Evaluation: Principles and Practice (3ie Working Paper 3) — International Initiative for Impact Evaluation (3ie), 2009.White. Why credible impact evaluation opens the causal chain with mixed evidence rather than reporting an estimate alone.Cited by: Mixed methods and secondary data in evaluation, Randomised controlled trials, Contribution analysis, Process tracing for evaluation, Qualitative comparative analysis (QCA)
- Toward Distinguishing Empowerment Evaluation and Placing It in a Larger Context — Evaluation Practice 18(2), 1997.Patton's critical examination: whether empowerment evaluation is distinguishable in practice from other participatory and collaborative approaches.Cited by: Participatory and Empowerment Evaluation
- Towards Evidence-Informed Adaptive Management: A Roadmap for Development and Humanitarian Organisations — ODI Working Paper 565, Overseas Development Institute, n.d..The adaptive-management agenda into which DE feeds in development practice; publication year to be confirmed from the title page.Cited by: Developmental Evaluation
- Towards Evidence-Informed Adaptive Management: A Roadmap for Development and Humanitarian Organisations (Working Paper 565) — Overseas Development Institute (ODI), 2019.Hernandez, Ramalingam & Wild. The case for pairing adaptation with a documented evidence trail, and its practical trade-offs.Cited by: Learning agendas, adaptive management and after-action reviews
- UN Women Evaluation Handbook: How to Manage Gender-Responsive Evaluation — UN Women Independent Evaluation Service, 2022.The UN system's operational guidance for gender-responsive evaluation, end to end — questions, process, team, methods and use. 2022 edition; first issued 2015.Cited by: Feminist and Gender-Responsive Evaluation, Qualitative data collection: interviews, focus groups, observation, Research ethics, do no harm and informed consent in M&E
- Understanding Process Tracing — PS: Political Science & Politics, 44(4), 823–830, 2011.Collier — the standard exposition of the four evidence tests.Cited by: Process tracing for evaluation
- UNDP Evaluation Guidelines — UNDP Independent Evaluation Office, 2021.Agency-level guidance on evaluation planning, quality and management responses across the programme cycle.Cited by: Mid-term reviews, Writing and judging evaluation reports, Dashboards and data visualisation for M&E, Evaluation use and the management response, Evaluation Quality Assessment and Meta-Evaluation
- UNEG Quality Checklist for Evaluation Reports — United Nations Evaluation Group, 2010.The UN system's item-by-item checklist for judging a report's completeness and quality.Cited by: Writing and judging evaluation reports, The UNEG Norms and Standards for Evaluation, Evaluation Quality Assessment and Meta-Evaluation
- Using Evidence in Policy and Practice: Lessons from Africa — Routledge (open access), 2020.Goldman & Pabari (eds.). Case evidence on what makes learning actually change decisions in African government and programme settings.Cited by: Learning agendas, adaptive management and after-action reviews, The African Evaluation Principles and Made in Africa Evaluation, Results-based management (RBM), National M&E systems
- Using Randomization in Development Economics Research: A Toolkit — NBER Technical Working Paper 333, 2006.Duflo, Glennerster & Kremer — the standard technical treatment of randomisation designs, power and inference.Cited by: Randomised controlled trials
- Using Synthetic Controls: Feasibility, Data Requirements, and Methodological Aspects — Journal of Economic Literature, 59(2), 391–425, 2021.Abadie — the method's originator on when synthetic control is and is not appropriate, and the checklist this page follows.Cited by: Synthetic control methods
- Using the Balanced Scorecard as a Strategic Management System — Harvard Business Review, July–August 2007 (reprint of the 1996 article), 2007.The scorecard's second act: from measurement instrument to strategy-management system linking long-term strategy with short-term action.Cited by: The Balanced Scorecard and strategy maps
- Utilisation-Focused Evaluation (approach page) — BetterEvaluation (Global Evaluation Initiative), n.d..Living overview of Patton's approach, its rationale and step sequence. Accessed 18 August 2026.Cited by: Utilization-Focused Evaluation, Developmental Evaluation, Evaluation use and the management response
- Utilization-Focused Evaluation (U-FE) Checklist — Michael Quinn Patton, hosted by BetterEvaluation, 2013.Patton's own operational checklist for running a UFE, updating his 2002 version. The closest thing to a canonical step sequence.Cited by: Utilization-Focused Evaluation, Evaluation use and the management response
- What is an OKR? Definition and Examples — WhatMatters.com (John Doerr's OKR resource), maintained continuously.The canonical formula — 'I will (Objective) as measured by (Key Results)' — and the method's lineage from Andy Grove at Intel through Doerr to Google.Cited by: OKRs and KPI cascades in institutions
- What's Trending in Difference-in-Differences? A Synthesis of the Recent Econometrics Literature — arXiv (published in Journal of Econometrics 235(2), 2218–2244), 2023.Roth, Sant'Anna, Bilinski & Poe — the survey of staggered-adoption and robust-inference advances.Cited by: Difference-in-differences, Synthetic control methods
- World Bank Group Evaluation Principles — World Bank Group & Independent Evaluation Group, 2019.An institution-level statement that utility and learning, not report production, are evaluation's purpose.Cited by: Evaluation use and the management response, Evaluation Quality Assessment and Meta-Evaluation