Evidence and context
An evaluation can ask the right question and still lack the evidence needed to answer it convincingly. That distinction matters when we describe the need for better program records.
Raifman and colleagues assessed the methodological quality of 37 randomly selected program evaluations from five major global health funders. Two researchers rated relevance, validity and reliability, resolving disagreements together. Most evaluations asked relevant questions, but the authors identified substantial weaknesses in the data and methods used. Raifman et al., 2018, Journal of Development Effectiveness
The assessment concerns this sample and its historical setting. It is not a current failure rate for all aid or UN programs. It does not show that the evaluated programs achieved nothing or that an evidence-capture tool would have corrected their methods.
Our product interpretation is that source availability and methodological quality should be treated as separate requirements. A program history could help a reviewer locate observations, inspect timing and understand changes. It cannot create a valid comparison group or a reliable measure that was never collected.
The strongest program record makes that boundary clear. It allows a reviewer to see what is available, what the record can support and what remains unresolved.
Implications for research and practice
Relevant evaluation questions need credible data and methods. The quality of a report cannot be judged from its conclusions or length alone.
Scope and limitations
This is a selective narrative brief based on the linked publications, not a systematic review or a new empirical study. Illustrative scenarios and practice implications are editorial analysis. Findings should be interpreted within the settings, methods and observation periods of their original sources.