Meta-Evaluation
Updated 2026-08-11
INTRODUCTION
English translation pending.
CORE DEFINITION
A practice from program evaluation and measurement theory, closely connected to Goodhart's law. The core proposition is that evaluation is itself a fallible product subject to bias, method error, and gaming, so it requires its own quality control. The key qualification is that the check must cover use as well as design, since a sound measurement can still cause harm when applied to decisions it was never meant to support.
SCAFFOLDING EFFECT
Reduce cognitive load
- Criterion check: confirm the metrics actually measure the intended construct. - Gaming scan: ask how people would behave if they optimized the metric directly. - Use review: check whether the results are being applied to decisions they can support.
Anchor fast decisions
Metrics are proxies for the construct of interest, and the gap between proxy and construct creates room for distortion. Once a measure becomes a target, effort shifts toward the measure rather than the underlying goal, so the indicator improves while the outcome does not. Meta-evaluation closes the loop by auditing the proxy, the incentives it creates, and the decisions it feeds.
MINIMUM ACTION
In progress 0/1Practice this model in one real situation:
account_treeGenealogyexpand_more
menu_bookReferencesexpand_more
Source support: Explicit
- eric.ed.govhttps://eric.ed.gov/?id=EJ916544verified
PRIVATE NOTES · Only visible to you
SAVED Q&A
ENTRY Q&A · Private saving available
Ask with a clear boundary
thinkingmodels answers from published entry context only.
Your question is sent to thinkingmodels. The answer uses public entry context only.
RELATED MODELS