Every government programme starts with a promise: reduce poverty, improve literacy, clean up rivers, deliver jobs. But how do we know whether that promise is being kept? This is where the design of policy evaluation becomes critical. A well-designed evaluation is far more than a report card; it is a structured process of listing programme goals, measuring what has actually been achieved, and suggesting course corrections so that performance aligns with original intentions. Yet designing an evaluation well is surprisingly hard, and doing it poorly can be worse than not doing it at all.
Table of Contents
- What does it mean to design an evaluation?
- The three-step logic of evaluation design
- Setting clear goals: the starting point
- Outputs versus outcomes
- Methods for measuring achievement
- Efficiency evaluation
- Impact evaluation
- Process evaluation
- Formative versus summative evaluation
- The institutional architecture for evaluation
- Evaluation as a conflict zone
- The measurement problem in the public sector
- The three pillars of effective evaluation design
- Adequate information
- Adequate resources
- Political will
- A case in point: MGNREGA
- From findings to feedback: closing the loop
What does it mean to design an evaluation?
Designing an evaluation is the deliberate planning that happens before data collection begins. It answers three foundational questions. What are we trying to measure? How will we measure it credibly? And what will we do with the findings? The Development Monitoring and Evaluation Office defines evaluation as a rigorous and independent assessment of ongoing or completed activities that determines the extent to which stated objectives are being achieved, and contributes to decision-making.
The key word is design. Without an intentional blueprint, evaluations often turn into vague exercises in output-counting: how many toilets built, how many students enrolled, how many cylinders distributed. These numbers matter, but they rarely tell us whether the policy actually changed lives. A thoughtfully designed evaluation goes deeper and asks whether the intended transformation is taking place.
The three-step logic of evaluation design
A sound evaluation design usually follows a clear logical sequence. The first step is articulating programme goals in measurable terms. A scheme that claims to “empower women” must translate that abstract goal into observable indicators such as changes in income, asset ownership, or decision-making authority at home. The second step is measuring achievements against those indicators using reliable data. The third step is suggesting changes, whether refinements to delivery, reallocation of resources, or, in rare cases, termination of the programme.
Setting clear goals: the starting point
Every evaluation begins with the question, “What is success?” This sounds obvious, but it is frequently the weakest link in Indian policy design. When goals are vague or politically ambitious, evaluators struggle to know what to measure. The Swachh Bharat Abhiyan offers a telling example: the programme aimed to eliminate open defecation, but analysts found that building toilets did not automatically change behaviour. In many areas, newly constructed toilets went unused because of cultural factors, maintenance issues, or water scarcity. These were outcomes the original design had not anticipated.
A rigorous evaluation would have captured both the output (toilets built) and the outcome (toilets used). Designing for that distinction requires evaluators to list not just the stated goals but also the implicit behavioural assumptions behind them.
Outputs versus outcomes
Policy outputs are the direct decisions and deliverables produced by implementers. Outcomes, however, describe what actually happens to the target group. A housing scheme may have an output of 10,000 houses constructed, but its outcome is whether families have secure, dignified shelter that improves their quality of life. Evaluation design must force this distinction, because programmes often meet output targets while missing outcome goals entirely.
Methods for measuring achievement
Once goals are defined, the next design choice is methodology. Different questions demand different tools, and picking the wrong method can lead to misleading conclusions.
Efficiency evaluation
Efficiency evaluation examines the relationship between inputs and outputs. Are we getting the maximum possible output from the resources invested? Could the same results be achieved more cheaply? Cost-effectiveness analysis, a common tool here, compares alternatives to identify the lowest cost for a given level of effectiveness, or the maximum effectiveness for a given cost.
Impact evaluation
Impact evaluation looks at the broader, longer-term effects of a policy on the target population and society. It tries to isolate what changed because of the intervention, separating genuine policy effects from other influences. Rigorous impact evaluations often rely on counterfactual designs such as randomised controlled trials or quasi-experimental methods, which compare beneficiaries with similar non-beneficiaries.
Process evaluation
Process evaluation, sometimes called implementation evaluation, assesses how a programme is being delivered. It asks whether frontline staff are following guidelines, whether beneficiaries are being correctly identified, and whether bottlenecks are slowing delivery. A structural analysis of MGNREGA implementation, for instance, has shown that factors such as strategic communication, planning, resources, and well-defined networks are critical to implementation success.
Formative versus summative evaluation
A final methodological distinction is between formative evaluations, conducted during implementation to guide mid-course corrections, and summative evaluations, conducted at the end to judge overall effectiveness. Both serve different purposes. Formative evaluations help programmes adapt in real time; summative evaluations inform decisions about continuation, scaling, or termination.
The institutional architecture for evaluation
Good evaluation design cannot sit in a vacuum. It needs institutions with the mandate, independence, and expertise to commission and conduct credible assessments. The Development Monitoring and Evaluation Office was constituted in September 2015 as an attached office of NITI Aayog, merging the erstwhile Programme Evaluation Organization and Independent Evaluation Office. Its mandate is to monitor and evaluate the implementation of government programmes, identify resource needs, and strengthen the likelihood of success and scope of delivery.
To ensure third-party credibility, the government has made evaluation of Centrally Sponsored Schemes and Central Sector schemes mandatory before they come up for fresh appraisal, with DMEO responsible for conducting independent third-party evaluation in a time-bound manner. This is an important design principle: evaluators must be insulated from the ministries whose programmes they are assessing, otherwise the findings risk being shaped by conflicts of interest.
Evaluation as a conflict zone
Here is an uncomfortable truth often glossed over in textbooks: evaluations are political events, not just technical ones. A well-designed evaluation can threaten careers, redistribute budgets, or even end a flagship scheme. Unsurprisingly, this creates friction.
When negative results are likely to trigger policy termination or deep restructuring, stakeholders with a vested interest in the status quo often resist rigorous evaluation. They may delay data sharing, dispute methodology, or lobby for softer framing of findings. This is why policy evaluation is sometimes described as a neglected area of the policy process, facing multiple problems, challenges, and dilemmas.
The measurement problem in the public sector
Measuring results in the public sector is genuinely difficult. Long-term goals in health, education, and rural development relate to quality of life and human capability, which are elusive qualities to quantify when an evaluation must be delivered quickly. How do you measure whether an education policy has improved the “quality” of learning? Test scores offer one indicator, but they may not capture critical thinking, creativity, or social skills. Similarly, assessing whether a community programme has genuinely enhanced empowerment or social cohesion is notoriously hard.
This measurement challenge can produce a phenomenon called goal displacement, where programmes start focusing on easily measured outputs rather than harder-to-measure outcomes, simply because that is what evaluators can capture.
The three pillars of effective evaluation design
A carefully designed evaluation needs three ingredients working together: information, resources, and political will. The absence of any one of them can hollow out even the most elegant evaluation plan.
Adequate information
Evaluation designs collapse without baseline data. If you do not know what rural literacy looked like before a scheme began, you cannot credibly claim the scheme improved it. Policy evaluation in India faces several challenges including lack of baseline data, difficulty in establishing causation, political sensitivities around findings, and limited evaluation capacity within government. Addressing this requires investment in administrative data systems, surveys, and management information systems that track inputs, outputs, and outcomes over time.
Adequate resources
Evaluations cost money. High-quality designs need trained personnel, field investigators, statistical analysis, and time. When evaluations are squeezed into tight budgets or unrealistic timelines, shortcuts become inevitable. Sample sizes shrink, methods are simplified, and findings lose credibility. Strengthening evaluation capacity at both central and state levels is therefore not a luxury but a precondition.
Political will
Perhaps the most important ingredient is political will, the genuine appetite among leaders to know the truth, even when it is inconvenient. Without this, evaluations become symbolic exercises. A national evaluation policy is widely seen as essential to institutionalise this culture and protect evaluators from pressure to soften findings.
A case in point: MGNREGA
The Mahatma Gandhi National Rural Employment Guarantee Act is one of the world’s largest social protection schemes and has been the subject of extensive evaluation. The findings are nuanced. On one hand, research suggests the scheme has benefited the poorest households, especially Dalits and women, by providing a safety net, raising rural wages, and reducing dependence on high-caste employers. On the other, a parliamentary standing committee has flagged serious concerns, including nominal wages that discourage participation, delayed wage payments, and weak enforcement of the provision for unemployment allowance.
This is exactly what a well-designed evaluation should produce: an honest, mixed picture that helps policymakers refine rather than abandon a programme. The challenge is translating such findings into action, which brings us back to political will.
From findings to feedback: closing the loop
The ultimate test of an evaluation design is whether its findings actually change anything. Evaluations that sit on shelves gathering dust represent wasted public money. Good design therefore includes a dissemination and utilisation plan from the start: who will receive the findings, in what format, by when, and what decision points they will inform. When this feedback loop is alive, evaluation stops being a compliance ritual and becomes a genuine engine of policy learning.
What do you think? If an evaluation of a popular welfare scheme reveals serious design flaws, should the government prioritise correcting them or protecting the scheme from political backlash? And in a country as diverse as ours, can a single evaluation framework really capture what “success” looks like across different states and communities?
References
- https://dmeo.gov.in/evaluation
- https://idronline.org/monitoring-and-evaluation-public-policies-rct-india/
- https://www.researchgate.net/publication/381390070_THE_PROCESS_OF_POLICY_FORMULATION_AND_IMPLEMENTATION_OF_MGNREGA
- https://dmeo.gov.in/content/who-we-are
- https://polsci.institute/public-policy-administration-india/public-policy-process-stages/
- https://prsindia.org/policy/report-summaries/critical-evaluation-of-mgnrega
Leave a Reply