Aller au contenu

Baseline, Midline and Endline Studies: Differences, Uses and Examples

Baseline, midline and endline studies are measurement points in an evaluation cycle. Their value comes from consistent indicators, comparable samples and a design that matches the decisions the programme needs to make.

Why these three studies matter

A programme can be busy, well funded and widely implemented without anyone being able to say with confidence whether conditions have changed. Baseline, midline and endline studies create structured measurement points that help programmes move from activity reporting to evidence about results. They are especially useful when the same core indicators can be measured consistently over time.

The three terms are related, but they are not interchangeable. A baseline establishes the starting position before an intervention, or as close to the start as is practical. A midline examines progress while implementation is still under way. An endline measures the situation at or near the end of the intervention. Together, they can describe the direction and scale of change, while the broader evaluation design determines how confidently that change can be attributed to the programme.

Baseline study: establishing the starting point

A baseline answers a simple but consequential question: what did the situation look like before the programme had time to influence it? It may measure household conditions, service access, knowledge, behaviours, institutional capacity, market outcomes or any other indicator linked to the programme theory of change.

A strong baseline does more than produce percentages. It defines indicators precisely, documents the sampling approach, tests tools, records contextual factors and creates a dataset that can be reproduced at later rounds. If the endline cannot recreate the baseline measurement, comparison becomes weak even when both studies appear technically sound on their own.

Midline study: learning while there is still time to act

A midline is not merely a smaller endline. Its strategic purpose is learning. It can show whether implementation is reaching intended groups, whether early outcomes are moving in the expected direction and whether barriers are emerging. This makes the midline particularly valuable for adaptive management.

For example, an education programme may discover at midline that attendance has improved but learning outcomes have not. That finding can trigger investigation into teaching quality, dosage, materials or household constraints before the programme closes. Midline studies are most useful when findings are delivered quickly enough to influence decisions.

Endline study: measuring the final position

An endline repeats key baseline measures at the end of the intervention and examines what has changed. It can assess outcome levels, variation across groups and locations, implementation experience and sustainability prospects. If the evaluation includes a credible comparison or counterfactual, the endline may contribute to stronger conclusions about programme effects.

Without a counterfactual, a baseline-to-endline difference should not automatically be called impact. Other events may have affected the same indicators. Good reporting separates observed change from causal attribution and explains what the evaluation design can and cannot support.

Design principles that make comparisons credible

Comparability is the discipline that connects the three rounds. Keep indicator definitions stable unless there is a documented reason to change them. Use equivalent question wording and response options, preserve the intended population and sampling logic, and manage seasonality where it could influence results. Changes in data-collection mode can also create mode effects, so any shift from face-to-face to phone or online collection should be tested and documented.

Programme teams should also decide early whether they need repeated cross-sectional samples or a panel that follows the same respondents. Panels can measure individual-level change but require stronger tracking and attrition management. Repeated cross-sections are often operationally simpler and can remain population-representative when sampled well.

What should be included in the study plan?

A practical measurement plan should specify the theory of change, evaluation questions, indicators, data sources, sampling frame, sample size, disaggregation requirements, fieldwork method, quality-control protocol, analysis plan and reporting timetable. It should also identify risks such as insecurity, migration, inaccessible communities or programme contamination.

When these elements are agreed before the first round, later comparisons become faster, cleaner and more defensible. The goal is not three disconnected reports; it is one measurement architecture observed at meaningful points in time.

Choosing indicators that survive across rounds

A baseline is only useful for later comparison if its indicators are defined in a way that can be repeated. That means specifying the numerator, denominator, unit of analysis, reference period and source for every core measure before fieldwork begins. Terms that appear simple, such as employment, access, participation or satisfaction, can be interpreted differently across teams and survey rounds.

An indicator reference sheet reduces that ambiguity by stating exactly what counts and what does not. It also helps analysts distinguish programme indicators from contextual variables that are useful for interpretation but are not themselves outcome measures. When indicators are likely to be disaggregated by sex, age, location, disability status or another characteristic, the study should plan those variables consistently from the beginning rather than trying to reconstruct them at endline.

Managing timing, seasonality and external change

Timing can materially affect the meaning of a baseline-to-endline comparison. Agricultural income, food security, school attendance, disease patterns, employment and household expenditure may vary by season even when a programme has had no effect. Ideally, repeated rounds are scheduled at comparable points in the seasonal calendar.

When this is impossible, the evaluation should document the difference and analyse whether seasonality could explain part of the observed change. The same principle applies to external shocks. Elections, inflation, conflict, disease outbreaks, policy changes, major infrastructure projects or market disruptions can affect outcomes independently of the intervention. A strong study records these contextual events and uses qualitative evidence, administrative data or comparison groups where possible to understand their influence.

What to do when a perfect baseline is not available

Programmes sometimes begin before a baseline can be completed. This does not make evaluation impossible, but it does change the design. Teams may reconstruct selected pre-programme information from reliable administrative records, earlier surveys or other documented sources.

They may also use retrospective questions for clearly memorable events, while acknowledging recall bias. In other cases, a comparison group, contribution analysis, outcome harvesting or Most Significant Change can provide useful evidence about results without pretending to recreate a baseline that never existed. The key is to match the method to the claim. If the available evidence supports a conclusion about contribution, progress or perceived change, the report should say exactly that rather than presenting it as a stronger causal impact estimate.

From measurement rounds to programme learning

The value of baseline, midline and endline work increases when the three rounds are treated as a learning system rather than separate contracts. Baseline findings can refine targeting and implementation assumptions. Midline findings can trigger corrective action, identify groups being left behind or reveal indicators that are not moving as expected.

Endline findings can then assess final outcomes and explain which implementation choices appear most important. A useful learning process includes short decision briefs, management discussions and a clear record of actions taken in response to evidence. This makes evaluation part of programme management instead of an exercise completed mainly for reporting. It also leaves future teams with a documented chain from starting conditions, through adaptation, to final results.

How Surveysphere Africa can support

Planning a baseline, midline or endline study across one or more African markets? Surveysphere Africa supports evaluation design, fieldwork, data quality, analysis and reporting from the first measurement round to the final learning product. Talk to our research team

Planning a baseline, midline or endline study?

Tell us about your programme, the indicators you need to track and your timeline. We can help you design comparable measurement rounds from baseline to endline.

Talk to our research team