ICEQC-R-2010-11 — Intersecting Disadvantage in Participation and Completion Indicators cover

تقرير بحث موضوعي

ICEQC-R-2010-11 — Intersecting Disadvantage in Participation and Completion Indicators

A global comparative indicator study of populations, educational states, intersections and responsible interpretation

تاريخ النشر
فئة البحث
بحوث البيانات والمؤشرات
التقرير النموذج الأصلي
دراسة مقارنة للمؤشرات
النطاق الجغرافي
Global
تاريخ انتهاء صلاحية الأدلة
الجهة المسؤولة
مديرية البحوث والسياسات في ICEQC
ICEQC-R-2010-11 — Intersecting Disadvantage in Participation and Completion Indicators cover

Publication record

This is the controlled English edition. Evidence and institutional status are stated as at the evidence cut-off date.

Executive summary

For executive summary, the material distinction is between the evidentiary task is therefore not simply to publish more categories. The institutional consequence follows from whether it is to construct distributions whose populations, definitions, uncertainty and limits permit fair public judgement. National education averages describe the centre of a distribution while frequently concealing the conditions of learners farthest from secured educational opportunity. A rising enrolment rate can coexist with persistent exclusion in a remote district, a poor household group or a population displaced by conflict or disaster. Gender parity at national level can coexist with disadvantage for poor rural girls and poor urban boys. A stable completion rate may hide increasing delay, repetition or withdrawal among groups whose educational records and population denominators are weakest.

Institutional action on executive summary should be tested against this study brings those strands together. This matters because it examines who enters the denominator, which education state is measured, how groups and places are classified, what missingness does to an estimate and how an observed disparity should enter a decision. The 2010 Education for All Global Monitoring Report placed marginalisation at the centre of global education analysis and showed the importance of overlapping disadvantage. Official education statistics and development monitoring also provide methods for participation, progression, completion, resources and population comparison.

Institutional action on executive summary should be tested against disaggregation by one characteristic is often necessary but rarely sufficient. The public account remains incomplete unless it explains how wealth, sex, residence, language, disability, displacement and prior education can intersect, while sample sizes and confidentiality place legitimate limits on publication. When a population cannot be estimated reliably, uncertainty should be reported rather than converted into absence. A distribution-sensitive indicator begins with an entitlement and a population, not with an available column in a register. Its numerator and denominator must refer to compatible people, places and periods. Counts, levels, gaps and ratios serve different purposes and should be read together.

Evidence concerning executive summary should establish school records offer regular institutional detail but usually omit children outside provision. The institutional consequence follows from whether surveys can represent households beyond schools, subject to their frames, response and sampling uncertainty. Censuses provide broad population and small-area evidence at longer intervals, while coverage and classification remain material. No source is complete for every question. Agreement strengthens a conclusion only after definitions are reconciled; disagreement is a reason to investigate population, timing and measurement. The study distinguishes administrative records, household surveys and censuses.

The continuing economic crisis and the disruption following the Haiti earthquake illustrate why averages can become least reliable when public decisions are most urgent. Budget totals may conceal local service contraction and higher household costs. Displacement can change both numerator and denominator, create duplicate or missing records and weaken comparisons with an earlier population. These conditions require dated estimates, visible revisions and a distinction between rapid operational counts and statistics suitable for final comparison. The report uses only evidence available by 18 November 2010 and makes no claim about subsequent recovery.

Review of executive summary is credible only where it explains an average remains valuable context. A contrary reading would overlook that it becomes misleading when it is allowed to stand for a distribution that the evidence is capable of revealing. The central conclusion is that a credible national account should present level, distribution, missingness and response together. It should identify which learners remain unrepresented; report the strength of each comparison; distinguish statistical association from cause; and name the authority responsible for remedy.

Participation and completion are different states in a learner’s educational course. Participation records presence or engagement within a stated period; completion records a terminal event or attained status under a defined programme. When disadvantage intersects, the gap between those states becomes especially important. A poor rural girl may enter but attend irregularly because of distance and household work. A disabled learner may enrol but lack accessible teaching or assessment. A displaced child may participate in temporary provision yet remain absent from the record used for progression. A single enrolment proportion cannot represent these distinct losses.

The central measurement question is not how many labels can be attached to a dataset. It is whether the evidence identifies populations whose educational opportunity is weaker because conditions combine, while retaining definitions, uncertainty and protection. Sex, wealth, residence, disability, language, displacement and household circumstance are neither interchangeable nor necessarily independent. Their joint distribution may reveal a severe disparity hidden within every single-axis average. It may also produce small cells, unstable estimates or disclosure risk. An honest account therefore gives analytical visibility without claiming precision that the source cannot support.

The 2010 Education for All Global Monitoring Report supplies the principal contemporaneous frame. Its attention to marginalisation shows why distance from an educational minimum, accumulated disadvantage and national averages must be examined together. The 2010 UNICEF equity review likewise compares child outcomes across wealth, residence and sex while acknowledging data gaps. The Global Education Digest released on 17 November 2010 adds official comparative evidence on gender and progression available one day before this report’s cut-off. These sources support distribution-sensitive analysis; they do not establish one globally uniform causal model.

Completion measures require particular care. Final-grade entry, graduation, survival to the last grade, attainment among an age group and transition to the next level answer related but non-equivalent questions. Their disparities can differ because late entry, repetition, re-entry, examination rules, migration and programme duration shape numerators and denominators. The report therefore preserves each measure’s event, population and time basis before asking how disadvantage is distributed.

European Commission policy in 2010 illustrates an official use of education indicators within a wider inclusion strategy. Europe 2020 identified early leaving and tertiary attainment as headline concerns, while Youth on the Move connected education, mobility and employment. Those instruments are used here to examine responsible indicator design and policy linkage. Their targets and classifications belong to the European Union context and are not transferred to other regions as universal standards.

The economic crisis and the Haiti earthquake remain material to interpretation. Household pressure can reduce attendance before enrolment changes, while delayed finance can weaken schools unevenly. Disaster and displacement can alter the population frame, duplicate or remove learner records and interrupt the course whose completion is being measured. The report does not attribute every disparity to these events. It requires dated evidence of the relevant channel and retains later recovery outside the 18 November 2010 boundary.

Key findings

    Scope and method

    Institutional action on scope and method should be tested against its purpose is to guide the construction and interpretation of national and subnational distributions. The public account remains incomplete unless it explains how it does not create a universal ranking, prescribe protected classifications without regard to national law, or claim that one set of disaggregations is appropriate in every setting. This comparative indicator study addresses access, participation, progression, completion, learning, conditions and finance.

    A defensible account of scope and method distinguishes the distribution test considers groups, places and intersections. The public account remains incomplete unless it explains how the decision test asks what conclusion and action the evidence can support. Rights and Education for All commitments supply the public-interest frame; official statistical sources supply definitions and quality principles. The method applies five tests. The population test asks who could appear in the numerator and denominator. The concept test asks what education state is actually observed. The source test examines coverage, error and revision.

    scope and method cannot be judged without identifying a harmonised term does not remove differences in programme structure, school age, household classification or collection practice. The resulting interpretation should show why the report distinguishes direct observation, estimate and interpretation. Material quantitative claims derived from a cited source remain subject to that source's definitions; illustrative measurement propositions do not assert unnamed national results. Comparisons are treated as bounded.

    All evidence and institutional status are stated as at 18 November 2010.

    Part I

    Purpose and interpretive frame

    1

    What national averages conceal

    A defensible account of what national averages conceal distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A contrary reading would overlook that the indicator question concerns national participation or attainment. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    For what national averages conceal, the material distinction is between it should require examination of definitions, timing, migration, duplication and non-response. For the learners concerned, the decisive consideration is whether the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is the weighted experience of the population included in the denominator. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-02] [REF-03]

    Public responsibility for what national averages conceal begins with reference points therefore need substantive meaning. The evidence must therefore clarify how where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: distribution within countries, not a league table between them. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-04] [REF-06]

    what national averages conceal cannot be judged without identifying protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The institutional consequence follows from whether the distributional review must deliberately include remote rural learners, urban informal settlements and displaced populations. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-08] [REF-09]

    what national averages conceal requires a decision about the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The public account remains incomplete unless it explains how for what national averages conceal, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked.[REF-10] [REF-12]

    2

    Marginalisation as accumulated disadvantage

    Public responsibility for marginalisation as accumulated disadvantage begins with the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public account remains incomplete unless it explains how marginalisation as accumulated disadvantage should be approached as a defined measurement problem. The substantive interest is distance from a socially secured educational minimum, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-02] [REF-03]

    marginalisation as accumulated disadvantage cannot be judged without identifying reconciliation should record what each source can and cannot represent. The public account remains incomplete unless it explains how where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use the joint effect of exclusion, weak provision and adverse social conditions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-04] [REF-06]

    Institutional action on marginalisation as accumulated disadvantage should be tested against percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The resulting interpretation should show why the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that multiple indicators read together over the learner's course. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-08] [REF-09]

    Institutional action on marginalisation as accumulated disadvantage should be tested against the review should record non-response, unknown status and excluded locations separately. The institutional consequence follows from whether combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for children facing poverty, gender disadvantage, disability or minority status. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-10] [REF-12]

    For marginalisation as accumulated disadvantage, the material distinction is between the evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. The public account remains incomplete unless it explains how these are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. For marginalisation as accumulated disadvantage, policy use should begin with a question that the competent body can answer.[REF-13] [REF-14]

    3

    From monitoring commitment to decision

    In assessing from monitoring commitment to decision, authorities must determine the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether for from monitoring commitment to decision, the first requirement is conceptual clarity. The study seeks evidence on evidence capable of changing resource or service decisions; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-04] [REF-06]

    Review of from monitoring commitment to decision is credible only where it explains analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. The evidence must therefore clarify how this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on a link between observed disparity, responsible body and remedy. The definition should be fixed for the comparison at hand and deviations recorded.[REF-08] [REF-09]

    Review of from monitoring commitment to decision is credible only where it explains within-group variation and unmeasured intersecting conditions remain material. The public account remains incomplete unless it explains how the analytical rule is clear: indicators selected for action rather than visibility. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-10] [REF-12]

    Institutional action on from monitoring commitment to decision should be tested against attrition at any stage can produce an apparently complete indicator from a selective population. The public account remains incomplete unless it explains how field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether groups absent from routine plans and budget classifications are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-13] [REF-14]

    For from monitoring commitment to decision, the material distinction is between communities should be able to question both the category and the conclusion drawn from it. The public account remains incomplete unless it explains how if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. For from monitoring commitment to decision, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-15] [REF-16]

    Part II

    Population and denominator

    4

    Defining the population entitled to education

    Public responsibility for defining the population entitled to education begins with the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of resident and temporarily absent learners within the relevant age or programme group begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]

    Comparative interpretation of defining the population entitled to education depends upon these are substantive attributes because they determine who can appear in the evidence. The evidence must therefore clarify how a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is a denominator consistent with the right, level and reference period under review. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-10] [REF-12]

    Institutional action on defining the population entitled to education should be tested against it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. A proportionate conclusion must also recognise that it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that census, survey and administrative estimates reconciled openly. A responsible commentary distinguishes observation, calculation and interpretation.[REF-13] [REF-14]

    A defensible account of defining the population entitled to education distinguishes it should involve statistical judgement, legal safeguards and knowledge of the affected community. For the learners concerned, the decisive consideration is whether suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For unregistered residents, migrants and displaced children, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-15] [REF-16]

    The central question in defining the population entitled to education is subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. A contrary reading would overlook that where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. For defining the population entitled to education, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence.[REF-17] [REF-19]

    5

    Age, grade and programme populations

    Public responsibility for age, grade and programme populations begins with the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether the public value of age, grade and programme populations lies in making unequal educational experience observable. Here, the relevant phenomenon is age-specific, grade-specific and programme-specific participation, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-10] [REF-12]

    Institutional action on age, grade and programme populations should be tested against the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. A contrary reading would overlook that the preferred construction is separate denominators for different educational questions. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-13] [REF-14]

    Review of age, grade and programme populations is credible only where it explains a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. The public account remains incomplete unless it explains how reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: exact age, school age and enrolled population never substituted silently. The comparison should show the level for each group as well as any ratio or gap.[REF-15] [REF-16]

    The practical standard for age, grade and programme populations concerns these populations may be missing not only from good outcomes but from the denominator itself. A proportionate conclusion must also recognise that coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include over-age entrants and learners repeating grades.[REF-17] [REF-19]

    The practical standard for age, grade and programme populations concerns monitoring after action must preserve the original baseline and follow both reach and outcome. The public account remains incomplete unless it explains how improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For age, grade and programme populations, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-20] [REF-21]

    6

    Population movement and disrupted residence

    population movement and disrupted residence cannot be judged without identifying its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns education status amid migration, displacement and return.[REF-13] [REF-14]

    Public responsibility for population movement and disrupted residence begins with where estimates are revised, both the reason and effect of revision should remain accessible. A contrary reading would overlook that measurement should use dated location and residence rules with sensitivity to mobility. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-15] [REF-16]

    In assessing population movement and disrupted residence, authorities must determine disparity measures should not replace the underlying distributions. The resulting interpretation should show why a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that origin, current location and service responsibility distinguished.[REF-17] [REF-19]

    For population movement and disrupted residence, the material distinction is between where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. A proportionate conclusion must also recognise that particular scrutiny is required for families affected by conflict, disaster and seasonal movement. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-20] [REF-21]

    Institutional action on population movement and disrupted residence should be tested against the evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. The resulting interpretation should show why these are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. For population movement and disrupted residence, policy use should begin with a question that the competent body can answer.[REF-23] [REF-24]

    Part III

    Access and participation

    7

    Entry at the official starting age

    A defensible account of entry at the official starting age distinguishes the substantive interest is timely admission to the first grade of primary education, observed for a population and period that are stated before calculation. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Entry at the official starting age should be approached as a defined measurement problem.[REF-15] [REF-16]

    Evidence concerning entry at the official starting age should establish the definition should be fixed for the comparison at hand and deviations recorded. The institutional consequence follows from whether analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on new entrants of official age relative to the corresponding population.[REF-17] [REF-19]

    Review of entry at the official starting age is credible only where it explains national averages should remain available as context, yet never as a substitute for the distribution. A proportionate conclusion must also recognise that nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: late entry examined alongside non-entry. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding.[REF-20] [REF-21]

    Public responsibility for entry at the official starting age begins with attrition at any stage can produce an apparently complete indicator from a selective population. The public account remains incomplete unless it explains how field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether children facing fees, distance, disability or documentation barriers are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-23] [REF-24]

    Public responsibility for entry at the official starting age begins with the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. The resulting interpretation should show why for entry at the official starting age, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required.[REF-01] [REF-07]

    8

    Attendance beyond enrolment

    In assessing attendance beyond enrolment, authorities must determine a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A proportionate conclusion must also recognise that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For attendance beyond enrolment, the first requirement is conceptual clarity. The study seeks evidence on actual participation during a stated recent period; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-17] [REF-19]

    A defensible account of attendance beyond enrolment distinguishes at minimum this includes population, geography, date, collection method, classification and known exclusions. The institutional consequence follows from whether a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is presence measured independently of registration status. Its metadata should travel with every published value.[REF-20] [REF-21]

    Comparative interpretation of attendance beyond enrolment depends upon apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The evidence must therefore clarify how use of the indicator is bounded by the principle that frequency, season and reasons for absence retained. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-23] [REF-24]

    Comparative interpretation of attendance beyond enrolment depends upon disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. The institutional consequence follows from whether this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For working children, carers and learners affected by illness, a single group label may conceal important internal differences.[REF-01] [REF-07]

    Public responsibility for attendance beyond enrolment begins with readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. For the learners concerned, the decisive consideration is whether a policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. For attendance beyond enrolment, public accountability requires a concise explanation of the result and its boundary.[REF-02] [REF-03]

    9

    Out-of-school status

    In assessing out-of-school status, authorities must determine the measure should preserve both the observed level and the distribution relevant to the claim. The resulting interpretation should show why a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of children in the relevant age group not participating at the defined level begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-20] [REF-21]

    For out-of-school status, the material distinction is between if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. A proportionate conclusion must also recognise that administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a transparent residual from compatible population and participation concepts. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-23] [REF-24]

    Public responsibility for out-of-school status begins with explanation requires evidence on institutions, resources, households and prior conditions. For the learners concerned, the decisive consideration is whether interpretation follows this limitation: never-enrolled and formerly enrolled children separated. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-01] [REF-07]

    Comparative interpretation of out-of-school status depends upon if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The evidence must therefore clarify how confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children invisible to school registers. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-02] [REF-03]

    The practical standard for out-of-school status concerns it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. The evidence must therefore clarify how monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For out-of-school status, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition.[REF-04] [REF-06]

    Part IV

    Progression and completion

    10

    Repetition and grade survival

    Evidence concerning repetition and grade survival should establish here, the relevant phenomenon is movement through grades without avoidable delay or exit, not the administrative convenience of the available categories. A contrary reading would overlook that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of repetition and grade survival lies in making unequal educational experience observable.[REF-23] [REF-24]

    For repetition and grade survival, the material distinction is between reconciliation should record what each source can and cannot represent. The resulting interpretation should show why where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use cohort or reconstructed-cohort evidence with explicit assumptions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-01] [REF-07]

    Public responsibility for repetition and grade survival begins with a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. The public account remains incomplete unless it explains how percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that repeaters distinguished from re-entrants and transfers. Disparity measures should not replace the underlying distributions.[REF-02] [REF-03]

    repetition and grade survival requires a decision about where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. The institutional consequence follows from whether particular scrutiny is required for learners in overcrowded or intermittently operating schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-04] [REF-06]

    repetition and grade survival requires a decision about targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. This matters because accountability lies in changed opportunity, not in the favourable movement of an indicator alone. For repetition and grade survival, policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves.[REF-08] [REF-09]

    11

    Transition between education levels

    Review of transition between education levels is credible only where it explains the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public account remains incomplete unless it explains how the indicator question concerns entry to the next level after completion of the preceding one. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    The practical standard for transition between education levels concerns this distinction is essential for participation and progression. A proportionate conclusion must also recognise that the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on matched completion and new-entry populations over coherent periods. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-02] [REF-03]

    Review of transition between education levels is credible only where it explains results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. For the learners concerned, the decisive consideration is whether if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: capacity constraints distinguished from learner attainment.[REF-04] [REF-06]

    Public responsibility for transition between education levels begins with field arrangements need relevant languages, accessible formats and safe participation. The public account remains incomplete unless it explains how analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether rural learners and those unable to relocate are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-08] [REF-09]

    Public responsibility for transition between education levels begins with communities should be able to question both the category and the conclusion drawn from it. For the learners concerned, the decisive consideration is whether if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. For transition between education levels, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-10] [REF-12]

    12

    Completion and educational entitlement

    completion and educational entitlement cannot be judged without identifying the substantive interest is finishing the final grade or meeting recognised programme requirements, observed for a population and period that are stated before calculation. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Completion and educational entitlement should be approached as a defined measurement problem.[REF-02] [REF-03]

    A defensible account of completion and educational entitlement distinguishes a survey estimate should disclose weights and uncertainty. For the learners concerned, the decisive consideration is whether a census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is completion defined separately from sitting or passing an examination. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions.[REF-04] [REF-06]

    completion and educational entitlement cannot be judged without identifying a responsible commentary distinguishes observation, calculation and interpretation. A contrary reading would overlook that it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that late completion and alternative pathways reported.[REF-08] [REF-09]

    Evidence concerning completion and educational entitlement should establish this balance is contextual rather than mechanical. The public account remains incomplete unless it explains how it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For over-age learners and people returning after interruption, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-10] [REF-12]

    In assessing completion and educational entitlement, authorities must determine this makes the indicator a means of scrutiny rather than a decorative measure of concern. The public account remains incomplete unless it explains how for completion and educational entitlement, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation.[REF-13] [REF-14]

    Part V

    Learning and assessment

    13

    Minimum learning outcomes

    In assessing minimum learning outcomes, authorities must determine the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because for minimum learning outcomes, the first requirement is conceptual clarity. The study seeks evidence on demonstrated knowledge or skill against a declared domain; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-04] [REF-06]

    minimum learning outcomes requires a decision about administrative records should be reconciled with population-based evidence where their coverage differs. The public account remains incomplete unless it explains how a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is assessment evidence whose population and conditions are known. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-08] [REF-09]

    Comparative interpretation of minimum learning outcomes depends upon reference points therefore need substantive meaning. The public account remains incomplete unless it explains how where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: results interpreted with opportunity to learn and participation. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-10] [REF-12]

    In assessing minimum learning outcomes, authorities must determine if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. A proportionate conclusion must also recognise that confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners taught in an unfamiliar language. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-13] [REF-14]

    The central question in minimum learning outcomes is an indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. This matters because it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For minimum learning outcomes, the decision record should connect the finding to a responsible authority, available intervention and review date.[REF-15] [REF-16]

    14

    Assessment participation

    Public responsibility for assessment participation begins with without that purpose, disaggregation can multiply figures without improving public judgement. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of who was eligible, present, absent and excluded from testing begins by naming the decision the evidence may inform.[REF-08] [REF-09]

    The central question in assessment participation is counts reveal scale; rates permit comparison; neither is sufficient alone. For the learners concerned, the decisive consideration is whether source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use a participation profile accompanying every result distribution. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-10] [REF-12]

    For assessment participation, the material distinction is between policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. This matters because the governing caution is that non-participation never treated as low attainment or ignored. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-13] [REF-14]

    In assessing assessment participation, authorities must determine combining unknown observations with the majority group biases both estimates and obscures the weakness. The public account remains incomplete unless it explains how where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners with disabilities and remote candidates. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-15] [REF-16]

    Public responsibility for assessment participation begins with urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. A proportionate conclusion must also recognise that targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. For assessment participation, policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-17] [REF-19]

    15

    Distribution of achievement

    Public responsibility for distribution of achievement begins with a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. For the learners concerned, the decisive consideration is whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of distribution of achievement lies in making unequal educational experience observable. Here, the relevant phenomenon is variation across the full score or proficiency distribution, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-10] [REF-12]

    The practical standard for distribution of achievement concerns the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. The institutional consequence follows from whether when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on percentiles and threshold shares beside a mean. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-13] [REF-14]

    Public responsibility for distribution of achievement begins with national averages should remain available as context, yet never as a substitute for the distribution. The institutional consequence follows from whether nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: uncertainty and scale properties stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding.[REF-15] [REF-16]

    Institutional action on distribution of achievement should be tested against attrition at any stage can produce an apparently complete indicator from a selective population. The public account remains incomplete unless it explains how field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether learners concentrated below a minimum proficiency threshold are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-17] [REF-19]

    distribution of achievement requires a decision about local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. The institutional consequence follows from whether communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. For distribution of achievement, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable.[REF-20] [REF-21]

    Part VI

    Gender and household resources

    16

    Gender parity and its limits

    gender parity and its limits cannot be judged without identifying the measure should preserve both the observed level and the distribution relevant to the claim. The institutional consequence follows from whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns differences between girls and boys in access, progression and learning. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-13] [REF-14]

    Evidence concerning gender parity and its limits should establish at minimum this includes population, geography, date, collection method, classification and known exclusions. The institutional consequence follows from whether a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is female-to-male ratios read with levels and absolute gaps. Its metadata should travel with every published value.[REF-15] [REF-16]

    gender parity and its limits cannot be judged without identifying a responsible commentary distinguishes observation, calculation and interpretation. The institutional consequence follows from whether it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that parity not confused with adequacy for either group.[REF-17] [REF-19]

    Comparative interpretation of gender parity and its limits depends upon it should involve statistical judgement, legal safeguards and knowledge of the affected community. A proportionate conclusion must also recognise that suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For girls in poor rural households and boys exposed to hazardous work, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-20] [REF-21]

    In assessing gender parity and its limits, authorities must determine where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This matters because this makes the indicator a means of scrutiny rather than a decorative measure of concern. For gender parity and its limits, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions.[REF-23] [REF-24]

    17

    Household wealth gradients

    household wealth gradients cannot be judged without identifying the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Household wealth gradients should be approached as a defined measurement problem. The substantive interest is education outcomes across relative household resource groups, observed for a population and period that are stated before calculation.[REF-15] [REF-16]

    A defensible account of household wealth gradients distinguishes it should require examination of definitions, timing, migration, duplication and non-response. A proportionate conclusion must also recognise that the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a documented asset or consumption classification within each setting. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-17] [REF-19]

    household wealth gradients requires a decision about a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. For the learners concerned, the decisive consideration is whether reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: wealth ranks not assumed equivalent across countries or time. The comparison should show the level for each group as well as any ratio or gap.[REF-20] [REF-21]

    household wealth gradients requires a decision about protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The public account remains incomplete unless it explains how the distributional review must deliberately include children in the poorest quintile and those near classification boundaries. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-23] [REF-24]

    Public responsibility for household wealth gradients begins with monitoring after action must preserve the original baseline and follow both reach and outcome. The resulting interpretation should show why improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For household wealth gradients, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-01] [REF-07]

    18

    Costs borne by households

    Institutional action on costs borne by households should be tested against a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The institutional consequence follows from whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For costs borne by households, the first requirement is conceptual clarity. The study seeks evidence on fees, materials, transport, clothing and foregone labour; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-17] [REF-19]

    costs borne by households requires a decision about where estimates are revised, both the reason and effect of revision should remain accessible. The resulting interpretation should show why measurement should use participation read against direct and indirect education costs. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-20] [REF-21]

    Institutional action on costs borne by households should be tested against disparity measures should not replace the underlying distributions. The resulting interpretation should show why a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that nominal fee abolition checked against remaining expenditure.[REF-23] [REF-24]

    costs borne by households cannot be judged without identifying the review should record non-response, unknown status and excluded locations separately. A contrary reading would overlook that combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for large families and households hit by economic crisis. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-01] [REF-07]

    costs borne by households requires a decision about targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. The public account remains incomplete unless it explains how accountability lies in changed opportunity, not in the favourable movement of an indicator alone. For costs borne by households, policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves.[REF-02] [REF-03]

    Part VII

    Place and service geography

    19

    Rural and urban residence

    Institutional action on rural and urban residence should be tested against the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of education outcomes by a declared settlement classification begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-20] [REF-21]

    Review of rural and urban residence is credible only where it explains the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. This matters because when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on residence linked to service availability and travel conditions. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-23] [REF-24]

    Public responsibility for rural and urban residence begins with if a conclusion changes under a reasonable specification, that instability is part of the finding. The evidence must therefore clarify how national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: national rural definitions preserved and comparison limitations stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-01] [REF-07]

    Evidence concerning rural and urban residence should establish missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. For the learners concerned, the decisive consideration is whether an adequate equity account asks whether remote villages, pastoral populations and peri-urban settlements are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-02] [REF-03]

    Review of rural and urban residence is credible only where it explains the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. This matters because for rural and urban residence, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required.[REF-04] [REF-06]

    20

    Subnational administrative disparity

    Review of subnational administrative disparity is credible only where it explains the measure should preserve both the observed level and the distribution relevant to the claim. The resulting interpretation should show why a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of subnational administrative disparity lies in making unequal educational experience observable. Here, the relevant phenomenon is variation between provinces, districts or comparable areas, not the administrative convenience of the available categories.[REF-23] [REF-24]

    subnational administrative disparity cannot be judged without identifying a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. This matters because operationally, the measure is area estimates with population size and precision. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence.[REF-01] [REF-07]

    subnational administrative disparity requires a decision about it does not assign cause from a cross-sectional difference. The evidence must therefore clarify how apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that administrative rankings not mistaken for causal explanations. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-02] [REF-03]

    Institutional action on subnational administrative disparity should be tested against disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. A contrary reading would overlook that this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For small districts and areas with incomplete reporting, a single group label may conceal important internal differences.[REF-04] [REF-06]

    The central question in subnational administrative disparity is subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. The public account remains incomplete unless it explains how where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. For subnational administrative disparity, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence.[REF-08] [REF-09]

    21

    Distance, isolation and transport

    Institutional action on distance, isolation and transport should be tested against the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that the indicator question concerns physical accessibility of the nearest appropriate service. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    The practical standard for distance, isolation and transport concerns a discrepancy is not resolved by selecting the more favourable source. The evidence must therefore clarify how it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is travel time, route safety and seasonal interruption rather than straight-line distance alone. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-02] [REF-03]

    The central question in distance, isolation and transport is a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. The evidence must therefore clarify how reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: household reports and facility mapping reconciled. The comparison should show the level for each group as well as any ratio or gap.[REF-04] [REF-06]

    The practical standard for distance, isolation and transport concerns if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The institutional consequence follows from whether confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners with limited mobility and communities cut off seasonally. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-08] [REF-09]

    The practical standard for distance, isolation and transport concerns monitoring after action must preserve the original baseline and follow both reach and outcome. The resulting interpretation should show why improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For distance, isolation and transport, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-10] [REF-12]

    Part VIII

    Disability, language and identity

    22

    Disability-sensitive education data

    For disability-sensitive education data, the material distinction is between the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how disability-sensitive education data should be approached as a defined measurement problem. The substantive interest is participation and learning by functional difficulty and support requirement, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-02] [REF-03]

    Evidence concerning disability-sensitive education data should establish source coverage must be tested before sources are combined. The public account remains incomplete unless it explains how school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use questions designed for comparable reporting without defining the child by diagnosis alone. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone.[REF-04] [REF-06]

    disability-sensitive education data requires a decision about a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. The public account remains incomplete unless it explains how percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that identification, environment and accommodation kept analytically distinct. Disparity measures should not replace the underlying distributions.[REF-08] [REF-09]

    Institutional action on disability-sensitive education data should be tested against where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. This matters because where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners whose impairments are not recorded by schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness.[REF-10] [REF-12]

    A defensible account of disability-sensitive education data distinguishes accountability lies in changed opportunity, not in the favourable movement of an indicator alone. A contrary reading would overlook that for disability-sensitive education data, policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold.[REF-13] [REF-14]

    23

    Language of home and instruction

    In assessing language of home and instruction, authorities must determine the study seeks evidence on alignment between learner language, teaching and assessment; it does not infer a learner's circumstances from a national or regional mean. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For language of home and instruction, the first requirement is conceptual clarity.[REF-04] [REF-06]

    The central question in language of home and instruction is analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. For the learners concerned, the decisive consideration is whether this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on language categories reflecting local use and instructional practice. The definition should be fixed for the comparison at hand and deviations recorded.[REF-08] [REF-09]

    language of home and instruction cannot be judged without identifying if a conclusion changes under a reasonable specification, that instability is part of the finding. A contrary reading would overlook that national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: small groups not erased through broad national labels. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-10] [REF-12]

    The central question in language of home and instruction is missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The public account remains incomplete unless it explains how an adequate equity account asks whether minority-language and multilingual learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-13] [REF-14]

    A defensible account of language of home and instruction distinguishes the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A contrary reading would overlook that for language of home and instruction, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required.[REF-15] [REF-16]

    24

    Ethnicity, indigeneity and protected identity

    Review of ethnicity, indigeneity and protected identity is credible only where it explains a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A proportionate conclusion must also recognise that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of disparity associated with historically excluded identity groups begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]

    The central question in ethnicity, indigeneity and protected identity is at minimum this includes population, geography, date, collection method, classification and known exclusions. The resulting interpretation should show why a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is lawful, voluntary and contextually meaningful classification. Its metadata should travel with every published value.[REF-10] [REF-12]

    In assessing ethnicity, indigeneity and protected identity, authorities must determine it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. The institutional consequence follows from whether it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that self-identification protected and non-response reported. A responsible commentary distinguishes observation, calculation and interpretation.[REF-13] [REF-14]

    A defensible account of ethnicity, indigeneity and protected identity distinguishes this balance is contextual rather than mechanical. For the learners concerned, the decisive consideration is whether it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For communities exposed to discrimination or forced assimilation, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-15] [REF-16]

    Review of ethnicity, indigeneity and protected identity is credible only where it explains subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. The resulting interpretation should show why where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. For ethnicity, indigeneity and protected identity, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence.[REF-17] [REF-19]

    Part IX

    Conflict, disaster and mobility

    25

    Education under conflict and insecurity

    Institutional action on education under conflict and insecurity should be tested against the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The resulting interpretation should show why the public value of education under conflict and insecurity lies in making unequal educational experience observable. Here, the relevant phenomenon is access, attendance and learning where violence alters service and movement, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-10] [REF-12]

    The practical standard for education under conflict and insecurity concerns the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. This matters because the preferred construction is location- and time-specific observation with explicit coverage gaps. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-13] [REF-14]

    The central question in education under conflict and insecurity is reference points therefore need substantive meaning. The resulting interpretation should show why where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: absence caused by insecurity distinguished from ordinary dropout. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-15] [REF-16]

    Institutional action on education under conflict and insecurity should be tested against coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. The resulting interpretation should show why if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners in insecure areas and host communities. These populations may be missing not only from good outcomes but from the denominator itself.[REF-17] [REF-19]

    Public responsibility for education under conflict and insecurity begins with improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The evidence must therefore clarify how the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For education under conflict and insecurity, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome.[REF-20] [REF-21]

    27

    Refugees, displaced persons and migrants

    For refugees, displaced persons and migrants, the material distinction is between the substantive interest is educational participation across changing legal and residential situations, observed for a population and period that are stated before calculation. A contrary reading would overlook that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Refugees, displaced persons and migrants should be approached as a defined measurement problem.[REF-15] [REF-16]

    Review of refugees, displaced persons and migrants is credible only where it explains this distinction is essential for participation and progression. The public account remains incomplete unless it explains how the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on status, origin, current location and service access recorded separately. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-17] [REF-19]

    Public responsibility for refugees, displaced persons and migrants begins with nor should a group estimate be read as a description of every member. The public account remains incomplete unless it explains how within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: mobility never converted into duplicate enrolment or unexplained disappearance. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-20] [REF-21]

    The central question in refugees, displaced persons and migrants is analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The resulting interpretation should show why missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether undocumented migrants, refugees and internally displaced learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-23] [REF-24]

    refugees, displaced persons and migrants cannot be judged without identifying the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. The institutional consequence follows from whether for refugees, displaced persons and migrants, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required.[REF-01] [REF-07]

    Part X

    School conditions and teachers

    28

    Teacher availability and distribution

    Review of teacher availability and distribution is credible only where it explains the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that for teacher availability and distribution, the first requirement is conceptual clarity. The study seeks evidence on access to competent teaching across schools and subjects; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-17] [REF-19]

    In assessing teacher availability and distribution, authorities must determine a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A contrary reading would overlook that a survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is teachers present and assigned relative to learner and curriculum need. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions.[REF-20] [REF-21]

    For teacher availability and distribution, the material distinction is between a responsible commentary distinguishes observation, calculation and interpretation. For the learners concerned, the decisive consideration is whether it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that payroll totals not substituted for classroom availability.[REF-23] [REF-24]

    Public responsibility for teacher availability and distribution begins with this balance is contextual rather than mechanical. The public account remains incomplete unless it explains how it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For schools serving poor, remote or displaced communities, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-01] [REF-07]

    Review of teacher availability and distribution is credible only where it explains a policy conclusion should be no broader than that evidence. The resulting interpretation should show why subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. For teacher availability and distribution, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified.[REF-02] [REF-03]

    29

    Class size, multi-grade teaching and time

    Review of class size, multi-grade teaching and time is credible only where it explains a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of the instructional conditions experienced by learners begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-20] [REF-21]

    Review of class size, multi-grade teaching and time is credible only where it explains if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. The resulting interpretation should show why administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is class organisation, scheduled time and delivered time considered together. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-23] [REF-24]

    Evidence concerning class size, multi-grade teaching and time should establish explanation requires evidence on institutions, resources, households and prior conditions. The resulting interpretation should show why interpretation follows this limitation: simple pupil-teacher ratios not treated as a complete quality measure. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-01] [REF-07]

    In assessing class size, multi-grade teaching and time, authorities must determine if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The evidence must therefore clarify how confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include early grades and mixed-age classes. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-02] [REF-03]

    In assessing class size, multi-grade teaching and time, authorities must determine improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The evidence must therefore clarify how the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For class size, multi-grade teaching and time, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome.[REF-04] [REF-06]

    30

    Materials, facilities and basic services

    A defensible account of materials, facilities and basic services distinguishes a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of materials, facilities and basic services lies in making unequal educational experience observable. Here, the relevant phenomenon is usable learning resources and safe, accessible school conditions, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-23] [REF-24]

    Review of materials, facilities and basic services is credible only where it explains reconciliation should record what each source can and cannot represent. For the learners concerned, the decisive consideration is whether where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use availability joined to condition, accessibility and regular use. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-01] [REF-07]

    A defensible account of materials, facilities and basic services distinguishes the comparison should identify the reference category but avoid presenting it as a natural norm. The resulting interpretation should show why policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that delivery counts checked against learner access. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-02] [REF-03]

    Institutional action on materials, facilities and basic services should be tested against their circumstances may alter access to enumeration, classification and the service being measured. The resulting interpretation should show why the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in temporary or damaged premises.[REF-04] [REF-06]

    Public responsibility for materials, facilities and basic services begins with urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. The public account remains incomplete unless it explains how targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. For materials, facilities and basic services, policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-08] [REF-09]

    Part XI

    Finance and distribution

    31

    Public spending by level and function

    A defensible account of public spending by level and function distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A contrary reading would overlook that the indicator question concerns resources assigned to education purposes across the system. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    public spending by level and function requires a decision about when several sources exist, consistency is evidence to consider, not proof that common error is absent. The evidence must therefore clarify how a defensible statistic would be based on expenditure classified by level, recurrent or capital use and responsible body. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-02] [REF-03]

    Review of public spending by level and function is credible only where it explains national averages should remain available as context, yet never as a substitute for the distribution. The public account remains incomplete unless it explains how nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: budgets, commitments and actual expenditure distinguished. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding.[REF-04] [REF-06]

    A defensible account of public spending by level and function distinguishes attrition at any stage can produce an apparently complete indicator from a selective population. The resulting interpretation should show why field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether basic education services under fiscal pressure are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-08] [REF-09]

    For public spending by level and function, the material distinction is between if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The resulting interpretation should show why the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. For public spending by level and function, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-10] [REF-12]

    32

    Incidence of education spending

    Institutional action on incidence of education spending should be tested against a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The resulting interpretation should show why the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Incidence of education spending should be approached as a defined measurement problem. The substantive interest is who benefits from publicly financed places and services, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-02] [REF-03]

    The practical standard for incidence of education spending concerns a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. The institutional consequence follows from whether operationally, the measure is unit resources combined with participation across population groups. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence.[REF-04] [REF-06]

    Public responsibility for incidence of education spending begins with apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. This matters because use of the indicator is bounded by the principle that benefit estimates not treated as household income. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-08] [REF-09]

    Institutional action on incidence of education spending should be tested against disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. A proportionate conclusion must also recognise that this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For groups excluded before public spending can reach them, a single group label may conceal important internal differences.[REF-10] [REF-12]

    incidence of education spending requires a decision about subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. A proportionate conclusion must also recognise that where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. For incidence of education spending, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence.[REF-13] [REF-14]

    33

    Protecting equity during fiscal constraint

    protecting equity during fiscal constraint cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The evidence must therefore clarify how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For protecting equity during fiscal constraint, the first requirement is conceptual clarity. The study seeks evidence on whether reductions or delays fall disproportionately on weaker services; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-04] [REF-06]

    In assessing protecting equity during fiscal constraint, authorities must determine numerator, denominator, reference date, unit and exclusions should appear together. The resulting interpretation should show why if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is dated finance and service indicators read together.[REF-08] [REF-09]

    The practical standard for protecting equity during fiscal constraint concerns where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. A contrary reading would overlook that statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: national totals tested against subnational allocation and household costs. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning.[REF-10] [REF-12]

    For protecting equity during fiscal constraint, the material distinction is between confidentiality is essential, especially where identity or status creates risk. The institutional consequence follows from whether protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include poor households and institutions with little financial reserve. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-13] [REF-14]

    Public responsibility for protecting equity during fiscal constraint begins with monitoring after action must preserve the original baseline and follow both reach and outcome. A contrary reading would overlook that improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For protecting equity during fiscal constraint, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-15] [REF-16]

    Part XII

    Data sources and measurement error

    34

    Administrative records

    Review of administrative records is credible only where it explains the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public account remains incomplete unless it explains how a distribution-sensitive account of regular learner, staff, facility and finance information begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-08] [REF-09]

    Review of administrative records is credible only where it explains where estimates are revised, both the reason and effect of revision should remain accessible. The resulting interpretation should show why measurement should use clear definitions, reporting coverage and revision history. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-10] [REF-12]

    The practical standard for administrative records concerns policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The institutional consequence follows from whether the governing caution is that non-reporting institutions kept visible in aggregates. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-13] [REF-14]

    Review of administrative records is credible only where it explains where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. The institutional consequence follows from whether particular scrutiny is required for small, private, non-formal and emergency providers. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-15] [REF-16]

    Comparative interpretation of administrative records depends upon accountability lies in changed opportunity, not in the favourable movement of an indicator alone. The evidence must therefore clarify how for administrative records, policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold.[REF-17] [REF-19]

    35

    Household surveys

    In assessing household surveys, authorities must determine here, the relevant phenomenon is population-based evidence beyond enrolled learners, not the administrative convenience of the available categories. The institutional consequence follows from whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of household surveys lies in making unequal educational experience observable.[REF-10] [REF-12]

    Public responsibility for household surveys begins with the definition should be fixed for the comparison at hand and deviations recorded. The public account remains incomplete unless it explains how analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on probability samples, weights and field dates documented.[REF-13] [REF-14]

    The practical standard for household surveys concerns results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. This matters because if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: sampling and non-response uncertainty carried into group comparisons.[REF-15] [REF-16]

    A defensible account of household surveys distinguishes analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. A contrary reading would overlook that missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether small minorities and mobile households are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-17] [REF-19]

    household surveys requires a decision about communities should be able to question both the category and the conclusion drawn from it. A proportionate conclusion must also recognise that if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. For household surveys, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-20] [REF-21]

    36

    Censuses and population frames

    Comparative interpretation of censuses and population frames depends upon the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns broad population coverage and small-area denominators. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-13] [REF-14]

    Review of censuses and population frames is credible only where it explains its metadata should travel with every published value. For the learners concerned, the decisive consideration is whether at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is enumeration date, usual residence and institutional coverage stated.[REF-15] [REF-16]

    The central question in censuses and population frames is apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. A proportionate conclusion must also recognise that use of the indicator is bounded by the principle that long intervals and under-enumeration acknowledged. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-17] [REF-19]

    A defensible account of censuses and population frames distinguishes the public report can still state that a disparity was examined, whether action is required and which body will monitor it. The evidence must therefore clarify how for homeless, displaced and geographically isolated people, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-20] [REF-21]

    A defensible account of censuses and population frames distinguishes where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. The resulting interpretation should show why this makes the indicator a means of scrutiny rather than a decorative measure of concern. For censuses and population frames, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions.[REF-23] [REF-24]

    Part XIII

    Disaggregation and intersection

    37

    Single-axis disaggregation

    Comparative interpretation of single-axis disaggregation depends upon the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because single-axis disaggregation should be approached as a defined measurement problem. The substantive interest is separate reporting by sex, wealth, residence or another characteristic, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-15] [REF-16]

    Evidence concerning single-axis disaggregation should establish it should require examination of definitions, timing, migration, duplication and non-response. A contrary reading would overlook that the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is levels, gaps and denominators shown for each category. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-17] [REF-19]

    Evidence concerning single-axis disaggregation should establish a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. The evidence must therefore clarify how reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: one axis not presented as a complete account of marginalisation. The comparison should show the level for each group as well as any ratio or gap.[REF-20] [REF-21]

    The central question in single-axis disaggregation is protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The resulting interpretation should show why the distributional review must deliberately include groups whose disadvantage lies on another unmeasured dimension. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-23] [REF-24]

    A defensible account of single-axis disaggregation distinguishes the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The institutional consequence follows from whether for single-axis disaggregation, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked.[REF-01] [REF-07]

    38

    Intersecting categories

    In assessing intersecting categories, authorities must determine the study seeks evidence on joint distributions such as sex by wealth and residence; it does not infer a learner's circumstances from a national or regional mean. For the learners concerned, the decisive consideration is whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For intersecting categories, the first requirement is conceptual clarity.[REF-17] [REF-19]

    Public responsibility for intersecting categories begins with counts reveal scale; rates permit comparison; neither is sufficient alone. For the learners concerned, the decisive consideration is whether source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use pre-specified combinations with sufficient observations. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-20] [REF-21]

    In assessing intersecting categories, authorities must determine policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The institutional consequence follows from whether the governing caution is that empty or unstable cells reported honestly. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-23] [REF-24]

    In assessing intersecting categories, authorities must determine their circumstances may alter access to enumeration, classification and the service being measured. The resulting interpretation should show why the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for poor rural girls, disabled learners in remote areas and displaced minorities.[REF-01] [REF-07]

    Comparative interpretation of intersecting categories depends upon urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. A proportionate conclusion must also recognise that targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. For intersecting categories, policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-02] [REF-03]

    39

    Small numbers, disclosure and reliability

    Evidence concerning small numbers, disclosure and reliability should establish the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of useful detail without unreliable estimates or identification begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-20] [REF-21]

    Review of small numbers, disclosure and reliability is credible only where it explains the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. The resulting interpretation should show why when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on suppression, aggregation or qualitative evidence chosen proportionately. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-23] [REF-24]

    A defensible account of small numbers, disclosure and reliability distinguishes nor should a group estimate be read as a description of every member. The institutional consequence follows from whether within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: confidentiality decisions separated from claims that no disparity exists. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-01] [REF-07]

    Evidence concerning small numbers, disclosure and reliability should establish attrition at any stage can produce an apparently complete indicator from a selective population. For the learners concerned, the decisive consideration is whether field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether small communities and learners with rare characteristics are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-02] [REF-03]

    Review of small numbers, disclosure and reliability is credible only where it explains local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. For the learners concerned, the decisive consideration is whether communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. For small numbers, disclosure and reliability, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable.[REF-04] [REF-06]

    Part XIV

    Comparison, uncertainty and change

    40

    Comparing unlike systems

    Comparative interpretation of comparing unlike systems depends upon the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether the public value of comparing unlike systems lies in making unequal educational experience observable. Here, the relevant phenomenon is cross-country patterns based on harmonised but bounded concepts, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-23] [REF-24]

    Review of comparing unlike systems is credible only where it explains a census figure should disclose enumeration rules. The institutional consequence follows from whether these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is metadata tests before numerical comparison. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-01] [REF-07]

    comparing unlike systems requires a decision about it does not assign cause from a cross-sectional difference. The institutional consequence follows from whether apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that differences in programme structure and classification remain visible. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-02] [REF-03]

    In assessing comparing unlike systems, authorities must determine it should involve statistical judgement, legal safeguards and knowledge of the affected community. The public account remains incomplete unless it explains how suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For countries with incomplete or rapidly changing systems, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-04] [REF-06]

    The central question in comparing unlike systems is where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. A proportionate conclusion must also recognise that this makes the indicator a means of scrutiny rather than a decorative measure of concern. For comparing unlike systems, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions.[REF-08] [REF-09]

    41

    Sampling error and other uncertainty

    Public responsibility for sampling error and other uncertainty begins with the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns the range of values reasonably compatible with the observations. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-01] [REF-07]

    Institutional action on sampling error and other uncertainty should be tested against a discrepancy is not resolved by selecting the more favourable source. This matters because it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is standard errors, design effects and data-quality qualifications. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-02] [REF-03]

    Institutional action on sampling error and other uncertainty should be tested against reference points therefore need substantive meaning. The public account remains incomplete unless it explains how where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: rank differences smaller than uncertainty not interpreted. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-04] [REF-06]

    For sampling error and other uncertainty, the material distinction is between protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. This matters because the distributional review must deliberately include small disaggregated populations. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-08] [REF-09]

    Institutional action on sampling error and other uncertainty should be tested against the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. A contrary reading would overlook that for sampling error and other uncertainty, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked.[REF-10] [REF-12]

    Part XV

    Responsible interpretation and action

    43

    Reading disparity without blaming learners

    reading disparity without blaming learners cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The public account remains incomplete unless it explains how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For reading disparity without blaming learners, the first requirement is conceptual clarity. The study seeks evidence on institutional and social conditions associated with unequal outcomes; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-04] [REF-06]

    In assessing reading disparity without blaming learners, authorities must determine when several sources exist, consistency is evidence to consider, not proof that common error is absent. A contrary reading would overlook that a defensible statistic would be based on descriptive findings separated from causal claims. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-08] [REF-09]

    Comparative interpretation of reading disparity without blaming learners depends upon within-group variation and unmeasured intersecting conditions remain material. This matters because the analytical rule is clear: group identity never treated as a mechanism by itself. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-10] [REF-12]

    Review of reading disparity without blaming learners is credible only where it explains field arrangements need relevant languages, accessible formats and safe participation. A proportionate conclusion must also recognise that analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether communities subject to stigma are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-13] [REF-14]

    A defensible account of reading disparity without blaming learners distinguishes local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. A proportionate conclusion must also recognise that communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. For reading disparity without blaming learners, a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable.[REF-15] [REF-16]

    44

    Turning evidence into equitable policy

    In assessing turning evidence into equitable policy, authorities must determine the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of a decision rule linking disparity to service, finance or legal responsibility begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]

    In assessing turning evidence into equitable policy, authorities must determine its metadata should travel with every published value. The evidence must therefore clarify how at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is baseline, intended reach, implementation evidence and review date.[REF-10] [REF-12]

    turning evidence into equitable policy requires a decision about it does not assign cause from a cross-sectional difference. A contrary reading would overlook that apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that targets accompanied by distributional safeguards. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-13] [REF-14]

    Evidence concerning turning evidence into equitable policy should establish disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. The evidence must therefore clarify how this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For learners farthest below a secured minimum, a single group label may conceal important internal differences.[REF-15] [REF-16]

    Evidence concerning turning evidence into equitable policy should establish this makes the indicator a means of scrutiny rather than a decorative measure of concern. The public account remains incomplete unless it explains how for turning evidence into equitable policy, public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation.[REF-17] [REF-19]

    45

    A public account beyond the average

    Public responsibility for a public account beyond the average begins with the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of a public account beyond the average lies in making unequal educational experience observable. Here, the relevant phenomenon is a concise national statement of level, distribution, missingness and remedy, not the administrative convenience of the available categories.[REF-10] [REF-12]

    In assessing a public account beyond the average, authorities must determine it should require examination of definitions, timing, migration, duplication and non-response. The resulting interpretation should show why the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is national totals presented beside selected group and place indicators. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-13] [REF-14]

    A defensible account of a public account beyond the average distinguishes statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. The public account remains incomplete unless it explains how explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: progress claims bounded by evidence coverage and unresolved gaps. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-15] [REF-16]

    Review of a public account beyond the average is credible only where it explains if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. This matters because confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include every learner otherwise hidden by a successful average. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-17] [REF-19]

    The practical standard for a public account beyond the average concerns monitoring after action must preserve the original baseline and follow both reach and outcome. The institutional consequence follows from whether improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. For a public account beyond the average, the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-20] [REF-21]

    Part XVI

    Applied distributional analysis

    46

    Composite measures and the loss of meaning

    Comparative interpretation of composite measures and the loss of meaning depends upon where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The institutional consequence follows from whether the analytical purpose is combining several dimensions into a summary measure. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises participation, completion, learning and school conditions, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-01] [REF-03]

    Evidence concerning composite measures and the loss of meaning should establish documentation should also state which observations are direct, which are estimated and which are unavailable. For the learners concerned, the decisive consideration is whether missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern weights, normalisation and substitution. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable.[REF-02] [REF-04]

    In assessing composite measures and the loss of meaning, authorities must determine a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. For the learners concerned, the decisive consideration is whether direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is a low value in one dimension being concealed by a high value in another. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold.[REF-05] [REF-06]

    For composite measures and the loss of meaning, the material distinction is between censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A proportionate conclusion must also recognise that a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. For composite measures and the loss of meaning, source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response.[REF-08] [REF-12]

    Evidence concerning composite measures and the loss of meaning should establish learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. The public account remains incomplete unless it explains how analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. For composite measures and the loss of meaning, equity review asks who can disappear during construction of the measure.[REF-13] [REF-16]

    Evidence concerning composite measures and the loss of meaning should establish otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The institutional consequence follows from whether the policy user is policy-makers deciding whether a summary aids or obscures resource allocation. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-17] [REF-19]

    The central question in composite measures and the loss of meaning is authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. The evidence must therefore clarify how feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. For composite measures and the loss of meaning, participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions.[REF-21] [REF-23]

    Institutional action on composite measures and the loss of meaning should be tested against the objective is not maximum numerical output. A contrary reading would overlook that it is a trustworthy connection between unequal educational experience, public responsibility and corrective action. For composite measures and the loss of meaning, final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence.[REF-23] [REF-24]

    47

    Decomposing an observed education gap

    Review of decomposing an observed education gap is credible only where it explains the evidence base comprises within-group and between-group differences, each of which describes a different feature of educational opportunity. This matters because a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is examining how a national disparity is distributed across places and population groups. That purpose should be written before the calculation because method follows the intended inference.[REF-02] [REF-04]

    Institutional action on decomposing an observed education gap should be tested against missing values should remain missing unless an explicit estimation method and its effect are shown. A contrary reading would overlook that the principal methodological questions concern population shares, outcome levels and overlapping membership. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable.[REF-05] [REF-06]

    The central question in decomposing an observed education gap is a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. For the learners concerned, the decisive consideration is whether direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is treating a descriptive decomposition as proof of cause. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold.[REF-08] [REF-12]

    The central question in decomposing an observed education gap is a record of disagreement should identify plausible sources and the decision consequence. For the learners concerned, the decisive consideration is whether if the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. For decomposing an observed education gap, source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity.[REF-13] [REF-16]

    The practical standard for decomposing an observed education gap concerns the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. The institutional consequence follows from whether for decomposing an observed education gap, equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity.[REF-17] [REF-19]

    decomposing an observed education gap requires a decision about credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. The institutional consequence follows from whether every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities locating where further enquiry and action are warranted. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-21] [REF-23]

    Comparative interpretation of decomposing an observed education gap depends upon the two forms answer different questions. A contrary reading would overlook that authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. For decomposing an observed education gap, participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier.[REF-23] [REF-24]

    Comparative interpretation of decomposing an observed education gap depends upon it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. The institutional consequence follows from whether changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. For decomposing an observed education gap, final reporting should give a clear institutional judgement.[REF-01] [REF-03]

    48

    Setting distribution-sensitive targets

    The practical standard for setting distribution-sensitive targets concerns where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. A contrary reading would overlook that the analytical purpose is expressing progress as improvement in secured minimums and unjustified gaps. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises the national level, the least-served group and the lower tail of the distribution, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-05] [REF-06]

    The practical standard for setting distribution-sensitive targets concerns documentation should also state which observations are direct, which are estimated and which are unavailable. For the learners concerned, the decisive consideration is whether missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern baseline stability, ambition and safeguards against exclusion. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable.[REF-08] [REF-12]

    Comparative interpretation of setting distribution-sensitive targets depends upon analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. The resulting interpretation should show why a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is meeting a mean target while abandoning those farthest behind. Avoiding it requires the underlying counts and distributions to remain visible beside any summary.[REF-13] [REF-16]

    A defensible account of setting distribution-sensitive targets distinguishes surveys can represent households beyond formal education, subject to sample size, field access and response. The evidence must therefore clarify how censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. For setting distribution-sensitive targets, source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-17] [REF-19]

    Evidence concerning setting distribution-sensitive targets should establish analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. The institutional consequence follows from whether small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. For setting distribution-sensitive targets, equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins.[REF-21] [REF-23]

    For setting distribution-sensitive targets, the material distinction is between credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. A contrary reading would overlook that every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is governments linking national commitments to subnational delivery. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-23] [REF-24]

    For setting distribution-sensitive targets, the material distinction is between people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. For the learners concerned, the decisive consideration is whether for setting distribution-sensitive targets, participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry.[REF-01] [REF-03]

    For setting distribution-sensitive targets, the material distinction is between it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. The institutional consequence follows from whether changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. For setting distribution-sensitive targets, final reporting should give a clear institutional judgement.[REF-02] [REF-04]

    49

    Linking learners to service geography

    A defensible account of linking learners to service geography distinguishes the evidence base comprises settlement populations, travel conditions, schools, teachers and programme levels, each of which describes a different feature of educational opportunity. For the learners concerned, the decisive consideration is whether a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is relating participation and learning to the location and capacity of education services. That purpose should be written before the calculation because method follows the intended inference.[REF-08] [REF-12]

    linking learners to service geography requires a decision about missing values should remain missing unless an explicit estimation method and its effect are shown. The evidence must therefore clarify how the principal methodological questions concern geographical scale, boundary effects and facility catchments. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable.[REF-13] [REF-16]

    Evidence concerning linking learners to service geography should establish analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. This matters because a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is assuming the nearest mapped institution is accessible or appropriate. Avoiding it requires the underlying counts and distributions to remain visible beside any summary.[REF-17] [REF-19]

    Comparative interpretation of linking learners to service geography depends upon a record of disagreement should identify plausible sources and the decision consequence. A contrary reading would overlook that if the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. For linking learners to service geography, source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity.[REF-21] [REF-23]

    In assessing linking learners to service geography, authorities must determine the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. The resulting interpretation should show why for linking learners to service geography, equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity.[REF-23] [REF-24]

    Comparative interpretation of linking learners to service geography depends upon otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. For the learners concerned, the decisive consideration is whether the policy user is planners choosing sites, transport support and teacher deployment. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-01] [REF-03]

    The central question in linking learners to service geography is the two forms answer different questions. The evidence must therefore clarify how authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. For linking learners to service geography, participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier.[REF-02] [REF-04]

    The practical standard for linking learners to service geography concerns it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. For the learners concerned, the decisive consideration is whether changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. For linking learners to service geography, final reporting should give a clear institutional judgement.[REF-05] [REF-06]

    50

    Reconciling conflicting sources

    In assessing reconciling conflicting sources, authorities must determine the national total supplies context, while the distribution tests whether the total is shared. For the learners concerned, the decisive consideration is whether where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is interpreting differences between administrative, survey and census estimates. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises coverage, timing, concepts and reporting incentives, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement.[REF-13] [REF-16]

    reconciling conflicting sources cannot be judged without identifying a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. A contrary reading would overlook that if the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a documented comparison of population and variable definitions. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion.[REF-17] [REF-19]

    Comparative interpretation of reconciling conflicting sources depends upon avoiding it requires the underlying counts and distributions to remain visible beside any summary. The evidence must therefore clarify how analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is averaging incompatible estimates into an apparently precise figure.[REF-21] [REF-23]

    Review of reconciling conflicting sources is credible only where it explains surveys can represent households beyond formal education, subject to sample size, field access and response. This matters because censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. For reconciling conflicting sources, source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-23] [REF-24]

    reconciling conflicting sources requires a decision about the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. The public account remains incomplete unless it explains how for reconciling conflicting sources, equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity.[REF-01] [REF-03]

    Evidence concerning reconciling conflicting sources should establish credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. The public account remains incomplete unless it explains how every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is statistical authorities issuing one bounded account with visible uncertainty. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-02] [REF-04]

    The practical standard for reconciling conflicting sources concerns authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. The public account remains incomplete unless it explains how feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. For reconciling conflicting sources, participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions.[REF-05] [REF-06]

    The central question in reconciling conflicting sources is it is a trustworthy connection between unequal educational experience, public responsibility and corrective action. The evidence must therefore clarify how for reconciling conflicting sources, final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output.[REF-08] [REF-12]

    51

    Monitoring marginalisation during severe disruption

    Evidence concerning monitoring marginalisation during severe disruption should establish a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. For the learners concerned, the decisive consideration is whether the national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is maintaining useful distributional evidence when populations and services move rapidly. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises rapid counts, restored administrative returns and household evidence, each of which describes a different feature of educational opportunity.[REF-17] [REF-19]

    Institutional action on monitoring marginalisation during severe disruption should be tested against they determine how strongly one component, place or population can influence the conclusion. The resulting interpretation should show why a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern dated estimates, revision practice and minimal essential classifications. Choices on these matters are not neutral presentation details.[REF-21] [REF-23]

    Review of monitoring marginalisation during severe disruption is credible only where it explains avoiding it requires the underlying counts and distributions to remain visible beside any summary. The evidence must therefore clarify how analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is using an unstable emergency denominator to assert durable improvement.[REF-23] [REF-24]

    A defensible account of monitoring marginalisation during severe disruption distinguishes administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. The resulting interpretation should show why surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. For monitoring marginalisation during severe disruption, source quality should be considered dimension by dimension.[REF-01] [REF-03]

    Review of monitoring marginalisation during severe disruption is credible only where it explains the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. The evidence must therefore clarify how for monitoring marginalisation during severe disruption, equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity.[REF-02] [REF-04]

    Comparative interpretation of monitoring marginalisation during severe disruption depends upon the certainty required depends upon the consequence. The resulting interpretation should show why credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities protecting access while rebuilding regular statistics. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-05] [REF-06]

    Evidence concerning monitoring marginalisation during severe disruption should establish feedback needs a recorded route into classification, service design or further enquiry. A proportionate conclusion must also recognise that people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. For monitoring marginalisation during severe disruption, participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-08] [REF-12]

    In assessing monitoring marginalisation during severe disruption, authorities must determine a concise account can carry considerable density when every figure retains its population and consequence. This matters because the objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. For monitoring marginalisation during severe disruption, final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress.[REF-13] [REF-16]

    52

    Communicating uncertainty without losing urgency

    Comparative interpretation of communicating uncertainty without losing urgency depends upon where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. The resulting interpretation should show why where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is explaining what is known strongly enough to justify action and what remains unresolved. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises point estimates, ranges, quality statements and missing populations, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-21] [REF-23]

    communicating uncertainty without losing urgency cannot be judged without identifying they determine how strongly one component, place or population can influence the conclusion. The evidence must therefore clarify how a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern plain institutional language joined to exact metadata. Choices on these matters are not neutral presentation details.[REF-23] [REF-24]

    Comparative interpretation of communicating uncertainty without losing urgency depends upon a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. This matters because direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is presenting caution as a reason for inaction or urgency as a reason for overstatement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold.[REF-01] [REF-03]

    Public responsibility for communicating uncertainty without losing urgency begins with a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A proportionate conclusion must also recognise that a record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. For communicating uncertainty without losing urgency, source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage.[REF-02] [REF-04]

    communicating uncertainty without losing urgency cannot be judged without identifying analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. The resulting interpretation should show why small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. For communicating uncertainty without losing urgency, equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins.[REF-05] [REF-06]

    The practical standard for communicating uncertainty without losing urgency concerns credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. This matters because every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is the public, affected communities and responsible decision-makers. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-08] [REF-12]

    The practical standard for communicating uncertainty without losing urgency concerns people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. The public account remains incomplete unless it explains how for communicating uncertainty without losing urgency, participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry.[REF-13] [REF-16]

    communicating uncertainty without losing urgency requires a decision about the objective is not maximum numerical output. For the learners concerned, the decisive consideration is whether it is a trustworthy connection between unequal educational experience, public responsibility and corrective action. For communicating uncertainty without losing urgency, final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence.[REF-17] [REF-19]

    53

    A national marginalisation profile

    Review of a national marginalisation profile is credible only where it explains where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. This matters because where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is assembling a concise recurring account beyond the national average. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises population, access, progression, learning, conditions, finance and unresolved evidence gaps, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-23] [REF-24]

    Evidence concerning a national marginalisation profile should establish if the headline conclusion changes materially, the range of results should be reported. A contrary reading would overlook that sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a stable core with context-specific distributions. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications.[REF-01] [REF-03]

    For a national marginalisation profile, the material distinction is between comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. This matters because no result should be described as equitable solely because one relative measure improved. The main interpretive danger is creating an encyclopaedia of indicators without decision priority. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-02] [REF-04]

    a national marginalisation profile requires a decision about surveys can represent households beyond formal education, subject to sample size, field access and response. The institutional consequence follows from whether censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. For a national marginalisation profile, source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-05] [REF-06]

    Evidence concerning a national marginalisation profile should establish these duties do not justify silence about a serious disparity. A contrary reading would overlook that the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. For a national marginalisation profile, equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability.[REF-08] [REF-12]

    a national marginalisation profile cannot be judged without identifying otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The public account remains incomplete unless it explains how the policy user is parliament, ministries, local authorities and communities reviewing educational equity. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-13] [REF-16]

    a national marginalisation profile requires a decision about people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. This matters because for a national marginalisation profile, participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry.[REF-17] [REF-19]

    Institutional action on a national marginalisation profile should be tested against it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. The evidence must therefore clarify how changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. For a national marginalisation profile, final reporting should give a clear institutional judgement.[REF-21] [REF-23]

    Part XVII

    Foundations for measuring completion

    54

    Completion is the operative international commitment

    The Millennium Declaration commits governments to ensure that, by 2015, children everywhere will be able to complete a full course of primary schooling. The Dakar Framework similarly calls for free and compulsory primary education of good quality and completion by all children, with particular attention to girls, children in difficult circumstances and members of ethnic minorities. The formulation contains four elements: a universal population, an opportunity to participate, progression through a defined course and completion of that course.[REF-25] [REF-26]

    Institutional action on completion is the operative international commitment should be tested against “Children everywhere” requires a population denominator and an account of children outside registered provision. The public account remains incomplete unless it explains how “Able to complete” concerns opportunity as well as observed outcome; current non-completion may reflect access barriers, cost, disability, insecurity, language, institutional supply or poor quality. “A full course” requires an identified beginning and end, a classification of the programme and a rule for curriculum or grade completion. “By 2015” requires a time series capable of distinguishing progress from statistical revision. Each element creates a measurement requirement.

    Review of completion is the operative international commitment is credible only where it explains a system can expand grade-one entry rapidly while losing many pupils before the end; its enrolment ratio may improve years before completion does. For the learners concerned, the decisive consideration is whether a high net enrolment ratio can coexist with late dropout, repetition or weak final-grade progression. An enrolment ratio addresses only part of this object. A net enrolment ratio estimates the participation of children in the official age group for a level. A gross enrolment ratio compares enrolment of all ages with the official-age population. Both can show the reach of current provision, subject to their definitions. Neither shows directly whether entrants remain through the final grade.

    completion is the operative international commitment cannot be judged without identifying these forms answer different questions. The resulting interpretation should show why completion is not merely a more advanced point on the same scale. It joins a stock measure of pupils present at a time with a terminal event or status. That change introduces questions of cohort, time and evidence. If the measure uses pupils in the last grade during one school year, it records a cross-section of final-grade flow. If it follows children who entered together, it records a cohort trajectory. If it asks young people whether they have completed primary education, it records attained status.

    The international commitment also concerns boys and girls alike. A total rate can meet a numerical target while a material group remains excluded if another group is overrepresented in the proxy numerator or if the rate exceeds 100 per cent. Equality is therefore part of the construct, not an optional analytical supplement. The same reasoning extends, where data and safeguards permit, to location, poverty, disability, language, minority status, displacement and other nationally relevant circumstances.[REF-33]

    55

    What constitutes a full course of primary schooling

    Primary education must be identified by programme content and position in the national structure, not by an institution's name alone. ISCED 1997 defines the primary level as normally beginning between ages five and seven, lasting four to six years and providing a sound basic education in reading, writing and mathematics together with an elementary understanding of other subjects. National structures depart from the modal pattern. Some integrate primary and lower-secondary grades; some divide primary provision between institutions; some have cycles of different duration.[REF-27]

    In assessing what constitutes a full course of primary schooling, authorities must determine a title such as “basic school”, “elementary school” or “first cycle” is not sufficient classification evidence. The public account remains incomplete unless it explains how for comparative measurement, the “full course” is the complete nationally defined programme mapped to ISCED level 1 for the reference year. The map is to identify the official entry age, nominal duration, first and final grade, graduation or promotion rule, and any programme variants included. It should record whether special, accelerated, non-formal or private programmes are counted and on what equivalence basis.

    Institutional action on what constitutes a full course of primary schooling should be tested against a measure based on final-grade entry will be available earlier and more widely than one based on certified completion, but it will include some pupils who later leave or fail. The resulting interpretation should show why a measure based on transition to secondary education will omit completers who do not continue and may be distorted by limited secondary places. Completion can be evidenced in several ways. A pupil may be recorded as a new entrant to the last grade; may remain enrolled at a prescribed census date; may finish the final school year; may satisfy an attendance rule; may pass an examination; may receive a certificate; or may enter the next level. These events do not occur at the same time and are not interchangeable.

    Evidence concerning what constitutes a full course of primary schooling should establish the public meaning of the terms should follow their measurement basis. A contrary reading would overlook that the national measurement standard should designate one event for each published indicator. Where final-grade new entrants are used as the international proxy, the publication should say so. The word “graduates” must be reserved for a result based on the national graduation rule. “Completed primary” in household data must be tied to a question and coding instruction that identifies the last grade or credential.

    what constitutes a full course of primary schooling cannot be judged without identifying joining the old and new values without a break would imply a continuous measure of an unchanged object. The resulting interpretation should show why the correct publication identifies the two structures, reports them separately during overlap, explains the denominator rule and establishes the first year in which the revised definition can support a comparable trend. A reform of cycle length requires explicit treatment. Suppose a system changes primary education from six to seven grades. During transition, pupils may complete under both structures, the official final-grade entrance age changes, and cohort exposure differs.

    56

    Four meanings commonly carried by the word completion

    In assessing four meanings commonly carried by the word completion, authorities must determine the first is the **gross intake ratio to the last grade**, the present international proxy. A proportionate conclusion must also recognise that it relates new entrants to the last grade, regardless of age, to the population of the official age for entry to that grade. The second is a **cohort completion probability**, the share of a defined group of entrants who eventually complete. The third is an **age-group attainment rate**, the share of persons in an age band who have completed primary education. The fourth is a **graduation ratio**, newly certified completers relative to a specified population. A fifth measure, survival to a particular grade, is often discussed with completion but represents progression to that grade under a reconstructed or observed cohort. The term “completion rate” is used for at least four distinct measures.

    four meanings commonly carried by the word completion requires a decision about its denominator can be obtained from population estimates. The institutional consequence follows from whether it is comparatively timely and can be disaggregated by sex where both components permit. Its principal weakness is that numerator members are not necessarily the same age as denominator members and may not complete the final year. It is a gross rate in the statistical sense, even where the shorter title “primary completion rate” is used. The proxy is attractive because ministries commonly collect enrolment and repeater data by grade each year.

    A cohort rate corresponds more closely to the probability that an entrant completes, but a true observed cohort requires pupil identifiers or carefully linked records over several years. In their absence, reconstructed cohort methods infer progression from two consecutive years of grade enrolment and repetition. The 2004 World Development Indicators notes the simplifying assumptions: pupils who leave do not return, promotion, repetition and dropout rates remain constant over the period, and the rates apply equally to all pupils in a grade. Migration, re-entry, grade skipping and transfer can violate those assumptions.[REF-28]

    A defensible account of four meanings commonly carried by the word completion distinguishes it can be disaggregated by household wealth or parental characteristics, but it depends on sample size, accurate reporting and stable classification. This matters because an age-group attainment rate addresses the population result directly. For example, a household survey may estimate the proportion of 15- to 19-year-olds who completed primary school. The chosen age band should allow most late entrants and repeaters sufficient time to finish while remaining close enough to current conditions to be policy relevant. The measure includes past experience rather than only the current school flow.

    The practical standard for four meanings commonly carried by the word completion concerns passing a selective examination can measure achievement under a national standard, but using it as the sole completion event may conflate completion of the course with success in selection for the next stage. The public account remains incomplete unless it explains how a graduation ratio is institutionally precise where certification is universal and records are complete. It becomes less comparable where completion does not require an external credential, where examination rules differ, or where certificates are withheld for fees or administrative reasons.

    The practical standard for four meanings commonly carried by the word completion concerns a high graduation ratio with weak transition may identify limited secondary capacity rather than primary failure. The public account remains incomplete unless it explains how these measures form a family and do not compete for a single label. Agreement among them strengthens confidence. Divergence provides information. A high final-grade intake proxy with a lower young-person attainment rate may indicate recent expansion, an inflated denominator relationship, final-year loss or differences in definition. A high cohort survival estimate with a low population attainment rate may indicate exclusion at entry.

    57

    Completion as event, status and opportunity

    The practical standard for completion as event, status and opportunity concerns confusion among these dimensions leads to claims that the data cannot support. The resulting interpretation should show why an event measure records that completion occurred during a period. A status measure records that a person had completed by the date of observation. An opportunity measure examines whether the system provided the conditions necessary to enter, continue and finish.

    Annual administrative data are strongest on events and current states within registered provision. They can count final-grade pupils, repeaters, examination candidates and graduates. A population survey is strongest on status and participation among individuals, including those outside school. Evidence of opportunity requires a wider account: whether a place was available, direct and indirect costs were manageable, the route was physically and socially accessible, instruction was acceptable, and provision could adapt to a child's circumstances without lowering the entitlement. The right-to-education framework of availability, accessibility, acceptability and adaptability provides an appropriate organising discipline.[REF-34]

    In assessing completion as event, status and opportunity, authorities must determine an increase following abolition of fees may represent genuine access and flow improvement, but it may also bring large classes and high repetition if teacher and classroom supply do not expand. This matters because the completion measure records the eventual terminal flow; it does not reveal the mechanism without supporting evidence. The target measure should remain an outcome indicator. It must not be redefined to include every condition of opportunity. Nevertheless, policy interpretation should ask whether a low or unequal result originates in supply, demand, progression, quality or measurement.

    Institutional action on completion as event, status and opportunity should be tested against a published rate must be accompanied by a bounded diagnostic profile capable of locating failure. This matters because the profile need not contain every education statistic. It should contain measures that connect the population to entry, entry to progression, progression to completion and completion to educational substance. This distinction protects both accountability and diagnosis. Governments can be held to a clear result without implying that the result alone explains performance.

    Part XVIII

    The terminal-grade proxy and its limits

    58

    The numerator: new entrants to the final grade

    A defensible account of the numerator: new entrants to the final grade distinguishes it is an administratively feasible approximation. A contrary reading would overlook that its accuracy depends on three records: the identification of the final primary grade, a complete count of pupils enrolled in that grade, and a valid distinction between new entrants and repeaters. The numerator used in the current proxy is generally final-grade enrolment less repeaters. This subtraction is intended to count pupils entering the last grade for the first time.

    Public responsibility for the numerator: new entrants to the final grade begins with if the national system includes a mainstream six-year cycle and an authorised accelerated programme, the numerator cannot simply combine their terminal grades without an equivalence and double-counting rule. The resulting interpretation should show why if private or community provision follows a different grade structure, coverage must be declared. The international classification map should precede aggregation. The first record can fail during structural reform or where programme variants end at different grades.

    Public responsibility for the numerator: new entrants to the final grade begins with systems are to state whether the count is opening enrolment, enrolment on a census date, average attendance or another measure. This matters because changes in the census date can alter the numerator independently of educational progress. The enrolment count should use a specified reference date. Registers opened at the start of the year may include pupils who never attend or leave before the census. A late census may exclude early leavers but may also miss mobile pupils.

    the numerator: new entrants to the final grade requires a decision about a pupil who repeats only part of a programme may not fit the administrative category. A contrary reading would overlook that instructions and validation should address these cases. Repeater status should refer to the same grade in the previous school year, not simply to age or examination failure. A pupil returning after an interruption may be recorded as a repeater in one system and a re-entrant in another. A pupil who transfers and re-registers can be counted twice where records are institution-based.

    Comparative interpretation of the numerator: new entrants to the final grade depends upon age-grade distributions, previous-year enrolment and promotion records can provide plausibility checks. This matters because large year-to-year changes in the repeater share require explanation before the completion proxy is interpreted. The subtraction method can produce biased results where repeater reporting is weak. Schools may have an incentive to under-report repetition if it is treated as failure, or to over-report it if resources follow pupil-years.

    For the numerator: new entrants to the final grade, the material distinction is between some will leave, repeat or fail before the end of the year. The public account remains incomplete unless it explains how the size of that difference depends on the national programme and cannot be assumed constant across countries. Where end-of-year records are available, systems should publish the relationship between first-time final-grade entrants, pupils completing the year, examination candidates, successful candidates and certified graduates. The international proxy can then retain its established form while its local meaning becomes more precise. Final-grade entrants are not final-grade completers.

    59

    The denominator: population at the official age for entering the final grade

    A defensible account of the denominator: population at the official age for entering the final grade distinguishes the completion proxy compares first-time final-grade entrants of all ages with the estimated population aged eleven. A proportionate conclusion must also recognise that this is a reference-population device; it is not a count of the same children as the numerator. The denominator is derived from the official entry age and nominal duration of primary education. If entry is officially at age six and the programme lasts six years without repetition, the official age for entering the final grade is ordinarily eleven.

    A defensible account of the denominator: population at the official age for entering the final grade distinguishes small errors matter because the denominator is a narrow age group. A proportionate conclusion must also recognise that a revised census may change both the current rate and the historical series even though school records do not change. Population estimates between censuses depend on fertility, mortality and migration assumptions. In settings with weak civil registration, an old census, rapid displacement or substantial migration, the single-year age estimate can be uncertain.

    Public responsibility for the denominator: population at the official age for entering the final grade begins with late entry and repetition make final-grade pupils older than the denominator age. The institutional consequence follows from whether early entry produces the opposite effect. A temporary enrolment campaign can bring several ages into grade one together; years later, the final-grade numerator may represent an unusually large mixed-age flow. The gross proxy correctly records that flow but must not be interpreted as the percentage of children of the official age who completed. The official-age rule may be a poor description of actual age progression.

    The central question in the denominator: population at the official age for entering the final grade is where the national statistical office and an international source use different series, both values and the reason for the difference must be documented. For the learners concerned, the decisive consideration is whether denominator governance should include an identified official population series, the date and source of its base census, projection method, revision policy and geographic boundary. Education authorities are not to select among population estimates according to which produces the preferred rate.

    Evidence concerning the denominator: population at the official age for entering the final grade should establish the numerator is ordinarily assigned by residence if the denominator is residence-based; if school location is used, the mismatch must be explicit. This matters because small-area results may require pooled years or counts alongside rates. For subnational rates, migration and small populations can produce pronounced volatility. Boarding schools and cross-boundary attendance can place pupils in one jurisdiction while their household population is counted in another.

    60

    Why the rate can exceed 100 per cent

    A result above 100 per cent is mathematically possible because the numerator and denominator are not a matched cohort. The numerator includes new entrants to the final grade regardless of age, while the denominator contains one official-age population. Late entrants, repeaters earlier in the cycle and early entrants can cause the terminal flow to exceed the size of that age group. Population underestimation, different reference dates, migration and incomplete repeater subtraction can add to the effect.[REF-28]

    why the rate can exceed 100 per cent cannot be judged without identifying it indicates that the gross proxy has reached a boundary at which its percentage form is easily misunderstood. The evidence must therefore clarify how the underlying counts remain informative. Publication should retain the observed rate, mark it as a gross measure, and provide a diagnostic note. The result is not evidence that more than all children complete. Nor does it necessarily demonstrate error.

    A defensible account of why the rate can exceed 100 per cent distinguishes a system moving from 118 to 104 may be resolving an age backlog while maintaining broad access; truncation would show no change. The evidence must therefore clarify how a comparison that treats 103 and 100 as equivalent may conceal different age and data conditions. For target assessment, a supplementary rule may classify values near or above 100 as consistent with universal terminal flow, but the source value should remain visible and must not substitute for population attainment evidence. Truncating the value at 100 removes information and can produce a false plateau.

    Institutional action on why the rate can exceed 100 per cent should be tested against the investigation may confirm that the result reflects a temporary cohort bulge, identify a correctable data defect, or show that the denominator needs revision. The resulting interpretation should show why an investigation is to examine the age distribution of final-grade entrants, repetition by grade, the date and coverage of enrolment, the population projection, recent census revisions, migration, changes in cycle structure and the inclusion of programmes outside the main system. Where age is missing for many pupils, the limitation itself must be reported.

    61

    The upper-bound property

    World Development Indicators 2004 states that data limitations preclude adjustment for pupils who leave during the final year and that the proxy should therefore be treated as an upper-bound estimate of actual primary completion. This qualification has direct implications for reporting. The measure is more accurately described as entry to the last grade than as verified completion.[REF-28]

    In assessing the upper-bound property, authorities must determine in a system with near-universal completion of the final year, it may be small. The evidence must therefore clarify how in a system with a high-stakes examination, seasonal withdrawal, conflict or substantial final-grade repetition, it may be material. A universal adjustment factor would create false comparability. National systems are to measure the final-grade flow where records permit. The size of the upper-bound difference is not known from the proxy itself.

    For the upper-bound property, the material distinction is between it then identifies pupils present at the reference census, pupils completing the prescribed instructional period, pupils eligible for any terminal assessment, assessment participants, successful candidates, certified completers and entrants to the next level. For the learners concerned, the decisive consideration is whether not every system uses each stage. The point is to identify the events that exist and prevent movement between them from disappearing inside one label. The preferred flow account begins with first-time entrants to the final grade.

    For the upper-bound property, the material distinction is between the statistical report should distinguish loss during the final year from failure to obtain a separate credential. For the learners concerned, the decisive consideration is whether a lower observed number of graduates does not automatically mean the proxy is invalid. Certification may cover only part of provision or may require a selective examination that is not itself the definition of completing the course. The relevant comparison depends on national rules.

    62

    A controlled formula and publication notation

    Review of a controlled formula and publication notation is credible only where it explains for year *t*, let (E_{g,t}) denote all pupils enrolled in the nationally defined last primary grade *g* at the official census date, (R_{g,t}) denote pupils repeating that grade, and (P_{a,t}) denote the population of official age *a* for entering grade *g*. For the learners concerned, the decisive consideration is whether the proxy is:. The conclusion for intersecting disadvantage in participation and completion indicators should connect the finding to the competent authority, the affected learners and the evidence required at review.

    **PCRP(t) = 100 × [E(g,t) − R(g,t)] / P(a,t)**

    a controlled formula and publication notation cannot be judged without identifying its purpose is to keep the method visible while the measures are compared. The evidence must therefore clarify how the notation PCRP is used in this report to distinguish the proxy from observed cohort or population-attainment completion measures. It is not proposed as a replacement for established international terminology.

    Institutional action on a controlled formula and publication notation should be tested against if school data refer to the school year ending in 2003 and population data to mid-2002, the publication must follow the stated international reference convention and identify it. The resulting interpretation should show why if the numerator covers the de facto population while the denominator uses usual residence, the mismatch requires examination. If the final grade changes within the year, the applicable structure must be recorded. The formula requires alignment of level, year, territory and population concept.

    Institutional action on a controlled formula and publication notation should be tested against the denominator must also be published. The resulting interpretation should show why a rate without components cannot be independently checked and makes revision difficult to understand. Sex-specific rates require sex-specific numerators and denominators; a sex distribution of the numerator alone is not a rate. The numerator is to be published as a count, together with the final-grade enrolment and repeater count from which it is derived.

    Part XIX

    Distribution of completion across populations

    63

    Sex-specific measurement

    For sex-specific measurement, the material distinction is between reporting the percentage of final-grade entrants who are girls is insufficient because it does not relate either group to its population. The institutional consequence follows from whether the international commitment explicitly includes girls and boys. A total completion proxy must therefore be accompanied by sex-specific numerators, denominators and rates.

    The central question in sex-specific measurement is the difference in percentage points and the ratio between rates can both be informative, but neither should replace the component values. The institutional consequence follows from whether the sex-specific proxy inherits all properties of the total measure. A male or female rate can exceed 100 per cent; age patterns and repetition may differ; population estimates by single year and sex can be unstable in small populations.

    Public responsibility for sex-specific measurement begins with a narrow final-grade gap can coexist with large exclusion at entry if only a selected group of girls enters and those who enter progress strongly. This matters because conversely, parity at entry can be lost through work, safety, early marriage, pregnancy, cost or institutional practice. Analysis should locate the stage at which disparity emerges. Entry, attendance, repetition, survival and final-year completion may show different patterns.

    The 2005 deadline for eliminating disparity in primary and secondary education gives this analysis immediate significance. Available global evidence indicates that the target will not be achieved everywhere. The measurement response must be precise about the level, indicator and direction of disparity and should avoid treating national parity as evidence that all girls and boys complete.[REF-28] [REF-29]

    64

    Geography and territorial inequality

    For geography and territorial inequality, the material distinction is between geographic reporting should use administrative areas meaningful for policy while preserving comparability over time. A proportionate conclusion must also recognise that boundary changes must be recorded and, where possible, historical data recalculated to stable units. National averages can conceal regions with low school supply, weak progression or unreliable data.

    Public responsibility for geography and territorial inequality begins with where pupil residence is not available, school-location results must be labelled as service-area measures. The institutional consequence follows from whether maps can be useful but must not imply precision beyond the data and should show missing or low-coverage areas distinctly rather than colouring them as average. Rates need residence-consistent numerators and denominators.

    The practical standard for geography and territorial inequality concerns geographic inequality is not a diagnosis by itself. A contrary reading would overlook that counts matter alongside rates. A small district with a very low rate and a populous district with a moderately low rate require different scale responses. Travel distance, population density, language and conflict exposure may explain different mechanisms.

    65

    Household economic circumstance

    household economic circumstance requires a decision about household surveys can estimate completion or attainment by wealth proxies, consumption or another economic classification. This matters because these measures identify distribution but require careful interpretation because wealth indices are relative to the survey population and may not be comparable across countries or survey rounds. Administrative school records rarely contain a reliable household-income measure.

    Public responsibility for household economic circumstance begins with reporting attainment across successive ages can distinguish delayed from absent completion while recognising that delay itself imposes costs and increases risk of exit. The evidence must therefore clarify how the age band should allow for delayed completion. Poorer children may enter later and repeat more often; measuring at the official completion age can exaggerate permanent inequality by counting delay as final non-completion.

    Economic analysis should connect results to direct fees, informal charges, uniforms, materials, transport, meals and the opportunity cost of children's work. It must not infer household preference from non-attendance without examining these constraints. The United Nations Millennium Project places education within a wider investment framework that includes health, nutrition, water, transport and public capacity; completion policy should reflect those interdependencies.[REF-32]

    66

    Disability and functional exclusion

    For disability and functional exclusion, the material distinction is between absence of a disaggregated rate cannot be interpreted as absence of inequality. The institutional consequence follows from whether the measurement plan should first examine whether the population is visible and whether classifications support rather than obstruct educational access. Many administrative systems do not identify children with disabilities consistently, and household questions may rely on narrow medical labels.

    Evidence concerning disability and functional exclusion should establish household data may provide a broader population estimate, but question design and stigma affect reporting. A proportionate conclusion must also recognise that rates based only on children already placed in special provision will exclude those outside education and those included in mainstream schools without recorded status. School records must not require a diagnosis as the sole basis for identifying support needs.

    Review of disability and functional exclusion is credible only where it explains qualitative and service-access evidence may be necessary until a valid rate is available. This matters because confidentiality is particularly important in small communities. Publication should state the definition, data source and coverage and should avoid comparing categories that differ materially between systems.

    67

    Language, ethnicity and minority status

    Public responsibility for language, ethnicity and minority status begins with national law, consultation and statistical confidentiality should govern the approach. A proportionate conclusion must also recognise that language of instruction and minority status can affect entry, comprehension, repetition and completion. Data may be necessary to identify inequality, but collection can create risk where identity is contested or where small groups can be recognised.

    Comparative interpretation of language, ethnicity and minority status depends upon household self-identification can provide a different perspective. The public account remains incomplete unless it explains how neither must be converted automatically into an explanation of performance. Curriculum access, teacher language capacity, location, poverty and discrimination may interact. Administrative language categories may reflect the language of the school rather than the learner.

    language, ethnicity and minority status cannot be judged without identifying an aggregate described as universal must not silently exclude populations for whom no classification or return exists. For the learners concerned, the decisive consideration is whether where a category cannot be collected safely, geographic, school-language or qualitative evidence may provide a limited alternative. The limitation is to be explicit.

    68

    Rural, remote and mobile populations

    Evidence concerning rural, remote and mobile populations should establish children may enter late, attend intermittently or move between sites. The resulting interpretation should show why school-census definitions designed for stable institutions can lose both pupils and provision. Distance, seasonal movement and low population density can make a conventional graded school difficult to provide.

    The practical standard for rural, remote and mobile populations concerns flexible calendars or mobile provision may support completion, but reduced instructional entitlement must not be accepted merely because delivery differs. The institutional consequence follows from whether records must follow learners where possible and reconcile transfers. The completion framework should identify valid alternative routes and their equivalence to the national primary course.

    rural, remote and mobile populations requires a decision about coverage limitations must be stated beside the result. The institutional consequence follows from whether household surveys need frames capable of representing mobile and remote populations. Where this is not possible, targeted enumeration or administrative mapping should supplement the national estimate.

    69

    Intersection and small populations

    intersection and small populations requires a decision about analysis should examine intersections where there is a policy reason and sufficient evidence. The resulting interpretation should show why single-dimension averages can conceal compounded exclusion. Rural girls from a minority-language group may face a different completion pattern from the national female or rural average.

    In assessing intersection and small populations, authorities must determine suppression must not erase the existence of the group; a note can state that evidence was collected but cannot be published at that level. This matters because small cells create statistical and confidentiality risks. A highly disaggregated table can produce unstable rates and identify individuals. Minimum reporting thresholds, pooling and controlled access to microdata must be established before publication.

    Institutional action on intersection and small populations should be tested against a pre-specified analytical plan reduces the risk of selecting only striking results. The public account remains incomplete unless it explains how intersectional analysis must be selective. The purpose is to identify a mechanism or distribution requiring action, not to generate every possible cross-tabulation.

    70

    Equity rules for target assessment

    Comparative interpretation of equity rules for target assessment depends upon the standard is a universal opportunity, not an average offset. The evidence must therefore clarify how a national completion target must not be declared achieved solely because the aggregate proxy reaches 100 per cent. At minimum, the result must be examined by sex and territory and reconciled with population attainment. Material groups or areas with no data must be identified.

    In assessing equity rules for target assessment, authorities must determine sampling error, population estimates and gross-rate properties make such a rule unsound. For the learners concerned, the decisive consideration is whether it requires evidence that no material disparity remains unexamined and that differences are within a stated tolerance or subject to a funded corrective plan. This does not require every subgroup estimate to equal precisely 100.

    In assessing equity rules for target assessment, authorities must determine this form is more informative than a binary label. The institutional consequence follows from whether the target-assessment statement should separate three judgements: whether terminal flow is consistent with universality; whether population evidence supports near-universal attainment; and whether material distributional gaps remain. A country can meet one judgement and not the others.

    References

    1. REF-01

      Education for All Global Monitoring Report Team. Reaching the Marginalized — EFA Global Monitoring Report 2010. 2010.

      Principal contemporaneous analysis of intersecting disadvantage and education marginalisation.

      https://unesdoc.unesco.org/ark:/48223/pf0000186606
    2. REF-02

      UNESCO Institute for Statistics. Global Education Digest 2010: Comparing Education Statistics Across the World. 2010.

      Comparative education statistics, definitions and limitations.

      https://uis.unesco.org/sites/default/files/documents/global-education-digest-2010-comparing-education-statistics-across-the-world-en.pdf
    3. REF-03

      UNESCO Institute for Statistics. Education Indicators: Technical Guidelines. 2009.

      Definitions and interpretation of participation, progression, completion and resource indicators.

      https://uis.unesco.org/sites/default/files/documents/education-indicators-technical-guidelines-en_0.pdf
    4. REF-04

      United Nations. The Millennium Development Goals Report 2010. 2010.

      Global and regional monitoring of primary education, gender, poverty and related development conditions.

      https://www.un.org/millenniumgoals/pdf/MDG%20Report%202010%20En%20r15%20-low%20res%2020100615%20-.pdf
    5. REF-05

      United Nations Development Programme. Human Development Report 2010: The Real Wealth of Nations — Pathways to Human Development. 2010.

      Distribution-sensitive human development concepts and evidence available before the cut-off.

      https://hdr.undp.org/content/human-development-report-2010
    6. REF-06

      UNICEF. Progress for Children: Achieving the MDGs with Equity, Number 9. 2010.

      Equity-focused child indicators and comparison between population groups.

      https://www.unicef.org/reports/progress-children-no-9
    7. REF-07

      World Education Forum. The Dakar Framework for Action: Education for All — Meeting Our Collective Commitments. 2000.

      Commitments to equitable access, quality, measurable outcomes and accountable national planning.

      https://unesdoc.unesco.org/ark:/48223/pf0000121147
    8. REF-08

      United Nations General Assembly. Convention on the Rights of the Child. 1989.

      Rights concerning non-discrimination, identity, education and development.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-child
    9. REF-09

      United Nations Committee on Economic, Social and Cultural Rights. General Comment No. 13: The Right to Education. 1999.

      Interpretation of availability, accessibility, acceptability and adaptability.

      https://www.refworld.org/legal/general/cescr/1999/en/37937
    10. REF-10

      United Nations General Assembly. Convention on the Rights of Persons with Disabilities. 2006.

      Non-discrimination, accessibility, inclusive education and disability data safeguards.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-persons-disabilities
    11. REF-11

      United Nations. Guiding Principles on Internal Displacement. 1998.

      Principles relevant to protection, documentation, education and non-discrimination of displaced persons.

      https://www.ohchr.org/en/special-procedures/sr-internally-displaced-persons/international-standards
    12. REF-12

      UNESCO and UNICEF. A Human Rights-Based Approach to Education for All. 2007.

      Rights-based planning, equality, participation, accountability and education quality.

      https://unesdoc.unesco.org/ark:/48223/pf0000154861
    13. REF-13

      Education for All Global Monitoring Report Team. Overcoming Inequality: Why Governance Matters — EFA Global Monitoring Report 2009. 2008.

      Governance, finance and unequal educational opportunity.

      https://unesdoc.unesco.org/ark:/48223/pf0000177683
    14. REF-14

      World Bank. World Development Report 2006: Equity and Development. 2005.

      Concepts of unequal opportunity, institutions and equitable public action.

      https://documents.worldbank.org/curated/en/435331468127174418/pdf/322040World0Development0Report02006.pdf
    15. REF-15

      World Bank. Safeguarding Education During Economic Crisis. 2009.

      Risks to budgets, households, participation and long-term human development during economic crisis.

      https://documents1.worldbank.org/curated/en/489131468340200911/pdf/485120WP0Avert10Box338912B01PUBLIC1.pdf
    16. REF-16

      Organisation for Economic Co-operation and Development. Education at a Glance 2010: OECD Indicators. 2010.

      Comparative participation, progression, expenditure and outcomes evidence with system-level metadata.

      https://doi.org/10.1787/eag-2010-en
    17. REF-17

      European Commission. Europe 2020: A Strategy for Smart, Sustainable and Inclusive Growth. 2010.

      Contemporaneous European policy context for education, inclusion, employment and headline indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC2020
    18. REF-18

      European Commission. Youth on the Move: An Initiative to Unleash the Potential of Young People to Achieve Smart, Sustainable and Inclusive Growth in the European Union. 2010.

      European education, mobility, attainment and youth inclusion policy context.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0477
    19. REF-19

      European Commission. A Renewed Commitment to Social Europe: Reinforcing the Open Method of Coordination for Social Protection and Social Inclusion. 2008.

      Social inclusion monitoring, common objectives and context-sensitive indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52008DC0418
    20. REF-20

      United Nations General Assembly. Resolution 64/250: Assistance to Haiti in the Aftermath of the Recent Earthquake. 2010.

      Contemporaneous recognition of humanitarian and reconstruction needs and national leadership.

      https://undocs.org/A/RES/64/250
    21. REF-21

      United Nations Office for the Coordination of Humanitarian Affairs. Haiti Revised Humanitarian Appeal. 2010.

      Displacement, service disruption and humanitarian education context.

      https://reliefweb.int/report/haiti/haiti-revised-humanitarian-appeal-2010
    22. REF-22

      European Commission. European Union Response to the Earthquake in Haiti. 2010.

      European humanitarian and recovery support, coordination and Haitian ownership.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0056
    23. REF-23

      United Nations Economic and Social Council. Principles and Recommendations for Population and Housing Censuses, Revision 2. 2008.

      Official principles for population coverage, definitions, classifications and data quality.

      https://unstats.un.org/unsd/demographic-social/Standards-and-Methods/files/Principles_and_Recommendations/Population-and-Housing-Censuses/Series_M67Rev2-E.pdf
    24. REF-24

      United Nations Statistics Division. Designing Household Survey Samples: Practical Guidelines. 2005.

      Sample design, estimation, precision and non-response guidance.

      https://unstats.un.org/unsd/demographic/sources/surveys/Handbook23June05.pdf
    25. REF-25

      World Education Forum. The Dakar Framework for Action: Education for All — Meeting Our Collective Commitments. 2000. ED-2000/WS/27.

      The universal-primary-education commitment, the quality requirement and the obligation to monitor participation and learning with attention to inequality.

      https://unesdoc.unesco.org/ark:/48223/pf0000121147
    26. REF-26

      United Nations General Assembly. United Nations Millennium Declaration. 2000. A/RES/55/2.

      The undertaking that children everywhere should be able to complete a full course of primary schooling by 2015.

      https://digitallibrary.un.org/record/422015
    27. REF-27

      United Nations Educational, Scientific and Cultural Organization. International Standard Classification of Education: ISCED 1997. 1997. BPE-98/WS/1.

      International classification principles for identifying primary education and managing differences in programme duration and structure.

      https://uis.unesco.org/sites/default/files/documents/international-standard-classification-of-education-1997-en_0.pdf
    28. REF-28

      World Bank. World Development Indicators 2004. 2004. Education efficiency, table 2.12 and methodological notes.

      Contemporaneous definition, proxy construction, limitations and international data for primary completion and related progression indicators.

      https://documents.worldbank.org/curated/en/517231468762935046/pdf/289690PAPER0WDI02004.pdf
    29. REF-29

      International Monetary Fund and World Bank. Global Monitoring Report 2004: Policies and Actions for Achieving the Millennium Development Goals and Related Outcomes. 2004.

      Regional progress estimates, observed and required rates of improvement, data-availability constraints and policy interpretation of primary completion.

      https://www.imf.org/external/np/pdr/gmr/eng/2004/041604.pdf
    30. REF-32

      United Nations Millennium Project. Investing in Development: A Practical Plan to Achieve the Millennium Development Goals. 2005.

      The development-planning context available before the cut-off, including the need to connect education targets with finance, health, nutrition, water, transport and administrative capacity.

      https://digitallibrary.un.org/record/564468
    31. REF-33

      United Nations. Convention on the Rights of the Child. 1989. A/RES/44/25. Articles 2, 28 and 29.

      The rights basis for non-discrimination, access, regular attendance, reduction of dropout and the substantive aims of education.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-child
    32. REF-34

      United Nations Committee on Economic, Social and Cultural Rights. General Comment No. 13: The Right to Education. 1999. E/C.12/1999/10.

      Interpretive framework of availability, accessibility, acceptability and adaptability relevant to the meaning of a completed primary education.

      https://docstore.ohchr.org/SelfServices/FilesHandler.ashx?enc=4slQ6QSmlBEDzFEovLCuW1AVC1NkPsgUedPlF1vfPMJb2C7KRvOaewo5P54LEjsHEpeN01Dr2U7Zw%2BK5%2F3WZKUclog1%2BBe3TC8O6zK4NNSgWPJ0yZhtq61OlL