Thematic Research Report

ICEQC-R-2014-09 — Positioning Quality Education within the Post-2015 Development Framework

A global policy synthesis on educational substance, equity, measurement and public responsibility

Publication date
Research category
Industry Policy and Regional Regulatory Analysis
Report archetype
Policy and Regulatory Synthesis
Geographic scope
Global
Evidence cut-off date
Responsible body
ICEQC Research and Policy Directorate
International Council for Education Quality Certification

ICEQC-R-2014-09

Positioning Quality Education within the Post-2015 Development Framework

A global policy synthesis on educational substance, equity, measurement and public responsibility

Publication date
Evidence cut-off date
Publication type
Thematic Research Report
Authoritative language
EN

Publication record

This is the controlled English edition. Evidence and institutional status are stated as at the evidence cut-off date.

Executive summary

Quality education should occupy a constitutive, rather than supplementary, position in the development framework being considered for the period beyond 2015. Access, equity, learning, teaching conditions and recognised progression describe connected aspects of one public obligation. A framework that treats access as the goal and quality as an optional qualification risks counting participation without examining its educational substance. A framework that treats measured learning as the whole goal risks excluding children who never enter, do not attend regularly or are absent from assessment. The policy task is to express a concise global commitment without separating conditions that jointly determine whether the right to education is secured.

This report proposes that quality be positioned through three linked requirements. First, the framework should state the learner-facing entitlement: meaningful participation, an enabling environment, competent teaching, a relevant curriculum, learning and recognised completion. Second, it should preserve distribution, so that progress cannot be inferred solely from a national mean. Third, it should link every principal claim to a competent public authority, resources and later review. The resulting architecture remains capable of international comparison while leaving national and regional bodies room to sequence action in accordance with diagnosed barriers.

No later intergovernmental settlement is presumed. The analysis is confined to commitments, official evidence and institutional positions available by 25 September 2014. Its purpose is to clarify the standards against which possible formulations could be judged, not to present a subsequent decision as though it had already occurred.

The formulation of an education goal beyond 2015 presents a double public obligation. The goal must be universal enough to express an equal claim to education and precise enough to expose unequal progress; it must also permit regions and countries to act upon different starting points without converting those differences into weaker entitlements. A common global framework should therefore establish the educational substance owed to every learner, while national and regional plans identify the barriers, institutional capacities and financing measures most relevant to securing it. Universality concerns the right and its minimum substance. Differentiation concerns the route, pace, sequencing and additional effort required to realise that right.

Regional consultation should not be reduced to the selection of separate numerical targets. Its first value is diagnostic: it can reveal where rapid expansion has outpaced instructional capacity, where conflict and displacement have damaged continuity, where demographic change has altered demand, where language and location structure exclusion, and where post-basic transitions or adult literacy require greater attention. Its second value is institutional: regional bodies can sustain comparable definitions, peer review, mobility arrangements and cross-border learning without displacing national responsibility. A global goal gains credibility when these differentiated conditions are stated openly within a common rights-based account.

This study treats access, participation, progression, completion, learning, teaching conditions, inclusion and public accountability as connected elements of educational quality. It rejects a choice between access and learning, or between a global measure and local relevance. Entry without sustained participation is incomplete; completion without meaningful learning is an inadequate claim; learning measures that omit excluded children are distributionally unsound; and aggregate progress without a visible account of disadvantaged populations cannot establish universal advance. The framework below identifies the evidentiary disciplines required to hold these elements together.

In assessing executive summary, authorities must determine the evidentiary task is therefore not simply to publish more categories. For the learners concerned, the decisive consideration is whether it is to construct distributions whose populations, definitions, uncertainty and limits permit fair public judgement. National education averages describe the centre of a distribution while frequently concealing the conditions of learners farthest from secured educational opportunity. A rising enrolment rate can coexist with persistent exclusion in a remote district, a poor household group or a population displaced by conflict or disaster. Gender parity at national level can coexist with disadvantage for poor rural girls and poor urban boys. A stable completion rate may hide increasing delay, repetition or withdrawal among groups whose educational records and population denominators are weakest.

Evidence concerning executive summary should establish official education statistics and development monitoring also provide methods for participation, progression, completion, resources and population comparison. A contrary reading would overlook that this study brings those strands together. It examines who enters the denominator, which education state is measured, how groups and places are classified, what missingness does to an estimate and how an observed disparity should enter a decision. The 2010 Education for All Global Monitoring Report placed marginalisation at the centre of global education analysis and showed the importance of overlapping disadvantage.

executive summary requires a decision about wealth, sex, residence, language, disability, displacement and prior education can intersect, while sample sizes and confidentiality place legitimate limits on publication. The institutional consequence follows from whether when a population cannot be estimated reliably, uncertainty should be reported rather than converted into absence. A distribution-sensitive indicator begins with an entitlement and a population, not with an available column in a register. Its numerator and denominator must refer to compatible people, places and periods. Counts, levels, gaps and ratios serve different purposes and should be read together. Disaggregation by one characteristic is often necessary but rarely sufficient.

The central question in executive summary is agreement strengthens a conclusion only after definitions are reconciled; disagreement is a reason to investigate population, timing and measurement. A contrary reading would overlook that the study distinguishes administrative records, household surveys and censuses. School records offer regular institutional detail but usually omit children outside provision. Surveys can represent households beyond schools, subject to their frames, response and sampling uncertainty. Censuses provide broad population and small-area evidence at longer intervals, while coverage and classification remain material. No source is complete for every question.

The continuing economic crisis and the disruption following the Haiti earthquake illustrate why averages can become least reliable when public decisions are most urgent. Budget totals may conceal local service contraction and higher household costs. Displacement can change both numerator and denominator, create duplicate or missing records and weaken comparisons with an earlier population. These conditions require dated estimates, visible revisions and a distinction between rapid operational counts and statistics suitable for final comparison. The report uses only evidence available by 20 March 2014 and makes no claim about subsequent recovery.

executive summary cannot be judged without identifying it becomes misleading when it is allowed to stand for a distribution that the evidence is capable of revealing. A proportionate conclusion must also recognise that the central conclusion is that a credible national account should present level, distribution, missingness and response together. It should identify which learners remain unrepresented; report the strength of each comparison; distinguish statistical association from cause; and name the authority responsible for remedy. An average remains valuable context.

Key findings

    Scope and method

    In assessing scope and method, authorities must determine it does not create a universal ranking, prescribe protected classifications without regard to national law, or claim that one set of disaggregations is appropriate in every setting. This matters because this comparative indicator study addresses access, participation, progression, completion, learning, conditions and finance. Its purpose is to guide the construction and interpretation of national and subnational distributions.

    Public responsibility for scope and method begins with the decision test asks what conclusion and action the evidence can support. The public account remains incomplete unless it explains how rights and Education for All commitments supply the public-interest frame; official statistical sources supply definitions and quality principles. The method applies five tests. The population test asks who could appear in the numerator and denominator. The concept test asks what education state is actually observed. The source test examines coverage, error and revision. The distribution test considers groups, places and intersections.

    scope and method cannot be judged without identifying material quantitative claims derived from a cited source remain subject to that source's definitions; illustrative measurement propositions do not assert unnamed national results. The public account remains incomplete unless it explains how comparisons are treated as bounded. A harmonised term does not remove differences in programme structure, school age, household classification or collection practice. The report distinguishes direct observation, estimate and interpretation.

    All evidence and institutional status are stated as at 25 September 2014. The study does not anticipate later intergovernmental decisions or outcomes.

    Part I

    Purpose and interpretive frame

    1

    What national averages conceal

    A defensible account of what national averages conceal distinguishes the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns national participation or attainment. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-01] [REF-07]

    In assessing what national averages conceal, authorities must determine numerator, denominator, reference date, unit and exclusions should appear together. A proportionate conclusion must also recognise that if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is the weighted experience of the population included in the denominator.[REF-02] [REF-03]

    Evidence concerning what national averages conceal should establish explanation requires evidence on institutions, resources, households and prior conditions. The resulting interpretation should show why interpretation follows this limitation: distribution within countries, not a league table between them. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-04] [REF-06]

    For what national averages conceal, the material distinction is between coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. A proportionate conclusion must also recognise that if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include remote rural learners, urban informal settlements and displaced populations. These populations may be missing not only from good outcomes but from the denominator itself.[REF-08] [REF-09]

    A defensible account of what national averages conceal distinguishes an indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. The institutional consequence follows from whether it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date.[REF-10] [REF-12]

    2

    Marginalisation as accumulated disadvantage

    A defensible account of marginalisation as accumulated disadvantage distinguishes a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Marginalisation as accumulated disadvantage should be approached as a defined measurement problem. The substantive interest is distance from a socially secured educational minimum, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-02] [REF-03]

    The central question in marginalisation as accumulated disadvantage is where estimates are revised, both the reason and effect of revision should remain accessible. The institutional consequence follows from whether measurement should use the joint effect of exclusion, weak provision and adverse social conditions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-04] [REF-06]

    Comparative interpretation of marginalisation as accumulated disadvantage depends upon disparity measures should not replace the underlying distributions. The resulting interpretation should show why a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that multiple indicators read together over the learner's course.[REF-08] [REF-09]

    A defensible account of marginalisation as accumulated disadvantage distinguishes the review should record non-response, unknown status and excluded locations separately. A contrary reading would overlook that combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for children facing poverty, gender disadvantage, disability or minority status. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-10] [REF-12]

    marginalisation as accumulated disadvantage requires a decision about accountability lies in changed opportunity, not in the favourable movement of an indicator alone. The evidence must therefore clarify how policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold.[REF-13] [REF-14]

    3

    From monitoring commitment to decision

    The practical standard for from monitoring commitment to decision concerns the study seeks evidence on evidence capable of changing resource or service decisions; it does not infer a learner's circumstances from a national or regional mean. The resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For from monitoring commitment to decision, the first requirement is conceptual clarity.[REF-04] [REF-06]

    Public responsibility for from monitoring commitment to decision begins with the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. The evidence must therefore clarify how when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on a link between observed disparity, responsible body and remedy. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-08] [REF-09]

    For from monitoring commitment to decision, the material distinction is between within-group variation and unmeasured intersecting conditions remain material. This matters because the analytical rule is clear: indicators selected for action rather than visibility. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-10] [REF-12]

    A defensible account of from monitoring commitment to decision distinguishes analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The public account remains incomplete unless it explains how missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether groups absent from routine plans and budget classifications are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-13] [REF-14]

    The practical standard for from monitoring commitment to decision concerns if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The institutional consequence follows from whether the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-15] [REF-16]

    Part II

    Population and denominator

    4

    Defining the population entitled to education

    Evidence concerning defining the population entitled to education should establish the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of resident and temporarily absent learners within the relevant age or programme group begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]

    Institutional action on defining the population entitled to education should be tested against a survey estimate should disclose weights and uncertainty. A contrary reading would overlook that a census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is a denominator consistent with the right, level and reference period under review. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions.[REF-10] [REF-12]

    The practical standard for defining the population entitled to education concerns a responsible commentary distinguishes observation, calculation and interpretation. The evidence must therefore clarify how it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that census, survey and administrative estimates reconciled openly.[REF-13] [REF-14]

    In assessing defining the population entitled to education, authorities must determine it should involve statistical judgement, legal safeguards and knowledge of the affected community. The evidence must therefore clarify how suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For unregistered residents, migrants and displaced children, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-15] [REF-16]

    defining the population entitled to education requires a decision about a policy conclusion should be no broader than that evidence. The public account remains incomplete unless it explains how subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified.[REF-17] [REF-19]

    5

    Age, grade and programme populations

    Institutional action on age, grade and programme populations should be tested against here, the relevant phenomenon is age-specific, grade-specific and programme-specific participation, not the administrative convenience of the available categories. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of age, grade and programme populations lies in making unequal educational experience observable.[REF-10] [REF-12]

    A defensible account of age, grade and programme populations distinguishes a discrepancy is not resolved by selecting the more favourable source. The resulting interpretation should show why it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is separate denominators for different educational questions. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-13] [REF-14]

    The central question in age, grade and programme populations is explanation requires evidence on institutions, resources, households and prior conditions. A proportionate conclusion must also recognise that interpretation follows this limitation: exact age, school age and enrolled population never substituted silently. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-15] [REF-16]

    In assessing age, grade and programme populations, authorities must determine these populations may be missing not only from good outcomes but from the denominator itself. A contrary reading would overlook that coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include over-age entrants and learners repeating grades.[REF-17] [REF-19]

    The practical standard for age, grade and programme populations concerns the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The evidence must therefore clarify how the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked.[REF-20] [REF-21]

    6

    Population movement and disrupted residence

    A defensible account of population movement and disrupted residence distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether the indicator question concerns education status amid migration, displacement and return. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-13] [REF-14]

    The central question in population movement and disrupted residence is school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. This matters because reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use dated location and residence rules with sensitivity to mobility. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined.[REF-15] [REF-16]

    population movement and disrupted residence requires a decision about the comparison should identify the reference category but avoid presenting it as a natural norm. This matters because policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that origin, current location and service responsibility distinguished. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-17] [REF-19]

    population movement and disrupted residence requires a decision about combining unknown observations with the majority group biases both estimates and obscures the weakness. This matters because where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for families affected by conflict, disaster and seasonal movement. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-20] [REF-21]

    Review of population movement and disrupted residence is credible only where it explains the evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. The resulting interpretation should show why these are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer.[REF-23] [REF-24]

    Part III

    Access and participation

    7

    Entry at the official starting age

    Comparative interpretation of entry at the official starting age depends upon the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public account remains incomplete unless it explains how entry at the official starting age should be approached as a defined measurement problem. The substantive interest is timely admission to the first grade of primary education, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-15] [REF-16]

    A defensible account of entry at the official starting age distinguishes analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. The evidence must therefore clarify how this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on new entrants of official age relative to the corresponding population. The definition should be fixed for the comparison at hand and deviations recorded.[REF-17] [REF-19]

    Public responsibility for entry at the official starting age begins with national averages should remain available as context, yet never as a substitute for the distribution. This matters because nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: late entry examined alongside non-entry. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding.[REF-20] [REF-21]

    For entry at the official starting age, the material distinction is between missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. This matters because an adequate equity account asks whether children facing fees, distance, disability or documentation barriers are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-23] [REF-24]

    A defensible account of entry at the official starting age distinguishes local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. For the learners concerned, the decisive consideration is whether communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable.[REF-01] [REF-07]

    8

    Attendance beyond enrolment

    The central question in attendance beyond enrolment is the measure should preserve both the observed level and the distribution relevant to the claim. For the learners concerned, the decisive consideration is whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For attendance beyond enrolment, the first requirement is conceptual clarity. The study seeks evidence on actual participation during a stated recent period; it does not infer a learner's circumstances from a national or regional mean.[REF-17] [REF-19]

    attendance beyond enrolment cannot be judged without identifying a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. For the learners concerned, the decisive consideration is whether operationally, the measure is presence measured independently of registration status. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence.[REF-20] [REF-21]

    The practical standard for attendance beyond enrolment concerns it does not assign cause from a cross-sectional difference. The institutional consequence follows from whether apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that frequency, season and reasons for absence retained. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-23] [REF-24]

    Public responsibility for attendance beyond enrolment begins with it should involve statistical judgement, legal safeguards and knowledge of the affected community. The institutional consequence follows from whether suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For working children, carers and learners affected by illness, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-01] [REF-07]

    For attendance beyond enrolment, the material distinction is between readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. The institutional consequence follows from whether for attendance beyond enrolment, the material distinction is between a policy conclusion should be no broader than that evidence. For the learners concerned, the decisive consideration is whether subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary.[REF-02] [REF-03]

    9

    Out-of-school status

    out-of-school status cannot be judged without identifying the measure should preserve both the observed level and the distribution relevant to the claim. For the learners concerned, the decisive consideration is whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of children in the relevant age group not participating at the defined level begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-20] [REF-21]

    For out-of-school status, the material distinction is between administrative records should be reconciled with population-based evidence where their coverage differs. The public account remains incomplete unless it explains how a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a transparent residual from compatible population and participation concepts. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-23] [REF-24]

    Institutional action on out-of-school status should be tested against a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. This matters because reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: never-enrolled and formerly enrolled children separated. The comparison should show the level for each group as well as any ratio or gap.[REF-01] [REF-07]

    out-of-school status cannot be judged without identifying confidentiality is essential, especially where identity or status creates risk. A contrary reading would overlook that protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children invisible to school registers. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-02] [REF-03]

    Comparative interpretation of out-of-school status depends upon the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. This matters because the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked.[REF-04] [REF-06]

    Part IV

    Progression and completion

    10

    Repetition and grade survival

    Evidence concerning repetition and grade survival should establish the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether the public value of repetition and grade survival lies in making unequal educational experience observable. Here, the relevant phenomenon is movement through grades without avoidable delay or exit, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-23] [REF-24]

    A defensible account of repetition and grade survival distinguishes source coverage must be tested before sources are combined. The public account remains incomplete unless it explains how school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use cohort or reconstructed-cohort evidence with explicit assumptions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone.[REF-01] [REF-07]

    repetition and grade survival cannot be judged without identifying a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. This matters because percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that repeaters distinguished from re-entrants and transfers. Disparity measures should not replace the underlying distributions.[REF-02] [REF-03]

    Evidence concerning repetition and grade survival should establish combining unknown observations with the majority group biases both estimates and obscures the weakness. The institutional consequence follows from whether where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in overcrowded or intermittently operating schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-04] [REF-06]

    repetition and grade survival cannot be judged without identifying the evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. This matters because these are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer.[REF-08] [REF-09]

    11

    Transition between education levels

    In assessing transition between education levels, authorities must determine the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether the indicator question concerns entry to the next level after completion of the preceding one. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    The practical standard for transition between education levels concerns this distinction is essential for participation and progression. For the learners concerned, the decisive consideration is whether the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on matched completion and new-entry populations over coherent periods. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-02] [REF-03]

    A defensible account of transition between education levels distinguishes results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. The evidence must therefore clarify how if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: capacity constraints distinguished from learner attainment.[REF-04] [REF-06]

    The practical standard for transition between education levels concerns missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. A contrary reading would overlook that an adequate equity account asks whether rural learners and those unable to relocate are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-08] [REF-09]

    The central question in transition between education levels is communities should be able to question both the category and the conclusion drawn from it. The resulting interpretation should show why if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-10] [REF-12]

    12

    Completion and educational entitlement

    Review of completion and educational entitlement is credible only where it explains the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that completion and educational entitlement should be approached as a defined measurement problem. The substantive interest is finishing the final grade or meeting recognised programme requirements, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-02] [REF-03]

    The central question in completion and educational entitlement is at minimum this includes population, geography, date, collection method, classification and known exclusions. A contrary reading would overlook that a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is completion defined separately from sitting or passing an examination. Its metadata should travel with every published value.[REF-04] [REF-06]

    Comparative interpretation of completion and educational entitlement depends upon a responsible commentary distinguishes observation, calculation and interpretation. The resulting interpretation should show why it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that late completion and alternative pathways reported.[REF-08] [REF-09]

    Institutional action on completion and educational entitlement should be tested against disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. A contrary reading would overlook that this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For over-age learners and people returning after interruption, a single group label may conceal important internal differences.[REF-10] [REF-12]

    The central question in completion and educational entitlement is readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. The institutional consequence follows from whether a policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary.[REF-13] [REF-14]

    Part V

    Learning and assessment

    13

    Minimum learning outcomes

    The central question in minimum learning outcomes is a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A proportionate conclusion must also recognise that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For minimum learning outcomes, the first requirement is conceptual clarity. The study seeks evidence on demonstrated knowledge or skill against a declared domain; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-04] [REF-06]

    The central question in minimum learning outcomes is it should require examination of definitions, timing, migration, duplication and non-response. For the learners concerned, the decisive consideration is whether the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is assessment evidence whose population and conditions are known. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-08] [REF-09]

    The practical standard for minimum learning outcomes concerns statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. This matters because explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: results interpreted with opportunity to learn and participation. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-10] [REF-12]

    Evidence concerning minimum learning outcomes should establish these populations may be missing not only from good outcomes but from the denominator itself. This matters because coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. The central question in minimum learning outcomes is if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. This matters because confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners taught in an unfamiliar language.[REF-13] [REF-14]

    For minimum learning outcomes, the material distinction is between improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The public account remains incomplete unless it explains how the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome.[REF-15] [REF-16]

    14

    Assessment participation

    For assessment participation, the material distinction is between the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of who was eligible, present, absent and excluded from testing begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]

    For assessment participation, the material distinction is between this formulation requires the reporting body to preserve the population base and the observation period beside the result. The resulting interpretation should show why counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use a participation profile accompanying every result distribution.[REF-10] [REF-12]

    Evidence concerning assessment participation should establish the comparison should identify the reference category but avoid presenting it as a natural norm. The institutional consequence follows from whether policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that non-participation never treated as low attainment or ignored. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-13] [REF-14]

    assessment participation requires a decision about combining unknown observations with the majority group biases both estimates and obscures the weakness. For the learners concerned, the decisive consideration is whether where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners with disabilities and remote candidates. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-15] [REF-16]

    Review of assessment participation is credible only where it explains the evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. The evidence must therefore clarify how these are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer.[REF-17] [REF-19]

    15

    Distribution of achievement

    distribution of achievement requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because the public value of distribution of achievement lies in making unequal educational experience observable. Here, the relevant phenomenon is variation across the full score or proficiency distribution, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-10] [REF-12]

    distribution of achievement cannot be judged without identifying the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. The evidence must therefore clarify how when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on percentiles and threshold shares beside a mean. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-13] [REF-14]

    A defensible account of distribution of achievement distinguishes nor should a group estimate be read as a description of every member. This matters because within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: uncertainty and scale properties stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-15] [REF-16]

    Comparative interpretation of distribution of achievement depends upon analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. For the learners concerned, the decisive consideration is whether missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether learners concentrated below a minimum proficiency threshold are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-17] [REF-19]

    distribution of achievement cannot be judged without identifying communities should be able to question both the category and the conclusion drawn from it. The resulting interpretation should show why if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-20] [REF-21]

    Part VI

    Gender and household resources

    16

    Gender parity and its limits

    For gender parity and its limits, the material distinction is between its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. A contrary reading would overlook that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns differences between girls and boys in access, progression and learning.[REF-13] [REF-14]

    The practical standard for gender parity and its limits concerns a census figure should disclose enumeration rules. A proportionate conclusion must also recognise that these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is female-to-male ratios read with levels and absolute gaps. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-15] [REF-16]

    In assessing gender parity and its limits, authorities must determine apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. For the learners concerned, the decisive consideration is whether use of the indicator is bounded by the principle that parity not confused with adequacy for either group. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-17] [REF-19]

    The central question in gender parity and its limits is this balance is contextual rather than mechanical. This matters because it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For girls in poor rural households and boys exposed to hazardous work, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-20] [REF-21]

    Public responsibility for gender parity and its limits begins with readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. The public account remains incomplete unless it explains how a policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary.[REF-23] [REF-24]

    17

    Household wealth gradients

    household wealth gradients cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Household wealth gradients should be approached as a defined measurement problem. The substantive interest is education outcomes across relative household resource groups, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-15] [REF-16]

    Comparative interpretation of household wealth gradients depends upon administrative records should be reconciled with population-based evidence where their coverage differs. The public account remains incomplete unless it explains how a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a documented asset or consumption classification within each setting. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-17] [REF-19]

    Review of household wealth gradients is credible only where it explains a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. The public account remains incomplete unless it explains how reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: wealth ranks not assumed equivalent across countries or time. The comparison should show the level for each group as well as any ratio or gap.[REF-20] [REF-21]

    Evidence concerning household wealth gradients should establish protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. A contrary reading would overlook that the distributional review must deliberately include children in the poorest quintile and those near classification boundaries. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-23] [REF-24]

    household wealth gradients requires a decision about monitoring after action must preserve the original baseline and follow both reach and outcome. For the learners concerned, the decisive consideration is whether improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-01] [REF-07]

    18

    Costs borne by households

    Review of costs borne by households is credible only where it explains a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For costs borne by households, the first requirement is conceptual clarity. The study seeks evidence on fees, materials, transport, clothing and foregone labour; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-17] [REF-19]

    costs borne by households requires a decision about counts reveal scale; rates permit comparison; neither is sufficient alone. The evidence must therefore clarify how source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use participation read against direct and indirect education costs. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-20] [REF-21]

    For costs borne by households, the material distinction is between the comparison should identify the reference category but avoid presenting it as a natural norm. The evidence must therefore clarify how policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that nominal fee abolition checked against remaining expenditure. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-23] [REF-24]

    costs borne by households requires a decision about where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. A contrary reading would overlook that particular scrutiny is required for large families and households hit by economic crisis. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-01] [REF-07]

    In assessing costs borne by households, authorities must determine urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. The evidence must therefore clarify how targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-02] [REF-03]

    Part VII

    Place and service geography

    19

    Rural and urban residence

    rural and urban residence requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that a distribution-sensitive account of education outcomes by a declared settlement classification begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]

    The practical standard for rural and urban residence concerns when several sources exist, consistency is evidence to consider, not proof that common error is absent. The resulting interpretation should show why a defensible statistic would be based on residence linked to service availability and travel conditions. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-23] [REF-24]

    Evidence concerning rural and urban residence should establish nor should a group estimate be read as a description of every member. For the learners concerned, the decisive consideration is whether within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: national rural definitions preserved and comparison limitations stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-01] [REF-07]

    A defensible account of rural and urban residence distinguishes analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The public account remains incomplete unless it explains how missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether remote villages, pastoral populations and peri-urban settlements are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-02] [REF-03]

    rural and urban residence cannot be judged without identifying the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. The evidence must therefore clarify how a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required.[REF-04] [REF-06]

    20

    Subnational administrative disparity

    subnational administrative disparity requires a decision about the measure should preserve both the observed level and the distribution relevant to the claim. The resulting interpretation should show why a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of subnational administrative disparity lies in making unequal educational experience observable. Here, the relevant phenomenon is variation between provinces, districts or comparable areas, not the administrative convenience of the available categories.[REF-23] [REF-24]

    Institutional action on subnational administrative disparity should be tested against a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. This matters because a survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is area estimates with population size and precision. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions.[REF-01] [REF-07]

    Public responsibility for subnational administrative disparity begins with it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. For the learners concerned, the decisive consideration is whether it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that administrative rankings not mistaken for causal explanations. A responsible commentary distinguishes observation, calculation and interpretation.[REF-02] [REF-03]

    subnational administrative disparity requires a decision about disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. The public account remains incomplete unless it explains how this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For small districts and areas with incomplete reporting, a single group label may conceal important internal differences.[REF-04] [REF-06]

    Review of subnational administrative disparity is credible only where it explains this makes the indicator a means of scrutiny rather than a decorative measure of concern. The resulting interpretation should show why public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation.[REF-08] [REF-09]

    21

    Distance, isolation and transport

    distance, isolation and transport cannot be judged without identifying the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns physical accessibility of the nearest appropriate service. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-01] [REF-07]

    The central question in distance, isolation and transport is numerator, denominator, reference date, unit and exclusions should appear together. The institutional consequence follows from whether if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is travel time, route safety and seasonal interruption rather than straight-line distance alone.[REF-02] [REF-03]

    distance, isolation and transport requires a decision about explanation requires evidence on institutions, resources, households and prior conditions. The public account remains incomplete unless it explains how interpretation follows this limitation: household reports and facility mapping reconciled. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-04] [REF-06]

    Comparative interpretation of distance, isolation and transport depends upon protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The institutional consequence follows from whether the distributional review must deliberately include learners with limited mobility and communities cut off seasonally. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-08] [REF-09]

    The practical standard for distance, isolation and transport concerns monitoring after action must preserve the original baseline and follow both reach and outcome. For the learners concerned, the decisive consideration is whether improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-10] [REF-12]

    Part VIII

    Disability, language and identity

    22

    Disability-sensitive education data

    Comparative interpretation of disability-sensitive education data depends upon the substantive interest is participation and learning by functional difficulty and support requirement, observed for a population and period that are stated before calculation. The institutional consequence follows from whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Disability-sensitive education data should be approached as a defined measurement problem.[REF-02] [REF-03]

    The practical standard for disability-sensitive education data concerns counts reveal scale; rates permit comparison; neither is sufficient alone. This matters because source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use questions designed for comparable reporting without defining the child by diagnosis alone. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-04] [REF-06]

    disability-sensitive education data requires a decision about percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The evidence must therefore clarify how the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that identification, environment and accommodation kept analytically distinct. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-08] [REF-09]

    Review of disability-sensitive education data is credible only where it explains where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. The institutional consequence follows from whether where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners whose impairments are not recorded by schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness.[REF-10] [REF-12]

    In assessing disability-sensitive education data, authorities must determine these are different decisions and require different certainty. The evidence must therefore clarify how urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure.[REF-13] [REF-14]

    23

    Language of home and instruction

    Review of language of home and instruction is credible only where it explains the study seeks evidence on alignment between learner language, teaching and assessment; it does not infer a learner's circumstances from a national or regional mean. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For language of home and instruction, the first requirement is conceptual clarity.[REF-04] [REF-06]

    language of home and instruction requires a decision about analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. A proportionate conclusion must also recognise that this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on language categories reflecting local use and instructional practice. The definition should be fixed for the comparison at hand and deviations recorded.[REF-08] [REF-09]

    For language of home and instruction, the material distinction is between results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. The public account remains incomplete unless it explains how if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: small groups not erased through broad national labels.[REF-10] [REF-12]

    Comparative interpretation of language of home and instruction depends upon missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. For the learners concerned, the decisive consideration is whether an adequate equity account asks whether minority-language and multilingual learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-13] [REF-14]

    language of home and instruction requires a decision about if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The institutional consequence follows from whether the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-15] [REF-16]

    24

    Ethnicity, indigeneity and protected identity

    Evidence concerning ethnicity, indigeneity and protected identity should establish a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of disparity associated with historically excluded identity groups begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]

    The practical standard for ethnicity, indigeneity and protected identity concerns these are substantive attributes because they determine who can appear in the evidence. The public account remains incomplete unless it explains how a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is lawful, voluntary and contextually meaningful classification. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-10] [REF-12]

    A defensible account of ethnicity, indigeneity and protected identity distinguishes it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. A contrary reading would overlook that it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that self-identification protected and non-response reported. A responsible commentary distinguishes observation, calculation and interpretation.[REF-13] [REF-14]

    ethnicity, indigeneity and protected identity requires a decision about the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For the learners concerned, the decisive consideration is whether for communities exposed to discrimination or forced assimilation, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-15] [REF-16]

    Institutional action on ethnicity, indigeneity and protected identity should be tested against subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. The institutional consequence follows from whether where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence.[REF-17] [REF-19]

    Part IX

    Conflict, disaster and mobility

    25

    Education under conflict and insecurity

    A defensible account of education under conflict and insecurity distinguishes the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of education under conflict and insecurity lies in making unequal educational experience observable. Here, the relevant phenomenon is access, attendance and learning where violence alters service and movement, not the administrative convenience of the available categories.[REF-10] [REF-12]

    education under conflict and insecurity requires a decision about the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. For the learners concerned, the decisive consideration is whether the preferred construction is location- and time-specific observation with explicit coverage gaps. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-13] [REF-14]

    Comparative interpretation of education under conflict and insecurity depends upon a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. A proportionate conclusion must also recognise that reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: absence caused by insecurity distinguished from ordinary dropout. The comparison should show the level for each group as well as any ratio or gap.[REF-15] [REF-16]

    The practical standard for education under conflict and insecurity concerns these populations may be missing not only from good outcomes but from the denominator itself. A proportionate conclusion must also recognise that coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners in insecure areas and host communities.[REF-17] [REF-19]

    education under conflict and insecurity requires a decision about monitoring after action must preserve the original baseline and follow both reach and outcome. This matters because improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-20] [REF-21]

    27

    Refugees, displaced persons and migrants

    In assessing refugees, displaced persons and migrants, authorities must determine the substantive interest is educational participation across changing legal and residential situations, observed for a population and period that are stated before calculation. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Refugees, displaced persons and migrants should be approached as a defined measurement problem.[REF-15] [REF-16]

    The practical standard for refugees, displaced persons and migrants concerns when several sources exist, consistency is evidence to consider, not proof that common error is absent. The evidence must therefore clarify how a defensible statistic would be based on status, origin, current location and service access recorded separately. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-17] [REF-19]

    Evidence concerning refugees, displaced persons and migrants should establish nor should a group estimate be read as a description of every member. The public account remains incomplete unless it explains how within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: mobility never converted into duplicate enrolment or unexplained disappearance. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-20] [REF-21]

    A defensible account of refugees, displaced persons and migrants distinguishes missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. This matters because an adequate equity account asks whether undocumented migrants, refugees and internally displaced learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-23] [REF-24]

    For refugees, displaced persons and migrants, the material distinction is between local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. The evidence must therefore clarify how communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable.[REF-01] [REF-07]

    Part X

    School conditions and teachers

    28

    Teacher availability and distribution

    Evidence concerning teacher availability and distribution should establish the measure should preserve both the observed level and the distribution relevant to the claim. The institutional consequence follows from whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For teacher availability and distribution, the first requirement is conceptual clarity. The study seeks evidence on access to competent teaching across schools and subjects; it does not infer a learner's circumstances from a national or regional mean.[REF-17] [REF-19]

    The central question in teacher availability and distribution is a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A proportionate conclusion must also recognise that a survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is teachers present and assigned relative to learner and curriculum need. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions.[REF-20] [REF-21]

    A defensible account of teacher availability and distribution distinguishes it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. A proportionate conclusion must also recognise that it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that payroll totals not substituted for classroom availability. A responsible commentary distinguishes observation, calculation and interpretation.[REF-23] [REF-24]

    The practical standard for teacher availability and distribution concerns it should involve statistical judgement, legal safeguards and knowledge of the affected community. A proportionate conclusion must also recognise that suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For schools serving poor, remote or displaced communities, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-01] [REF-07]

    Comparative interpretation of teacher availability and distribution depends upon this makes the indicator a means of scrutiny rather than a decorative measure of concern. A proportionate conclusion must also recognise that public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation.[REF-02] [REF-03]

    29

    Class size, multi-grade teaching and time

    A defensible account of class size, multi-grade teaching and time distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because a distribution-sensitive account of the instructional conditions experienced by learners begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]

    The central question in class size, multi-grade teaching and time is the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. This matters because the preferred construction is class organisation, scheduled time and delivered time considered together. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-23] [REF-24]

    Review of class size, multi-grade teaching and time is credible only where it explains explanation requires evidence on institutions, resources, households and prior conditions. The resulting interpretation should show why interpretation follows this limitation: simple pupil-teacher ratios not treated as a complete quality measure. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-01] [REF-07]

    Review of class size, multi-grade teaching and time is credible only where it explains protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The institutional consequence follows from whether the distributional review must deliberately include early grades and mixed-age classes. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-02] [REF-03]

    The practical standard for class size, multi-grade teaching and time concerns it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. The resulting interpretation should show why monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition.[REF-04] [REF-06]

    30

    Materials, facilities and basic services

    Evidence concerning materials, facilities and basic services should establish here, the relevant phenomenon is usable learning resources and safe, accessible school conditions, not the administrative convenience of the available categories. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of materials, facilities and basic services lies in making unequal educational experience observable.[REF-23] [REF-24]

    Comparative interpretation of materials, facilities and basic services depends upon counts reveal scale; rates permit comparison; neither is sufficient alone. This matters because source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use availability joined to condition, accessibility and regular use. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-01] [REF-07]

    For materials, facilities and basic services, the material distinction is between the comparison should identify the reference category but avoid presenting it as a natural norm. For the learners concerned, the decisive consideration is whether policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that delivery counts checked against learner access. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-02] [REF-03]

    The central question in materials, facilities and basic services is combining unknown observations with the majority group biases both estimates and obscures the weakness. For the learners concerned, the decisive consideration is whether where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in temporary or damaged premises. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-04] [REF-06]

    materials, facilities and basic services cannot be judged without identifying targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. For the learners concerned, the decisive consideration is whether accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves.[REF-08] [REF-09]

    Part XI

    Finance and distribution

    31

    Public spending by level and function

    Institutional action on public spending by level and function should be tested against the measure should preserve both the observed level and the distribution relevant to the claim. A contrary reading would overlook that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns resources assigned to education purposes across the system. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-01] [REF-07]

    In assessing public spending by level and function, authorities must determine analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. For the learners concerned, the decisive consideration is whether this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on expenditure classified by level, recurrent or capital use and responsible body. The definition should be fixed for the comparison at hand and deviations recorded.[REF-02] [REF-03]

    public spending by level and function requires a decision about national averages should remain available as context, yet never as a substitute for the distribution. A contrary reading would overlook that nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: budgets, commitments and actual expenditure distinguished. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding.[REF-04] [REF-06]

    public spending by level and function cannot be judged without identifying field arrangements need relevant languages, accessible formats and safe participation. The resulting interpretation should show why analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether basic education services under fiscal pressure are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-08] [REF-09]

    The central question in public spending by level and function is the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. The public account remains incomplete unless it explains how a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required.[REF-10] [REF-12]

    32

    Incidence of education spending

    The practical standard for incidence of education spending concerns the substantive interest is who benefits from publicly financed places and services, observed for a population and period that are stated before calculation. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Incidence of education spending should be approached as a defined measurement problem.[REF-02] [REF-03]

    Evidence concerning incidence of education spending should establish a survey estimate should disclose weights and uncertainty. The institutional consequence follows from whether a census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is unit resources combined with participation across population groups. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions.[REF-04] [REF-06]

    Comparative interpretation of incidence of education spending depends upon a responsible commentary distinguishes observation, calculation and interpretation. This matters because it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that benefit estimates not treated as household income.[REF-08] [REF-09]

    incidence of education spending cannot be judged without identifying the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For the learners concerned, the decisive consideration is whether for groups excluded before public spending can reach them, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-10] [REF-12]

    incidence of education spending requires a decision about readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. This matters because a policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary.[REF-13] [REF-14]

    33

    Protecting equity during fiscal constraint

    In assessing protecting equity during fiscal constraint, authorities must determine the study seeks evidence on whether reductions or delays fall disproportionately on weaker services; it does not infer a learner's circumstances from a national or regional mean. For the learners concerned, the decisive consideration is whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For protecting equity during fiscal constraint, the first requirement is conceptual clarity.[REF-04] [REF-06]

    The practical standard for protecting equity during fiscal constraint concerns it should require examination of definitions, timing, migration, duplication and non-response. The institutional consequence follows from whether the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is dated finance and service indicators read together. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-08] [REF-09]

    For protecting equity during fiscal constraint, the material distinction is between reference points therefore need substantive meaning. The institutional consequence follows from whether where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: national totals tested against subnational allocation and household costs. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-10] [REF-12]

    Review of protecting equity during fiscal constraint is credible only where it explains if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. A contrary reading would overlook that for protecting equity during fiscal constraint, the material distinction is between confidentiality is essential, especially where identity or status creates risk. The institutional consequence follows from whether protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include poor households and institutions with little financial reserve. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-13] [REF-14]

    The practical standard for protecting equity during fiscal constraint concerns the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. A proportionate conclusion must also recognise that the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked.[REF-15] [REF-16]

    Part XII

    Data sources and measurement error

    34

    Administrative records

    The practical standard for administrative records concerns the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A contrary reading would overlook that a distribution-sensitive account of regular learner, staff, facility and finance information begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-08] [REF-09]

    Evidence concerning administrative records should establish counts reveal scale; rates permit comparison; neither is sufficient alone. This matters because source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use clear definitions, reporting coverage and revision history. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-10] [REF-12]

    administrative records cannot be judged without identifying the comparison should identify the reference category but avoid presenting it as a natural norm. This matters because policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that non-reporting institutions kept visible in aggregates. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-13] [REF-14]

    The practical standard for administrative records concerns their circumstances may alter access to enumeration, classification and the service being measured. The institutional consequence follows from whether the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for small, private, non-formal and emergency providers.[REF-15] [REF-16]

    For administrative records, the material distinction is between urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. The resulting interpretation should show why targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-17] [REF-19]

    35

    Household surveys

    household surveys cannot be judged without identifying here, the relevant phenomenon is population-based evidence beyond enrolled learners, not the administrative convenience of the available categories. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of household surveys lies in making unequal educational experience observable.[REF-10] [REF-12]

    household surveys requires a decision about the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. For the learners concerned, the decisive consideration is whether when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on probability samples, weights and field dates documented. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-13] [REF-14]

    Public responsibility for household surveys begins with if a conclusion changes under a reasonable specification, that instability is part of the finding. The public account remains incomplete unless it explains how national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: sampling and non-response uncertainty carried into group comparisons. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-15] [REF-16]

    household surveys requires a decision about missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The resulting interpretation should show why an adequate equity account asks whether small minorities and mobile households are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-17] [REF-19]

    Institutional action on household surveys should be tested against the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. The institutional consequence follows from whether a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required.[REF-20] [REF-21]

    36

    Censuses and population frames

    censuses and population frames cannot be judged without identifying the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns broad population coverage and small-area denominators. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-13] [REF-14]

    A defensible account of censuses and population frames distinguishes at minimum this includes population, geography, date, collection method, classification and known exclusions. The resulting interpretation should show why a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is enumeration date, usual residence and institutional coverage stated. Its metadata should travel with every published value.[REF-15] [REF-16]

    Evidence concerning censuses and population frames should establish apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. A proportionate conclusion must also recognise that use of the indicator is bounded by the principle that long intervals and under-enumeration acknowledged. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-17] [REF-19]

    In assessing censuses and population frames, authorities must determine this balance is contextual rather than mechanical. A proportionate conclusion must also recognise that it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For homeless, displaced and geographically isolated people, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-20] [REF-21]

    Public responsibility for censuses and population frames begins with where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. A contrary reading would overlook that this makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions.[REF-23] [REF-24]

    Part XIII

    Disaggregation and intersection

    37

    Single-axis disaggregation

    single-axis disaggregation cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The public account remains incomplete unless it explains how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Single-axis disaggregation should be approached as a defined measurement problem. The substantive interest is separate reporting by sex, wealth, residence or another characteristic, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-15] [REF-16]

    Review of single-axis disaggregation is credible only where it explains if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. A proportionate conclusion must also recognise that administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is levels, gaps and denominators shown for each category. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-17] [REF-19]

    Institutional action on single-axis disaggregation should be tested against the comparison should show the level for each group as well as any ratio or gap. The public account remains incomplete unless it explains how a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: one axis not presented as a complete account of marginalisation.[REF-20] [REF-21]

    single-axis disaggregation requires a decision about protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The institutional consequence follows from whether the distributional review must deliberately include groups whose disadvantage lies on another unmeasured dimension. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-23] [REF-24]

    Evidence concerning single-axis disaggregation should establish it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. This matters because monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition.[REF-01] [REF-07]

    38

    Intersecting categories

    The central question in intersecting categories is the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For intersecting categories, the first requirement is conceptual clarity. The study seeks evidence on joint distributions such as sex by wealth and residence; it does not infer a learner's circumstances from a national or regional mean.[REF-17] [REF-19]

    Comparative interpretation of intersecting categories depends upon counts reveal scale; rates permit comparison; neither is sufficient alone. The institutional consequence follows from whether source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use pre-specified combinations with sufficient observations. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-20] [REF-21]

    The central question in intersecting categories is percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The public account remains incomplete unless it explains how the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that empty or unstable cells reported honestly. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-23] [REF-24]

    Public responsibility for intersecting categories begins with their circumstances may alter access to enumeration, classification and the service being measured. A contrary reading would overlook that the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for poor rural girls, disabled learners in remote areas and displaced minorities.[REF-01] [REF-07]

    Review of intersecting categories is credible only where it explains targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. A contrary reading would overlook that accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves.[REF-02] [REF-03]

    39

    Small numbers, disclosure and reliability

    A defensible account of small numbers, disclosure and reliability distinguishes without that purpose, disaggregation can multiply figures without improving public judgement. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of useful detail without unreliable estimates or identification begins by naming the decision the evidence may inform.[REF-20] [REF-21]

    For small numbers, disclosure and reliability, the material distinction is between the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. The evidence must therefore clarify how when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on suppression, aggregation or qualitative evidence chosen proportionately. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-23] [REF-24]

    small numbers, disclosure and reliability cannot be judged without identifying results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. A contrary reading would overlook that if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: confidentiality decisions separated from claims that no disparity exists.[REF-01] [REF-07]

    The practical standard for small numbers, disclosure and reliability concerns analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. This matters because missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether small communities and learners with rare characteristics are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-02] [REF-03]

    In assessing small numbers, disclosure and reliability, authorities must determine if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The institutional consequence follows from whether the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-04] [REF-06]

    Part XIV

    Comparison, uncertainty and change

    40

    Comparing unlike systems

    For comparing unlike systems, the material distinction is between a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The institutional consequence follows from whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of comparing unlike systems lies in making unequal educational experience observable. Here, the relevant phenomenon is cross-country patterns based on harmonised but bounded concepts, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-23] [REF-24]

    comparing unlike systems cannot be judged without identifying at minimum this includes population, geography, date, collection method, classification and known exclusions. For the learners concerned, the decisive consideration is whether a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is metadata tests before numerical comparison. Its metadata should travel with every published value.[REF-01] [REF-07]

    Evidence concerning comparing unlike systems should establish it does not assign cause from a cross-sectional difference. A contrary reading would overlook that apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that differences in programme structure and classification remain visible. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-02] [REF-03]

    The practical standard for comparing unlike systems concerns it should involve statistical judgement, legal safeguards and knowledge of the affected community. A contrary reading would overlook that suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For countries with incomplete or rapidly changing systems, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-04] [REF-06]

    Public responsibility for comparing unlike systems begins with this makes the indicator a means of scrutiny rather than a decorative measure of concern. For the learners concerned, the decisive consideration is whether public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation.[REF-08] [REF-09]

    41

    Sampling error and other uncertainty

    sampling error and other uncertainty requires a decision about its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns the range of values reasonably compatible with the observations.[REF-01] [REF-07]

    A defensible account of sampling error and other uncertainty distinguishes if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. For the learners concerned, the decisive consideration is whether administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is standard errors, design effects and data-quality qualifications. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-02] [REF-03]

    In assessing sampling error and other uncertainty, authorities must determine statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. A contrary reading would overlook that explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: rank differences smaller than uncertainty not interpreted. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-04] [REF-06]

    The central question in sampling error and other uncertainty is if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. For the learners concerned, the decisive consideration is whether confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include small disaggregated populations. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-08] [REF-09]

    Review of sampling error and other uncertainty is credible only where it explains it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. The public account remains incomplete unless it explains how monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition.[REF-10] [REF-12]

    Part XV

    Responsible interpretation and action

    43

    Reading disparity without blaming learners

    reading disparity without blaming learners cannot be judged without identifying the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how for reading disparity without blaming learners, the first requirement is conceptual clarity. The study seeks evidence on institutional and social conditions associated with unequal outcomes; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-04] [REF-06]

    For reading disparity without blaming learners, the material distinction is between the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. A proportionate conclusion must also recognise that when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on descriptive findings separated from causal claims. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-08] [REF-09]

    Comparative interpretation of reading disparity without blaming learners depends upon nor should a group estimate be read as a description of every member. For the learners concerned, the decisive consideration is whether within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: group identity never treated as a mechanism by itself. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-10] [REF-12]

    reading disparity without blaming learners cannot be judged without identifying missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. A contrary reading would overlook that an adequate equity account asks whether communities subject to stigma are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-13] [REF-14]

    Institutional action on reading disparity without blaming learners should be tested against communities should be able to question both the category and the conclusion drawn from it. A proportionate conclusion must also recognise that if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-15] [REF-16]

    44

    Turning evidence into equitable policy

    turning evidence into equitable policy requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A contrary reading would overlook that a distribution-sensitive account of a decision rule linking disparity to service, finance or legal responsibility begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-08] [REF-09]

    Evidence concerning turning evidence into equitable policy should establish a survey estimate should disclose weights and uncertainty. The institutional consequence follows from whether a census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is baseline, intended reach, implementation evidence and review date. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions.[REF-10] [REF-12]

    Comparative interpretation of turning evidence into equitable policy depends upon a responsible commentary distinguishes observation, calculation and interpretation. A proportionate conclusion must also recognise that it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that targets accompanied by distributional safeguards.[REF-13] [REF-14]

    Comparative interpretation of turning evidence into equitable policy depends upon this balance is contextual rather than mechanical. The evidence must therefore clarify how it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For learners farthest below a secured minimum, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-15] [REF-16]

    Institutional action on turning evidence into equitable policy should be tested against subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. The public account remains incomplete unless it explains how where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence.[REF-17] [REF-19]

    45

    A public account beyond the average

    For a public account beyond the average, the material distinction is between a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The resulting interpretation should show why the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of a public account beyond the average lies in making unequal educational experience observable. Here, the relevant phenomenon is a concise national statement of level, distribution, missingness and remedy, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-10] [REF-12]

    Institutional action on a public account beyond the average should be tested against it should require examination of definitions, timing, migration, duplication and non-response. The resulting interpretation should show why the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is national totals presented beside selected group and place indicators. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-13] [REF-14]

    Evidence concerning a public account beyond the average should establish reference points therefore need substantive meaning. The institutional consequence follows from whether where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: progress claims bounded by evidence coverage and unresolved gaps. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-15] [REF-16]

    Public responsibility for a public account beyond the average begins with coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. This matters because if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include every learner otherwise hidden by a successful average. These populations may be missing not only from good outcomes but from the denominator itself.[REF-17] [REF-19]

    Comparative interpretation of a public account beyond the average depends upon monitoring after action must preserve the original baseline and follow both reach and outcome. The resulting interpretation should show why improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-20] [REF-21]

    Part XVI

    Applied distributional analysis

    46

    Composite measures and the loss of meaning

    In assessing composite measures and the loss of meaning, authorities must determine that purpose should be written before the calculation because method follows the intended inference. The public account remains incomplete unless it explains how the evidence base comprises participation, completion, learning and school conditions, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is combining several dimensions into a summary measure.[REF-01] [REF-03]

    Evidence concerning composite measures and the loss of meaning should establish if the headline conclusion changes materially, the range of results should be reported. The public account remains incomplete unless it explains how sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern weights, normalisation and substitution. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications.[REF-02] [REF-04]

    The practical standard for composite measures and the loss of meaning concerns analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A proportionate conclusion must also recognise that a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is a low value in one dimension being concealed by a high value in another. Avoiding it requires the underlying counts and distributions to remain visible beside any summary.[REF-05] [REF-06]

    Comparative interpretation of composite measures and the loss of meaning depends upon a record of disagreement should identify plausible sources and the decision consequence. A proportionate conclusion must also recognise that if the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity.[REF-08] [REF-12]

    A defensible account of composite measures and the loss of meaning distinguishes learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. The public account remains incomplete unless it explains how analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure.[REF-13] [REF-16]

    Institutional action on composite measures and the loss of meaning should be tested against every response should name the expected population reach and the later observation that will test it. A contrary reading would overlook that otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is policy-makers deciding whether a summary aids or obscures resource allocation. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves.[REF-17] [REF-19]

    In assessing composite measures and the loss of meaning, authorities must determine feedback needs a recorded route into classification, service design or further enquiry. A proportionate conclusion must also recognise that people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-21] [REF-23]

    For composite measures and the loss of meaning, the material distinction is between changes to definitions, boundaries or population estimates should appear at the point where a series changes. This matters because earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review.[REF-23] [REF-24]

    47

    Decomposing an observed education gap

    Review of decomposing an observed education gap is credible only where it explains the national total supplies context, while the distribution tests whether the total is shared. The resulting interpretation should show why where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is examining how a national disparity is distributed across places and population groups. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises within-group and between-group differences, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement.[REF-02] [REF-04]

    The practical standard for decomposing an observed education gap concerns sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. The institutional consequence follows from whether documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern population shares, outcome levels and overlapping membership. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported.[REF-05] [REF-06]

    decomposing an observed education gap cannot be judged without identifying avoiding it requires the underlying counts and distributions to remain visible beside any summary. A proportionate conclusion must also recognise that analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is treating a descriptive decomposition as proof of cause.[REF-08] [REF-12]

    In assessing decomposing an observed education gap, authorities must determine a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. The institutional consequence follows from whether a record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage.[REF-13] [REF-16]

    Institutional action on decomposing an observed education gap should be tested against these duties do not justify silence about a serious disparity. This matters because the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability.[REF-17] [REF-19]

    Institutional action on decomposing an observed education gap should be tested against otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The resulting interpretation should show why the policy user is authorities locating where further enquiry and action are warranted. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-21] [REF-23]

    decomposing an observed education gap requires a decision about the two forms answer different questions. For the learners concerned, the decisive consideration is whether authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. The central question in decomposing an observed education gap is people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. A proportionate conclusion must also recognise that participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier.[REF-23] [REF-24]

    In assessing decomposing an observed education gap, authorities must determine a concise account can carry considerable density when every figure retains its population and consequence. The resulting interpretation should show why the objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress.[REF-01] [REF-03]

    48

    Setting distribution-sensitive targets

    A defensible account of setting distribution-sensitive targets distinguishes where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The resulting interpretation should show why the analytical purpose is expressing progress as improvement in secured minimums and unjustified gaps. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises the national level, the least-served group and the lower tail of the distribution, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-05] [REF-06]

    setting distribution-sensitive targets requires a decision about choices on these matters are not neutral presentation details. A contrary reading would overlook that they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern baseline stability, ambition and safeguards against exclusion.[REF-08] [REF-12]

    setting distribution-sensitive targets requires a decision about no result should be described as equitable solely because one relative measure improved. This matters because the main interpretive danger is meeting a mean target while abandoning those farthest behind. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity.[REF-13] [REF-16]

    Comparative interpretation of setting distribution-sensitive targets depends upon censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. The evidence must therefore clarify how a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response.[REF-17] [REF-19]

    setting distribution-sensitive targets cannot be judged without identifying learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. The evidence must therefore clarify how analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure.[REF-21] [REF-23]

    setting distribution-sensitive targets cannot be judged without identifying every response should name the expected population reach and the later observation that will test it. The evidence must therefore clarify how otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is governments linking national commitments to subnational delivery. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves.[REF-23] [REF-24]

    Comparative interpretation of setting distribution-sensitive targets depends upon feedback needs a recorded route into classification, service design or further enquiry. A proportionate conclusion must also recognise that people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-01] [REF-03]

    The central question in setting distribution-sensitive targets is a concise account can carry considerable density when every figure retains its population and consequence. The resulting interpretation should show why the objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress.[REF-02] [REF-04]

    49

    Linking learners to service geography

    A defensible account of linking learners to service geography distinguishes the evidence base comprises settlement populations, travel conditions, schools, teachers and programme levels, each of which describes a different feature of educational opportunity. The resulting interpretation should show why a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is relating participation and learning to the location and capacity of education services. That purpose should be written before the calculation because method follows the intended inference.[REF-08] [REF-12]

    In assessing linking learners to service geography, authorities must determine they determine how strongly one component, place or population can influence the conclusion. The institutional consequence follows from whether a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern geographical scale, boundary effects and facility catchments. Choices on these matters are not neutral presentation details.[REF-13] [REF-16]

    Public responsibility for linking learners to service geography begins with comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. The resulting interpretation should show why no result should be described as equitable solely because one relative measure improved. The main interpretive danger is assuming the nearest mapped institution is accessible or appropriate. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-17] [REF-19]

    Institutional action on linking learners to service geography should be tested against surveys can represent households beyond formal education, subject to sample size, field access and response. This matters because censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-21] [REF-23]

    Institutional action on linking learners to service geography should be tested against the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. A proportionate conclusion must also recognise that equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity.[REF-23] [REF-24]

    In assessing linking learners to service geography, authorities must determine credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. The evidence must therefore clarify how every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is planners choosing sites, transport support and teacher deployment. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-01] [REF-03]

    Public responsibility for linking learners to service geography begins with feedback needs a recorded route into classification, service design or further enquiry. This matters because people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-02] [REF-04]

    linking learners to service geography cannot be judged without identifying the objective is not maximum numerical output. For the learners concerned, the decisive consideration is whether it is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence.[REF-05] [REF-06]

    50

    Reconciling conflicting sources

    Institutional action on reconciling conflicting sources should be tested against where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. A contrary reading would overlook that where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is interpreting differences between administrative, survey and census estimates. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises coverage, timing, concepts and reporting incentives, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-13] [REF-16]

    A defensible account of reconciling conflicting sources distinguishes choices on these matters are not neutral presentation details. The public account remains incomplete unless it explains how they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a documented comparison of population and variable definitions.[REF-17] [REF-19]

    Review of reconciling conflicting sources is credible only where it explains no result should be described as equitable solely because one relative measure improved. For the learners concerned, the decisive consideration is whether the main interpretive danger is averaging incompatible estimates into an apparently precise figure. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity.[REF-21] [REF-23]

    Public responsibility for reconciling conflicting sources begins with a record of disagreement should identify plausible sources and the decision consequence. The evidence must therefore clarify how if the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity.[REF-23] [REF-24]

    The central question in reconciling conflicting sources is analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. For the learners concerned, the decisive consideration is whether small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins.[REF-01] [REF-03]

    Review of reconciling conflicting sources is credible only where it explains otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The public account remains incomplete unless it explains how the policy user is statistical authorities issuing one bounded account with visible uncertainty. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-02] [REF-04]

    Review of reconciling conflicting sources is credible only where it explains the two forms answer different questions. This matters because authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier.[REF-05] [REF-06]

    The central question in reconciling conflicting sources is it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. The institutional consequence follows from whether changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement.[REF-08] [REF-12]

    51

    Monitoring marginalisation during severe disruption

    Review of monitoring marginalisation during severe disruption is credible only where it explains that purpose should be written before the calculation because method follows the intended inference. For the learners concerned, the decisive consideration is whether the evidence base comprises rapid counts, restored administrative returns and household evidence, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is maintaining useful distributional evidence when populations and services move rapidly.[REF-17] [REF-19]

    Review of monitoring marginalisation during severe disruption is credible only where it explains a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. This matters because if the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern dated estimates, revision practice and minimal essential classifications. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion.[REF-21] [REF-23]

    Institutional action on monitoring marginalisation during severe disruption should be tested against a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. The public account remains incomplete unless it explains how direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is using an unstable emergency denominator to assert durable improvement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold.[REF-23] [REF-24]

    For monitoring marginalisation during severe disruption, the material distinction is between surveys can represent households beyond formal education, subject to sample size, field access and response. The public account remains incomplete unless it explains how censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-01] [REF-03]

    Institutional action on monitoring marginalisation during severe disruption should be tested against analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. A contrary reading would overlook that small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins.[REF-02] [REF-04]

    monitoring marginalisation during severe disruption requires a decision about credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. The resulting interpretation should show why every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities protecting access while rebuilding regular statistics. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-05] [REF-06]

    A defensible account of monitoring marginalisation during severe disruption distinguishes feedback needs a recorded route into classification, service design or further enquiry. The resulting interpretation should show why people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-08] [REF-12]

    For monitoring marginalisation during severe disruption, the material distinction is between changes to definitions, boundaries or population estimates should appear at the point where a series changes. The resulting interpretation should show why earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review.[REF-13] [REF-16]

    52

    Communicating uncertainty without losing urgency

    communicating uncertainty without losing urgency cannot be judged without identifying the national total supplies context, while the distribution tests whether the total is shared. The resulting interpretation should show why where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is explaining what is known strongly enough to justify action and what remains unresolved. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises point estimates, ranges, quality statements and missing populations, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement.[REF-21] [REF-23]

    Public responsibility for communicating uncertainty without losing urgency begins with missing values should remain missing unless an explicit estimation method and its effect are shown. A proportionate conclusion must also recognise that the principal methodological questions concern plain institutional language joined to exact metadata. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable.[REF-23] [REF-24]

    communicating uncertainty without losing urgency requires a decision about direction must therefore be joined to adequacy. The evidence must therefore clarify how comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is presenting caution as a reason for inaction or urgency as a reason for overstatement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens.[REF-01] [REF-03]

    A defensible account of communicating uncertainty without losing urgency distinguishes a record of disagreement should identify plausible sources and the decision consequence. A proportionate conclusion must also recognise that if the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity.[REF-02] [REF-04]

    communicating uncertainty without losing urgency cannot be judged without identifying the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. This matters because equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity.[REF-05] [REF-06]

    The central question in communicating uncertainty without losing urgency is evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. For the learners concerned, the decisive consideration is whether the certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is the public, affected communities and responsible decision-makers.[REF-08] [REF-12]

    Public responsibility for communicating uncertainty without losing urgency begins with consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The public account remains incomplete unless it explains how the two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss.[REF-13] [REF-16]

    Institutional action on communicating uncertainty without losing urgency should be tested against it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. The resulting interpretation should show why changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement.[REF-17] [REF-19]

    53

    A national marginalisation profile

    The central question in a national marginalisation profile is where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. For the learners concerned, the decisive consideration is whether where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is assembling a concise recurring account beyond the national average. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises population, access, progression, learning, conditions, finance and unresolved evidence gaps, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-23] [REF-24]

    Evidence concerning a national marginalisation profile should establish they determine how strongly one component, place or population can influence the conclusion. The public account remains incomplete unless it explains how a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a stable core with context-specific distributions. Choices on these matters are not neutral presentation details.[REF-01] [REF-03]

    a national marginalisation profile requires a decision about comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. A contrary reading would overlook that no result should be described as equitable solely because one relative measure improved. The main interpretive danger is creating an encyclopaedia of indicators without decision priority. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-02] [REF-04]

    a national marginalisation profile requires a decision about a record of disagreement should identify plausible sources and the decision consequence. A contrary reading would overlook that if the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity.[REF-05] [REF-06]

    The central question in a national marginalisation profile is small cells require protection against disclosure and caution about statistical stability. The resulting interpretation should show why these duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete.[REF-08] [REF-12]

    Institutional action on a national marginalisation profile should be tested against otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The resulting interpretation should show why the policy user is parliament, ministries, local authorities and communities reviewing educational equity. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-13] [REF-16]

    The practical standard for a national marginalisation profile concerns the two forms answer different questions. The evidence must therefore clarify how authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier.[REF-17] [REF-19]

    Institutional action on a national marginalisation profile should be tested against changes to definitions, boundaries or population estimates should appear at the point where a series changes. For the learners concerned, the decisive consideration is whether earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review.[REF-21] [REF-23]

    Part XVII

    Operational interpretation of inclusive and equitable quality

    54

    Inclusion as presence, participation and belonging

    The interpretive question is whether every learner can enter, attend, participate and remain within a recognised educational route. The term should be applied to a stated population, education level and decision. Broad commitment provides direction, but operational use requires observable conditions. A national average cannot establish that the condition exists for every school or learner. The standard should identify the learner-facing entitlement before selecting an indicator or administrative instrument.[REF-01] [REF-07]

    The principal misinterpretation is enrolment alone being treated as inclusion. Avoiding it requires distinctions between eligibility, offer, access, participation, educational quality, learning and recognised progression. Each is necessary for some claims and none automatically proves the rest. A learner can be enrolled yet unable to participate; present yet taught an inaccessible curriculum; learning yet denied recognised continuation. The judgement should locate the first substantive break.[REF-03] [REF-04]

    Relevant evidence comprises admission, accessibility, participation, voice, support and progression. Administrative records, assessments, observation and learner evidence should be used within scope. Missing populations remain visible. Counts show scale; rates permit comparison; distributions reveal unequal conditions. A result should carry its date, population, geography and limitations. Where a source does not observe the relevant condition, the report should not infer it from a nearby measure.[REF-05] [REF-06]

    The governing action is to state the excluded population and practical barrier. This requires an accountable body, authority, resources, deadline and review. Local adaptation may be justified by language, disability, geography or prior opportunity, but the common substantive entitlement should remain clear. Differentiation is not a lower standard when it removes an irrelevant barrier or supplies the additional condition needed for equal participation.[REF-08] [REF-12]

    A defensible account of inclusion as presence, participation and belonging distinguishes these conditions can intersect. For the learners concerned, the decisive consideration is whether group evidence should support institutional correction rather than define individual worth or ability. Small numbers require confidentiality and caution about precision. They do not justify omission from the duty to examine access and quality. The standard should be tested for learners facing poverty, disability, remote residence, minority language, displacement, gendered barriers or weak prior provision.[REF-10] [REF-13]

    In assessing inclusion as presence, participation and belonging, authorities must determine review should separate delivery, use and consequence and should preserve adverse findings. This matters because where correction is needed, it may concern admission, staffing, curriculum, accessibility, support, assessment, records or finance. The remedy should match the demonstrated mechanism. Public assurance should be no broader than the evidence. A policy, budget or facility can establish authorisation or input; learner experience and outcome establish different propositions.[REF-01] [REF-07]

    In assessing inclusion as presence, participation and belonging, authorities must determine temporary activity, a favourable average or the end of finance does not close unresolved learner needs. A proportionate conclusion must also recognise that the final record should identify remaining limitation, responsible body and next evidence date. This makes inclusive and equitable quality education a reviewable public standard rather than an unbounded description. Completion means that the ordinary system can sustain the interpreted condition and correct foreseeable departures.[REF-03] [REF-04]

    55

    Equity as substantive fairness

    The interpretive question is whether resources, rules and support respond fairly to different barriers and starting conditions. The term should be applied to a stated population, education level and decision. Broad commitment provides direction, but operational use requires observable conditions. A national average cannot establish that the condition exists for every school or learner. The standard should identify the learner-facing entitlement before selecting an indicator or administrative instrument.[REF-03] [REF-04]

    The principal misinterpretation is identical treatment being presumed equitable. Avoiding it requires distinctions between eligibility, offer, access, participation, educational quality, learning and recognised progression. Each is necessary for some claims and none automatically proves the rest. A learner can be enrolled yet unable to participate; present yet taught an inaccessible curriculum; learning yet denied recognised continuation. The judgement should locate the first substantive break.[REF-05] [REF-06]

    Relevant evidence comprises levels, gaps, needs, additional cost and distributional effect. Administrative records, assessments, observation and learner evidence should be used within scope. Missing populations remain visible. Counts show scale; rates permit comparison; distributions reveal unequal conditions. A result should carry its date, population, geography and limitations. Where a source does not observe the relevant condition, the report should not infer it from a nearby measure.[REF-08] [REF-12]

    The governing action is to justify differentiation by the common entitlement protected. This requires an accountable body, authority, resources, deadline and review. Local adaptation may be justified by language, disability, geography or prior opportunity, but the common substantive entitlement should remain clear. Differentiation is not a lower standard when it removes an irrelevant barrier or supplies the additional condition needed for equal participation.[REF-10] [REF-13]

    equity as substantive fairness requires a decision about these conditions can intersect. The evidence must therefore clarify how group evidence should support institutional correction rather than define individual worth or ability. Small numbers require confidentiality and caution about precision. They do not justify omission from the duty to examine access and quality. The standard should be tested for learners facing poverty, disability, remote residence, minority language, displacement, gendered barriers or weak prior provision.[REF-01] [REF-07]

    Review of equity as substantive fairness is credible only where it explains where correction is needed, it may concern admission, staffing, curriculum, accessibility, support, assessment, records or finance. The resulting interpretation should show why the remedy should match the demonstrated mechanism. Public assurance should be no broader than the evidence. A policy, budget or facility can establish authorisation or input; learner experience and outcome establish different propositions. Review should separate delivery, use and consequence and should preserve adverse findings.[REF-03] [REF-04]

    Comparative interpretation of equity as substantive fairness depends upon this makes inclusive and equitable quality education a reviewable public standard rather than an unbounded description. The institutional consequence follows from whether completion means that the ordinary system can sustain the interpreted condition and correct foreseeable departures. Temporary activity, a favourable average or the end of finance does not close unresolved learner needs. The final record should identify remaining limitation, responsible body and next evidence date.[REF-05] [REF-06]

    56

    Quality as educational substance and conditions

    The interpretive question is whether education offers credible knowledge, capability, teaching, assessment and recognised progression. The term should be applied to a stated population, education level and decision. Broad commitment provides direction, but operational use requires observable conditions. A national average cannot establish that the condition exists for every school or learner. The standard should identify the learner-facing entitlement before selecting an indicator or administrative instrument.[REF-05] [REF-06]

    The principal misinterpretation is test scores or inputs alone standing for the whole educational experience. Avoiding it requires distinctions between eligibility, offer, access, participation, educational quality, learning and recognised progression. Each is necessary for some claims and none automatically proves the rest. A learner can be enrolled yet unable to participate; present yet taught an inaccessible curriculum; learning yet denied recognised continuation. The judgement should locate the first substantive break.[REF-08] [REF-12]

    Relevant evidence comprises curriculum, teachers, time, materials, learning evidence and wellbeing. Administrative records, assessments, observation and learner evidence should be used within scope. Missing populations remain visible. Counts show scale; rates permit comparison; distributions reveal unequal conditions. A result should carry its date, population, geography and limitations. Where a source does not observe the relevant condition, the report should not infer it from a nearby measure.[REF-10] [REF-13]

    The governing action is to connect quality judgement to learner use and public purpose. This requires an accountable body, authority, resources, deadline and review. Local adaptation may be justified by language, disability, geography or prior opportunity, but the common substantive entitlement should remain clear. Differentiation is not a lower standard when it removes an irrelevant barrier or supplies the additional condition needed for equal participation.[REF-01] [REF-07]

    quality as educational substance and conditions requires a decision about small numbers require confidentiality and caution about precision. The evidence must therefore clarify how they do not justify omission from the duty to examine access and quality. The standard should be tested for learners facing poverty, disability, remote residence, minority language, displacement, gendered barriers or weak prior provision. These conditions can intersect. Group evidence should support institutional correction rather than define individual worth or ability.[REF-03] [REF-04]

    Evidence concerning quality as educational substance and conditions should establish a policy, budget or facility can establish authorisation or input; learner experience and outcome establish different propositions. The institutional consequence follows from whether review should separate delivery, use and consequence and should preserve adverse findings. Where correction is needed, it may concern admission, staffing, curriculum, accessibility, support, assessment, records or finance. The remedy should match the demonstrated mechanism. Public assurance should be no broader than the evidence.[REF-05] [REF-06]

    The central question in quality as educational substance and conditions is this makes inclusive and equitable quality education a reviewable public standard rather than an unbounded description. A proportionate conclusion must also recognise that completion means that the ordinary system can sustain the interpreted condition and correct foreseeable departures. Temporary activity, a favourable average or the end of finance does not close unresolved learner needs. The final record should identify remaining limitation, responsible body and next evidence date.[REF-08] [REF-12]

    57

    The combined minimum threshold

    The interpretive question is whether inclusion, equity and quality are jointly satisfied at a defensible minimum. The term should be applied to a stated population, education level and decision. Broad commitment provides direction, but operational use requires observable conditions. A national average cannot establish that the condition exists for every school or learner. The standard should identify the learner-facing entitlement before selecting an indicator or administrative instrument.[REF-08] [REF-12]

    The principal misinterpretation is a strong result in one dimension compensating silently for failure in another. Avoiding it requires distinctions between eligibility, offer, access, participation, educational quality, learning and recognised progression. Each is necessary for some claims and none automatically proves the rest. A learner can be enrolled yet unable to participate; present yet taught an inaccessible curriculum; learning yet denied recognised continuation. The judgement should locate the first substantive break.[REF-10] [REF-13]

    Relevant evidence comprises access, conditions, learning, distribution and remedy considered together. Administrative records, assessments, observation and learner evidence should be used within scope. Missing populations remain visible. Counts show scale; rates permit comparison; distributions reveal unequal conditions. A result should carry its date, population, geography and limitations. Where a source does not observe the relevant condition, the report should not infer it from a nearby measure.[REF-01] [REF-07]

    The governing action is to refuse a composite assurance that conceals a failed essential condition. This requires an accountable body, authority, resources, deadline and review. Local adaptation may be justified by language, disability, geography or prior opportunity, but the common substantive entitlement should remain clear. Differentiation is not a lower standard when it removes an irrelevant barrier or supplies the additional condition needed for equal participation.[REF-03] [REF-04]

    Evidence concerning the combined minimum threshold should establish they do not justify omission from the duty to examine access and quality. The public account remains incomplete unless it explains how the standard should be tested for learners facing poverty, disability, remote residence, minority language, displacement, gendered barriers or weak prior provision. These conditions can intersect. Group evidence should support institutional correction rather than define individual worth or ability. Small numbers require confidentiality and caution about precision.[REF-05] [REF-06]

    The practical standard for the combined minimum threshold concerns where correction is needed, it may concern admission, staffing, curriculum, accessibility, support, assessment, records or finance. The public account remains incomplete unless it explains how the remedy should match the demonstrated mechanism. Public assurance should be no broader than the evidence. A policy, budget or facility can establish authorisation or input; learner experience and outcome establish different propositions. Review should separate delivery, use and consequence and should preserve adverse findings.[REF-08] [REF-12]

    Review of the combined minimum threshold is credible only where it explains the final record should identify remaining limitation, responsible body and next evidence date. This matters because this makes inclusive and equitable quality education a reviewable public standard rather than an unbounded description. Completion means that the ordinary system can sustain the interpreted condition and correct foreseeable departures. Temporary activity, a favourable average or the end of finance does not close unresolved learner needs.[REF-10] [REF-13]

    58

    Operational judgement and public accountability

    The interpretive question is how authorities decide, evidence and correct departures from the standard. The term should be applied to a stated population, education level and decision. Broad commitment provides direction, but operational use requires observable conditions. A national average cannot establish that the condition exists for every school or learner. The standard should identify the learner-facing entitlement before selecting an indicator or administrative instrument.[REF-10] [REF-13]

    The principal misinterpretation is aspirational language lacking ownership, thresholds or review. Avoiding it requires distinctions between eligibility, offer, access, participation, educational quality, learning and recognised progression. Each is necessary for some claims and none automatically proves the rest. A learner can be enrolled yet unable to participate; present yet taught an inaccessible curriculum; learning yet denied recognised continuation. The judgement should locate the first substantive break.[REF-01] [REF-07]

    Relevant evidence comprises population, baseline, competent body, measure, evidence and next date. Administrative records, assessments, observation and learner evidence should be used within scope. Missing populations remain visible. Counts show scale; rates permit comparison; distributions reveal unequal conditions. A result should carry its date, population, geography and limitations. Where a source does not observe the relevant condition, the report should not infer it from a nearby measure.[REF-03] [REF-04]

    The governing action is to publish a bounded decision and implement practical correction. This requires an accountable body, authority, resources, deadline and review. Local adaptation may be justified by language, disability, geography or prior opportunity, but the common substantive entitlement should remain clear. Differentiation is not a lower standard when it removes an irrelevant barrier or supplies the additional condition needed for equal participation.[REF-05] [REF-06]

    The practical standard for operational judgement and public accountability concerns small numbers require confidentiality and caution about precision. A contrary reading would overlook that they do not justify omission from the duty to examine access and quality. The standard should be tested for learners facing poverty, disability, remote residence, minority language, displacement, gendered barriers or weak prior provision. These conditions can intersect. Group evidence should support institutional correction rather than define individual worth or ability.[REF-08] [REF-12]

    Public assurance should be no broader than the evidence. A policy, budget or facility can establish authorisation or input; learner experience and outcome establish different propositions. Review should separate delivery, use and consequence and should preserve adverse findings. Where correction is needed, it may concern admission, staffing, curriculum, accessibility, support, assessment, records or finance. The remedy should match the demonstrated mechanism.[REF-10] [REF-13]

    Completion means that the ordinary system can sustain the interpreted condition and correct foreseeable departures. Temporary activity, a favourable average or the end of finance does not close unresolved learner needs. The final record should identify remaining limitation, responsible body and next evidence date. This makes inclusive and equitable quality education a reviewable public standard rather than an unbounded description.[REF-01] [REF-07]

    Part XVIII

    Regional priorities within a universal goal

    59

    A common entitlement with differentiated routes

    A universal goal should state the educational entitlement in terms capable of application to every learner: entry at the appropriate stage, sustained participation, completion of a meaningful course, demonstrable learning, safe and enabling conditions, and recognised opportunity to progress. Regional differentiation should not qualify that entitlement. It should identify the conditions that make its achievement more difficult and the public measures required in response. This distinction protects against two opposite errors: a global statement so general that it cannot discipline policy, and separate regional standards that make a learner's claim depend upon place.[REF-01] [REF-07] [REF-08]

    Starting points nevertheless matter. A system seeking to bring a large out-of-school population into primary education faces a different immediate allocation problem from a system with near-universal entry but severe inequality in completion or learning. A region experiencing rapid growth of the school-age population requires sustained expansion of classrooms and teachers; a region with demographic contraction may need to reorganise provision without abandoning remote communities. Common indicators should reveal those differences, while national milestones state the feasible sequence of action. The purpose of differentiation is to direct sufficient effort, not to normalise a lower outcome.[REF-02] [REF-03] [REF-12]

    Regional priorities should be expressed through a limited set of diagnosed barriers. These may include distance, poverty, disability, language, gendered expectations, conflict, displacement, weak civil registration, teacher shortages, insecure financing or unequal transition to secondary education. Each priority should name the affected population, the responsible authority, the educational consequence and the evidence by which improvement will be judged. A list of concerns without ownership does not constitute an implementation framework. Nor does a regional average reveal whether the barrier has been removed for the populations most exposed to it.[REF-04] [REF-06] [REF-10]

    The global account should therefore distinguish the universal measure from the contextual explanation. A completion rate, for example, may be defined consistently, while the explanation for non-completion draws upon regional evidence concerning household costs, seasonal work, displacement, school supply or the language of instruction. Comparable measures permit collective scrutiny; contextual explanation prevents crude ranking. The two functions should be joined in reporting but not confused. An authority should not alter the indicator merely because the result is unfavourable, and a global review should not prescribe the same remedy where the mechanisms differ.[REF-05] [REF-09] [REF-13]

    Milestones should preserve ambition across the full distribution. Concentrating support on learners already near completion may improve a national value while leaving the least-served groups unchanged. Regional plans should consequently combine an aggregate trajectory with a minimum floor and an explicit distributional test. Where reliable group estimates are unavailable, the immediate obligation is to improve coverage and use service evidence cautiously, not to treat the unobserved population as having attained the target. A universal goal becomes credible only when progress among learners facing the greatest barriers is a constitutive part of progress, not a supplementary narrative.[REF-11] [REF-14] [REF-24]

    Regional institutions can add value by agreeing compatible concepts, convening peer examination, supporting countries with limited statistical capacity and identifying cross-border concerns. Their contribution should complement the responsibility of national authorities to finance provision, regulate institutions and report results. Regional comparison is most constructive when it asks which institutional arrangements have secured a defined educational condition under comparable constraints. It is least constructive when it ranks unlike systems without disclosing population, programme or measurement differences. The common goal should encourage disciplined learning between systems while preserving the right of each population to question its own public authorities.[REF-16] [REF-18] [REF-20]

    60

    Priorities for sub-Saharan Africa and other rapidly expanding systems

    In countries where the number of learners entering school continues to rise quickly, the immediate priority is to secure expansion without allowing instructional conditions to deteriorate. Enrolment growth must be read beside teacher recruitment, deployment, attendance, classroom use, learning time, materials and progression. A place offered is not yet a complete educational opportunity if the school cannot sustain teaching. Planning should therefore use demographic projections, local school-age populations and teacher requirements together, with special attention to districts where national expansion has not reached the learner.[REF-01] [REF-02] [REF-23]

    Primary entry remains indispensable, but progression and completion require equal attention. Late entry, repetition and temporary withdrawal can produce overcrowded lower grades and conceal the distance between registration and a completed cycle. Cohort evidence should distinguish promotion, repetition, transfer, re-entry and permanent exit where records permit. Household evidence is needed because administrative returns describe learners in institutions better than children who remain outside them. The priority is not to substitute one source for another but to reconcile them sufficiently to understand where participation is lost and which authority can act.[REF-03] [REF-05] [REF-24]

    Teacher policy is central to the credibility of any universal goal. National totals may coexist with severe shortages in remote, poor or insecure areas. Recruitment must therefore be accompanied by preparation, fair deployment, professional support, suitable housing where relevant, and measures to retain qualified personnel. Reliance on unprepared personnel can maintain nominal staffing at the cost of instructional quality; rigid requirements can also delay urgent expansion where supported entry routes would be more appropriate. The defensible course is to protect minimum competence while establishing a financed path to recognised qualification and continuing support.[REF-10] [REF-12] [REF-14]

    Language and locality shape participation. Where children begin education in a language unfamiliar to them, attendance and apparent attainment may reflect the accessibility of instruction as much as underlying capability. Policy should consider the language of early learning, the availability of teachers and materials, transition between languages, and the participation of communities in defining workable arrangements. Remote and sparsely populated communities may require smaller schools, multigrade provision, transport or other adaptations. Such measures should be judged by the common entitlement they secure, not dismissed because their organisation differs from an urban norm.[REF-06] [REF-08] [REF-13]

    Financing must reflect the additional cost of reaching underserved populations. Equal per-learner allocations can reproduce inequality where distance, disability, language or weak infrastructure make provision more costly. Formulae and grants should disclose the population factors they recognise and be checked against actual receipt by schools. Household costs also matter: fee abolition does not remove expenditure on transport, clothing, materials or the opportunity cost of attendance. Public reporting should link announced allocation, timely transfer, school receipt and learner-facing condition, while avoiding an assumption that expenditure alone proves educational improvement.[REF-15] [REF-17] [REF-20]

    Regional review should thus place expansion, completion, teaching capacity and equity in one account. The strongest evidence will show whether districts with the greatest initial deficit gained teachers, instructional time and sustained participation, and whether learning opportunities improved without excluding children from assessment. Emergency measures may be required in periods of disruption, but they should be dated and connected to recovery of ordinary provision. The universal goal should make the scale of the task visible without representing rapid expansion as an excuse for weak educational substance.[REF-04] [REF-09] [REF-21]

    61

    Priorities for South and West Asia

    The regional priority is to join unfinished access commitments with a more exact account of participation, gender, social stratification and learning. National progress can conceal deep differences associated with wealth, rural residence, disability, language, caste or other locally relevant forms of disadvantage. Single-axis comparison is insufficient where these conditions overlap. The public account should retain national direction while showing the absolute number and educational experience of learners in the least-served groups. Where classifications are sensitive, safeguards should protect individuals without removing the institutional obligation to identify exclusion.[REF-01] [REF-04] [REF-08]

    Gender parity should be interpreted as one test within a broader equality duty. A ratio close to parity can coexist with low participation for both girls and boys, and a national value can conceal opposite patterns across wealth groups or localities. Reporting should therefore show underlying levels, timely entry, attendance, transition and completion. Policy explanation should examine safety, sanitation, distance, household labour, early marriage and social expectations where evidence supports them. It should not treat gender identity itself as the cause of an observed difference. The purpose is to identify remediable institutions and conditions.[REF-03] [REF-06] [REF-10]

    Large and diverse systems require subnational evidence capable of guiding allocation. Provincial or district averages may still conceal communities that differ in school supply or language. Administrative completeness, population estimates and survey coverage should be examined before disparities are compared. Migration between rural and urban areas can change both the demand for places and the validity of residence-based denominators. A dated account of population movement, new school capacity and service responsibility is necessary, particularly in informal settlements where households may fall outside ordinary planning categories.[REF-02] [REF-05] [REF-23]

    Learning improvement should be grounded in opportunity to learn. Assessment can identify weakness only among learners who are represented, present and able to engage with the instrument. Results should therefore be read beside curriculum exposure, teacher attendance, instructional language and prior participation. When low results trigger support, the response should strengthen teaching rather than create incentives to exclude lower-performing pupils. Where sample or school-based assessment omits children outside provision, that coverage limitation should be stated prominently in any claim about regional or national learning.[REF-11] [REF-12] [REF-14]

    The secondary transition demands growing attention as primary participation expands. Capacity, cost, distance, selection and the organisation of academic and technical routes affect who progresses. A universal goal should not define success solely at initial entry if learners encounter an avoidable barrier at the next stage. It should show completion of the preceding level, availability of a suitable onward place, take-up and later retention. Technical and vocational routes should carry transparent entry requirements and recognised progression so that differentiation does not become a means of directing disadvantaged learners into programmes with weak standing.[REF-07] [REF-13] [REF-16]

    Regional cooperation can improve evidence on comparable participation and learning while respecting programme differences. Common concepts should be accompanied by national metadata on age, duration, curriculum and collection practice. Peer examination should consider whether public measures reached disadvantaged groups and whether reported change reflects better services rather than revised coverage. The resulting regional priority is not a separate entitlement but a disciplined concentration on persistent internal inequality, the quality of classroom provision and the continuity from primary entry through recognised completion.[REF-18] [REF-20] [REF-24]

    62

    Priorities for Arab States and conflict-affected education systems

    Where conflict, political instability or displacement disrupts provision, continuity and protection become immediate educational priorities. Yet emergency access should remain connected to a recognised course of learning, competent teaching, safe participation and a route back into ordinary provision. Temporary learning spaces may be necessary, but their existence does not alone establish educational continuity. Authorities and supporting organisations should record the population served, duration, curriculum, teacher arrangements and recognised status of learning so that urgent measures do not become an unexamined lower track.[REF-09] [REF-21] [REF-22]

    Population denominators require particular caution. Displacement changes residence, duplicates or removes records, and can leave learners outside both origin and host planning systems. Counts from reception points, schools and households may refer to different populations and dates. A responsible account should retain those differences, update estimates visibly and separate rapid figures used for immediate allocation from values suitable for later comparison. Where population size remains uncertain, ranges and coverage statements are preferable to spurious precision. The absence of a reliable denominator must not be interpreted as the absence of educational need.[REF-02] [REF-03] [REF-23]

    Host systems need financing and authority proportionate to the additional population they serve. A universal goal should make displaced and refugee learners visible without encouraging separate structures where inclusion in public provision is feasible. Planning should address classrooms, teachers, language support, records, certification and the needs of host communities. Provision that benefits only one population can intensify local grievance; provision that ignores the additional barriers faced by displaced learners can reproduce exclusion. Equity requires differentiated support directed toward a common substantive opportunity.[REF-04] [REF-08] [REF-20]

    Recognition of prior learning and portable records are central to continuity. Learners may lack documents, have completed part of a programme under another curriculum, or move again before assessment. Admission rules should distinguish the need to establish an appropriate placement from requirements that become impossible barriers. Provisional placement, diagnostic assessment and later documentation can protect continuity where supported by public rules. Certification should state what has been completed and provide a credible route to further education; otherwise attendance may not translate into recognised progression.[REF-07] [REF-13] [REF-16]

    Gender, disability and poverty can intensify the effect of insecurity. Safe travel, suitable facilities, accessible communication, psychosocial support and the protection of learners and personnel may be prerequisites for participation. Evidence should not assume that registration demonstrates regular attendance or that assessment results represent learners who could not reach the school. Community evidence can reveal these barriers, but collection must avoid creating additional risk. The governing test is whether the arrangement restores a dependable educational relationship and whether an accountable body can correct interruption or exclusion.[REF-10] [REF-12] [REF-14]

    Regional cooperation is especially valuable where displacement crosses borders. Compatible records, recognition arrangements, exchange of population information and support to host authorities can reduce discontinuity. Cooperation must comply with protection and confidentiality duties; individual information should not be disclosed merely for statistical completeness. The regional priority should be measured by continuity, inclusion, recognised learning and the restoration of ordinary public responsibility, rather than by the number of short-term activities reported.[REF-05] [REF-18] [REF-24]

    63

    Priorities for Latin America and the Caribbean

    Where average participation has expanded, the policy question increasingly concerns unequal completion, learning and transition within systems. National progress should be disaggregated by wealth, rural and urban location, language, ethnicity or other relevant conditions, with care for rights and confidentiality. The purpose is to reveal institutional barriers and unequal provision, not to attach educational deficit to the identity of a population. Public accounts should show levels and absolute numbers as well as gaps, because relative convergence can coexist with serious deprivation across all groups.[REF-01] [REF-04] [REF-08]

    Remote rural areas and urban informal settlements pose different planning challenges. Sparse populations can make conventional school organisation costly, while rapidly growing settlements may not appear promptly in administrative maps or population projections. Local service evidence should therefore be reconciled with census and household information. Transport, multigrade teaching, flexible calendars and new school capacity may each be appropriate in particular settings, provided that adaptations protect learning time, qualified support and recognised completion. The global goal should judge the educational condition secured rather than the administrative form through which it is delivered.[REF-02] [REF-03] [REF-23]

    Secondary education and the transition to work or further study deserve an explicit place in regional planning. Expansion can be inequitable when poorer learners enter programmes with weaker resources, less recognised certification or limited progression. Reporting should distinguish access to different programme types, completion and subsequent opportunity. Technical education should respond to public and learner purposes beyond immediate placement, including foundational knowledge, citizenship and the capacity for further learning. Regulation should ensure that public and private provision supplies accurate information and credentials with clear standing.[REF-07] [REF-13] [REF-16]

    Learning assessment should support improvement at school and system levels without narrowing curriculum or penalising institutions serving disadvantaged populations. Results require contextual evidence on prior opportunity, resources, language and attendance. Public reporting should preserve uncertainty and avoid causal claims based solely on associations. The strongest use is to identify where curricular expectations are not being secured, allocate support, examine teaching conditions and review whether the intended population benefited. A favourable regional mean cannot substitute for evidence that the lower part of the distribution advanced.[REF-10] [REF-11] [REF-14]

    Education financing should be examined for distributional effect. Decentralised responsibilities can bring decisions closer to communities but may reproduce differences in local revenue or administrative capacity unless national transfers compensate. Funding rules should make explicit how poverty, remoteness, disability and school size are recognised. Expenditure figures should be linked to staffing, materials, infrastructure and learner experience. Community participation can improve scrutiny, but it should not transfer the cost or statutory duty of provision to households least able to bear it.[REF-12] [REF-15] [REF-20]

    Regional mechanisms can strengthen comparison on completion, learning and equity while supporting exchange on bilingual education, rural provision and secondary transition. Their legitimacy rests on transparent definitions and attention to internal distributions. The priority is to convert broad participation gains into sustained and recognised learning for populations whom averages still conceal. This requires stable public responsibility, remedial action based on diagnosed barriers and reporting that allows communities to see whether announced reforms changed the conditions of their schools.[REF-05] [REF-18] [REF-24]

    64

    Priorities for Europe, North America and other mature systems

    High aggregate participation does not remove the need for a universal goal. It shifts attention toward early leaving, unequal learning, migrant and minority inclusion, disability, adult skills and transitions through secondary, vocational and higher education. A mature system can reproduce exclusion through selection, residential segregation, inaccessible provision or routes with unequal standing. Regional reporting should therefore preserve the common entitlement while examining distribution across schools, programmes and places. Near-universal entry is an achievement, not evidence that every learner receives an equitable quality education.[REF-01] [REF-07] [REF-16]

    Participation should be traced through completion and recognised progression. Administrative definitions of early leaving, completion and qualification need sufficient stability for comparison, while national programme differences remain visible. A learner's movement between general, technical and vocational routes should not be recorded as success if the receiving route offers no credible qualification or onward opportunity. Information, guidance and recognition arrangements should enable informed choice. Public authorities should examine whether selection and financing concentrate disadvantaged learners in institutions with weaker teaching conditions or expectations.[REF-03] [REF-13] [REF-18]

    Migration requires responsive admission, language support and recognition of prior education. Residence status, documentation or unfamiliar credentials can interrupt participation even where places exist. Placement should be timely and based on relevant educational evidence, with additional support where language or interrupted schooling requires it. Group disparities may reveal barriers, but identity should not be treated as a causal explanation. Analysis should examine school allocation, curriculum access, teacher expectations, discrimination and household conditions, while protecting personal information and avoiding categories that expose individuals to harm.[REF-04] [REF-08] [REF-20]

    Disability inclusion requires attention to accessibility, reasonable support, curriculum participation, assessment and transition. Counting learners in a mainstream institution cannot establish inclusion if they receive little instructional time or remain excluded from the curriculum. Separate provision also requires scrutiny of educational substance and recognised progression. Data collection should make disability visible without relying on classifications so narrow that many functional barriers disappear. The policy test is whether the ordinary system anticipates diverse needs, supplies appropriate support and corrects denial of participation.[REF-06] [REF-10] [REF-12]

    Fiscal constraint should not be allowed to obscure distributional choices. Aggregate education expenditure may remain stable while local services, staffing or household costs change. Review should identify which levels and populations bear reductions, whether short-term economies create longer-term exclusion, and whether support is concentrated where need is greatest. Efficiency is a legitimate public objective when it protects educational substance; it is not established by lower expenditure alone. Any reorganisation of schools or programmes should assess travel, accessibility, staffing and continuity for affected learners before claiming improvement.[REF-15] [REF-17] [REF-19]

    Regional comparison can use common concepts for attainment, early leaving, skills and participation, but indicators should be accompanied by programme and population metadata. Peer learning is strongest when it links observed difference to institutional design and preserves uncertainty. The regional priority is to confront persistent inequality within well-resourced systems and to show how learning, inclusion and recognised progression are secured across the distribution. The universal goal remains relevant precisely because educational exclusion can persist behind favourable averages and established administrative capacity.[REF-05] [REF-14] [REF-24]

    65

    Regional accountability and the global public record

    The global framework should contain a concise set of common measures and a disciplined space for regional priorities. The common measures should cover population, participation, completion, learning, teaching conditions and distribution. Regional accounts should explain the principal barriers, selected milestones, public measures and evidentiary limitations. They should not create a parallel record in which a region can claim success against a weaker definition. Every regional statement should be traceable to the same substantive entitlement, even where the indicator, sequencing or institutional arrangement reflects different circumstances.[REF-01] [REF-03] [REF-07]

    Baseline choice is consequential. A recent value may improve relevance but conceal earlier deterioration; an older baseline may be unreliable or unrepresentative of present borders and populations. Reports should retain the baseline selected, state revisions and show whether change reflects service improvement, demographic movement or altered coverage. Milestones should be judged against both pace and remaining distance. Countries with strong initial levels still bear duties toward excluded populations, while countries with severe initial deficits require support proportionate to the scale and cost of securing the common minimum.[REF-02] [REF-05] [REF-23]

    Regional averages should never be the sole unit of judgement. Differences between countries can be smaller than differences within them, and the regional mean may be dominated by large populations. Results should show country and internal distributions where evidence permits, accompanied by absolute numbers and uncertainty. Small states, territories and displaced populations require explicit treatment so that their experience is not lost through weighting. Where comparable evidence is absent, the gap should be assigned for improvement and supported through regional statistical cooperation.[REF-04] [REF-08] [REF-24]

    Review should distinguish commitment, adoption, delivery, participation and consequence. A declaration can establish political agreement; legislation can create authority; a budget can authorise resources; none alone proves changed educational opportunity. Regional bodies should report which stage is evidenced and should preserve adverse or incomplete findings. Independent scrutiny, civil society participation and access to underlying definitions strengthen the public account. Participation should be organised so that affected communities can question both the categories used and the conclusions drawn without bearing responsibility for correcting failures that belong to public authorities.[REF-10] [REF-12] [REF-20]

    International support should follow the diagnosed constraint. Financial assistance may be necessary where national revenue cannot meet rapid expansion or emergency demand; technical cooperation may strengthen teacher preparation, data systems or curriculum; regional agreements may improve recognition and continuity. Support should not fragment national responsibility through disconnected activities. It should be aligned with a public plan, disclose its duration and financing, and establish how functions will be sustained. The measure of cooperation is the educational condition secured, especially for learners whom existing institutions have served least well.[REF-13] [REF-16] [REF-18]

    As at 3 April 2014, the defensible conclusion is procedural and substantive rather than predictive. A universal education goal beyond 2015 should join equal entitlement, measurable educational substance, visible distribution, differentiated implementation and public responsibility. Later intergovernmental choices cannot be presumed. The present task is to establish what any credible formulation would need to protect: no learner outside the population account, no access claim detached from participation and learning, no regional priority converted into a lesser right, and no favourable aggregate accepted without evidence of who benefited.[REF-06] [REF-14] [REF-21]

    Part XIX

    Positioning quality in the post-2015 framework

    66

    Quality as a property of the education commitment

    Quality should not appear only in explanatory text attached to an access target. Its position must be strong enough to affect what is counted as educational progress. Participation is indispensable because teaching cannot benefit learners who are excluded, yet participation alone does not establish that the intended education has been provided. The framework should therefore join access and educational substance in its principal formulation and carry that relationship into target definitions. A country should be able to show both who participates and what conditions and learning that participation makes possible.[REF-01] [REF-03] [REF-07]

    The term quality must also remain broader than one outcome measure. Learning evidence is essential, but quality includes the curriculum offered, the competence and support of teachers, instructional time, safety, accessibility, language, materials, assessment and recognised progression. These conditions are not interchangeable. A favourable score among assessed learners cannot compensate for the exclusion of others; adequate resources do not prove that teaching is effective; completion does not prove that the course carried recognised educational value. A sound public account states which proposition each item of evidence supports and avoids converting a partial observation into a comprehensive assurance.[REF-05] [REF-10] [REF-12]

    Relevance should be interpreted through public purpose and the learner's capacity to participate in society, continue learning and pursue a livelihood with dignity. It does not justify reducing education to an immediate labour-market signal or allowing local adaptation to remove access to shared knowledge. National curricula may express different histories, languages and priorities while remaining subject to rights, equality and the requirement that qualifications support further opportunity. The global framework need not prescribe one curriculum. It should require that curriculum and assessment be coherent, publicly justified and accessible to the populations entitled to them.[REF-08] [REF-13] [REF-16]

    Quality has a distribution. A national average is incapable of establishing that educational substance reaches remote schools, disadvantaged households, minority-language communities, learners with disabilities or populations affected by displacement. The framework should therefore require disaggregation selected for a clear equality question and safeguarded against disclosure or stigma. Group results should be interpreted alongside counts, population coverage and uncertainty. Where a population is missing from the evidence, the omission should appear as an evidentiary limitation and an obligation to improve coverage, never as tacit evidence of satisfactory provision.[REF-02] [REF-04] [REF-24]

    Improvement should be judged against a secured minimum and continuing advancement. A relative gap can narrow because the better-served group deteriorates, while an aggregate level can rise because resources are concentrated on learners closest to a threshold. Neither pattern establishes equitable quality. Targets should protect the direction of the national level, the minimum condition for every group and the narrowing of unjustified disparities. The framework should not demand a composite score that conceals failure in one essential dimension; a transparent set of related measures is more faithful to the educational judgement.[REF-06] [REF-11] [REF-14]

    Positioning quality in this way gives it practical consequences. Plans must identify the responsible body, diagnosed constraint, measure, financing and review date. Reports must distinguish commitment, delivery, learner use and outcome. Corrective action should match the mechanism shown by evidence: staffing where instructional capacity is weak, language support where curriculum access is impaired, transport where distance prevents attendance, accessibility where facilities exclude, and household assistance where cost is decisive. Quality becomes operational when public authorities can be required to explain and correct the condition, not merely affirm the term.[REF-15] [REF-18] [REF-20]

    67

    A balanced target and indicator structure

    The target structure should follow the learner's educational course without becoming an inventory of every administrative variable. Entry, sustained participation, progression, completion, learning and transition form a coherent sequence. Measures at one stage should not be used as proxies for another. Enrolment can show recorded access but not attendance; attendance can show presence but not the content of instruction; completion can show an administrative status but not necessarily learning; assessment can show performance among represented learners but not whether all entitled learners had an opportunity to be assessed.[REF-02] [REF-03] [REF-05]

    A small number of headline measures may be necessary for global communication. Their limitations should be controlled through a wider national account, not hidden. Every headline value should carry a precise population, numerator, denominator, reference period and treatment of missing observations. The accompanying account should show relevant distributions, source coverage and any revision. Where administrative and household sources differ, the discrepancy should be examined for population, timing and definition before one figure is preferred. Global simplicity is legitimate only when it rests upon national evidence capable of testing the claim.[REF-01] [REF-10] [REF-23]

    Learning indicators require particular safeguards. The domain assessed, age or grade population, language, participation rate and treatment of non-response determine the meaning of a result. School-based assessments may provide strong evidence about enrolled pupils but cannot represent children outside school without additional population information. Grade-based comparisons may be affected by repetition, late entry and selection. A framework should encourage learning evidence while refusing an inference broader than the assessed population and content. The proper response to weak participation is to improve inclusion and reporting, not to award silence the value of success.[REF-04] [REF-11] [REF-14]

    Teacher indicators should connect numbers with competence, deployment and support. A national pupil-teacher ratio can conceal shortages by subject, level and locality and can overlook teacher absence or multiple shifts. Qualification categories may not be comparable if requirements differ. Reports should therefore state national definitions and show whether underserved schools receive prepared personnel, instructional leadership and continuing support. Teacher conditions are both an educational and distributional issue: unstable employment, excessive workload or unsafe placement can weaken continuity precisely where learners face the greatest barriers.[REF-06] [REF-12] [REF-20]

    Equity should not be confined to a parity index. Levels, absolute gaps, ratios and counts answer different questions and should be read together. A parity value near one can coexist with severe deprivation for all groups, and a small percentage-point difference may concern a large number of learners. The choice of comparison should be justified by the entitlement and decision. Intersections may be material where poverty, sex, residence, disability, language or displacement combine, but small samples and confidentiality require careful analysis rather than indiscriminate publication.[REF-07] [REF-08] [REF-24]

    The indicator structure should resist perverse incentives. If accountability rewards only a mean score, institutions may narrow curriculum or exclude lower-performing learners. If it rewards only completion, non-completion may be reclassified. Independent checks, stable definitions and publication of participation can reduce these risks. The framework should treat unexpected improvement as a reason for confirmation rather than suspicion, while making changes in definition and coverage visible. Credibility depends on measures that can be checked and corrected without turning schools or learners into instruments of ranking detached from their circumstances.[REF-09] [REF-13] [REF-18]

    68

    Means of implementation and differentiated responsibility

    An ambitious education commitment requires a credible account of means. Public finance, trained personnel, institutional authority, information and time are not ancillary matters; they determine whether the entitlement can be secured. National governments retain primary responsibility for organising and financing education, even when implementation is decentralised or international support is substantial. Plans should state the recurrent cost of teachers, facilities, materials, inclusion measures and evidence, rather than relying on temporary activity whose continuation is uncertain.[REF-01] [REF-15] [REF-20]

    Equal entitlement may require unequal allocation. Remote schools, learners with disabilities, minority-language provision, displaced populations and communities with weak infrastructure can cost more to serve well. Funding formulae should recognise relevant need transparently and be checked against actual school receipt and learner-facing conditions. Equal per-capita amounts can preserve unequal opportunity if they disregard those costs. Conversely, an additional allocation is not evidence of effectiveness by itself; review should establish whether resources arrived, were usable and changed the barrier they were intended to address.[REF-04] [REF-08] [REF-12]

    International cooperation should respond to diagnosed national and regional constraints. Financial support may be central where revenue and demographic pressure are misaligned; technical cooperation may strengthen teacher preparation, curriculum, data quality or planning; cross-border agreements may protect recognition and continuity. Support should operate within a public plan and disclose duration, conditions and responsibility for sustaining essential functions. Fragmented initiatives can impose parallel reporting and weaken national institutions. The proper measure is whether cooperation expands dependable public capacity and improves the educational condition of learners least well served.[REF-02] [REF-13] [REF-16]

    Capacity should not be treated as a fixed national attribute. It can be built through stable institutions, professional learning, regional cooperation and careful use of existing administrative and survey systems. Requirements should be proportionate to the decision: a rapid operational count may guide emergency provision, while a final comparative estimate needs stronger coverage and validation. Countries should not be excluded from the global account because evidence is weak; the weakness should be described and addressed. At the same time, uncertainty cannot justify claims more precise than the available information.[REF-03] [REF-10] [REF-23]

    Differentiated milestones can reflect different starting points without creating differentiated rights. Countries with large access deficits may prioritise rapid expansion alongside minimum teaching conditions; those with high participation may emphasise learning disparities, early leaving or adult education. Every trajectory should protect universal inclusion, public quality and transparent review. Ambition should be assessed by remaining distance, rate of improvement and the distribution of benefit. A country beginning from a low level requires support proportionate to the task, while a country with a favourable mean remains responsible for populations hidden behind it.[REF-05] [REF-14] [REF-24]

    Means of implementation should appear in the public review, not only in planning documents. Reports should identify authorised and executed expenditure, staffing, delivery of essential inputs, participation and educational consequence. This sequence prevents a budget announcement from standing for changed opportunity and enables explanation where delivery fails. It also distinguishes a resource constraint from a problem of authority, design or implementation. Such clarity supports proportionate remedy and allows national legislatures, communities and international partners to judge whether commitments are matched by sustained institutional action.[REF-06] [REF-17] [REF-18]

    69

    Public accountability before the framework is settled

    As of the evidence cut-off, the framework beyond 2015 remains under formulation. Policy analysis must therefore distinguish proposals, institutional positions and existing commitments from an adopted settlement. It may evaluate whether an option protects access, quality, equity and accountability, but it should not describe future wording or machinery as decided. This temporal discipline is substantive: presenting a proposal as an outcome can narrow public deliberation, misstate institutional authority and hide choices that remain open. Every cited proposition should be dated and attributed to the body competent to make it.[REF-01] [REF-07] [REF-20]

    The present review can nonetheless establish tests for credibility. A defensible formulation should identify the population entitled to education, the substance of the opportunity, the measures that support public judgement and the authority responsible for remedy. It should preserve both global comparability and national explanation. It should ensure that groups facing disadvantage remain visible and that progress does not depend on deterioration elsewhere. It should connect targets to resources and later review. These tests do not anticipate a negotiated text; they arise from existing rights commitments and the evidentiary limits demonstrated by official education monitoring.[REF-03] [REF-08] [REF-12]

    Accountability should operate at several levels without becoming diffuse. Schools and local authorities can report conditions and correct immediate barriers; national governments set law, finance and system standards; regional bodies can support compatible definitions and peer examination; international institutions can maintain the common account and coordinate support. Each role should be stated. A chain of reporting in which every level transmits information but no level owns correction is not accountability. The public record should identify who can alter the relevant condition and when evidence of the response will next be examined.[REF-04] [REF-10] [REF-18]

    Participation by teachers, learners and communities strengthens the interpretation of evidence. Administrative categories may miss irregular attendance, inaccessible instruction, informal cost or discriminatory treatment. Consultation can reveal these mechanisms and test whether a proposed remedy is workable. It should not be used to transfer statutory duties or financing burdens to affected communities. Nor should participation be reduced to endorsement of a predetermined plan. Authorities should record material concerns, explain their decision and provide a route for later challenge where the educational condition remains unresolved.[REF-05] [REF-13] [REF-24]

    Public communication should resist both excessive aggregation and excessive complexity. A concise statement can present the national level, distribution, missingness, principal action and next review date. Technical definitions should remain accessible so that independent readers can reproduce or question the claim. Corrections and revised estimates should be visible. A country or institution strengthens credibility by acknowledging uncertainty and outstanding barriers; confidence is weakened when a favourable headline is protected by changing definitions or withholding the population on which it rests.[REF-02] [REF-11] [REF-23]

    The policy position at 25 September 2014 is consequently clear without presuming later outcomes. Quality education should be integral to the principal commitment, measurable through related evidence rather than a single proxy, distributed equitably, financed according to need and subject to named public responsibility. Access, learning and inclusion are not rival priorities to be traded away in drafting. They are connected tests of whether educational opportunity is real. Any future framework should be judged by its capacity to hold those tests together for every learner and every jurisdiction.[REF-06] [REF-14] [REF-16]

    References

    1. REF-01

      Education for All Global Monitoring Report Team. Reaching the Marginalized — EFA Global Monitoring Report 2010. 2010.

      Principal contemporaneous analysis of intersecting disadvantage and education marginalisation.

      https://unesdoc.unesco.org/ark:/48223/pf0000186606
    2. REF-02

      UNESCO Institute for Statistics. Global Education Digest 2010: Comparing Education Statistics Across the World. 2010.

      Comparative education statistics, definitions and limitations.

      https://uis.unesco.org/sites/default/files/documents/global-education-digest-2010-comparing-education-statistics-across-the-world-en.pdf
    3. REF-03

      UNESCO Institute for Statistics. Education Indicators: Technical Guidelines. 2009.

      Definitions and interpretation of participation, progression, completion and resource indicators.

      https://uis.unesco.org/sites/default/files/documents/education-indicators-technical-guidelines-en_0.pdf
    4. REF-04

      United Nations. The Millennium Development Goals Report 2010. 2010.

      Global and regional monitoring of primary education, gender, poverty and related development conditions.

      https://www.un.org/millenniumgoals/pdf/MDG%20Report%202010%20En%20r15%20-low%20res%2020100615%20-.pdf
    5. REF-05

      United Nations Development Programme. Human Development Report 2010: The Real Wealth of Nations — Pathways to Human Development. 2010.

      Distribution-sensitive human development concepts and evidence available before the cut-off.

      https://hdr.undp.org/content/human-development-report-2010
    6. REF-06

      UNICEF. Progress for Children: Achieving the MDGs with Equity, Number 9. 2010.

      Equity-focused child indicators and comparison between population groups.

      https://www.unicef.org/reports/progress-children-no-9
    7. REF-07

      World Education Forum. The Dakar Framework for Action: Education for All — Meeting Our Collective Commitments. 2000.

      Commitments to equitable access, quality, measurable outcomes and accountable national planning.

      https://unesdoc.unesco.org/ark:/48223/pf0000121147
    8. REF-08

      United Nations General Assembly. Convention on the Rights of the Child. 1989.

      Rights concerning non-discrimination, identity, education and development.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-child
    9. REF-09

      United Nations Committee on Economic, Social and Cultural Rights. General Comment No. 13: The Right to Education. 1999.

      Interpretation of availability, accessibility, acceptability and adaptability.

      https://www.refworld.org/legal/general/cescr/1999/en/37937
    10. REF-10

      United Nations General Assembly. Convention on the Rights of Persons with Disabilities. 2006.

      Non-discrimination, accessibility, inclusive education and disability data safeguards.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-persons-disabilities
    11. REF-11

      United Nations. Guiding Principles on Internal Displacement. 1998.

      Principles relevant to protection, documentation, education and non-discrimination of displaced persons.

      https://www.ohchr.org/en/special-procedures/sr-internally-displaced-persons/international-standards
    12. REF-12

      UNESCO and UNICEF. A Human Rights-Based Approach to Education for All. 2007.

      Rights-based planning, equality, participation, accountability and education quality.

      https://unesdoc.unesco.org/ark:/48223/pf0000154861
    13. REF-13

      Education for All Global Monitoring Report Team. Overcoming Inequality: Why Governance Matters — EFA Global Monitoring Report 2009. 2008.

      Governance, finance and unequal educational opportunity.

      https://unesdoc.unesco.org/ark:/48223/pf0000177683
    14. REF-14

      World Bank. World Development Report 2006: Equity and Development. 2005.

      Concepts of unequal opportunity, institutions and equitable public action.

      https://documents.worldbank.org/curated/en/435331468127174418/pdf/322040World0Development0Report02006.pdf
    15. REF-15

      World Bank. Safeguarding Education During Economic Crisis. 2009.

      Risks to budgets, households, participation and long-term human development during economic crisis.

      https://documents1.worldbank.org/curated/en/489131468340200911/pdf/485120WP0Avert10Box338912B01PUBLIC1.pdf
    16. REF-16

      Organisation for Economic Co-operation and Development. Education at a Glance 2010: OECD Indicators. 2010.

      Comparative participation, progression, expenditure and outcomes evidence with system-level metadata.

      https://doi.org/10.1787/eag-2010-en
    17. REF-17

      European Commission. Europe 2020: A Strategy for Smart, Sustainable and Inclusive Growth. 2010.

      Contemporaneous European policy context for education, inclusion, employment and headline indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC2020
    18. REF-18

      European Commission. Youth on the Move: An Initiative to Unleash the Potential of Young People to Achieve Smart, Sustainable and Inclusive Growth in the European Union. 2010.

      European education, mobility, attainment and youth inclusion policy context.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0477
    19. REF-19

      European Commission. A Renewed Commitment to Social Europe: Reinforcing the Open Method of Coordination for Social Protection and Social Inclusion. 2008.

      Social inclusion monitoring, common objectives and context-sensitive indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52008DC0418
    20. REF-20

      United Nations General Assembly. Resolution 64/250: Assistance to Haiti in the Aftermath of the Recent Earthquake. 2010.

      Contemporaneous recognition of humanitarian and reconstruction needs and national leadership.

      https://undocs.org/A/RES/64/250
    21. REF-21

      United Nations Office for the Coordination of Humanitarian Affairs. Haiti Revised Humanitarian Appeal. 2010.

      Displacement, service disruption and humanitarian education context.

      https://reliefweb.int/report/haiti/haiti-revised-humanitarian-appeal-2010
    22. REF-22

      European Commission. European Union Response to the Earthquake in Haiti. 2010.

      European humanitarian and recovery support, coordination and Haitian ownership.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0056
    23. REF-23

      United Nations Economic and Social Council. Principles and Recommendations for Population and Housing Censuses, Revision 2. 2008.

      Official principles for population coverage, definitions, classifications and data quality.

      https://unstats.un.org/unsd/demographic-social/Standards-and-Methods/files/Principles_and_Recommendations/Population-and-Housing-Censuses/Series_M67Rev2-E.pdf
    24. REF-24

      United Nations Statistics Division. Designing Household Survey Samples: Practical Guidelines. 2005.

      Sample design, estimation, precision and non-response guidance.

      https://unstats.un.org/unsd/demographic/sources/surveys/Handbook23June05.pdf