ICEQC-R-2017-05 — Transparent Responsibilities for Governments, Institutions, Teachers and Families cover

专题研究报告

ICEQC-R-2017-05 — Transparent Responsibilities for Governments, Institutions, Teachers and Families

A global standards-interpretive study of public duty, professional responsibility, participation and remedy

发布日期
研究类别
标准解读
报告类型
标准解释研究
地理范围
Global
证据截止日期
负责机构
国际教育质量认证委员会研究与政策司
ICEQC-R-2017-05 — Transparent Responsibilities for Governments, Institutions, Teachers and Families cover

Publication record

This is the controlled English edition. Evidence and institutional status are stated as at the evidence cut-off date.

Executive summary

Key findings

    Scope and method

    Part II

    Purpose and interpretive frame

    1

    What national averages conceal

    what national averages conceal requires a decision about the measure should preserve both the observed level and the distribution relevant to the claim. The institutional consequence follows from whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns national participation or attainment. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-01] [REF-07]

    what national averages conceal requires a decision about numerator, denominator, reference date, unit and exclusions should appear together. For the learners concerned, the decisive consideration is whether if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is the weighted experience of the population included in the denominator.[REF-02] [REF-03]

    what national averages conceal cannot be judged without identifying where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. The public account remains incomplete unless it explains how statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: distribution within countries, not a league table between them. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning.[REF-04] [REF-06]

    For what national averages conceal, the material distinction is between coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. A contrary reading would overlook that if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include remote rural learners, urban informal settlements and displaced populations. These populations may be missing not only from good outcomes but from the denominator itself.[REF-08] [REF-09]

    Public responsibility for what national averages conceal begins with monitoring after action must preserve the original baseline and follow both reach and outcome. A proportionate conclusion must also recognise that improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-10] [REF-12]

    2

    Marginalisation as accumulated disadvantage

    In assessing marginalisation as accumulated disadvantage, authorities must determine the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Marginalisation as accumulated disadvantage should be approached as a defined measurement problem. The substantive interest is distance from a socially secured educational minimum, observed for a population and period that are stated before calculation.[REF-02] [REF-03]

    Review of marginalisation as accumulated disadvantage is credible only where it explains school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. This matters because reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use the joint effect of exclusion, weak provision and adverse social conditions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined.[REF-04] [REF-06]

    marginalisation as accumulated disadvantage requires a decision about disparity measures should not replace the underlying distributions. This matters because a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that multiple indicators read together over the learner's course.[REF-08] [REF-09]

    Public responsibility for marginalisation as accumulated disadvantage begins with their circumstances may alter access to enumeration, classification and the service being measured. This matters because the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for children facing poverty, gender disadvantage, disability or minority status.[REF-10] [REF-12]

    Comparative interpretation of marginalisation as accumulated disadvantage depends upon these are different decisions and require different certainty. The public account remains incomplete unless it explains how urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure.[REF-13] [REF-14]

    3

    From monitoring commitment to decision

    In assessing from monitoring commitment to decision, authorities must determine the study seeks evidence on evidence capable of changing resource or service decisions; it does not infer a learner's circumstances from a national or regional mean. For the learners concerned, the decisive consideration is whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For from monitoring commitment to decision, the first requirement is conceptual clarity.[REF-04] [REF-06]

    In assessing from monitoring commitment to decision, authorities must determine the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. For the learners concerned, the decisive consideration is whether when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on a link between observed disparity, responsible body and remedy. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-08] [REF-09]

    For from monitoring commitment to decision, the material distinction is between if a conclusion changes under a reasonable specification, that instability is part of the finding. The evidence must therefore clarify how national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: indicators selected for action rather than visibility. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-10] [REF-12]

    The practical standard for from monitoring commitment to decision concerns attrition at any stage can produce an apparently complete indicator from a selective population. The institutional consequence follows from whether field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether groups absent from routine plans and budget classifications are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-13] [REF-14]

    Review of from monitoring commitment to decision is credible only where it explains communities should be able to question both the category and the conclusion drawn from it. The institutional consequence follows from whether if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-15] [REF-16]

    Part III

    Population and denominator

    4

    Defining the population entitled to education

    defining the population entitled to education requires a decision about the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of resident and temporarily absent learners within the relevant age or programme group begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]

    The central question in defining the population entitled to education is a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. The institutional consequence follows from whether a survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is a denominator consistent with the right, level and reference period under review. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions.[REF-10] [REF-12]

    Review of defining the population entitled to education is credible only where it explains apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The institutional consequence follows from whether use of the indicator is bounded by the principle that census, survey and administrative estimates reconciled openly. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-13] [REF-14]

    Institutional action on defining the population entitled to education should be tested against it should involve statistical judgement, legal safeguards and knowledge of the affected community. A proportionate conclusion must also recognise that suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For unregistered residents, migrants and displaced children, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-15] [REF-16]

    A defensible account of defining the population entitled to education distinguishes where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This matters because this makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions.[REF-17] [REF-19]

    5

    Age, grade and programme populations

    Comparative interpretation of age, grade and programme populations depends upon the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The resulting interpretation should show why the public value of age, grade and programme populations lies in making unequal educational experience observable. Here, the relevant phenomenon is age-specific, grade-specific and programme-specific participation, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-10] [REF-12]

    A defensible account of age, grade and programme populations distinguishes the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. A proportionate conclusion must also recognise that the preferred construction is separate denominators for different educational questions. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-13] [REF-14]

    For age, grade and programme populations, the material distinction is between the comparison should show the level for each group as well as any ratio or gap. The evidence must therefore clarify how a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: exact age, school age and enrolled population never substituted silently.[REF-15] [REF-16]

    The central question in age, grade and programme populations is coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. This matters because if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include over-age entrants and learners repeating grades. These populations may be missing not only from good outcomes but from the denominator itself.[REF-17] [REF-19]

    6

    Population movement and disrupted residence

    Review of population movement and disrupted residence is credible only where it explains a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A proportionate conclusion must also recognise that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns education status amid migration, displacement and return. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-13] [REF-14]

    Evidence concerning population movement and disrupted residence should establish this formulation requires the reporting body to preserve the population base and the observation period beside the result. The evidence must therefore clarify how counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use dated location and residence rules with sensitivity to mobility.[REF-15] [REF-16]

    population movement and disrupted residence cannot be judged without identifying a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. A proportionate conclusion must also recognise that percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that origin, current location and service responsibility distinguished. Disparity measures should not replace the underlying distributions.[REF-17] [REF-19]

    Evidence concerning population movement and disrupted residence should establish where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. A proportionate conclusion must also recognise that where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for families affected by conflict, disaster and seasonal movement. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness.[REF-20] [REF-21]

    Part IV

    Access and participation

    7

    Entry at the official starting age

    For entry at the official starting age, the material distinction is between a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The resulting interpretation should show why the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Entry at the official starting age should be approached as a defined measurement problem. The substantive interest is timely admission to the first grade of primary education, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-15] [REF-16]

    In assessing entry at the official starting age, authorities must determine the definition should be fixed for the comparison at hand and deviations recorded. The public account remains incomplete unless it explains how analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on new entrants of official age relative to the corresponding population.[REF-17] [REF-19]

    The central question in entry at the official starting age is results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. For the learners concerned, the decisive consideration is whether if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: late entry examined alongside non-entry.[REF-20] [REF-21]

    entry at the official starting age requires a decision about attrition at any stage can produce an apparently complete indicator from a selective population. A proportionate conclusion must also recognise that field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether children facing fees, distance, disability or documentation barriers are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-23] [REF-24]

    8

    Attendance beyond enrolment

    Review of attendance beyond enrolment is credible only where it explains the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether for attendance beyond enrolment, the first requirement is conceptual clarity. The study seeks evidence on actual participation during a stated recent period; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-17] [REF-19]

    Comparative interpretation of attendance beyond enrolment depends upon a census figure should disclose enumeration rules. This matters because these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is presence measured independently of registration status. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-20] [REF-21]

    Institutional action on attendance beyond enrolment should be tested against apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. A proportionate conclusion must also recognise that use of the indicator is bounded by the principle that frequency, season and reasons for absence retained. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-23] [REF-24]

    For attendance beyond enrolment, the material distinction is between the public report can still state that a disparity was examined, whether action is required and which body will monitor it. The institutional consequence follows from whether for working children, carers and learners affected by illness, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-01] [REF-07]

    9

    Out-of-school status

    The practical standard for out-of-school status concerns without that purpose, disaggregation can multiply figures without improving public judgement. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of children in the relevant age group not participating at the defined level begins by naming the decision the evidence may inform.[REF-20] [REF-21]

    out-of-school status cannot be judged without identifying the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The institutional consequence follows from whether the preferred construction is a transparent residual from compatible population and participation concepts. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-23] [REF-24]

    A defensible account of out-of-school status distinguishes where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. The resulting interpretation should show why statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: never-enrolled and formerly enrolled children separated. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning.[REF-01] [REF-07]

    Institutional action on out-of-school status should be tested against if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. A contrary reading would overlook that confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children invisible to school registers. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-02] [REF-03]

    Part V

    Progression and completion

    10

    Repetition and grade survival

    Comparative interpretation of repetition and grade survival depends upon a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. For the learners concerned, the decisive consideration is whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of repetition and grade survival lies in making unequal educational experience observable. Here, the relevant phenomenon is movement through grades without avoidable delay or exit, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-23] [REF-24]

    A defensible account of repetition and grade survival distinguishes where estimates are revised, both the reason and effect of revision should remain accessible. A proportionate conclusion must also recognise that measurement should use cohort or reconstructed-cohort evidence with explicit assumptions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-01] [REF-07]

    For repetition and grade survival, the material distinction is between a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. For the learners concerned, the decisive consideration is whether percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that repeaters distinguished from re-entrants and transfers. Disparity measures should not replace the underlying distributions.[REF-02] [REF-03]

    The central question in repetition and grade survival is where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. The resulting interpretation should show why particular scrutiny is required for learners in overcrowded or intermittently operating schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-04] [REF-06]

    11

    Transition between education levels

    The central question in transition between education levels is the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how the indicator question concerns entry to the next level after completion of the preceding one. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    In assessing transition between education levels, authorities must determine the definition should be fixed for the comparison at hand and deviations recorded. For the learners concerned, the decisive consideration is whether analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on matched completion and new-entry populations over coherent periods.[REF-02] [REF-03]

    transition between education levels cannot be judged without identifying results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. For the learners concerned, the decisive consideration is whether if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: capacity constraints distinguished from learner attainment.[REF-04] [REF-06]

    Comparative interpretation of transition between education levels depends upon attrition at any stage can produce an apparently complete indicator from a selective population. This matters because field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether rural learners and those unable to relocate are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-08] [REF-09]

    12

    Completion and educational entitlement

    Institutional action on completion and educational entitlement should be tested against a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The public account remains incomplete unless it explains how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Completion and educational entitlement should be approached as a defined measurement problem. The substantive interest is finishing the final grade or meeting recognised programme requirements, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-02] [REF-03]

    The practical standard for completion and educational entitlement concerns a census figure should disclose enumeration rules. This matters because these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is completion defined separately from sitting or passing an examination. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-04] [REF-06]

    The practical standard for completion and educational entitlement concerns it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. The public account remains incomplete unless it explains how it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that late completion and alternative pathways reported. A responsible commentary distinguishes observation, calculation and interpretation.[REF-08] [REF-09]

    For completion and educational entitlement, the material distinction is between it should involve statistical judgement, legal safeguards and knowledge of the affected community. This matters because suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For over-age learners and people returning after interruption, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-10] [REF-12]

    Part VI

    Learning and assessment

    13

    Minimum learning outcomes

    The practical standard for minimum learning outcomes concerns the study seeks evidence on demonstrated knowledge or skill against a declared domain; it does not infer a learner's circumstances from a national or regional mean. The institutional consequence follows from whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For minimum learning outcomes, the first requirement is conceptual clarity.[REF-04] [REF-06]

    The practical standard for minimum learning outcomes concerns if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. A proportionate conclusion must also recognise that administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is assessment evidence whose population and conditions are known. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-08] [REF-09]

    The practical standard for minimum learning outcomes concerns the comparison should show the level for each group as well as any ratio or gap. The resulting interpretation should show why a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: results interpreted with opportunity to learn and participation.[REF-10] [REF-12]

    Review of minimum learning outcomes is credible only where it explains these populations may be missing not only from good outcomes but from the denominator itself. A contrary reading would overlook that coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners taught in an unfamiliar language.[REF-13] [REF-14]

    14

    Assessment participation

    assessment participation requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because a distribution-sensitive account of who was eligible, present, absent and excluded from testing begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-08] [REF-09]

    The central question in assessment participation is reconciliation should record what each source can and cannot represent. This matters because where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use a participation profile accompanying every result distribution. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-10] [REF-12]

    Review of assessment participation is credible only where it explains disparity measures should not replace the underlying distributions. For the learners concerned, the decisive consideration is whether a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that non-participation never treated as low attainment or ignored.[REF-13] [REF-14]

    The central question in assessment participation is where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. This matters because where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners with disabilities and remote candidates. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness.[REF-15] [REF-16]

    15

    Distribution of achievement

    Comparative interpretation of distribution of achievement depends upon a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A proportionate conclusion must also recognise that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of distribution of achievement lies in making unequal educational experience observable. Here, the relevant phenomenon is variation across the full score or proficiency distribution, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-10] [REF-12]

    distribution of achievement cannot be judged without identifying this distinction is essential for participation and progression. A proportionate conclusion must also recognise that the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on percentiles and threshold shares beside a mean. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-13] [REF-14]

    The central question in distribution of achievement is results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. The evidence must therefore clarify how if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: uncertainty and scale properties stated.[REF-15] [REF-16]

    The central question in distribution of achievement is attrition at any stage can produce an apparently complete indicator from a selective population. For the learners concerned, the decisive consideration is whether field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether learners concentrated below a minimum proficiency threshold are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-17] [REF-19]

    Part VII

    Gender and household resources

    16

    Gender parity and its limits

    A defensible account of gender parity and its limits distinguishes its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns differences between girls and boys in access, progression and learning.[REF-13] [REF-14]

    Institutional action on gender parity and its limits should be tested against its metadata should travel with every published value. The evidence must therefore clarify how at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is female-to-male ratios read with levels and absolute gaps.[REF-15] [REF-16]

    Evidence concerning gender parity and its limits should establish it does not assign cause from a cross-sectional difference. The institutional consequence follows from whether apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that parity not confused with adequacy for either group. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-17] [REF-19]

    For gender parity and its limits, the material distinction is between suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The institutional consequence follows from whether the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For girls in poor rural households and boys exposed to hazardous work, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-20] [REF-21]

    17

    Household wealth gradients

    For household wealth gradients, the material distinction is between the substantive interest is education outcomes across relative household resource groups, observed for a population and period that are stated before calculation. The resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Household wealth gradients should be approached as a defined measurement problem.[REF-15] [REF-16]

    In assessing household wealth gradients, authorities must determine if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. A contrary reading would overlook that administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a documented asset or consumption classification within each setting. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-17] [REF-19]

    In assessing household wealth gradients, authorities must determine statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. This matters because explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: wealth ranks not assumed equivalent across countries or time. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-20] [REF-21]

    Evidence concerning household wealth gradients should establish if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The public account remains incomplete unless it explains how confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children in the poorest quintile and those near classification boundaries. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-23] [REF-24]

    18

    Costs borne by households

    Comparative interpretation of costs borne by households depends upon the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because for costs borne by households, the first requirement is conceptual clarity. The study seeks evidence on fees, materials, transport, clothing and foregone labour; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-17] [REF-19]

    A defensible account of costs borne by households distinguishes counts reveal scale; rates permit comparison; neither is sufficient alone. The evidence must therefore clarify how source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use participation read against direct and indirect education costs. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-20] [REF-21]

    costs borne by households requires a decision about the comparison should identify the reference category but avoid presenting it as a natural norm. The resulting interpretation should show why policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that nominal fee abolition checked against remaining expenditure. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-23] [REF-24]

    The practical standard for costs borne by households concerns the review should record non-response, unknown status and excluded locations separately. The resulting interpretation should show why combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for large families and households hit by economic crisis. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-01] [REF-07]

    Part VIII

    Place and service geography

    19

    Rural and urban residence

    Comparative interpretation of rural and urban residence depends upon without that purpose, disaggregation can multiply figures without improving public judgement. The resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of education outcomes by a declared settlement classification begins by naming the decision the evidence may inform.[REF-20] [REF-21]

    Public responsibility for rural and urban residence begins with the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. For the learners concerned, the decisive consideration is whether when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on residence linked to service availability and travel conditions. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-23] [REF-24]

    Institutional action on rural and urban residence should be tested against results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. The evidence must therefore clarify how if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: national rural definitions preserved and comparison limitations stated.[REF-01] [REF-07]

    Public responsibility for rural and urban residence begins with attrition at any stage can produce an apparently complete indicator from a selective population. A proportionate conclusion must also recognise that field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether remote villages, pastoral populations and peri-urban settlements are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-02] [REF-03]

    20

    Subnational administrative disparity

    subnational administrative disparity requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether the public value of subnational administrative disparity lies in making unequal educational experience observable. Here, the relevant phenomenon is variation between provinces, districts or comparable areas, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-23] [REF-24]

    Review of subnational administrative disparity is credible only where it explains at minimum this includes population, geography, date, collection method, classification and known exclusions. The resulting interpretation should show why a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is area estimates with population size and precision. Its metadata should travel with every published value.[REF-01] [REF-07]

    For subnational administrative disparity, the material distinction is between a responsible commentary distinguishes observation, calculation and interpretation. The evidence must therefore clarify how it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that administrative rankings not mistaken for causal explanations.[REF-02] [REF-03]

    The practical standard for subnational administrative disparity concerns this balance is contextual rather than mechanical. For the learners concerned, the decisive consideration is whether it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For small districts and areas with incomplete reporting, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-04] [REF-06]

    21

    Distance, isolation and transport

    A defensible account of distance, isolation and transport distinguishes a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns physical accessibility of the nearest appropriate service. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-01] [REF-07]

    Evidence concerning distance, isolation and transport should establish the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The evidence must therefore clarify how the preferred construction is travel time, route safety and seasonal interruption rather than straight-line distance alone. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-02] [REF-03]

    The central question in distance, isolation and transport is explanation requires evidence on institutions, resources, households and prior conditions. The institutional consequence follows from whether interpretation follows this limitation: household reports and facility mapping reconciled. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-04] [REF-06]

    distance, isolation and transport requires a decision about if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. For the learners concerned, the decisive consideration is whether confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners with limited mobility and communities cut off seasonally. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-08] [REF-09]

    Part IX

    Disability, language and identity

    22

    Disability-sensitive education data

    Review of disability-sensitive education data is credible only where it explains the measure should preserve both the observed level and the distribution relevant to the claim. For the learners concerned, the decisive consideration is whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Disability-sensitive education data should be approached as a defined measurement problem. The substantive interest is participation and learning by functional difficulty and support requirement, observed for a population and period that are stated before calculation.[REF-02] [REF-03]

    disability-sensitive education data cannot be judged without identifying source coverage must be tested before sources are combined. For the learners concerned, the decisive consideration is whether school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use questions designed for comparable reporting without defining the child by diagnosis alone. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone.[REF-04] [REF-06]

    Public responsibility for disability-sensitive education data begins with disparity measures should not replace the underlying distributions. The evidence must therefore clarify how a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that identification, environment and accommodation kept analytically distinct.[REF-08] [REF-09]

    A defensible account of disability-sensitive education data distinguishes where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. This matters because where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners whose impairments are not recorded by schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness.[REF-10] [REF-12]

    23

    Language of home and instruction

    The practical standard for language of home and instruction concerns the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how for language of home and instruction, the first requirement is conceptual clarity. The study seeks evidence on alignment between learner language, teaching and assessment; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-04] [REF-06]

    The practical standard for language of home and instruction concerns analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This matters because this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on language categories reflecting local use and instructional practice. The definition should be fixed for the comparison at hand and deviations recorded.[REF-08] [REF-09]

    In assessing language of home and instruction, authorities must determine if a conclusion changes under a reasonable specification, that instability is part of the finding. The resulting interpretation should show why national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: small groups not erased through broad national labels. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-10] [REF-12]

    The practical standard for language of home and instruction concerns field arrangements need relevant languages, accessible formats and safe participation. The public account remains incomplete unless it explains how analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether minority-language and multilingual learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-13] [REF-14]

    24

    Ethnicity, indigeneity and protected identity

    In assessing ethnicity, indigeneity and protected identity, authorities must determine without that purpose, disaggregation can multiply figures without improving public judgement. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of disparity associated with historically excluded identity groups begins by naming the decision the evidence may inform.[REF-08] [REF-09]

    The practical standard for ethnicity, indigeneity and protected identity concerns a survey estimate should disclose weights and uncertainty. This matters because a census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is lawful, voluntary and contextually meaningful classification. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions.[REF-10] [REF-12]

    Review of ethnicity, indigeneity and protected identity is credible only where it explains it does not assign cause from a cross-sectional difference. A proportionate conclusion must also recognise that apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that self-identification protected and non-response reported. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-13] [REF-14]

    Public responsibility for ethnicity, indigeneity and protected identity begins with suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. This matters because the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For communities exposed to discrimination or forced assimilation, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-15] [REF-16]

    Part X

    Conflict, disaster and mobility

    25

    Education under conflict and insecurity

    The central question in education under conflict and insecurity is a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. This matters because the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of education under conflict and insecurity lies in making unequal educational experience observable. Here, the relevant phenomenon is access, attendance and learning where violence alters service and movement, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-10] [REF-12]

    education under conflict and insecurity cannot be judged without identifying if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. The resulting interpretation should show why administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is location- and time-specific observation with explicit coverage gaps. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-13] [REF-14]

    education under conflict and insecurity cannot be judged without identifying the comparison should show the level for each group as well as any ratio or gap. The public account remains incomplete unless it explains how a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: absence caused by insecurity distinguished from ordinary dropout.[REF-15] [REF-16]

    Institutional action on education under conflict and insecurity should be tested against coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. A contrary reading would overlook that if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners in insecure areas and host communities. These populations may be missing not only from good outcomes but from the denominator itself.[REF-17] [REF-19]

    27

    Refugees, displaced persons and migrants

    Comparative interpretation of refugees, displaced persons and migrants depends upon a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The public account remains incomplete unless it explains how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Refugees, displaced persons and migrants should be approached as a defined measurement problem. The substantive interest is educational participation across changing legal and residential situations, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-15] [REF-16]

    The central question in refugees, displaced persons and migrants is analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. The institutional consequence follows from whether this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on status, origin, current location and service access recorded separately. The definition should be fixed for the comparison at hand and deviations recorded.[REF-17] [REF-19]

    The central question in refugees, displaced persons and migrants is within-group variation and unmeasured intersecting conditions remain material. A proportionate conclusion must also recognise that the analytical rule is clear: mobility never converted into duplicate enrolment or unexplained disappearance. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-20] [REF-21]

    refugees, displaced persons and migrants cannot be judged without identifying missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. This matters because an adequate equity account asks whether undocumented migrants, refugees and internally displaced learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-23] [REF-24]

    Part XI

    School conditions and teachers

    28

    Teacher availability and distribution

    For teacher availability and distribution, the material distinction is between the study seeks evidence on access to competent teaching across schools and subjects; it does not infer a learner's circumstances from a national or regional mean. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For teacher availability and distribution, the first requirement is conceptual clarity.[REF-17] [REF-19]

    Institutional action on teacher availability and distribution should be tested against its metadata should travel with every published value. For the learners concerned, the decisive consideration is whether at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is teachers present and assigned relative to learner and curriculum need.[REF-20] [REF-21]

    Review of teacher availability and distribution is credible only where it explains a responsible commentary distinguishes observation, calculation and interpretation. This matters because it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that payroll totals not substituted for classroom availability.[REF-23] [REF-24]

    Comparative interpretation of teacher availability and distribution depends upon disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This matters because this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For schools serving poor, remote or displaced communities, a single group label may conceal important internal differences.[REF-01] [REF-07]

    29

    Class size, multi-grade teaching and time

    Public responsibility for class size, multi-grade teaching and time begins with the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how a distribution-sensitive account of the instructional conditions experienced by learners begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]

    class size, multi-grade teaching and time requires a decision about if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. The public account remains incomplete unless it explains how administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is class organisation, scheduled time and delivered time considered together. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-23] [REF-24]

    For class size, multi-grade teaching and time, the material distinction is between explanation requires evidence on institutions, resources, households and prior conditions. A contrary reading would overlook that interpretation follows this limitation: simple pupil-teacher ratios not treated as a complete quality measure. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-01] [REF-07]

    Review of class size, multi-grade teaching and time is credible only where it explains coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. This matters because if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include early grades and mixed-age classes. These populations may be missing not only from good outcomes but from the denominator itself.[REF-02] [REF-03]

    30

    Materials, facilities and basic services

    Evidence concerning materials, facilities and basic services should establish a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. This matters because the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of materials, facilities and basic services lies in making unequal educational experience observable. Here, the relevant phenomenon is usable learning resources and safe, accessible school conditions, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-23] [REF-24]

    Comparative interpretation of materials, facilities and basic services depends upon school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. The institutional consequence follows from whether reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use availability joined to condition, accessibility and regular use. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined.[REF-01] [REF-07]

    The central question in materials, facilities and basic services is disparity measures should not replace the underlying distributions. The institutional consequence follows from whether a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that delivery counts checked against learner access.[REF-02] [REF-03]

    materials, facilities and basic services cannot be judged without identifying the review should record non-response, unknown status and excluded locations separately. The institutional consequence follows from whether combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in temporary or damaged premises. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-04] [REF-06]

    Part XII

    Finance and distribution

    31

    Public spending by level and function

    The central question in public spending by level and function is the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because the indicator question concerns resources assigned to education purposes across the system. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    Comparative interpretation of public spending by level and function depends upon analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. A contrary reading would overlook that this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on expenditure classified by level, recurrent or capital use and responsible body. The definition should be fixed for the comparison at hand and deviations recorded.[REF-02] [REF-03]

    A defensible account of public spending by level and function distinguishes if a conclusion changes under a reasonable specification, that instability is part of the finding. The resulting interpretation should show why national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: budgets, commitments and actual expenditure distinguished. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-04] [REF-06]

    For public spending by level and function, the material distinction is between missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. This matters because an adequate equity account asks whether basic education services under fiscal pressure are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-08] [REF-09]

    32

    Incidence of education spending

    incidence of education spending requires a decision about the substantive interest is who benefits from publicly financed places and services, observed for a population and period that are stated before calculation. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Incidence of education spending should be approached as a defined measurement problem.[REF-02] [REF-03]

    A defensible account of incidence of education spending distinguishes these are substantive attributes because they determine who can appear in the evidence. The evidence must therefore clarify how a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is unit resources combined with participation across population groups. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-04] [REF-06]

    The practical standard for incidence of education spending concerns it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. The institutional consequence follows from whether it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that benefit estimates not treated as household income. A responsible commentary distinguishes observation, calculation and interpretation.[REF-08] [REF-09]

    A defensible account of incidence of education spending distinguishes this balance is contextual rather than mechanical. The evidence must therefore clarify how it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For groups excluded before public spending can reach them, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-10] [REF-12]

    33

    Protecting equity during fiscal constraint

    The central question in protecting equity during fiscal constraint is the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For protecting equity during fiscal constraint, the first requirement is conceptual clarity. The study seeks evidence on whether reductions or delays fall disproportionately on weaker services; it does not infer a learner's circumstances from a national or regional mean.[REF-04] [REF-06]

    The central question in protecting equity during fiscal constraint is numerator, denominator, reference date, unit and exclusions should appear together. A contrary reading would overlook that if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is dated finance and service indicators read together.[REF-08] [REF-09]

    Review of protecting equity during fiscal constraint is credible only where it explains a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. For the learners concerned, the decisive consideration is whether reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: national totals tested against subnational allocation and household costs. The comparison should show the level for each group as well as any ratio or gap.[REF-10] [REF-12]

    A defensible account of protecting equity during fiscal constraint distinguishes confidentiality is essential, especially where identity or status creates risk. The evidence must therefore clarify how protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include poor households and institutions with little financial reserve. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-13] [REF-14]

    Part XIII

    Data sources and measurement error

    34

    Administrative records

    Review of administrative records is credible only where it explains the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because a distribution-sensitive account of regular learner, staff, facility and finance information begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-08] [REF-09]

    administrative records requires a decision about school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. The resulting interpretation should show why reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use clear definitions, reporting coverage and revision history. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined.[REF-10] [REF-12]

    Evidence concerning administrative records should establish the comparison should identify the reference category but avoid presenting it as a natural norm. The institutional consequence follows from whether policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that non-reporting institutions kept visible in aggregates. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-13] [REF-14]

    administrative records requires a decision about where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. For the learners concerned, the decisive consideration is whether particular scrutiny is required for small, private, non-formal and emergency providers. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-15] [REF-16]

    35

    Household surveys

    Review of household surveys is credible only where it explains the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of household surveys lies in making unequal educational experience observable. Here, the relevant phenomenon is population-based evidence beyond enrolled learners, not the administrative convenience of the available categories.[REF-10] [REF-12]

    household surveys requires a decision about when several sources exist, consistency is evidence to consider, not proof that common error is absent. The evidence must therefore clarify how a defensible statistic would be based on probability samples, weights and field dates documented. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-13] [REF-14]

    In assessing household surveys, authorities must determine if a conclusion changes under a reasonable specification, that instability is part of the finding. The evidence must therefore clarify how national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: sampling and non-response uncertainty carried into group comparisons. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-15] [REF-16]

    Public responsibility for household surveys begins with attrition at any stage can produce an apparently complete indicator from a selective population. A proportionate conclusion must also recognise that field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether small minorities and mobile households are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-17] [REF-19]

    36

    Censuses and population frames

    The practical standard for censuses and population frames concerns the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether the indicator question concerns broad population coverage and small-area denominators. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-13] [REF-14]

    In assessing censuses and population frames, authorities must determine a census figure should disclose enumeration rules. For the learners concerned, the decisive consideration is whether these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is enumeration date, usual residence and institutional coverage stated. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-15] [REF-16]

    Comparative interpretation of censuses and population frames depends upon it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. The public account remains incomplete unless it explains how it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that long intervals and under-enumeration acknowledged. A responsible commentary distinguishes observation, calculation and interpretation.[REF-17] [REF-19]

    In assessing censuses and population frames, authorities must determine it should involve statistical judgement, legal safeguards and knowledge of the affected community. The evidence must therefore clarify how suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For homeless, displaced and geographically isolated people, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-20] [REF-21]

    Part XIV

    Disaggregation and intersection

    37

    Single-axis disaggregation

    Review of single-axis disaggregation is credible only where it explains the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Single-axis disaggregation should be approached as a defined measurement problem. The substantive interest is separate reporting by sex, wealth, residence or another characteristic, observed for a population and period that are stated before calculation.[REF-15] [REF-16]

    Evidence concerning single-axis disaggregation should establish it should require examination of definitions, timing, migration, duplication and non-response. The evidence must therefore clarify how the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is levels, gaps and denominators shown for each category. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-17] [REF-19]

    Evidence concerning single-axis disaggregation should establish the comparison should show the level for each group as well as any ratio or gap. A contrary reading would overlook that a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: one axis not presented as a complete account of marginalisation.[REF-20] [REF-21]

    single-axis disaggregation cannot be judged without identifying if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. This matters because confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include groups whose disadvantage lies on another unmeasured dimension. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-23] [REF-24]

    38

    Intersecting categories

    A defensible account of intersecting categories distinguishes a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The institutional consequence follows from whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For intersecting categories, the first requirement is conceptual clarity. The study seeks evidence on joint distributions such as sex by wealth and residence; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-17] [REF-19]

    Public responsibility for intersecting categories begins with reconciliation should record what each source can and cannot represent. For the learners concerned, the decisive consideration is whether where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use pre-specified combinations with sufficient observations. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-20] [REF-21]

    The practical standard for intersecting categories concerns a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. A contrary reading would overlook that percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that empty or unstable cells reported honestly. Disparity measures should not replace the underlying distributions.[REF-23] [REF-24]

    Public responsibility for intersecting categories begins with combining unknown observations with the majority group biases both estimates and obscures the weakness. This matters because where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for poor rural girls, disabled learners in remote areas and displaced minorities. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-01] [REF-07]

    39

    Small numbers, disclosure and reliability

    small numbers, disclosure and reliability cannot be judged without identifying the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because a distribution-sensitive account of useful detail without unreliable estimates or identification begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]

    Evidence concerning small numbers, disclosure and reliability should establish this distinction is essential for participation and progression. A proportionate conclusion must also recognise that the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on suppression, aggregation or qualitative evidence chosen proportionately. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-23] [REF-24]

    For small numbers, disclosure and reliability, the material distinction is between nor should a group estimate be read as a description of every member. This matters because within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: confidentiality decisions separated from claims that no disparity exists. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-01] [REF-07]

    In assessing small numbers, disclosure and reliability, authorities must determine analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The evidence must therefore clarify how missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether small communities and learners with rare characteristics are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-02] [REF-03]

    Part XV

    Comparison, uncertainty and change

    40

    Comparing unlike systems

    The central question in comparing unlike systems is the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of comparing unlike systems lies in making unequal educational experience observable. Here, the relevant phenomenon is cross-country patterns based on harmonised but bounded concepts, not the administrative convenience of the available categories.[REF-23] [REF-24]

    For comparing unlike systems, the material distinction is between these are substantive attributes because they determine who can appear in the evidence. The institutional consequence follows from whether a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is metadata tests before numerical comparison. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-01] [REF-07]

    Comparative interpretation of comparing unlike systems depends upon it does not assign cause from a cross-sectional difference. The institutional consequence follows from whether apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that differences in programme structure and classification remain visible. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-02] [REF-03]

    In assessing comparing unlike systems, authorities must determine the public report can still state that a disparity was examined, whether action is required and which body will monitor it. The evidence must therefore clarify how for countries with incomplete or rapidly changing systems, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-04] [REF-06]

    41

    Sampling error and other uncertainty

    A defensible account of sampling error and other uncertainty distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because the indicator question concerns the range of values reasonably compatible with the observations. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    Review of sampling error and other uncertainty is credible only where it explains the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. A proportionate conclusion must also recognise that the preferred construction is standard errors, design effects and data-quality qualifications. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-02] [REF-03]

    Evidence concerning sampling error and other uncertainty should establish where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. A proportionate conclusion must also recognise that statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: rank differences smaller than uncertainty not interpreted. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning.[REF-04] [REF-06]

    Review of sampling error and other uncertainty is credible only where it explains these populations may be missing not only from good outcomes but from the denominator itself. A proportionate conclusion must also recognise that coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include small disaggregated populations.[REF-08] [REF-09]

    Part XVI

    Responsible interpretation and action

    43

    Reading disparity without blaming learners

    Comparative interpretation of reading disparity without blaming learners depends upon the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For reading disparity without blaming learners, the first requirement is conceptual clarity. The study seeks evidence on institutional and social conditions associated with unequal outcomes; it does not infer a learner's circumstances from a national or regional mean.[REF-04] [REF-06]

    For reading disparity without blaming learners, the material distinction is between when several sources exist, consistency is evidence to consider, not proof that common error is absent. The institutional consequence follows from whether a defensible statistic would be based on descriptive findings separated from causal claims. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-08] [REF-09]

    reading disparity without blaming learners cannot be judged without identifying if a conclusion changes under a reasonable specification, that instability is part of the finding. This matters because national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: group identity never treated as a mechanism by itself. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-10] [REF-12]

    Comparative interpretation of reading disparity without blaming learners depends upon field arrangements need relevant languages, accessible formats and safe participation. A contrary reading would overlook that analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether communities subject to stigma are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-13] [REF-14]

    44

    Turning evidence into equitable policy

    Comparative interpretation of turning evidence into equitable policy depends upon a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of a decision rule linking disparity to service, finance or legal responsibility begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]

    In assessing turning evidence into equitable policy, authorities must determine its metadata should travel with every published value. The resulting interpretation should show why at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is baseline, intended reach, implementation evidence and review date.[REF-10] [REF-12]

    The central question in turning evidence into equitable policy is it does not assign cause from a cross-sectional difference. The evidence must therefore clarify how apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that targets accompanied by distributional safeguards. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-13] [REF-14]

    Comparative interpretation of turning evidence into equitable policy depends upon the public report can still state that a disparity was examined, whether action is required and which body will monitor it. This matters because for learners farthest below a secured minimum, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-15] [REF-16]

    45

    A public account beyond the average

    Public responsibility for a public account beyond the average begins with the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how the public value of a public account beyond the average lies in making unequal educational experience observable. Here, the relevant phenomenon is a concise national statement of level, distribution, missingness and remedy, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-10] [REF-12]

    A defensible account of a public account beyond the average distinguishes a discrepancy is not resolved by selecting the more favourable source. For the learners concerned, the decisive consideration is whether it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is national totals presented beside selected group and place indicators. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-13] [REF-14]

    Review of a public account beyond the average is credible only where it explains the comparison should show the level for each group as well as any ratio or gap. A proportionate conclusion must also recognise that a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: progress claims bounded by evidence coverage and unresolved gaps.[REF-15] [REF-16]

    Evidence concerning a public account beyond the average should establish these populations may be missing not only from good outcomes but from the denominator itself. For the learners concerned, the decisive consideration is whether coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include every learner otherwise hidden by a successful average.[REF-17] [REF-19]

    Part XVII

    Applied distributional analysis

    46

    Composite measures and the loss of meaning

    The practical standard for composite measures and the loss of meaning concerns where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. For the learners concerned, the decisive consideration is whether the analytical purpose is combining several dimensions into a summary measure. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises participation, completion, learning and school conditions, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-01] [REF-03]

    For composite measures and the loss of meaning, the material distinction is between documentation should also state which observations are direct, which are estimated and which are unavailable. The public account remains incomplete unless it explains how missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern weights, normalisation and substitution. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable.[REF-02] [REF-04]

    composite measures and the loss of meaning cannot be judged without identifying analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A proportionate conclusion must also recognise that a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is a low value in one dimension being concealed by a high value in another. Avoiding it requires the underlying counts and distributions to remain visible beside any summary.[REF-05] [REF-06]

    Review of composite measures and the loss of meaning is credible only where it explains censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. For the learners concerned, the decisive consideration is whether a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response.[REF-08] [REF-12]

    composite measures and the loss of meaning cannot be judged without identifying analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. This matters because small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins.[REF-13] [REF-16]

    The central question in composite measures and the loss of meaning is every response should name the expected population reach and the later observation that will test it. For the learners concerned, the decisive consideration is whether otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is policy-makers deciding whether a summary aids or obscures resource allocation. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves.[REF-17] [REF-19]

    A defensible account of composite measures and the loss of meaning distinguishes the two forms answer different questions. A contrary reading would overlook that authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier.[REF-21] [REF-23]

    The central question in composite measures and the loss of meaning is it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. This matters because changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement.[REF-23] [REF-24]

    47

    Decomposing an observed education gap

    Review of decomposing an observed education gap is credible only where it explains where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. A contrary reading would overlook that where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is examining how a national disparity is distributed across places and population groups. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises within-group and between-group differences, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-02] [REF-04]

    Comparative interpretation of decomposing an observed education gap depends upon choices on these matters are not neutral presentation details. The institutional consequence follows from whether they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern population shares, outcome levels and overlapping membership.[REF-05] [REF-06]

    Review of decomposing an observed education gap is credible only where it explains avoiding it requires the underlying counts and distributions to remain visible beside any summary. A contrary reading would overlook that analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is treating a descriptive decomposition as proof of cause.[REF-08] [REF-12]

    Public responsibility for decomposing an observed education gap begins with credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. A proportionate conclusion must also recognise that every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities locating where further enquiry and action are warranted. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-21] [REF-23]

    48

    Setting distribution-sensitive targets

    Public responsibility for setting distribution-sensitive targets begins with where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. The resulting interpretation should show why where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is expressing progress as improvement in secured minimums and unjustified gaps. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises the national level, the least-served group and the lower tail of the distribution, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-05] [REF-06]

    Evidence concerning setting distribution-sensitive targets should establish sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. A proportionate conclusion must also recognise that documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern baseline stability, ambition and safeguards against exclusion. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported.[REF-08] [REF-12]

    Institutional action on setting distribution-sensitive targets should be tested against comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. The evidence must therefore clarify how no result should be described as equitable solely because one relative measure improved. The main interpretive danger is meeting a mean target while abandoning those farthest behind. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-13] [REF-16]

    The practical standard for setting distribution-sensitive targets concerns every response should name the expected population reach and the later observation that will test it. The public account remains incomplete unless it explains how otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is governments linking national commitments to subnational delivery. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves.[REF-23] [REF-24]

    49

    Linking learners to service geography

    For linking learners to service geography, the material distinction is between that purpose should be written before the calculation because method follows the intended inference. The institutional consequence follows from whether the evidence base comprises settlement populations, travel conditions, schools, teachers and programme levels, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is relating participation and learning to the location and capacity of education services.[REF-08] [REF-12]

    The practical standard for linking learners to service geography concerns they determine how strongly one component, place or population can influence the conclusion. The public account remains incomplete unless it explains how a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern geographical scale, boundary effects and facility catchments. Choices on these matters are not neutral presentation details.[REF-13] [REF-16]

    For linking learners to service geography, the material distinction is between no result should be described as equitable solely because one relative measure improved. The resulting interpretation should show why the main interpretive danger is assuming the nearest mapped institution is accessible or appropriate. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity.[REF-17] [REF-19]

    Review of linking learners to service geography is credible only where it explains evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The institutional consequence follows from whether the certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is planners choosing sites, transport support and teacher deployment.[REF-01] [REF-03]

    50

    Reconciling conflicting sources

    A defensible account of reconciling conflicting sources distinguishes the evidence base comprises coverage, timing, concepts and reporting incentives, each of which describes a different feature of educational opportunity. For the learners concerned, the decisive consideration is whether a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is interpreting differences between administrative, survey and census estimates. That purpose should be written before the calculation because method follows the intended inference.[REF-13] [REF-16]

    A defensible account of reconciling conflicting sources distinguishes missing values should remain missing unless an explicit estimation method and its effect are shown. A proportionate conclusion must also recognise that the principal methodological questions concern a documented comparison of population and variable definitions. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable.[REF-17] [REF-19]

    The practical standard for reconciling conflicting sources concerns a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. For the learners concerned, the decisive consideration is whether direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is averaging incompatible estimates into an apparently precise figure. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold.[REF-21] [REF-23]

    Institutional action on reconciling conflicting sources should be tested against evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The evidence must therefore clarify how the certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is statistical authorities issuing one bounded account with visible uncertainty.[REF-02] [REF-04]

    51

    Monitoring marginalisation during severe disruption

    The central question in monitoring marginalisation during severe disruption is that purpose should be written before the calculation because method follows the intended inference. For the learners concerned, the decisive consideration is whether the evidence base comprises rapid counts, restored administrative returns and household evidence, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is maintaining useful distributional evidence when populations and services move rapidly.[REF-17] [REF-19]

    Review of monitoring marginalisation during severe disruption is credible only where it explains they determine how strongly one component, place or population can influence the conclusion. A contrary reading would overlook that a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern dated estimates, revision practice and minimal essential classifications. Choices on these matters are not neutral presentation details.[REF-21] [REF-23]

    Evidence concerning monitoring marginalisation during severe disruption should establish comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. The evidence must therefore clarify how no result should be described as equitable solely because one relative measure improved. The main interpretive danger is using an unstable emergency denominator to assert durable improvement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-23] [REF-24]

    Review of monitoring marginalisation during severe disruption is credible only where it explains credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. A proportionate conclusion must also recognise that every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities protecting access while rebuilding regular statistics. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-05] [REF-06]

    52

    Communicating uncertainty without losing urgency

    communicating uncertainty without losing urgency requires a decision about where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The institutional consequence follows from whether the analytical purpose is explaining what is known strongly enough to justify action and what remains unresolved. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises point estimates, ranges, quality statements and missing populations, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-21] [REF-23]

    Comparative interpretation of communicating uncertainty without losing urgency depends upon choices on these matters are not neutral presentation details. The institutional consequence follows from whether they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern plain institutional language joined to exact metadata.[REF-23] [REF-24]

    The central question in communicating uncertainty without losing urgency is comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. A proportionate conclusion must also recognise that no result should be described as equitable solely because one relative measure improved. The main interpretive danger is presenting caution as a reason for inaction or urgency as a reason for overstatement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-01] [REF-03]

    Evidence concerning communicating uncertainty without losing urgency should establish credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. A contrary reading would overlook that every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is the public, affected communities and responsible decision-makers. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-08] [REF-12]

    53

    A national marginalisation profile

    Comparative interpretation of a national marginalisation profile depends upon where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. The resulting interpretation should show why where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is assembling a concise recurring account beyond the national average. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises population, access, progression, learning, conditions, finance and unresolved evidence gaps, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-23] [REF-24]

    The practical standard for a national marginalisation profile concerns they determine how strongly one component, place or population can influence the conclusion. The evidence must therefore clarify how a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a stable core with context-specific distributions. Choices on these matters are not neutral presentation details.[REF-01] [REF-03]

    A defensible account of a national marginalisation profile distinguishes no result should be described as equitable solely because one relative measure improved. The institutional consequence follows from whether the main interpretive danger is creating an encyclopaedia of indicators without decision priority. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity.[REF-02] [REF-04]

    a national marginalisation profile cannot be judged without identifying otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. For the learners concerned, the decisive consideration is whether the policy user is parliament, ministries, local authorities and communities reviewing educational equity. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-13] [REF-16]

    Part XVIII

    Extended disparity interpretation

    54

    Minimum comparison threshold

    A defensible account of minimum comparison threshold distinguishes opportunity to learn, participation and exclusions remain material. A proportionate conclusion must also recognise that a learning-disparity comparison requires a declared construct, represented population, assessment conditions, scale, uncertainty and distribution. National means should be accompanied by lower-tail, threshold and group evidence.[REF-02] [REF-03]

    55

    Within-system interpretation

    Institutional action on within-system interpretation should be tested against counts, levels, absolute gaps and ratios answer different questions. A proportionate conclusion must also recognise that missing learners and non-participating schools must remain visible. Within-system gaps should preserve place, group and institutional context without assigning cause from identity.[REF-01] [REF-06]

    56

    Between-system interpretation

    Evidence concerning between-system interpretation should establish for the learners concerned, the decisive consideration is whether harmonisation does not remove substantive system difference. This matters because between-system comparison requires metadata tests for curriculum, age, language, sampling and assessment. The central question in between-system interpretation is rank differences smaller than uncertainty should not support categorical conclusions.[REF-02] [REF-03]

    Part XIX

    Extended analysis of learning disparities

    57

    Assessment participation and the represented learning population

    The central question in assessment participation and the represented learning population is a result calculated only for participating learners may describe their performance accurately while failing to describe the educational system's full learner population. A proportionate conclusion must also recognise that a learning distribution is defined partly by who participates in the assessment. The target population, eligible population, sampled population and assessed population should be reported separately. Learners can be absent because of illness, displacement, school non-attendance, language barriers, disability, conflict, administrative exclusion or ordinary sampling loss. These routes have different meanings.[REF-01] [REF-02] [REF-03]

    Evidence concerning assessment participation and the represented learning population should establish if institutions outside the frame differ systematically from included schools, a high learner response rate within sampled schools cannot repair the coverage limitation. The evidence must therefore clarify how similarly, replacement of inaccessible schools can preserve sample size while changing the population represented. Reports should describe replacement rules and show any material effect on geography or institutional type. Participation should therefore accompany every reported mean, proficiency share or percentile. School exclusion and within-school absence need separate observation where the design permits.

    The central question in assessment participation and the represented learning population is where a numerical adjustment is not defensible, the limitation remains a substantive finding. The resulting interpretation should show why the absence of learning evidence for a population requiring education attention should not be interpreted as evidence of no disparity. Non-participation should not be assigned a score. Coding an absent learner as below threshold invents performance evidence; removing every absence without comment creates a different bias. Sensitivity analysis can examine plausible bounds or compare known characteristics of participants and non-participants.

    In assessing assessment participation and the represented learning population, authorities must determine where an adapted form changes the construct materially, separate interpretation may be necessary. The institutional consequence follows from whether where it removes an irrelevant barrier, results may belong in the common distribution. The report should explain that judgement rather than treat all adaptations as either incomparable or automatically identical. Accommodation and language arrangements affect participation and result validity. An assessment may permit attendance yet fail to elicit the intended construct if the format, communication or response mode is inaccessible. Exemption practices should be reported by reason and learner group.[REF-05] [REF-10] [REF-12]

    assessment participation and the represented learning population requires a decision about the responsible education body should identify which population remains unrepresented and when better evidence will be available. A contrary reading would overlook that a national learning claim should be no broader than that coverage. Public interpretation should connect participation to policy. A low response among remote schools may require field and service improvement; absence among out-of-school children may require a population-based study and re-entry action; assessment exclusion may require accessible design. These are not corrections to be made solely through statistical weighting.

    58

    Scale, threshold and distribution comparability

    In assessing scale, threshold and distribution comparability, authorities must determine a numerical score has meaning through the tasks, response model, scoring rules and population for which interpretation has been supported. This matters because equal numerical differences should not be assumed to represent equal educational differences unless the scale warrants that inference. A threshold adds a substantive judgement about the knowledge or capability learners should demonstrate; it should not be selected merely because it divides the sample conveniently. Comparisons require a clear statement of what the assessment scale represents.[REF-02] [REF-03]

    The practical standard for scale, threshold and distribution comparability concerns the lower part of the distribution is especially important for minimum learning opportunity, while the upper part may reveal whether expansion has altered advanced performance. A proportionate conclusion must also recognise that these summaries should remain tied to uncertainty and assessment coverage. Means provide one description of the centre and can conceal change elsewhere. The same mean can accompany a compressed distribution, a wide lower tail or polarisation. Reports should therefore consider percentiles, threshold shares and dispersion where technically sound.

    A defensible account of scale, threshold and distribution comparability distinguishes if standards are reset, the break should be visible at the year of change and parallel results reported where possible. For the learners concerned, the decisive consideration is whether descriptions such as basic, adequate or advanced should be treated as definitions under the assessment, not universal attributes of a learner or education system. Threshold comparisons need stable standard-setting and clear labels. A change in the percentage above a threshold may reflect learning, scale revision, task composition or population change.

    The central question in scale, threshold and distribution comparability is review should examine translation, adaptation, differential item behaviour and opportunity to learn, while avoiding the claim that every detected difference invalidates the whole assessment. The evidence must therefore clarify how some comparisons may remain defensible at a broad domain while narrower subscales do not. The permissible inference should be stated accordingly. Cross-language and cross-cultural comparability require evidence, not an assumption that translation has preserved difficulty and meaning. Task familiarity, curriculum exposure and response conventions can affect results.

    Public responsibility for scale, threshold and distribution comparability begins with small rank movement can follow changes in participating systems or sampling variation. For the learners concerned, the decisive consideration is whether public reporting should avoid categorical language when intervals overlap or scale linkage is weak. The useful question is whether the evidence indicates a material disparity requiring enquiry, what population is affected and which educational condition might be changed. League order does not answer those questions. A rank is usually less informative than the estimated difference, uncertainty and distribution.

    59

    Opportunity to learn and interpretation of achievement gaps

    A defensible account of opportunity to learn and interpretation of achievement gaps distinguishes opportunity does not determine performance completely, and its measurement is imperfect. A proportionate conclusion must also recognise that it nevertheless prevents a learning gap from being attributed solely to learners or households when institutions supplied different educational conditions. Achievement evidence should be interpreted with opportunity to learn. This includes curriculum entitlement, content actually taught, instructional time, teacher availability, language, materials, attendance and access to support.[REF-01] [REF-13] [REF-14]

    opportunity to learn and interpretation of achievement gaps requires a decision about teacher reports, schedules, classroom observation and learner work can add evidence, each with limitations. A contrary reading would overlook that teacher self-report may be influenced by recall or expectations; observation covers a short period; work samples are selective. Agreement across sources strengthens interpretation when dates and populations align. Contradictions can identify local variation or weak measurement and should not be resolved by choosing the account most favourable to the system. The official curriculum establishes intended opportunity, not delivered instruction.

    Public responsibility for opportunity to learn and interpretation of achievement gaps begins with time is therefore one component of the explanatory evidence, not a conversion factor for predicted score. A proportionate conclusion must also recognise that instructional time should distinguish scheduled, delivered and attended time. Closures, teacher absence, shortened shifts and late entry can reduce delivered or usable time. Total hours also conceal subject allocation and teaching quality. A learner receiving more hours of poorly organised instruction does not necessarily have greater opportunity in the intended domain.

    Evidence concerning opportunity to learn and interpretation of achievement gaps should establish national resource averages can obscure concentration of weak provision among the same learners whose scores form the lower tail. The institutional consequence follows from whether resource indicators should be connected to use. Textbooks delivered to a school may not be available in the relevant language, grade or classroom. Teacher qualifications recorded administratively may not match subject assignment. Facilities may exist but be inaccessible or unsafe. These conditions should be examined at the level where the learning evidence was collected.

    The practical standard for opportunity to learn and interpretation of achievement gaps concerns where learners received less curriculum exposure, a response may include additional teaching, staff deployment, accessible materials or revised pacing. The institutional consequence follows from whether later assessment should test whether opportunity and learning changed. The original disparity and the remedial conditions should remain visible rather than being erased by a new cohort average. Policy conclusions should avoid treating opportunity indicators as excuses for low expectations. Their purpose is to identify conditions within public responsibility and to design support.

    60

    Decomposing disparities within education systems

    Institutional action on decomposing disparities within education systems should be tested against a large within-school component can coexist with institutional inequality and should not be read as evidence that schools are irrelevant. For the learners concerned, the decisive consideration is whether a national disparity can reflect differences between regions, schools, classrooms and learners. Decomposition can describe where variation is concentrated, provided it is not treated as proof of cause. A large between-school component may indicate segregation, resource distribution or residential pattern; it does not identify which mechanism operates.[REF-01] [REF-09] [REF-14]

    For decomposing disparities within education systems, the material distinction is between learners are commonly nested within classes and schools, while policies may operate through districts or providers. The institutional consequence follows from whether standard errors and models should respect clustering. Small schools and sparsely populated areas may require special treatment, but removal changes the population represented. Reports should state exclusions and avoid presenting a modelled residual as direct observation. The level of analysis should correspond to the sampling and decision structure.

    Evidence concerning decomposing disparities within education systems should establish they should not replace the unadjusted learner outcome or become a definitive quality rank. A proportionate conclusion must also recognise that adjustment for a condition influenced by the school can also remove part of the very effect under review. The purpose and causal assumptions need explanation. Group composition affects school comparisons. A raw school mean combines prior opportunity, intake, mobility, attendance and current teaching. Adjusted measures can answer bounded questions but depend on variables and assumptions.

    decomposing disparities within education systems cannot be judged without identifying maps and rankings, where used elsewhere, should not expose small communities or imply a boundary creates the disparity. The public account remains incomplete unless it explains how geographical decomposition should preserve absolute numbers and service context. A small district with a severe gap may need urgent support even though it contributes little to national variance. A populous area with a modest gap may represent many learners. Policy priority should consider educational severity, population, rights and feasibility rather than statistical contribution alone.

    For decomposing disparities within education systems, the material distinction is between between-region evidence may lead to allocation review; between-school evidence may lead to staffing, admissions or support enquiry; within-school evidence may lead to classroom, language or accessibility review. This matters because each hypothesis requires additional evidence. The decomposition locates questions; it does not authorise blame. Follow-up should state the condition examined, action taken and later learning evidence. The appropriate outcome is a decision agenda.

    61

    Bounded conclusions between education systems

    Evidence concerning bounded conclusions between education systems should establish metadata review should precede numerical comparison. The institutional consequence follows from whether where a material difference cannot be reconciled, the systems may still be described separately without a common rank. Between-system comparison serves public learning when it identifies patterns, plausible questions and alternative institutional arrangements. It becomes misleading when harmonised labels conceal different programme structures, ages, curricula, languages, participation or assessment conditions.[REF-02] [REF-03] [REF-16]

    Public responsibility for bounded conclusions between education systems begins with group composition and population coverage matter. A contrary reading would overlook that an apparent national advantage may not extend to poor, rural, minority-language or disabled learners. Reports should place distributional evidence beside the system result and avoid using nationality as an explanation. Country and system averages should not be interpreted as attributes of every school or learner. Within-system distributions can overlap substantially even where means differ.

    Review of bounded conclusions between education systems is credible only where it explains a later higher score should not automatically be described as system improvement where the represented population changed materially. The public account remains incomplete unless it explains how temporal comparison requires stable linkage. Changes in curriculum, assessment mode, participation, sampling frame or system boundaries can create discontinuity. A linked scale can support trend if common items or other methods preserve meaning and security, but linkage error should accompany the estimate.

    bounded conclusions between education systems cannot be judged without identifying the comparison can identify an option, not guarantee its effect. A proportionate conclusion must also recognise that pilots should state the mechanism and review evidence. Adoption based solely on rank proximity or reputation substitutes imitation for analysis. Policy borrowing should attend to authority, capacity, sequence and context. A practice associated with high performance elsewhere may depend upon teacher preparation, finance, curriculum coherence or social conditions absent in the receiving system.

    Comparative interpretation of bounded conclusions between education systems depends upon policy consideration proposes further enquiry or action under national authority. The evidence must therefore clarify how causal judgement requires additional design. This separation permits strong public concern about a learning disparity without false certainty about its source and protects education systems from both complacency and unsupported prescription. The final international statement should distinguish three levels. Recorded fact describes the observed results under stated methods. Interpretation identifies patterns and limitations.

    62

    Uncertainty, materiality and the duty to respond

    uncertainty, materiality and the duty to respond requires a decision about the public report should identify which uncertainty could alter the decision and which does not affect the direction of urgent protection. The public account remains incomplete unless it explains how uncertainty should qualify a disparity claim without neutralising it. Sampling error, non-response, scale linkage, classification and model choice each affect the range of defensible conclusions. They should be described separately because they support different remedies. A larger sample may reduce sampling error; it does not correct systematic exclusion or an invalid construct. More decimal places cannot repair weak coverage.[REF-02] [REF-03]

    The practical standard for uncertainty, materiality and the duty to respond concerns a small estimated difference may be precise but have limited practical consequence, while an uncertain large gap affecting a protected minimum may require immediate enquiry. A proportionate conclusion must also recognise that materiality should consider the knowledge or capability involved, the number of learners, distribution, duration and consequences for later progression. The threshold for action also depends upon reversibility. Additional diagnostic support can be introduced and reviewed more readily than a high-stakes classification of schools or learners. Statistical significance should not substitute for educational materiality.

    For uncertainty, materiality and the duty to respond, the material distinction is between where a group disparity persists under several definitions, investigate institutional mechanisms without attributing cause to identity. A contrary reading would overlook that every response should name the competent body, intended population, resources and review date. A remedy should match the evidence. Where participation is selective, improve coverage and examine barriers. Where opportunity to learn differs, address time, teachers, curriculum, language, accessibility or materials. Where scale comparability is weak, improve the assessment before publishing ranks.

    For uncertainty, materiality and the duty to respond, the material distinction is between a national evidence system demonstrates strength when it can acknowledge uncertainty, improve measurement and amend policy. The resulting interpretation should show why the final test is whether learners receive better educational opportunity and whether remaining disparities continue to be visible rather than whether one annual figure becomes more favourable. Later evidence should be capable of changing the conclusion. Reports should preserve the original estimate, method and limitation, then state why revision occurred. Silent replacement makes apparent improvement impossible to distinguish from correction.

    Part XX

    Targeted improvement planning

    63

    Defining a persistent learning gap

    Institutional action on defining a persistent learning gap should be tested against the plan should state which observation is direct, which is estimated and which remains unknown. The resulting interpretation should show why the improvement question concerns a sustained disparity in a declared learning domain, population and period. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity. The baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined.[REF-01] [REF-02]

    In assessing defining a persistent learning gap, authorities must determine achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. The institutional consequence follows from whether authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is a fluctuating score or one cohort difference being treated as proof of persistence. Avoiding it requires evidence on both learning and the opportunity supplied.[REF-03] [REF-06]

    Institutional action on defining a persistent learning gap should be tested against administrative records can describe staffing and participation while omitting non-enrolled learners. A proportionate conclusion must also recognise that assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes baseline, assessment coverage, uncertainty, distribution and opportunity to learn. Each source should be used within scope.[REF-08] [REF-09]

    In assessing defining a persistent learning gap, authorities must determine local adaptation should remain possible within a common substantive condition. The public account remains incomplete unless it explains how any departure should be recorded with its reason and expected learner consequence. The governing action is to select a gap that is educationally material and within public influence. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism.[REF-13] [REF-14]

    Review of defining a persistent learning gap is credible only where it explains a plan can improve its average by reaching learners closest to a threshold while leaving those farthest behind. The public account remains incomplete unless it explains how targets should therefore include the least-served position and protection against exclusion. Small groups require confidentiality and careful precision, not disappearance from review. Distributional review should ask who is eligible, offered support, participates, receives the intended intensity and demonstrates a later response. Gender, household resources, disability, language, residence and prior opportunity may intersect.[REF-17] [REF-19]

    The central question in defining a persistent learning gap is a plan relying on a few exceptional individuals is not institutionally secure. The resulting interpretation should show why leadership should route obstacles to bodies able to change staffing, finance, curriculum or assessment rather than leaving every correction to the classroom. Professional capability is central. Teachers need subject knowledge, worked examples, diagnostic interpretation and protected time to collaborate. Moderation should examine evidence and reasoning rather than force identical decisions for unlike cases. Staff workload and turnover should be monitored.[REF-23] [REF-24]

    In assessing defining a persistent learning gap, authorities must determine where the intervention is ineffective, adaptation or cessation is responsible improvement. The resulting interpretation should show why where benefit depends on temporary support, institutionalisation requires recurrent finance and ordinary ownership. Completion is demonstrated by stronger learning opportunity and a functioning correction route, not by the end of a project. The public account should distinguish authorisation, delivery, use and learning consequence. It should report limitations, adverse effects and unresolved learners. A favourable later result may reflect population or assessment change and should be tested against the baseline metadata.[REF-01] [REF-02]

    64

    From diagnosis to an intervention hypothesis

    Evidence concerning from diagnosis to an intervention hypothesis should establish the plan should state which observation is direct, which is estimated and which remains unknown. A proportionate conclusion must also recognise that the improvement question concerns an explicit account of the condition expected to change learning. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity. The baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined.[REF-03] [REF-06]

    Institutional action on from diagnosis to an intervention hypothesis should be tested against achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. A contrary reading would overlook that authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is group identity or low performance being mistaken for a causal explanation. Avoiding it requires evidence on both learning and the opportunity supplied.[REF-08] [REF-09]

    Public responsibility for from diagnosis to an intervention hypothesis begins with agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The institutional consequence follows from whether the required evidence includes curriculum exposure, teaching practice, language, time, materials and support. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers.[REF-13] [REF-14]

    For from diagnosis to an intervention hypothesis, the material distinction is between any departure should be recorded with its reason and expected learner consequence. A proportionate conclusion must also recognise that the governing action is to state the mechanism and plausible alternatives before choosing activity. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition.[REF-17] [REF-19]

    65

    Designing the targeted plan

    Institutional action on designing the targeted plan should be tested against if the assessment population excludes learners most at risk, the gap is not adequately defined. The evidence must therefore clarify how the plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns a bounded sequence connecting resources and actions to learner-facing change. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity. The baseline should preserve the underlying distribution and absolute numbers.[REF-08] [REF-09]

    designing the targeted plan requires a decision about authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. The resulting interpretation should show why alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is a list of activities replacing a coherent implementation logic. Avoiding it requires evidence on both learning and the opportunity supplied. Achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms.[REF-13] [REF-14]

    In assessing designing the targeted plan, authorities must determine learner and teacher accounts can identify barriers. This matters because agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes responsibility, staff capability, learner support, milestones and correction. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence.[REF-17] [REF-19]

    A defensible account of designing the targeted plan distinguishes teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. For the learners concerned, the decisive consideration is whether local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to choose a feasible intensity and protect the common educational entitlement. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced.[REF-23] [REF-24]

    66

    Resourcing equitable implementation

    Review of resourcing equitable implementation is credible only where it explains the plan should state which observation is direct, which is estimated and which remains unknown. For the learners concerned, the decisive consideration is whether the improvement question concerns the staff, time, materials, accessibility and finance required by the selected mechanism. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity. The baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined.[REF-13] [REF-14]

    A defensible account of resourcing equitable implementation distinguishes this discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. For the learners concerned, the decisive consideration is whether the principal error is weak institutions being expected to implement with the same nominal allocation. Avoiding it requires evidence on both learning and the opportunity supplied. Achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. Authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them.[REF-17] [REF-19]

    resourcing equitable implementation requires a decision about agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. This matters because the required evidence includes recurrent cost, teacher workload, additional need and household burden. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers.[REF-23] [REF-24]

    Comparative interpretation of resourcing equitable implementation depends upon teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. The evidence must therefore clarify how local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to direct greater support where barriers and implementation costs are greater. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced.[REF-01] [REF-02]

    67

    Monitoring reach, quality and learning response

    monitoring reach, quality and learning response cannot be judged without identifying the baseline should preserve the underlying distribution and absolute numbers. A proportionate conclusion must also recognise that if the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns evidence that the intended learners received the intervention as designed and benefited educationally. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity.[REF-17] [REF-19]

    Institutional action on monitoring reach, quality and learning response should be tested against avoiding it requires evidence on both learning and the opportunity supplied. A contrary reading would overlook that achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. Authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is participation counts being treated as learning evidence.[REF-23] [REF-24]

    Public responsibility for monitoring reach, quality and learning response begins with agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The evidence must therefore clarify how the required evidence includes eligibility, offer, take-up, dosage, teaching quality, work and assessment. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers.[REF-01] [REF-02]

    Public responsibility for monitoring reach, quality and learning response begins with implementation should identify who acts, with what authority, resources and deadline. A proportionate conclusion must also recognise that dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to combine timely implementation evidence with valid learning review.[REF-03] [REF-06]

    68

    Adaptation, institutionalisation and exit

    adaptation, institutionalisation and exit requires a decision about a national average is insufficient when the plan addresses a local or group disparity. The public account remains incomplete unless it explains how the baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns reasoned decisions to continue, change, scale or end the plan. It should be stated with the learner population, educational domain, geography and evidence period.[REF-23] [REF-24]

    adaptation, institutionalisation and exit cannot be judged without identifying this discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. For the learners concerned, the decisive consideration is whether the principal error is temporary measures persisting without benefit or disappearing before durable capability exists. Avoiding it requires evidence on both learning and the opportunity supplied. Achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. Authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them.[REF-01] [REF-02]

    The central question in adaptation, institutionalisation and exit is administrative records can describe staffing and participation while omitting non-enrolled learners. The institutional consequence follows from whether assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes thresholds, adverse effects, unresolved cases, recurrent ownership and later evidence. Each source should be used within scope.[REF-03] [REF-06]

    adaptation, institutionalisation and exit cannot be judged without identifying teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. This matters because local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to retain useful capability while ending ineffective or inequitable arrangements. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced.[REF-08] [REF-09]

    Part XXI

    National translation after adoption of the 2030 Agenda

    69

    Fixing the contemporaneous institutional baseline

    fixing the contemporaneous institutional baseline cannot be judged without identifying the Statistical Commission decision provides a contemporaneous measurement starting point while leaving future refinements and later settlements to later records. The evidence must therefore clarify how national planning should begin from the institutional position that existed on 24 June 2017. The Incheon Declaration expressed the education community's commitment to inclusive and equitable quality education and lifelong learning; Addis supplied a financing framework; the 2030 Agenda established Goal 4 and its targets; and the Education 2030 Framework for Action supplied implementation guidance.[REF-25] [REF-26] [REF-27] [REF-28] [REF-29]

    Institutional action on fixing the contemporaneous institutional baseline should be tested against adoption establishes an agreed direction and permits governments to begin alignment. For the learners concerned, the decisive consideration is whether it does not prove that national law, plans, budgets, information or delivery arrangements already satisfy the commitment. Each country should identify which obligations and targets can be acted upon under existing authority, which require legislative or administrative change, and which depend on clarification through later competent decisions. Unsettled detail should be recorded as such rather than filled by anticipation. The distinction between adoption and implementation is essential.[REF-07] [REF-12] [REF-27]

    Public responsibility for fixing the contemporaneous institutional baseline begins with this avoids two risks: abandoning useful evidence because terminology changed, and claiming continuity where a new target has broader scope or different educational substance. The institutional consequence follows from whether the baseline should preserve existing education commitments and national evidence. The new agenda does not erase the right to education, the unfinished Education for All undertaking, programme structures or established statistical series. National authorities should map the newly adopted targets against those instruments and against the education stages, populations and institutions already in law and plans.[REF-08] [REF-09] [REF-25]

    In assessing fixing the contemporaneous institutional baseline, authorities must determine for each relevant proposition, it should record the adopting body, date, legal or policy status, national authority, present implementing instrument and unresolved question. The evidence must therefore clarify how this is not an administrative inventory for its own sake. It prevents a proposed measure from being represented as an obligation, a declaration from being treated as proof of delivery, and a later decision from being projected backwards. The register should be revised openly as competent bodies act. A dated commitment register can support institutional accuracy.[REF-17] [REF-19] [REF-27]

    fixing the contemporaneous institutional baseline cannot be judged without identifying it also establishes a reliable point from which later implementation can be judged. The public account remains incomplete unless it explains how public communication should use the same discipline. Governments can state that the 2030 Agenda has been adopted and that national alignment is commencing. They should not state that every indicator, national milestone or implementation mechanism has already been internationally settled. A precise account strengthens credibility because it makes clear which choices belong to national democratic and administrative processes and which follow directly from adopted commitments.[REF-25] [REF-26] [REF-27]

    70

    Selecting priorities without narrowing the commitment

    For selecting priorities without narrowing the commitment, the material distinction is between a national plan cannot improve every condition simultaneously, yet it should retain a complete map of early childhood, primary and secondary education, technical and vocational learning, tertiary participation, adult learning, relevant skills, equality, literacy, learning environments, scholarships and teachers as they appear in the adopted targets. The resulting interpretation should show why a first improvement priority should be chosen because evidence shows a serious and remediable break, not because other elements have ceased to matter. The breadth of Goal 4 requires sequencing, not selective abandonment.[REF-25] [REF-27]

    selecting priorities without narrowing the commitment requires a decision about the entitlement test asks which population and educational condition are at stake. For the learners concerned, the decisive consideration is whether the consequence test asks the scale and severity of the denial. The actionability test asks whether a competent authority has a plausible means of change. The equity test asks whether the measure will reach learners farthest from the secured opportunity. A priority that scores highly on visibility but weakly on consequence or equity should be reconsidered. The reasons for selection should be published beside the evidence and known limitations. Priority selection should apply four tests.[REF-01] [REF-03] [REF-12]

    Comparative interpretation of selecting priorities without narrowing the commitment depends upon it should also distinguish poor data from satisfactory conditions. A proportionate conclusion must also recognise that where the least-served population is weakly observed, strengthening coverage can itself become an immediate priority while urgent service evidence supports proportionate protection. National averages should not determine the sequence alone. A moderate national gap may conceal acute failure in one district or population; a large aggregate shortfall may require broad system expansion alongside targeted support. Analysis should retain the national level, absolute number affected, subnational and group distributions and the minimum educational floor.[REF-02] [REF-06] [REF-24]

    Comparative interpretation of selecting priorities without narrowing the commitment depends upon expansion of secondary places may depend on primary completion, trained staff and facilities. The evidence must therefore clarify how adult learning may require flexible provision, recognition and learner support. The plan should make these dependencies visible and decide which must precede or accompany the selected measure. Otherwise a high-level commitment can be converted into an isolated activity unable to change the learner-facing condition. Dependencies should influence sequence. An assessment reform cannot improve learning without curriculum alignment, teacher capability and participation.[REF-03] [REF-16] [REF-18]

    Public responsibility for selecting priorities without narrowing the commitment begins with revision is not a retreat from ambition when reasons and consequences are published. For the learners concerned, the decisive consideration is whether it is a condition of responsible improvement. The complete commitment map should remain in view so that repeated concentration on one readily measured target does not produce silent neglect of lifelong learning, equality, educational quality or populations outside formal schooling. Selection should remain revisable. New evidence may show that the diagnosed mechanism was wrong, that another population is more severely affected or that implementation capacity is insufficient.[REF-08] [REF-25] [REF-27]

    71

    From global target to national improvement proposition

    In assessing from global target to national improvement proposition, authorities must determine for example, a commitment to equitable quality education is too broad to guide one implementation decision; a proposition to improve regular attendance for a defined remote population through transport, staffing and calendar changes can be examined and corrected. This matters because the narrower proposition remains connected to the universal commitment and should not be mistaken for its completion. A national improvement proposition should translate a broad target into a bounded statement of change. It should name the population, present educational condition, institutional mechanism, responsible body, resources, time and evidence of success.[REF-07] [REF-12] [REF-27]

    from global target to national improvement proposition requires a decision about a low completion rate may reflect late entry, repetition, household cost, distance, school safety, language, disability exclusion, teacher shortage or unreliable records. For the learners concerned, the decisive consideration is whether these mechanisms require different responses. A government should compare administrative, household, assessment and local service evidence and state where inference remains uncertain. Consultation with teachers, learners and communities can identify mechanisms, but it should not replace representative population evidence when prevalence is claimed. Diagnosis should precede instrument choice.[REF-01] [REF-03] [REF-24]

    Review of from global target to national improvement proposition is credible only where it explains additional materials will not improve learning if teachers lack time or knowledge to use them; professional guidance will not improve attendance where transport is decisive; a new indicator will not correct exclusion without authority and resources. The evidence must therefore clarify how the plan should identify necessary dependencies and foreseeable adverse effects. It should specify which part of the hypothesis is established by evidence and which remains to be tested during implementation. The intervention hypothesis should explain how the proposed measure changes the barrier.[REF-06] [REF-12] [REF-16]

    from global target to national improvement proposition cannot be judged without identifying it is not legitimate where it narrows the entitled population, lowers expectations for disadvantaged groups or converts learning into attendance alone. For the learners concerned, the decisive consideration is whether the public record should show the relationship between the global target, national definition and selected measure, including any material difference from an existing series. National adaptation should preserve educational substance. Targets may need national definitions for programme levels, age groups, language and institutional responsibility. Adaptation is legitimate where it makes the commitment operational and comparable over time.[REF-02] [REF-10] [REF-25]

    For from global target to national improvement proposition, the material distinction is between review must distinguish theory failure, implementation failure, measurement weakness and insufficient time. For the learners concerned, the decisive consideration is whether the proposition should end with a decision rule. Authorities should state what evidence would justify continuation, expansion, adaptation or cessation and when that decision will occur. A pilot that continues because it attracts support rather than because it changes the intended condition is not an improvement method. Conversely, an intervention should not be abandoned merely because early outcomes are uncertain where delivery has not reached intended intensity.[REF-08] [REF-11] [REF-23]

    72

    Aligning authority, finance and professional capability

    The practical standard for aligning authority, finance and professional capability concerns one public owner should remain answerable for whether the learner-facing condition changes. A proportionate conclusion must also recognise that implementation requires a chain of competent authority. National policy may set the priority, but regional administrations, municipalities, schools, training institutions or other bodies may control staffing, facilities and learner support. The plan should allocate each function to the body able to perform it and identify escalation where local authority is insufficient. Coordination should not allow responsibility to become diffuse.[REF-12] [REF-17] [REF-19]

    In assessing aligning authority, finance and professional capability, authorities must determine the Addis Ababa Action Agenda places national action within a broader financing context, but an international commitment does not supply a national cost estimate. The institutional consequence follows from whether authorities should identify recurrent and capital requirements, the source and timing of funds, distribution rules and the conditions for continuity after temporary support. Finance should cover the complete intervention rather than visible start-up items. Staff preparation, salaries, accessible materials, transport, maintenance, guidance, assessment, evidence and review may all be necessary.[REF-15] [REF-26]

    The practical standard for aligning authority, finance and professional capability concerns equal per-learner funding can reproduce inequality where remoteness, disability, language, insecurity or weak infrastructure makes adequate provision more expensive. The institutional consequence follows from whether formulae should state which need factors are recognised and should be checked against actual receipt. Announced expenditure is not proof of delivery: review should trace authorization, transfer, institutional use and learner consequence. Where households continue to bear material cost, formal fee policy should not be represented as full accessibility. Allocation should respond to unequal cost and starting capacity.[REF-06] [REF-10] [REF-13]

    The central question in aligning authority, finance and professional capability is staffing and turnover should be examined in the locations expected to implement first. The evidence must therefore clarify how a plan dependent on exceptional individuals or uncompensated workload is not institutionally secure and can deepen disparity between strong and weak institutions. Teachers and institutional leaders require capability proportionate to the change. A new curriculum, assessment or inclusion expectation needs more than notification. Professional learning should provide subject substance, practical examples, time for collaboration and a route to support.[REF-01] [REF-18]

    aligning authority, finance and professional capability cannot be judged without identifying its success should be judged by national capability and equitable learner opportunity, not the duration or visibility of the supporting activity. The public account remains incomplete unless it explains how international cooperation should strengthen ordinary national capacity. External support may finance initial expansion, evidence, technical work or regional learning, but roles, conditions and exit should be clear. Parallel activities and reporting arrangements can fragment public authority. Assistance should align with a national improvement proposition, use compatible records and establish how essential functions will enter recurrent provision.[REF-22] [REF-26] [REF-27]

    73

    Monitoring delivery, reach and educational consequence

    A defensible account of monitoring delivery, reach and educational consequence distinguishes a budget can be executed without materials arriving, a programme can operate without reaching disadvantaged learners, and participation can rise without improvement in learning or progression. For the learners concerned, the decisive consideration is whether monitoring should follow the causal sequence of the improvement proposition. Inputs show whether resources and staff were available; delivery evidence shows whether the measure operated; reach shows which eligible learners participated and with what intensity; educational evidence shows whether the intended condition changed. These stages should not be collapsed.[REF-03] [REF-06] [REF-12]

    Evidence concerning monitoring delivery, reach and educational consequence should establish comparable trends are valuable, but continuity should not be asserted where concepts differ materially. A contrary reading would overlook that the baseline should retain numerator, denominator, population, date, geography, definition and exclusions. Where a new target requires a new measure, authorities should preserve the preceding series and identify any break. Apparent improvement caused by revised population estimates, wider institutional reporting or changed assessment participation should be separated from educational change.[REF-02] [REF-23] [REF-24]

    Review of monitoring delivery, reach and educational consequence is credible only where it explains it should therefore report the least-served position and protect against exclusion or deterioration. The resulting interpretation should show why disaggregation must remain lawful, meaningful and safe. Small numbers may require controlled access or combined reporting, but the affected population and public responsibility should not disappear. Equity monitoring should show eligibility, offer, take-up, attendance, completion and outcome for relevant groups and places. A plan can improve its average by reaching learners already closest to the desired condition.[REF-01] [REF-10] [REF-13]

    Evidence concerning monitoring delivery, reach and educational consequence should establish a favourable mean among tested learners cannot represent those outside school or absent from assessment. A contrary reading would overlook that classroom observation, work samples and learner accounts can illuminate mechanism without establishing national prevalence. Agreement across sources strengthens a conclusion only after definitions and dates align; disagreement should guide investigation rather than selective reporting. Learning evidence requires population coverage and opportunity to learn. Assessment results should identify the domain, eligible population, participation and exclusions.[REF-03] [REF-24]

    Public responsibility for monitoring delivery, reach and educational consequence begins with the next review date and responsible authority should be visible. This matters because this makes monitoring a means of correction rather than an obligation to produce favourable figures. The adopted agenda gains national credibility when evidence can alter action and expose populations who remain underserved. Public reporting should connect the result to a decision. A concise account should state what was implemented, who received it, what changed, what remains uncertain, which adverse effects occurred and whether the measure will continue, adapt, expand or cease.[REF-08] [REF-25] [REF-27]

    74

    First national decisions after adoption

    For first national decisions after adoption, the material distinction is between governments can designate a competent coordinating authority, preserve existing sector responsibilities, assemble a dated commitment register and commission a baseline review. The public account remains incomplete unless it explains how the review should map every adopted education target against national law, plans, budgets and evidence. It should identify urgent gaps and unsettled definitions separately. This creates a disciplined bridge between global adoption and national action. The immediate national decision is to establish governance for alignment without pretending that implementation detail is complete.[REF-25] [REF-27]

    Review of first national decisions after adoption is credible only where it explains the plan should guard against choosing only learners and institutions most likely to produce rapid favourable results. This matters because a first priority should be small enough for accountable action and important enough to change educational opportunity. It should name the population and condition, state why it takes precedence, and preserve the wider commitment map. Where credible evidence reveals severe exclusion or harm, protective action need not await a perfect estimate. Longer-term allocation, however, should be reviewed as coverage improves.[REF-01] [REF-06] [REF-12]

    first national decisions after adoption requires a decision about a financing gap should be described rather than hidden through reduced educational substance or transfer of cost to poor households. The evidence must therefore clarify how the first budget decision should identify recurrent implications. Temporary finance may permit testing, but teachers, learner support, accessible facilities, maintenance and evidence cannot be sustained by an announcement. The national authority should state the source, timing and distribution of finance and how external cooperation relates to ordinary provision.[REF-15] [REF-26]

    In assessing first national decisions after adoption, authorities must determine accurate status is a condition of accountable planning, not a reason for delay. The institutional consequence follows from whether the first public report should be candid about chronology. It can state that three major texts relevant to education and financing had been adopted by the cut-off, including the 2030 Agenda on the preceding day. It should also state that later implementation instruments are outside the record. This protects the difference between an adopted commitment, national policy choice and future institutional clarification.[REF-25] [REF-26] [REF-27]

    The practical standard for first national decisions after adoption concerns it should examine authority, delivery, reach, professional capability, finance, learner experience and early educational consequence. The resulting interpretation should show why it should record adverse findings and adapt the measure where the hypothesis or delivery proves weak. The transition from global commitment to national improvement is complete only when ordinary institutions can sustain the changed condition and correct foreseeable departures; the end of a project or reporting period is not evidence of that result. The first review should test whether institutions learned, not merely whether a plan was issued.[REF-08] [REF-12] [REF-23]

    Part XXII

    Quality and equity within the contemporaneous indicator framework

    75

    The hierarchy from goal to observation

    the hierarchy from goal to observation cannot be judged without identifying authority and inference narrow at each stage. The institutional consequence follows from whether a measure can support judgement on part of a target without becoming the target's complete meaning. The indicator framework should be read through a hierarchy of public meaning. Goal 4 states the overarching commitment; each target identifies an educational change or condition; an indicator selects one observable aspect; a national measure applies definitions, sources and calculations; and a released value describes a population and period.[REF-27] [REF-28] [REF-29]

    Comparative interpretation of the hierarchy from goal to observation depends upon a learning indicator can provide important evidence on one domain and population, but cannot establish free access, regular participation, safety, curricular breadth or recognised progression. For the learners concerned, the decisive consideration is whether a parity measure can reveal a relative relationship while concealing low levels for both groups or exclusion from the denominator. Interpretation should therefore state the specific proposition supported and name the remaining parts of the target that require other evidence. This distinction is essential for quality and equity because both are multidimensional.[REF-01] [REF-03] [REF-12]

    In assessing the hierarchy from goal to observation, authorities must determine a responsible national record should retain the indicator version, metadata, source and calculation used at each release. This matters because where subsequent refinement changes the population or concept, the effect should be assessed and the series broken when continuity is not defensible. The Statistical Commission's March 2016 treatment of the proposed framework as a practical starting point supports use and continuing refinement. It does not justify presenting every definition as fixed for all later periods.[REF-16] [REF-23] [REF-29]

    Evidence concerning the hierarchy from goal to observation should establish the common measure supports international orientation; the national measure supports domestic decision. The institutional consequence follows from whether neither should be preferred merely because it produces a more favourable result. A correspondence statement should identify population, education level, event, reference period and material divergence so that readers can see whether two values answer the same question. National measures may add detail relevant to law, programme structure and known barriers, provided their relation to the global concept is documented.[REF-02] [REF-19] [REF-28]

    A defensible account of the hierarchy from goal to observation distinguishes these are different findings with different responsible bodies and remedies. For the learners concerned, the decisive consideration is whether public reporting should preserve this hierarchy in its language. It should not say that a goal was achieved because one indicator improved, or that an indicator failed because a source is temporarily unavailable. An unavailable measure identifies a statistical gap; an unfavourable value identifies an observed condition; an incomplete target judgement requires a broader evidentiary account.[REF-08] [REF-17] [REF-29]

    76

    Interpreting quality beyond one outcome

    A defensible account of interpreting quality beyond one outcome distinguishes no single item establishes the whole. This matters because inputs can be present but unused; learning can be observed among a selected population; completion can carry weak or uncertain educational value. The framework should therefore be interpreted as a related set rather than a search for one quality proxy. Quality concerns the educational substance and conditions through which learners participate, learn, complete and progress. Its evidence can include curriculum, teachers, instructional time, facilities, safety, accessibility, assessment, learner work and recognised qualifications.[REF-09] [REF-12] [REF-28]

    The central question in interpreting quality beyond one outcome is school-based assessment cannot represent children outside provision without an explicit population model. The public account remains incomplete unless it explains how participation and exclusions should accompany the result so that improved coverage is not mistaken for deteriorating learning or vice versa. Learning measures require a specified domain, target population, instrument, language, administration and proficiency threshold. Grade-based results describe learners who reached the grade and participated under stated conditions. Age-based results may represent a broader population while using different educational correspondence.[REF-03] [REF-24]

    interpreting quality beyond one outcome requires a decision about a report should state which part of curriculum is observed and should not infer general school quality from a limited instrument. The institutional consequence follows from whether where a target concerns citizenship or sustainable development, policy presence and learner capability should remain separate stages. Relevant and effective learning outcomes should not be narrowed to whatever domain is easiest to compare. National curricula and public purposes remain material. Comparative measures can establish a common bounded domain, while national evidence addresses additional knowledge, capabilities and progression.[REF-15] [REF-24] [REF-27]

    Institutional action on interpreting quality beyond one outcome should be tested against counts should be interpreted with preparation, assignment, attendance, workload and support. A contrary reading would overlook that a national ratio can conceal shortage by subject, level or locality. Qualification definitions vary and need national metadata. Reporting should connect teacher evidence to the learners and institutions exposed to the condition rather than assume that a national workforce total describes classroom opportunity uniformly. Teachers are both a means-of-implementation concern and a condition of quality across targets.[REF-04] [REF-18] [REF-28]

    Public responsibility for interpreting quality beyond one outcome begins with if facilities exclude learners with disabilities, an aggregate infrastructure count is insufficient. A proportionate conclusion must also recognise that the public account should connect the observed condition to the competent authority, resource and later review without claiming attribution beyond the evidence. Quality interpretation should lead to action. If low learning coincides with weak participation, response should not concentrate solely on assessed pupils. If teacher shortage is local, a national recruitment measure may be too broad or too slow.[REF-10] [REF-13] [REF-17]

    77

    Equity, parity and minimum educational levels

    Public responsibility for equity, parity and minimum educational levels begins with a parity index can improve because a disadvantaged group advanced, because the reference group deteriorated or because both changed. The evidence must therefore clarify how only the underlying levels reveal which occurred. A small relative gap can coexist with severe deprivation across all groups. Equity is substantive fairness in opportunity, support and result. Indicator interpretation should therefore retain levels, absolute numbers, gaps, ratios and distance from an educational minimum. These quantities answer different questions.[REF-05] [REF-06] [REF-14]

    Review of equity, parity and minimum educational levels is credible only where it explains more detail is not automatically better: small cells can be unstable or identifying. A proportionate conclusion must also recognise that restricted analysis, pooled estimates or descriptive service evidence may support action while public reports preserve confidentiality. Disaggregation should begin from a public question and an entitled population. Sex, household resources, residence, disability, language, migration, displacement and indigenous or other nationally relevant status may reveal different barriers. Categories should be lawful, respectful, stable enough for comparison and safe to publish.[REF-07] [REF-10] [REF-24]

    Comparative interpretation of equity, parity and minimum educational levels depends upon cross-classification should be selected for a plausible mechanism and adequate precision rather than produced mechanically. A proportionate conclusion must also recognise that the report should show the population share represented and state where a group estimate is unavailable. An unobserved population should never be coded as equality or merged silently with the majority. Intersections can reveal exclusion hidden by single categories. The experience of poor rural girls may differ from the national female average, and disability may interact with language or displacement.[REF-01] [REF-13] [REF-23]

    Review of equity, parity and minimum educational levels is credible only where it explains if disadvantaged learners are more likely to leave before the observed stage, a narrowing outcome gap among those remaining can coexist with wider system exclusion. This matters because reports should relate outcome distributions to entry, attendance, progression and participation in the measurement itself. The equity test should also consider who enters the outcome population. Assessment and completion measures are conditional on prior access and progression.[REF-02] [REF-03] [REF-16]

    Evidence concerning equity, parity and minimum educational levels should establish a distribution-sensitive milestone should identify the least-served population, expected level and absolute number, and should review reach as well as outcome. The evidence must therefore clarify how equity becomes accountable when a material gap is connected to an authority and a remedy, not when it is merely displayed. Targets should protect a minimum condition and prevent deterioration. Concentrating support on learners closest to a threshold can raise a national result while leaving those farthest behind unchanged.[REF-08] [REF-12] [REF-28]

    78

    Refugee and migrant inclusion after the New York Declaration

    A defensible account of refugee and migrant inclusion after the new york declaration distinguishes its adoption does not establish that every learner entered education or that later implementation succeeded. A proportionate conclusion must also recognise that indicator interpretation should identify whether refugees and migrants are within population denominators, school registers, surveys, assessments and public reporting, and should state relevant status and residence rules. The New York Declaration is contemporaneous evidence that refugee and migrant children should be included within education policy and evidence.[REF-30]

    For refugee and migrant inclusion after the new york declaration, the material distinction is between registration, border, shelter, household and school sources may refer to different populations and dates and may contain duplicate or missing persons. The institutional consequence follows from whether a rapid operational count can guide service allocation when dated and revised; it should not be represented as a final participation rate without a suitable denominator. Country of origin, present residence and service responsibility should remain distinct where each affects interpretation. Population movement makes denominators time-sensitive.[REF-11] [REF-20] [REF-23]

    Evidence concerning refugee and migrant inclusion after the new york declaration should establish national reports should show whether refugee and migrant learners are integrated into public provision, served through temporary arrangements or absent from the recorded system. This matters because education participation should be traced through admission, attendance, instructional time, learning, completion and recognised progression. Documentation, language, prior curriculum, disability, household cost and movement can interrupt this course. A place formally available is not evidence of usable access when those conditions are unresolved.[REF-09] [REF-12] [REF-30]

    Review of refugee and migrant inclusion after the new york declaration is credible only where it explains diagnostic placement can support continuity but should not become an exclusionary high-stakes barrier. This matters because results should be interpreted with opportunity to learn and participation. Records and certification should support movement and progression while protecting personal information and avoiding disclosure that creates risk. Assessment and qualification evidence need continuity safeguards. Learners may lack prior records, enter at a different grade or study in an unfamiliar language.[REF-10] [REF-24]

    Institutional action on refugee and migrant inclusion after the new york declaration should be tested against additional support directed only to one population may create parallel structures, while equal treatment without language or bridging support may reproduce exclusion. The evidence must therefore clarify how equity requires differentiated resources to secure a common substantive opportunity, monitored across displaced, migrant and host learners. Host-community conditions belong in the same account. Rapid enrolment can change class size, teacher workload and materials for existing learners.[REF-06] [REF-14] [REF-30]

    79

    Responsible use before later settlement

    Evidence concerning responsible use before later settlement should establish later General Assembly action, later tier classifications, technical revisions, estimates and results cannot be projected backwards. The evidence must therefore clarify how a country may use a contemporaneous candidate or national measure, but should identify its status and avoid implying that future arrangements had already been settled. As at the cut-off, indicator use should remain tied to the March 2016 Statistical Commission decision and the official metadata available by that date.[REF-29]

    The practical standard for responsible use before later settlement concerns a value released in 2016 may describe an earlier school year, survey or assessment. The resulting interpretation should show why it should not be labelled a 2016 educational condition merely because of release date. Historical reconstruction should use the definitions and sources available for the relevant period, disclose estimation and avoid filling gaps with later results. When a later revision occurs, the original baseline and reason for change should remain accessible. Baseline construction should preserve actual reference periods.[REF-03] [REF-16] [REF-23]

    Evidence concerning responsible use before later settlement should establish similar indicator labels can conceal differences in age, programme, threshold, population coverage or collection. The evidence must therefore clarify how metadata should travel with the value, and apparent differences should be tested against those features before ranking systems. A global aggregate may be useful for orientation while remaining unable to describe internal distributions or countries with missing evidence. Comparison should be conditional on conceptual correspondence.[REF-02] [REF-19] [REF-24]

    A defensible account of responsible use before later settlement distinguishes unexpected improvement is not evidence of manipulation by itself, but it should be confirmed against population, delivery and measurement change before broad policy claims are made. The resulting interpretation should show why indicators should not create incentives to exclude difficult cases, narrow curriculum or reclassify non-completion. Participation, source coverage and definition stability provide safeguards. Independent statistical authority, release calendars and visible revision strengthen trust.[REF-08] [REF-17]

    Public responsibility for responsible use before later settlement begins with it should state the observed value, population, distribution, uncertainty, interpretation, responsible authority and next review. The resulting interpretation should show why if evidence is inadequate, the decision may be stronger collection rather than immediate intervention; if credible evidence shows serious exclusion, proportionate action need not wait for perfect precision. In either case, the indicator serves public judgement rather than becoming a detached mark of compliance. The public statement should conclude with a bounded decision.[REF-12] [REF-28] [REF-29]

    80

    Minimum interpretive record

    minimum interpretive record cannot be judged without identifying it should state population, numerator, denominator, reference period, geography, source, disaggregation, uncertainty and known exclusions. This matters because where the measure is non-numeric, it should state the institutional condition, evidence and judgement rule. The minimum interpretive record should identify the Goal 4 target, the specific proposition measured, the indicator status at 24 June 2017 and the national correspondence.[REF-27] [REF-28] [REF-29]

    Evidence concerning minimum interpretive record should establish equity claims should show group levels and absolute numbers beside gaps or ratios and should identify the minimum condition protected. A proportionate conclusion must also recognise that a composite score should not be used where it conceals failure in an essential dimension or permits success among represented learners to compensate for exclusion from the population. Quality claims should identify which educational domain and population are observed and which conditions remain outside the measure.[REF-01] [REF-05] [REF-12]

    For minimum interpretive record, the material distinction is between it should identify changes in definition or coverage and preserve revision history. This matters because missing, unknown, not applicable and suppressed are different states. When public disclosure is unsafe, the affected population and institutional action should remain visible even if the detailed value is restricted. The record should distinguish direct observation, calculation, estimate and interpretation.[REF-08] [REF-10] [REF-23]

    In assessing minimum interpretive record, authorities must determine the report should state whether the next action concerns evidence, authority, finance or implementation because these require different remedies. For the learners concerned, the decisive consideration is whether an equity or quality finding should be linked to a competent body, available measure, resource requirement and review date. Statistical offices protect methodological integrity; education authorities act on service and policy; institutions correct delivery within their mandate. Coordination should not diffuse ownership.[REF-17] [REF-19] [REF-28]

    The practical standard for minimum interpretive record concerns this is the minimum standard for evidence that can support comparison without narrowing the universal commitment. A proportionate conclusion must also recognise that at 24 June 2017, a credible interpretation treats the indicator framework as a practical measurement starting point subordinate to the full education targets. It uses the New York Declaration within its adopted scope, excludes later settlements and results, protects unrepresented populations, and refuses to infer quality or equity from a single convenient value.[REF-29] [REF-30]

    Part XXIII

    Allocation of responsibility and remedy

    81

    Government responsibility

    government responsibility cannot be judged without identifying public authorities determine compulsory and free provision, curriculum and qualification rules, teacher establishments, resource distribution, data safeguards and complaint mechanisms. The public account remains incomplete unless it explains how decentralisation may allocate functions, but it does not remove the obligation to ensure that every competent body has authority and resources and that gaps between jurisdictions do not leave learners without remedy. Government responsibility begins with the legal and financial conditions of the right to education.[REF-08] [REF-09] [REF-12]

    A defensible account of government responsibility distinguishes a national standard should identify what institutions and learners may expect and how additional support is allocated where disability, language, poverty, remoteness or displacement creates greater cost. For the learners concerned, the decisive consideration is whether equal nominal rules are not sufficient when they produce unequal use. National authorities should define the substantive minimum while permitting context-sensitive implementation. Availability, accessibility, acceptability and adaptability require provision, non-discrimination, educational quality and responsiveness.[REF-10] [REF-13] [REF-27]

    For government responsibility, the material distinction is between a national budget total cannot show whether understaffed or inaccessible institutions received the resources required. A proportionate conclusion must also recognise that where responsibility is shared with international partners or private providers, public authority should retain standards, protection and continuity and should disclose conditions that affect the service. Finance should be transparent from authorization through transfer, receipt and learner-facing use. Governments should publish the basis of allocations, timing and treatment of need.[REF-17] [REF-19] [REF-26]

    Review of government responsibility is credible only where it explains missing groups should remain visible; personal information should be protected; revisions should be public. A contrary reading would overlook that a lack of data does not extinguish a duty, but it can require a proportionate improvement in evidence alongside immediate protection where credible information shows exclusion. Government also carries responsibility for information capable of revealing unequal opportunity. Statistical offices and ministries should maintain population, administrative, assessment and finance evidence with professional integrity.[REF-03] [REF-23] [REF-24]

    Review of government responsibility is credible only where it explains government review should identify recurrent complaints and correct the system condition rather than resolve identical cases one by one without institutional learning. The public account remains incomplete unless it explains how the remedy structure should be navigable. Learners, families, institutions and teachers need to know which body can address admission, staffing, accessibility, discrimination, qualification or finance and within what period. Appeals should not require specialist knowledge unavailable to affected communities.[REF-05] [REF-12]

    82

    Institutional responsibility

    The central question in institutional responsibility is institutional plans should identify the eligible population, service conditions and staff roles and should publish information sufficient for learners and families to understand what is offered and how to challenge a failure. This matters because education institutions are responsible for implementing law and public standards within their delegated authority. They organize admission, attendance, teaching, assessment, safeguarding, records, support and progression.[REF-05] [REF-15]

    The practical standard for institutional responsibility concerns it cannot confer a qualification without competent recognition, promise a service dependent on unfunded external action or exclude a learner because another public body failed to provide documentation or support. The resulting interpretation should show why where a material dependency lies elsewhere, the institution should escalate it and protect continuity as far as lawfully possible. The unresolved condition should remain in the plan and public account. An institution should not claim authority it lacks.[REF-09] [REF-12]

    Institutional action on institutional responsibility should be tested against timetables, grouping, discipline, assessment, language and communication can create barriers even where national rules are formally equal. The evidence must therefore clarify how leaders should examine participation and outcome distributions and seek learner, teacher and family evidence. Group difference should initiate enquiry into institutional mechanisms, not blame or lower expectations. Institutional responsibility includes equitable organization.[REF-01] [REF-13]

    Public responsibility for institutional responsibility begins with performance information should be used to identify conditions and support, not to attach sole responsibility to a classroom for finance, staffing or population differences beyond its control. This matters because leadership should document decisions, adverse effects and requested system action. Institutions should support professional judgement and correction. Teachers need curriculum guidance, planning time, materials and routes to obtain specialist help.[REF-17] [REF-18]

    Institutional action on institutional responsibility should be tested against confidentiality should be protected. The public account remains incomplete unless it explains how the institution should state outcome, reasons and further appeal and should use recurring evidence to revise ordinary practice. Complaints and review should be safe and practical. Families and learners should be able to raise concerns without retaliation, and staff should be able to report unsafe or unworkable conditions.[REF-08] [REF-10]

    83

    Teacher professional responsibility

    In assessing teacher professional responsibility, authorities must determine it should not be used to make teachers individually liable for system shortages, household poverty or infrastructure decisions they cannot change. This matters because teachers hold professional responsibility for planning and delivering curriculum, assessing learning, maintaining inclusive participation and using evidence within their competence. This responsibility is real but bounded by training, time, class conditions, materials and authority.[REF-15] [REF-18]

    teacher professional responsibility requires a decision about observation and learner evidence should address bounded practice and context. The public account remains incomplete unless it explains how one result or visit should not stand for comprehensive competence, and evaluation should distinguish individual practice from shared institutional conditions. Professional standards should be clear enough for support and fair review. Teachers should know the intended learning, assessment principles, safeguarding duties, confidentiality requirements and route for referral.[REF-03] [REF-16]

    Institutional action on teacher professional responsibility should be tested against high-stakes decisions should have moderation and protection against conflict of interest. A contrary reading would overlook that assessment responsibility includes accurate, fair and accessible judgement. Teachers should use multiple evidence where appropriate, apply criteria consistently and avoid deficit assumptions based on identity. Learners should receive feedback and a route to challenge factual or procedural error.[REF-08] [REF-12]

    The central question in teacher professional responsibility is a referral does not end educational responsibility; continued participation and follow-up remain material. The public account remains incomplete unless it explains how teachers also have a duty to identify barriers and escalate them. Persistent absence, exclusion, unsafe facilities, disability needs, language barriers or protection concerns may require action beyond the classroom. Teachers should have a documented route to institutional or specialist response and should not be asked to conduct investigations beyond professional competence.[REF-10] [REF-13]

    Professional accountability should be accompanied by due process and learning. A concern should state the practice, evidence, expected standard, support and review. Serious misconduct may require formal action, but weak results alone do not establish misconduct. Improvement is sustained when teachers can examine evidence collectively, obtain assistance and see system obstacles addressed.[REF-17] [REF-19]

    84

    Family and learner participation without transferred liability

    The central question in family and learner participation without transferred liability is a family cannot be held responsible for distance, unaffordable charges, inaccessible facilities or missing teachers created by public arrangements. A contrary reading would overlook that families contribute essential knowledge about participation, health, language, prior learning, cost and safety. They support attendance and learning within their circumstances and can question institutional decisions. Their participation improves diagnosis and legitimacy, but it does not replace the State's responsibility to provide accessible and acceptable education.[REF-05] [REF-12]

    family and learner participation without transferred liability requires a decision about meetings and consultation should account for work, care, language, disability and travel so that participation is not limited to those with the greatest resources. The evidence must therefore clarify how communication should state rights, programme expectations, attendance, assessment, support, cost and complaint routes in accessible language and format. Information is not participation when families cannot ask questions or influence a decision.[REF-10] [REF-13]

    Evidence concerning family and learner participation without transferred liability should establish institutions should record how the evidence affected the decision. A proportionate conclusion must also recognise that learners should be heard on teaching, safety, discrimination, workload and support in ways appropriate to age and protection. Participation should be voluntary and should not expose personal circumstances publicly. Learner views are evidence of experience and mechanism; they should not be represented as population prevalence without a suitable design.[REF-08] [REF-15]

    The central question in family and learner participation without transferred liability is voluntary contribution can become an informal condition of admission or participation, excluding poor families. The resulting interpretation should show why public reports should distinguish statutory charge, optional contribution and actual household expenditure. Community activity can strengthen provision but should not finance essential functions through unequal ability to pay. Household contribution to cost requires scrutiny.[REF-06] [REF-17]

    family and learner participation without transferred liability cannot be judged without identifying the response should be proportionate, protect the learner and match the demonstrated mechanism. The public account remains incomplete unless it explains how where attendance is weak, authorities should examine the barrier before assigning fault. Communication and family support may be appropriate, but punitive action can worsen exclusion when the cause is transport, safety, disability, work, care or school quality.[REF-01] [REF-14]

    85

    Shared action and undivided accountability

    The central question in shared action and undivided accountability is admission may involve civil registration, social services and schools; disability inclusion may involve health and accessibility bodies; refugee education may involve origin, host and international authorities; transition to work may involve education and labour institutions. The evidence must therefore clarify how coordination should define tasks and information while preserving one accountable owner for the learner-facing result. Many education conditions require shared action.[REF-20] [REF-30]

    Public responsibility for shared action and undivided accountability begins with delays and disagreement should remain visible rather than be described as stakeholder complexity without remedy. The resulting interpretation should show why shared plans should identify dependency, authority, resource, deadline and escalation. A committee or memorandum does not prove delivery. Each body should report the action within its control and the lead body should report whether the combined condition changed.[REF-17] [REF-19]

    shared action and undivided accountability requires a decision about bodies should specify purpose, minimum fields, access, retention and correction. This matters because aggregate evidence may be sufficient for planning, while individual referral requires consent or another lawful basis and safeguards. Information sharing should follow necessity and protection. Coordination does not authorize unrestricted transfer of personal data.[REF-08] [REF-10]

    Comparative interpretation of shared action and undivided accountability depends upon consultation should not allow the best-organized actor to define the public interest or permit government to present consensus where material objections remain. The public account remains incomplete unless it explains how decisions should state reasons and how affected populations were represented. Power differences matter. Families, learners and teachers may participate in a forum without controlling finance or regulation.[REF-05] [REF-12]

    Institutional action on shared action and undivided accountability should be tested against subsequent review should examine whether coordination reduced delay and changed educational opportunity, not merely whether meetings occurred. A proportionate conclusion must also recognise that undivided accountability means that the learner can obtain a decision even when several bodies contribute. The lead authority should not redirect the person indefinitely between offices. Internal allocation can be complex; the public route should be simple.[REF-09] [REF-16]

    86

    Minimum responsibility statement

    Institutional action on minimum responsibility statement should be tested against a responsibility statement should identify the educational entitlement, eligible population, government authority, institutional duty, teacher professional role, family and learner participation, required resources, evidence and review date. The resulting interpretation should show why it should distinguish mandatory function, discretionary contribution and action outside each actor's authority. The conclusion for transparent responsibilities for governments, institutions, teachers and families should connect the finding to the competent authority, affected learners and evidence required at review.[REF-08] [REF-12]

    Comparative interpretation of minimum responsibility statement depends upon admission, finance, staffing, curriculum, assessment, accessibility, safeguarding and records may require different owners. For the learners concerned, the decisive consideration is whether a general reference to stakeholders is not sufficient. Where authority is missing or contested, government should resolve it and protect learners during the interval. The statement should attach every material finding to the body capable of remedy.[REF-09] [REF-17]

    minimum responsibility statement requires a decision about corrections should reach the audience that received the original claim. The resulting interpretation should show why transparency should be proportionate and safe. Performance and complaint information should show the condition and response without exposing personal information or stigmatizing groups. Missing evidence, unresolved cases and adverse findings should remain visible.[REF-10] [REF-13]

    Comparative interpretation of minimum responsibility statement depends upon a policy can establish authority, a budget can authorize finance and an institution can document activity; none alone proves changed opportunity. This matters because the responsible body should state what is known, what remains uncertain and the next evidence needed. Review should distinguish authorization, delivery, use and educational consequence.[REF-03] [REF-16]

    A defensible account of minimum responsibility statement distinguishes teachers exercise supported professional responsibility;. A proportionate conclusion must also recognise that families and learners participate without assuming public liability;. and coordination must end in a named decision and remedy. At 24 June 2017, the minimum interpretation is clear: governments retain the public duty;. institutions organize accountable provision;.[REF-05] [REF-12] [REF-27]

    References

    1. REF-01

      Education for All Global Monitoring Report Team. Reaching the Marginalized — EFA Global Monitoring Report 2010. 2010.

      Principal contemporaneous analysis of intersecting disadvantage and education marginalisation.

      https://unesdoc.unesco.org/ark:/48223/pf0000186606
    2. REF-02

      UNESCO Institute for Statistics. Global Education Digest 2010: Comparing Education Statistics Across the World. 2010.

      Comparative education statistics, definitions and limitations.

      https://uis.unesco.org/sites/default/files/documents/global-education-digest-2010-comparing-education-statistics-across-the-world-en.pdf
    3. REF-03

      UNESCO Institute for Statistics. Education Indicators: Technical Guidelines. 2009.

      Definitions and interpretation of participation, progression, completion and resource indicators.

      https://uis.unesco.org/sites/default/files/documents/education-indicators-technical-guidelines-en_0.pdf
    4. REF-04

      United Nations. The Millennium Development Goals Report 2010. 2010.

      Global and regional monitoring of primary education, gender, poverty and related development conditions.

      https://www.un.org/millenniumgoals/pdf/MDG%20Report%202010%20En%20r15%20-low%20res%2020100615%20-.pdf
    5. REF-05

      United Nations Development Programme. Human Development Report 2010: The Real Wealth of Nations — Pathways to Human Development. 2010.

      Distribution-sensitive human development concepts and evidence available before the cut-off.

      https://hdr.undp.org/content/human-development-report-2010
    6. REF-06

      UNICEF. Progress for Children: Achieving the MDGs with Equity, Number 9. 2010.

      Equity-focused child indicators and comparison between population groups.

      https://www.unicef.org/reports/progress-children-no-9
    7. REF-07

      World Education Forum. The Dakar Framework for Action: Education for All — Meeting Our Collective Commitments. 2000.

      Commitments to equitable access, quality, measurable outcomes and accountable national planning.

      https://unesdoc.unesco.org/ark:/48223/pf0000121147
    8. REF-08

      United Nations General Assembly. Convention on the Rights of the Child. 1989.

      Rights concerning non-discrimination, identity, education and development.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-child
    9. REF-09

      United Nations Committee on Economic, Social and Cultural Rights. General Comment No. 13: The Right to Education. 1999.

      Interpretation of availability, accessibility, acceptability and adaptability.

      https://www.refworld.org/legal/general/cescr/1999/en/37937
    10. REF-10

      United Nations General Assembly. Convention on the Rights of Persons with Disabilities. 2006.

      Non-discrimination, accessibility, inclusive education and disability data safeguards.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-persons-disabilities
    11. REF-11

      United Nations. Guiding Principles on Internal Displacement. 1998.

      Principles relevant to protection, documentation, education and non-discrimination of displaced persons.

      https://www.ohchr.org/en/special-procedures/sr-internally-displaced-persons/international-standards
    12. REF-12

      UNESCO and UNICEF. A Human Rights-Based Approach to Education for All. 2007.

      Rights-based planning, equality, participation, accountability and education quality.

      https://unesdoc.unesco.org/ark:/48223/pf0000154861
    13. REF-13

      Education for All Global Monitoring Report Team. Overcoming Inequality: Why Governance Matters — EFA Global Monitoring Report 2009. 2008.

      Governance, finance and unequal educational opportunity.

      https://unesdoc.unesco.org/ark:/48223/pf0000177683
    14. REF-14

      World Bank. World Development Report 2006: Equity and Development. 2005.

      Concepts of unequal opportunity, institutions and equitable public action.

      https://documents.worldbank.org/curated/en/435331468127174418/pdf/322040World0Development0Report02006.pdf
    15. REF-15

      World Bank. Safeguarding Education During Economic Crisis. 2009.

      Risks to budgets, households, participation and long-term human development during economic crisis.

      https://documents1.worldbank.org/curated/en/489131468340200911/pdf/485120WP0Avert10Box338912B01PUBLIC1.pdf
    16. REF-16

      Organisation for Economic Co-operation and Development. Education at a Glance 2010: OECD Indicators. 2010.

      Comparative participation, progression, expenditure and outcomes evidence with system-level metadata.

      https://doi.org/10.1787/eag-2010-en
    17. REF-17

      European Commission. Europe 2020: A Strategy for Smart, Sustainable and Inclusive Growth. 2010.

      Contemporaneous European policy context for education, inclusion, employment and headline indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC2020
    18. REF-18

      European Commission. Youth on the Move: An Initiative to Unleash the Potential of Young People to Achieve Smart, Sustainable and Inclusive Growth in the European Union. 2010.

      European education, mobility, attainment and youth inclusion policy context.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0477
    19. REF-19

      European Commission. A Renewed Commitment to Social Europe: Reinforcing the Open Method of Coordination for Social Protection and Social Inclusion. 2008.

      Social inclusion monitoring, common objectives and context-sensitive indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52008DC0418
    20. REF-20

      United Nations General Assembly. Resolution 64/250: Assistance to Haiti in the Aftermath of the Recent Earthquake. 2010.

      Contemporaneous recognition of humanitarian and reconstruction needs and national leadership.

      https://undocs.org/A/RES/64/250
    21. REF-21

      United Nations Office for the Coordination of Humanitarian Affairs. Haiti Revised Humanitarian Appeal. 2010.

      Displacement, service disruption and humanitarian education context.

      https://reliefweb.int/report/haiti/haiti-revised-humanitarian-appeal-2010
    22. REF-22

      European Commission. European Union Response to the Earthquake in Haiti. 2010.

      European humanitarian and recovery support, coordination and Haitian ownership.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0056
    23. REF-23

      United Nations Economic and Social Council. Principles and Recommendations for Population and Housing Censuses, Revision 2. 2008.

      Official principles for population coverage, definitions, classifications and data quality.

      https://unstats.un.org/unsd/demographic-social/Standards-and-Methods/files/Principles_and_Recommendations/Population-and-Housing-Censuses/Series_M67Rev2-E.pdf
    24. REF-24

      United Nations Statistics Division. Designing Household Survey Samples: Practical Guidelines. 2005.

      Sample design, estimation, precision and non-response guidance.

      https://unstats.un.org/unsd/demographic/sources/surveys/Handbook23June05.pdf
    25. REF-25

      World Education Forum 2015. Incheon Declaration: Education 2030 — Towards Inclusive and Equitable Quality Education and Lifelong Learning for All. 2015.

      Education commitments adopted at Incheon and their contemporaneous institutional status.

      https://unesdoc.unesco.org/ark:/48223/pf0000233137
    26. REF-26

      United Nations General Assembly. Addis Ababa Action Agenda of the Third International Conference on Financing for Development. 2015.

      Adopted global financing framework and relevant principles for domestic public finance and international cooperation.

      https://undocs.org/A/RES/69/313
    27. REF-27

      United Nations General Assembly. Transforming Our World: The 2030 Agenda for Sustainable Development. 2015.

      Agenda adopted on 25 September 2015, including Goal 4 and its education targets, as available at the evidence cut-off.

      https://undocs.org/A/RES/70/1
    28. REF-28

      UNESCO and the World Education Forum 2015 co-convening agencies. Education 2030 Framework for Action. 2015.

      Implementation guidance adopted at the 4 November 2015 high-level meeting; used without later indicator arrangements, statistics or results.

      https://www.unesco.org/en/articles/education-2030-framework-action-be-formally-adopted-and-launched
    29. REF-29

      United Nations Statistical Commission. Report on the Forty-Seventh Session (8–11 March 2016). 2016.

      Contemporaneous Statistical Commission decision treating the proposed global indicator framework as a practical starting point, subject to future refinement.

      https://unstats.un.org/UNSDWebsite/statcom/session_47/documents/2016-34-FinalReport-E.pdf
    30. REF-30

      United Nations General Assembly. New York Declaration for Refugees and Migrants. 2016.

      Adopted commitment relevant to access to education for refugee and migrant children within the cut-off.

      https://undocs.org/A/RES/71/1