ICEQC-R-2017-12 — Global Education Quality Annual Report 2017: Accountability, Transparency and Shared Responsibility cover

Documento técnico anual

ICEQC-R-2017-12 — Global Education Quality Annual Report 2017: Accountability, Transparency and Shared Responsibility

Integrated annual monitoring of standards, indicators, improvement and public responsibility

Fecha de publicación
Categoría de investigación
Integrated Annual Monitoring across the Four Approved Research Domains
Informe arquetipo
Monitoreo integrado anual
Ámbito geográfico
Global with regional differentiation
Fecha límite para la presentación de pruebas
Organismo responsable
Dirección de Investigación y Políticas de ICEQC
ICEQC-R-2017-12 — Global Education Quality Annual Report 2017: Accountability, Transparency and Shared Responsibility cover

Publication record

This is the controlled English edition. Evidence and institutional status are stated as at the evidence cut-off date.

Executive summary

Institutional action on executive summary should be tested against governments remain responsible for the right to education and for a coherent publicly accountable system. This matters because local bodies, institutions, teachers, families, communities, international organisations and non-state providers hold different duties and capabilities. Shared responsibility is credible only where each consequential decision has an identifiable authority, applicable standard, resource, evidence and route to correction. The 2017 education-quality account is organised around responsibility.

The Global Education Monitoring Report 2017/8, published before this annual report's cut-off, supplies the year's principal thematic analysis. It treats accountability as a means to educational ends and warns against disproportionate blame for conditions beyond an actor's control. This annual report uses that analysis as a contemporaneous institutional reference. It does not import statistics, policies or outcomes published after 10 December 2017.

The practical standard for executive summary concerns publication can support scrutiny only where affected people can understand the claim and where a competent authority can respond. The institutional consequence follows from whether transparency is not the release of information without context. A public account should distinguish legal commitment, authorised finance, activity, service, reach and educational result. It should state definitions, population, geography, date, missing evidence and uncertainty.

The central question in executive summary is a common label should not conceal unequal service, and local discretion should not weaken a public minimum. The resulting interpretation should show why the standards domain requires clarity about entitlement, public authority and minimum conditions. National commitments need translation into rules governing access, curriculum, teachers, facilities, assessment, recognition and remedy. Standards should protect equity while permitting contextual delivery.

Comparative interpretation of executive summary depends upon a funding pledge is not a commitment; a commitment is not receipt; expenditure is not learner-facing service. A contrary reading would overlook that measures should preserve denominators, coverage and revision. National averages should be accompanied by evidence capable of revealing populations missing from the account. The indicator domain requires separation of levels, distributions and transitions. Enrolment is not attendance; attendance is not completion; completion is not learning.

Public responsibility for executive summary begins with evaluation should identify the population, criterion and evidence; institutional response should name authority, resource and intended operating change; follow-up should verify delivery, reach and result. For the learners concerned, the decisive consideration is whether additional reporting, meetings or training events do not establish improvement. Where a cause lies beyond institutional control, the dependency should be escalated rather than converted into local blame. The improvement domain requires a trace from finding to correction.

executive summary cannot be judged without identifying local bodies should translate population and place into service plans. This matters because institutions should deliver and review teaching and support. Coordination should make responsibility visible and should protect professional competence, personal information and access to remedy. The policy and regulatory domain requires alignment across levels. Central authorities should establish law, finance, equalisation and national reporting.

In assessing executive summary, authorities must determine constraint does not remove responsibility; it changes the necessary response. The institutional consequence follows from whether the body controlling the missing condition should be identified, and interim protection should continue while correction is sought. Accountability under resource constraint should focus on the complete educational condition. A body should not be held responsible for an outcome when staffing, facilities, finance or legal authority lie elsewhere.

Public responsibility for executive summary begins with appeals, announcements, pledges, signed agreements, receipts, allocations, disbursements, expenditure and results are distinct. The resulting interpretation should show why public reporting should preserve currency, period, restrictions and intended population. Crisis-affected and displaced learners require continuity, qualified teachers, records and recognised progression, not only temporary space. Education aid illustrates the transparency problem.

In assessing executive summary, authorities must determine it reinforces shared responsibility concerning refugees and migrants, including education. The resulting interpretation should show why its status should not be inflated into proof of later delivery. Likewise, the launch of Education Cannot Wait establishes an institutional initiative, while any finance, reach or result requires separate dated evidence. The New York Declaration is an adopted commitment within the evidence period.

Institutional action on executive summary should be tested against learners and families can reveal inaccessible rules, delay and service failure; teachers and local officials can identify contradictory duties and unavailable conditions. The public account remains incomplete unless it explains how participation should be safe, accessible and capable of influencing an authorised decision. Low complaint volume should not be treated automatically as absence of harm. Participation and complaint are essential controls.

In assessing executive summary, authorities must determine it should show unresolved limitations and transfer them to a continuing route. The institutional consequence follows from whether accountability then becomes an institutional capacity to explain and correct, rather than a contest over who can avoid blame. The central annual recommendation is a public responsibility record linking each material commitment to the entitled population, competent authority, partner contributions, finance, service milestone, distributional test, evidence and review date.

Key findings

    Scope and method

    Part I

    Annual accountability frame

    1

    Institutional status and continuity of responsibility

    Comparative interpretation of institutional status and continuity of responsibility depends upon existing responsibilities under human-rights, refugee and humanitarian law remain applicable according to their terms. The evidence must therefore clarify how national authorities retain the primary education role; international support should strengthen rather than obscure that responsibility. The Summit was a high-level global meeting convened by the United Nations Secretary-General. Its political commitments should be reported accurately without being treated as new binding law.[REF-29] [REF-30]

    Institutional action on institutional status and continuity of responsibility should be tested against public reporting should distinguish a Summit-wide record, a government pledge, an agency initiative and a financed agreement. This matters because the Chair's Summary is the Chair's account of the Summit. It can establish themes and announcements, but not universal consent to every statement made by participants. Individual commitments require their own issuing authority and evidence.

    2

    Population, movement and denominators

    population, movement and denominators cannot be judged without identifying each observes a different population. A proportionate conclusion must also recognise that duplicate and missing records should be expected where families move. The estimate should state reference date, place, status definition and revision. An emergency population changes over time and can be counted through registration, border, shelter, household, community and administrative sources.

    A defensible account of population, movement and denominators distinguishes a denominator restricted to applicants or camp residents cannot establish participation for an entire displaced population. This matters because host-community population change also matters where schools absorb rapid enrolment. Children outside registration remain within educational concern.

    3

    Safe and usable access

    The central question in safe and usable access is distance, transport, documentation, language, household cost, disability, gender-related risk and work or care duties can prevent attendance. For the learners concerned, the decisive consideration is whether the response should identify the barrier rather than classify non-attendance as lack of demand. A place is usable only where the learner can reach it safely and participate with necessary accommodation.

    safe and usable access requires a decision about reports should distinguish verified incident, reported concern and general exposure. The evidence must therefore clarify how protection information should be collected only for a clear purpose and handled under suitable safeguards. School safety includes the site, route, building, water and sanitation, safeguarding and conduct of personnel.

    Part I

    Purpose and interpretive frame

    1

    Institutional status and continuity of responsibility

    Comparative interpretation of institutional status and continuity of responsibility depends upon existing responsibilities under human-rights, refugee and humanitarian law remain applicable according to their terms. The evidence must therefore clarify how national authorities retain the primary education role; international support should strengthen rather than obscure that responsibility. The Summit was a high-level global meeting convened by the United Nations Secretary-General. Its political commitments should be reported accurately without being treated as new binding law.[REF-29] [REF-30]

    Institutional action on institutional status and continuity of responsibility should be tested against public reporting should distinguish a Summit-wide record, a government pledge, an agency initiative and a financed agreement. This matters because the Chair's Summary is the Chair's account of the Summit. It can establish themes and announcements, but not universal consent to every statement made by participants. Individual commitments require their own issuing authority and evidence.

    2

    Population, movement and denominators

    population, movement and denominators cannot be judged without identifying each observes a different population. A proportionate conclusion must also recognise that duplicate and missing records should be expected where families move. The estimate should state reference date, place, status definition and revision. An emergency population changes over time and can be counted through registration, border, shelter, household, community and administrative sources.

    A defensible account of population, movement and denominators distinguishes a denominator restricted to applicants or camp residents cannot establish participation for an entire displaced population. This matters because host-community population change also matters where schools absorb rapid enrolment. Children outside registration remain within educational concern.

    3

    Safe and usable access

    The central question in safe and usable access is distance, transport, documentation, language, household cost, disability, gender-related risk and work or care duties can prevent attendance. For the learners concerned, the decisive consideration is whether the response should identify the barrier rather than classify non-attendance as lack of demand. A place is usable only where the learner can reach it safely and participate with necessary accommodation.

    safe and usable access requires a decision about reports should distinguish verified incident, reported concern and general exposure. The evidence must therefore clarify how protection information should be collected only for a clear purpose and handled under suitable safeguards. School safety includes the site, route, building, water and sanitation, safeguarding and conduct of personnel.

    Part II

    Population and denominator

    4

    Teaching and learning continuity

    For teaching and learning continuity, the material distinction is between a receiving system may need bridging or language support while retaining a route to recognised national education. The evidence must therefore clarify how parallel curricula and records can impede later movement if compatibility is not planned. Curriculum continuity requires an account of prior learning, present language, intended progression and available time.

    Public responsibility for teaching and learning continuity begins with observation should focus on a bounded practice and lead to support. A contrary reading would overlook that one crisis-affected lesson cannot establish general teacher performance. Teachers may require concise guidance, joint planning, materials and coaching for mixed-age, multilingual or overcrowded classes.

    5

    Assessment, records and recognition

    Review of assessment, records and recognition is credible only where it explains missing certificates should not create an absolute bar to entry. For the learners concerned, the decisive consideration is whether provisional placement can draw on learner work, dialogue, prior history and observed performance, with a review date. Diagnostic evidence should guide placement and teaching without becoming an unannounced exclusion test.

    assessment, records and recognition cannot be judged without identifying reconstructed entries require provenance. A contrary reading would overlook that high-consequence certification needs stronger evidence than provisional entry, but institutional disruption should not be transferred entirely to the learner. Records should identify learner, programme, dates, curriculum, attendance, assessed learning and decision status.

    6

    Financing and the new fund

    financing and the new fund requires a decision about school-year recruitment, teacher payment and procurement have deadlines. A proportionate conclusion must also recognise that late disbursement can be recorded as financial execution while failing the intended educational period. Multi-year crises require a recurrent horizon. Emergency finance should align time with educational need.

    A defensible account of financing and the new fund distinguishes the Education Cannot Wait launch supplied a stated institutional ambition, not evidence of disbursement or effect. For the learners concerned, the decisive consideration is whether future monitoring should identify contributors, commitments, cash received, allocation criteria, implementing authority, populations reached, service quality and learning or continuity evidence. The conclusion for global education quality annual report 2017 should connect the finding to the competent authority, affected learners and evidence required at review.[REF-31]

    Part III

    Access and participation

    7

    Accountability and exit from temporary arrangements

    Institutional action on accountability and exit from temporary arrangements should be tested against temporary sites, condensed curricula, volunteer teaching or parallel records should not become a permanent lower standard for affected populations. The public account remains incomplete unless it explains how exit can mean integration, recognised transfer, return or closure with responsibility handed to an authorised body. Every emergency measure needs an owner, review date and transition condition.

    Review of accountability and exit from temporary arrangements is credible only where it explains a project ending does not establish educational recovery. The public account remains incomplete unless it explains how completion is a sustained, recognised route forward. Public reports should state unresolved populations and limitations. A favourable attendance rate among reached learners does not erase those never offered a place.

    8

    Attendance beyond enrolment

    In assessing attendance beyond enrolment, authorities must determine the study seeks evidence on actual participation during a stated recent period; it does not infer a learner's circumstances from a national or regional mean. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For attendance beyond enrolment, the first requirement is conceptual clarity.[REF-17] [REF-19]

    Public responsibility for attendance beyond enrolment begins with these are substantive attributes because they determine who can appear in the evidence. The public account remains incomplete unless it explains how a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is presence measured independently of registration status. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-20] [REF-21]

    Review of attendance beyond enrolment is credible only where it explains apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. A contrary reading would overlook that use of the indicator is bounded by the principle that frequency, season and reasons for absence retained. In assessing attendance beyond enrolment, authorities must determine a responsible commentary distinguishes observation, calculation and interpretation. A proportionate conclusion must also recognise that it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-23] [REF-24]

    A defensible account of attendance beyond enrolment distinguishes it should involve statistical judgement, legal safeguards and knowledge of the affected community. The resulting interpretation should show why suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For working children, carers and learners affected by illness, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-01] [REF-07]

    9

    Out-of-school status

    In assessing out-of-school status, authorities must determine without that purpose, disaggregation can multiply figures without improving public judgement. The resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of children in the relevant age group not participating at the defined level begins by naming the decision the evidence may inform.[REF-20] [REF-21]

    Public responsibility for out-of-school status begins with the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. A proportionate conclusion must also recognise that the preferred construction is a transparent residual from compatible population and participation concepts. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-23] [REF-24]

    out-of-school status requires a decision about the comparison should show the level for each group as well as any ratio or gap. For the learners concerned, the decisive consideration is whether a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: never-enrolled and formerly enrolled children separated.[REF-01] [REF-07]

    out-of-school status requires a decision about protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. A contrary reading would overlook that the distributional review must deliberately include children invisible to school registers. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-02] [REF-03]

    Part IV

    Progression and completion

    10

    Repetition and grade survival

    Evidence concerning repetition and grade survival should establish here, the relevant phenomenon is movement through grades without avoidable delay or exit, not the administrative convenience of the available categories. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of repetition and grade survival lies in making unequal educational experience observable.[REF-23] [REF-24]

    Institutional action on repetition and grade survival should be tested against this formulation requires the reporting body to preserve the population base and the observation period beside the result. The public account remains incomplete unless it explains how counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use cohort or reconstructed-cohort evidence with explicit assumptions.[REF-01] [REF-07]

    repetition and grade survival requires a decision about disparity measures should not replace the underlying distributions. This matters because a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The practical standard for repetition and grade survival concerns the comparison should identify the reference category but avoid presenting it as a natural norm. A contrary reading would overlook that policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that repeaters distinguished from re-entrants and transfers.[REF-02] [REF-03]

    Institutional action on repetition and grade survival should be tested against their circumstances may alter access to enumeration, classification and the service being measured. This matters because the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in overcrowded or intermittently operating schools.[REF-04] [REF-06]

    11

    Transition between education levels

    A defensible account of transition between education levels distinguishes its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns entry to the next level after completion of the preceding one.[REF-01] [REF-07]

    Review of transition between education levels is credible only where it explains analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This matters because this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on matched completion and new-entry populations over coherent periods. The definition should be fixed for the comparison at hand and deviations recorded.[REF-02] [REF-03]

    In assessing transition between education levels, authorities must determine national averages should remain available as context, yet never as a substitute for the distribution. This matters because nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: capacity constraints distinguished from learner attainment. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding.[REF-04] [REF-06]

    Public responsibility for transition between education levels begins with an adequate equity account asks whether rural learners and those unable to relocate are represented at each stage: population frame, collection, valid response, classification, analysis and publication. The public account remains incomplete unless it explains how attrition at any stage can produce an apparently complete indicator from a selective population. Public responsibility for transition between education levels begins with field arrangements need relevant languages, accessible formats and safe participation. The public account remains incomplete unless it explains how analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed.[REF-08] [REF-09]

    12

    Completion and educational entitlement

    Evidence concerning completion and educational entitlement should establish the substantive interest is finishing the final grade or meeting recognised programme requirements, observed for a population and period that are stated before calculation. The resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Completion and educational entitlement should be approached as a defined measurement problem.[REF-02] [REF-03]

    Review of completion and educational entitlement is credible only where it explains a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. For the learners concerned, the decisive consideration is whether operationally, the measure is completion defined separately from sitting or passing an examination. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence.[REF-04] [REF-06]

    For completion and educational entitlement, the material distinction is between it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. The resulting interpretation should show why it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that late completion and alternative pathways reported. A responsible commentary distinguishes observation, calculation and interpretation.[REF-08] [REF-09]

    For completion and educational entitlement, the material distinction is between suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. This matters because the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For over-age learners and people returning after interruption, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-10] [REF-12]

    Part V

    Learning and assessment

    13

    Minimum learning outcomes

    Evidence concerning minimum learning outcomes should establish a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The resulting interpretation should show why the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For minimum learning outcomes, the first requirement is conceptual clarity. The study seeks evidence on demonstrated knowledge or skill against a declared domain; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-04] [REF-06]

    A defensible account of minimum learning outcomes distinguishes the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. A proportionate conclusion must also recognise that the preferred construction is assessment evidence whose population and conditions are known. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-08] [REF-09]

    Comparative interpretation of minimum learning outcomes depends upon a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. The evidence must therefore clarify how reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: results interpreted with opportunity to learn and participation. The comparison should show the level for each group as well as any ratio or gap.[REF-10] [REF-12]

    minimum learning outcomes requires a decision about protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The institutional consequence follows from whether the distributional review must deliberately include learners taught in an unfamiliar language. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-13] [REF-14]

    14

    Assessment participation

    Institutional action on assessment participation should be tested against the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how a distribution-sensitive account of who was eligible, present, absent and excluded from testing begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-08] [REF-09]

    Institutional action on assessment participation should be tested against school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. The evidence must therefore clarify how reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use a participation profile accompanying every result distribution. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined.[REF-10] [REF-12]

    Evidence concerning assessment participation should establish policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The public account remains incomplete unless it explains how the governing caution is that non-participation never treated as low attainment or ignored. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-13] [REF-14]

    For assessment participation, the material distinction is between their circumstances may alter access to enumeration, classification and the service being measured. The public account remains incomplete unless it explains how the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners with disabilities and remote candidates.[REF-15] [REF-16]

    15

    Distribution of achievement

    For distribution of achievement, the material distinction is between here, the relevant phenomenon is variation across the full score or proficiency distribution, not the administrative convenience of the available categories. The institutional consequence follows from whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of distribution of achievement lies in making unequal educational experience observable.[REF-10] [REF-12]

    A defensible account of distribution of achievement distinguishes the definition should be fixed for the comparison at hand and deviations recorded. This matters because analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on percentiles and threshold shares beside a mean.[REF-13] [REF-14]

    The central question in distribution of achievement is nor should a group estimate be read as a description of every member. The resulting interpretation should show why within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: uncertainty and scale properties stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-15] [REF-16]

    Evidence concerning distribution of achievement should establish missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. A contrary reading would overlook that an adequate equity account asks whether learners concentrated below a minimum proficiency threshold are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-17] [REF-19]

    Part VI

    Gender and household resources

    16

    Gender parity and its limits

    Comparative interpretation of gender parity and its limits depends upon the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether the indicator question concerns differences between girls and boys in access, progression and learning. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-13] [REF-14]

    Institutional action on gender parity and its limits should be tested against a census figure should disclose enumeration rules. The institutional consequence follows from whether these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is female-to-male ratios read with levels and absolute gaps. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-15] [REF-16]

    gender parity and its limits requires a decision about it does not assign cause from a cross-sectional difference. The institutional consequence follows from whether apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that parity not confused with adequacy for either group. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-17] [REF-19]

    For gender parity and its limits, the material distinction is between this balance is contextual rather than mechanical. The institutional consequence follows from whether it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For girls in poor rural households and boys exposed to hazardous work, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-20] [REF-21]

    17

    Household wealth gradients

    The practical standard for household wealth gradients concerns the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether household wealth gradients should be approached as a defined measurement problem. The substantive interest is education outcomes across relative household resource groups, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-15] [REF-16]

    household wealth gradients requires a decision about it should require examination of definitions, timing, migration, duplication and non-response. The evidence must therefore clarify how the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a documented asset or consumption classification within each setting. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-17] [REF-19]

    Review of household wealth gradients is credible only where it explains explanation requires evidence on institutions, resources, households and prior conditions. This matters because interpretation follows this limitation: wealth ranks not assumed equivalent across countries or time. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-20] [REF-21]

    household wealth gradients cannot be judged without identifying coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. The public account remains incomplete unless it explains how if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children in the poorest quintile and those near classification boundaries. These populations may be missing not only from good outcomes but from the denominator itself.[REF-23] [REF-24]

    18

    Costs borne by households

    For costs borne by households, the material distinction is between the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A contrary reading would overlook that for costs borne by households, the first requirement is conceptual clarity. The study seeks evidence on fees, materials, transport, clothing and foregone labour; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-17] [REF-19]

    Review of costs borne by households is credible only where it explains school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. The public account remains incomplete unless it explains how reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use participation read against direct and indirect education costs. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined.[REF-20] [REF-21]

    The practical standard for costs borne by households concerns the comparison should identify the reference category but avoid presenting it as a natural norm. The resulting interpretation should show why policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that nominal fee abolition checked against remaining expenditure. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-23] [REF-24]

    The central question in costs borne by households is their circumstances may alter access to enumeration, classification and the service being measured. The evidence must therefore clarify how the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for large families and households hit by economic crisis.[REF-01] [REF-07]

    Part VII

    Place and service geography

    19

    Rural and urban residence

    A defensible account of rural and urban residence distinguishes without that purpose, disaggregation can multiply figures without improving public judgement. The institutional consequence follows from whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of education outcomes by a declared settlement classification begins by naming the decision the evidence may inform.[REF-20] [REF-21]

    Institutional action on rural and urban residence should be tested against analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. The evidence must therefore clarify how this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on residence linked to service availability and travel conditions. The definition should be fixed for the comparison at hand and deviations recorded.[REF-23] [REF-24]

    The central question in rural and urban residence is national averages should remain available as context, yet never as a substitute for the distribution. The institutional consequence follows from whether nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: national rural definitions preserved and comparison limitations stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding.[REF-01] [REF-07]

    A defensible account of rural and urban residence distinguishes analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The institutional consequence follows from whether missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether remote villages, pastoral populations and peri-urban settlements are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-02] [REF-03]

    20

    Subnational administrative disparity

    subnational administrative disparity requires a decision about here, the relevant phenomenon is variation between provinces, districts or comparable areas, not the administrative convenience of the available categories. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of subnational administrative disparity lies in making unequal educational experience observable.[REF-23] [REF-24]

    A defensible account of subnational administrative disparity distinguishes a census figure should disclose enumeration rules. This matters because these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is area estimates with population size and precision. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-01] [REF-07]

    A defensible account of subnational administrative disparity distinguishes a responsible commentary distinguishes observation, calculation and interpretation. The public account remains incomplete unless it explains how it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that administrative rankings not mistaken for causal explanations.[REF-02] [REF-03]

    Comparative interpretation of subnational administrative disparity depends upon the public report can still state that a disparity was examined, whether action is required and which body will monitor it. The evidence must therefore clarify how for small districts and areas with incomplete reporting, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-04] [REF-06]

    21

    Distance, isolation and transport

    In assessing distance, isolation and transport, authorities must determine its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The institutional consequence follows from whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns physical accessibility of the nearest appropriate service.[REF-01] [REF-07]

    distance, isolation and transport cannot be judged without identifying the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The evidence must therefore clarify how the preferred construction is travel time, route safety and seasonal interruption rather than straight-line distance alone. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-02] [REF-03]

    Evidence concerning distance, isolation and transport should establish a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. The resulting interpretation should show why reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: household reports and facility mapping reconciled. The comparison should show the level for each group as well as any ratio or gap.[REF-04] [REF-06]

    The central question in distance, isolation and transport is protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The institutional consequence follows from whether the distributional review must deliberately include learners with limited mobility and communities cut off seasonally. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-08] [REF-09]

    Part VIII

    Disability, language and identity

    22

    Disability-sensitive education data

    Public responsibility for disability-sensitive education data begins with a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. For the learners concerned, the decisive consideration is whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Disability-sensitive education data should be approached as a defined measurement problem. The substantive interest is participation and learning by functional difficulty and support requirement, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-02] [REF-03]

    A defensible account of disability-sensitive education data distinguishes where estimates are revised, both the reason and effect of revision should remain accessible. A contrary reading would overlook that measurement should use questions designed for comparable reporting without defining the child by diagnosis alone. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-04] [REF-06]

    Public responsibility for disability-sensitive education data begins with disparity measures should not replace the underlying distributions. The resulting interpretation should show why a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that identification, environment and accommodation kept analytically distinct.[REF-08] [REF-09]

    Institutional action on disability-sensitive education data should be tested against where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. A proportionate conclusion must also recognise that where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners whose impairments are not recorded by schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness.[REF-10] [REF-12]

    23

    Language of home and instruction

    For language of home and instruction, the material distinction is between the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For language of home and instruction, the first requirement is conceptual clarity. The study seeks evidence on alignment between learner language, teaching and assessment; it does not infer a learner's circumstances from a national or regional mean.[REF-04] [REF-06]

    In assessing language of home and instruction, authorities must determine this distinction is essential for participation and progression. The institutional consequence follows from whether the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on language categories reflecting local use and instructional practice. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-08] [REF-09]

    In assessing language of home and instruction, authorities must determine national averages should remain available as context, yet never as a substitute for the distribution. The evidence must therefore clarify how nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: small groups not erased through broad national labels. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding.[REF-10] [REF-12]

    language of home and instruction cannot be judged without identifying missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The public account remains incomplete unless it explains how an adequate equity account asks whether minority-language and multilingual learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-13] [REF-14]

    24

    Ethnicity, indigeneity and protected identity

    A defensible account of ethnicity, indigeneity and protected identity distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that a distribution-sensitive account of disparity associated with historically excluded identity groups begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-08] [REF-09]

    ethnicity, indigeneity and protected identity cannot be judged without identifying its metadata should travel with every published value. The public account remains incomplete unless it explains how at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is lawful, voluntary and contextually meaningful classification.[REF-10] [REF-12]

    Comparative interpretation of ethnicity, indigeneity and protected identity depends upon apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The resulting interpretation should show why use of the indicator is bounded by the principle that self-identification protected and non-response reported. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-13] [REF-14]

    ethnicity, indigeneity and protected identity requires a decision about disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. The resulting interpretation should show why this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For communities exposed to discrimination or forced assimilation, a single group label may conceal important internal differences.[REF-15] [REF-16]

    Part IX

    Conflict, disaster and mobility

    25

    Education under conflict and insecurity

    The practical standard for education under conflict and insecurity concerns here, the relevant phenomenon is access, attendance and learning where violence alters service and movement, not the administrative convenience of the available categories. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of education under conflict and insecurity lies in making unequal educational experience observable.[REF-10] [REF-12]

    Institutional action on education under conflict and insecurity should be tested against a discrepancy is not resolved by selecting the more favourable source. The public account remains incomplete unless it explains how it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is location- and time-specific observation with explicit coverage gaps. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-13] [REF-14]

    The practical standard for education under conflict and insecurity concerns reference points therefore need substantive meaning. A proportionate conclusion must also recognise that where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: absence caused by insecurity distinguished from ordinary dropout. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-15] [REF-16]

    Review of education under conflict and insecurity is credible only where it explains protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. A proportionate conclusion must also recognise that the distributional review must deliberately include learners in insecure areas and host communities. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-17] [REF-19]

    27

    Refugees, displaced persons and migrants

    For refugees, displaced persons and migrants, the material distinction is between the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Refugees, displaced persons and migrants should be approached as a defined measurement problem. The substantive interest is educational participation across changing legal and residential situations, observed for a population and period that are stated before calculation.[REF-15] [REF-16]

    Comparative interpretation of refugees, displaced persons and migrants depends upon the definition should be fixed for the comparison at hand and deviations recorded. A contrary reading would overlook that analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on status, origin, current location and service access recorded separately.[REF-17] [REF-19]

    The central question in refugees, displaced persons and migrants is results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. A contrary reading would overlook that if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: mobility never converted into duplicate enrolment or unexplained disappearance.[REF-20] [REF-21]

    A defensible account of refugees, displaced persons and migrants distinguishes attrition at any stage can produce an apparently complete indicator from a selective population. This matters because field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether undocumented migrants, refugees and internally displaced learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-23] [REF-24]

    Part X

    School conditions and teachers

    28

    Teacher availability and distribution

    Comparative interpretation of teacher availability and distribution depends upon the measure should preserve both the observed level and the distribution relevant to the claim. For the learners concerned, the decisive consideration is whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For teacher availability and distribution, the first requirement is conceptual clarity. The study seeks evidence on access to competent teaching across schools and subjects; it does not infer a learner's circumstances from a national or regional mean.[REF-17] [REF-19]

    The practical standard for teacher availability and distribution concerns a census figure should disclose enumeration rules. The institutional consequence follows from whether these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is teachers present and assigned relative to learner and curriculum need. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-20] [REF-21]

    teacher availability and distribution requires a decision about a responsible commentary distinguishes observation, calculation and interpretation. The public account remains incomplete unless it explains how it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that payroll totals not substituted for classroom availability.[REF-23] [REF-24]

    For teacher availability and distribution, the material distinction is between disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. A contrary reading would overlook that this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For schools serving poor, remote or displaced communities, a single group label may conceal important internal differences.[REF-01] [REF-07]

    29

    Class size, multi-grade teaching and time

    Institutional action on class size, multi-grade teaching and time should be tested against a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of the instructional conditions experienced by learners begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-20] [REF-21]

    Institutional action on class size, multi-grade teaching and time should be tested against numerator, denominator, reference date, unit and exclusions should appear together. For the learners concerned, the decisive consideration is whether if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is class organisation, scheduled time and delivered time considered together.[REF-23] [REF-24]

    Comparative interpretation of class size, multi-grade teaching and time depends upon reference points therefore need substantive meaning. The resulting interpretation should show why where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: simple pupil-teacher ratios not treated as a complete quality measure. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-01] [REF-07]

    In assessing class size, multi-grade teaching and time, authorities must determine coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. A proportionate conclusion must also recognise that if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include early grades and mixed-age classes. These populations may be missing not only from good outcomes but from the denominator itself.[REF-02] [REF-03]

    30

    Materials, facilities and basic services

    In assessing materials, facilities and basic services, authorities must determine the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of materials, facilities and basic services lies in making unequal educational experience observable. Here, the relevant phenomenon is usable learning resources and safe, accessible school conditions, not the administrative convenience of the available categories.[REF-23] [REF-24]

    materials, facilities and basic services requires a decision about where estimates are revised, both the reason and effect of revision should remain accessible. For the learners concerned, the decisive consideration is whether measurement should use availability joined to condition, accessibility and regular use. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-01] [REF-07]

    Review of materials, facilities and basic services is credible only where it explains policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. This matters because the governing caution is that delivery counts checked against learner access. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-02] [REF-03]

    materials, facilities and basic services requires a decision about the review should record non-response, unknown status and excluded locations separately. The evidence must therefore clarify how combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in temporary or damaged premises. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-04] [REF-06]

    Part XI

    Finance and distribution

    31

    Public spending by level and function

    The central question in public spending by level and function is the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns resources assigned to education purposes across the system. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-01] [REF-07]

    Evidence concerning public spending by level and function should establish the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. The resulting interpretation should show why when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on expenditure classified by level, recurrent or capital use and responsible body. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-02] [REF-03]

    Review of public spending by level and function is credible only where it explains results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. The institutional consequence follows from whether if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: budgets, commitments and actual expenditure distinguished.[REF-04] [REF-06]

    Review of public spending by level and function is credible only where it explains attrition at any stage can produce an apparently complete indicator from a selective population. The evidence must therefore clarify how field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether basic education services under fiscal pressure are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-08] [REF-09]

    32

    Incidence of education spending

    The central question in incidence of education spending is the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Incidence of education spending should be approached as a defined measurement problem. The substantive interest is who benefits from publicly financed places and services, observed for a population and period that are stated before calculation.[REF-02] [REF-03]

    Public responsibility for incidence of education spending begins with a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. This matters because a survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is unit resources combined with participation across population groups. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions.[REF-04] [REF-06]

    incidence of education spending requires a decision about it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. For the learners concerned, the decisive consideration is whether it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that benefit estimates not treated as household income. A responsible commentary distinguishes observation, calculation and interpretation.[REF-08] [REF-09]

    A defensible account of incidence of education spending distinguishes suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. For the learners concerned, the decisive consideration is whether the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For groups excluded before public spending can reach them, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-10] [REF-12]

    33

    Protecting equity during fiscal constraint

    Evidence concerning protecting equity during fiscal constraint should establish a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For protecting equity during fiscal constraint, the first requirement is conceptual clarity. The study seeks evidence on whether reductions or delays fall disproportionately on weaker services; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-04] [REF-06]

    Review of protecting equity during fiscal constraint is credible only where it explains administrative records should be reconciled with population-based evidence where their coverage differs. A proportionate conclusion must also recognise that a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is dated finance and service indicators read together. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-08] [REF-09]

    In assessing protecting equity during fiscal constraint, authorities must determine statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. The resulting interpretation should show why explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: national totals tested against subnational allocation and household costs. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-10] [REF-12]

    The central question in protecting equity during fiscal constraint is protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The institutional consequence follows from whether the distributional review must deliberately include poor households and institutions with little financial reserve. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-13] [REF-14]

    Part XII

    Data sources and measurement error

    34

    Administrative records

    For administrative records, the material distinction is between without that purpose, disaggregation can multiply figures without improving public judgement. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of regular learner, staff, facility and finance information begins by naming the decision the evidence may inform.[REF-08] [REF-09]

    A defensible account of administrative records distinguishes school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. A proportionate conclusion must also recognise that reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use clear definitions, reporting coverage and revision history. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined.[REF-10] [REF-12]

    A defensible account of administrative records distinguishes percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The public account remains incomplete unless it explains how the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that non-reporting institutions kept visible in aggregates. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-13] [REF-14]

    The practical standard for administrative records concerns combining unknown observations with the majority group biases both estimates and obscures the weakness. For the learners concerned, the decisive consideration is whether where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for small, private, non-formal and emergency providers. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-15] [REF-16]

    35

    Household surveys

    The practical standard for household surveys concerns here, the relevant phenomenon is population-based evidence beyond enrolled learners, not the administrative convenience of the available categories. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of household surveys lies in making unequal educational experience observable.[REF-10] [REF-12]

    In assessing household surveys, authorities must determine the definition should be fixed for the comparison at hand and deviations recorded. A contrary reading would overlook that analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on probability samples, weights and field dates documented.[REF-13] [REF-14]

    Institutional action on household surveys should be tested against nor should a group estimate be read as a description of every member. The resulting interpretation should show why within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: sampling and non-response uncertainty carried into group comparisons. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-15] [REF-16]

    household surveys requires a decision about missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The institutional consequence follows from whether an adequate equity account asks whether small minorities and mobile households are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-17] [REF-19]

    36

    Censuses and population frames

    censuses and population frames requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that the indicator question concerns broad population coverage and small-area denominators. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-13] [REF-14]

    Public responsibility for censuses and population frames begins with these are substantive attributes because they determine who can appear in the evidence. For the learners concerned, the decisive consideration is whether a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is enumeration date, usual residence and institutional coverage stated. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-15] [REF-16]

    censuses and population frames cannot be judged without identifying it does not assign cause from a cross-sectional difference. This matters because apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that long intervals and under-enumeration acknowledged. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-17] [REF-19]

    Public responsibility for censuses and population frames begins with disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. For the learners concerned, the decisive consideration is whether this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For homeless, displaced and geographically isolated people, a single group label may conceal important internal differences.[REF-20] [REF-21]

    Part XIII

    Disaggregation and intersection

    37

    Single-axis disaggregation

    single-axis disaggregation cannot be judged without identifying the substantive interest is separate reporting by sex, wealth, residence or another characteristic, observed for a population and period that are stated before calculation. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Single-axis disaggregation should be approached as a defined measurement problem.[REF-15] [REF-16]

    single-axis disaggregation requires a decision about a discrepancy is not resolved by selecting the more favourable source. A proportionate conclusion must also recognise that it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is levels, gaps and denominators shown for each category. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-17] [REF-19]

    The practical standard for single-axis disaggregation concerns statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. For the learners concerned, the decisive consideration is whether explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: one axis not presented as a complete account of marginalisation. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-20] [REF-21]

    single-axis disaggregation requires a decision about if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The institutional consequence follows from whether confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include groups whose disadvantage lies on another unmeasured dimension. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-23] [REF-24]

    38

    Intersecting categories

    Public responsibility for intersecting categories begins with the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For intersecting categories, the first requirement is conceptual clarity. The study seeks evidence on joint distributions such as sex by wealth and residence; it does not infer a learner's circumstances from a national or regional mean.[REF-17] [REF-19]

    Institutional action on intersecting categories should be tested against source coverage must be tested before sources are combined. For the learners concerned, the decisive consideration is whether school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use pre-specified combinations with sufficient observations. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone.[REF-20] [REF-21]

    intersecting categories cannot be judged without identifying disparity measures should not replace the underlying distributions. A contrary reading would overlook that a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that empty or unstable cells reported honestly.[REF-23] [REF-24]

    In assessing intersecting categories, authorities must determine where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. For the learners concerned, the decisive consideration is whether particular scrutiny is required for poor rural girls, disabled learners in remote areas and displaced minorities. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-01] [REF-07]

    39

    Small numbers, disclosure and reliability

    Comparative interpretation of small numbers, disclosure and reliability depends upon without that purpose, disaggregation can multiply figures without improving public judgement. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of useful detail without unreliable estimates or identification begins by naming the decision the evidence may inform.[REF-20] [REF-21]

    The central question in small numbers, disclosure and reliability is this distinction is essential for participation and progression. The public account remains incomplete unless it explains how the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on suppression, aggregation or qualitative evidence chosen proportionately. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-23] [REF-24]

    The practical standard for small numbers, disclosure and reliability concerns within-group variation and unmeasured intersecting conditions remain material. The evidence must therefore clarify how the analytical rule is clear: confidentiality decisions separated from claims that no disparity exists. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-01] [REF-07]

    A defensible account of small numbers, disclosure and reliability distinguishes missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The public account remains incomplete unless it explains how an adequate equity account asks whether small communities and learners with rare characteristics are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-02] [REF-03]

    Part XIV

    Comparison, uncertainty and change

    40

    Comparing unlike systems

    Comparative interpretation of comparing unlike systems depends upon a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. For the learners concerned, the decisive consideration is whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of comparing unlike systems lies in making unequal educational experience observable. Here, the relevant phenomenon is cross-country patterns based on harmonised but bounded concepts, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-23] [REF-24]

    The central question in comparing unlike systems is at minimum this includes population, geography, date, collection method, classification and known exclusions. The institutional consequence follows from whether a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is metadata tests before numerical comparison. Its metadata should travel with every published value.[REF-01] [REF-07]

    Comparative interpretation of comparing unlike systems depends upon a responsible commentary distinguishes observation, calculation and interpretation. A contrary reading would overlook that it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that differences in programme structure and classification remain visible.[REF-02] [REF-03]

    Public responsibility for comparing unlike systems begins with the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For the learners concerned, the decisive consideration is whether for countries with incomplete or rapidly changing systems, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-04] [REF-06]

    41

    Sampling error and other uncertainty

    sampling error and other uncertainty requires a decision about the resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns the range of values reasonably compatible with the observations. For sampling error and other uncertainty, the material distinction is between its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-01] [REF-07]

    Review of sampling error and other uncertainty is credible only where it explains it should require examination of definitions, timing, migration, duplication and non-response. The resulting interpretation should show why the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is standard errors, design effects and data-quality qualifications. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-02] [REF-03]

    Review of sampling error and other uncertainty is credible only where it explains the comparison should show the level for each group as well as any ratio or gap. The resulting interpretation should show why a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: rank differences smaller than uncertainty not interpreted.[REF-04] [REF-06]

    A defensible account of sampling error and other uncertainty distinguishes coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. For the learners concerned, the decisive consideration is whether if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include small disaggregated populations. These populations may be missing not only from good outcomes but from the denominator itself.[REF-08] [REF-09]

    Part XV

    Responsible interpretation and action

    43

    Reading disparity without blaming learners

    Evidence concerning reading disparity without blaming learners should establish the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether for reading disparity without blaming learners, the first requirement is conceptual clarity. The study seeks evidence on institutional and social conditions associated with unequal outcomes; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-04] [REF-06]

    Institutional action on reading disparity without blaming learners should be tested against the definition should be fixed for the comparison at hand and deviations recorded. The institutional consequence follows from whether analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on descriptive findings separated from causal claims.[REF-08] [REF-09]

    Public responsibility for reading disparity without blaming learners begins with results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. The public account remains incomplete unless it explains how if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: group identity never treated as a mechanism by itself.[REF-10] [REF-12]

    The practical standard for reading disparity without blaming learners concerns analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. For the learners concerned, the decisive consideration is whether missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether communities subject to stigma are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-13] [REF-14]

    44

    Turning evidence into equitable policy

    Institutional action on turning evidence into equitable policy should be tested against a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The public account remains incomplete unless it explains how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of a decision rule linking disparity to service, finance or legal responsibility begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]

    For turning evidence into equitable policy, the material distinction is between a survey estimate should disclose weights and uncertainty. The evidence must therefore clarify how a census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is baseline, intended reach, implementation evidence and review date. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions.[REF-10] [REF-12]

    turning evidence into equitable policy cannot be judged without identifying it does not assign cause from a cross-sectional difference. A contrary reading would overlook that apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that targets accompanied by distributional safeguards. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-13] [REF-14]

    The practical standard for turning evidence into equitable policy concerns the public report can still state that a disparity was examined, whether action is required and which body will monitor it. A contrary reading would overlook that for learners farthest below a secured minimum, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-15] [REF-16]

    45

    A public account beyond the average

    Institutional action on a public account beyond the average should be tested against the measure should preserve both the observed level and the distribution relevant to the claim. For the learners concerned, the decisive consideration is whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of a public account beyond the average lies in making unequal educational experience observable. Here, the relevant phenomenon is a concise national statement of level, distribution, missingness and remedy, not the administrative convenience of the available categories.[REF-10] [REF-12]

    In assessing a public account beyond the average, authorities must determine a discrepancy is not resolved by selecting the more favourable source. For the learners concerned, the decisive consideration is whether it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is national totals presented beside selected group and place indicators. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-13] [REF-14]

    The practical standard for a public account beyond the average concerns a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. For the learners concerned, the decisive consideration is whether reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: progress claims bounded by evidence coverage and unresolved gaps. The comparison should show the level for each group as well as any ratio or gap.[REF-15] [REF-16]

    Evidence concerning a public account beyond the average should establish confidentiality is essential, especially where identity or status creates risk. A contrary reading would overlook that protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include every learner otherwise hidden by a successful average. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-17] [REF-19]

    Part XVI

    Applied distributional analysis

    46

    Composite measures and the loss of meaning

    Institutional action on composite measures and the loss of meaning should be tested against the national total supplies context, while the distribution tests whether the total is shared. The institutional consequence follows from whether where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is combining several dimensions into a summary measure. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises participation, completion, learning and school conditions, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement.[REF-01] [REF-03]

    Evidence concerning composite measures and the loss of meaning should establish choices on these matters are not neutral presentation details. For the learners concerned, the decisive consideration is whether they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern weights, normalisation and substitution.[REF-02] [REF-04]

    For composite measures and the loss of meaning, the material distinction is between avoiding it requires the underlying counts and distributions to remain visible beside any summary. A contrary reading would overlook that analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is a low value in one dimension being concealed by a high value in another.[REF-05] [REF-06]

    In assessing composite measures and the loss of meaning, authorities must determine surveys can represent households beyond formal education, subject to sample size, field access and response. The institutional consequence follows from whether censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-08] [REF-12]

    Public responsibility for composite measures and the loss of meaning begins with these duties do not justify silence about a serious disparity. The institutional consequence follows from whether the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability.[REF-13] [REF-16]

    composite measures and the loss of meaning cannot be judged without identifying every response should name the expected population reach and the later observation that will test it. A contrary reading would overlook that otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is policy-makers deciding whether a summary aids or obscures resource allocation. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves.[REF-17] [REF-19]

    Evidence concerning composite measures and the loss of meaning should establish feedback needs a recorded route into classification, service design or further enquiry. The institutional consequence follows from whether people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-21] [REF-23]

    Review of composite measures and the loss of meaning is credible only where it explains changes to definitions, boundaries or population estimates should appear at the point where a series changes. The institutional consequence follows from whether earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review.[REF-23] [REF-24]

    47

    Decomposing an observed education gap

    The practical standard for decomposing an observed education gap concerns where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. For the learners concerned, the decisive consideration is whether the analytical purpose is examining how a national disparity is distributed across places and population groups. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises within-group and between-group differences, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-02] [REF-04]

    For decomposing an observed education gap, the material distinction is between a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. A proportionate conclusion must also recognise that if the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern population shares, outcome levels and overlapping membership. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion.[REF-05] [REF-06]

    The central question in decomposing an observed education gap is analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. The institutional consequence follows from whether a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is treating a descriptive decomposition as proof of cause. Avoiding it requires the underlying counts and distributions to remain visible beside any summary.[REF-08] [REF-12]

    A defensible account of decomposing an observed education gap distinguishes credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. For the learners concerned, the decisive consideration is whether every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities locating where further enquiry and action are warranted. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-21] [REF-23]

    48

    Setting distribution-sensitive targets

    setting distribution-sensitive targets requires a decision about the national total supplies context, while the distribution tests whether the total is shared. This matters because where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is expressing progress as improvement in secured minimums and unjustified gaps. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises the national level, the least-served group and the lower tail of the distribution, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement.[REF-05] [REF-06]

    Public responsibility for setting distribution-sensitive targets begins with documentation should also state which observations are direct, which are estimated and which are unavailable. The resulting interpretation should show why missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern baseline stability, ambition and safeguards against exclusion. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable.[REF-08] [REF-12]

    Institutional action on setting distribution-sensitive targets should be tested against avoiding it requires the underlying counts and distributions to remain visible beside any summary. The evidence must therefore clarify how analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is meeting a mean target while abandoning those farthest behind.[REF-13] [REF-16]

    Review of setting distribution-sensitive targets is credible only where it explains evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The evidence must therefore clarify how the certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is governments linking national commitments to subnational delivery.[REF-23] [REF-24]

    49

    Linking learners to service geography

    Public responsibility for linking learners to service geography begins with that purpose should be written before the calculation because method follows the intended inference. A proportionate conclusion must also recognise that the evidence base comprises settlement populations, travel conditions, schools, teachers and programme levels, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is relating participation and learning to the location and capacity of education services.[REF-08] [REF-12]

    linking learners to service geography requires a decision about they determine how strongly one component, place or population can influence the conclusion. The institutional consequence follows from whether a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern geographical scale, boundary effects and facility catchments. Choices on these matters are not neutral presentation details.[REF-13] [REF-16]

    For linking learners to service geography, the material distinction is between avoiding it requires the underlying counts and distributions to remain visible beside any summary. A proportionate conclusion must also recognise that analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is assuming the nearest mapped institution is accessible or appropriate.[REF-17] [REF-19]

    Comparative interpretation of linking learners to service geography depends upon the certainty required depends upon the consequence. This matters because credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is planners choosing sites, transport support and teacher deployment. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-01] [REF-03]

    50

    Reconciling conflicting sources

    In assessing reconciling conflicting sources, authorities must determine that purpose should be written before the calculation because method follows the intended inference. The public account remains incomplete unless it explains how the evidence base comprises coverage, timing, concepts and reporting incentives, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is interpreting differences between administrative, survey and census estimates.[REF-13] [REF-16]

    Institutional action on reconciling conflicting sources should be tested against documentation should also state which observations are direct, which are estimated and which are unavailable. This matters because missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a documented comparison of population and variable definitions. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable.[REF-17] [REF-19]

    Evidence concerning reconciling conflicting sources should establish direction must therefore be joined to adequacy. The resulting interpretation should show why comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. Review of reconciling conflicting sources is credible only where it explains no result should be described as equitable solely because one relative measure improved. For the learners concerned, the decisive consideration is whether the main interpretive danger is averaging incompatible estimates into an apparently precise figure. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens.[REF-21] [REF-23]

    Comparative interpretation of reconciling conflicting sources depends upon evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The institutional consequence follows from whether the certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is statistical authorities issuing one bounded account with visible uncertainty.[REF-02] [REF-04]

    51

    Monitoring marginalisation during severe disruption

    Public responsibility for monitoring marginalisation during severe disruption begins with where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. A proportionate conclusion must also recognise that the analytical purpose is maintaining useful distributional evidence when populations and services move rapidly. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises rapid counts, restored administrative returns and household evidence, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-17] [REF-19]

    Comparative interpretation of monitoring marginalisation during severe disruption depends upon if the headline conclusion changes materially, the range of results should be reported. The public account remains incomplete unless it explains how sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern dated estimates, revision practice and minimal essential classifications. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications.[REF-21] [REF-23]

    A defensible account of monitoring marginalisation during severe disruption distinguishes analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. The resulting interpretation should show why a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is using an unstable emergency denominator to assert durable improvement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary.[REF-23] [REF-24]

    Evidence concerning monitoring marginalisation during severe disruption should establish credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. The evidence must therefore clarify how every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities protecting access while rebuilding regular statistics. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-05] [REF-06]

    52

    Communicating uncertainty without losing urgency

    For communicating uncertainty without losing urgency, the material distinction is between where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. A contrary reading would overlook that where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is explaining what is known strongly enough to justify action and what remains unresolved. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises point estimates, ranges, quality statements and missing populations, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-21] [REF-23]

    Comparative interpretation of communicating uncertainty without losing urgency depends upon missing values should remain missing unless an explicit estimation method and its effect are shown. The resulting interpretation should show why the principal methodological questions concern plain institutional language joined to exact metadata. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable.[REF-23] [REF-24]

    communicating uncertainty without losing urgency cannot be judged without identifying avoiding it requires the underlying counts and distributions to remain visible beside any summary. The evidence must therefore clarify how analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is presenting caution as a reason for inaction or urgency as a reason for overstatement.[REF-01] [REF-03]

    In assessing communicating uncertainty without losing urgency, authorities must determine credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. The evidence must therefore clarify how every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is the public, affected communities and responsible decision-makers. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-08] [REF-12]

    53

    A national marginalisation profile

    a national marginalisation profile cannot be judged without identifying that purpose should be written before the calculation because method follows the intended inference. A proportionate conclusion must also recognise that the evidence base comprises population, access, progression, learning, conditions, finance and unresolved evidence gaps, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is assembling a concise recurring account beyond the national average.[REF-23] [REF-24]

    Review of a national marginalisation profile is credible only where it explains they determine how strongly one component, place or population can influence the conclusion. A contrary reading would overlook that a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a stable core with context-specific distributions. Choices on these matters are not neutral presentation details.[REF-01] [REF-03]

    Evidence concerning a national marginalisation profile should establish comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. A contrary reading would overlook that no result should be described as equitable solely because one relative measure improved. The main interpretive danger is creating an encyclopaedia of indicators without decision priority. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-02] [REF-04]

    Review of a national marginalisation profile is credible only where it explains every response should name the expected population reach and the later observation that will test it. A proportionate conclusion must also recognise that otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is parliament, ministries, local authorities and communities reviewing educational equity. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves.[REF-13] [REF-16]

    Part XVII

    Extended disparity interpretation

    54

    Minimum comparison threshold

    The central question in minimum comparison threshold is national means should be accompanied by lower-tail, threshold and group evidence. The resulting interpretation should show why opportunity to learn, participation and exclusions remain material. A learning-disparity comparison requires a declared construct, represented population, assessment conditions, scale, uncertainty and distribution.[REF-02] [REF-03]

    55

    Within-system interpretation

    A defensible account of within-system interpretation distinguishes counts, levels, absolute gaps and ratios answer different questions. A proportionate conclusion must also recognise that the central question in within-system interpretation is missing learners and non-participating schools must remain visible. A proportionate conclusion must also recognise that within-system gaps should preserve place, group and institutional context without assigning cause from identity.[REF-01] [REF-06]

    56

    Between-system interpretation

    A defensible account of between-system interpretation distinguishes rank differences smaller than uncertainty should not support categorical conclusions. The institutional consequence follows from whether harmonisation does not remove substantive system difference. Between-system comparison requires metadata tests for curriculum, age, language, sampling and assessment.[REF-02] [REF-03]

    Part XVIII

    Extended analysis of learning disparities

    57

    Assessment participation and the represented learning population

    Institutional action on assessment participation and the represented learning population should be tested against the target population, eligible population, sampled population and assessed population should be reported separately. The institutional consequence follows from whether learners can be absent because of illness, displacement, school non-attendance, language barriers, disability, conflict, administrative exclusion or ordinary sampling loss. These routes have different meanings. A result calculated only for participating learners may describe their performance accurately while failing to describe the educational system's full learner population. A learning distribution is defined partly by who participates in the assessment.[REF-01] [REF-02] [REF-03]

    Public responsibility for assessment participation and the represented learning population begins with school exclusion and within-school absence need separate observation where the design permits. A contrary reading would overlook that if institutions outside the frame differ systematically from included schools, a high learner response rate within sampled schools cannot repair the coverage limitation. Similarly, replacement of inaccessible schools can preserve sample size while changing the population represented. Reports should describe replacement rules and show any material effect on geography or institutional type. Participation should therefore accompany every reported mean, proficiency share or percentile.

    A defensible account of assessment participation and the represented learning population distinguishes coding an absent learner as below threshold invents performance evidence; removing every absence without comment creates a different bias. For the learners concerned, the decisive consideration is whether sensitivity analysis can examine plausible bounds or compare known characteristics of participants and non-participants. Where a numerical adjustment is not defensible, the limitation remains a substantive finding. The absence of learning evidence for a population requiring education attention should not be interpreted as evidence of no disparity. Non-participation should not be assigned a score.

    Evidence concerning assessment participation and the represented learning population should establish exemption practices should be reported by reason and learner group. A proportionate conclusion must also recognise that where an adapted form changes the construct materially, separate interpretation may be necessary. Where it removes an irrelevant barrier, results may belong in the common distribution. The report should explain that judgement rather than treat all adaptations as either incomparable or automatically identical. Accommodation and language arrangements affect participation and result validity. An assessment may permit attendance yet fail to elicit the intended construct if the format, communication or response mode is inaccessible.[REF-05] [REF-10] [REF-12]

    For assessment participation and the represented learning population, the material distinction is between a low response among remote schools may require field and service improvement; absence among out-of-school children may require a population-based study and re-entry action; assessment exclusion may require accessible design. This matters because these are not corrections to be made solely through statistical weighting. The responsible education body should identify which population remains unrepresented and when better evidence will be available. A national learning claim should be no broader than that coverage. Public interpretation should connect participation to policy.

    58

    Scale, threshold and distribution comparability

    Public responsibility for scale, threshold and distribution comparability begins with a numerical score has meaning through the tasks, response model, scoring rules and population for which interpretation has been supported. The resulting interpretation should show why equal numerical differences should not be assumed to represent equal educational differences unless the scale warrants that inference. A threshold adds a substantive judgement about the knowledge or capability learners should demonstrate; it should not be selected merely because it divides the sample conveniently. Comparisons require a clear statement of what the assessment scale represents.[REF-02] [REF-03]

    Public responsibility for scale, threshold and distribution comparability begins with the same mean can accompany a compressed distribution, a wide lower tail or polarisation. The resulting interpretation should show why reports should therefore consider percentiles, threshold shares and dispersion where technically sound. The lower part of the distribution is especially important for minimum learning opportunity, while the upper part may reveal whether expansion has altered advanced performance. These summaries should remain tied to uncertainty and assessment coverage. Means provide one description of the centre and can conceal change elsewhere.

    Public responsibility for scale, threshold and distribution comparability begins with if standards are reset, the break should be visible at the year of change and parallel results reported where possible. For the learners concerned, the decisive consideration is whether descriptions such as basic, adequate or advanced should be treated as definitions under the assessment, not universal attributes of a learner or education system. Threshold comparisons need stable standard-setting and clear labels. A change in the percentage above a threshold may reflect learning, scale revision, task composition or population change.

    In assessing scale, threshold and distribution comparability, authorities must determine the permissible inference should be stated accordingly. For the learners concerned, the decisive consideration is whether cross-language and cross-cultural comparability require evidence, not an assumption that translation has preserved difficulty and meaning. Task familiarity, curriculum exposure and response conventions can affect results. Review should examine translation, adaptation, differential item behaviour and opportunity to learn, while avoiding the claim that every detected difference invalidates the whole assessment. Some comparisons may remain defensible at a broad domain while narrower subscales do not.

    The central question in scale, threshold and distribution comparability is league order does not answer those questions. The institutional consequence follows from whether a rank is usually less informative than the estimated difference, uncertainty and distribution. Small rank movement can follow changes in participating systems or sampling variation. Public reporting should avoid categorical language when intervals overlap or scale linkage is weak. The useful question is whether the evidence indicates a material disparity requiring enquiry, what population is affected and which educational condition might be changed.

    59

    Opportunity to learn and interpretation of achievement gaps

    Comparative interpretation of opportunity to learn and interpretation of achievement gaps depends upon opportunity does not determine performance completely, and its measurement is imperfect. A proportionate conclusion must also recognise that it nevertheless prevents a learning gap from being attributed solely to learners or households when institutions supplied different educational conditions. Achievement evidence should be interpreted with opportunity to learn. This includes curriculum entitlement, content actually taught, instructional time, teacher availability, language, materials, attendance and access to support.[REF-01] [REF-13] [REF-14]

    opportunity to learn and interpretation of achievement gaps cannot be judged without identifying teacher self-report may be influenced by recall or expectations; observation covers a short period; work samples are selective. The evidence must therefore clarify how agreement across sources strengthens interpretation when dates and populations align. Contradictions can identify local variation or weak measurement and should not be resolved by choosing the account most favourable to the system. The official curriculum establishes intended opportunity, not delivered instruction. Teacher reports, schedules, classroom observation and learner work can add evidence, each with limitations.

    Review of opportunity to learn and interpretation of achievement gaps is credible only where it explains closures, teacher absence, shortened shifts and late entry can reduce delivered or usable time. The public account remains incomplete unless it explains how total hours also conceal subject allocation and teaching quality. A learner receiving more hours of poorly organised instruction does not necessarily have greater opportunity in the intended domain. Time is therefore one component of the explanatory evidence, not a conversion factor for predicted score. Instructional time should distinguish scheduled, delivered and attended time.

    The central question in opportunity to learn and interpretation of achievement gaps is these conditions should be examined at the level where the learning evidence was collected. The resulting interpretation should show why national resource averages can obscure concentration of weak provision among the same learners whose scores form the lower tail. Resource indicators should be connected to use. Textbooks delivered to a school may not be available in the relevant language, grade or classroom. Teacher qualifications recorded administratively may not match subject assignment. Facilities may exist but be inaccessible or unsafe.

    A defensible account of opportunity to learn and interpretation of achievement gaps distinguishes the original disparity and the remedial conditions should remain visible rather than being erased by a new cohort average. For the learners concerned, the decisive consideration is whether policy conclusions should avoid treating opportunity indicators as excuses for low expectations. Their purpose is to identify conditions within public responsibility and to design support. Where learners received less curriculum exposure, a response may include additional teaching, staff deployment, accessible materials or revised pacing. Later assessment should test whether opportunity and learning changed.

    60

    Decomposing disparities within education systems

    A defensible account of decomposing disparities within education systems distinguishes a large within-school component can coexist with institutional inequality and should not be read as evidence that schools are irrelevant. The public account remains incomplete unless it explains how a national disparity can reflect differences between regions, schools, classrooms and learners. Decomposition can describe where variation is concentrated, provided it is not treated as proof of cause. A large between-school component may indicate segregation, resource distribution or residential pattern; it does not identify which mechanism operates.[REF-01] [REF-09] [REF-14]

    The practical standard for decomposing disparities within education systems concerns reports should state exclusions and avoid presenting a modelled residual as direct observation. This matters because the level of analysis should correspond to the sampling and decision structure. Learners are commonly nested within classes and schools, while policies may operate through districts or providers. Standard errors and models should respect clustering. Small schools and sparsely populated areas may require special treatment, but removal changes the population represented.

    For decomposing disparities within education systems, the material distinction is between adjustment for a condition influenced by the school can also remove part of the very effect under review. This matters because the purpose and causal assumptions need explanation. Group composition affects school comparisons. A raw school mean combines prior opportunity, intake, mobility, attendance and current teaching. Adjusted measures can answer bounded questions but depend on variables and assumptions. They should not replace the unadjusted learner outcome or become a definitive quality rank.

    In assessing decomposing disparities within education systems, authorities must determine maps and rankings, where used elsewhere, should not expose small communities or imply a boundary creates the disparity. This matters because geographical decomposition should preserve absolute numbers and service context. A small district with a severe gap may need urgent support even though it contributes little to national variance. A populous area with a modest gap may represent many learners. Policy priority should consider educational severity, population, rights and feasibility rather than statistical contribution alone.

    A defensible account of decomposing disparities within education systems distinguishes between-region evidence may lead to allocation review; between-school evidence may lead to staffing, admissions or support enquiry; within-school evidence may lead to classroom, language or accessibility review. A contrary reading would overlook that each hypothesis requires additional evidence. The decomposition locates questions; it does not authorise blame. Follow-up should state the condition examined, action taken and later learning evidence. The appropriate outcome is a decision agenda.

    61

    Bounded conclusions between education systems

    Comparative interpretation of bounded conclusions between education systems depends upon metadata review should precede numerical comparison. The resulting interpretation should show why where a material difference cannot be reconciled, the systems may still be described separately without a common rank. Between-system comparison serves public learning when it identifies patterns, plausible questions and alternative institutional arrangements. It becomes misleading when harmonised labels conceal different programme structures, ages, curricula, languages, participation or assessment conditions.[REF-02] [REF-03] [REF-16]

    bounded conclusions between education systems requires a decision about within-system distributions can overlap substantially even where means differ. For the learners concerned, the decisive consideration is whether group composition and population coverage matter. An apparent national advantage may not extend to poor, rural, minority-language or disabled learners. Reports should place distributional evidence beside the system result and avoid using nationality as an explanation. Country and system averages should not be interpreted as attributes of every school or learner.

    The practical standard for bounded conclusions between education systems concerns a linked scale can support trend if common items or other methods preserve meaning and security, but linkage error should accompany the estimate. A proportionate conclusion must also recognise that a later higher score should not automatically be described as system improvement where the represented population changed materially. Temporal comparison requires stable linkage. Changes in curriculum, assessment mode, participation, sampling frame or system boundaries can create discontinuity.

    bounded conclusions between education systems requires a decision about pilots should state the mechanism and review evidence. This matters because adoption based solely on rank proximity or reputation substitutes imitation for analysis. Policy borrowing should attend to authority, capacity, sequence and context. A practice associated with high performance elsewhere may depend upon teacher preparation, finance, curriculum coherence or social conditions absent in the receiving system. The comparison can identify an option, not guarantee its effect.

    Review of bounded conclusions between education systems is credible only where it explains interpretation identifies patterns and limitations. A proportionate conclusion must also recognise that policy consideration proposes further enquiry or action under national authority. Causal judgement requires additional design. This separation permits strong public concern about a learning disparity without false certainty about its source and protects education systems from both complacency and unsupported prescription. The final international statement should distinguish three levels. Recorded fact describes the observed results under stated methods.

    62

    Uncertainty, materiality and the duty to respond

    The practical standard for uncertainty, materiality and the duty to respond concerns the public report should identify which uncertainty could alter the decision and which does not affect the direction of urgent protection. For the learners concerned, the decisive consideration is whether uncertainty should qualify a disparity claim without neutralising it. Sampling error, non-response, scale linkage, classification and model choice each affect the range of defensible conclusions. They should be described separately because they support different remedies. A larger sample may reduce sampling error; it does not correct systematic exclusion or an invalid construct. More decimal places cannot repair weak coverage.[REF-02] [REF-03]

    Evidence concerning uncertainty, materiality and the duty to respond should establish the threshold for action also depends upon reversibility. The evidence must therefore clarify how additional diagnostic support can be introduced and reviewed more readily than a high-stakes classification of schools or learners. Statistical significance should not substitute for educational materiality. A small estimated difference may be precise but have limited practical consequence, while an uncertain large gap affecting a protected minimum may require immediate enquiry. Materiality should consider the knowledge or capability involved, the number of learners, distribution, duration and consequences for later progression.

    Comparative interpretation of uncertainty, materiality and the duty to respond depends upon where participation is selective, improve coverage and examine barriers. The institutional consequence follows from whether where opportunity to learn differs, address time, teachers, curriculum, language, accessibility or materials. Where scale comparability is weak, improve the assessment before publishing ranks. Where a group disparity persists under several definitions, investigate institutional mechanisms without attributing cause to identity. Every response should name the competent body, intended population, resources and review date. A remedy should match the evidence.

    uncertainty, materiality and the duty to respond cannot be judged without identifying a national evidence system demonstrates strength when it can acknowledge uncertainty, improve measurement and amend policy. The institutional consequence follows from whether the final test is whether learners receive better educational opportunity and whether remaining disparities continue to be visible rather than whether one annual figure becomes more favourable. Later evidence should be capable of changing the conclusion. Reports should preserve the original estimate, method and limitation, then state why revision occurred. Silent replacement makes apparent improvement impossible to distinguish from correction.

    Part XIX

    Targeted improvement planning

    63

    Defining a persistent learning gap

    Evidence concerning defining a persistent learning gap should establish the baseline should preserve the underlying distribution and absolute numbers. The resulting interpretation should show why if the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns a sustained disparity in a declared learning domain, population and period. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity.[REF-01] [REF-02]

    defining a persistent learning gap requires a decision about achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. For the learners concerned, the decisive consideration is whether authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is a fluctuating score or one cohort difference being treated as proof of persistence. Avoiding it requires evidence on both learning and the opportunity supplied.[REF-03] [REF-06]

    A defensible account of defining a persistent learning gap distinguishes assessment provides bounded learning evidence subject to coverage and validity. A proportionate conclusion must also recognise that observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes baseline, assessment coverage, uncertainty, distribution and opportunity to learn. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners.[REF-08] [REF-09]

    In assessing defining a persistent learning gap, authorities must determine teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. The public account remains incomplete unless it explains how local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to select a gap that is educationally material and within public influence. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced.[REF-13] [REF-14]

    The central question in defining a persistent learning gap is gender, household resources, disability, language, residence and prior opportunity may intersect. A proportionate conclusion must also recognise that a plan can improve its average by reaching learners closest to a threshold while leaving those farthest behind. Targets should therefore include the least-served position and protection against exclusion. Small groups require confidentiality and careful precision, not disappearance from review. Distributional review should ask who is eligible, offered support, participates, receives the intended intensity and demonstrates a later response.[REF-17] [REF-19]

    Comparative interpretation of defining a persistent learning gap depends upon leadership should route obstacles to bodies able to change staffing, finance, curriculum or assessment rather than leaving every correction to the classroom. The public account remains incomplete unless it explains how professional capability is central. Teachers need subject knowledge, worked examples, diagnostic interpretation and protected time to collaborate. Moderation should examine evidence and reasoning rather than force identical decisions for unlike cases. Staff workload and turnover should be monitored. A plan relying on a few exceptional individuals is not institutionally secure.[REF-23] [REF-24]

    Evidence concerning defining a persistent learning gap should establish a favourable later result may reflect population or assessment change and should be tested against the baseline metadata. This matters because where the intervention is ineffective, adaptation or cessation is responsible improvement. Where benefit depends on temporary support, institutionalisation requires recurrent finance and ordinary ownership. Completion is demonstrated by stronger learning opportunity and a functioning correction route, not by the end of a project. The public account should distinguish authorisation, delivery, use and learning consequence. It should report limitations, adverse effects and unresolved learners.[REF-01] [REF-02]

    64

    From diagnosis to an intervention hypothesis

    For from diagnosis to an intervention hypothesis, the material distinction is between the baseline should preserve the underlying distribution and absolute numbers. The institutional consequence follows from whether if the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns an explicit account of the condition expected to change learning. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity.[REF-03] [REF-06]

    A defensible account of from diagnosis to an intervention hypothesis distinguishes avoiding it requires evidence on both learning and the opportunity supplied. For the learners concerned, the decisive consideration is whether achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. Authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is group identity or low performance being mistaken for a causal explanation.[REF-08] [REF-09]

    A defensible account of from diagnosis to an intervention hypothesis distinguishes agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The institutional consequence follows from whether the required evidence includes curriculum exposure, teaching practice, language, time, materials and support. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers.[REF-13] [REF-14]

    For from diagnosis to an intervention hypothesis, the material distinction is between implementation should identify who acts, with what authority, resources and deadline. The institutional consequence follows from whether dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to state the mechanism and plausible alternatives before choosing activity.[REF-17] [REF-19]

    65

    Designing the targeted plan

    Public responsibility for designing the targeted plan begins with the plan should state which observation is direct, which is estimated and which remains unknown. A contrary reading would overlook that the improvement question concerns a bounded sequence connecting resources and actions to learner-facing change. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity. The baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined.[REF-08] [REF-09]

    The practical standard for designing the targeted plan concerns authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. The institutional consequence follows from whether alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is a list of activities replacing a coherent implementation logic. Avoiding it requires evidence on both learning and the opportunity supplied. Achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms.[REF-13] [REF-14]

    Institutional action on designing the targeted plan should be tested against each source should be used within scope. For the learners concerned, the decisive consideration is whether administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes responsibility, staff capability, learner support, milestones and correction.[REF-17] [REF-19]

    A defensible account of designing the targeted plan distinguishes dependencies should be sequenced. A contrary reading would overlook that teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to choose a feasible intensity and protect the common educational entitlement. Implementation should identify who acts, with what authority, resources and deadline.[REF-23] [REF-24]

    66

    Resourcing equitable implementation

    The central question in resourcing equitable implementation is a national average is insufficient when the plan addresses a local or group disparity. The public account remains incomplete unless it explains how the baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns the staff, time, materials, accessibility and finance required by the selected mechanism. It should be stated with the learner population, educational domain, geography and evidence period.[REF-13] [REF-14]

    Comparative interpretation of resourcing equitable implementation depends upon achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. The institutional consequence follows from whether authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is weak institutions being expected to implement with the same nominal allocation. Avoiding it requires evidence on both learning and the opportunity supplied.[REF-17] [REF-19]

    A defensible account of resourcing equitable implementation distinguishes agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. For the learners concerned, the decisive consideration is whether the required evidence includes recurrent cost, teacher workload, additional need and household burden. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers.[REF-23] [REF-24]

    The practical standard for resourcing equitable implementation concerns any departure should be recorded with its reason and expected learner consequence. A proportionate conclusion must also recognise that the governing action is to direct greater support where barriers and implementation costs are greater. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition.[REF-01] [REF-02]

    67

    Monitoring reach, quality and learning response

    Comparative interpretation of monitoring reach, quality and learning response depends upon the plan should state which observation is direct, which is estimated and which remains unknown. This matters because the improvement question concerns evidence that the intended learners received the intervention as designed and benefited educationally. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity. The baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined.[REF-17] [REF-19]

    A defensible account of monitoring reach, quality and learning response distinguishes alternative explanations should remain open until evidence discriminates among them. The institutional consequence follows from whether this discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is participation counts being treated as learning evidence. Avoiding it requires evidence on both learning and the opportunity supplied. Achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. Authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support.[REF-23] [REF-24]

    In assessing monitoring reach, quality and learning response, authorities must determine agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. A proportionate conclusion must also recognise that the required evidence includes eligibility, offer, take-up, dosage, teaching quality, work and assessment. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers.[REF-01] [REF-02]

    Public responsibility for monitoring reach, quality and learning response begins with implementation should identify who acts, with what authority, resources and deadline. A contrary reading would overlook that dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to combine timely implementation evidence with valid learning review.[REF-03] [REF-06]

    68

    Adaptation, institutionalisation and exit

    Evidence concerning adaptation, institutionalisation and exit should establish a national average is insufficient when the plan addresses a local or group disparity. The public account remains incomplete unless it explains how the baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns reasoned decisions to continue, change, scale or end the plan. It should be stated with the learner population, educational domain, geography and evidence period.[REF-23] [REF-24]

    Review of adaptation, institutionalisation and exit is credible only where it explains authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. A contrary reading would overlook that alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is temporary measures persisting without benefit or disappearing before durable capability exists. Avoiding it requires evidence on both learning and the opportunity supplied. Achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms.[REF-01] [REF-02]

    In assessing adaptation, institutionalisation and exit, authorities must determine agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. A contrary reading would overlook that the required evidence includes thresholds, adverse effects, unresolved cases, recurrent ownership and later evidence. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers.[REF-03] [REF-06]

    Review of adaptation, institutionalisation and exit is credible only where it explains local adaptation should remain possible within a common substantive condition. The institutional consequence follows from whether any departure should be recorded with its reason and expected learner consequence. The governing action is to retain useful capability while ending ineffective or inequitable arrangements. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism.[REF-08] [REF-09]

    Part XX

    National translation after adoption of the 2030 Agenda

    69

    Fixing the contemporaneous institutional baseline

    Institutional planning should begin from the position established by 10 December 2017. The Incheon Declaration expressed the education community's commitment to inclusive and equitable quality education and lifelong learning; Addis established a wider financing framework; the General Assembly adopted the 2030 Agenda; and the Education 2030 Framework for Action supplied an implementation reference available by the cut-off. These texts provide direction without proving that national or institutional delivery has occurred.[REF-25] [REF-26] [REF-27] [REF-28]

    fixing the contemporaneous institutional baseline requires a decision about each country should identify which obligations and targets can be acted upon under existing authority, which require legislative or administrative change, and which depend on clarification through later competent decisions. A proportionate conclusion must also recognise that unsettled detail should be recorded as such rather than filled by anticipation. The distinction between adoption and implementation is essential. Adoption establishes an agreed direction and permits governments to begin alignment. It does not prove that national law, plans, budgets, information or delivery arrangements already satisfy the commitment.[REF-07] [REF-12] [REF-27]

    The practical standard for fixing the contemporaneous institutional baseline concerns national authorities should map the newly adopted targets against those instruments and against the education stages, populations and institutions already in law and plans. The evidence must therefore clarify how this avoids two risks: abandoning useful evidence because terminology changed, and claiming continuity where a new target has broader scope or different educational substance. The baseline should preserve existing education commitments and national evidence. The new agenda does not erase the right to education, the unfinished Education for All undertaking, programme structures or established statistical series.[REF-08] [REF-09] [REF-25]

    Institutional action on fixing the contemporaneous institutional baseline should be tested against this is not an administrative inventory for its own sake. This matters because it prevents a proposed measure from being represented as an obligation, a declaration from being treated as proof of delivery, and a later decision from being projected backwards. The register should be revised openly as competent bodies act. A dated commitment register can support institutional accuracy. For each relevant proposition, it should record the adopting body, date, legal or policy status, national authority, present implementing instrument and unresolved question.[REF-17] [REF-19] [REF-27]

    Institutional action on fixing the contemporaneous institutional baseline should be tested against it also establishes a reliable point from which later implementation can be judged. This matters because public communication should use the same discipline. Governments can state that the 2030 Agenda has been adopted and that national alignment is commencing. They should not state that every indicator, national milestone or implementation mechanism has already been internationally settled. A precise account strengthens credibility because it makes clear which choices belong to national democratic and administrative processes and which follow directly from adopted commitments.[REF-25] [REF-26] [REF-27]

    70

    Selecting priorities without narrowing the commitment

    selecting priorities without narrowing the commitment cannot be judged without identifying a first improvement priority should be chosen because evidence shows a serious and remediable break, not because other elements have ceased to matter. A proportionate conclusion must also recognise that the breadth of Goal 4 requires sequencing, not selective abandonment. A national plan cannot improve every condition simultaneously, yet it should retain a complete map of early childhood, primary and secondary education, technical and vocational learning, tertiary participation, adult learning, relevant skills, equality, literacy, learning environments, scholarships and teachers as they appear in the adopted targets.[REF-25] [REF-27]

    For selecting priorities without narrowing the commitment, the material distinction is between the consequence test asks the scale and severity of the denial. For the learners concerned, the decisive consideration is whether the actionability test asks whether a competent authority has a plausible means of change. The equity test asks whether the measure will reach learners farthest from the secured opportunity. A priority that scores highly on visibility but weakly on consequence or equity should be reconsidered. The reasons for selection should be published beside the evidence and known limitations. Priority selection should apply four tests. The entitlement test asks which population and educational condition are at stake.[REF-01] [REF-03] [REF-12]

    Institutional action on selecting priorities without narrowing the commitment should be tested against analysis should retain the national level, absolute number affected, subnational and group distributions and the minimum educational floor. The public account remains incomplete unless it explains how it should also distinguish poor data from satisfactory conditions. Where the least-served population is weakly observed, strengthening coverage can itself become an immediate priority while urgent service evidence supports proportionate protection. National averages should not determine the sequence alone. A moderate national gap may conceal acute failure in one district or population; a large aggregate shortfall may require broad system expansion alongside targeted support.[REF-02] [REF-06] [REF-24]

    A defensible account of selecting priorities without narrowing the commitment distinguishes expansion of secondary places may depend on primary completion, trained staff and facilities. For the learners concerned, the decisive consideration is whether adult learning may require flexible provision, recognition and learner support. The plan should make these dependencies visible and decide which must precede or accompany the selected measure. Otherwise a high-level commitment can be converted into an isolated activity unable to change the learner-facing condition. Dependencies should influence sequence. An assessment reform cannot improve learning without curriculum alignment, teacher capability and participation.[REF-03] [REF-16] [REF-18]

    A defensible account of selecting priorities without narrowing the commitment distinguishes revision is not a retreat from ambition when reasons and consequences are published. The institutional consequence follows from whether it is a condition of responsible improvement. The complete commitment map should remain in view so that repeated concentration on one readily measured target does not produce silent neglect of lifelong learning, equality, educational quality or populations outside formal schooling. Selection should remain revisable. New evidence may show that the diagnosed mechanism was wrong, that another population is more severely affected or that implementation capacity is insufficient.[REF-08] [REF-25] [REF-27]

    71

    From global target to national improvement proposition

    Public responsibility for from global target to national improvement proposition begins with it should name the population, present educational condition, institutional mechanism, responsible body, resources, time and evidence of success. A contrary reading would overlook that for example, a commitment to equitable quality education is too broad to guide one implementation decision; a proposition to improve regular attendance for a defined remote population through transport, staffing and calendar changes can be examined and corrected. The narrower proposition remains connected to the universal commitment and should not be mistaken for its completion. A national improvement proposition should translate a broad target into a bounded statement of change.[REF-07] [REF-12] [REF-27]

    from global target to national improvement proposition requires a decision about consultation with teachers, learners and communities can identify mechanisms, but it should not replace representative population evidence when prevalence is claimed. A contrary reading would overlook that diagnosis should precede instrument choice. A low completion rate may reflect late entry, repetition, household cost, distance, school safety, language, disability exclusion, teacher shortage or unreliable records. These mechanisms require different responses. A government should compare administrative, household, assessment and local service evidence and state where inference remains uncertain.[REF-01] [REF-03] [REF-24]

    Evidence concerning from global target to national improvement proposition should establish additional materials will not improve learning if teachers lack time or knowledge to use them; professional guidance will not improve attendance where transport is decisive; a new indicator will not correct exclusion without authority and resources. For the learners concerned, the decisive consideration is whether the plan should identify necessary dependencies and foreseeable adverse effects. It should specify which part of the hypothesis is established by evidence and which remains to be tested during implementation. The intervention hypothesis should explain how the proposed measure changes the barrier.[REF-06] [REF-12] [REF-16]

    from global target to national improvement proposition cannot be judged without identifying the public record should show the relationship between the global target, national definition and selected measure, including any material difference from an existing series. This matters because national adaptation should preserve educational substance. Targets may need national definitions for programme levels, age groups, language and institutional responsibility. Adaptation is legitimate where it makes the commitment operational and comparable over time. It is not legitimate where it narrows the entitled population, lowers expectations for disadvantaged groups or converts learning into attendance alone.[REF-02] [REF-10] [REF-25]

    Institutional action on from global target to national improvement proposition should be tested against authorities should state what evidence would justify continuation, expansion, adaptation or cessation and when that decision will occur. The evidence must therefore clarify how a pilot that continues because it attracts support rather than because it changes the intended condition is not an improvement method. Conversely, an intervention should not be abandoned merely because early outcomes are uncertain where delivery has not reached intended intensity. Review must distinguish theory failure, implementation failure, measurement weakness and insufficient time. The proposition should end with a decision rule.[REF-08] [REF-11] [REF-23]

    72

    Aligning authority, finance and professional capability

    aligning authority, finance and professional capability requires a decision about one public owner should remain answerable for whether the learner-facing condition changes. This matters because implementation requires a chain of competent authority. National policy may set the priority, but regional administrations, municipalities, schools, training institutions or other bodies may control staffing, facilities and learner support. The plan should allocate each function to the body able to perform it and identify escalation where local authority is insufficient. Coordination should not allow responsibility to become diffuse.[REF-12] [REF-17] [REF-19]

    In assessing aligning authority, finance and professional capability, authorities must determine staff preparation, salaries, accessible materials, transport, maintenance, guidance, assessment, evidence and review may all be necessary. For the learners concerned, the decisive consideration is whether the Addis Ababa Action Agenda places national action within a broader financing context, but an international commitment does not supply a national cost estimate. Authorities should identify recurrent and capital requirements, the source and timing of funds, distribution rules and the conditions for continuity after temporary support. Finance should cover the complete intervention rather than visible start-up items.[REF-15] [REF-26]

    A defensible account of aligning authority, finance and professional capability distinguishes where households continue to bear material cost, formal fee policy should not be represented as full accessibility. For the learners concerned, the decisive consideration is whether allocation should respond to unequal cost and starting capacity. Equal per-learner funding can reproduce inequality where remoteness, disability, language, insecurity or weak infrastructure makes adequate provision more expensive. Formulae should state which need factors are recognised and should be checked against actual receipt. Announced expenditure is not proof of delivery: review should trace authorization, transfer, institutional use and learner consequence.[REF-06] [REF-10] [REF-13]

    Evidence concerning aligning authority, finance and professional capability should establish a plan dependent on exceptional individuals or uncompensated workload is not institutionally secure and can deepen disparity between strong and weak institutions. The resulting interpretation should show why teachers and institutional leaders require capability proportionate to the change. A new curriculum, assessment or inclusion expectation needs more than notification. Professional learning should provide subject substance, practical examples, time for collaboration and a route to support. Staffing and turnover should be examined in the locations expected to implement first.[REF-01] [REF-18]

    In assessing aligning authority, finance and professional capability, authorities must determine external support may finance initial expansion, evidence, technical work or regional learning, but roles, conditions and exit should be clear. The resulting interpretation should show why parallel activities and reporting arrangements can fragment public authority. Assistance should align with a national improvement proposition, use compatible records and establish how essential functions will enter recurrent provision. Its success should be judged by national capability and equitable learner opportunity, not the duration or visibility of the supporting activity. International cooperation should strengthen ordinary national capacity.[REF-22] [REF-26] [REF-27]

    73

    Monitoring delivery, reach and educational consequence

    monitoring delivery, reach and educational consequence requires a decision about these stages should not be collapsed. A proportionate conclusion must also recognise that a budget can be executed without materials arriving, a programme can operate without reaching disadvantaged learners, and participation can rise without improvement in learning or progression. Monitoring should follow the causal sequence of the improvement proposition. Inputs show whether resources and staff were available; delivery evidence shows whether the measure operated; reach shows which eligible learners participated and with what intensity; educational evidence shows whether the intended condition changed.[REF-03] [REF-06] [REF-12]

    In assessing monitoring delivery, reach and educational consequence, authorities must determine where a new target requires a new measure, authorities should preserve the preceding series and identify any break. The public account remains incomplete unless it explains how apparent improvement caused by revised population estimates, wider institutional reporting or changed assessment participation should be separated from educational change. Comparable trends are valuable, but continuity should not be asserted where concepts differ materially. The baseline should retain numerator, denominator, population, date, geography, definition and exclusions.[REF-02] [REF-23] [REF-24]

    A defensible account of monitoring delivery, reach and educational consequence distinguishes it should therefore report the least-served position and protect against exclusion or deterioration. A contrary reading would overlook that disaggregation must remain lawful, meaningful and safe. Small numbers may require controlled access or combined reporting, but the affected population and public responsibility should not disappear. Equity monitoring should show eligibility, offer, take-up, attendance, completion and outcome for relevant groups and places. A plan can improve its average by reaching learners already closest to the desired condition.[REF-01] [REF-10] [REF-13]

    monitoring delivery, reach and educational consequence requires a decision about agreement across sources strengthens a conclusion only after definitions and dates align; disagreement should guide investigation rather than selective reporting. The public account remains incomplete unless it explains how learning evidence requires population coverage and opportunity to learn. Assessment results should identify the domain, eligible population, participation and exclusions. A favourable mean among tested learners cannot represent those outside school or absent from assessment. Classroom observation, work samples and learner accounts can illuminate mechanism without establishing national prevalence.[REF-03] [REF-24]

    A defensible account of monitoring delivery, reach and educational consequence distinguishes the next review date and responsible authority should be visible. A contrary reading would overlook that this makes monitoring a means of correction rather than an obligation to produce favourable figures. The adopted agenda gains national credibility when evidence can alter action and expose populations who remain underserved. Public reporting should connect the result to a decision. A concise account should state what was implemented, who received it, what changed, what remains uncertain, which adverse effects occurred and whether the measure will continue, adapt, expand or cease.[REF-08] [REF-25] [REF-27]

    74

    First national decisions after adoption

    A defensible account of first national decisions after adoption distinguishes this creates a disciplined bridge between global adoption and national action. For the learners concerned, the decisive consideration is whether the immediate national decision is to establish governance for alignment without pretending that implementation detail is complete. Governments can designate a competent coordinating authority, preserve existing sector responsibilities, assemble a dated commitment register and commission a baseline review. The review should map every adopted education target against national law, plans, budgets and evidence. It should identify urgent gaps and unsettled definitions separately.[REF-25] [REF-27]

    Public responsibility for first national decisions after adoption begins with where credible evidence reveals severe exclusion or harm, protective action need not await a perfect estimate. This matters because longer-term allocation, however, should be reviewed as coverage improves. The plan should guard against choosing only learners and institutions most likely to produce rapid favourable results. A first priority should be small enough for accountable action and important enough to change educational opportunity. It should name the population and condition, state why it takes precedence, and preserve the wider commitment map.[REF-01] [REF-06] [REF-12]

    Comparative interpretation of first national decisions after adoption depends upon the national authority should state the source, timing and distribution of finance and how external cooperation relates to ordinary provision. The public account remains incomplete unless it explains how a financing gap should be described rather than hidden through reduced educational substance or transfer of cost to poor households. The first budget decision should identify recurrent implications. Temporary finance may permit testing, but teachers, learner support, accessible facilities, maintenance and evidence cannot be sustained by an announcement.[REF-15] [REF-26]

    Public responsibility for first national decisions after adoption begins with accurate status is a condition of accountable planning, not a reason for delay. This matters because the first public report should be candid about chronology. It can state that the principal education, financing and implementation texts named in this report were available by the cut-off. It should also state that later indicator settlements, technical revisions and results are outside the record. This protects the difference between an adopted commitment, a national policy choice, an institutional plan and later statistical clarification.[REF-25] [REF-26] [REF-27] [REF-28]

    In assessing first national decisions after adoption, authorities must determine the transition from global commitment to national improvement is complete only when ordinary institutions can sustain the changed condition and correct foreseeable departures; the end of a project or reporting period is not evidence of that result. The public account remains incomplete unless it explains how the first review should test whether institutions learned, not merely whether a plan was issued. It should examine authority, delivery, reach, professional capability, finance, learner experience and early educational consequence. It should record adverse findings and adapt the measure where the hypothesis or delivery proves weak.[REF-08] [REF-12] [REF-23]

    Part XXI

    Institutional improvement plans aligned with Education 2030

    75

    Institutional mandate and scope

    For institutional mandate and scope, the material distinction is between a school, training provider, university, local administration or other body does not implement the whole global agenda alone. The institutional consequence follows from whether it contributes within its mandate, while national authorities retain responsibilities for law, finance, curriculum, workforce and equitable system provision. The plan should therefore map each selected Education 2030 priority to the institutional function that can materially influence it and identify dependencies requiring action at another level. An institutional improvement plan should begin by identifying the authority under which the institution acts and the population for whom it is responsible.[REF-12] [REF-25] [REF-28]

    The central question in institutional mandate and scope is the statement should name the eligible population, current condition, intended change and period. The public account remains incomplete unless it explains how a global target supplies direction, but the institutional proposition must be narrow enough for authority, resources and evidence to be assigned. Scope should be expressed as a learner-facing condition rather than a broad aspiration. An institution may seek more regular participation, stronger foundational learning, safer and more accessible facilities, better transition, improved teacher support or a more equitable adult-learning offer.[REF-07] [REF-27] [REF-28]

    Comparative interpretation of institutional mandate and scope depends upon a plan concerned with learning cannot ignore assessment participation or opportunity to learn; a plan concerned with enrolment cannot treat registration as regular attendance; a plan concerned with completion should examine the educational substance and recognised value of the programme. A contrary reading would overlook that selecting one priority does not authorize deterioration in another essential condition. Review of institutional mandate and scope is credible only where it explains safeguards should state which minimums must be protected during implementation. For the learners concerned, the decisive consideration is whether institutional plans should preserve the relationship among access, equity and quality.[REF-01] [REF-03] [REF-09]

    Review of institutional mandate and scope is credible only where it explains any requested exception should identify the legal basis, expected consequence and competent approving body. A contrary reading would overlook that the plan should distinguish obligations from discretionary methods. Rights, national requirements and adopted policy establish conditions that the institution must respect. The choice of timetable, professional-learning arrangement, local partnership or diagnostic method may permit adaptation. This distinction allows institutions to respond to context without treating national variation or resource constraint as authority to lower the substantive entitlement of disadvantaged learners.[REF-08] [REF-10] [REF-28]

    institutional mandate and scope cannot be judged without identifying the decision record should show how evidence and consultation altered the priority. A contrary reading would overlook that where the institution lacks authority over a material dependency, the plan should route the issue to the responsible authority and retain it as an unresolved risk rather than quietly narrow the objective. Governance should identify one accountable owner and the roles of staff, learners, communities and partner bodies. Participation can reveal barriers and test whether proposed measures are workable; it should not transfer public responsibility to those affected by weak provision.[REF-17] [REF-19] [REF-25]

    76

    Diagnostic review and priority selection

    diagnostic review and priority selection requires a decision about entry, attendance, progression, completion, learning and transition should be mapped together with staffing, instructional time, facilities, accessibility and learner support. This matters because the purpose is to locate the first material break rather than compile every available statistic. Administrative records describe registered learners and delivery; household or community evidence may reveal people outside the institution; assessment and work samples describe bounded aspects of learning. Diagnostic review should begin with the institution's eligible population and educational course.[REF-02] [REF-03] [REF-24]

    Institutional action on diagnostic review and priority selection should be tested against small groups require caution about precision and disclosure, but the service barrier should remain visible. A contrary reading would overlook that where population estimates are weak, qualitative and case evidence may justify immediate correction without being represented as a prevalence estimate. The review should compare level, distribution and trend. A stable institutional mean can conceal deterioration among one programme or learner group. Sex, household constraint, disability, language, residence, displacement and prior opportunity may be relevant, subject to lawful collection and protection.[REF-06] [REF-10] [REF-13]

    Institutional action on diagnostic review and priority selection should be tested against low attendance may be associated with transport, cost, safety, calendar, health, discrimination or inaccessible instruction. The institutional consequence follows from whether weak learning may reflect interrupted participation, limited curriculum exposure, teacher absence, language, assessment design or inadequate support. A descriptive difference cannot rank these explanations. The institution should gather the smallest additional evidence needed for a practical decision and should identify explanations requiring action outside its authority. Diagnosis should distinguish symptom from mechanism.[REF-01] [REF-05] [REF-12]

    The practical standard for diagnostic review and priority selection concerns ease of measurement alone is not a sufficient selection criterion. A proportionate conclusion must also recognise that priority selection should consider severity, scale, equity, feasibility and dependency. A condition affecting fewer learners may warrant immediate action if the consequence is severe or violates a minimum entitlement. Conversely, a large but weakly measured difference may first require stronger evidence. The plan should explain why the selected priority takes precedence, which learners are expected to benefit and how the rest of the institutional responsibility remains monitored.[REF-07] [REF-16] [REF-28]

    A defensible account of diagnostic review and priority selection distinguishes where the institution changes its record system or assessment during implementation, overlap evidence should be retained and breaks marked. A proportionate conclusion must also recognise that a baseline reconstructed after results are known is vulnerable to selective interpretation. The plan should approve the baseline and revision rule before judging change. The baseline should be preserved at the point of decision. Population, definitions, source completeness, date, uncertainty and any exclusions should travel with the starting value.[REF-03] [REF-23] [REF-24]

    77

    Improvement proposition and implementation design

    Review of improvement proposition and implementation design is credible only where it explains it should connect an action to an institutional mechanism: additional instructional support to a documented curriculum gap; revised attendance arrangements to a known timing or transport barrier; accessible materials and assessment to exclusion of learners with disabilities; professional support to weak subject teaching. A proportionate conclusion must also recognise that the central question in improvement proposition and implementation design is activities should not be listed without the reason they are expected to alter learner opportunity. A proportionate conclusion must also recognise that the improvement proposition should state how a defined measure is expected to change the diagnosed condition.[REF-10] [REF-12] [REF-28]

    A defensible account of improvement proposition and implementation design distinguishes new materials require procurement, accessible formats, distribution and maintenance. The public account remains incomplete unless it explains how learner support may require eligibility rules, communication and protection of personal information. A plan should identify the latest point at which a missing dependency can be corrected without compromising delivery. Where a dependency lies with another authority, agreement and escalation should precede large-scale implementation. Dependencies should be sequenced. Teacher guidance may require curriculum clarification, planning time, examples and leadership support.[REF-17] [REF-18] [REF-25]

    In assessing improvement proposition and implementation design, authorities must determine attendance at one professional meeting, receipt of materials or registration in support does not prove sustained use. The public account remains incomplete unless it explains how monitoring should capture actual exposure and reasons for incomplete delivery while avoiding burdens so heavy that they displace teaching or service. Implementation intensity should be specified. An institution should know which learners or staff receive the measure, how frequently, for how long and with what expected standard. Without intensity, non-response cannot be distinguished from weak delivery.[REF-03] [REF-06] [REF-16]

    In assessing improvement proposition and implementation design, authorities must determine the plan should test which learners can take up the measure and provide the additional condition required for substantive access. A contrary reading would overlook that local adaptation is desirable where it removes a barrier while protecting the common educational objective. It becomes inequitable where learners facing disadvantage receive reduced content, weaker staff or an unrecognised route. Equity should be built into eligibility, outreach and adaptation. An apparently equal offer may be unusable because of cost, time, language, disability, safety or location.[REF-08] [REF-13] [REF-14]

    The practical standard for improvement proposition and implementation design concerns additional assessment can narrow curriculum or increase exclusion; targeted grouping can stigmatise learners; extended time can increase household costs; intensive support for one cohort can divert staff from another. The resulting interpretation should show why the plan should identify indicators or qualitative evidence that would reveal these effects and the authority able to intervene. Improvement is not established by movement of the selected measure if another essential condition deteriorates. Foreseeable adverse effects should be recorded.[REF-01] [REF-11] [REF-23]

    78

    Resources and professional capability

    Public responsibility for resources and professional capability begins with the institution should identify which costs fall within its budget, which require higher-level allocation and which are presently unfunded. The public account remains incomplete unless it explains how an unfunded dependency should remain visible in the risk statement. The resource plan should cover the complete recurrent service: personnel, preparation, instructional time, accessible materials, facilities, learner support, assessment, coordination, evidence and review. Capital expenditure or a temporary grant may be necessary but cannot establish sustained capability by itself.[REF-15] [REF-26] [REF-28]

    In assessing resources and professional capability, authorities must determine additional finance should be justified by the common educational condition it protects, not by a permanent assumption of lower capability among the affected learners. The institutional consequence follows from whether distributional finance matters within institutions as well as across systems. Programmes serving learners with greater need may require more staff time, specialist support, transport or accessible materials. Equal departmental or per-learner amounts can reproduce unequal opportunity. Allocation reasons should be documented and reviewed against actual receipt and use.[REF-06] [REF-10] [REF-13]

    Evidence concerning resources and professional capability should establish teachers, trainers and institutional leaders need subject knowledge, diagnostic skill, practical examples and time to collaborate. The resulting interpretation should show why where the plan alters curriculum, assessment or inclusion arrangements, staff should have supported opportunities to understand the educational reasoning. One briefing does not establish readiness for sustained change. The plan should monitor participation, usable learning, workload, turnover and access to later assistance. Professional capability is a central implementation condition.[REF-01] [REF-18] [REF-25]

    In assessing resources and professional capability, authorities must determine conversely, higher-level constraints should not prevent local correction that is lawful and feasible. For the learners concerned, the decisive consideration is whether the plan should identify decision thresholds, escalation routes and the time within which the responsible body will respond. Leadership should create a route from classroom or service evidence to institutional and system correction. Staff should not be held responsible for changing transport, staffing establishment, qualification rules or infrastructure beyond their authority.[REF-12] [REF-17] [REF-19]

    For resources and professional capability, the material distinction is between agreements should state the educational purpose, authority, resource duration, records, safeguards and handover. The institutional consequence follows from whether the measure of partnership is whether the institution can sustain equitable quality and correction after exceptional support ends. External partnership should strengthen ordinary capability. Universities, civil society, employers or international bodies may contribute knowledge, facilities or finance, but programme status, data protection and continuity should be clear. Parallel arrangements can create fragmented standards or reporting burdens.[REF-22] [REF-26] [REF-28]

    79

    Evidence, adaptation and institutionalisation

    Institutional action on evidence, adaptation and institutionalisation should be tested against approval of the plan establishes authority; expenditure demonstrates a resource transaction; implementation records show activity; learner evidence shows participation or change. The public account remains incomplete unless it explains how none automatically proves the next. The review schedule should collect evidence proportionate to each claim and identify the population represented. A favourable result should not be broadened beyond the learners, content and period observed. Review should distinguish authorization, delivery, reach, educational quality and consequence.[REF-03] [REF-16] [REF-24]

    In assessing evidence, adaptation and institutionalisation, authorities must determine the review should disaggregate where lawful and precise and should retain absolute numbers. The public account remains incomplete unless it explains how an improved average does not establish that learners farthest from the baseline benefited. The plan should define a minimum floor or least-served test alongside its aggregate objective. Reach should be examined from eligibility through offer, take-up, participation and intended intensity. Non-participation can reveal communication, cost, timing, safety or accessibility barriers.[REF-01] [REF-06] [REF-13]

    In assessing evidence, adaptation and institutionalisation, authorities must determine contradiction should initiate enquiry into coverage, implementation or measurement rather than selection of the most favourable source. A proportionate conclusion must also recognise that educational consequence may require several forms of evidence. Assessment can show a bounded learning domain; attendance and completion records show participation; observation and work samples can explain teaching and learner activity; interviews can identify experience and mechanism. Agreement strengthens the conclusion only after dates and definitions align.[REF-02] [REF-12] [REF-23]

    Institutional action on evidence, adaptation and institutionalisation should be tested against ending an ineffective measure is responsible improvement when learner safeguards and alternative provision are addressed. The public account remains incomplete unless it explains how adaptation should follow a stated decision rule. Weak outcomes may reflect an incorrect hypothesis, incomplete delivery, insufficient intensity, adverse conditions, measurement weakness or inadequate time. These explanations call for different action. The review should record which is supported, what changes and how the baseline or comparison remains interpretable.[REF-08] [REF-11] [REF-28]

    In assessing evidence, adaptation and institutionalisation, authorities must determine completion means that the institution can sustain the improved learner-facing condition and correct foreseeable departure, not that the plan period has ended. The resulting interpretation should show why institutionalisation requires ordinary authority, recurrent finance, professional ownership and a functioning correction route. Continued activity after a pilot is not sufficient if it depends on exceptional staff or external funds. The final decision should state which function enters ordinary provision, which limitation remains, who owns it and when evidence will next be examined.[REF-17] [REF-25] [REF-26]

    Part XII

    Aid, crisis and shared international responsibility

    Part I

    Emergency education after Istanbul

    1

    Institutional status and continuity of responsibility

    Comparative interpretation of institutional status and continuity of responsibility depends upon existing responsibilities under human-rights, refugee and humanitarian law remain applicable according to their terms. The evidence must therefore clarify how national authorities retain the primary education role; international support should strengthen rather than obscure that responsibility. The Summit was a high-level global meeting convened by the United Nations Secretary-General. Its political commitments should be reported accurately without being treated as new binding law.[REF-29] [REF-30]

    Institutional action on institutional status and continuity of responsibility should be tested against public reporting should distinguish a Summit-wide record, a government pledge, an agency initiative and a financed agreement. This matters because the Chair's Summary is the Chair's account of the Summit. It can establish themes and announcements, but not universal consent to every statement made by participants. Individual commitments require their own issuing authority and evidence.

    2

    Population, movement and denominators

    population, movement and denominators cannot be judged without identifying each observes a different population. A proportionate conclusion must also recognise that duplicate and missing records should be expected where families move. The estimate should state reference date, place, status definition and revision. An emergency population changes over time and can be counted through registration, border, shelter, household, community and administrative sources.

    A defensible account of population, movement and denominators distinguishes a denominator restricted to applicants or camp residents cannot establish participation for an entire displaced population. This matters because host-community population change also matters where schools absorb rapid enrolment. Children outside registration remain within educational concern.

    3

    Safe and usable access

    The central question in safe and usable access is distance, transport, documentation, language, household cost, disability, gender-related risk and work or care duties can prevent attendance. For the learners concerned, the decisive consideration is whether the response should identify the barrier rather than classify non-attendance as lack of demand. A place is usable only where the learner can reach it safely and participate with necessary accommodation.

    safe and usable access requires a decision about reports should distinguish verified incident, reported concern and general exposure. The evidence must therefore clarify how protection information should be collected only for a clear purpose and handled under suitable safeguards. School safety includes the site, route, building, water and sanitation, safeguarding and conduct of personnel.

    4

    Teaching and learning continuity

    For teaching and learning continuity, the material distinction is between a receiving system may need bridging or language support while retaining a route to recognised national education. The evidence must therefore clarify how parallel curricula and records can impede later movement if compatibility is not planned. Curriculum continuity requires an account of prior learning, present language, intended progression and available time.

    Public responsibility for teaching and learning continuity begins with observation should focus on a bounded practice and lead to support. A contrary reading would overlook that one crisis-affected lesson cannot establish general teacher performance. Teachers may require concise guidance, joint planning, materials and coaching for mixed-age, multilingual or overcrowded classes.

    5

    Assessment, records and recognition

    Review of assessment, records and recognition is credible only where it explains missing certificates should not create an absolute bar to entry. For the learners concerned, the decisive consideration is whether provisional placement can draw on learner work, dialogue, prior history and observed performance, with a review date. Diagnostic evidence should guide placement and teaching without becoming an unannounced exclusion test.

    assessment, records and recognition cannot be judged without identifying reconstructed entries require provenance. A contrary reading would overlook that high-consequence certification needs stronger evidence than provisional entry, but institutional disruption should not be transferred entirely to the learner. Records should identify learner, programme, dates, curriculum, attendance, assessed learning and decision status.

    6

    Financing and the new fund

    financing and the new fund requires a decision about school-year recruitment, teacher payment and procurement have deadlines. A proportionate conclusion must also recognise that late disbursement can be recorded as financial execution while failing the intended educational period. Multi-year crises require a recurrent horizon. Emergency finance should align time with educational need.

    A defensible account of financing and the new fund distinguishes the Education Cannot Wait launch supplied a stated institutional ambition, not evidence of disbursement or effect. For the learners concerned, the decisive consideration is whether future monitoring should identify contributors, commitments, cash received, allocation criteria, implementing authority, populations reached, service quality and learning or continuity evidence. The conclusion for global education quality annual report 2017 should connect the finding to the competent authority, affected learners and evidence required at review.[REF-31]

    7

    Accountability and exit from temporary arrangements

    Institutional action on accountability and exit from temporary arrangements should be tested against temporary sites, condensed curricula, volunteer teaching or parallel records should not become a permanent lower standard for affected populations. The public account remains incomplete unless it explains how exit can mean integration, recognised transfer, return or closure with responsibility handed to an authorised body. Every emergency measure needs an owner, review date and transition condition.

    Review of accountability and exit from temporary arrangements is credible only where it explains a project ending does not establish educational recovery. The public account remains incomplete unless it explains how completion is a sustained, recognised route forward. Public reports should state unresolved populations and limitations. A favourable attendance rate among reached learners does not erase those never offered a place.

    Part XXII

    Comparative financing evidence for protracted crises

    80

    Estimating the complete educational requirement

    Evidence concerning estimating the complete educational requirement should establish reports should retain the source and uncertainty and show host-community learners where additional demand changes their educational conditions. The resulting interpretation should show why an education-financing requirement should begin with the affected population and the service standard. Population estimates need a reference date, geography, age or programme group, displacement or other status definition and coverage statement. In a protracted crisis, people move, registrations lapse and multiple agencies may count the same household. A point estimate without revision history can be misleading.[REF-20] [REF-21] [REF-23]

    Comparative interpretation of estimating the complete educational requirement depends upon not every intervention needs every item in the same form, but omissions should be reasoned. The public account remains incomplete unless it explains how a space-and-supplies estimate can support immediate reopening; it cannot represent the complete cost of sustained quality education if salaries, support and progression are absent. The service specification should separate safe access, instructional space, teachers, learning time, curriculum, materials, language support, disability accommodation, water and sanitation, protection referral, assessment, records and recognised transition.[REF-09] [REF-12] [REF-32]

    estimating the complete educational requirement requires a decision about start-up and recurrent cost should be separated, as should national currency and foreign-exchange assumptions where material. This matters because shared costs such as administration, assessment and data should be allocated transparently rather than omitted or distributed by an unexplained percentage. Unit costs should state the unit, duration, price basis, population and service content. A per-learner annual requirement can support comparison only when programme duration and included items are similar.[REF-03] [REF-16] [REF-26]

    Comparative interpretation of estimating the complete educational requirement depends upon it should not become an undisclosed addition. A proportionate conclusion must also recognise that the estimate should identify the risk addressed, calculation and conditions for release. Scenario ranges can be more informative than one exact amount when population and access may change. Each scenario should retain a common minimum educational condition so that the lower figure does not silently imply a lower standard for learners facing the most severe crisis. Contingency is legitimate where insecurity, access and movement make delivery uncertain.[REF-05] [REF-14] [REF-29]

    In assessing estimating the complete educational requirement, authorities must determine a rising requirement is not evidence of financial mismanagement if the population grew or the initial service specification was incomplete; a falling requirement is not necessarily efficiency if access contracted. The resulting interpretation should show why the public account should distinguish changed need from changed measurement and changed ambition. Needs should be revised when population, prices, programme design or delivery access changes. Revision should preserve the preceding estimate, reason and effect.[REF-08] [REF-23] [REF-24]

    81

    Defining and interpreting the financing gap

    defining and interpreting the financing gap cannot be judged without identifying earmarked finance cannot automatically fill an unrelated requirement; a multi-year commitment should not be compared with a one-year need without stating the temporal conversion. A contrary reading would overlook that the financing gap is the difference between a defensible requirement and finance available for the same population, service and period. It is not the difference between an appeal headline and any amount mentioned in its vicinity. The numerator and denominator of the comparison must be conceptually aligned.[REF-03] [REF-26] [REF-29]

    Review of defining and interpreting the financing gap is credible only where it explains these sources carry different authority, timing and conditions. A contrary reading would overlook that household expenditure should not be treated as desirable gap closure where cost excludes poor learners. Double counting can arise when one contribution passes through several agencies or when both commitment and transfer are summed. A funding registry should preserve the origin, intermediary, recipient and final purpose. Available finance should distinguish domestic public allocation, humanitarian contribution, development assistance, private or philanthropic contribution and household payment where relevant.[REF-15] [REF-17] [REF-31]

    The central question in defining and interpreting the financing gap is distributional reporting should show which learners and institutions bear the shortfall. This matters because it should avoid inferring that a population with no costed plan has no gap; absence of an estimate is an evidence deficiency, not evidence of adequate finance. The gap should be reported by service component and geography where evidence permits. An aggregate can conceal a fully financed capital item beside an unfunded teacher requirement or an accessible urban programme beside an underserved remote population.[REF-01] [REF-06] [REF-13]

    Comparative interpretation of defining and interpreting the financing gap depends upon if both requirement and available finance are ranges or subject to material revision, the resulting gap should not carry greater precision. The resulting interpretation should show why reports can present a central estimate and bounded alternatives, identify the dominant uncertainty and test whether the conclusion changes. A gap clearly material under every plausible specification may justify urgent action even if its exact magnitude is unsettled. Uncertainty belongs in the gap.[REF-23] [REF-24]

    In assessing defining and interpreting the financing gap, authorities must determine the strongest comparative use is to reveal which components are systematically omitted and which financing arrangements support continuity under comparable conditions. The evidence must therefore clarify how gap comparisons between crises require caution. Cost, programme structure, public capacity, access, population composition and the role of existing national provision differ. A per-learner gap can support orientation when the included service is disclosed; it should not rank need or performance without context.[REF-16] [REF-19] [REF-30]

    82

    Predictability and the time profile of finance

    Review of predictability and the time profile of finance is credible only where it explains a pledge with no schedule is less usable for planning than an enforceable appropriation or signed agreement with known timing, even if its nominal amount is larger. A proportionate conclusion must also recognise that the baseline should therefore report amount, status, period, conditions, expected transfer dates, currency and recipient. It should distinguish an indicative multi-year intention from finance legally or administratively available. Predictability concerns whether the competent institution can plan, contract, staff and sustain a defined educational service.[REF-26] [REF-29] [REF-31]

    Review of predictability and the time profile of finance is credible only where it explains this makes a foreseeable interruption visible early enough for correction. The institutional consequence follows from whether education carries recurrent obligations. Teacher contracts, rent, transport, language support, materials, assessment and records require continuity over school years. Short grants can create interruptions or repeated procurement and recruitment. A financing profile should show the share of the service covered by month, term or year and identify the date at which essential functions become unfunded.[REF-12] [REF-15]

    Comparative interpretation of predictability and the time profile of finance depends upon delay reasons should be classified without assuming that every control is unnecessary; fiduciary assurance and speed must both protect learner interest. The institutional consequence follows from whether timing should be assessed against the education calendar and operational lead times. Funds arriving after enrolment, teacher recruitment or procurement deadlines may be counted as received while failing to protect continuity. Disbursement timeliness should therefore be measured from the decision point relevant to the service, not only from the donor's commitment date.[REF-17] [REF-18] [REF-29]

    predictability and the time profile of finance cannot be judged without identifying reports should identify transfers between cost categories, reasons and learner consequence. The evidence must therefore clarify how the test is whether adaptation maintained the intended educational condition under changed circumstances. Flexibility is also material. A tightly earmarked grant can leave a critical support or host-community need unfunded even where total resources appear adequate. Flexibility should operate within a public plan, authority and safeguards rather than permit unrecorded substitution.[REF-08] [REF-30]

    A defensible account of predictability and the time profile of finance distinguishes a programme that remains open through repeated emergency extensions without a stated continuity arrangement is not predictable merely because support has historically been renewed. The resulting interpretation should show why exit and handover are part of predictability. Temporary international finance should identify whether the service will transfer to national or local budgets, continue under renewed assistance, adapt or close. The plan should state responsibilities, recurrent cost and evidence required before the transition.[REF-25] [REF-26] [REF-28]

    83

    From pledge to learner-facing use

    from pledge to learner-facing use cannot be judged without identifying each state requires its own evidence and date. The institutional consequence follows from whether movement should not be inferred. A press release can establish an announcement; a signed agreement can establish a commitment; a bank or accounting record can establish transfer or expenditure; only service evidence can establish that learners received the intended condition. Financial reporting should use a state model: announced amount, pledge, formal commitment, transfer, recipient receipt, obligation, expenditure, delivered input, usable service and educational consequence.[REF-29] [REF-30] [REF-31]

    Institutional action on from pledge to learner-facing use should be tested against commitments made in one currency and spent in another should state the conversion date and method. This matters because in-kind contributions should be valued separately and should identify the item, quantity, condition and service use. Counting the procurement value of unusable or delayed material as full learner benefit would overstate delivery. Where valuation is uncertain, physical quantities and distribution may be more informative than a precise monetary equivalent. Currency and valuation can alter the chain.[REF-03] [REF-23]

    from pledge to learner-facing use cannot be judged without identifying funds may pass from donor to pooled mechanism, agency, national authority and provider. The public account remains incomplete unless it explains how the aggregate should count the contribution once while reporting administrative and programme expenditure at the relevant stages. Subgrants should retain purpose, period and recipient. Confidential or security-sensitive information may require controlled access, but public totals should still describe coverage and method. Implementing partners should be linked without double counting.[REF-10] [REF-24]

    A defensible account of from pledge to learner-facing use distinguishes these distinctions allow finance to be connected to educational continuity without claiming that expenditure alone produced learning. The institutional consequence follows from whether learner-facing use should be observed through staffing, operational time, facilities, materials, accommodation and support actually available. Procurement completion is not the same as arrival; arrival is not the same as accessible use. Teacher positions should be distinguished from filled and attended posts. A temporary classroom should be distinguished from safe instructional time.[REF-02] [REF-32]

    Evidence concerning from pledge to learner-facing use should establish descriptive improvement after expenditure can support operational judgement but does not by timing alone establish attribution. A contrary reading would overlook that causal claims require a credible comparison; bounded service evidence remains sufficient to test whether the financed input reached the intended learners. Educational consequence belongs later in the chain and needs suitable evidence. Attendance, progression, learning and recognised transition are influenced by finance and by implementation, insecurity, population movement and prior opportunity.[REF-01] [REF-16]

    84

    Equity, host systems and distributional incidence

    In assessing equity, host systems and distributional incidence, authorities must determine a separate programme may address urgent barriers while creating parallel standards or weak recognition. The resulting interpretation should show why the distributional account should show how finance affects the common public service and any differentiated support required to secure an equal substantive opportunity. Financing should be assessed across displaced, crisis-affected and host populations. Additional enrolment can reduce class time, materials or teacher availability for existing learners if host systems receive no compensating resource.[REF-06] [REF-12] [REF-20]

    The practical standard for equity, host systems and distributional incidence concerns disability accommodation, minority-language instruction, remote access, protection, interrupted prior learning and sparse populations may require more resources. For the learners concerned, the decisive consideration is whether equal per-learner finance can reproduce inequality where these costs differ. The factors, weights and evidence should be transparent and tested against actual distribution. Additional allocation should protect the common entitlement rather than institutionalise lower expectations for the affected population. Allocation formulae should recognise relevant additional cost.[REF-09] [REF-10] [REF-13]

    Evidence concerning equity, host systems and distributional incidence should establish assistance should be reported by eligibility, offer, receipt and use rather than announced coverage alone. A contrary reading would overlook that household incidence is material. Crisis-affected families may face transport, materials, examination, documentation and opportunity costs even where tuition is formally free. A funding plan that transfers cost to households can appear balanced while participation contracts among the poorest. Household evidence should be read alongside service and price information, subject to sampling and residence limitations.[REF-14] [REF-15] [REF-24]

    equity, host systems and distributional incidence requires a decision about adolescent girls may face safety or care barriers; boys may enter hazardous work; older learners may lack recognised routes; young children may be absent from emergency registries. A proportionate conclusion must also recognise that funding should follow diagnosed barriers without treating identity as the cause. The review should test reach and learner experience and should protect sensitive information. Gender and age patterns can differ across displacement and host settings.[REF-08] [REF-21]

    Public responsibility for equity, host systems and distributional incidence begins with where detailed financial allocation cannot be linked directly to learners, institutional and geographic distribution can provide a bounded intermediate measure. A contrary reading would overlook that the report should state the allocation assumption and avoid implying individual benefit from an aggregate transfer. Incidence should be measured at the level where the service is experienced. National or response-wide totals can conceal unfunded districts, schools or programme levels.[REF-03] [REF-16] [REF-19]

    85

    Efficiency, quality and the risk of false economy

    Evidence concerning efficiency, quality and the risk of false economy should establish comparative analysis should state the service content, population and outcome before interpreting cost differences. This matters because efficiency should mean securing the greatest defensible educational value from resources while protecting rights, quality and equity. A lower cost per learner is not efficient if instructional time, teacher preparation, safety or recognition is reduced below an acceptable condition. Likewise, high expenditure is not proof of quality.[REF-09] [REF-12] [REF-16]

    Comparative interpretation of efficiency, quality and the risk of false economy depends upon costs should be assigned by function and examined for whether they enable safe, accountable delivery. A contrary reading would overlook that overhead and coordination costs require substantive review. Excessive fragmentation can multiply management and reporting, while inadequate coordination can produce duplication, geographic gaps or incompatible records. A simple low-overhead ratio can also be misleading if necessary protection, monitoring or technical support is classified as administration.[REF-29] [REF-30]

    Evidence concerning efficiency, quality and the risk of false economy should establish local procurement may shorten delivery or support continuity but requires fair standards and capacity. A proportionate conclusion must also recognise that the comparison should use lifecycle and service costs where these materially affect the educational condition. Procurement decisions should consider suitability, maintenance, transport and usable life. The cheapest unit can be false economy if it is inaccessible, unsafe, incompatible or impossible to maintain.[REF-03] [REF-22]

    Evidence concerning efficiency, quality and the risk of false economy should establish stable, prepared personnel are central to educational continuity. This matters because short-term or unpaid arrangements can reduce the recorded budget while increasing turnover, absence and weak teaching. Compensation, preparation, supervision, safety and workload should be included. Where emergency personnel enter through accelerated routes, a financed path to recognised preparation and continuing support should be established. Teacher cost should not be treated as a residual.[REF-01] [REF-18] [REF-32]

    Review of efficiency, quality and the risk of false economy is credible only where it explains the responsible body should state whether a cost difference reflects price, population, service design, delay, waste or genuine improvement. The institutional consequence follows from whether savings should be traced to their effect on the complete requirement and should not be counted twice as both lower expenditure and an unfunded gap. A revision should preserve adverse evidence and the reason for changing allocation or delivery. Efficiency findings should lead to correction, not only comparison.[REF-08] [REF-23]

    86

    Minimum public account as at launch

    minimum public account as at launch requires a decision about it does not comprise later receipts, disbursement, geographic reach or education results. A contrary reading would overlook that those belong to later observation periods. A new institution should not be judged as though it has completed delivery, but neither should future performance be presumed from its launch. At the launch stage of Education Cannot Wait, the minimum defensible record comprises the institutional announcement, stated purpose, governance information officially available and any contemporaneous commitments that can be evidenced.[REF-31]

    Review of minimum public account as at launch is credible only where it explains amounts should use aligned populations, periods and service definitions. The public account remains incomplete unless it explains how pledges, commitments, receipts and expenditure should be reported separately. In-kind contributions and currency conversion should be explained where material. For each crisis-financing account, the public record should state affected population, educational requirement, cost method, available finance by source and status, gap, timing, distribution, implementing authority and revision date.[REF-03] [REF-23] [REF-26]

    minimum public account as at launch requires a decision about provisional counts can support allocation when dated and revised; they should not later be presented as final statistics without review. A proportionate conclusion must also recognise that the record should name the least-observed population and the most consequential unfunded service. Missing evidence should have an owner, financed improvement action and expected date. Uncertainty should qualify language but should not suspend urgent protection where credible evidence shows serious exclusion or interruption.[REF-20] [REF-21] [REF-24]

    A defensible account of minimum public account as at launch distinguishes a donor may revise a commitment, a pooled mechanism may allocate, a government may authorize teachers, and a provider may deliver instruction. The institutional consequence follows from whether coordination is necessary, but remedy should not disappear between levels. The report should connect each material gap or delay to the body with authority and state the next decision date. Accountability should identify who can change finance and who can change delivery.[REF-17] [REF-19] [REF-30]

    As at 10 December 2017, the Summit, fund launch and New York Declaration are contemporaneous institutional facts, while any result still requires its own dated evidence. The defensible conclusion is therefore a financing standard: assess the complete need, preserve every funding state, test predictability and distribution, trace learner-facing use and delay all result claims until suitable later evidence exists.[REF-29] [REF-30] [REF-31]

    References

    1. REF-01

      Education for All Global Monitoring Report Team. Reaching the Marginalized — EFA Global Monitoring Report 2010. 2010.

      Principal contemporaneous analysis of intersecting disadvantage and education marginalisation.

      https://unesdoc.unesco.org/ark:/48223/pf0000186606
    2. REF-02

      UNESCO Institute for Statistics. Global Education Digest 2010: Comparing Education Statistics Across the World. 2010.

      Comparative education statistics, definitions and limitations.

      https://uis.unesco.org/sites/default/files/documents/global-education-digest-2010-comparing-education-statistics-across-the-world-en.pdf
    3. REF-03

      UNESCO Institute for Statistics. Education Indicators: Technical Guidelines. 2009.

      Definitions and interpretation of participation, progression, completion and resource indicators.

      https://uis.unesco.org/sites/default/files/documents/education-indicators-technical-guidelines-en_0.pdf
    4. REF-04

      United Nations. The Millennium Development Goals Report 2010. 2010.

      Global and regional monitoring of primary education, gender, poverty and related development conditions.

      https://www.un.org/millenniumgoals/pdf/MDG%20Report%202010%20En%20r15%20-low%20res%2020100615%20-.pdf
    5. REF-05

      United Nations Development Programme. Human Development Report 2010: The Real Wealth of Nations — Pathways to Human Development. 2010.

      Distribution-sensitive human development concepts and evidence available before the cut-off.

      https://hdr.undp.org/content/human-development-report-2010
    6. REF-06

      UNICEF. Progress for Children: Achieving the MDGs with Equity, Number 9. 2010.

      Equity-focused child indicators and comparison between population groups.

      https://www.unicef.org/reports/progress-children-no-9
    7. REF-07

      World Education Forum. The Dakar Framework for Action: Education for All — Meeting Our Collective Commitments. 2000.

      Commitments to equitable access, quality, measurable outcomes and accountable national planning.

      https://unesdoc.unesco.org/ark:/48223/pf0000121147
    8. REF-08

      United Nations General Assembly. Convention on the Rights of the Child. 1989.

      Rights concerning non-discrimination, identity, education and development.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-child
    9. REF-09

      United Nations Committee on Economic, Social and Cultural Rights. General Comment No. 13: The Right to Education. 1999.

      Interpretation of availability, accessibility, acceptability and adaptability.

      https://www.refworld.org/legal/general/cescr/1999/en/37937
    10. REF-10

      United Nations General Assembly. Convention on the Rights of Persons with Disabilities. 2006.

      Non-discrimination, accessibility, inclusive education and disability data safeguards.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-persons-disabilities
    11. REF-11

      United Nations. Guiding Principles on Internal Displacement. 1998.

      Principles relevant to protection, documentation, education and non-discrimination of displaced persons.

      https://www.ohchr.org/en/special-procedures/sr-internally-displaced-persons/international-standards
    12. REF-12

      UNESCO and UNICEF. A Human Rights-Based Approach to Education for All. 2007.

      Rights-based planning, equality, participation, accountability and education quality.

      https://unesdoc.unesco.org/ark:/48223/pf0000154861
    13. REF-13

      Education for All Global Monitoring Report Team. Overcoming Inequality: Why Governance Matters — EFA Global Monitoring Report 2009. 2008.

      Governance, finance and unequal educational opportunity.

      https://unesdoc.unesco.org/ark:/48223/pf0000177683
    14. REF-14

      World Bank. World Development Report 2006: Equity and Development. 2005.

      Concepts of unequal opportunity, institutions and equitable public action.

      https://documents.worldbank.org/curated/en/435331468127174418/pdf/322040World0Development0Report02006.pdf
    15. REF-15

      World Bank. Safeguarding Education During Economic Crisis. 2009.

      Risks to budgets, households, participation and long-term human development during economic crisis.

      https://documents1.worldbank.org/curated/en/489131468340200911/pdf/485120WP0Avert10Box338912B01PUBLIC1.pdf
    16. REF-16

      Organisation for Economic Co-operation and Development. Education at a Glance 2010: OECD Indicators. 2010.

      Comparative participation, progression, expenditure and outcomes evidence with system-level metadata.

      https://doi.org/10.1787/eag-2010-en
    17. REF-17

      European Commission. Europe 2020: A Strategy for Smart, Sustainable and Inclusive Growth. 2010.

      Contemporaneous European policy context for education, inclusion, employment and headline indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC2020
    18. REF-18

      European Commission. Youth on the Move: An Initiative to Unleash the Potential of Young People to Achieve Smart, Sustainable and Inclusive Growth in the European Union. 2010.

      European education, mobility, attainment and youth inclusion policy context.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0477
    19. REF-19

      European Commission. A Renewed Commitment to Social Europe: Reinforcing the Open Method of Coordination for Social Protection and Social Inclusion. 2008.

      Social inclusion monitoring, common objectives and context-sensitive indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52008DC0418
    20. REF-20

      United Nations General Assembly. Resolution 64/250: Assistance to Haiti in the Aftermath of the Recent Earthquake. 2010.

      Contemporaneous recognition of humanitarian and reconstruction needs and national leadership.

      https://undocs.org/A/RES/64/250
    21. REF-21

      United Nations Office for the Coordination of Humanitarian Affairs. Haiti Revised Humanitarian Appeal. 2010.

      Displacement, service disruption and humanitarian education context.

      https://reliefweb.int/report/haiti/haiti-revised-humanitarian-appeal-2010
    22. REF-22

      European Commission. European Union Response to the Earthquake in Haiti. 2010.

      European humanitarian and recovery support, coordination and Haitian ownership.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0056
    23. REF-23

      United Nations Economic and Social Council. Principles and Recommendations for Population and Housing Censuses, Revision 2. 2008.

      Official principles for population coverage, definitions, classifications and data quality.

      https://unstats.un.org/unsd/demographic-social/Standards-and-Methods/files/Principles_and_Recommendations/Population-and-Housing-Censuses/Series_M67Rev2-E.pdf
    24. REF-24

      United Nations Statistics Division. Designing Household Survey Samples: Practical Guidelines. 2005.

      Sample design, estimation, precision and non-response guidance.

      https://unstats.un.org/unsd/demographic/sources/surveys/Handbook23June05.pdf
    25. REF-25

      World Education Forum 2015. Incheon Declaration: Education 2030 — Towards Inclusive and Equitable Quality Education and Lifelong Learning for All. 2015.

      Education commitments adopted at Incheon and their contemporaneous institutional status.

      https://unesdoc.unesco.org/ark:/48223/pf0000233137
    26. REF-26

      United Nations General Assembly. Addis Ababa Action Agenda of the Third International Conference on Financing for Development. 2015.

      Adopted global financing framework and relevant principles for domestic public finance and international cooperation.

      https://undocs.org/A/RES/69/313
    27. REF-27

      United Nations General Assembly. Transforming Our World: The 2030 Agenda for Sustainable Development. 2015.

      Agenda adopted on 25 September 2015, including Goal 4 and its education targets, as available at the evidence cut-off.

      https://undocs.org/A/RES/70/1
    28. REF-28

      UNESCO. Education 2030: Incheon Declaration and Framework for Action for the Implementation of Sustainable Development Goal 4. 2015.

      Framework for Action available before the cut-off and relevant to institutional planning, coordination, equity and review.

      https://unesdoc.unesco.org/ark:/48223/pf0000245656
    29. REF-29

      United Nations Secretary-General. One Humanity: Shared Responsibility — Report of the Secretary-General for the World Humanitarian Summit. 2016. A/70/709.

      Secretary-General’s pre-Summit Agenda for Humanity and proposals for humanitarian responsibility.

      https://docs.un.org/A/70/709
    30. REF-30

      Chair of the World Humanitarian Summit. Chair’s Summary: Standing Up for Humanity — Committing to Action. 2016.

      Official Chair’s account of the May 2016 Summit discussions and commitments, used with its status as a summary.

      https://agendaforhumanity.org/sites/default/files/resources/2017/Jul/chairs_summary.pdf
    31. REF-31

      United Nations. New Fund Launched at UN Humanitarian Summit to Address Education in Crisis Zones. 2016.

      Contemporaneous official account of the Education Cannot Wait launch and stated aims; not evidence of later finance or results.

      https://www.un.org/sustainabledevelopment/blog/2016/05/new-fund-launched-at-un-humanitarian-summit-to-address-education-in-crisis-zones/
    32. REF-32

      United Nations General Assembly. The Right to Education in Emergency Situations. 2010. A/RES/64/290.

      Recognition of education as an integral element of humanitarian response and calls for access, continuity, protection and support.

      https://docs.un.org/A/RES/64/290
    33. REF-33

      United Nations General Assembly. New York Declaration for Refugees and Migrants, Resolution 71/1. 2016.

      Contemporaneous commitment relevant to shared responsibility, refugee and migrant education, and the reporting boundary between commitment and result.

      https://undocs.org/A/RES/71/1
    34. REF-34

      Global Education Monitoring Report Team. Accountability in Education: Meeting Our Commitments — Global Education Monitoring Report 2017/8. 2017.

      Principal contemporaneous analysis of accountability relations, responsibility, transparency, participation and unintended effects in education.

      https://unesdoc.unesco.org/ark:/48223/pf0000259338