
Publication record
This is the controlled English edition. Evidence and institutional status are stated as at the evidence cut-off date.
Executive summary
Institutional action on executive summary should be tested against teachers require useful evidence, curriculum guidance, time, materials, manageable groups and specialist support. For the learners concerned, the decisive consideration is whether wider variation in learner attainment should be addressed as an instructional and system-capacity question, not as a reason to lower entitlement or blame learners and teachers. Disruption, uneven access, prior exclusion, language, disability, mobility and different learning histories can widen the range within a class.
A defensible account of executive summary distinguishes teacher observation, learner work, questioning and bounded tasks can supply complementary evidence. A contrary reading would overlook that diagnostic evidence should identify current knowledge and next learning needs. It should be brief enough to preserve teaching time, accessible to the learner and aligned with intended curriculum. A single score should not become a permanent ability label.
In assessing executive summary, authorities must determine teachers need examples of sequencing and adaptation while retaining professional judgement for the learners present. A contrary reading would overlook that curriculum guidance should distinguish essential foundations, later progression and opportunities to revisit learning. Prioritisation should preserve coherence and should not reduce the curriculum indefinitely to a narrow tested subset.
Institutional action on executive summary should be tested against learners should move as evidence changes and should retain access to rich subject learning. A contrary reading would overlook that additional teaching should supplement rather than become an inferior parallel curriculum. Grouping can support targeted teaching when it is flexible, evidence-informed and reviewed. Fixed low groups can limit opportunity and expectations.
executive summary cannot be judged without identifying authorities should identify which tasks can be simplified, which require additional staff and which services lie beyond the teacher's role. This matters because more assessment without time for response can worsen the problem. Teacher workload is a delivery condition. Preparing several pathways, contacting learners, interpreting diagnostic evidence and coordinating support take time.
A defensible account of executive summary distinguishes accessible formats, language support, worked examples, guided practice and extension can help. The resulting interpretation should show why digital materials should not be assumed usable without devices, connectivity and teacher integration. Materials should permit different starting points while aiming toward common educational purposes.
A defensible account of executive summary distinguishes support should strengthen the institution rather than create a separate service with no continuing ownership. The evidence must therefore clarify how additional staff can support smaller groups, tutoring, language learning, disability access and learner contact. Roles, qualifications, supervision and safeguarding should be clear. Temporary personnel and short grants need a handover.
executive summary cannot be judged without identifying attendance at a course does not establish changed teaching. A contrary reading would overlook that collaborative time can help teachers interpret evidence and share approaches, but meetings need a defined educational purpose. Professional learning should be connected to current work and followed by planning, observation or learner evidence.
The practical standard for executive summary concerns this matters because a favourable average should not close the finding where a substantial group remains unreached. For the learners concerned, the decisive consideration is whether the improvement record should state the learner need, intended population, teaching response, enabling conditions, responsible authority, delivery evidence and review date. Review of executive summary is credible only where it explains follow-up should examine reach, learner work, progression and unequal effects.
Institutional action on executive summary should be tested against authorities should examine access, participation, instructional opportunity, learner work, teacher evidence, assessment conditions and populations missing from the available record. For the learners concerned, the decisive consideration is whether learning recovery should begin with evidence of current educational need and the feasibility of an effective response. Disruption did not affect every learner, subject, institution or place equally. A single estimate of average loss cannot determine priority.
A defensible account of executive summary distinguishes a contrary reading would overlook that recovery should not define every difference from a prior cohort as a deficit caused by closure, nor should it ignore pre-existing exclusion. The evidence must therefore clarify how the baseline should distinguish learning already insecure before disruption, learning not taught or practised during disruption, and current capability observed after return or continued remote provision. Review of executive summary is credible only where it explains these categories require different responses.
The practical standard for executive summary concerns a prioritisation rule should be public and should retain the wider curriculum. A contrary reading would overlook that priority should combine educational consequence, number and distribution of learners affected, responsiveness to available action, time sensitivity and feasible capacity. Foundational knowledge may warrant early attention because later learning depends on it, while transition, certification, language support and safeguarding may also be urgent.
For executive summary, the material distinction is between results should identify support needs rather than attach blame to learners, teachers or institutions for system-wide disruption. For the learners concerned, the decisive consideration is whether assessment should be proportionate to the decision. Teacher observation, learner work and bounded diagnostic tasks may guide immediate instruction. High-consequence testing should account for unequal opportunity and changed assessment conditions.
In assessing executive summary, authorities must determine plans should state the minimum delivery conditions and which public body supplies them. A proportionate conclusion must also recognise that feasibility includes teacher time, competence, materials, instructional time, class composition, learner attendance, facilities, finance and institutional authority. An intervention effective under intensive external support may not be feasible at scale.
The central question in executive summary is prioritisation should not become a way to normalise inadequate finance or abandon entitlements. The evidence must therefore clarify how authorities should distinguish temporary sequencing from permanent narrowing and should report unmet need alongside funded action. Finance constrained recovery choices. Education Finance Watch 2021 is within the evidence period and documents pressure on education budgets.
Public responsibility for executive summary begins with extended time can exclude learners with work or care duties; digital support can exclude those without access; accelerated content can disadvantage learners requiring language or disability support. A proportionate conclusion must also recognise that targeted resources may be necessary to secure an equal educational opportunity. Equity review should ask who can use the selected response.
For executive summary, the material distinction is between completion of training, procurement or a timetable change is not recovery. For the learners concerned, the decisive consideration is whether follow-up should examine reach and learning under comparable conditions and should state uncertainty. The improvement record should name the finding, population, selected response, responsible authority, resource, intended operating change, delivery evidence, learner evidence and review date.
At the 30 March 2022 cutoff, later global reports and recommendations published in the final months of 2021 were unavailable and are excluded. No later recovery statistics or results are used. The report establishes an evidence-and-feasibility method for decisions available by the cutoff.
executive summary requires a decision about authorities should preserve entitlement, learner contact, teaching, support, records, progression and remedy while applying proportionate health-related rules issued by competent bodies. A contrary reading would overlook that education continuity during COVID-19 is a continuing public duty under altered and uncertain conditions. Closure, remote provision, partial reopening and renewed interruption can coexist across places and levels.
The central question in executive summary is local discretion may be necessary for conditions and capacity, but common minimum rights, transparency and escalation should remain. A contrary reading would overlook that regulation should identify its purpose, legal authority, covered institutions, duration, evidence and review. Education bodies should translate applicable public-health requirements into feasible institutional arrangements without claiming clinical competence they do not possess.
In assessing executive summary, authorities must determine a uniform rule can have unequal effects and may require accommodation or a supported alternative. A contrary reading would overlook that proportionality requires that a measure be connected to a legitimate aim, suitable for that aim, no broader than necessary and reviewed as evidence changes. Consequences for access, disability, transport, staffing, safeguarding and household burden should be examined.
executive summary requires a decision about the institutional consequence follows from whether continuity measures should be described according to actual service. The resulting interpretation should show why provision of a platform, broadcast or paper package does not establish learner access, teacher interaction, sustained participation or learning. Authorities should monitor reach and barriers without using surveillance or personal information beyond authorised need. Review of executive summary is credible only where it explains institutions require resources and a route to report conditions they cannot correct alone.
Evidence concerning executive summary should establish distributional evidence, accessible communication and participation should inform decisions. The public account remains incomplete unless it explains how the 2020 Global Education Monitoring Report on inclusion is within the evidence period and reinforces the need to identify learners excluded by ordinary arrangements. Emergency measures should not make already marginalised populations analytically invisible.
Comparative interpretation of executive summary depends upon institutions can organise timetables and support within their capacity; central and local authorities remain responsible for standards, finance, staffing and cross-sector dependencies. The evidence must therefore clarify how accountability should follow authority and control. Teachers can adapt instruction and maintain contact within their professional role; they cannot supply household connectivity, public transport or specialist services.
Public responsibility for executive summary begins with high-consequence decisions should not assume that all learners received the same instructional access. The institutional consequence follows from whether diagnostic evidence can guide teaching, while certification and transition require transparent rules and review. Changes in curriculum, assessment or population should remain visible in trend reporting. Assessment and progression rules should respond to unequal opportunity without abandoning credibility.
For executive summary, the material distinction is between confidence should not be claimed from communication activity alone; it is supported when institutions act consistently and acknowledge limitations. A proportionate conclusion must also recognise that public confidence depends on dated, accessible reasons and correction. Authorities should explain what changed, why, which evidence informed the decision, what institutions and families must do, and where an exception or complaint can be considered.
At the 30 March 2022 evidence cut-off, education systems continued to experience closure, reopening and disrupted provision associated with COVID-19. The report uses only official material available by that date and does not import later health guidance, statistics or results.
executive summary requires a decision about it is an authorised decision about whether and under what conditions learners and staff can resume in-person education. The evidence must therefore clarify how the responsible authority should consider public-health advice available at the time, legal competence, institutional readiness, protection, transport, water and sanitation, staffing, learner access and the capacity to detect and correct unequal effects. Education bodies should not invent clinical standards, while health advice should be translated into feasible educational arrangements by competent authorities. Reopening is not a single global date or a return to an unchanged service.
The central question in executive summary is learners may differ in health vulnerability, disability, transport, household responsibilities, access to information, experience of disruption and ability to use an adjusted timetable. The public account remains incomplete unless it explains how staff availability may vary. A nominally common rule can exclude people who cannot comply without accommodation or who depend on services that remain unavailable. Inclusion should be assessed before a general announcement.
executive summary requires a decision about uncertainty should be stated. For the learners concerned, the decisive consideration is whether an announcement unsupported by local readiness can weaken confidence and transfer risk to institutions and families. The planning record should distinguish minimum conditions, sequencing options and unresolved risks. It should identify who decides, what evidence is current, which institutions are covered, what local discretion exists and when the decision will be reviewed.
In assessing executive summary, authorities must determine publication of a rule does not establish that it reached or could be used by affected communities. The institutional consequence follows from whether public confidence depends on candour, consistency and a route for questions. Authorities should explain the educational and protection aims, evidence considered, responsibilities of institutions and families, and conditions that could change the decision. Information should be accessible across language, disability and communications access.
Evidence concerning executive summary should establish participation in an online or broadcast offer should not be treated automatically as equivalent to in-person instructional access. A proportionate conclusion must also recognise that the reopening baseline should identify missed contact and populations absent from the available evidence. Continuity during closure remains part of reopening readiness. Institutions should preserve learner contact, records, safeguarding routes, assessment decisions and support without assuming that remote provision reached everyone.
The practical standard for executive summary concerns teachers need preparation, materials, time and clear referral arrangements. The evidence must therefore clarify how they should not be required to make health judgements beyond their competence. Teaching plans should protect curriculum coherence while acknowledging lost time and altered conditions. Immediate diagnostic work can identify current learning and support needs, but high-consequence testing during disrupted access may amplify inequity.
The practical standard for executive summary concerns a central standard should be matched by finance and supply. This matters because institutions should not be declared non-compliant for conditions controlled by another public authority without a correction route. Facilities require a local readiness account: usable space, water, sanitation, cleaning, ventilation as then understood, safe routes, accessibility and arrangements for learners who need assistance.
The practical standard for executive summary concerns criteria should be public, rights-respecting and reviewed for unequal effect. The institutional consequence follows from whether prioritisation should not create an indefinite lower entitlement for learners with disabilities, displaced learners, remote communities or those unable to use remote alternatives. Each temporary arrangement needs an end point and learner-facing support. Sequencing may be necessary where capacity cannot support full return.
executive summary requires a decision about absence should trigger contact and support, not automatic sanction, while safeguarding concerns remain subject to competent action. For the learners concerned, the decisive consideration is whether attendance reporting should distinguish enrolment, expected attendance under the adjusted arrangement, actual presence, authorised absence and inability to access. Comparisons with earlier attendance require caution because timetables and denominators may differ.
Public responsibility for executive summary begins with international guidance and experience published after the cutoff remain outside the record. The institutional consequence follows from whether no claim of safe or successful reopening can be made from an adopted plan alone. Evidence should follow implementation, reach, reported incidents, continuity, participation and educational condition. This report establishes governance, continuity and review tests.
A defensible account of executive summary distinguishes national authorities remain responsible for translating goals into lawful, equitable and financed delivery. The institutional consequence follows from whether system alignment for learning requires a coherent chain from educational entitlement and curriculum goals to teachers, materials, finance, institutional practice, assessment and correction. The World Development Report 2018 distinguishes schooling from learning and calls for assessment, action on evidence and alignment of actors. Those propositions are used here as a contemporaneous policy reference, not as a universal administrative model.
The central question in executive summary is alignment should therefore be tested around learner-facing service, not the formal consistency of plans. This matters because misalignment arises when goals, incentives and resources point in different directions. A curriculum may require forms of teaching for which teachers lack time or materials; an assessment may reward only a narrow portion of the curriculum; finance may reach institutions late or under categories that cannot answer the identified need; accountability may attach consequence to actors who cannot control the relevant condition.
Public responsibility for executive summary begins with measures should support diagnosis and improvement while preserving levels, distributions, coverage and uncertainty. A contrary reading would overlook that learning goals should remain connected to the right to education and to broad educational purposes. Attention to foundational learning is necessary, but it should not narrow entitlement to what can be measured easily. Participation, progression, subject knowledge, transferable capability, inclusion and recognised completion remain material.
Evidence concerning executive summary should establish a national allocation does not prove that the intended condition reached classrooms. The evidence must therefore clarify how finance is aligned when it supplies the conditions required by the chosen goals. Authorities should trace appropriation, release, transfer, institutional receipt, expenditure and service. Teacher preparation, instructional materials, accessibility, assessment, learner support and maintenance have different cost structures.
The central question in executive summary is central bodies may set standards and equalisation; local bodies may plan places and distribute resources; institutions organise teaching and support; teachers exercise professional judgement. For the learners concerned, the decisive consideration is whether the responsibility map should name the decision-maker, contribution, deadline, evidence and escalation route. Holding one actor answerable for a system-wide failure can weaken candour and encourage avoidance. Delivery depends on authority at several levels.
At the 30 March 2022 cut-off, the Global Compact on Refugees and New York Declaration are available according to their adopted status. They do not establish later implementation results.
The central question in executive summary is central, local and institutional bodies should know which decisions they control, which standards bind them, which resources they receive and to whom they must explain performance. For the learners concerned, the decisive consideration is whether shared work is necessary, but a general declaration that all actors are responsible can leave no authority answerable for a failed service. Accountability should connect a defined duty to evidence, public explanation, review and correction. Public accountability in education begins with assigned responsibility.
Institutional action on executive summary should be tested against the essential requirement is an authority map showing who decides, who contributes, who reviews and who can remedy failure. The public account remains incomplete unless it explains how central authorities ordinarily hold responsibility for the legal framework, national standards, equalisation, system finance and an intelligible public account. Local authorities may plan places, allocate staff or resources, supervise services and coordinate with communities. Institutions deliver teaching, learner support and daily safeguards under applicable authority. These roles vary by jurisdiction.
The central question in executive summary is transfers, staffing, procurement, service availability and learner experience require distinct evidence. A proportionate conclusion must also recognise that decentralisation does not transfer responsibility credibly unless duty, authority, finance and information move together. A local body cannot be held to account for a service when it lacks the lawful authority or resources to provide it. Conversely, central government should not treat a national allocation or policy as proof of local delivery.
A defensible account of executive summary distinguishes teachers and leaders can be answerable for work within their competence and control; they should not bear sole responsibility for curriculum design, class size, infrastructure, poverty or unavailable specialist services. The institutional consequence follows from whether review of executive summary is credible only where it explains measures that attach consequence to outcomes beyond institutional control can encourage avoidance, selective admission or manipulation. The resulting interpretation should show why review should examine the conditions under which work was undertaken. Institutional accountability should preserve professional judgement.
Evidence concerning executive summary should establish a law establishes duty; a budget establishes authorised finance; expenditure shows use of resources; an institutional record may show delivery; learner evidence may show access, experience or outcome. A proportionate conclusion must also recognise that no state automatically proves the next. Reports should date evidence, define populations, preserve uncertainty and identify missing groups. Public reporting should separate commitment, authorisation, activity, service and result.
Institutional action on executive summary should be tested against learners, families, teachers and communities may reveal barriers absent from routine records. For the learners concerned, the decisive consideration is whether accessible information, consultation, complaint and reasoned response are distinct functions. Participation does not replace representative authority, technical evidence or protection of personal information. Participation strengthens accountability when it can alter a decision.
Institutional action on executive summary should be tested against central authorities should ensure common minimum entitlements and equalisation; local bodies should identify geographic and service barriers; institutions should examine participation and support. The evidence must therefore clarify how missing populations and uncertain denominators should remain visible rather than disappear from a favourable aggregate. Equity requires distributional evidence. National averages can improve while disadvantaged populations remain excluded.
Review of executive summary is credible only where it explains multiple overlapping reporting requirements can consume time without improving public knowledge. The evidence must therefore clarify how information should be collected once where feasible, governed by a clear definition and used for an identified decision. Accountability mechanisms should be proportionate. High-consequence duties warrant formal reasons, independent review and remedy. Routine professional adaptation may need a lighter record.
Review of executive summary is credible only where it explains dependencies beyond local or institutional control should be escalated and remain open. The evidence must therefore clarify how unresolved learner cases require a continuing route even when a programme or review closes. Correction is the final test. A complaint count, inspection finding or poor result is not resolved by issuing another instruction. The competent authority should identify cause, select a response, resource it and verify the changed condition.
executive summary cannot be judged without identifying the 2017/18 Global Education Monitoring Report and World Development Report 2018 are within the evidence period and inform the analysis of responsibility and alignment. This matters because their recommendations do not replace national authority or establish later implementation results. The conclusion for supporting teachers to address wider variation in learner attainment should connect the finding to the competent authority, affected learners and evidence required at review.
Key findings
Scope and method
Part I
The accountability chain
What national averages conceal
what national averages conceal requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A contrary reading would overlook that the indicator question concerns national participation or attainment. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]
what national averages conceal cannot be judged without identifying a discrepancy is not resolved by selecting the more favourable source. The resulting interpretation should show why it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is the weighted experience of the population included in the denominator. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-02] [REF-03]
For what national averages conceal, the material distinction is between where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. For the learners concerned, the decisive consideration is whether statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: distribution within countries, not a league table between them. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning.[REF-04] [REF-06]
The central question in what national averages conceal is if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The institutional consequence follows from whether confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include remote rural learners, urban informal settlements and displaced populations. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-08] [REF-09]
In assessing what national averages conceal, authorities must determine the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The public account remains incomplete unless it explains how the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked.[REF-10] [REF-12]
Marginalisation as accumulated disadvantage
A defensible account of marginalisation as accumulated disadvantage distinguishes the substantive interest is distance from a socially secured educational minimum, observed for a population and period that are stated before calculation. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Marginalisation as accumulated disadvantage should be approached as a defined measurement problem.[REF-02] [REF-03]
The central question in marginalisation as accumulated disadvantage is this formulation requires the reporting body to preserve the population base and the observation period beside the result. For the learners concerned, the decisive consideration is whether counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use the joint effect of exclusion, weak provision and adverse social conditions.[REF-04] [REF-06]
A defensible account of marginalisation as accumulated disadvantage distinguishes disparity measures should not replace the underlying distributions. The evidence must therefore clarify how a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that multiple indicators read together over the learner's course.[REF-08] [REF-09]
A defensible account of marginalisation as accumulated disadvantage distinguishes the review should record non-response, unknown status and excluded locations separately. A proportionate conclusion must also recognise that combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for children facing poverty, gender disadvantage, disability or minority status. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-10] [REF-12]
marginalisation as accumulated disadvantage cannot be judged without identifying urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. A contrary reading would overlook that targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-13] [REF-14]
From monitoring commitment to decision
The central question in from monitoring commitment to decision is a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The evidence must therefore clarify how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For from monitoring commitment to decision, the first requirement is conceptual clarity. The study seeks evidence on evidence capable of changing resource or service decisions; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-04] [REF-06]
from monitoring commitment to decision requires a decision about the definition should be fixed for the comparison at hand and deviations recorded. The evidence must therefore clarify how analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on a link between observed disparity, responsible body and remedy.[REF-08] [REF-09]
from monitoring commitment to decision cannot be judged without identifying nor should a group estimate be read as a description of every member. The resulting interpretation should show why within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: indicators selected for action rather than visibility. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-10] [REF-12]
Evidence concerning from monitoring commitment to decision should establish missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The public account remains incomplete unless it explains how an adequate equity account asks whether groups absent from routine plans and budget classifications are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-13] [REF-14]
In assessing from monitoring commitment to decision, authorities must determine local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. A contrary reading would overlook that communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable.[REF-15] [REF-16]
Defining the population entitled to education
defining the population entitled to education cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The public account remains incomplete unless it explains how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of resident and temporarily absent learners within the relevant age or programme group begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]
The practical standard for defining the population entitled to education concerns its metadata should travel with every published value. The resulting interpretation should show why at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is a denominator consistent with the right, level and reference period under review.[REF-10] [REF-12]
The practical standard for defining the population entitled to education concerns apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The institutional consequence follows from whether use of the indicator is bounded by the principle that census, survey and administrative estimates reconciled openly. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-13] [REF-14]
Public responsibility for defining the population entitled to education begins with suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. This matters because the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For unregistered residents, migrants and displaced children, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-15] [REF-16]
The practical standard for defining the population entitled to education concerns readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A proportionate conclusion must also recognise that a policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary.[REF-17] [REF-19]
Part II
Accountability by level
Age, grade and programme populations
The practical standard for age, grade and programme populations concerns the measure should preserve both the observed level and the distribution relevant to the claim. The resulting interpretation should show why a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. Evidence concerning age, grade and programme populations should establish the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether the public value of age, grade and programme populations lies in making unequal educational experience observable. Here, the relevant phenomenon is age-specific, grade-specific and programme-specific participation, not the administrative convenience of the available categories.[REF-10] [REF-12]
For age, grade and programme populations, the material distinction is between administrative records should be reconciled with population-based evidence where their coverage differs. The evidence must therefore clarify how a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is separate denominators for different educational questions. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-13] [REF-14]
Public responsibility for age, grade and programme populations begins with explanation requires evidence on institutions, resources, households and prior conditions. This matters because interpretation follows this limitation: exact age, school age and enrolled population never substituted silently. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-15] [REF-16]
Comparative interpretation of age, grade and programme populations depends upon confidentiality is essential, especially where identity or status creates risk. The public account remains incomplete unless it explains how protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include over-age entrants and learners repeating grades. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-17] [REF-19]
Population movement and disrupted residence
The central question in population movement and disrupted residence is the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The resulting interpretation should show why the indicator question concerns education status amid migration, displacement and return. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-13] [REF-14]
Public responsibility for population movement and disrupted residence begins with this formulation requires the reporting body to preserve the population base and the observation period beside the result. A proportionate conclusion must also recognise that counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use dated location and residence rules with sensitivity to mobility.[REF-15] [REF-16]
A defensible account of population movement and disrupted residence distinguishes policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. A proportionate conclusion must also recognise that the governing caution is that origin, current location and service responsibility distinguished. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-17] [REF-19]
population movement and disrupted residence requires a decision about combining unknown observations with the majority group biases both estimates and obscures the weakness. The resulting interpretation should show why where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for families affected by conflict, disaster and seasonal movement. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-20] [REF-21]
Entry at the official starting age
Evidence concerning entry at the official starting age should establish a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The institutional consequence follows from whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Entry at the official starting age should be approached as a defined measurement problem. The substantive interest is timely admission to the first grade of primary education, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-15] [REF-16]
entry at the official starting age requires a decision about when several sources exist, consistency is evidence to consider, not proof that common error is absent. The institutional consequence follows from whether a defensible statistic would be based on new entrants of official age relative to the corresponding population. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-17] [REF-19]
A defensible account of entry at the official starting age distinguishes nor should a group estimate be read as a description of every member. A proportionate conclusion must also recognise that within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: late entry examined alongside non-entry. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-20] [REF-21]
The central question in entry at the official starting age is analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. For the learners concerned, the decisive consideration is whether missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether children facing fees, distance, disability or documentation barriers are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-23] [REF-24]
Attendance beyond enrolment
A defensible account of attendance beyond enrolment distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The resulting interpretation should show why for attendance beyond enrolment, the first requirement is conceptual clarity. The study seeks evidence on actual participation during a stated recent period; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-17] [REF-19]
Evidence concerning attendance beyond enrolment should establish at minimum this includes population, geography, date, collection method, classification and known exclusions. The resulting interpretation should show why a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is presence measured independently of registration status. Its metadata should travel with every published value.[REF-20] [REF-21]
Review of attendance beyond enrolment is credible only where it explains apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The resulting interpretation should show why use of the indicator is bounded by the principle that frequency, season and reasons for absence retained. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-23] [REF-24]
attendance beyond enrolment cannot be judged without identifying suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The resulting interpretation should show why the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For working children, carers and learners affected by illness, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-01] [REF-07]
Part III
Public reporting and participation
Out-of-school status
In assessing out-of-school status, authorities must determine the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that a distribution-sensitive account of children in the relevant age group not participating at the defined level begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]
The practical standard for out-of-school status concerns numerator, denominator, reference date, unit and exclusions should appear together. The evidence must therefore clarify how if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a transparent residual from compatible population and participation concepts.[REF-23] [REF-24]
The central question in out-of-school status is explanation requires evidence on institutions, resources, households and prior conditions. A proportionate conclusion must also recognise that interpretation follows this limitation: never-enrolled and formerly enrolled children separated. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-01] [REF-07]
Evidence concerning out-of-school status should establish these populations may be missing not only from good outcomes but from the denominator itself. This matters because coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children invisible to school registers.[REF-02] [REF-03]
Repetition and grade survival
repetition and grade survival cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. This matters because the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of repetition and grade survival lies in making unequal educational experience observable. Here, the relevant phenomenon is movement through grades without avoidable delay or exit, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-23] [REF-24]
Evidence concerning repetition and grade survival should establish where estimates are revised, both the reason and effect of revision should remain accessible. This matters because measurement should use cohort or reconstructed-cohort evidence with explicit assumptions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-01] [REF-07]
repetition and grade survival requires a decision about policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. For the learners concerned, the decisive consideration is whether the governing caution is that repeaters distinguished from re-entrants and transfers. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-02] [REF-03]
The practical standard for repetition and grade survival concerns combining unknown observations with the majority group biases both estimates and obscures the weakness. The evidence must therefore clarify how where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in overcrowded or intermittently operating schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-04] [REF-06]
Transition between education levels
transition between education levels cannot be judged without identifying its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns entry to the next level after completion of the preceding one.[REF-01] [REF-07]
Institutional action on transition between education levels should be tested against the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. For the learners concerned, the decisive consideration is whether when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on matched completion and new-entry populations over coherent periods. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-02] [REF-03]
The central question in transition between education levels is nor should a group estimate be read as a description of every member. The institutional consequence follows from whether within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: capacity constraints distinguished from learner attainment. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-04] [REF-06]
transition between education levels requires a decision about analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The resulting interpretation should show why missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether rural learners and those unable to relocate are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-08] [REF-09]
Part I
Purpose and interpretive frame
What national averages conceal
what national averages conceal requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A contrary reading would overlook that the indicator question concerns national participation or attainment. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]
what national averages conceal cannot be judged without identifying a discrepancy is not resolved by selecting the more favourable source. The resulting interpretation should show why it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is the weighted experience of the population included in the denominator. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-02] [REF-03]
For what national averages conceal, the material distinction is between where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. For the learners concerned, the decisive consideration is whether statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: distribution within countries, not a league table between them. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning.[REF-04] [REF-06]
The central question in what national averages conceal is if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The institutional consequence follows from whether confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include remote rural learners, urban informal settlements and displaced populations. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-08] [REF-09]
In assessing what national averages conceal, authorities must determine the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The public account remains incomplete unless it explains how the decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked.[REF-10] [REF-12]
Marginalisation as accumulated disadvantage
A defensible account of marginalisation as accumulated disadvantage distinguishes the substantive interest is distance from a socially secured educational minimum, observed for a population and period that are stated before calculation. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Marginalisation as accumulated disadvantage should be approached as a defined measurement problem.[REF-02] [REF-03]
The central question in marginalisation as accumulated disadvantage is this formulation requires the reporting body to preserve the population base and the observation period beside the result. For the learners concerned, the decisive consideration is whether counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use the joint effect of exclusion, weak provision and adverse social conditions.[REF-04] [REF-06]
A defensible account of marginalisation as accumulated disadvantage distinguishes disparity measures should not replace the underlying distributions. The evidence must therefore clarify how a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that multiple indicators read together over the learner's course.[REF-08] [REF-09]
A defensible account of marginalisation as accumulated disadvantage distinguishes the review should record non-response, unknown status and excluded locations separately. A proportionate conclusion must also recognise that combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for children facing poverty, gender disadvantage, disability or minority status. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-10] [REF-12]
marginalisation as accumulated disadvantage cannot be judged without identifying urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. A contrary reading would overlook that targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-13] [REF-14]
From monitoring commitment to decision
The central question in from monitoring commitment to decision is a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The evidence must therefore clarify how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For from monitoring commitment to decision, the first requirement is conceptual clarity. The study seeks evidence on evidence capable of changing resource or service decisions; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-04] [REF-06]
from monitoring commitment to decision requires a decision about the definition should be fixed for the comparison at hand and deviations recorded. The evidence must therefore clarify how analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on a link between observed disparity, responsible body and remedy.[REF-08] [REF-09]
from monitoring commitment to decision cannot be judged without identifying nor should a group estimate be read as a description of every member. The resulting interpretation should show why within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: indicators selected for action rather than visibility. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-10] [REF-12]
Evidence concerning from monitoring commitment to decision should establish missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The public account remains incomplete unless it explains how an adequate equity account asks whether groups absent from routine plans and budget classifications are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-13] [REF-14]
In assessing from monitoring commitment to decision, authorities must determine local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. A contrary reading would overlook that communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable.[REF-15] [REF-16]
Part II
Population and denominator
Defining the population entitled to education
defining the population entitled to education cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The public account remains incomplete unless it explains how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of resident and temporarily absent learners within the relevant age or programme group begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]
The practical standard for defining the population entitled to education concerns its metadata should travel with every published value. The resulting interpretation should show why at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is a denominator consistent with the right, level and reference period under review.[REF-10] [REF-12]
The practical standard for defining the population entitled to education concerns apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The institutional consequence follows from whether use of the indicator is bounded by the principle that census, survey and administrative estimates reconciled openly. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-13] [REF-14]
Public responsibility for defining the population entitled to education begins with suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. This matters because the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For unregistered residents, migrants and displaced children, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-15] [REF-16]
The practical standard for defining the population entitled to education concerns readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A proportionate conclusion must also recognise that a policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary.[REF-17] [REF-19]
Age, grade and programme populations
The practical standard for age, grade and programme populations concerns the measure should preserve both the observed level and the distribution relevant to the claim. The resulting interpretation should show why a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. Evidence concerning age, grade and programme populations should establish the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether the public value of age, grade and programme populations lies in making unequal educational experience observable. Here, the relevant phenomenon is age-specific, grade-specific and programme-specific participation, not the administrative convenience of the available categories.[REF-10] [REF-12]
For age, grade and programme populations, the material distinction is between administrative records should be reconciled with population-based evidence where their coverage differs. The evidence must therefore clarify how a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is separate denominators for different educational questions. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-13] [REF-14]
Public responsibility for age, grade and programme populations begins with explanation requires evidence on institutions, resources, households and prior conditions. This matters because interpretation follows this limitation: exact age, school age and enrolled population never substituted silently. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-15] [REF-16]
Comparative interpretation of age, grade and programme populations depends upon confidentiality is essential, especially where identity or status creates risk. The public account remains incomplete unless it explains how protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include over-age entrants and learners repeating grades. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-17] [REF-19]
Population movement and disrupted residence
The central question in population movement and disrupted residence is the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The resulting interpretation should show why the indicator question concerns education status amid migration, displacement and return. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-13] [REF-14]
Public responsibility for population movement and disrupted residence begins with this formulation requires the reporting body to preserve the population base and the observation period beside the result. A proportionate conclusion must also recognise that counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use dated location and residence rules with sensitivity to mobility.[REF-15] [REF-16]
A defensible account of population movement and disrupted residence distinguishes policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. A proportionate conclusion must also recognise that the governing caution is that origin, current location and service responsibility distinguished. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-17] [REF-19]
population movement and disrupted residence requires a decision about combining unknown observations with the majority group biases both estimates and obscures the weakness. The resulting interpretation should show why where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for families affected by conflict, disaster and seasonal movement. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-20] [REF-21]
Part III
Access and participation
Entry at the official starting age
Evidence concerning entry at the official starting age should establish a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The institutional consequence follows from whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Entry at the official starting age should be approached as a defined measurement problem. The substantive interest is timely admission to the first grade of primary education, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-15] [REF-16]
entry at the official starting age requires a decision about when several sources exist, consistency is evidence to consider, not proof that common error is absent. The institutional consequence follows from whether a defensible statistic would be based on new entrants of official age relative to the corresponding population. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-17] [REF-19]
A defensible account of entry at the official starting age distinguishes nor should a group estimate be read as a description of every member. A proportionate conclusion must also recognise that within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: late entry examined alongside non-entry. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-20] [REF-21]
The central question in entry at the official starting age is analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. For the learners concerned, the decisive consideration is whether missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether children facing fees, distance, disability or documentation barriers are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-23] [REF-24]
Attendance beyond enrolment
A defensible account of attendance beyond enrolment distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The resulting interpretation should show why for attendance beyond enrolment, the first requirement is conceptual clarity. The study seeks evidence on actual participation during a stated recent period; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-17] [REF-19]
Evidence concerning attendance beyond enrolment should establish at minimum this includes population, geography, date, collection method, classification and known exclusions. The resulting interpretation should show why a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is presence measured independently of registration status. Its metadata should travel with every published value.[REF-20] [REF-21]
Review of attendance beyond enrolment is credible only where it explains apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The resulting interpretation should show why use of the indicator is bounded by the principle that frequency, season and reasons for absence retained. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-23] [REF-24]
attendance beyond enrolment cannot be judged without identifying suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The resulting interpretation should show why the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For working children, carers and learners affected by illness, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-01] [REF-07]
Out-of-school status
In assessing out-of-school status, authorities must determine the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that a distribution-sensitive account of children in the relevant age group not participating at the defined level begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]
The practical standard for out-of-school status concerns numerator, denominator, reference date, unit and exclusions should appear together. The evidence must therefore clarify how if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a transparent residual from compatible population and participation concepts.[REF-23] [REF-24]
The central question in out-of-school status is explanation requires evidence on institutions, resources, households and prior conditions. A proportionate conclusion must also recognise that interpretation follows this limitation: never-enrolled and formerly enrolled children separated. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-01] [REF-07]
Evidence concerning out-of-school status should establish these populations may be missing not only from good outcomes but from the denominator itself. This matters because coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children invisible to school registers.[REF-02] [REF-03]
Part IV
Progression and completion
Repetition and grade survival
repetition and grade survival cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. This matters because the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of repetition and grade survival lies in making unequal educational experience observable. Here, the relevant phenomenon is movement through grades without avoidable delay or exit, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-23] [REF-24]
Evidence concerning repetition and grade survival should establish where estimates are revised, both the reason and effect of revision should remain accessible. This matters because measurement should use cohort or reconstructed-cohort evidence with explicit assumptions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-01] [REF-07]
repetition and grade survival requires a decision about policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. For the learners concerned, the decisive consideration is whether the governing caution is that repeaters distinguished from re-entrants and transfers. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-02] [REF-03]
The practical standard for repetition and grade survival concerns combining unknown observations with the majority group biases both estimates and obscures the weakness. The evidence must therefore clarify how where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in overcrowded or intermittently operating schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-04] [REF-06]
Transition between education levels
transition between education levels cannot be judged without identifying its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns entry to the next level after completion of the preceding one.[REF-01] [REF-07]
Institutional action on transition between education levels should be tested against the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. For the learners concerned, the decisive consideration is whether when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on matched completion and new-entry populations over coherent periods. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-02] [REF-03]
The central question in transition between education levels is nor should a group estimate be read as a description of every member. The institutional consequence follows from whether within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: capacity constraints distinguished from learner attainment. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-04] [REF-06]
transition between education levels requires a decision about analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The resulting interpretation should show why missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether rural learners and those unable to relocate are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-08] [REF-09]
Completion and educational entitlement
A defensible account of completion and educational entitlement distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether completion and educational entitlement should be approached as a defined measurement problem. The substantive interest is finishing the final grade or meeting recognised programme requirements, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-02] [REF-03]
For completion and educational entitlement, the material distinction is between its metadata should travel with every published value. The public account remains incomplete unless it explains how at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is completion defined separately from sitting or passing an examination.[REF-04] [REF-06]
In assessing completion and educational entitlement, authorities must determine it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. The public account remains incomplete unless it explains how it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that late completion and alternative pathways reported. A responsible commentary distinguishes observation, calculation and interpretation.[REF-08] [REF-09]
Comparative interpretation of completion and educational entitlement depends upon disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. A contrary reading would overlook that this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For over-age learners and people returning after interruption, a single group label may conceal important internal differences.[REF-10] [REF-12]
Part V
Learning and assessment
Minimum learning outcomes
Evidence concerning minimum learning outcomes should establish the measure should preserve both the observed level and the distribution relevant to the claim. The institutional consequence follows from whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For minimum learning outcomes, the first requirement is conceptual clarity. The study seeks evidence on demonstrated knowledge or skill against a declared domain; it does not infer a learner's circumstances from a national or regional mean.[REF-04] [REF-06]
The practical standard for minimum learning outcomes concerns numerator, denominator, reference date, unit and exclusions should appear together. For the learners concerned, the decisive consideration is whether if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is assessment evidence whose population and conditions are known.[REF-08] [REF-09]
minimum learning outcomes cannot be judged without identifying reference points therefore need substantive meaning. For the learners concerned, the decisive consideration is whether where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: results interpreted with opportunity to learn and participation. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-10] [REF-12]
minimum learning outcomes cannot be judged without identifying if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The resulting interpretation should show why confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners taught in an unfamiliar language. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-13] [REF-14]
Assessment participation
Review of assessment participation is credible only where it explains the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of who was eligible, present, absent and excluded from testing begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]
Institutional action on assessment participation should be tested against source coverage must be tested before sources are combined. This matters because school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use a participation profile accompanying every result distribution. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone.[REF-10] [REF-12]
A defensible account of assessment participation distinguishes policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The institutional consequence follows from whether the governing caution is that non-participation never treated as low attainment or ignored. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-13] [REF-14]
Evidence concerning assessment participation should establish the review should record non-response, unknown status and excluded locations separately. For the learners concerned, the decisive consideration is whether combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners with disabilities and remote candidates. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-15] [REF-16]
Distribution of achievement
distribution of achievement cannot be judged without identifying a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The resulting interpretation should show why the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of distribution of achievement lies in making unequal educational experience observable. Here, the relevant phenomenon is variation across the full score or proficiency distribution, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-10] [REF-12]
The practical standard for distribution of achievement concerns the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. A contrary reading would overlook that when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on percentiles and threshold shares beside a mean. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-13] [REF-14]
For distribution of achievement, the material distinction is between nor should a group estimate be read as a description of every member. The public account remains incomplete unless it explains how within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: uncertainty and scale properties stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-15] [REF-16]
For distribution of achievement, the material distinction is between missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The institutional consequence follows from whether an adequate equity account asks whether learners concentrated below a minimum proficiency threshold are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-17] [REF-19]
Part VI
Gender and household resources
Gender parity and its limits
Evidence concerning gender parity and its limits should establish its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns differences between girls and boys in access, progression and learning.[REF-13] [REF-14]
A defensible account of gender parity and its limits distinguishes a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. The public account remains incomplete unless it explains how a survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is female-to-male ratios read with levels and absolute gaps. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions.[REF-15] [REF-16]
Public responsibility for gender parity and its limits begins with apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. A contrary reading would overlook that use of the indicator is bounded by the principle that parity not confused with adequacy for either group. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-17] [REF-19]
Public responsibility for gender parity and its limits begins with it should involve statistical judgement, legal safeguards and knowledge of the affected community. The evidence must therefore clarify how suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For girls in poor rural households and boys exposed to hazardous work, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-20] [REF-21]
Household wealth gradients
Comparative interpretation of household wealth gradients depends upon the substantive interest is education outcomes across relative household resource groups, observed for a population and period that are stated before calculation. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The practical standard for household wealth gradients concerns the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether household wealth gradients should be approached as a defined measurement problem.[REF-15] [REF-16]
Comparative interpretation of household wealth gradients depends upon numerator, denominator, reference date, unit and exclusions should appear together. A contrary reading would overlook that if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a documented asset or consumption classification within each setting.[REF-17] [REF-19]
In assessing household wealth gradients, authorities must determine the comparison should show the level for each group as well as any ratio or gap. A proportionate conclusion must also recognise that a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: wealth ranks not assumed equivalent across countries or time.[REF-20] [REF-21]
The central question in household wealth gradients is if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The resulting interpretation should show why confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children in the poorest quintile and those near classification boundaries. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-23] [REF-24]
Costs borne by households
A defensible account of costs borne by households distinguishes the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For costs borne by households, the first requirement is conceptual clarity. The study seeks evidence on fees, materials, transport, clothing and foregone labour; it does not infer a learner's circumstances from a national or regional mean.[REF-17] [REF-19]
costs borne by households cannot be judged without identifying reconciliation should record what each source can and cannot represent. A proportionate conclusion must also recognise that where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use participation read against direct and indirect education costs. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-20] [REF-21]
Comparative interpretation of costs borne by households depends upon disparity measures should not replace the underlying distributions. The evidence must therefore clarify how a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that nominal fee abolition checked against remaining expenditure.[REF-23] [REF-24]
Institutional action on costs borne by households should be tested against their circumstances may alter access to enumeration, classification and the service being measured. This matters because the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for large families and households hit by economic crisis.[REF-01] [REF-07]
Part VII
Place and service geography
Rural and urban residence
In assessing rural and urban residence, authorities must determine the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether a distribution-sensitive account of education outcomes by a declared settlement classification begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]
The practical standard for rural and urban residence concerns analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. The institutional consequence follows from whether this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on residence linked to service availability and travel conditions. The definition should be fixed for the comparison at hand and deviations recorded.[REF-23] [REF-24]
rural and urban residence requires a decision about if a conclusion changes under a reasonable specification, that instability is part of the finding. The evidence must therefore clarify how national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: national rural definitions preserved and comparison limitations stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-01] [REF-07]
Institutional action on rural and urban residence should be tested against attrition at any stage can produce an apparently complete indicator from a selective population. A proportionate conclusion must also recognise that field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether remote villages, pastoral populations and peri-urban settlements are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-02] [REF-03]
Subnational administrative disparity
For subnational administrative disparity, the material distinction is between a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The evidence must therefore clarify how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of subnational administrative disparity lies in making unequal educational experience observable. Here, the relevant phenomenon is variation between provinces, districts or comparable areas, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-23] [REF-24]
Review of subnational administrative disparity is credible only where it explains these are substantive attributes because they determine who can appear in the evidence. This matters because a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is area estimates with population size and precision. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-01] [REF-07]
Review of subnational administrative disparity is credible only where it explains it does not assign cause from a cross-sectional difference. For the learners concerned, the decisive consideration is whether apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that administrative rankings not mistaken for causal explanations. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-02] [REF-03]
subnational administrative disparity requires a decision about the public report can still state that a disparity was examined, whether action is required and which body will monitor it. A contrary reading would overlook that for small districts and areas with incomplete reporting, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-04] [REF-06]
Distance, isolation and transport
Review of distance, isolation and transport is credible only where it explains a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The evidence must therefore clarify how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns physical accessibility of the nearest appropriate service. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-01] [REF-07]
In assessing distance, isolation and transport, authorities must determine the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The resulting interpretation should show why the preferred construction is travel time, route safety and seasonal interruption rather than straight-line distance alone. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-02] [REF-03]
Evidence concerning distance, isolation and transport should establish statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. This matters because explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: household reports and facility mapping reconciled. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-04] [REF-06]
Comparative interpretation of distance, isolation and transport depends upon confidentiality is essential, especially where identity or status creates risk. The evidence must therefore clarify how protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners with limited mobility and communities cut off seasonally. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-08] [REF-09]
Part VIII
Disability, language and identity
Disability-sensitive education data
A defensible account of disability-sensitive education data distinguishes the substantive interest is participation and learning by functional difficulty and support requirement, observed for a population and period that are stated before calculation. For the learners concerned, the decisive consideration is whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Disability-sensitive education data should be approached as a defined measurement problem.[REF-02] [REF-03]
The practical standard for disability-sensitive education data concerns school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. A contrary reading would overlook that reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use questions designed for comparable reporting without defining the child by diagnosis alone. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined.[REF-04] [REF-06]
disability-sensitive education data cannot be judged without identifying percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The resulting interpretation should show why the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that identification, environment and accommodation kept analytically distinct. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-08] [REF-09]
Public responsibility for disability-sensitive education data begins with combining unknown observations with the majority group biases both estimates and obscures the weakness. A proportionate conclusion must also recognise that where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners whose impairments are not recorded by schools. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-10] [REF-12]
Language of home and instruction
Evidence concerning language of home and instruction should establish the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public account remains incomplete unless it explains how for language of home and instruction, the first requirement is conceptual clarity. The study seeks evidence on alignment between learner language, teaching and assessment; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-04] [REF-06]
language of home and instruction cannot be judged without identifying the definition should be fixed for the comparison at hand and deviations recorded. For the learners concerned, the decisive consideration is whether analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on language categories reflecting local use and instructional practice.[REF-08] [REF-09]
Institutional action on language of home and instruction should be tested against within-group variation and unmeasured intersecting conditions remain material. The resulting interpretation should show why the analytical rule is clear: small groups not erased through broad national labels. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-10] [REF-12]
Public responsibility for language of home and instruction begins with analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The institutional consequence follows from whether missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether minority-language and multilingual learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-13] [REF-14]
Ethnicity, indigeneity and protected identity
Comparative interpretation of ethnicity, indigeneity and protected identity depends upon a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of disparity associated with historically excluded identity groups begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]
For ethnicity, indigeneity and protected identity, the material distinction is between these are substantive attributes because they determine who can appear in the evidence. A contrary reading would overlook that a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is lawful, voluntary and contextually meaningful classification. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-10] [REF-12]
For ethnicity, indigeneity and protected identity, the material distinction is between apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The resulting interpretation should show why use of the indicator is bounded by the principle that self-identification protected and non-response reported. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-13] [REF-14]
Evidence concerning ethnicity, indigeneity and protected identity should establish suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The resulting interpretation should show why the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For communities exposed to discrimination or forced assimilation, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-15] [REF-16]
Part IX
Conflict, disaster and mobility
Education under conflict and insecurity
The practical standard for education under conflict and insecurity concerns a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of education under conflict and insecurity lies in making unequal educational experience observable. Here, the relevant phenomenon is access, attendance and learning where violence alters service and movement, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-10] [REF-12]
Comparative interpretation of education under conflict and insecurity depends upon it should require examination of definitions, timing, migration, duplication and non-response. The public account remains incomplete unless it explains how the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is location- and time-specific observation with explicit coverage gaps. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-13] [REF-14]
education under conflict and insecurity requires a decision about explanation requires evidence on institutions, resources, households and prior conditions. The resulting interpretation should show why interpretation follows this limitation: absence caused by insecurity distinguished from ordinary dropout. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-15] [REF-16]
education under conflict and insecurity requires a decision about if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. The evidence must therefore clarify how confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners in insecure areas and host communities. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-17] [REF-19]
Refugees, displaced persons and migrants
Evidence concerning refugees, displaced persons and migrants should establish the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Refugees, displaced persons and migrants should be approached as a defined measurement problem. The substantive interest is educational participation across changing legal and residential situations, observed for a population and period that are stated before calculation.[REF-15] [REF-16]
The central question in refugees, displaced persons and migrants is the resulting interpretation should show why this distinction is essential for participation and progression. The evidence must therefore clarify how the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on status, origin, current location and service access recorded separately. The definition should be fixed for the comparison at hand and deviations recorded. Institutional action on refugees, displaced persons and migrants should be tested against analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-17] [REF-19]
Institutional action on refugees, displaced persons and migrants should be tested against within-group variation and unmeasured intersecting conditions remain material. The resulting interpretation should show why the analytical rule is clear: mobility never converted into duplicate enrolment or unexplained disappearance. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-20] [REF-21]
Comparative interpretation of refugees, displaced persons and migrants depends upon field arrangements need relevant languages, accessible formats and safe participation. The institutional consequence follows from whether analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether undocumented migrants, refugees and internally displaced learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-23] [REF-24]
Part X
School conditions and teachers
Teacher availability and distribution
Review of teacher availability and distribution is credible only where it explains a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. For the learners concerned, the decisive consideration is whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For teacher availability and distribution, the first requirement is conceptual clarity. The study seeks evidence on access to competent teaching across schools and subjects; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-17] [REF-19]
The practical standard for teacher availability and distribution concerns its metadata should travel with every published value. A contrary reading would overlook that at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is teachers present and assigned relative to learner and curriculum need.[REF-20] [REF-21]
Public responsibility for teacher availability and distribution begins with use of the indicator is bounded by the principle that payroll totals not substituted for classroom availability. A contrary reading would overlook that review of teacher availability and distribution is credible only where it explains a responsible commentary distinguishes observation, calculation and interpretation. This matters because it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere.[REF-23] [REF-24]
For teacher availability and distribution, the material distinction is between the public report can still state that a disparity was examined, whether action is required and which body will monitor it. The institutional consequence follows from whether for schools serving poor, remote or displaced communities, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-01] [REF-07]
Class size, multi-grade teaching and time
A defensible account of class size, multi-grade teaching and time distinguishes a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. This matters because the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of the instructional conditions experienced by learners begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-20] [REF-21]
Institutional action on class size, multi-grade teaching and time should be tested against administrative records should be reconciled with population-based evidence where their coverage differs. A proportionate conclusion must also recognise that a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is class organisation, scheduled time and delivered time considered together. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-23] [REF-24]
In assessing class size, multi-grade teaching and time, authorities must determine a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. The evidence must therefore clarify how reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: simple pupil-teacher ratios not treated as a complete quality measure. The comparison should show the level for each group as well as any ratio or gap.[REF-01] [REF-07]
A defensible account of class size, multi-grade teaching and time distinguishes protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. A contrary reading would overlook that the distributional review must deliberately include early grades and mixed-age classes. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-02] [REF-03]
Materials, facilities and basic services
The central question in materials, facilities and basic services is here, the relevant phenomenon is usable learning resources and safe, accessible school conditions, not the administrative convenience of the available categories. The public account remains incomplete unless it explains how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of materials, facilities and basic services lies in making unequal educational experience observable.[REF-23] [REF-24]
A defensible account of materials, facilities and basic services distinguishes reconciliation should record what each source can and cannot represent. The resulting interpretation should show why where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use availability joined to condition, accessibility and regular use. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-01] [REF-07]
The practical standard for materials, facilities and basic services concerns the comparison should identify the reference category but avoid presenting it as a natural norm. The public account remains incomplete unless it explains how policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that delivery counts checked against learner access. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation.[REF-02] [REF-03]
Public responsibility for materials, facilities and basic services begins with combining unknown observations with the majority group biases both estimates and obscures the weakness. The resulting interpretation should show why where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in temporary or damaged premises. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-04] [REF-06]
Part XI
Finance and distribution
Public spending by level and function
For public spending by level and function, the material distinction is between the measure should preserve both the observed level and the distribution relevant to the claim. The resulting interpretation should show why a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns resources assigned to education purposes across the system. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-01] [REF-07]
The central question in public spending by level and function is analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. For the learners concerned, the decisive consideration is whether this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on expenditure classified by level, recurrent or capital use and responsible body. The definition should be fixed for the comparison at hand and deviations recorded.[REF-02] [REF-03]
Evidence concerning public spending by level and function should establish nor should a group estimate be read as a description of every member. For the learners concerned, the decisive consideration is whether within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: budgets, commitments and actual expenditure distinguished. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-04] [REF-06]
The central question in public spending by level and function is field arrangements need relevant languages, accessible formats and safe participation. For the learners concerned, the decisive consideration is whether analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether basic education services under fiscal pressure are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-08] [REF-09]
Incidence of education spending
incidence of education spending cannot be judged without identifying the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Incidence of education spending should be approached as a defined measurement problem. The substantive interest is who benefits from publicly financed places and services, observed for a population and period that are stated before calculation.[REF-02] [REF-03]
Public responsibility for incidence of education spending begins with a census figure should disclose enumeration rules. A contrary reading would overlook that these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is unit resources combined with participation across population groups. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-04] [REF-06]
In assessing incidence of education spending, authorities must determine apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The public account remains incomplete unless it explains how use of the indicator is bounded by the principle that benefit estimates not treated as household income. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-08] [REF-09]
A defensible account of incidence of education spending distinguishes suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The evidence must therefore clarify how the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For groups excluded before public spending can reach them, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-10] [REF-12]
Protecting equity during fiscal constraint
The practical standard for protecting equity during fiscal constraint concerns the measure should preserve both the observed level and the distribution relevant to the claim. The institutional consequence follows from whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For protecting equity during fiscal constraint, the first requirement is conceptual clarity. The study seeks evidence on whether reductions or delays fall disproportionately on weaker services; it does not infer a learner's circumstances from a national or regional mean.[REF-04] [REF-06]
For protecting equity during fiscal constraint, the material distinction is between numerator, denominator, reference date, unit and exclusions should appear together. The public account remains incomplete unless it explains how if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is dated finance and service indicators read together.[REF-08] [REF-09]
Institutional action on protecting equity during fiscal constraint should be tested against the comparison should show the level for each group as well as any ratio or gap. A contrary reading would overlook that a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: national totals tested against subnational allocation and household costs.[REF-10] [REF-12]
The central question in protecting equity during fiscal constraint is confidentiality is essential, especially where identity or status creates risk. A contrary reading would overlook that protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include poor households and institutions with little financial reserve. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-13] [REF-14]
Part XII
Data sources and measurement error
Administrative records
Institutional action on administrative records should be tested against the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of regular learner, staff, facility and finance information begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]
Institutional action on administrative records should be tested against this formulation requires the reporting body to preserve the population base and the observation period beside the result. A contrary reading would overlook that counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use clear definitions, reporting coverage and revision history.[REF-10] [REF-12]
administrative records requires a decision about percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. This matters because the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that non-reporting institutions kept visible in aggregates. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-13] [REF-14]
For administrative records, the material distinction is between their circumstances may alter access to enumeration, classification and the service being measured. The evidence must therefore clarify how the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for small, private, non-formal and emergency providers.[REF-15] [REF-16]
Household surveys
Review of household surveys is credible only where it explains here, the relevant phenomenon is population-based evidence beyond enrolled learners, not the administrative convenience of the available categories. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of household surveys lies in making unequal educational experience observable.[REF-10] [REF-12]
A defensible account of household surveys distinguishes analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This matters because this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on probability samples, weights and field dates documented. The definition should be fixed for the comparison at hand and deviations recorded.[REF-13] [REF-14]
In assessing household surveys, authorities must determine nor should a group estimate be read as a description of every member. This matters because within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: sampling and non-response uncertainty carried into group comparisons. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-15] [REF-16]
The practical standard for household surveys concerns field arrangements need relevant languages, accessible formats and safe participation. A proportionate conclusion must also recognise that analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether small minorities and mobile households are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-17] [REF-19]
Censuses and population frames
The central question in censuses and population frames is the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public account remains incomplete unless it explains how the indicator question concerns broad population coverage and small-area denominators. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-13] [REF-14]
Public responsibility for censuses and population frames begins with a survey estimate should disclose weights and uncertainty. This matters because a census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is enumeration date, usual residence and institutional coverage stated. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions.[REF-15] [REF-16]
A defensible account of censuses and population frames distinguishes it does not assign cause from a cross-sectional difference. The resulting interpretation should show why apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that long intervals and under-enumeration acknowledged. Evidence concerning censuses and population frames should establish a responsible commentary distinguishes observation, calculation and interpretation. The resulting interpretation should show why it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-17] [REF-19]
The central question in censuses and population frames is this balance is contextual rather than mechanical. A contrary reading would overlook that it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For homeless, displaced and geographically isolated people, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-20] [REF-21]
Part XIII
Disaggregation and intersection
Single-axis disaggregation
Public responsibility for single-axis disaggregation begins with the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Single-axis disaggregation should be approached as a defined measurement problem. The substantive interest is separate reporting by sex, wealth, residence or another characteristic, observed for a population and period that are stated before calculation.[REF-15] [REF-16]
Comparative interpretation of single-axis disaggregation depends upon it should require examination of definitions, timing, migration, duplication and non-response. A proportionate conclusion must also recognise that the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is levels, gaps and denominators shown for each category. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-17] [REF-19]
Evidence concerning single-axis disaggregation should establish explanation requires evidence on institutions, resources, households and prior conditions. The public account remains incomplete unless it explains how interpretation follows this limitation: one axis not presented as a complete account of marginalisation. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose.[REF-20] [REF-21]
For single-axis disaggregation, the material distinction is between if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. A proportionate conclusion must also recognise that confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include groups whose disadvantage lies on another unmeasured dimension. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete.[REF-23] [REF-24]
Intersecting categories
Review of intersecting categories is credible only where it explains a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The public account remains incomplete unless it explains how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For intersecting categories, the first requirement is conceptual clarity. The study seeks evidence on joint distributions such as sex by wealth and residence; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-17] [REF-19]
Public responsibility for intersecting categories begins with source coverage must be tested before sources are combined. For the learners concerned, the decisive consideration is whether school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use pre-specified combinations with sufficient observations. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone.[REF-20] [REF-21]
intersecting categories requires a decision about percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. For the learners concerned, the decisive consideration is whether the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that empty or unstable cells reported honestly. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-23] [REF-24]
A defensible account of intersecting categories distinguishes where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. This matters because where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for poor rural girls, disabled learners in remote areas and displaced minorities. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness.[REF-01] [REF-07]
Small numbers, disclosure and reliability
small numbers, disclosure and reliability requires a decision about the measure should preserve both the observed level and the distribution relevant to the claim. A contrary reading would overlook that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of useful detail without unreliable estimates or identification begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-20] [REF-21]
For small numbers, disclosure and reliability, the material distinction is between analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. The evidence must therefore clarify how this distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on suppression, aggregation or qualitative evidence chosen proportionately. The definition should be fixed for the comparison at hand and deviations recorded.[REF-23] [REF-24]
The practical standard for small numbers, disclosure and reliability concerns nor should a group estimate be read as a description of every member. This matters because within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: confidentiality decisions separated from claims that no disparity exists. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-01] [REF-07]
For small numbers, disclosure and reliability, the material distinction is between attrition at any stage can produce an apparently complete indicator from a selective population. The public account remains incomplete unless it explains how field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether small communities and learners with rare characteristics are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-02] [REF-03]
Part XIV
Comparison, uncertainty and change
Comparing unlike systems
Evidence concerning comparing unlike systems should establish here, the relevant phenomenon is cross-country patterns based on harmonised but bounded concepts, not the administrative convenience of the available categories. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of comparing unlike systems lies in making unequal educational experience observable.[REF-23] [REF-24]
Institutional action on comparing unlike systems should be tested against at minimum this includes population, geography, date, collection method, classification and known exclusions. A proportionate conclusion must also recognise that a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is metadata tests before numerical comparison. Its metadata should travel with every published value.[REF-01] [REF-07]
The central question in comparing unlike systems is a responsible commentary distinguishes observation, calculation and interpretation. For the learners concerned, the decisive consideration is whether it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that differences in programme structure and classification remain visible.[REF-02] [REF-03]
comparing unlike systems cannot be judged without identifying the public report can still state that a disparity was examined, whether action is required and which body will monitor it. The resulting interpretation should show why for countries with incomplete or rapidly changing systems, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-04] [REF-06]
Sampling error and other uncertainty
The central question in sampling error and other uncertainty is its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. A contrary reading would overlook that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns the range of values reasonably compatible with the observations.[REF-01] [REF-07]
A defensible account of sampling error and other uncertainty distinguishes the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. For the learners concerned, the decisive consideration is whether the preferred construction is standard errors, design effects and data-quality qualifications. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-02] [REF-03]
sampling error and other uncertainty cannot be judged without identifying the comparison should show the level for each group as well as any ratio or gap. The evidence must therefore clarify how a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: rank differences smaller than uncertainty not interpreted.[REF-04] [REF-06]
For sampling error and other uncertainty, the material distinction is between protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The evidence must therefore clarify how the distributional review must deliberately include small disaggregated populations. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-08] [REF-09]
Trends and breaks in series
In assessing trends and breaks in series, authorities must determine a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that institutional action on trends and breaks in series should be tested against the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The institutional consequence follows from whether trends and breaks in series should be approached as a defined measurement problem. The substantive interest is change over time under stable definitions and populations, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-02] [REF-03]
Review of trends and breaks in series is credible only where it explains source coverage must be tested before sources are combined. A proportionate conclusion must also recognise that school returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use parallel reporting where methods or boundaries change. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone.[REF-04] [REF-06]
The practical standard for trends and breaks in series concerns disparity measures should not replace the underlying distributions. A proportionate conclusion must also recognise that a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that revisions and discontinuities marked at the point of change.[REF-08] [REF-09]
In assessing trends and breaks in series, authorities must determine the review should record non-response, unknown status and excluded locations separately. The resulting interpretation should show why combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for groups entering the data only after improved coverage. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-10] [REF-12]
Part XV
Responsible interpretation and action
Reading disparity without blaming learners
Institutional action on reading disparity without blaming learners should be tested against the study seeks evidence on institutional and social conditions associated with unequal outcomes; it does not infer a learner's circumstances from a national or regional mean. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For reading disparity without blaming learners, the first requirement is conceptual clarity.[REF-04] [REF-06]
reading disparity without blaming learners cannot be judged without identifying the definition should be fixed for the comparison at hand and deviations recorded. A contrary reading would overlook that analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on descriptive findings separated from causal claims.[REF-08] [REF-09]
reading disparity without blaming learners cannot be judged without identifying if a conclusion changes under a reasonable specification, that instability is part of the finding. The evidence must therefore clarify how national averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: group identity never treated as a mechanism by itself. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved.[REF-10] [REF-12]
The central question in reading disparity without blaming learners is missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The resulting interpretation should show why an adequate equity account asks whether communities subject to stigma are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-13] [REF-14]
Turning evidence into equitable policy
Comparative interpretation of turning evidence into equitable policy depends upon the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of a decision rule linking disparity to service, finance or legal responsibility begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]
Public responsibility for turning evidence into equitable policy begins with these are substantive attributes because they determine who can appear in the evidence. The institutional consequence follows from whether a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is baseline, intended reach, implementation evidence and review date. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-10] [REF-12]
A defensible account of turning evidence into equitable policy distinguishes apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The resulting interpretation should show why use of the indicator is bounded by the principle that targets accompanied by distributional safeguards. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-13] [REF-14]
Public responsibility for turning evidence into equitable policy begins with this balance is contextual rather than mechanical. A contrary reading would overlook that it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For learners farthest below a secured minimum, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-15] [REF-16]
A public account beyond the average
The practical standard for a public account beyond the average concerns the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of a public account beyond the average lies in making unequal educational experience observable. Here, the relevant phenomenon is a concise national statement of level, distribution, missingness and remedy, not the administrative convenience of the available categories.[REF-10] [REF-12]
Institutional action on a public account beyond the average should be tested against a discrepancy is not resolved by selecting the more favourable source. The institutional consequence follows from whether it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is national totals presented beside selected group and place indicators. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-13] [REF-14]
The practical standard for a public account beyond the average concerns where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. The institutional consequence follows from whether statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: progress claims bounded by evidence coverage and unresolved gaps. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning.[REF-15] [REF-16]
a public account beyond the average cannot be judged without identifying confidentiality is essential, especially where identity or status creates risk. A contrary reading would overlook that protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include every learner otherwise hidden by a successful average. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-17] [REF-19]
Part XVI
Applied distributional analysis
Composite measures and the loss of meaning
In assessing composite measures and the loss of meaning, authorities must determine the resulting interpretation should show why where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The resulting interpretation should show why the analytical purpose is combining several dimensions into a summary measure. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises participation, completion, learning and school conditions, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. In assessing composite measures and the loss of meaning, authorities must determine where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-01] [REF-03]
The central question in composite measures and the loss of meaning is if the headline conclusion changes materially, the range of results should be reported. This matters because sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern weights, normalisation and substitution. In assessing composite measures and the loss of meaning, authorities must determine choices on these matters are not neutral presentation details. A proportionate conclusion must also recognise that they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications.[REF-02] [REF-04]
composite measures and the loss of meaning cannot be judged without identifying direction must therefore be joined to adequacy. This matters because comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. In assessing composite measures and the loss of meaning, authorities must determine no result should be described as equitable solely because one relative measure improved. The resulting interpretation should show why the main interpretive danger is a low value in one dimension being concealed by a high value in another. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens.[REF-05] [REF-06]
composite measures and the loss of meaning requires a decision about surveys can represent households beyond formal education, subject to sample size, field access and response. This matters because censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-08] [REF-12]
Evidence concerning composite measures and the loss of meaning should establish small cells require protection against disclosure and caution about statistical stability. The public account remains incomplete unless it explains how these duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete.[REF-13] [REF-16]
Review of composite measures and the loss of meaning is credible only where it explains the certainty required depends upon the consequence. The public account remains incomplete unless it explains how credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is policy-makers deciding whether a summary aids or obscures resource allocation. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-17] [REF-19]
For composite measures and the loss of meaning, the material distinction is between local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. For the learners concerned, the decisive consideration is whether consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. In assessing composite measures and the loss of meaning, authorities must determine feedback needs a recorded route into classification, service design or further enquiry. The public account remains incomplete unless it explains how people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy.[REF-21] [REF-23]
For composite measures and the loss of meaning, the material distinction is between it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. The evidence must therefore clarify how changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement.[REF-23] [REF-24]
Decomposing an observed education gap
The practical standard for decomposing an observed education gap concerns that purpose should be written before the calculation because method follows the intended inference. A proportionate conclusion must also recognise that the evidence base comprises within-group and between-group differences, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is examining how a national disparity is distributed across places and population groups.[REF-02] [REF-04]
The practical standard for decomposing an observed education gap concerns sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. The resulting interpretation should show why documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern population shares, outcome levels and overlapping membership. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported.[REF-05] [REF-06]
Public responsibility for decomposing an observed education gap begins with no result should be described as equitable solely because one relative measure improved. The institutional consequence follows from whether the main interpretive danger is treating a descriptive decomposition as proof of cause. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity.[REF-08] [REF-12]
The practical standard for decomposing an observed education gap concerns the policy user is authorities locating where further enquiry and action are warranted. The evidence must therefore clarify how evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. decomposing an observed education gap requires a decision about every response should name the expected population reach and the later observation that will test it. A contrary reading would overlook that otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners.[REF-21] [REF-23]
Setting distribution-sensitive targets
The central question in setting distribution-sensitive targets is the national total supplies context, while the distribution tests whether the total is shared. A proportionate conclusion must also recognise that where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is expressing progress as improvement in secured minimums and unjustified gaps. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises the national level, the least-served group and the lower tail of the distribution, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement.[REF-05] [REF-06]
The practical standard for setting distribution-sensitive targets concerns missing values should remain missing unless an explicit estimation method and its effect are shown. A proportionate conclusion must also recognise that the principal methodological questions concern baseline stability, ambition and safeguards against exclusion. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable.[REF-08] [REF-12]
The central question in setting distribution-sensitive targets is comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. For the learners concerned, the decisive consideration is whether no result should be described as equitable solely because one relative measure improved. The main interpretive danger is meeting a mean target while abandoning those farthest behind. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-13] [REF-16]
Comparative interpretation of setting distribution-sensitive targets depends upon the certainty required depends upon the consequence. The evidence must therefore clarify how credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is governments linking national commitments to subnational delivery. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-23] [REF-24]
Linking learners to service geography
Public responsibility for linking learners to service geography begins with where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The resulting interpretation should show why the analytical purpose is relating participation and learning to the location and capacity of education services. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises settlement populations, travel conditions, schools, teachers and programme levels, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-08] [REF-12]
The central question in linking learners to service geography is a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. For the learners concerned, the decisive consideration is whether if the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern geographical scale, boundary effects and facility catchments. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion.[REF-13] [REF-16]
linking learners to service geography cannot be judged without identifying avoiding it requires the underlying counts and distributions to remain visible beside any summary. The evidence must therefore clarify how analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is assuming the nearest mapped institution is accessible or appropriate.[REF-17] [REF-19]
Public responsibility for linking learners to service geography begins with otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. A proportionate conclusion must also recognise that the policy user is planners choosing sites, transport support and teacher deployment. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-01] [REF-03]
Reconciling conflicting sources
The practical standard for reconciling conflicting sources concerns a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. A contrary reading would overlook that the national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is interpreting differences between administrative, survey and census estimates. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises coverage, timing, concepts and reporting incentives, each of which describes a different feature of educational opportunity.[REF-13] [REF-16]
In assessing reconciling conflicting sources, authorities must determine they determine how strongly one component, place or population can influence the conclusion. A proportionate conclusion must also recognise that a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a documented comparison of population and variable definitions. Choices on these matters are not neutral presentation details.[REF-17] [REF-19]
Review of reconciling conflicting sources is credible only where it explains comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. The resulting interpretation should show why no result should be described as equitable solely because one relative measure improved. The main interpretive danger is averaging incompatible estimates into an apparently precise figure. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-21] [REF-23]
The practical standard for reconciling conflicting sources concerns otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. A contrary reading would overlook that the policy user is statistical authorities issuing one bounded account with visible uncertainty. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it.[REF-02] [REF-04]
Monitoring marginalisation during severe disruption
Review of monitoring marginalisation during severe disruption is credible only where it explains where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The public account remains incomplete unless it explains how the analytical purpose is maintaining useful distributional evidence when populations and services move rapidly. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises rapid counts, restored administrative returns and household evidence, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-17] [REF-19]
Review of monitoring marginalisation during severe disruption is credible only where it explains choices on these matters are not neutral presentation details. The resulting interpretation should show why they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern dated estimates, revision practice and minimal essential classifications.[REF-21] [REF-23]
Public responsibility for monitoring marginalisation during severe disruption begins with analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A contrary reading would overlook that a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is using an unstable emergency denominator to assert durable improvement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary.[REF-23] [REF-24]
Evidence concerning monitoring marginalisation during severe disruption should establish the certainty required depends upon the consequence. The public account remains incomplete unless it explains how credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities protecting access while rebuilding regular statistics. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-05] [REF-06]
Communicating uncertainty without losing urgency
Comparative interpretation of communicating uncertainty without losing urgency depends upon a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. A proportionate conclusion must also recognise that the national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is explaining what is known strongly enough to justify action and what remains unresolved. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises point estimates, ranges, quality statements and missing populations, each of which describes a different feature of educational opportunity.[REF-21] [REF-23]
communicating uncertainty without losing urgency requires a decision about missing values should remain missing unless an explicit estimation method and its effect are shown. This matters because the principal methodological questions concern plain institutional language joined to exact metadata. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable.[REF-23] [REF-24]
Review of communicating uncertainty without losing urgency is credible only where it explains comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. A proportionate conclusion must also recognise that no result should be described as equitable solely because one relative measure improved. The main interpretive danger is presenting caution as a reason for inaction or urgency as a reason for overstatement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy.[REF-01] [REF-03]
communicating uncertainty without losing urgency cannot be judged without identifying the certainty required depends upon the consequence. The public account remains incomplete unless it explains how credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is the public, affected communities and responsible decision-makers. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-08] [REF-12]
A national marginalisation profile
The practical standard for a national marginalisation profile concerns where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. This matters because where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is assembling a concise recurring account beyond the national average. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises population, access, progression, learning, conditions, finance and unresolved evidence gaps, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-23] [REF-24]
In assessing a national marginalisation profile, authorities must determine choices on these matters are not neutral presentation details. This matters because they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a stable core with context-specific distributions.[REF-01] [REF-03]
Evidence concerning a national marginalisation profile should establish avoiding it requires the underlying counts and distributions to remain visible beside any summary. The public account remains incomplete unless it explains how analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is creating an encyclopaedia of indicators without decision priority.[REF-02] [REF-04]
Evidence concerning a national marginalisation profile should establish evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The public account remains incomplete unless it explains how the certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is parliament, ministries, local authorities and communities reviewing educational equity.[REF-13] [REF-16]
Part XVII
Extended disparity interpretation
Minimum comparison threshold
Comparative interpretation of minimum comparison threshold depends upon opportunity to learn, participation and exclusions remain material. A contrary reading would overlook that a learning-disparity comparison requires a declared construct, represented population, assessment conditions, scale, uncertainty and distribution. National means should be accompanied by lower-tail, threshold and group evidence.[REF-02] [REF-03]
Within-system interpretation
In assessing within-system interpretation, authorities must determine counts, levels, absolute gaps and ratios answer different questions. For the learners concerned, the decisive consideration is whether missing learners and non-participating schools must remain visible. Within-system gaps should preserve place, group and institutional context without assigning cause from identity.[REF-01] [REF-06]
Between-system interpretation
Public responsibility for between-system interpretation begins with harmonisation does not remove substantive system difference. The evidence must therefore clarify how between-system comparison requires metadata tests for curriculum, age, language, sampling and assessment. Rank differences smaller than uncertainty should not support categorical conclusions.[REF-02] [REF-03]
Part XVIII
Extended analysis of learning disparities
Assessment participation and the represented learning population
A defensible account of assessment participation and the represented learning population distinguishes the target population, eligible population, sampled population and assessed population should be reported separately. For the learners concerned, the decisive consideration is whether learners can be absent because of illness, displacement, school non-attendance, language barriers, disability, conflict, administrative exclusion or ordinary sampling loss. These routes have different meanings. A result calculated only for participating learners may describe their performance accurately while failing to describe the educational system's full learner population. A learning distribution is defined partly by who participates in the assessment.[REF-01] [REF-02] [REF-03]
Comparative interpretation of assessment participation and the represented learning population depends upon reports should describe replacement rules and show any material effect on geography or institutional type. A contrary reading would overlook that participation should therefore accompany every reported mean, proficiency share or percentile. School exclusion and within-school absence need separate observation where the design permits. If institutions outside the frame differ systematically from included schools, a high learner response rate within sampled schools cannot repair the coverage limitation. Similarly, replacement of inaccessible schools can preserve sample size while changing the population represented.
Review of assessment participation and the represented learning population is credible only where it explains where a numerical adjustment is not defensible, the limitation remains a substantive finding. The resulting interpretation should show why the absence of learning evidence for a population requiring education attention should not be interpreted as evidence of no disparity. Non-participation should not be assigned a score. Coding an absent learner as below threshold invents performance evidence; removing every absence without comment creates a different bias. Sensitivity analysis can examine plausible bounds or compare known characteristics of participants and non-participants.
assessment participation and the represented learning population cannot be judged without identifying where it removes an irrelevant barrier, results may belong in the common distribution. The institutional consequence follows from whether the report should explain that judgement rather than treat all adaptations as either incomparable or automatically identical. Accommodation and language arrangements affect participation and result validity. An assessment may permit attendance yet fail to elicit the intended construct if the format, communication or response mode is inaccessible. Exemption practices should be reported by reason and learner group. Where an adapted form changes the construct materially, separate interpretation may be necessary.[REF-05] [REF-10] [REF-12]
assessment participation and the represented learning population cannot be judged without identifying the responsible education body should identify which population remains unrepresented and when better evidence will be available. The public account remains incomplete unless it explains how a national learning claim should be no broader than that coverage. Public interpretation should connect participation to policy. A low response among remote schools may require field and service improvement; absence among out-of-school children may require a population-based study and re-entry action; assessment exclusion may require accessible design. These are not corrections to be made solely through statistical weighting.
Scale, threshold and distribution comparability
Comparative interpretation of scale, threshold and distribution comparability depends upon a numerical score has meaning through the tasks, response model, scoring rules and population for which interpretation has been supported. The resulting interpretation should show why equal numerical differences should not be assumed to represent equal educational differences unless the scale warrants that inference. A threshold adds a substantive judgement about the knowledge or capability learners should demonstrate; it should not be selected merely because it divides the sample conveniently. Comparisons require a clear statement of what the assessment scale represents.[REF-02] [REF-03]
Comparative interpretation of scale, threshold and distribution comparability depends upon reports should therefore consider percentiles, threshold shares and dispersion where technically sound. The public account remains incomplete unless it explains how the lower part of the distribution is especially important for minimum learning opportunity, while the upper part may reveal whether expansion has altered advanced performance. These summaries should remain tied to uncertainty and assessment coverage. Means provide one description of the centre and can conceal change elsewhere. The same mean can accompany a compressed distribution, a wide lower tail or polarisation.
scale, threshold and distribution comparability cannot be judged without identifying a change in the percentage above a threshold may reflect learning, scale revision, task composition or population change. The resulting interpretation should show why if standards are reset, the break should be visible at the year of change and parallel results reported where possible. Descriptions such as basic, adequate or advanced should be treated as definitions under the assessment, not universal attributes of a learner or education system. Threshold comparisons need stable standard-setting and clear labels.
Institutional action on scale, threshold and distribution comparability should be tested against some comparisons may remain defensible at a broad domain while narrower subscales do not. A proportionate conclusion must also recognise that the permissible inference should be stated accordingly. Cross-language and cross-cultural comparability require evidence, not an assumption that translation has preserved difficulty and meaning. Task familiarity, curriculum exposure and response conventions can affect results. Review should examine translation, adaptation, differential item behaviour and opportunity to learn, while avoiding the claim that every detected difference invalidates the whole assessment.
Review of scale, threshold and distribution comparability is credible only where it explains public reporting should avoid categorical language when intervals overlap or scale linkage is weak. A contrary reading would overlook that the useful question is whether the evidence indicates a material disparity requiring enquiry, what population is affected and which educational condition might be changed. League order does not answer those questions. A rank is usually less informative than the estimated difference, uncertainty and distribution. Small rank movement can follow changes in participating systems or sampling variation.
Opportunity to learn and interpretation of achievement gaps
Comparative interpretation of opportunity to learn and interpretation of achievement gaps depends upon opportunity does not determine performance completely, and its measurement is imperfect. The institutional consequence follows from whether it nevertheless prevents a learning gap from being attributed solely to learners or households when institutions supplied different educational conditions. Achievement evidence should be interpreted with opportunity to learn. This includes curriculum entitlement, content actually taught, instructional time, teacher availability, language, materials, attendance and access to support.[REF-01] [REF-13] [REF-14]
Institutional action on opportunity to learn and interpretation of achievement gaps should be tested against teacher self-report may be influenced by recall or expectations; observation covers a short period; work samples are selective. The resulting interpretation should show why agreement across sources strengthens interpretation when dates and populations align. Contradictions can identify local variation or weak measurement and should not be resolved by choosing the account most favourable to the system. The official curriculum establishes intended opportunity, not delivered instruction. Teacher reports, schedules, classroom observation and learner work can add evidence, each with limitations.
Institutional action on opportunity to learn and interpretation of achievement gaps should be tested against time is therefore one component of the explanatory evidence, not a conversion factor for predicted score. The resulting interpretation should show why instructional time should distinguish scheduled, delivered and attended time. Closures, teacher absence, shortened shifts and late entry can reduce delivered or usable time. Total hours also conceal subject allocation and teaching quality. A learner receiving more hours of poorly organised instruction does not necessarily have greater opportunity in the intended domain.
Evidence concerning opportunity to learn and interpretation of achievement gaps should establish these conditions should be examined at the level where the learning evidence was collected. The evidence must therefore clarify how national resource averages can obscure concentration of weak provision among the same learners whose scores form the lower tail. Resource indicators should be connected to use. Textbooks delivered to a school may not be available in the relevant language, grade or classroom. Teacher qualifications recorded administratively may not match subject assignment. Facilities may exist but be inaccessible or unsafe.
Public responsibility for opportunity to learn and interpretation of achievement gaps begins with where learners received less curriculum exposure, a response may include additional teaching, staff deployment, accessible materials or revised pacing. This matters because later assessment should test whether opportunity and learning changed. The original disparity and the remedial conditions should remain visible rather than being erased by a new cohort average. Policy conclusions should avoid treating opportunity indicators as excuses for low expectations. Their purpose is to identify conditions within public responsibility and to design support.
Decomposing disparities within education systems
Comparative interpretation of decomposing disparities within education systems depends upon decomposition can describe where variation is concentrated, provided it is not treated as proof of cause. The public account remains incomplete unless it explains how a large between-school component may indicate segregation, resource distribution or residential pattern; it does not identify which mechanism operates. A large within-school component can coexist with institutional inequality and should not be read as evidence that schools are irrelevant. A national disparity can reflect differences between regions, schools, classrooms and learners.[REF-01] [REF-09] [REF-14]
Institutional action on decomposing disparities within education systems should be tested against learners are commonly nested within classes and schools, while policies may operate through districts or providers. The resulting interpretation should show why standard errors and models should respect clustering. Small schools and sparsely populated areas may require special treatment, but removal changes the population represented. Reports should state exclusions and avoid presenting a modelled residual as direct observation. The level of analysis should correspond to the sampling and decision structure.
For decomposing disparities within education systems, the material distinction is between a raw school mean combines prior opportunity, intake, mobility, attendance and current teaching. The evidence must therefore clarify how adjusted measures can answer bounded questions but depend on variables and assumptions. They should not replace the unadjusted learner outcome or become a definitive quality rank. Adjustment for a condition influenced by the school can also remove part of the very effect under review. The purpose and causal assumptions need explanation. Group composition affects school comparisons.
In assessing decomposing disparities within education systems, authorities must determine policy priority should consider educational severity, population, rights and feasibility rather than statistical contribution alone. This matters because maps and rankings, where used elsewhere, should not expose small communities or imply a boundary creates the disparity. Geographical decomposition should preserve absolute numbers and service context. A small district with a severe gap may need urgent support even though it contributes little to national variance. A populous area with a modest gap may represent many learners.
decomposing disparities within education systems requires a decision about follow-up should state the condition examined, action taken and later learning evidence. The institutional consequence follows from whether the appropriate outcome is a decision agenda. Between-region evidence may lead to allocation review; between-school evidence may lead to staffing, admissions or support enquiry; within-school evidence may lead to classroom, language or accessibility review. Each hypothesis requires additional evidence. The decomposition locates questions; it does not authorise blame.
Bounded conclusions between education systems
The practical standard for bounded conclusions between education systems concerns where a material difference cannot be reconciled, the systems may still be described separately without a common rank. The evidence must therefore clarify how between-system comparison serves public learning when it identifies patterns, plausible questions and alternative institutional arrangements. It becomes misleading when harmonised labels conceal different programme structures, ages, curricula, languages, participation or assessment conditions. Metadata review should precede numerical comparison.[REF-02] [REF-03] [REF-16]
Institutional action on bounded conclusions between education systems should be tested against an apparent national advantage may not extend to poor, rural, minority-language or disabled learners. This matters because reports should place distributional evidence beside the system result and avoid using nationality as an explanation. Country and system averages should not be interpreted as attributes of every school or learner. Within-system distributions can overlap substantially even where means differ. Group composition and population coverage matter.
In assessing bounded conclusions between education systems, authorities must determine a later higher score should not automatically be described as system improvement where the represented population changed materially. A proportionate conclusion must also recognise that temporal comparison requires stable linkage. Changes in curriculum, assessment mode, participation, sampling frame or system boundaries can create discontinuity. A linked scale can support trend if common items or other methods preserve meaning and security, but linkage error should accompany the estimate.
bounded conclusions between education systems cannot be judged without identifying a practice associated with high performance elsewhere may depend upon teacher preparation, finance, curriculum coherence or social conditions absent in the receiving system. The evidence must therefore clarify how the comparison can identify an option, not guarantee its effect. Pilots should state the mechanism and review evidence. Adoption based solely on rank proximity or reputation substitutes imitation for analysis. Policy borrowing should attend to authority, capacity, sequence and context.
The practical standard for bounded conclusions between education systems concerns this separation permits strong public concern about a learning disparity without false certainty about its source and protects education systems from both complacency and unsupported prescription. For the learners concerned, the decisive consideration is whether the final international statement should distinguish three levels. Recorded fact describes the observed results under stated methods. Interpretation identifies patterns and limitations. Policy consideration proposes further enquiry or action under national authority. Causal judgement requires additional design.
Uncertainty, materiality and the duty to respond
Review of uncertainty, materiality and the duty to respond is credible only where it explains sampling error, non-response, scale linkage, classification and model choice each affect the range of defensible conclusions. For the learners concerned, the decisive consideration is whether they should be described separately because they support different remedies. A larger sample may reduce sampling error; it does not correct systematic exclusion or an invalid construct. More decimal places cannot repair weak coverage. The public report should identify which uncertainty could alter the decision and which does not affect the direction of urgent protection. Uncertainty should qualify a disparity claim without neutralising it.[REF-02] [REF-03]
The practical standard for uncertainty, materiality and the duty to respond concerns additional diagnostic support can be introduced and reviewed more readily than a high-stakes classification of schools or learners. A proportionate conclusion must also recognise that statistical significance should not substitute for educational materiality. A small estimated difference may be precise but have limited practical consequence, while an uncertain large gap affecting a protected minimum may require immediate enquiry. Materiality should consider the knowledge or capability involved, the number of learners, distribution, duration and consequences for later progression. The threshold for action also depends upon reversibility.
Comparative interpretation of uncertainty, materiality and the duty to respond depends upon where opportunity to learn differs, address time, teachers, curriculum, language, accessibility or materials. The evidence must therefore clarify how where scale comparability is weak, improve the assessment before publishing ranks. Where a group disparity persists under several definitions, investigate institutional mechanisms without attributing cause to identity. Every response should name the competent body, intended population, resources and review date. A remedy should match the evidence. Where participation is selective, improve coverage and examine barriers.
Institutional action on uncertainty, materiality and the duty to respond should be tested against reports should preserve the original estimate, method and limitation, then state why revision occurred. The evidence must therefore clarify how silent replacement makes apparent improvement impossible to distinguish from correction. A national evidence system demonstrates strength when it can acknowledge uncertainty, improve measurement and amend policy. The final test is whether learners receive better educational opportunity and whether remaining disparities continue to be visible rather than whether one annual figure becomes more favourable. Later evidence should be capable of changing the conclusion.
Part XIX
Targeted improvement planning
Defining a persistent learning gap
Institutional action on defining a persistent learning gap should be tested against if the assessment population excludes learners most at risk, the gap is not adequately defined. For the learners concerned, the decisive consideration is whether the plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns a sustained disparity in a declared learning domain, population and period. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity. The baseline should preserve the underlying distribution and absolute numbers.[REF-01] [REF-02]
defining a persistent learning gap requires a decision about achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. The evidence must therefore clarify how authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is a fluctuating score or one cohort difference being treated as proof of persistence. Avoiding it requires evidence on both learning and the opportunity supplied.[REF-03] [REF-06]
Public responsibility for defining a persistent learning gap begins with observation and work samples can explain classroom conditions without estimating national prevalence. The resulting interpretation should show why learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes baseline, assessment coverage, uncertainty, distribution and opportunity to learn. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity.[REF-08] [REF-09]
Public responsibility for defining a persistent learning gap begins with local adaptation should remain possible within a common substantive condition. The public account remains incomplete unless it explains how any departure should be recorded with its reason and expected learner consequence. The governing action is to select a gap that is educationally material and within public influence. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism.[REF-13] [REF-14]
In assessing defining a persistent learning gap, authorities must determine gender, household resources, disability, language, residence and prior opportunity may intersect. The resulting interpretation should show why a plan can improve its average by reaching learners closest to a threshold while leaving those farthest behind. Targets should therefore include the least-served position and protection against exclusion. Small groups require confidentiality and careful precision, not disappearance from review. Distributional review should ask who is eligible, offered support, participates, receives the intended intensity and demonstrates a later response.[REF-17] [REF-19]
Institutional action on defining a persistent learning gap should be tested against teachers need subject knowledge, worked examples, diagnostic interpretation and protected time to collaborate. The public account remains incomplete unless it explains how moderation should examine evidence and reasoning rather than force identical decisions for unlike cases. Staff workload and turnover should be monitored. A plan relying on a few exceptional individuals is not institutionally secure. Leadership should route obstacles to bodies able to change staffing, finance, curriculum or assessment rather than leaving every correction to the classroom. Professional capability is central.[REF-23] [REF-24]
defining a persistent learning gap cannot be judged without identifying where the intervention is ineffective, adaptation or cessation is responsible improvement. A contrary reading would overlook that where benefit depends on temporary support, institutionalisation requires recurrent finance and ordinary ownership. Completion is demonstrated by stronger learning opportunity and a functioning correction route, not by the end of a project. The public account should distinguish authorisation, delivery, use and learning consequence. It should report limitations, adverse effects and unresolved learners. A favourable later result may reflect population or assessment change and should be tested against the baseline metadata.[REF-01] [REF-02]
From diagnosis to an intervention hypothesis
Comparative interpretation of from diagnosis to an intervention hypothesis depends upon a national average is insufficient when the plan addresses a local or group disparity. For the learners concerned, the decisive consideration is whether the baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns an explicit account of the condition expected to change learning. It should be stated with the learner population, educational domain, geography and evidence period.[REF-03] [REF-06]
Public responsibility for from diagnosis to an intervention hypothesis begins with authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. The resulting interpretation should show why alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is group identity or low performance being mistaken for a causal explanation. Avoiding it requires evidence on both learning and the opportunity supplied. Achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms.[REF-08] [REF-09]
Comparative interpretation of from diagnosis to an intervention hypothesis depends upon administrative records can describe staffing and participation while omitting non-enrolled learners. For the learners concerned, the decisive consideration is whether assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes curriculum exposure, teaching practice, language, time, materials and support. Each source should be used within scope.[REF-13] [REF-14]
from diagnosis to an intervention hypothesis cannot be judged without identifying dependencies should be sequenced. For the learners concerned, the decisive consideration is whether teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to state the mechanism and plausible alternatives before choosing activity. Implementation should identify who acts, with what authority, resources and deadline.[REF-17] [REF-19]
Designing the targeted plan
A defensible account of designing the targeted plan distinguishes a national average is insufficient when the plan addresses a local or group disparity. A contrary reading would overlook that the baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns a bounded sequence connecting resources and actions to learner-facing change. It should be stated with the learner population, educational domain, geography and evidence period.[REF-08] [REF-09]
Public responsibility for designing the targeted plan begins with authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. The evidence must therefore clarify how alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is a list of activities replacing a coherent implementation logic. Avoiding it requires evidence on both learning and the opportunity supplied. Achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms.[REF-13] [REF-14]
Comparative interpretation of designing the targeted plan depends upon administrative records can describe staffing and participation while omitting non-enrolled learners. This matters because assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes responsibility, staff capability, learner support, milestones and correction. Each source should be used within scope.[REF-17] [REF-19]
Institutional action on designing the targeted plan should be tested against implementation should identify who acts, with what authority, resources and deadline. A contrary reading would overlook that dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to choose a feasible intensity and protect the common educational entitlement.[REF-23] [REF-24]
Resourcing equitable implementation
Review of resourcing equitable implementation is credible only where it explains a national average is insufficient when the plan addresses a local or group disparity. For the learners concerned, the decisive consideration is whether the baseline should preserve the underlying distribution and absolute numbers. If the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns the staff, time, materials, accessibility and finance required by the selected mechanism. It should be stated with the learner population, educational domain, geography and evidence period.[REF-13] [REF-14]
The central question in resourcing equitable implementation is avoiding it requires evidence on both learning and the opportunity supplied. The evidence must therefore clarify how achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. Authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is weak institutions being expected to implement with the same nominal allocation.[REF-17] [REF-19]
For resourcing equitable implementation, the material distinction is between administrative records can describe staffing and participation while omitting non-enrolled learners. This matters because assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes recurrent cost, teacher workload, additional need and household burden. Each source should be used within scope.[REF-23] [REF-24]
The practical standard for resourcing equitable implementation concerns dependencies should be sequenced. The public account remains incomplete unless it explains how teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. Local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to direct greater support where barriers and implementation costs are greater. Implementation should identify who acts, with what authority, resources and deadline.[REF-01] [REF-02]
Monitoring reach, quality and learning response
For monitoring reach, quality and learning response, the material distinction is between if the assessment population excludes learners most at risk, the gap is not adequately defined. For the learners concerned, the decisive consideration is whether the plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns evidence that the intended learners received the intervention as designed and benefited educationally. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity. The baseline should preserve the underlying distribution and absolute numbers.[REF-17] [REF-19]
For monitoring reach, quality and learning response, the material distinction is between achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. This matters because authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is participation counts being treated as learning evidence. Avoiding it requires evidence on both learning and the opportunity supplied.[REF-23] [REF-24]
Institutional action on monitoring reach, quality and learning response should be tested against learner and teacher accounts can identify barriers. The public account remains incomplete unless it explains how agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes eligibility, offer, take-up, dosage, teaching quality, work and assessment. Each source should be used within scope. Administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence.[REF-01] [REF-02]
In assessing monitoring reach, quality and learning response, authorities must determine teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism. The resulting interpretation should show why local adaptation should remain possible within a common substantive condition. Any departure should be recorded with its reason and expected learner consequence. The governing action is to combine timely implementation evidence with valid learning review. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced.[REF-03] [REF-06]
Adaptation, institutionalisation and exit
Comparative interpretation of adaptation, institutionalisation and exit depends upon the baseline should preserve the underlying distribution and absolute numbers. For the learners concerned, the decisive consideration is whether if the assessment population excludes learners most at risk, the gap is not adequately defined. The plan should state which observation is direct, which is estimated and which remains unknown. The improvement question concerns reasoned decisions to continue, change, scale or end the plan. It should be stated with the learner population, educational domain, geography and evidence period. A national average is insufficient when the plan addresses a local or group disparity.[REF-23] [REF-24]
Public responsibility for adaptation, institutionalisation and exit begins with achievement differences may be associated with poverty, language, disability or residence, but those characteristics are not instructional mechanisms. The institutional consequence follows from whether authorities should examine curriculum, teacher availability, time, attendance, materials, assessment access and learner support. Alternative explanations should remain open until evidence discriminates among them. This discipline prevents a targeted plan from attaching deficit to learners instead of changing institutions. The principal error is temporary measures persisting without benefit or disappearing before durable capability exists. Avoiding it requires evidence on both learning and the opportunity supplied.[REF-01] [REF-02]
Institutional action on adaptation, institutionalisation and exit should be tested against each source should be used within scope. The resulting interpretation should show why administrative records can describe staffing and participation while omitting non-enrolled learners. Assessment provides bounded learning evidence subject to coverage and validity. Observation and work samples can explain classroom conditions without estimating national prevalence. Learner and teacher accounts can identify barriers. Agreement adds confidence after dates and definitions align; contradiction should guide further enquiry rather than selective reporting. The required evidence includes thresholds, adverse effects, unresolved cases, recurrent ownership and later evidence.[REF-03] [REF-06]
Review of adaptation, institutionalisation and exit is credible only where it explains local adaptation should remain possible within a common substantive condition. The public account remains incomplete unless it explains how any departure should be recorded with its reason and expected learner consequence. The governing action is to retain useful capability while ending ineffective or inequitable arrangements. Implementation should identify who acts, with what authority, resources and deadline. Dependencies should be sequenced. Teacher guidance without planning time, materials without accessible use, or tutoring without safe attendance cannot deliver the expected mechanism.[REF-08] [REF-09]
Part XX
National translation after adoption of the 2030 Agenda
Fixing the contemporaneous institutional baseline
Institutional planning should begin from the position established by 30 March 2022. The Incheon Declaration expressed the education community's commitment to inclusive and equitable quality education and lifelong learning; Addis established a wider financing framework; the General Assembly adopted the 2030 Agenda; and the Education 2030 Framework for Action supplied an implementation reference available by the cut-off. These texts provide direction without proving that national or institutional delivery has occurred.[REF-25] [REF-26] [REF-27] [REF-28]
A defensible account of fixing the contemporaneous institutional baseline distinguishes it does not prove that national law, plans, budgets, information or delivery arrangements already satisfy the commitment. This matters because each country should identify which obligations and targets can be acted upon under existing authority, which require legislative or administrative change, and which depend on clarification through later competent decisions. Unsettled detail should be recorded as such rather than filled by anticipation. The distinction between adoption and implementation is essential. Adoption establishes an agreed direction and permits governments to begin alignment.[REF-07] [REF-12] [REF-27]
Public responsibility for fixing the contemporaneous institutional baseline begins with national authorities should map the newly adopted targets against those instruments and against the education stages, populations and institutions already in law and plans. For the learners concerned, the decisive consideration is whether this avoids two risks: abandoning useful evidence because terminology changed, and claiming continuity where a new target has broader scope or different educational substance. The baseline should preserve existing education commitments and national evidence. The new agenda does not erase the right to education, the unfinished Education for All undertaking, programme structures or established statistical series.[REF-08] [REF-09] [REF-25]
Evidence concerning fixing the contemporaneous institutional baseline should establish it prevents a proposed measure from being represented as an obligation, a declaration from being treated as proof of delivery, and a later decision from being projected backwards. For the learners concerned, the decisive consideration is whether the register should be revised openly as competent bodies act. A dated commitment register can support institutional accuracy. For each relevant proposition, it should record the adopting body, date, legal or policy status, national authority, present implementing instrument and unresolved question. This is not an administrative inventory for its own sake.[REF-17] [REF-19] [REF-27]
fixing the contemporaneous institutional baseline cannot be judged without identifying governments can state that the 2030 Agenda has been adopted and that national alignment is commencing. For the learners concerned, the decisive consideration is whether they should not state that every indicator, national milestone or implementation mechanism has already been internationally settled. A precise account strengthens credibility because it makes clear which choices belong to national democratic and administrative processes and which follow directly from adopted commitments. It also establishes a reliable point from which later implementation can be judged. Public communication should use the same discipline.[REF-25] [REF-26] [REF-27]
Selecting priorities without narrowing the commitment
For selecting priorities without narrowing the commitment, the material distinction is between a national plan cannot improve every condition simultaneously, yet it should retain a complete map of early childhood, primary and secondary education, technical and vocational learning, tertiary participation, adult learning, relevant skills, equality, literacy, learning environments, scholarships and teachers as they appear in the adopted targets. The institutional consequence follows from whether a first improvement priority should be chosen because evidence shows a serious and remediable break, not because other elements have ceased to matter. The breadth of Goal 4 requires sequencing, not selective abandonment.[REF-25] [REF-27]
Institutional action on selecting priorities without narrowing the commitment should be tested against the equity test asks whether the measure will reach learners farthest from the secured opportunity. For the learners concerned, the decisive consideration is whether a priority that scores highly on visibility but weakly on consequence or equity should be reconsidered. The reasons for selection should be published beside the evidence and known limitations. Priority selection should apply four tests. The entitlement test asks which population and educational condition are at stake. The consequence test asks the scale and severity of the denial. The actionability test asks whether a competent authority has a plausible means of change.[REF-01] [REF-03] [REF-12]
The central question in selecting priorities without narrowing the commitment is it should also distinguish poor data from satisfactory conditions. This matters because where the least-served population is weakly observed, strengthening coverage can itself become an immediate priority while urgent service evidence supports proportionate protection. National averages should not determine the sequence alone. A moderate national gap may conceal acute failure in one district or population; a large aggregate shortfall may require broad system expansion alongside targeted support. Analysis should retain the national level, absolute number affected, subnational and group distributions and the minimum educational floor.[REF-02] [REF-06] [REF-24]
The central question in selecting priorities without narrowing the commitment is the plan should make these dependencies visible and decide which must precede or accompany the selected measure. The resulting interpretation should show why otherwise a high-level commitment can be converted into an isolated activity unable to change the learner-facing condition. Dependencies should influence sequence. An assessment reform cannot improve learning without curriculum alignment, teacher capability and participation. Expansion of secondary places may depend on primary completion, trained staff and facilities. Adult learning may require flexible provision, recognition and learner support.[REF-03] [REF-16] [REF-18]
Public responsibility for selecting priorities without narrowing the commitment begins with revision is not a retreat from ambition when reasons and consequences are published. The public account remains incomplete unless it explains how it is a condition of responsible improvement. The complete commitment map should remain in view so that repeated concentration on one readily measured target does not produce silent neglect of lifelong learning, equality, educational quality or populations outside formal schooling. Selection should remain revisable. New evidence may show that the diagnosed mechanism was wrong, that another population is more severely affected or that implementation capacity is insufficient.[REF-08] [REF-25] [REF-27]
From global target to national improvement proposition
from global target to national improvement proposition cannot be judged without identifying for example, a commitment to equitable quality education is too broad to guide one implementation decision; a proposition to improve regular attendance for a defined remote population through transport, staffing and calendar changes can be examined and corrected. This matters because the narrower proposition remains connected to the universal commitment and should not be mistaken for its completion. A national improvement proposition should translate a broad target into a bounded statement of change. It should name the population, present educational condition, institutional mechanism, responsible body, resources, time and evidence of success.[REF-07] [REF-12] [REF-27]
The practical standard for from global target to national improvement proposition concerns consultation with teachers, learners and communities can identify mechanisms, but it should not replace representative population evidence when prevalence is claimed. A proportionate conclusion must also recognise that diagnosis should precede instrument choice. A low completion rate may reflect late entry, repetition, household cost, distance, school safety, language, disability exclusion, teacher shortage or unreliable records. These mechanisms require different responses. A government should compare administrative, household, assessment and local service evidence and state where inference remains uncertain.[REF-01] [REF-03] [REF-24]
A defensible account of from global target to national improvement proposition distinguishes additional materials will not improve learning if teachers lack time or knowledge to use them; professional guidance will not improve attendance where transport is decisive; a new indicator will not correct exclusion without authority and resources. The resulting interpretation should show why the plan should identify necessary dependencies and foreseeable adverse effects. It should specify which part of the hypothesis is established by evidence and which remains to be tested during implementation. The intervention hypothesis should explain how the proposed measure changes the barrier.[REF-06] [REF-12] [REF-16]
Comparative interpretation of from global target to national improvement proposition depends upon it is not legitimate where it narrows the entitled population, lowers expectations for disadvantaged groups or converts learning into attendance alone. A proportionate conclusion must also recognise that the public record should show the relationship between the global target, national definition and selected measure, including any material difference from an existing series. National adaptation should preserve educational substance. Targets may need national definitions for programme levels, age groups, language and institutional responsibility. Adaptation is legitimate where it makes the commitment operational and comparable over time.[REF-02] [REF-10] [REF-25]
from global target to national improvement proposition requires a decision about a pilot that continues because it attracts support rather than because it changes the intended condition is not an improvement method. The public account remains incomplete unless it explains how conversely, an intervention should not be abandoned merely because early outcomes are uncertain where delivery has not reached intended intensity. Review must distinguish theory failure, implementation failure, measurement weakness and insufficient time. The proposition should end with a decision rule. Authorities should state what evidence would justify continuation, expansion, adaptation or cessation and when that decision will occur.[REF-08] [REF-11] [REF-23]
Monitoring delivery, reach and educational consequence
In assessing monitoring delivery, reach and educational consequence, authorities must determine these stages should not be collapsed. A proportionate conclusion must also recognise that a budget can be executed without materials arriving, a programme can operate without reaching disadvantaged learners, and participation can rise without improvement in learning or progression. Monitoring should follow the causal sequence of the improvement proposition. Inputs show whether resources and staff were available; delivery evidence shows whether the measure operated; reach shows which eligible learners participated and with what intensity; educational evidence shows whether the intended condition changed.[REF-03] [REF-06] [REF-12]
Review of monitoring delivery, reach and educational consequence is credible only where it explains apparent improvement caused by revised population estimates, wider institutional reporting or changed assessment participation should be separated from educational change. The evidence must therefore clarify how comparable trends are valuable, but continuity should not be asserted where concepts differ materially. The baseline should retain numerator, denominator, population, date, geography, definition and exclusions. Where a new target requires a new measure, authorities should preserve the preceding series and identify any break.[REF-02] [REF-23] [REF-24]
monitoring delivery, reach and educational consequence cannot be judged without identifying disaggregation must remain lawful, meaningful and safe. A proportionate conclusion must also recognise that small numbers may require controlled access or combined reporting, but the affected population and public responsibility should not disappear. Equity monitoring should show eligibility, offer, take-up, attendance, completion and outcome for relevant groups and places. A plan can improve its average by reaching learners already closest to the desired condition. It should therefore report the least-served position and protect against exclusion or deterioration.[REF-01] [REF-10] [REF-13]
The central question in monitoring delivery, reach and educational consequence is classroom observation, work samples and learner accounts can illuminate mechanism without establishing national prevalence. The evidence must therefore clarify how agreement across sources strengthens a conclusion only after definitions and dates align; disagreement should guide investigation rather than selective reporting. Learning evidence requires population coverage and opportunity to learn. Assessment results should identify the domain, eligible population, participation and exclusions. A favourable mean among tested learners cannot represent those outside school or absent from assessment.[REF-03] [REF-24]
Review of monitoring delivery, reach and educational consequence is credible only where it explains a proportionate conclusion must also recognise that the next review date and responsible authority should be visible. This matters because this makes monitoring a means of correction rather than an obligation to produce favourable figures. The adopted agenda gains national credibility when evidence can alter action and expose populations who remain underserved. Public reporting should connect the result to a decision. A defensible account of monitoring delivery, reach and educational consequence distinguishes a concise account should state what was implemented, who received it, what changed, what remains uncertain, which adverse effects occurred and whether the measure will continue, adapt, expand or cease.[REF-08] [REF-25] [REF-27]
First national decisions after adoption
For first national decisions after adoption, the material distinction is between governments can designate a competent coordinating authority, preserve existing sector responsibilities, assemble a dated commitment register and commission a baseline review. A contrary reading would overlook that the review should map every adopted education target against national law, plans, budgets and evidence. It should identify urgent gaps and unsettled definitions separately. This creates a disciplined bridge between global adoption and national action. The immediate national decision is to establish governance for alignment without pretending that implementation detail is complete.[REF-25] [REF-27]
In assessing first national decisions after adoption, authorities must determine it should name the population and condition, state why it takes precedence, and preserve the wider commitment map. The public account remains incomplete unless it explains how where credible evidence reveals severe exclusion or harm, protective action need not await a perfect estimate. Longer-term allocation, however, should be reviewed as coverage improves. The plan should guard against choosing only learners and institutions most likely to produce rapid favourable results. A first priority should be small enough for accountable action and important enough to change educational opportunity.[REF-01] [REF-06] [REF-12]
The central question in first national decisions after adoption is the national authority should state the source, timing and distribution of finance and how external cooperation relates to ordinary provision. A proportionate conclusion must also recognise that a financing gap should be described rather than hidden through reduced educational substance or transfer of cost to poor households. The first budget decision should identify recurrent implications. Temporary finance may permit testing, but teachers, learner support, accessible facilities, maintenance and evidence cannot be sustained by an announcement.[REF-15] [REF-26]
A defensible account of first national decisions after adoption distinguishes it can state that the principal education, financing and implementation texts named in this report were available by the cut-off. A contrary reading would overlook that it should also state that later indicator settlements, technical revisions and results are outside the record. This protects the difference between an adopted commitment, a national policy choice, an institutional plan and later statistical clarification. Accurate status is a condition of accountable planning, not a reason for delay. The first public report should be candid about chronology.[REF-25] [REF-26] [REF-27] [REF-28]
A defensible account of first national decisions after adoption distinguishes it should examine authority, delivery, reach, professional capability, finance, learner experience and early educational consequence. The public account remains incomplete unless it explains how it should record adverse findings and adapt the measure where the hypothesis or delivery proves weak. The transition from global commitment to national improvement is complete only when ordinary institutions can sustain the changed condition and correct foreseeable departures; the end of a project or reporting period is not evidence of that result. The first review should test whether institutions learned, not merely whether a plan was issued.[REF-08] [REF-12] [REF-23]
Part XXI
Institutional improvement plans aligned with Education 2030
Institutional mandate and scope
Institutional action on institutional mandate and scope should be tested against the plan should therefore map each selected Education 2030 priority to the institutional function that can materially influence it and identify dependencies requiring action at another level. The public account remains incomplete unless it explains how an institutional improvement plan should begin by identifying the authority under which the institution acts and the population for whom it is responsible. A school, training provider, university, local administration or other body does not implement the whole global agenda alone. It contributes within its mandate, while national authorities retain responsibilities for law, finance, curriculum, workforce and equitable system provision.[REF-12] [REF-25] [REF-28]
Institutional action on institutional mandate and scope should be tested against a global target supplies direction, but the institutional proposition must be narrow enough for authority, resources and evidence to be assigned. The public account remains incomplete unless it explains how scope should be expressed as a learner-facing condition rather than a broad aspiration. An institution may seek more regular participation, stronger foundational learning, safer and more accessible facilities, better transition, improved teacher support or a more equitable adult-learning offer. The statement should name the eligible population, current condition, intended change and period.[REF-07] [REF-27] [REF-28]
Comparative interpretation of institutional mandate and scope depends upon safeguards should state which minimums must be protected during implementation. The public account remains incomplete unless it explains how institutional plans should preserve the relationship among access, equity and quality. A plan concerned with learning cannot ignore assessment participation or opportunity to learn; a plan concerned with enrolment cannot treat registration as regular attendance; a plan concerned with completion should examine the educational substance and recognised value of the programme. Selecting one priority does not authorize deterioration in another essential condition.[REF-01] [REF-03] [REF-09]
The practical standard for institutional mandate and scope concerns this distinction allows institutions to respond to context without treating national variation or resource constraint as authority to lower the substantive entitlement of disadvantaged learners. A contrary reading would overlook that any requested exception should identify the legal basis, expected consequence and competent approving body. The plan should distinguish obligations from discretionary methods. Rights, national requirements and adopted policy establish conditions that the institution must respect. The choice of timetable, professional-learning arrangement, local partnership or diagnostic method may permit adaptation.[REF-08] [REF-10] [REF-28]
The central question in institutional mandate and scope is where the institution lacks authority over a material dependency, the plan should route the issue to the responsible authority and retain it as an unresolved risk rather than quietly narrow the objective. A proportionate conclusion must also recognise that governance should identify one accountable owner and the roles of staff, learners, communities and partner bodies. Participation can reveal barriers and test whether proposed measures are workable; it should not transfer public responsibility to those affected by weak provision. The decision record should show how evidence and consultation altered the priority.[REF-17] [REF-19] [REF-25]
Diagnostic review and priority selection
In assessing diagnostic review and priority selection, authorities must determine the purpose is to locate the first material break rather than compile every available statistic. The public account remains incomplete unless it explains how administrative records describe registered learners and delivery; household or community evidence may reveal people outside the institution; assessment and work samples describe bounded aspects of learning. Diagnostic review should begin with the institution's eligible population and educational course. Entry, attendance, progression, completion, learning and transition should be mapped together with staffing, instructional time, facilities, accessibility and learner support.[REF-02] [REF-03] [REF-24]
Institutional action on diagnostic review and priority selection should be tested against where population estimates are weak, qualitative and case evidence may justify immediate correction without being represented as a prevalence estimate. The institutional consequence follows from whether the review should compare level, distribution and trend. A stable institutional mean can conceal deterioration among one programme or learner group. Sex, household constraint, disability, language, residence, displacement and prior opportunity may be relevant, subject to lawful collection and protection. Small groups require caution about precision and disclosure, but the service barrier should remain visible.[REF-06] [REF-10] [REF-13]
For diagnostic review and priority selection, the material distinction is between low attendance may be associated with transport, cost, safety, calendar, health, discrimination or inaccessible instruction. This matters because weak learning may reflect interrupted participation, limited curriculum exposure, teacher absence, language, assessment design or inadequate support. A descriptive difference cannot rank these explanations. The institution should gather the smallest additional evidence needed for a practical decision and should identify explanations requiring action outside its authority. Diagnosis should distinguish symptom from mechanism.[REF-01] [REF-05] [REF-12]
In assessing diagnostic review and priority selection, authorities must determine ease of measurement alone is not a sufficient selection criterion. The resulting interpretation should show why priority selection should consider severity, scale, equity, feasibility and dependency. A condition affecting fewer learners may warrant immediate action if the consequence is severe or violates a minimum entitlement. Conversely, a large but weakly measured difference may first require stronger evidence. The plan should explain why the selected priority takes precedence, which learners are expected to benefit and how the rest of the institutional responsibility remains monitored.[REF-07] [REF-16] [REF-28]
Comparative interpretation of diagnostic review and priority selection depends upon where the institution changes its record system or assessment during implementation, overlap evidence should be retained and breaks marked. For the learners concerned, the decisive consideration is whether a baseline reconstructed after results are known is vulnerable to selective interpretation. The plan should approve the baseline and revision rule before judging change. The baseline should be preserved at the point of decision. Population, definitions, source completeness, date, uncertainty and any exclusions should travel with the starting value.[REF-03] [REF-23] [REF-24]
Improvement proposition and implementation design
In assessing improvement proposition and implementation design, authorities must determine it should connect an action to an institutional mechanism: additional instructional support to a documented curriculum gap; revised attendance arrangements to a known timing or transport barrier; accessible materials and assessment to exclusion of learners with disabilities; professional support to weak subject teaching. The resulting interpretation should show why activities should not be listed without the reason they are expected to alter learner opportunity. The improvement proposition should state how a defined measure is expected to change the diagnosed condition.[REF-10] [REF-12] [REF-28]
Public responsibility for improvement proposition and implementation design begins with learner support may require eligibility rules, communication and protection of personal information. The institutional consequence follows from whether a plan should identify the latest point at which a missing dependency can be corrected without compromising delivery. Where a dependency lies with another authority, agreement and escalation should precede large-scale implementation. Dependencies should be sequenced. Teacher guidance may require curriculum clarification, planning time, examples and leadership support. New materials require procurement, accessible formats, distribution and maintenance.[REF-17] [REF-18] [REF-25]
In assessing improvement proposition and implementation design, authorities must determine an institution should know which learners or staff receive the measure, how frequently, for how long and with what expected standard. The evidence must therefore clarify how without intensity, non-response cannot be distinguished from weak delivery. Attendance at one professional meeting, receipt of materials or registration in support does not prove sustained use. Monitoring should capture actual exposure and reasons for incomplete delivery while avoiding burdens so heavy that they displace teaching or service. Implementation intensity should be specified.[REF-03] [REF-06] [REF-16]
Review of improvement proposition and implementation design is credible only where it explains the plan should test which learners can take up the measure and provide the additional condition required for substantive access. A proportionate conclusion must also recognise that local adaptation is desirable where it removes a barrier while protecting the common educational objective. It becomes inequitable where learners facing disadvantage receive reduced content, weaker staff or an unrecognised route. Equity should be built into eligibility, outreach and adaptation. An apparently equal offer may be unusable because of cost, time, language, disability, safety or location.[REF-08] [REF-13] [REF-14]
improvement proposition and implementation design requires a decision about the plan should identify indicators or qualitative evidence that would reveal these effects and the authority able to intervene. A contrary reading would overlook that improvement is not established by movement of the selected measure if another essential condition deteriorates. Foreseeable adverse effects should be recorded. Additional assessment can narrow curriculum or increase exclusion; targeted grouping can stigmatise learners; extended time can increase household costs; intensive support for one cohort can divert staff from another.[REF-01] [REF-11] [REF-23]
Resources and professional capability
The central question in resources and professional capability is capital expenditure or a temporary grant may be necessary but cannot establish sustained capability by itself. For the learners concerned, the decisive consideration is whether the institution should identify which costs fall within its budget, which require higher-level allocation and which are presently unfunded. An unfunded dependency should remain visible in the risk statement. The resource plan should cover the complete recurrent service: personnel, preparation, instructional time, accessible materials, facilities, learner support, assessment, coordination, evidence and review.[REF-15] [REF-26] [REF-28]
resources and professional capability requires a decision about additional finance should be justified by the common educational condition it protects, not by a permanent assumption of lower capability among the affected learners. A contrary reading would overlook that distributional finance matters within institutions as well as across systems. Programmes serving learners with greater need may require more staff time, specialist support, transport or accessible materials. Equal departmental or per-learner amounts can reproduce unequal opportunity. Allocation reasons should be documented and reviewed against actual receipt and use.[REF-06] [REF-10] [REF-13]
Evidence concerning resources and professional capability should establish the plan should monitor participation, usable learning, workload, turnover and access to later assistance. A contrary reading would overlook that professional capability is a central implementation condition. Teachers, trainers and institutional leaders need subject knowledge, diagnostic skill, practical examples and time to collaborate. Where the plan alters curriculum, assessment or inclusion arrangements, staff should have supported opportunities to understand the educational reasoning. One briefing does not establish readiness for sustained change.[REF-01] [REF-18] [REF-25]
The practical standard for resources and professional capability concerns conversely, higher-level constraints should not prevent local correction that is lawful and feasible. This matters because the plan should identify decision thresholds, escalation routes and the time within which the responsible body will respond. Leadership should create a route from classroom or service evidence to institutional and system correction. Staff should not be held responsible for changing transport, staffing establishment, qualification rules or infrastructure beyond their authority.[REF-12] [REF-17] [REF-19]
The central question in resources and professional capability is the measure of partnership is whether the institution can sustain equitable quality and correction after exceptional support ends. This matters because external partnership should strengthen ordinary capability. Universities, civil society, employers or international bodies may contribute knowledge, facilities or finance, but programme status, data protection and continuity should be clear. Parallel arrangements can create fragmented standards or reporting burdens. Agreements should state the educational purpose, authority, resource duration, records, safeguards and handover.[REF-22] [REF-26] [REF-28]
Evidence, adaptation and institutionalisation
Review of evidence, adaptation and institutionalisation is credible only where it explains none automatically proves the next. This matters because the review schedule should collect evidence proportionate to each claim and identify the population represented. A favourable result should not be broadened beyond the learners, content and period observed. Review should distinguish authorization, delivery, reach, educational quality and consequence. Approval of the plan establishes authority; expenditure demonstrates a resource transaction; implementation records show activity; learner evidence shows participation or change.[REF-03] [REF-16] [REF-24]
Review of evidence, adaptation and institutionalisation is credible only where it explains non-participation can reveal communication, cost, timing, safety or accessibility barriers. The institutional consequence follows from whether the review should disaggregate where lawful and precise and should retain absolute numbers. An improved average does not establish that learners farthest from the baseline benefited. The plan should define a minimum floor or least-served test alongside its aggregate objective. Reach should be examined from eligibility through offer, take-up, participation and intended intensity.[REF-01] [REF-06] [REF-13]
Evidence concerning evidence, adaptation and institutionalisation should establish agreement strengthens the conclusion only after dates and definitions align. The institutional consequence follows from whether contradiction should initiate enquiry into coverage, implementation or measurement rather than selection of the most favourable source. Educational consequence may require several forms of evidence. Assessment can show a bounded learning domain; attendance and completion records show participation; observation and work samples can explain teaching and learner activity; interviews can identify experience and mechanism.[REF-02] [REF-12] [REF-23]
Public responsibility for evidence, adaptation and institutionalisation begins with these explanations call for different action. The institutional consequence follows from whether the review should record which is supported, what changes and how the baseline or comparison remains interpretable. Ending an ineffective measure is responsible improvement when learner safeguards and alternative provision are addressed. Adaptation should follow a stated decision rule. Weak outcomes may reflect an incorrect hypothesis, incomplete delivery, insufficient intensity, adverse conditions, measurement weakness or inadequate time.[REF-08] [REF-11] [REF-28]
Evidence concerning evidence, adaptation and institutionalisation should establish continued activity after a pilot is not sufficient if it depends on exceptional staff or external funds. For the learners concerned, the decisive consideration is whether the final decision should state which function enters ordinary provision, which limitation remains, who owns it and when evidence will next be examined. Completion means that the institution can sustain the improved learner-facing condition and correct foreseeable departure, not that the plan period has ended. Institutionalisation requires ordinary authority, recurrent finance, professional ownership and a functioning correction route.[REF-17] [REF-25] [REF-26]
References
- REF-01
Education for All Global Monitoring Report Team. Reaching the Marginalized — EFA Global Monitoring Report 2010. 2010.
Principal contemporaneous analysis of intersecting disadvantage and education marginalisation.
https://unesdoc.unesco.org/ark:/48223/pf0000186606 - REF-02
UNESCO Institute for Statistics. Global Education Digest 2010: Comparing Education Statistics Across the World. 2010.
Comparative education statistics, definitions and limitations.
https://uis.unesco.org/sites/default/files/documents/global-education-digest-2010-comparing-education-statistics-across-the-world-en.pdf - REF-03
UNESCO Institute for Statistics. Education Indicators: Technical Guidelines. 2009.
Definitions and interpretation of participation, progression, completion and resource indicators.
https://uis.unesco.org/sites/default/files/documents/education-indicators-technical-guidelines-en_0.pdf - REF-04
United Nations. The Millennium Development Goals Report 2010. 2010.
Global and regional monitoring of primary education, gender, poverty and related development conditions.
https://www.un.org/millenniumgoals/pdf/MDG%20Report%202010%20En%20r15%20-low%20res%2020100615%20-.pdf - REF-05
United Nations Development Programme. Human Development Report 2010: The Real Wealth of Nations — Pathways to Human Development. 2010.
Distribution-sensitive human development concepts and evidence available before the cut-off.
https://hdr.undp.org/content/human-development-report-2010 - REF-06
UNICEF. Progress for Children: Achieving the MDGs with Equity, Number 9. 2010.
Equity-focused child indicators and comparison between population groups.
https://www.unicef.org/reports/progress-children-no-9 - REF-07
World Education Forum. The Dakar Framework for Action: Education for All — Meeting Our Collective Commitments. 2000.
Commitments to equitable access, quality, measurable outcomes and accountable national planning.
https://unesdoc.unesco.org/ark:/48223/pf0000121147 - REF-08
United Nations General Assembly. Convention on the Rights of the Child. 1989.
Rights concerning non-discrimination, identity, education and development.
https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-child - REF-09
United Nations Committee on Economic, Social and Cultural Rights. General Comment No. 13: The Right to Education. 1999.
Interpretation of availability, accessibility, acceptability and adaptability.
https://www.refworld.org/legal/general/cescr/1999/en/37937 - REF-10
United Nations General Assembly. Convention on the Rights of Persons with Disabilities. 2006.
Non-discrimination, accessibility, inclusive education and disability data safeguards.
https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-persons-disabilities - REF-11
United Nations. Guiding Principles on Internal Displacement. 1998.
Principles relevant to protection, documentation, education and non-discrimination of displaced persons.
https://www.ohchr.org/en/special-procedures/sr-internally-displaced-persons/international-standards - REF-12
UNESCO and UNICEF. A Human Rights-Based Approach to Education for All. 2007.
Rights-based planning, equality, participation, accountability and education quality.
https://unesdoc.unesco.org/ark:/48223/pf0000154861 - REF-13
Education for All Global Monitoring Report Team. Overcoming Inequality: Why Governance Matters — EFA Global Monitoring Report 2009. 2008.
Governance, finance and unequal educational opportunity.
https://unesdoc.unesco.org/ark:/48223/pf0000177683 - REF-14
World Bank. World Development Report 2006: Equity and Development. 2005.
Concepts of unequal opportunity, institutions and equitable public action.
https://documents.worldbank.org/curated/en/435331468127174418/pdf/322040World0Development0Report02006.pdf - REF-15
World Bank. Safeguarding Education During Economic Crisis. 2009.
Risks to budgets, households, participation and long-term human development during economic crisis.
https://documents1.worldbank.org/curated/en/489131468340200911/pdf/485120WP0Avert10Box338912B01PUBLIC1.pdf - REF-16
Organisation for Economic Co-operation and Development. Education at a Glance 2010: OECD Indicators. 2010.
Comparative participation, progression, expenditure and outcomes evidence with system-level metadata.
https://doi.org/10.1787/eag-2010-en - REF-17
European Commission. Europe 2020: A Strategy for Smart, Sustainable and Inclusive Growth. 2010.
Contemporaneous European policy context for education, inclusion, employment and headline indicators.
https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC2020 - REF-18
European Commission. Youth on the Move: An Initiative to Unleash the Potential of Young People to Achieve Smart, Sustainable and Inclusive Growth in the European Union. 2010.
European education, mobility, attainment and youth inclusion policy context.
https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0477 - REF-19
European Commission. A Renewed Commitment to Social Europe: Reinforcing the Open Method of Coordination for Social Protection and Social Inclusion. 2008.
Social inclusion monitoring, common objectives and context-sensitive indicators.
https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52008DC0418 - REF-20
United Nations General Assembly. Resolution 64/250: Assistance to Haiti in the Aftermath of the Recent Earthquake. 2010.
Contemporaneous recognition of humanitarian and reconstruction needs and national leadership.
https://undocs.org/A/RES/64/250 - REF-21
United Nations Office for the Coordination of Humanitarian Affairs. Haiti Revised Humanitarian Appeal. 2010.
Displacement, service disruption and humanitarian education context.
https://reliefweb.int/report/haiti/haiti-revised-humanitarian-appeal-2010 - REF-22
European Commission. European Union Response to the Earthquake in Haiti. 2010.
European humanitarian and recovery support, coordination and Haitian ownership.
https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0056 - REF-23
United Nations Economic and Social Council. Principles and Recommendations for Population and Housing Censuses, Revision 2. 2008.
Official principles for population coverage, definitions, classifications and data quality.
https://unstats.un.org/unsd/demographic-social/Standards-and-Methods/files/Principles_and_Recommendations/Population-and-Housing-Censuses/Series_M67Rev2-E.pdf - REF-24
United Nations Statistics Division. Designing Household Survey Samples: Practical Guidelines. 2005.
Sample design, estimation, precision and non-response guidance.
https://unstats.un.org/unsd/demographic/sources/surveys/Handbook23June05.pdf - REF-25
World Education Forum 2015. Incheon Declaration: Education 2030 — Towards Inclusive and Equitable Quality Education and Lifelong Learning for All. 2015.
Education commitments adopted at Incheon and their contemporaneous institutional status.
https://unesdoc.unesco.org/ark:/48223/pf0000233137 - REF-26
United Nations General Assembly. Addis Ababa Action Agenda of the Third International Conference on Financing for Development. 2015.
Adopted global financing framework and relevant principles for domestic public finance and international cooperation.
https://undocs.org/A/RES/69/313 - REF-27
United Nations General Assembly. Transforming Our World: The 2030 Agenda for Sustainable Development. 2015.
Agenda adopted on 25 September 2015, including Goal 4 and its education targets, as available at the evidence cut-off.
https://undocs.org/A/RES/70/1 - REF-28
UNESCO. Education 2030: Incheon Declaration and Framework for Action for the Implementation of Sustainable Development Goal 4. 2015.
Framework for Action available before the cut-off and relevant to institutional planning, coordination, equity and review.
https://unesdoc.unesco.org/ark:/48223/pf0000245656 - REF-29
United Nations Secretary-General. One Humanity: Shared Responsibility — Report of the Secretary-General for the World Humanitarian Summit. 2016. A/70/709.
Secretary-General’s pre-Summit Agenda for Humanity and proposals for humanitarian responsibility.
https://docs.un.org/A/70/709 - REF-30
Chair of the World Humanitarian Summit. Chair’s Summary: Standing Up for Humanity — Committing to Action. 2016.
Official Chair’s account of the May 2016 Summit discussions and commitments, used with its status as a summary.
https://agendaforhumanity.org/sites/default/files/resources/2017/Jul/chairs_summary.pdf - REF-31
United Nations. New Fund Launched at UN Humanitarian Summit to Address Education in Crisis Zones. 2016.
Contemporaneous official account of the Education Cannot Wait launch and stated aims; not evidence of later finance or results.
https://www.un.org/sustainabledevelopment/blog/2016/05/new-fund-launched-at-un-humanitarian-summit-to-address-education-in-crisis-zones/ - REF-32
United Nations General Assembly. The Right to Education in Emergency Situations. 2010. A/RES/64/290.
Recognition of education as an integral element of humanitarian response and calls for access, continuity, protection and support.
https://docs.un.org/A/RES/64/290 - REF-33
United Nations General Assembly. New York Declaration for Refugees and Migrants, Resolution 71/1. 2016.
Contemporaneous commitment relevant to shared responsibility, refugee and migrant education, and the reporting boundary between commitment and result.
https://undocs.org/A/RES/71/1 - REF-34
Global Education Monitoring Report Team. Accountability in Education: Meeting Our Commitments — Global Education Monitoring Report 2017/8. 2017.
Principal contemporaneous analysis of accountability relations, responsibility, transparency, participation and unintended effects in education.
https://unesdoc.unesco.org/ark:/48223/pf0000259338 - REF-35
World Bank. World Development Report 2018: Learning to Realize Education's Promise. 2018.
Principal contemporaneous analysis of learning, measurement, evidence-led action and alignment of education actors.
https://www.worldbank.org/en/publication/wdr2018 - REF-36
United Nations General Assembly. Global Compact on Refugees, Report of the United Nations High Commissioner for Refugees, Part II, affirmed by Resolution 73/151. 2018.
Controlling contemporaneous compact on predictable and equitable responsibility-sharing, national arrangements and refugee inclusion, including education.
https://www.unhcr.org/media/global-compact-refugees-booklet - REF-37
UNESCO. With One in Five Learners Kept out of School, UNESCO Mobilizes Education Ministers to Face the COVID-19 Crisis. 2020.
Early contemporaneous record of rapidly expanding institutional closures and international education coordination.
https://www.unesco.org/en/articles/one-five-learners-kept-out-school-unesco-mobilizes-education-ministers-face-covid-19-crisis - REF-38
World Health Organization. WHO Director-General's Opening Remarks at the Media Briefing on COVID-19 — 11 March 2020. 2020.
Contemporaneous characterization of COVID-19 as a pandemic and call for comprehensive national action.
https://www.who.int/news-room/speeches/item/who-director-general-s-opening-remarks-at-the-media-briefing-on-covid-19---11-march-2020 - REF-39
United Nations. Shared Responsibility, Global Solidarity: Responding to the Socio-Economic Impacts of COVID-19. 2020.
Contemporaneous United Nations account of social and economic disruption and the need for rights-respecting coordinated action.
https://unsdg.un.org/resources/shared-responsibility-global-solidarity-responding-socio-economic-impacts-covid-19 - REF-40
Global Education Monitoring Report Team. Inclusion and Education: All Means All — Global Education Monitoring Report 2020. 2020.
Principal contemporaneous inclusion analysis available before the cutoff, used to interpret unequal access and populations omitted by disruption measures.
https://unesdoc.unesco.org/ark:/48223/pf0000373718 - REF-41
World Bank, UNESCO Global Education Monitoring Report Team and UNESCO Institute for Statistics. Education Finance Watch 2021. 2021.
Contemporaneous evidence on education-budget pressure and the need to align recovery priorities with feasible public finance.
https://www.worldbank.org/en/news/press-release/2021/02/22/two-thirds-of-poorer-countries-are-cutting-education-budgets-due-to-covid-19 - REF-42
UNESCO General Conference. Global Convention on the Recognition of Qualifications concerning Higher Education. 2019.
Adopted recognition instrument available by the cutoff; interpreted according to its pre-entry-into-force status.
https://www.unesco.org/en/legal-affairs/global-convention-recognition-qualifications-concerning-higher-education