ICEQC-R-2010-10 — Identifying Marginalisation in Education Data: Beyond National Averages cover

专题研究报告

ICEQC-R-2010-10 — Identifying Marginalisation in Education Data: Beyond National Averages

A global comparative indicator study of distributions, intersecting disadvantage and accountable use of disaggregated evidence

发布日期
研究类别
数据与指标研究
报告类型
指标比较研究
地理范围
Global
证据截止日期
负责机构
国际教育质量认证委员会研究与政策司
ICEQC-R-2010-10 — Identifying Marginalisation in Education Data: Beyond National Averages cover

Publication record

This is the controlled English edition. Evidence and institutional status are stated as at the evidence cut-off date.

Executive summary

Review of executive summary is credible only where it explains a stable completion rate may hide increasing delay, repetition or withdrawal among groups whose educational records and population denominators are weakest. A proportionate conclusion must also recognise that the evidentiary task is therefore not simply to publish more categories. It is to construct distributions whose populations, definitions, uncertainty and limits permit fair public judgement. National education averages describe the centre of a distribution while frequently concealing the conditions of learners farthest from secured educational opportunity. A rising enrolment rate can coexist with persistent exclusion in a remote district, a poor household group or a population displaced by conflict or disaster. Gender parity at national level can coexist with disadvantage for poor rural girls and poor urban boys.

For executive summary, the material distinction is between official education statistics and development monitoring also provide methods for participation, progression, completion, resources and population comparison. The evidence must therefore clarify how this study brings those strands together. It examines who enters the denominator, which education state is measured, how groups and places are classified, what missingness does to an estimate and how an observed disparity should enter a decision. The 2010 Education for All Global Monitoring Report placed marginalisation at the centre of global education analysis and showed the importance of overlapping disadvantage.

executive summary requires a decision about disaggregation by one characteristic is often necessary but rarely sufficient. The evidence must therefore clarify how wealth, sex, residence, language, disability, displacement and prior education can intersect, while sample sizes and confidentiality place legitimate limits on publication. When a population cannot be estimated reliably, uncertainty should be reported rather than converted into absence. A distribution-sensitive indicator begins with an entitlement and a population, not with an available column in a register. Its numerator and denominator must refer to compatible people, places and periods. Counts, levels, gaps and ratios serve different purposes and should be read together.

The practical standard for executive summary concerns censuses provide broad population and small-area evidence at longer intervals, while coverage and classification remain material. The public account remains incomplete unless it explains how no source is complete for every question. Agreement strengthens a conclusion only after definitions are reconciled; disagreement is a reason to investigate population, timing and measurement. The study distinguishes administrative records, household surveys and censuses. School records offer regular institutional detail but usually omit children outside provision. Surveys can represent households beyond schools, subject to their frames, response and sampling uncertainty.

The continuing economic crisis and the disruption following the Haiti earthquake illustrate why averages can become least reliable when public decisions are most urgent. Budget totals may conceal local service contraction and higher household costs. Displacement can change both numerator and denominator, create duplicate or missing records and weaken comparisons with an earlier population. These conditions require dated estimates, visible revisions and a distinction between rapid operational counts and statistics suitable for final comparison. The report uses only evidence available by 10 November 2010 and makes no claim about subsequent recovery.

In assessing executive summary, authorities must determine an average remains valuable context. For the learners concerned, the decisive consideration is whether it becomes misleading when it is allowed to stand for a distribution that the evidence is capable of revealing. The central conclusion is that a credible national account should present level, distribution, missingness and response together. It should identify which learners remain unrepresented; report the strength of each comparison; distinguish statistical association from cause; and name the authority responsible for remedy.

Key findings

    Scope and method

    Evidence concerning scope and method should establish it does not create a universal ranking, prescribe protected classifications without regard to national law, or claim that one set of disaggregations is appropriate in every setting. For the learners concerned, the decisive consideration is whether this comparative indicator study addresses access, participation, progression, completion, learning, conditions and finance. Its purpose is to guide the construction and interpretation of national and subnational distributions.

    In assessing scope and method, authorities must determine the source test examines coverage, error and revision. A proportionate conclusion must also recognise that the distribution test considers groups, places and intersections. The decision test asks what conclusion and action the evidence can support. Rights and Education for All commitments supply the public-interest frame; official statistical sources supply definitions and quality principles. The method applies five tests. The population test asks who could appear in the numerator and denominator. The concept test asks what education state is actually observed.

    scope and method requires a decision about the report distinguishes direct observation, estimate and interpretation. A proportionate conclusion must also recognise that material quantitative claims derived from a cited source remain subject to that source's definitions; illustrative measurement propositions do not assert unnamed national results. Comparisons are treated as bounded. A harmonised term does not remove differences in programme structure, school age, household classification or collection practice.

    All evidence and institutional status are stated as at 10 November 2010.

    Part I

    Purpose and interpretive frame

    1

    What national averages conceal

    what national averages conceal requires a decision about the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how the indicator question concerns national participation or attainment. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    The practical standard for what national averages conceal concerns administrative records should be reconciled with population-based evidence where their coverage differs. For the learners concerned, the decisive consideration is whether a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is the weighted experience of the population included in the denominator. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-02] [REF-03]

    Institutional action on what national averages conceal should be tested against statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. For the learners concerned, the decisive consideration is whether explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: distribution within countries, not a league table between them. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-04] [REF-06]

    A defensible account of what national averages conceal distinguishes protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The public account remains incomplete unless it explains how the distributional review must deliberately include remote rural learners, urban informal settlements and displaced populations. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-08] [REF-09]

    Comparative interpretation of what national averages conceal depends upon monitoring after action must preserve the original baseline and follow both reach and outcome. A proportionate conclusion must also recognise that improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-10] [REF-12]

    2

    Marginalisation as accumulated disadvantage

    The central question in marginalisation as accumulated disadvantage is a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The evidence must therefore clarify how the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Marginalisation as accumulated disadvantage should be approached as a defined measurement problem. The substantive interest is distance from a socially secured educational minimum, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-02] [REF-03]

    The central question in marginalisation as accumulated disadvantage is reconciliation should record what each source can and cannot represent. The institutional consequence follows from whether where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use the joint effect of exclusion, weak provision and adverse social conditions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-04] [REF-06]

    Public responsibility for marginalisation as accumulated disadvantage begins with percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. This matters because the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that multiple indicators read together over the learner's course. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-08] [REF-09]

    marginalisation as accumulated disadvantage cannot be judged without identifying where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. The institutional consequence follows from whether particular scrutiny is required for children facing poverty, gender disadvantage, disability or minority status. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-10] [REF-12]

    The practical standard for marginalisation as accumulated disadvantage concerns these are different decisions and require different certainty. The evidence must therefore clarify how urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure.[REF-13] [REF-14]

    3

    From monitoring commitment to decision

    The central question in from monitoring commitment to decision is the study seeks evidence on evidence capable of changing resource or service decisions; it does not infer a learner's circumstances from a national or regional mean. A proportionate conclusion must also recognise that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For from monitoring commitment to decision, the first requirement is conceptual clarity.[REF-04] [REF-06]

    For from monitoring commitment to decision, the material distinction is between this distinction is essential for participation and progression. For the learners concerned, the decisive consideration is whether the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on a link between observed disparity, responsible body and remedy. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-08] [REF-09]

    A defensible account of from monitoring commitment to decision distinguishes results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. A proportionate conclusion must also recognise that if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: indicators selected for action rather than visibility.[REF-10] [REF-12]

    In assessing from monitoring commitment to decision, authorities must determine attrition at any stage can produce an apparently complete indicator from a selective population. For the learners concerned, the decisive consideration is whether field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether groups absent from routine plans and budget classifications are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-13] [REF-14]

    Institutional action on from monitoring commitment to decision should be tested against communities should be able to question both the category and the conclusion drawn from it. The public account remains incomplete unless it explains how if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-15] [REF-16]

    Part II

    Population and denominator

    4

    Defining the population entitled to education

    A defensible account of defining the population entitled to education distinguishes without that purpose, disaggregation can multiply figures without improving public judgement. The institutional consequence follows from whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of resident and temporarily absent learners within the relevant age or programme group begins by naming the decision the evidence may inform.[REF-08] [REF-09]

    Review of defining the population entitled to education is credible only where it explains a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. The public account remains incomplete unless it explains how operationally, the measure is a denominator consistent with the right, level and reference period under review. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence.[REF-10] [REF-12]

    The central question in defining the population entitled to education is apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. This matters because use of the indicator is bounded by the principle that census, survey and administrative estimates reconciled openly. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-13] [REF-14]

    defining the population entitled to education cannot be judged without identifying suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. A contrary reading would overlook that the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For unregistered residents, migrants and displaced children, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community.[REF-15] [REF-16]

    defining the population entitled to education cannot be judged without identifying subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. This matters because where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence.[REF-17] [REF-19]

    5

    Age, grade and programme populations

    Evidence concerning age, grade and programme populations should establish a contrary reading would overlook that the public value of age, grade and programme populations lies in making unequal educational experience observable. A contrary reading would overlook that here, the relevant phenomenon is age-specific, grade-specific and programme-specific participation, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The central question in age, grade and programme populations is the report should state the educational consequence before choosing a gap, ratio, threshold or rank.[REF-10] [REF-12]

    The practical standard for age, grade and programme populations concerns the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The evidence must therefore clarify how the preferred construction is separate denominators for different educational questions. Numerator, denominator, reference date, unit and exclusions should appear together. The central question in age, grade and programme populations is if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. A proportionate conclusion must also recognise that administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response.[REF-13] [REF-14]

    Review of age, grade and programme populations is credible only where it explains statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. The resulting interpretation should show why explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: exact age, school age and enrolled population never substituted silently. The central question in age, grade and programme populations is the comparison should show the level for each group as well as any ratio or gap. This matters because a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-15] [REF-16]

    age, grade and programme populations cannot be judged without identifying these populations may be missing not only from good outcomes but from the denominator itself. The institutional consequence follows from whether the central question in age, grade and programme populations is coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. This matters because if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include over-age entrants and learners repeating grades.[REF-17] [REF-19]

    Evidence concerning age, grade and programme populations should establish improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. A proportionate conclusion must also recognise that the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome.[REF-20] [REF-21]

    6

    Population movement and disrupted residence

    In assessing population movement and disrupted residence, authorities must determine the measure should preserve both the observed level and the distribution relevant to the claim. The public account remains incomplete unless it explains how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns education status amid migration, displacement and return. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-13] [REF-14]

    population movement and disrupted residence requires a decision about where estimates are revised, both the reason and effect of revision should remain accessible. This matters because measurement should use dated location and residence rules with sensitivity to mobility. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-15] [REF-16]

    Evidence concerning population movement and disrupted residence should establish policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The resulting interpretation should show why the governing caution is that origin, current location and service responsibility distinguished. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-17] [REF-19]

    The central question in population movement and disrupted residence is their circumstances may alter access to enumeration, classification and the service being measured. The evidence must therefore clarify how the review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for families affected by conflict, disaster and seasonal movement.[REF-20] [REF-21]

    Comparative interpretation of population movement and disrupted residence depends upon urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. The resulting interpretation should show why targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-23] [REF-24]

    Part III

    Access and participation

    7

    Entry at the official starting age

    Institutional action on entry at the official starting age should be tested against the measure should preserve both the observed level and the distribution relevant to the claim. The institutional consequence follows from whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Entry at the official starting age should be approached as a defined measurement problem. The substantive interest is timely admission to the first grade of primary education, observed for a population and period that are stated before calculation.[REF-15] [REF-16]

    Institutional action on entry at the official starting age should be tested against the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. The evidence must therefore clarify how when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on new entrants of official age relative to the corresponding population. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-17] [REF-19]

    Public responsibility for entry at the official starting age begins with nor should a group estimate be read as a description of every member. The institutional consequence follows from whether within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: late entry examined alongside non-entry. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-20] [REF-21]

    In assessing entry at the official starting age, authorities must determine analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. A contrary reading would overlook that missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether children facing fees, distance, disability or documentation barriers are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-23] [REF-24]

    Evidence concerning entry at the official starting age should establish if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. A proportionate conclusion must also recognise that the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-01] [REF-07]

    8

    Attendance beyond enrolment

    For attendance beyond enrolment, the material distinction is between the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how for attendance beyond enrolment, the first requirement is conceptual clarity. The study seeks evidence on actual participation during a stated recent period; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-17] [REF-19]

    attendance beyond enrolment cannot be judged without identifying its metadata should travel with every published value. A contrary reading would overlook that at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is presence measured independently of registration status.[REF-20] [REF-21]

    attendance beyond enrolment cannot be judged without identifying a responsible commentary distinguishes observation, calculation and interpretation. A contrary reading would overlook that it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that frequency, season and reasons for absence retained.[REF-23] [REF-24]

    Evidence concerning attendance beyond enrolment should establish disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. The resulting interpretation should show why this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For working children, carers and learners affected by illness, a single group label may conceal important internal differences.[REF-01] [REF-07]

    For attendance beyond enrolment, the material distinction is between a policy conclusion should be no broader than that evidence. For the learners concerned, the decisive consideration is whether subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified.[REF-02] [REF-03]

    9

    Out-of-school status

    For out-of-school status, the material distinction is between without that purpose, disaggregation can multiply figures without improving public judgement. A contrary reading would overlook that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of children in the relevant age group not participating at the defined level begins by naming the decision the evidence may inform.[REF-20] [REF-21]

    The practical standard for out-of-school status concerns it should require examination of definitions, timing, migration, duplication and non-response. The resulting interpretation should show why the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a transparent residual from compatible population and participation concepts. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-23] [REF-24]

    A defensible account of out-of-school status distinguishes reference points therefore need substantive meaning. This matters because where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: never-enrolled and formerly enrolled children separated. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-01] [REF-07]

    Comparative interpretation of out-of-school status depends upon these populations may be missing not only from good outcomes but from the denominator itself. The evidence must therefore clarify how coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children invisible to school registers.[REF-02] [REF-03]

    Comparative interpretation of out-of-school status depends upon it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. The public account remains incomplete unless it explains how monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition.[REF-04] [REF-06]

    Part IV

    Progression and completion

    10

    Repetition and grade survival

    Review of repetition and grade survival is credible only where it explains the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A proportionate conclusion must also recognise that the public value of repetition and grade survival lies in making unequal educational experience observable. Here, the relevant phenomenon is movement through grades without avoidable delay or exit, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-23] [REF-24]

    repetition and grade survival cannot be judged without identifying where estimates are revised, both the reason and effect of revision should remain accessible. A contrary reading would overlook that measurement should use cohort or reconstructed-cohort evidence with explicit assumptions. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-01] [REF-07]

    For repetition and grade survival, the material distinction is between a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. A proportionate conclusion must also recognise that percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that repeaters distinguished from re-entrants and transfers. Disparity measures should not replace the underlying distributions.[REF-02] [REF-03]

    A defensible account of repetition and grade survival distinguishes the review should record non-response, unknown status and excluded locations separately. The evidence must therefore clarify how combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in overcrowded or intermittently operating schools. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-04] [REF-06]

    Comparative interpretation of repetition and grade survival depends upon targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. The institutional consequence follows from whether accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves.[REF-08] [REF-09]

    11

    Transition between education levels

    The central question in transition between education levels is its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns entry to the next level after completion of the preceding one.[REF-01] [REF-07] [REF-18]

    Public responsibility for transition between education levels begins with the definition should be fixed for the comparison at hand and deviations recorded. The public account remains incomplete unless it explains how analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on matched completion and new-entry populations over coherent periods.[REF-02] [REF-03]

    Public responsibility for transition between education levels begins with nor should a group estimate be read as a description of every member. The evidence must therefore clarify how within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: capacity constraints distinguished from learner attainment. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-04] [REF-06]

    The practical standard for transition between education levels concerns analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. The evidence must therefore clarify how missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether rural learners and those unable to relocate are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-08] [REF-09]

    Public responsibility for transition between education levels begins with communities should be able to question both the category and the conclusion drawn from it. The resulting interpretation should show why if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-10] [REF-12]

    12

    Completion and educational entitlement

    Institutional action on completion and educational entitlement should be tested against the measure should preserve both the observed level and the distribution relevant to the claim. A contrary reading would overlook that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Completion and educational entitlement should be approached as a defined measurement problem. The substantive interest is finishing the final grade or meeting recognised programme requirements, observed for a population and period that are stated before calculation.[REF-02] [REF-03]

    Institutional action on completion and educational entitlement should be tested against a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. The evidence must therefore clarify how a survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is completion defined separately from sitting or passing an examination. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions.[REF-04] [REF-06]

    The practical standard for completion and educational entitlement concerns it does not assign cause from a cross-sectional difference. The evidence must therefore clarify how apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that late completion and alternative pathways reported. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-08] [REF-09]

    completion and educational entitlement requires a decision about this balance is contextual rather than mechanical. For the learners concerned, the decisive consideration is whether it should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For over-age learners and people returning after interruption, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable.[REF-10] [REF-12]

    completion and educational entitlement requires a decision about a policy conclusion should be no broader than that evidence. The resulting interpretation should show why subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified.[REF-13] [REF-14]

    Part V

    Learning and assessment

    13

    Minimum learning outcomes

    Public responsibility for minimum learning outcomes begins with a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A proportionate conclusion must also recognise that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For minimum learning outcomes, the first requirement is conceptual clarity. The study seeks evidence on demonstrated knowledge or skill against a declared domain; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-04] [REF-06]

    Comparative interpretation of minimum learning outcomes depends upon administrative records should be reconciled with population-based evidence where their coverage differs. A contrary reading would overlook that a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is assessment evidence whose population and conditions are known. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-08] [REF-09]

    minimum learning outcomes cannot be judged without identifying a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. The institutional consequence follows from whether reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: results interpreted with opportunity to learn and participation. The comparison should show the level for each group as well as any ratio or gap.[REF-10] [REF-12]

    Public responsibility for minimum learning outcomes begins with confidentiality is essential, especially where identity or status creates risk. For the learners concerned, the decisive consideration is whether protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners taught in an unfamiliar language. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-13] [REF-14]

    Comparative interpretation of minimum learning outcomes depends upon an indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. The evidence must therefore clarify how it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date.[REF-15] [REF-16]

    14

    Assessment participation

    For assessment participation, the material distinction is between a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The institutional consequence follows from whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of who was eligible, present, absent and excluded from testing begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]

    The central question in assessment participation is reconciliation should record what each source can and cannot represent. For the learners concerned, the decisive consideration is whether where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use a participation profile accompanying every result distribution. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-10] [REF-12]

    Review of assessment participation is credible only where it explains policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The evidence must therefore clarify how the governing caution is that non-participation never treated as low attainment or ignored. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm.[REF-13] [REF-14]

    In assessing assessment participation, authorities must determine where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. The public account remains incomplete unless it explains how particular scrutiny is required for learners with disabilities and remote candidates. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-15] [REF-16]

    Comparative interpretation of assessment participation depends upon urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. The institutional consequence follows from whether targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty.[REF-17] [REF-19]

    15

    Distribution of achievement

    Institutional action on distribution of achievement should be tested against here, the relevant phenomenon is variation across the full score or proficiency distribution, not the administrative convenience of the available categories. For the learners concerned, the decisive consideration is whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of distribution of achievement lies in making unequal educational experience observable.[REF-10] [REF-12]

    distribution of achievement requires a decision about the definition should be fixed for the comparison at hand and deviations recorded. The evidence must therefore clarify how analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on percentiles and threshold shares beside a mean.[REF-13] [REF-14]

    Comparative interpretation of distribution of achievement depends upon within-group variation and unmeasured intersecting conditions remain material. The public account remains incomplete unless it explains how the analytical rule is clear: uncertainty and scale properties stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-15] [REF-16]

    Review of distribution of achievement is credible only where it explains field arrangements need relevant languages, accessible formats and safe participation. The evidence must therefore clarify how analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether learners concentrated below a minimum proficiency threshold are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-17] [REF-19]

    A defensible account of distribution of achievement distinguishes if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The evidence must therefore clarify how the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-20] [REF-21]

    Part VI

    Gender and household resources

    16

    Gender parity and its limits

    Public responsibility for gender parity and its limits begins with the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The resulting interpretation should show why the indicator question concerns differences between girls and boys in access, progression and learning. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-13] [REF-14]

    The central question in gender parity and its limits is a census figure should disclose enumeration rules. The public account remains incomplete unless it explains how these are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is female-to-male ratios read with levels and absolute gaps. The central question in gender parity and its limits is its metadata should travel with every published value. The resulting interpretation should show why at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty.[REF-15] [REF-16]

    Comparative interpretation of gender parity and its limits depends upon for the learners concerned, the decisive consideration is whether use of the indicator is bounded by the principle that parity not confused with adequacy for either group. A contrary reading would overlook that a responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. The central question in gender parity and its limits is apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere.[REF-17] [REF-19]

    For gender parity and its limits, the material distinction is between disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. The evidence must therefore clarify how this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. The central question in gender parity and its limits is suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. A proportionate conclusion must also recognise that the public report can still state that a disparity was examined, whether action is required and which body will monitor it. For girls in poor rural households and boys exposed to hazardous work, a single group label may conceal important internal differences.[REF-20] [REF-21]

    The central question in gender parity and its limits is readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. The resulting interpretation should show why a policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. The central question in gender parity and its limits is this makes the indicator a means of scrutiny rather than a decorative measure of concern. The institutional consequence follows from whether public accountability requires a concise explanation of the result and its boundary.[REF-23] [REF-24]

    17

    Household wealth gradients

    In assessing household wealth gradients, authorities must determine the measure should preserve both the observed level and the distribution relevant to the claim. A contrary reading would overlook that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Household wealth gradients should be approached as a defined measurement problem. The substantive interest is education outcomes across relative household resource groups, observed for a population and period that are stated before calculation.[REF-15] [REF-16]

    household wealth gradients requires a decision about a discrepancy is not resolved by selecting the more favourable source. The evidence must therefore clarify how it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is a documented asset or consumption classification within each setting. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-17] [REF-19]

    household wealth gradients cannot be judged without identifying where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. The evidence must therefore clarify how statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: wealth ranks not assumed equivalent across countries or time. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning.[REF-20] [REF-21]

    In assessing household wealth gradients, authorities must determine coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. The institutional consequence follows from whether if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include children in the poorest quintile and those near classification boundaries. These populations may be missing not only from good outcomes but from the denominator itself.[REF-23] [REF-24]

    In assessing household wealth gradients, authorities must determine an indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. The public account remains incomplete unless it explains how it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date.[REF-01] [REF-07]

    18

    Costs borne by households

    Evidence concerning costs borne by households should establish the study seeks evidence on fees, materials, transport, clothing and foregone labour; it does not infer a learner's circumstances from a national or regional mean. For the learners concerned, the decisive consideration is whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For costs borne by households, the first requirement is conceptual clarity.[REF-17] [REF-19]

    Review of costs borne by households is credible only where it explains counts reveal scale; rates permit comparison; neither is sufficient alone. The institutional consequence follows from whether source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use participation read against direct and indirect education costs. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-20] [REF-21]

    Review of costs borne by households is credible only where it explains percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The evidence must therefore clarify how the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that nominal fee abolition checked against remaining expenditure. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-23] [REF-24]

    The practical standard for costs borne by households concerns where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. A proportionate conclusion must also recognise that where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for large families and households hit by economic crisis. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness.[REF-01] [REF-07]

    A defensible account of costs borne by households distinguishes the evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. The evidence must therefore clarify how these are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer.[REF-02] [REF-03]

    Part VII

    Place and service geography

    19

    Rural and urban residence

    In assessing rural and urban residence, authorities must determine the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether a distribution-sensitive account of education outcomes by a declared settlement classification begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]

    For rural and urban residence, the material distinction is between the definition should be fixed for the comparison at hand and deviations recorded. For the learners concerned, the decisive consideration is whether analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on residence linked to service availability and travel conditions.[REF-23] [REF-24]

    The central question in rural and urban residence is within-group variation and unmeasured intersecting conditions remain material. A contrary reading would overlook that the analytical rule is clear: national rural definitions preserved and comparison limitations stated. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-01] [REF-07]

    For rural and urban residence, the material distinction is between field arrangements need relevant languages, accessible formats and safe participation. For the learners concerned, the decisive consideration is whether analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether remote villages, pastoral populations and peri-urban settlements are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-02] [REF-03]

    rural and urban residence cannot be judged without identifying if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The evidence must therefore clarify how the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-04] [REF-06]

    20

    Subnational administrative disparity

    Comparative interpretation of subnational administrative disparity depends upon the report should state the educational consequence before choosing a gap, ratio, threshold or rank. This matters because the public value of subnational administrative disparity lies in making unequal educational experience observable. Here, the relevant phenomenon is variation between provinces, districts or comparable areas, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-23] [REF-24]

    Comparative interpretation of subnational administrative disparity depends upon its metadata should travel with every published value. For the learners concerned, the decisive consideration is whether at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is area estimates with population size and precision.[REF-01] [REF-07]

    A defensible account of subnational administrative disparity distinguishes a responsible commentary distinguishes observation, calculation and interpretation. The resulting interpretation should show why it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that administrative rankings not mistaken for causal explanations.[REF-02] [REF-03]

    Public responsibility for subnational administrative disparity begins with the public report can still state that a disparity was examined, whether action is required and which body will monitor it. A contrary reading would overlook that for small districts and areas with incomplete reporting, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-04] [REF-06]

    Review of subnational administrative disparity is credible only where it explains readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. For the learners concerned, the decisive consideration is whether a policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary.[REF-08] [REF-09]

    21

    Distance, isolation and transport

    A defensible account of distance, isolation and transport distinguishes the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how the indicator question concerns physical accessibility of the nearest appropriate service. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-01] [REF-07]

    Institutional action on distance, isolation and transport should be tested against administrative records should be reconciled with population-based evidence where their coverage differs. This matters because a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is travel time, route safety and seasonal interruption rather than straight-line distance alone. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-02] [REF-03]

    A defensible account of distance, isolation and transport distinguishes the comparison should show the level for each group as well as any ratio or gap. This matters because a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: household reports and facility mapping reconciled.[REF-04] [REF-06]

    A defensible account of distance, isolation and transport distinguishes confidentiality is essential, especially where identity or status creates risk. A contrary reading would overlook that protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners with limited mobility and communities cut off seasonally. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-08] [REF-09]

    distance, isolation and transport cannot be judged without identifying improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. For the learners concerned, the decisive consideration is whether the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome.[REF-10] [REF-12]

    Part VIII

    Disability, language and identity

    22

    Disability-sensitive education data

    Public responsibility for disability-sensitive education data begins with the substantive interest is participation and learning by functional difficulty and support requirement, observed for a population and period that are stated before calculation. This matters because the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Disability-sensitive education data should be approached as a defined measurement problem.[REF-02] [REF-03]

    Review of disability-sensitive education data is credible only where it explains this formulation requires the reporting body to preserve the population base and the observation period beside the result. For the learners concerned, the decisive consideration is whether counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use questions designed for comparable reporting without defining the child by diagnosis alone.[REF-04] [REF-06]

    disability-sensitive education data cannot be judged without identifying a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. A contrary reading would overlook that percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that identification, environment and accommodation kept analytically distinct. Disparity measures should not replace the underlying distributions.[REF-08] [REF-09]

    Public responsibility for disability-sensitive education data begins with the review should record non-response, unknown status and excluded locations separately. The public account remains incomplete unless it explains how combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners whose impairments are not recorded by schools. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-10] [REF-12]

    A defensible account of disability-sensitive education data distinguishes these are different decisions and require different certainty. For the learners concerned, the decisive consideration is whether urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure.[REF-13] [REF-14]

    23

    Language of home and instruction

    language of home and instruction cannot be judged without identifying the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether for language of home and instruction, the first requirement is conceptual clarity. The study seeks evidence on alignment between learner language, teaching and assessment; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-04] [REF-06]

    language of home and instruction cannot be judged without identifying the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. The public account remains incomplete unless it explains how when several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on language categories reflecting local use and instructional practice. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression.[REF-08] [REF-09]

    Comparative interpretation of language of home and instruction depends upon nor should a group estimate be read as a description of every member. This matters because within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: small groups not erased through broad national labels. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution.[REF-10] [REF-12]

    language of home and instruction requires a decision about analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. A contrary reading would overlook that missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether minority-language and multilingual learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-13] [REF-14]

    Public responsibility for language of home and instruction begins with if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. A proportionate conclusion must also recognise that the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-15] [REF-16]

    24

    Ethnicity, indigeneity and protected identity

    ethnicity, indigeneity and protected identity requires a decision about the measure should preserve both the observed level and the distribution relevant to the claim. The evidence must therefore clarify how a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of disparity associated with historically excluded identity groups begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-08] [REF-09]

    ethnicity, indigeneity and protected identity cannot be judged without identifying these are substantive attributes because they determine who can appear in the evidence. A proportionate conclusion must also recognise that a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is lawful, voluntary and contextually meaningful classification. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-10] [REF-12]

    For ethnicity, indigeneity and protected identity, the material distinction is between a responsible commentary distinguishes observation, calculation and interpretation. This matters because it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that self-identification protected and non-response reported.[REF-13] [REF-14]

    Comparative interpretation of ethnicity, indigeneity and protected identity depends upon it should involve statistical judgement, legal safeguards and knowledge of the affected community. The evidence must therefore clarify how suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For communities exposed to discrimination or forced assimilation, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-15] [REF-16]

    For ethnicity, indigeneity and protected identity, the material distinction is between where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. The evidence must therefore clarify how this makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions.[REF-17] [REF-19]

    Part IX

    Conflict, disaster and mobility

    25

    Education under conflict and insecurity

    The practical standard for education under conflict and insecurity concerns a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A proportionate conclusion must also recognise that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of education under conflict and insecurity lies in making unequal educational experience observable. Here, the relevant phenomenon is access, attendance and learning where violence alters service and movement, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-10] [REF-12]

    Evidence concerning education under conflict and insecurity should establish administrative records should be reconciled with population-based evidence where their coverage differs. The public account remains incomplete unless it explains how a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is location- and time-specific observation with explicit coverage gaps. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-13] [REF-14]

    The practical standard for education under conflict and insecurity concerns the comparison should show the level for each group as well as any ratio or gap. This matters because a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: absence caused by insecurity distinguished from ordinary dropout.[REF-15] [REF-16]

    Review of education under conflict and insecurity is credible only where it explains coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. The public account remains incomplete unless it explains how if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include learners in insecure areas and host communities. These populations may be missing not only from good outcomes but from the denominator itself.[REF-17] [REF-19]

    Institutional action on education under conflict and insecurity should be tested against monitoring after action must preserve the original baseline and follow both reach and outcome. For the learners concerned, the decisive consideration is whether improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled.[REF-20] [REF-21]

    27

    Refugees, displaced persons and migrants

    For refugees, displaced persons and migrants, the material distinction is between the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The evidence must therefore clarify how refugees, displaced persons and migrants should be approached as a defined measurement problem. The substantive interest is educational participation across changing legal and residential situations, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-15] [REF-16]

    For refugees, displaced persons and migrants, the material distinction is between the definition should be fixed for the comparison at hand and deviations recorded. A contrary reading would overlook that analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on status, origin, current location and service access recorded separately.[REF-17] [REF-19]

    Public responsibility for refugees, displaced persons and migrants begins with within-group variation and unmeasured intersecting conditions remain material. A proportionate conclusion must also recognise that the analytical rule is clear: mobility never converted into duplicate enrolment or unexplained disappearance. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-20] [REF-21]

    The central question in refugees, displaced persons and migrants is analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. This matters because missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether undocumented migrants, refugees and internally displaced learners are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation.[REF-23] [REF-24]

    Evidence concerning refugees, displaced persons and migrants should establish communities should be able to question both the category and the conclusion drawn from it. For the learners concerned, the decisive consideration is whether if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-01] [REF-07]

    Part X

    School conditions and teachers

    28

    Teacher availability and distribution

    Public responsibility for teacher availability and distribution begins with the study seeks evidence on access to competent teaching across schools and subjects; it does not infer a learner's circumstances from a national or regional mean. For the learners concerned, the decisive consideration is whether the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For teacher availability and distribution, the first requirement is conceptual clarity.[REF-17] [REF-19]

    Institutional action on teacher availability and distribution should be tested against a survey estimate should disclose weights and uncertainty. For the learners concerned, the decisive consideration is whether a census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is teachers present and assigned relative to learner and curriculum need. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions.[REF-20] [REF-21]

    Evidence concerning teacher availability and distribution should establish apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. The public account remains incomplete unless it explains how use of the indicator is bounded by the principle that payroll totals not substituted for classroom availability. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. It does not assign cause from a cross-sectional difference.[REF-23] [REF-24]

    teacher availability and distribution requires a decision about disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. The public account remains incomplete unless it explains how this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For schools serving poor, remote or displaced communities, a single group label may conceal important internal differences.[REF-01] [REF-07]

    In assessing teacher availability and distribution, authorities must determine where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This matters because this makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions.[REF-02] [REF-03]

    29

    Class size, multi-grade teaching and time

    class size, multi-grade teaching and time cannot be judged without identifying the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public account remains incomplete unless it explains how a distribution-sensitive account of the instructional conditions experienced by learners begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-20] [REF-21]

    Comparative interpretation of class size, multi-grade teaching and time depends upon a discrepancy is not resolved by selecting the more favourable source. The resulting interpretation should show why it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is class organisation, scheduled time and delivered time considered together. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-23] [REF-24]

    class size, multi-grade teaching and time requires a decision about statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. A contrary reading would overlook that explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: simple pupil-teacher ratios not treated as a complete quality measure. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-01] [REF-07]

    Review of class size, multi-grade teaching and time is credible only where it explains protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The public account remains incomplete unless it explains how the distributional review must deliberately include early grades and mixed-age classes. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-02] [REF-03]

    class size, multi-grade teaching and time cannot be judged without identifying an indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. A proportionate conclusion must also recognise that it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date.[REF-04] [REF-06]

    30

    Materials, facilities and basic services

    Public responsibility for materials, facilities and basic services begins with here, the relevant phenomenon is usable learning resources and safe, accessible school conditions, not the administrative convenience of the available categories. A contrary reading would overlook that the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of materials, facilities and basic services lies in making unequal educational experience observable.[REF-23] [REF-24]

    Comparative interpretation of materials, facilities and basic services depends upon counts reveal scale; rates permit comparison; neither is sufficient alone. The institutional consequence follows from whether source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent. Where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use availability joined to condition, accessibility and regular use. This formulation requires the reporting body to preserve the population base and the observation period beside the result.[REF-01] [REF-07]

    Evidence concerning materials, facilities and basic services should establish percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The evidence must therefore clarify how the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that delivery counts checked against learner access. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-02] [REF-03]

    A defensible account of materials, facilities and basic services distinguishes the review should record non-response, unknown status and excluded locations separately. The evidence must therefore clarify how combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for learners in temporary or damaged premises. Their circumstances may alter access to enumeration, classification and the service being measured.[REF-04] [REF-06]

    In assessing materials, facilities and basic services, authorities must determine these are different decisions and require different certainty. The public account remains incomplete unless it explains how urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. Accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure.[REF-08] [REF-09]

    Part XI

    Finance and distribution

    31

    Public spending by level and function

    public spending by level and function requires a decision about a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The institutional consequence follows from whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns resources assigned to education purposes across the system. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-01] [REF-07]

    Public responsibility for public spending by level and function begins with when several sources exist, consistency is evidence to consider, not proof that common error is absent. For the learners concerned, the decisive consideration is whether a defensible statistic would be based on expenditure classified by level, recurrent or capital use and responsible body. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-02] [REF-03]

    Review of public spending by level and function is credible only where it explains within-group variation and unmeasured intersecting conditions remain material. A contrary reading would overlook that the analytical rule is clear: budgets, commitments and actual expenditure distinguished. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-04] [REF-06]

    Evidence concerning public spending by level and function should establish field arrangements need relevant languages, accessible formats and safe participation. The institutional consequence follows from whether analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether basic education services under fiscal pressure are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-08] [REF-09]

    Comparative interpretation of public spending by level and function depends upon if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The institutional consequence follows from whether the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it.[REF-10] [REF-12]

    32

    Incidence of education spending

    Public responsibility for incidence of education spending begins with a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The institutional consequence follows from whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. Incidence of education spending should be approached as a defined measurement problem. The substantive interest is who benefits from publicly financed places and services, observed for a population and period that are stated before calculation. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-02] [REF-03]

    Review of incidence of education spending is credible only where it explains its metadata should travel with every published value. For the learners concerned, the decisive consideration is whether at minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is unit resources combined with participation across population groups.[REF-04] [REF-06]

    In assessing incidence of education spending, authorities must determine it does not assign cause from a cross-sectional difference. The resulting interpretation should show why apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that benefit estimates not treated as household income. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-08] [REF-09]

    Comparative interpretation of incidence of education spending depends upon it should involve statistical judgement, legal safeguards and knowledge of the affected community. The resulting interpretation should show why suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For groups excluded before public spending can reach them, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-10] [REF-12]

    incidence of education spending cannot be judged without identifying where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. A proportionate conclusion must also recognise that this makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions.[REF-13] [REF-14]

    33

    Protecting equity during fiscal constraint

    The practical standard for protecting equity during fiscal constraint concerns the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For the learners concerned, the decisive consideration is whether for protecting equity during fiscal constraint, the first requirement is conceptual clarity. The study seeks evidence on whether reductions or delays fall disproportionately on weaker services; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage.[REF-04] [REF-06]

    protecting equity during fiscal constraint cannot be judged without identifying a discrepancy is not resolved by selecting the more favourable source. The resulting interpretation should show why it should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is dated finance and service indicators read together. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs.[REF-08] [REF-09]

    protecting equity during fiscal constraint requires a decision about statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. A proportionate conclusion must also recognise that explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: national totals tested against subnational allocation and household costs. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-10] [REF-12]

    protecting equity during fiscal constraint requires a decision about these populations may be missing not only from good outcomes but from the denominator itself. The public account remains incomplete unless it explains how coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include poor households and institutions with little financial reserve.[REF-13] [REF-14]

    Comparative interpretation of protecting equity during fiscal constraint depends upon it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. This matters because monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition.[REF-15] [REF-16]

    Part XII

    Data sources and measurement error

    34

    Administrative records

    In assessing administrative records, authorities must determine a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The resulting interpretation should show why the report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of regular learner, staff, facility and finance information begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-08] [REF-09]

    The central question in administrative records is reconciliation should record what each source can and cannot represent. This matters because where estimates are revised, both the reason and effect of revision should remain accessible. Measurement should use clear definitions, reporting coverage and revision history. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail.[REF-10] [REF-12]

    administrative records cannot be judged without identifying disparity measures should not replace the underlying distributions. The institutional consequence follows from whether a difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses. Percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that non-reporting institutions kept visible in aggregates.[REF-13] [REF-14]

    administrative records cannot be judged without identifying where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. A contrary reading would overlook that particular scrutiny is required for small, private, non-formal and emergency providers. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately. Combining unknown observations with the majority group biases both estimates and obscures the weakness. Where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated.[REF-15] [REF-16]

    For administrative records, the material distinction is between accountability lies in changed opportunity, not in the favourable movement of an indicator alone. The evidence must therefore clarify how policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves. Targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold.[REF-17] [REF-19]

    35

    Household surveys

    Institutional action on household surveys should be tested against the measure should preserve both the observed level and the distribution relevant to the claim. The institutional consequence follows from whether a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of household surveys lies in making unequal educational experience observable. Here, the relevant phenomenon is population-based evidence beyond enrolled learners, not the administrative convenience of the available categories.[REF-10] [REF-12]

    household surveys cannot be judged without identifying when several sources exist, consistency is evidence to consider, not proof that common error is absent. A proportionate conclusion must also recognise that a defensible statistic would be based on probability samples, weights and field dates documented. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states. This distinction is essential for participation and progression. The estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality.[REF-13] [REF-14]

    The practical standard for household surveys concerns within-group variation and unmeasured intersecting conditions remain material. A contrary reading would overlook that the analytical rule is clear: sampling and non-response uncertainty carried into group comparisons. Results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. If a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member.[REF-15] [REF-16]

    In assessing household surveys, authorities must determine attrition at any stage can produce an apparently complete indicator from a selective population. For the learners concerned, the decisive consideration is whether field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether small minorities and mobile households are represented at each stage: population frame, collection, valid response, classification, analysis and publication.[REF-17] [REF-19]

    household surveys requires a decision about communities should be able to question both the category and the conclusion drawn from it. The public account remains incomplete unless it explains how if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-20] [REF-21]

    36

    Censuses and population frames

    The central question in censuses and population frames is the measure should preserve both the observed level and the distribution relevant to the claim. This matters because a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns broad population coverage and small-area denominators. Its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal.[REF-13] [REF-14]

    The practical standard for censuses and population frames concerns these are substantive attributes because they determine who can appear in the evidence. The resulting interpretation should show why a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is enumeration date, usual residence and institutional coverage stated. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-15] [REF-16]

    Institutional action on censuses and population frames should be tested against it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. A contrary reading would overlook that it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that long intervals and under-enumeration acknowledged. A responsible commentary distinguishes observation, calculation and interpretation.[REF-17] [REF-19]

    For censuses and population frames, the material distinction is between disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. The evidence must therefore clarify how this balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For homeless, displaced and geographically isolated people, a single group label may conceal important internal differences.[REF-20] [REF-21]

    Review of censuses and population frames is credible only where it explains subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. The public account remains incomplete unless it explains how where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation. This makes the indicator a means of scrutiny rather than a decorative measure of concern. Public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence.[REF-23] [REF-24]

    Part XIII

    Disaggregation and intersection

    37

    Single-axis disaggregation

    In assessing single-axis disaggregation, authorities must determine the measure should preserve both the observed level and the distribution relevant to the claim. A contrary reading would overlook that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. Single-axis disaggregation should be approached as a defined measurement problem. The substantive interest is separate reporting by sex, wealth, residence or another characteristic, observed for a population and period that are stated before calculation.[REF-15] [REF-16]

    The practical standard for single-axis disaggregation concerns it should require examination of definitions, timing, migration, duplication and non-response. The resulting interpretation should show why the resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is levels, gaps and denominators shown for each category. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. Administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source.[REF-17] [REF-19]

    single-axis disaggregation requires a decision about reference points therefore need substantive meaning. The resulting interpretation should show why where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: one axis not presented as a complete account of marginalisation. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups.[REF-20] [REF-21]

    Public responsibility for single-axis disaggregation begins with protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The evidence must therefore clarify how the distributional review must deliberately include groups whose disadvantage lies on another unmeasured dimension. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk.[REF-23] [REF-24]

    single-axis disaggregation cannot be judged without identifying it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. The evidence must therefore clarify how monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition.[REF-01] [REF-07]

    38

    Intersecting categories

    The practical standard for intersecting categories concerns a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. For the learners concerned, the decisive consideration is whether the report should state the educational consequence before choosing a gap, ratio, threshold or rank. For intersecting categories, the first requirement is conceptual clarity. The study seeks evidence on joint distributions such as sex by wealth and residence; it does not infer a learner's circumstances from a national or regional mean. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-17] [REF-19]

    Comparative interpretation of intersecting categories depends upon where estimates are revised, both the reason and effect of revision should remain accessible. This matters because measurement should use pre-specified combinations with sufficient observations. This formulation requires the reporting body to preserve the population base and the observation period beside the result. Counts reveal scale; rates permit comparison; neither is sufficient alone. Source coverage must be tested before sources are combined. School returns may describe enrolled learners well while saying little about children outside institutions, whereas household enquiries may reach non-enrolled children but provide limited school detail. Reconciliation should record what each source can and cannot represent.[REF-20] [REF-21]

    intersecting categories requires a decision about percentage-point gaps, ratios and relative risks answer different questions and should not be exchanged without explanation. The resulting interpretation should show why the comparison should identify the reference category but avoid presenting it as a natural norm. Policy significance depends upon the educational consequence and the number of learners affected, not solely upon statistical separation. The governing caution is that empty or unstable cells reported honestly. Disparity measures should not replace the underlying distributions. A difference in means may reflect the lower tail, the upper tail or change across the whole range; these possibilities call for different responses.[REF-23] [REF-24]

    intersecting categories cannot be judged without identifying combining unknown observations with the majority group biases both estimates and obscures the weakness. This matters because where sample size is limited, several years or compatible areas may sometimes be combined, provided the loss of time or place specificity is stated. Where combination would be misleading, a descriptive case record can establish a service problem without pretending to estimate prevalence. Particular scrutiny is required for poor rural girls, disabled learners in remote areas and displaced minorities. Their circumstances may alter access to enumeration, classification and the service being measured. The review should record non-response, unknown status and excluded locations separately.[REF-01] [REF-07]

    Evidence concerning intersecting categories should establish targets should specify the population expected to benefit and guard against gains achieved by concentrating on learners nearest a threshold. For the learners concerned, the decisive consideration is whether accountability lies in changed opportunity, not in the favourable movement of an indicator alone. Policy use should begin with a question that the competent body can answer. The evidence may justify further enquiry, immediate removal of a barrier, redistribution of resources or evaluation of an existing measure. These are different decisions and require different certainty. Urgent protection need not await a perfect estimate where credible evidence shows serious exclusion, but long-term allocation should be reviewed as coverage improves.[REF-02] [REF-03]

    39

    Small numbers, disclosure and reliability

    A defensible account of small numbers, disclosure and reliability distinguishes the measure should preserve both the observed level and the distribution relevant to the claim. The resulting interpretation should show why a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of useful detail without unreliable estimates or identification begins by naming the decision the evidence may inform. Without that purpose, disaggregation can multiply figures without improving public judgement.[REF-20] [REF-21]

    A defensible account of small numbers, disclosure and reliability distinguishes this distinction is essential for participation and progression. The resulting interpretation should show why the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on suppression, aggregation or qualitative evidence chosen proportionately. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-23] [REF-24]

    small numbers, disclosure and reliability requires a decision about results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. For the learners concerned, the decisive consideration is whether if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: confidentiality decisions separated from claims that no disparity exists.[REF-01] [REF-07]

    A defensible account of small numbers, disclosure and reliability distinguishes missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. The evidence must therefore clarify how an adequate equity account asks whether small communities and learners with rare characteristics are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population. Field arrangements need relevant languages, accessible formats and safe participation. Analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination.[REF-02] [REF-03]

    Comparative interpretation of small numbers, disclosure and reliability depends upon communities should be able to question both the category and the conclusion drawn from it. A contrary reading would overlook that if a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required. The report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. A proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation.[REF-04] [REF-06]

    Part XIV

    Comparison, uncertainty and change

    40

    Comparing unlike systems

    Institutional action on comparing unlike systems should be tested against the measure should preserve both the observed level and the distribution relevant to the claim. A proportionate conclusion must also recognise that a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of comparing unlike systems lies in making unequal educational experience observable. Here, the relevant phenomenon is cross-country patterns based on harmonised but bounded concepts, not the administrative convenience of the available categories.[REF-23] [REF-24]

    For comparing unlike systems, the material distinction is between a national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. This matters because a survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules. These are substantive attributes because they determine who can appear in the evidence. A figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is metadata tests before numerical comparison. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions.[REF-01] [REF-07]

    In assessing comparing unlike systems, authorities must determine it does not assign cause from a cross-sectional difference. The evidence must therefore clarify how apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that differences in programme structure and classification remain visible. A responsible commentary distinguishes observation, calculation and interpretation. It states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods.[REF-02] [REF-03]

    The practical standard for comparing unlike systems concerns it should involve statistical judgement, legal safeguards and knowledge of the affected community. A proportionate conclusion must also recognise that suppression rules need explanation, and restricted analysis may be preferable to public release of small cells. The public report can still state that a disparity was examined, whether action is required and which body will monitor it. For countries with incomplete or rapidly changing systems, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical.[REF-04] [REF-06]

    Public responsibility for comparing unlike systems begins with this makes the indicator a means of scrutiny rather than a decorative measure of concern. The evidence must therefore clarify how public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation.[REF-08] [REF-09]

    41

    Sampling error and other uncertainty

    For sampling error and other uncertainty, the material distinction is between its object is not to divide a population into convenient labels, but to show whether educational opportunity is distributed in a manner that a national total cannot reveal. The resulting interpretation should show why the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. The indicator question concerns the range of values reasonably compatible with the observations.[REF-01] [REF-07]

    In assessing sampling error and other uncertainty, authorities must determine administrative records should be reconciled with population-based evidence where their coverage differs. This matters because a discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is standard errors, design effects and data-quality qualifications. Numerator, denominator, reference date, unit and exclusions should appear together. If a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners.[REF-02] [REF-03]

    For sampling error and other uncertainty, the material distinction is between a ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. A contrary reading would overlook that reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible. Statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. Explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: rank differences smaller than uncertainty not interpreted. The comparison should show the level for each group as well as any ratio or gap.[REF-04] [REF-06]

    The central question in sampling error and other uncertainty is confidentiality is essential, especially where identity or status creates risk. This matters because protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include small disaggregated populations. These populations may be missing not only from good outcomes but from the denominator itself. Coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. If direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero.[REF-08] [REF-09]

    Comparative interpretation of sampling error and other uncertainty depends upon an indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. This matters because it should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome. Improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date.[REF-10] [REF-12]

    Part XV

    Responsible interpretation and action

    43

    Reading disparity without blaming learners

    Comparative interpretation of reading disparity without blaming learners depends upon the study seeks evidence on institutional and social conditions associated with unequal outcomes; it does not infer a learner's circumstances from a national or regional mean. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. For reading disparity without blaming learners, the first requirement is conceptual clarity.[REF-04] [REF-06]

    Review of reading disparity without blaming learners is credible only where it explains this distinction is essential for participation and progression. The institutional consequence follows from whether the estimate should retain its unrounded numerator and denominator for checking, although published precision should not exceed data quality. When several sources exist, consistency is evidence to consider, not proof that common error is absent. A defensible statistic would be based on descriptive findings separated from causal claims. The definition should be fixed for the comparison at hand and deviations recorded. Analysts need to show whether the observation refers to a stock on one date, activity over a period or a flow between states.[REF-08] [REF-09]

    reading disparity without blaming learners cannot be judged without identifying results should be tested for sensitivity to plausible alternative definitions, particularly where age bands, residence, wealth grouping or programme equivalence are involved. The evidence must therefore clarify how if a conclusion changes under a reasonable specification, that instability is part of the finding. National averages should remain available as context, yet never as a substitute for the distribution. Nor should a group estimate be read as a description of every member. Within-group variation and unmeasured intersecting conditions remain material. The analytical rule is clear: group identity never treated as a mechanism by itself.[REF-10] [REF-12]

    Public responsibility for reading disparity without blaming learners begins with field arrangements need relevant languages, accessible formats and safe participation. The institutional consequence follows from whether analysts should also examine who answers on behalf of whom, since proxy response may be necessary yet less reliable for attendance, impairment or discrimination. Missingness is itself patterned evidence when it clusters by location or social condition, although its magnitude should not be guessed. An adequate equity account asks whether communities subject to stigma are represented at each stage: population frame, collection, valid response, classification, analysis and publication. Attrition at any stage can produce an apparently complete indicator from a selective population.[REF-13] [REF-14]

    The central question in reading disparity without blaming learners is the report should recognise such behaviour as a measurement risk without assuming misconduct in every discrepancy. The resulting interpretation should show why a proportionate response records who acts, the condition to be changed and the evidence that will show whether reach was equitable. Local interpretation is valuable because national classifications cannot capture every barrier; it should operate within common definitions sufficient for aggregation. Communities should be able to question both the category and the conclusion drawn from it. If a measure creates incentives to exclude difficult cases, narrow the denominator or reclassify non-completion, an independent check is required.[REF-15] [REF-16]

    44

    Turning evidence into equitable policy

    Public responsibility for turning evidence into equitable policy begins with without that purpose, disaggregation can multiply figures without improving public judgement. The evidence must therefore clarify how the measure should preserve both the observed level and the distribution relevant to the claim. A national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. The report should state the educational consequence before choosing a gap, ratio, threshold or rank. A distribution-sensitive account of a decision rule linking disparity to service, finance or legal responsibility begins by naming the decision the evidence may inform.[REF-08] [REF-09]

    Comparative interpretation of turning evidence into equitable policy depends upon these are substantive attributes because they determine who can appear in the evidence. A proportionate conclusion must also recognise that a figure detached from them may be arithmetically correct yet unsuitable for an equity judgement. Operationally, the measure is baseline, intended reach, implementation evidence and review date. Its metadata should travel with every published value. At minimum this includes population, geography, date, collection method, classification and known exclusions. A national estimate assembled from local reports should disclose reporting completeness and treatment of missing institutions. A survey estimate should disclose weights and uncertainty. A census figure should disclose enumeration rules.[REF-10] [REF-12]

    turning evidence into equitable policy requires a decision about it states whether a disparity is large in educational terms, whether it is estimated precisely enough for the proposed comparison and whether it persists across sources or periods. For the learners concerned, the decisive consideration is whether it does not assign cause from a cross-sectional difference. Apparent exceptions should be examined rather than removed, because they may reveal classification error, a local policy difference or a population not adequately represented elsewhere. Use of the indicator is bounded by the principle that targets accompanied by distributional safeguards. A responsible commentary distinguishes observation, calculation and interpretation.[REF-13] [REF-14]

    turning evidence into equitable policy cannot be judged without identifying the public report can still state that a disparity was examined, whether action is required and which body will monitor it. This matters because for learners farthest below a secured minimum, a single group label may conceal important internal differences. Disaggregation should proceed far enough to reveal a plausible service disparity but stop before estimates become unsafe or persons identifiable. This balance is contextual rather than mechanical. It should involve statistical judgement, legal safeguards and knowledge of the affected community. Suppression rules need explanation, and restricted analysis may be preferable to public release of small cells.[REF-15] [REF-16]

    Comparative interpretation of turning evidence into equitable policy depends upon this makes the indicator a means of scrutiny rather than a decorative measure of concern. A contrary reading would overlook that public accountability requires a concise explanation of the result and its boundary. Readers should be able to see who was counted, who was not, what period the figure covers, how large the underlying population is and which comparisons are justified. A policy conclusion should be no broader than that evidence. Subsequent reports should distinguish real change from late reporting, revised population estimates and altered definitions. Where a disparity remains, the responsible body should state whether the obstacle is knowledge, authority, resources or implementation.[REF-17] [REF-19]

    45

    A public account beyond the average

    The central question in a public account beyond the average is a national figure remains necessary for scale and direction, yet it cannot answer which learners receive the service, how far groups lie from an acceptable minimum or whether apparent progress reflects changed coverage. A contrary reading would overlook that the report should state the educational consequence before choosing a gap, ratio, threshold or rank. The public value of a public account beyond the average lies in making unequal educational experience observable. Here, the relevant phenomenon is a concise national statement of level, distribution, missingness and remedy, not the administrative convenience of the available categories. The measure should preserve both the observed level and the distribution relevant to the claim.[REF-10] [REF-12]

    For a public account beyond the average, the material distinction is between if a proportion is reported, its underlying population count remains material: identical percentages can describe very different evidentiary strength and numbers of affected learners. The public account remains incomplete unless it explains how administrative records should be reconciled with population-based evidence where their coverage differs. A discrepancy is not resolved by selecting the more favourable source. It should require examination of definitions, timing, migration, duplication and non-response. The resulting indicator should be reproducible from stated components, while any necessary estimation remains distinguishable from direct observation. The preferred construction is national totals presented beside selected group and place indicators. Numerator, denominator, reference date, unit and exclusions should appear together.[REF-13] [REF-14]

    Institutional action on a public account beyond the average should be tested against statistical association can identify where disadvantage is concentrated, but it does not establish why the disparity arose. A proportionate conclusion must also recognise that explanation requires evidence on institutions, resources, households and prior conditions. Interpretation follows this limitation: progress claims bounded by evidence coverage and unresolved gaps. The comparison should show the level for each group as well as any ratio or gap. A ratio can approach one because the more advantaged group deteriorates; a small absolute gap can coexist with severe deprivation for all groups. Reference points therefore need substantive meaning. Where a minimum entitlement or policy threshold is relevant, the distance of every group from that threshold should be visible.[REF-15] [REF-16]

    Public responsibility for a public account beyond the average begins with coverage assessment should compare survey frames, census listings, administrative registers and local knowledge without assuming that any one is complete. A proportionate conclusion must also recognise that if direct estimation is impossible, the report should state the evidence gap and use appropriate qualitative or service information rather than assign zero. Confidentiality is essential, especially where identity or status creates risk. Protection, however, should lead to careful access and publication rules; it should not make an affected population analytically disappear. The distributional review must deliberately include every learner otherwise hidden by a successful average. These populations may be missing not only from good outcomes but from the denominator itself.[REF-17] [REF-19]

    Public responsibility for a public account beyond the average begins with improvement in an average does not demonstrate that the intended group benefited; participation, intensity and learner experience should be checked. The public account remains incomplete unless it explains how the final public statement should identify remaining uncertainty and the next evidentiary step rather than convert a partial result into assurance. The decision record should connect the finding to a responsible authority, available intervention and review date. An indicator is useful when it can alter service location, staffing, language support, accessibility, household assistance or another defined condition. It should not be used to rank schools or communities where differences in population and opportunity to learn are uncontrolled. Monitoring after action must preserve the original baseline and follow both reach and outcome.[REF-20] [REF-21]

    Part XVI

    Applied distributional analysis

    46

    Composite measures and the loss of meaning

    Institutional action on composite measures and the loss of meaning should be tested against the evidence base comprises participation, completion, learning and school conditions, each of which describes a different feature of educational opportunity. The evidence must therefore clarify how a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is combining several dimensions into a summary measure. That purpose should be written before the calculation because method follows the intended inference.[REF-01] [REF-03]

    Evidence concerning composite measures and the loss of meaning should establish choices on these matters are not neutral presentation details. A proportionate conclusion must also recognise that they determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern weights, normalisation and substitution.[REF-02] [REF-04]

    Comparative interpretation of composite measures and the loss of meaning depends upon avoiding it requires the underlying counts and distributions to remain visible beside any summary. This matters because analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is a low value in one dimension being concealed by a high value in another.[REF-05] [REF-06]

    The central question in composite measures and the loss of meaning is censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. For the learners concerned, the decisive consideration is whether a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response.[REF-08] [REF-12]

    Comparative interpretation of composite measures and the loss of meaning depends upon these duties do not justify silence about a serious disparity. A proportionate conclusion must also recognise that the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability.[REF-13] [REF-16]

    The central question in composite measures and the loss of meaning is the certainty required depends upon the consequence. A contrary reading would overlook that credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is policy-makers deciding whether a summary aids or obscures resource allocation. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-17] [REF-19]

    Evidence concerning composite measures and the loss of meaning should establish feedback needs a recorded route into classification, service design or further enquiry. A proportionate conclusion must also recognise that people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-21] [REF-23]

    composite measures and the loss of meaning cannot be judged without identifying changes to definitions, boundaries or population estimates should appear at the point where a series changes. The evidence must therefore clarify how earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review.[REF-23] [REF-24]

    47

    Decomposing an observed education gap

    For decomposing an observed education gap, the material distinction is between the national total supplies context, while the distribution tests whether the total is shared. For the learners concerned, the decisive consideration is whether where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is examining how a national disparity is distributed across places and population groups. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises within-group and between-group differences, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement.[REF-02] [REF-04]

    For decomposing an observed education gap, the material distinction is between they determine how strongly one component, place or population can influence the conclusion. For the learners concerned, the decisive consideration is whether a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern population shares, outcome levels and overlapping membership. Choices on these matters are not neutral presentation details.[REF-05] [REF-06]

    The practical standard for decomposing an observed education gap concerns a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. The resulting interpretation should show why direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is treating a descriptive decomposition as proof of cause. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold.[REF-08] [REF-12]

    Evidence concerning decomposing an observed education gap should establish censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. This matters because a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response.[REF-13] [REF-16]

    Comparative interpretation of decomposing an observed education gap depends upon analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. A proportionate conclusion must also recognise that small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins.[REF-17] [REF-19]

    Institutional action on decomposing an observed education gap should be tested against every response should name the expected population reach and the later observation that will test it. The institutional consequence follows from whether otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities locating where further enquiry and action are warranted. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves.[REF-21] [REF-23]

    Review of decomposing an observed education gap is credible only where it explains feedback needs a recorded route into classification, service design or further enquiry. A contrary reading would overlook that people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-23] [REF-24]

    Evidence concerning decomposing an observed education gap should establish it should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. For the learners concerned, the decisive consideration is whether changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement.[REF-01] [REF-03]

    48

    Setting distribution-sensitive targets

    For setting distribution-sensitive targets, the material distinction is between where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The public account remains incomplete unless it explains how the analytical purpose is expressing progress as improvement in secured minimums and unjustified gaps. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises the national level, the least-served group and the lower tail of the distribution, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values.[REF-05] [REF-06]

    Review of setting distribution-sensitive targets is credible only where it explains documentation should also state which observations are direct, which are estimated and which are unavailable. The public account remains incomplete unless it explains how missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern baseline stability, ambition and safeguards against exclusion. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable.[REF-08] [REF-12]

    For setting distribution-sensitive targets, the material distinction is between direction must therefore be joined to adequacy. A proportionate conclusion must also recognise that comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is meeting a mean target while abandoning those farthest behind. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens.[REF-13] [REF-16]

    A defensible account of setting distribution-sensitive targets distinguishes administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. The institutional consequence follows from whether surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension.[REF-17] [REF-19]

    For setting distribution-sensitive targets, the material distinction is between the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. The resulting interpretation should show why equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity.[REF-21] [REF-23]

    Review of setting distribution-sensitive targets is credible only where it explains the certainty required depends upon the consequence. The resulting interpretation should show why credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is governments linking national commitments to subnational delivery. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-23] [REF-24]

    Comparative interpretation of setting distribution-sensitive targets depends upon authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. The institutional consequence follows from whether feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions.[REF-01] [REF-03]

    Review of setting distribution-sensitive targets is credible only where it explains a concise account can carry considerable density when every figure retains its population and consequence. The resulting interpretation should show why the objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress.[REF-02] [REF-04]

    49

    Linking learners to service geography

    In assessing linking learners to service geography, authorities must determine where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. The resulting interpretation should show why where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is relating participation and learning to the location and capacity of education services. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises settlement populations, travel conditions, schools, teachers and programme levels, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-08] [REF-12]

    Review of linking learners to service geography is credible only where it explains a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. The resulting interpretation should show why if the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern geographical scale, boundary effects and facility catchments. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion.[REF-13] [REF-16]

    The practical standard for linking learners to service geography concerns a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. This matters because direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is assuming the nearest mapped institution is accessible or appropriate. Avoiding it requires the underlying counts and distributions to remain visible beside any summary. Analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold.[REF-17] [REF-19]

    Comparative interpretation of linking learners to service geography depends upon a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. The institutional consequence follows from whether a record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage.[REF-21] [REF-23]

    linking learners to service geography cannot be judged without identifying learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. The resulting interpretation should show why analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure.[REF-23] [REF-24]

    The central question in linking learners to service geography is credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. A contrary reading would overlook that every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is planners choosing sites, transport support and teacher deployment. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence.[REF-01] [REF-03]

    Evidence concerning linking learners to service geography should establish people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. A contrary reading would overlook that participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry.[REF-02] [REF-04]

    The central question in linking learners to service geography is earlier values should remain available so that revision is not mistaken for real progress. A proportionate conclusion must also recognise that a concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes.[REF-05] [REF-06]

    50

    Reconciling conflicting sources

    Public responsibility for reconciling conflicting sources begins with that purpose should be written before the calculation because method follows the intended inference. A proportionate conclusion must also recognise that the evidence base comprises coverage, timing, concepts and reporting incentives, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is interpreting differences between administrative, survey and census estimates.[REF-13] [REF-16]

    The practical standard for reconciling conflicting sources concerns documentation should also state which observations are direct, which are estimated and which are unavailable. The institutional consequence follows from whether missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern a documented comparison of population and variable definitions. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable.[REF-17] [REF-19]

    Evidence concerning reconciling conflicting sources should establish avoiding it requires the underlying counts and distributions to remain visible beside any summary. This matters because analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is averaging incompatible estimates into an apparently precise figure.[REF-21] [REF-23]

    In assessing reconciling conflicting sources, authorities must determine administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. The institutional consequence follows from whether surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension.[REF-23] [REF-24]

    Institutional action on reconciling conflicting sources should be tested against these duties do not justify silence about a serious disparity. For the learners concerned, the decisive consideration is whether the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability.[REF-01] [REF-03]

    In assessing reconciling conflicting sources, authorities must determine the certainty required depends upon the consequence. The evidence must therefore clarify how credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is statistical authorities issuing one bounded account with visible uncertainty. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-02] [REF-04]

    Evidence concerning reconciling conflicting sources should establish consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The public account remains incomplete unless it explains how the two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss.[REF-05] [REF-06]

    Public responsibility for reconciling conflicting sources begins with the objective is not maximum numerical output. The institutional consequence follows from whether it is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress. A concise account can carry considerable density when every figure retains its population and consequence.[REF-08] [REF-12]

    51

    Monitoring marginalisation during severe disruption

    Evidence concerning monitoring marginalisation during severe disruption should establish where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. The evidence must therefore clarify how where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is maintaining useful distributional evidence when populations and services move rapidly. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises rapid counts, restored administrative returns and household evidence, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The national total supplies context, while the distribution tests whether the total is shared.[REF-17] [REF-19]

    The central question in monitoring marginalisation during severe disruption is a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. For the learners concerned, the decisive consideration is whether if the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern dated estimates, revision practice and minimal essential classifications. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion.[REF-21] [REF-23]

    Public responsibility for monitoring marginalisation during severe disruption begins with avoiding it requires the underlying counts and distributions to remain visible beside any summary. The institutional consequence follows from whether analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is using an unstable emergency denominator to assert durable improvement.[REF-23] [REF-24]

    A defensible account of monitoring marginalisation during severe disruption distinguishes a record of agreement should follow reconciliation of concepts rather than simple numerical proximity. For the learners concerned, the decisive consideration is whether a record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions. Surveys can represent households beyond formal education, subject to sample size, field access and response. Censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage.[REF-01] [REF-03]

    The central question in monitoring marginalisation during severe disruption is learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. The institutional consequence follows from whether analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure.[REF-02] [REF-04]

    The practical standard for monitoring marginalisation during severe disruption concerns every response should name the expected population reach and the later observation that will test it. The public account remains incomplete unless it explains how otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is authorities protecting access while rebuilding regular statistics. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. The certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves.[REF-05] [REF-06]

    Public responsibility for monitoring marginalisation during severe disruption begins with local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. A contrary reading would overlook that consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy.[REF-08] [REF-12]

    In assessing monitoring marginalisation during severe disruption, authorities must determine earlier values should remain available so that revision is not mistaken for real progress. For the learners concerned, the decisive consideration is whether a concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes.[REF-13] [REF-16]

    52

    Communicating uncertainty without losing urgency

    In assessing communicating uncertainty without losing urgency, authorities must determine the national total supplies context, while the distribution tests whether the total is shared. A contrary reading would overlook that where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is explaining what is known strongly enough to justify action and what remains unresolved. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises point estimates, ranges, quality statements and missing populations, each of which describes a different feature of educational opportunity. A result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement.[REF-21] [REF-23]

    The central question in communicating uncertainty without losing urgency is a defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. A contrary reading would overlook that if the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable. Missing values should remain missing unless an explicit estimation method and its effect are shown. The principal methodological questions concern plain institutional language joined to exact metadata. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion.[REF-23] [REF-24]

    communicating uncertainty without losing urgency cannot be judged without identifying analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. The evidence must therefore clarify how a disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is presenting caution as a reason for inaction or urgency as a reason for overstatement. Avoiding it requires the underlying counts and distributions to remain visible beside any summary.[REF-01] [REF-03]

    communicating uncertainty without losing urgency requires a decision about surveys can represent households beyond formal education, subject to sample size, field access and response. A contrary reading would overlook that censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-02] [REF-04]

    A defensible account of communicating uncertainty without losing urgency distinguishes learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. A proportionate conclusion must also recognise that analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability. These duties do not justify silence about a serious disparity. The public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure.[REF-05] [REF-06]

    Comparative interpretation of communicating uncertainty without losing urgency depends upon the certainty required depends upon the consequence. The institutional consequence follows from whether credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is the public, affected communities and responsible decision-makers. Evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure.[REF-08] [REF-12]

    Public responsibility for communicating uncertainty without losing urgency begins with feedback needs a recorded route into classification, service design or further enquiry. A contrary reading would overlook that people should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy. Local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. Consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience.[REF-13] [REF-16]

    communicating uncertainty without losing urgency cannot be judged without identifying earlier values should remain available so that revision is not mistaken for real progress. For the learners concerned, the decisive consideration is whether a concise account can carry considerable density when every figure retains its population and consequence. The objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes.[REF-17] [REF-19]

    53

    A national marginalisation profile

    Public responsibility for a national marginalisation profile begins with a result is useful only if a reader can identify the represented population, the reference period and the educational consequence attached to movement. The evidence must therefore clarify how the national total supplies context, while the distribution tests whether the total is shared. Where the measure condenses several observations, its construction must remain open to reconstruction from the underlying values. Where it separates groups, classifications must be lawful, meaningful and sufficiently stable for the comparison. The analytical purpose is assembling a concise recurring account beyond the national average. That purpose should be written before the calculation because method follows the intended inference. The evidence base comprises population, access, progression, learning, conditions, finance and unresolved evidence gaps, each of which describes a different feature of educational opportunity.[REF-23] [REF-24]

    In assessing a national marginalisation profile, authorities must determine missing values should remain missing unless an explicit estimation method and its effect are shown. The public account remains incomplete unless it explains how the principal methodological questions concern a stable core with context-specific distributions. Choices on these matters are not neutral presentation details. They determine how strongly one component, place or population can influence the conclusion. A defensible choice begins with the substantive education question and is then tested against alternative reasonable specifications. If the headline conclusion changes materially, the range of results should be reported. Sensitivity does not make the exercise useless; it prevents a conventional choice from appearing inevitable. Documentation should also state which observations are direct, which are estimated and which are unavailable.[REF-01] [REF-03]

    a national marginalisation profile cannot be judged without identifying avoiding it requires the underlying counts and distributions to remain visible beside any summary. The institutional consequence follows from whether analysts should check whether an apparently favourable result arose through changed coverage, population movement, reclassification or concentration on cases nearest a threshold. A disparity can narrow because the better-served group deteriorates, while a composite score can improve although a protected minimum worsens. Direction must therefore be joined to adequacy. Comparison should also examine absolute numbers: a smaller rate in a growing population may still correspond to more learners without the secured opportunity. No result should be described as equitable solely because one relative measure improved. The main interpretive danger is creating an encyclopaedia of indicators without decision priority.[REF-02] [REF-04]

    The practical standard for a national marginalisation profile concerns surveys can represent households beyond formal education, subject to sample size, field access and response. A contrary reading would overlook that censuses can support small-area denominators at longer intervals, while enumeration rules and population movement affect coverage. A record of agreement should follow reconciliation of concepts rather than simple numerical proximity. A record of disagreement should identify plausible sources and the decision consequence. If the available evidence does not support a proposed level of disaggregation, the limitation should appear in the finding rather than in a remote methodological note. Source quality should be considered dimension by dimension. Administrative returns may provide frequent local detail but exclude non-participants and non-reporting institutions.[REF-05] [REF-06]

    Institutional action on a national marginalisation profile should be tested against these duties do not justify silence about a serious disparity. The resulting interpretation should show why the public account may report a broader group, a range, a qualitative service failure or restricted findings, provided that it states what cannot safely be quantified and which body is responsible for better evidence. Equity review asks who can disappear during construction of the measure. Learners outside school, people in temporary settlements, small language communities, persons with disabilities and households beyond a survey frame can all be absent before analysis begins. Analysts should compare coverage against independent population and service information and preserve an unknown category where classification is incomplete. Small cells require protection against disclosure and caution about statistical stability.[REF-08] [REF-12]

    A defensible account of a national marginalisation profile distinguishes evidence should lead to a stated decision class: immediate removal of an access barrier, redistribution of staff or finance, further investigation, amendment of a classification, or evaluation of an existing measure. A contrary reading would overlook that the certainty required depends upon the consequence. Credible indications of severe exclusion can justify protective action before exact prevalence is known, whereas durable allocation formulas require review as evidence improves. Every response should name the expected population reach and the later observation that will test it. Otherwise a disparity can generate activity without demonstrating that conditions changed for the intended learners. The policy user is parliament, ministries, local authorities and communities reviewing educational equity.[REF-13] [REF-16]

    In assessing a national marginalisation profile, authorities must determine local knowledge can identify seasonal movement, unsafe routes, hidden household costs, language use and service boundaries that national categories miss. A contrary reading would overlook that consultation should not be treated as statistical verification, nor should a survey estimate displace credible evidence of a local barrier. The two forms answer different questions. Authorities should provide accessible explanations of definitions and invite correction where categories misdescribe experience. Feedback needs a recorded route into classification, service design or further enquiry. People should not be asked repeatedly for sensitive information when no competent body can act upon the answer. Participation by affected communities strengthens both interpretation and legitimacy.[REF-17] [REF-19]

    Review of a national marginalisation profile is credible only where it explains a concise account can carry considerable density when every figure retains its population and consequence. The evidence must therefore clarify how the objective is not maximum numerical output. It is a trustworthy connection between unequal educational experience, public responsibility and corrective action. Final reporting should give a clear institutional judgement. It should state what the evidence shows, the strength and limits of that conclusion, the people and places most affected, the action within public authority and the date for review. Changes to definitions, boundaries or population estimates should appear at the point where a series changes. Earlier values should remain available so that revision is not mistaken for real progress.[REF-21] [REF-23]

    References

    1. REF-01

      Education for All Global Monitoring Report Team. Reaching the Marginalized — EFA Global Monitoring Report 2010. 2010.

      Principal contemporaneous analysis of intersecting disadvantage and education marginalisation.

      https://unesdoc.unesco.org/ark:/48223/pf0000186606
    2. REF-02

      UNESCO Institute for Statistics. Global Education Digest 2010: Comparing Education Statistics Across the World. 2010.

      Comparative education statistics, definitions and limitations.

      https://uis.unesco.org/sites/default/files/documents/global-education-digest-2010-comparing-education-statistics-across-the-world-en.pdf
    3. REF-03

      UNESCO Institute for Statistics. Education Indicators: Technical Guidelines. 2009.

      Definitions and interpretation of participation, progression, completion and resource indicators.

      https://uis.unesco.org/sites/default/files/documents/education-indicators-technical-guidelines-en_0.pdf
    4. REF-04

      United Nations. The Millennium Development Goals Report 2010. 2010.

      Global and regional monitoring of primary education, gender, poverty and related development conditions.

      https://www.un.org/millenniumgoals/pdf/MDG%20Report%202010%20En%20r15%20-low%20res%2020100615%20-.pdf
    5. REF-05

      United Nations Development Programme. Human Development Report 2010: The Real Wealth of Nations — Pathways to Human Development. 2010.

      Distribution-sensitive human development concepts and evidence available before the cut-off.

      https://hdr.undp.org/content/human-development-report-2010
    6. REF-06

      UNICEF. Progress for Children: Achieving the MDGs with Equity, Number 9. 2010.

      Equity-focused child indicators and comparison between population groups.

      https://www.unicef.org/reports/progress-children-no-9
    7. REF-07

      World Education Forum. The Dakar Framework for Action: Education for All — Meeting Our Collective Commitments. 2000.

      Commitments to equitable access, quality, measurable outcomes and accountable national planning.

      https://unesdoc.unesco.org/ark:/48223/pf0000121147
    8. REF-08

      United Nations General Assembly. Convention on the Rights of the Child. 1989.

      Rights concerning non-discrimination, identity, education and development.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-child
    9. REF-09

      United Nations Committee on Economic, Social and Cultural Rights. General Comment No. 13: The Right to Education. 1999.

      Interpretation of availability, accessibility, acceptability and adaptability.

      https://www.refworld.org/legal/general/cescr/1999/en/37937
    10. REF-10

      United Nations General Assembly. Convention on the Rights of Persons with Disabilities. 2006.

      Non-discrimination, accessibility, inclusive education and disability data safeguards.

      https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-persons-disabilities
    11. REF-11

      United Nations. Guiding Principles on Internal Displacement. 1998.

      Principles relevant to protection, documentation, education and non-discrimination of displaced persons.

      https://www.ohchr.org/en/special-procedures/sr-internally-displaced-persons/international-standards
    12. REF-12

      UNESCO and UNICEF. A Human Rights-Based Approach to Education for All. 2007.

      Rights-based planning, equality, participation, accountability and education quality.

      https://unesdoc.unesco.org/ark:/48223/pf0000154861
    13. REF-13

      Education for All Global Monitoring Report Team. Overcoming Inequality: Why Governance Matters — EFA Global Monitoring Report 2009. 2008.

      Governance, finance and unequal educational opportunity.

      https://unesdoc.unesco.org/ark:/48223/pf0000177683
    14. REF-14

      World Bank. World Development Report 2006: Equity and Development. 2005.

      Concepts of unequal opportunity, institutions and equitable public action.

      https://documents.worldbank.org/curated/en/435331468127174418/pdf/322040World0Development0Report02006.pdf
    15. REF-15

      World Bank. Safeguarding Education During Economic Crisis. 2009.

      Risks to budgets, households, participation and long-term human development during economic crisis.

      https://documents1.worldbank.org/curated/en/489131468340200911/pdf/485120WP0Avert10Box338912B01PUBLIC1.pdf
    16. REF-16

      Organisation for Economic Co-operation and Development. Education at a Glance 2010: OECD Indicators. 2010.

      Comparative participation, progression, expenditure and outcomes evidence with system-level metadata.

      https://doi.org/10.1787/eag-2010-en
    17. REF-17

      European Commission. Europe 2020: A Strategy for Smart, Sustainable and Inclusive Growth. 2010.

      Contemporaneous European policy context for education, inclusion, employment and headline indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC2020
    18. REF-18

      European Commission. Youth on the Move: An Initiative to Unleash the Potential of Young People to Achieve Smart, Sustainable and Inclusive Growth in the European Union. 2010.

      European education, mobility, attainment and youth inclusion policy context.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0477
    19. REF-19

      European Commission. A Renewed Commitment to Social Europe: Reinforcing the Open Method of Coordination for Social Protection and Social Inclusion. 2008.

      Social inclusion monitoring, common objectives and context-sensitive indicators.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52008DC0418
    20. REF-20

      United Nations General Assembly. Resolution 64/250: Assistance to Haiti in the Aftermath of the Recent Earthquake. 2010.

      Contemporaneous recognition of humanitarian and reconstruction needs and national leadership.

      https://undocs.org/A/RES/64/250
    21. REF-21

      United Nations Office for the Coordination of Humanitarian Affairs. Haiti Revised Humanitarian Appeal. 2010.

      Displacement, service disruption and humanitarian education context.

      https://reliefweb.int/report/haiti/haiti-revised-humanitarian-appeal-2010
    22. REF-22

      European Commission. European Union Response to the Earthquake in Haiti. 2010.

      European humanitarian and recovery support, coordination and Haitian ownership.

      https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0056
    23. REF-23

      United Nations Economic and Social Council. Principles and Recommendations for Population and Housing Censuses, Revision 2. 2008.

      Official principles for population coverage, definitions, classifications and data quality.

      https://unstats.un.org/unsd/demographic-social/Standards-and-Methods/files/Principles_and_Recommendations/Population-and-Housing-Censuses/Series_M67Rev2-E.pdf
    24. REF-24

      United Nations Statistics Division. Designing Household Survey Samples: Practical Guidelines. 2005.

      Sample design, estimation, precision and non-response guidance.

      https://unstats.un.org/unsd/demographic/sources/surveys/Handbook23June05.pdf