ICEQC-R-2014-11 — Evidence and Attribution in Public Claims of Educational Improvement cover

Informe de Investigación Temática

ICEQC-R-2014-11 — Evidence and Attribution in Public Claims of Educational Improvement

A global comparative indicator study of change, contribution and responsible public interpretation

Fecha de publicación
Categoría de investigación
Investigación de datos e indicadores
Informe arquetipo
Estudio comparativo de indicadores
Ámbito geográfico
Global
Fecha límite para la presentación de pruebas
Organismo responsable
Dirección de Investigación y Políticas de ICEQC
ICEQC-R-2014-11 — Evidence and Attribution in Public Claims of Educational Improvement cover

Publication record

This is the controlled English edition. Evidence and institutional status are stated as at the evidence cut-off date.

Executive summary

Public claims of educational improvement influence confidence, finance, institutional judgement and policy continuation. Such claims should identify the outcome, population, period, baseline and evidence and should distinguish observed change, contribution and causal attribution.

Before-and-after improvement may reflect the intervention, wider trends, changed participation, measurement differences or concurrent policy. Conversely, imperfect causal evidence does not make all interpretation impossible. A disciplined contribution account can examine implementation, sequence, mechanism, alternative explanation, distribution and uncertainty.

This report examines fourteen domains through tests of definition, source, design, comparability, uncertainty, equity, implementation, alternative explanation, proportionality and correction. It connects the wording of a public claim with the strength of evidence supporting it.

The central conclusion is that claim precision is a public-interest safeguard. Results can be communicated clearly without overstating scope or causation, and correction should reach the audiences that received or relied on the original claim.

Key findings

  • Improvement claims should state outcome, population, period and comparison.
  • Announced intervention and verified implementation are different evidence.
  • Before-and-after change does not alone establish causal attribution.
  • Timing should support the proposed sequence from action to outcome.
  • Participation, attrition and missing populations can alter apparent results.
  • Average gains should be examined for distribution and equity.
  • Mechanism evidence can strengthen a bounded contribution account.
  • Pilots should not be described as durable system-wide results without evidence.
  • Sensitivity to reasonable definitions and models should be visible.
  • Public corrections should address the practical effect of the earlier claim.

Scope and method

Part I

Claim and decision purpose

1

Claim proposition

The relevant claim concerns the exact improvement, population, period and public decision represented by the statement. The principal risk is that broad language can move beyond the measured result and make correction difficult. Authorities should state the claim in a form that evidence can support or refute.[REF-05]

2

Definition

For claim and decision purpose, a public improvement claim depends on the exact improvement, population, period and public decision represented by the statement. The principal risk is that broad language can move beyond the measured result and make correction difficult. Evidence producers should therefore state the claim in a form that evidence can support or refute.

The definition test asks which intervention, outcome, population and period the statement concerns. Authorities should publish an exact claim and unit. The governing proposition is that improvement is not a self-defining measure. Every claim should identify outcome, population, period, source and material limitation.

Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses.

3

Evidence source

Evidence source is material because the exact improvement, population, period and public decision represented by the statement cannot be inferred from a favourable trend or completed activity. In this domain, broad language can move beyond the measured result and make correction difficult. The immediate safeguard is to state the claim in a form that evidence can support or refute.

A defensible account should establish which data and records support implementation and outcome and should identify provenance, coverage, date and quality, recognising that multiple reports from one source are not independent corroboration. Evidence inconsistent with the preferred explanation should remain visible.

Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

4

Design

The public-interest meaning of claim and decision purpose concerns the exact improvement, population, period and public decision represented by the statement. If broad language can move beyond the measured result and make correction difficult, a credible result can be converted into an unsupported institutional claim. The required course is to state the claim in a form that evidence can support or refute.

Under design, review concerns which comparison and identification strategy support attribution. Public bodies should explain assumptions and threats, subject to the rule that no design eliminates the need for contextual judgement. Reasons for selecting the measure and comparison should be stated.

Subgroup and local evidence should be capable of qualifying a headline average. Improvement should not be claimed universally where material populations deteriorated or remained unobserved.

5

Comparability

Analysis of claim and decision purpose should begin with the exact improvement, population, period and public decision represented by the statement. A foreseeable failure arises where broad language can move beyond the measured result and make correction difficult. Authorities should state the claim in a form that evidence can support or refute.[REF-10]

The comparability standard requires consideration of whether measures, populations and conditions are stable enough across observations. Analysts should document breaks and adjustments, because common labels can conceal changed instruments or participation. Revisions and breaks should be adjacent to the affected comparison.

The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.

6

Uncertainty

Claim and decision purpose requires explicit metadata because it concerns the exact improvement, population, period and public decision represented by the statement. The risk is that broad language can move beyond the measured result and make correction difficult. A competent producer should state the claim in a form that evidence can support or refute.

Review of uncertainty should determine which sampling, measurement, model and missing-data limits apply. It should publish intervals and qualitative constraints and acknowledge that precision of language should not exceed precision of evidence. Personal information should remain protected while unequal effect remains visible.

A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear.

7

Equity

The public purpose of analysis on claim and decision purpose is a responsible account of the exact improvement, population, period and public decision represented by the statement. Where broad language can move beyond the measured result and make correction difficult, public confidence and resource allocation can be distorted. Authorities should state the claim in a form that evidence can support or refute.

For equity, bodies should identify which populations received the intervention and benefit and should analyse distribution and access responsibly. Any conclusion must respect that average improvement does not establish equitable improvement.

The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes.

8

Implementation

implementation cannot be judged without identifying the implication for claim strength and correction should be recorded. A contrary reading would overlook that applied to claim and decision purpose, this requirement has a distinct attribution consequence. For claim and decision purpose, a public improvement claim depends on the exact improvement, population, period and public decision represented by the statement. The principal risk is that broad language can move beyond the measured result and make correction difficult. Evidence producers should therefore state the claim in a form that evidence can support or refute.

The implementation test asks whether intended activity reached the population at sufficient quality and intensity. Authorities should link finance, delivery, use and service condition. The governing proposition is that an announced or funded programme is not evidence of exposure. Every claim should identify outcome, population, period, source and material limitation.

In assessing implementation, authorities must determine the implication for claim strength and correction should be recorded. The institutional consequence follows from whether applied to claim and decision purpose, this requirement has a distinct attribution consequence. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses.

9

Alternative explanation

Alternative explanation is material because the exact improvement, population, period and public decision represented by the statement cannot be inferred from a favourable trend or completed activity. In this domain, broad language can move beyond the measured result and make correction difficult. The immediate safeguard is to state the claim in a form that evidence can support or refute.[REF-12]

A defensible account should establish which other changes could plausibly produce the finding and should test or acknowledge concurrent factors, recognising that chronology and association alone do not establish causation. Evidence inconsistent with the preferred explanation should remain visible.

For alternative explanation, the material distinction is between where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The resulting interpretation should show why the implication for claim strength and correction should be recorded. Applied to claim and decision purpose, this requirement has a distinct attribution consequence.

10

Proportionality

Public responsibility for proportionality begins with if broad language can move beyond the measured result and make correction difficult, a credible result can be converted into an unsupported institutional claim. A contrary reading would overlook that the required course is to state the claim in a form that evidence can support or refute. The implication for claim strength and correction should be recorded. Applied to claim and decision purpose, this requirement has a distinct attribution consequence. The public-interest meaning of claim and decision purpose concerns the exact improvement, population, period and public decision represented by the statement.

Under proportionality, review concerns how strength of evidence relates to the consequence and certainty of the claim. Public bodies should use bounded language and distinguish contribution from attribution, subject to the rule that higher-stakes claims require stronger support and clearer qualification. Reasons for selecting the measure and comparison should be stated.

For proportionality, the material distinction is between improvement should not be claimed universally where material populations deteriorated or remained unobserved. The evidence must therefore clarify how review of proportionality is credible only where it explains the implication for claim strength and correction should be recorded. The public account remains incomplete unless it explains how applied to claim and decision purpose, this requirement has a distinct attribution consequence. Subgroup and local evidence should be capable of qualifying a headline average.

11

Review and correction

review and correction cannot be judged without identifying the implication for claim strength and correction should be recorded. A contrary reading would overlook that applied to claim and decision purpose, this requirement has a distinct attribution consequence. Analysis of claim and decision purpose should begin with the exact improvement, population, period and public decision represented by the statement. A foreseeable failure arises where broad language can move beyond the measured result and make correction difficult. Authorities should state the claim in a form that evidence can support or refute.

The review and correction standard requires consideration of how new evidence, error or changed method revises the claim. Analysts should retain versions, reasons and direct correction, because quiet technical amendment is insufficient where the public relied on a stronger statement. Revisions and breaks should be adjacent to the affected comparison.

In assessing review and correction, authorities must determine the implication for claim strength and correction should be recorded. The public account remains incomplete unless it explains how applied to claim and decision purpose, this requirement has a distinct attribution consequence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.[REF-70]

Part II

Baseline and reference point

12

Claim proposition

The relevant claim concerns the prior condition against which change is judged. The principal risk is that a favourable endpoint can be selected without a comparable starting population or method. Authorities should define baseline, source, date and any break in series.[REF-10]

13

Definition

Baseline and reference point requires explicit metadata because it concerns the prior condition against which change is judged. The risk is that a favourable endpoint can be selected without a comparable starting population or method. A competent producer should define baseline, source, date and any break in series.

Review of definition should determine which intervention, outcome, population and period the statement concerns. It should publish an exact claim and unit and acknowledge that improvement is not a self-defining measure. Personal information should remain protected while unequal effect remains visible.

Comparative interpretation of definition depends upon responsible communication preserves value by making the evidentiary boundary clear. For the learners concerned, the decisive consideration is whether the implication for claim strength and correction should be recorded. Applied to baseline and reference point, this requirement has a distinct attribution consequence. A result can be important without supporting every explanation attached to it.

14

Evidence source

The public purpose of analysis on baseline and reference point is a responsible account of the prior condition against which change is judged. Where a favourable endpoint can be selected without a comparable starting population or method, public confidence and resource allocation can be distorted. Authorities should define baseline, source, date and any break in series.

For evidence source, bodies should identify which data and records support implementation and outcome and should identify provenance, coverage, date and quality. Any conclusion must respect that multiple reports from one source are not independent corroboration.

Review of evidence source is credible only where it explains the implication for claim strength and correction should be recorded. For the learners concerned, the decisive consideration is whether applied to baseline and reference point, this requirement has a distinct attribution consequence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes.

15

Design

For baseline and reference point, a public improvement claim depends on the prior condition against which change is judged. The principal risk is that a favourable endpoint can be selected without a comparable starting population or method. Evidence producers should therefore define baseline, source, date and any break in series.

The design test asks which comparison and identification strategy support attribution. Authorities should explain assumptions and threats. The governing proposition is that no design eliminates the need for contextual judgement. Every claim should identify outcome, population, period, source and material limitation.[REF-71]

A defensible account of design distinguishes public language should state which one the evidence addresses. The resulting interpretation should show why any departure should be disclosed with its likely effect on inference. Within baseline and reference point, the principle bears on the claim described in this part. Observed change, programme contribution and sole causal attribution are different propositions.

16

Comparability

Comparability is material because the prior condition against which change is judged cannot be inferred from a favourable trend or completed activity. In this domain, a favourable endpoint can be selected without a comparable starting population or method. The immediate safeguard is to define baseline, source, date and any break in series.

A defensible account should establish whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments, recognising that common labels can conceal changed instruments or participation. Evidence inconsistent with the preferred explanation should remain visible.

The practical standard for comparability concerns where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. For the learners concerned, the decisive consideration is whether any departure should be disclosed with its likely effect on inference. Within baseline and reference point, the principle bears on the claim described in this part.

17

Uncertainty

The public-interest meaning of baseline and reference point concerns the prior condition against which change is judged. If a favourable endpoint can be selected without a comparable starting population or method, a credible result can be converted into an unsupported institutional claim. The required course is to define baseline, source, date and any break in series.

Under uncertainty, review concerns which sampling, measurement, model and missing-data limits apply. Public bodies should publish intervals and qualitative constraints, subject to the rule that precision of language should not exceed precision of evidence. Reasons for selecting the measure and comparison should be stated.

The practical standard for uncertainty concerns subgroup and local evidence should be capable of qualifying a headline average. This matters because improvement should not be claimed universally where material populations deteriorated or remained unobserved. Any departure should be disclosed with its likely effect on inference. Within baseline and reference point, the principle bears on the claim described in this part.

18

Equity

Analysis of baseline and reference point should begin with the prior condition against which change is judged. A foreseeable failure arises where a favourable endpoint can be selected without a comparable starting population or method. Authorities should define baseline, source, date and any break in series.[REF-10]

The equity standard requires consideration of which populations received the intervention and benefit. Analysts should analyse distribution and access responsibly, because average improvement does not establish equitable improvement. Revisions and breaks should be adjacent to the affected comparison.

For equity, the material distinction is between any departure should be disclosed with its likely effect on inference. A proportionate conclusion must also recognise that within baseline and reference point, the principle bears on the claim described in this part. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.

19

Implementation

Evidence concerning implementation should establish baseline and reference point requires explicit metadata because it concerns the prior condition against which change is judged. A proportionate conclusion must also recognise that the risk is that a favourable endpoint can be selected without a comparable starting population or method. A competent producer should define baseline, source, date and any break in series. The implication for claim strength and correction should be recorded. Applied to baseline and reference point, this requirement has a distinct attribution consequence.

Review of implementation should determine whether intended activity reached the population at sufficient quality and intensity. It should link finance, delivery, use and service condition and acknowledge that an announced or funded programme is not evidence of exposure. Personal information should remain protected while unequal effect remains visible.

Evidence concerning implementation should establish a result can be important without supporting every explanation attached to it. A proportionate conclusion must also recognise that responsible communication preserves value by making the evidentiary boundary clear. Any departure should be disclosed with its likely effect on inference. Within baseline and reference point, the principle bears on the claim described in this part.

20

Alternative explanation

Evidence concerning alternative explanation should establish where a favourable endpoint can be selected without a comparable starting population or method, public confidence and resource allocation can be distorted. The public account remains incomplete unless it explains how authorities should define baseline, source, date and any break in series. The implication for claim strength and correction should be recorded. Applied to baseline and reference point, this requirement has a distinct attribution consequence. The public purpose of analysis on baseline and reference point is a responsible account of the prior condition against which change is judged.

For alternative explanation, bodies should identify which other changes could plausibly produce the finding and should test or acknowledge concurrent factors. Any conclusion must respect that chronology and association alone do not establish causation.

Review of alternative explanation is credible only where it explains claim strength should be revised as evidence changes. A proportionate conclusion must also recognise that any departure should be disclosed with its likely effect on inference. Within baseline and reference point, the principle bears on the claim described in this part. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion.

21

Proportionality

Public responsibility for proportionality begins with the principal risk is that a favourable endpoint can be selected without a comparable starting population or method. This matters because evidence producers should therefore define baseline, source, date and any break in series. The implication for claim strength and correction should be recorded. Applied to baseline and reference point, this requirement has a distinct attribution consequence. For baseline and reference point, a public improvement claim depends on the prior condition against which change is judged.[REF-73]

The proportionality test asks how strength of evidence relates to the consequence and certainty of the claim. Authorities should use bounded language and distinguish contribution from attribution. The governing proposition is that higher-stakes claims require stronger support and clearer qualification. Every claim should identify outcome, population, period, source and material limitation.

For baseline and reference point, analysts should test this safeguard against the design and population evidence. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses. The conclusion should retain the material limitation and responsible body.

22

Review and correction

Review and correction is material because the prior condition against which change is judged cannot be inferred from a favourable trend or completed activity. In this domain, a favourable endpoint can be selected without a comparable starting population or method. The immediate safeguard is to define baseline, source, date and any break in series.

A defensible account should establish how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction, recognising that quiet technical amendment is insufficient where the public relied on a stronger statement. Evidence inconsistent with the preferred explanation should remain visible.

For baseline and reference point, analysts should test this safeguard against the design and population evidence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The conclusion should retain the material limitation and responsible body.

Part III

Intervention and implementation

23

Claim proposition

The relevant claim concerns the policy, programme or institutional action alleged to have contributed to change. The principal risk is that announced design or expenditure can be treated as delivered exposure. Authorities should verify implementation, reach, intensity and material variation.[REF-12]

24

Definition

The public-interest meaning of intervention and implementation concerns the policy, programme or institutional action alleged to have contributed to change. If announced design or expenditure can be treated as delivered exposure, a credible result can be converted into an unsupported institutional claim. The required course is to verify implementation, reach, intensity and material variation.

Under definition, review concerns which intervention, outcome, population and period the statement concerns. Public bodies should publish an exact claim and unit, subject to the rule that improvement is not a self-defining measure. Reasons for selecting the measure and comparison should be stated.

Public responsibility for definition begins with the conclusion should retain the material limitation and responsible body. For the learners concerned, the decisive consideration is whether for intervention and implementation, analysts should test this safeguard against the design and population evidence. Subgroup and local evidence should be capable of qualifying a headline average. Improvement should not be claimed universally where material populations deteriorated or remained unobserved.[REF-05]

25

Evidence source

Analysis of intervention and implementation should begin with the policy, programme or institutional action alleged to have contributed to change. A foreseeable failure arises where announced design or expenditure can be treated as delivered exposure. Authorities should verify implementation, reach, intensity and material variation.

The evidence source standard requires consideration of which data and records support implementation and outcome. Analysts should identify provenance, coverage, date and quality, because multiple reports from one source are not independent corroboration. Revisions and breaks should be adjacent to the affected comparison.

Institutional action on evidence source should be tested against these conditions support different decisions and should not be collapsed into one judgement. The resulting interpretation should show why the conclusion should retain the material limitation and responsible body. For intervention and implementation, analysts should test this safeguard against the design and population evidence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence.

26

Design

Intervention and implementation requires explicit metadata because it concerns the policy, programme or institutional action alleged to have contributed to change. The risk is that announced design or expenditure can be treated as delivered exposure. A competent producer should verify implementation, reach, intensity and material variation.

Review of design should determine which comparison and identification strategy support attribution. It should explain assumptions and threats and acknowledge that no design eliminates the need for contextual judgement. Personal information should remain protected while unequal effect remains visible.

design requires a decision about the conclusion should retain the material limitation and responsible body. This matters because for intervention and implementation, analysts should test this safeguard against the design and population evidence. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear.

27

Comparability

The public purpose of analysis on intervention and implementation is a responsible account of the policy, programme or institutional action alleged to have contributed to change. Where announced design or expenditure can be treated as delivered exposure, public confidence and resource allocation can be distorted. Authorities should verify implementation, reach, intensity and material variation.

For comparability, bodies should identify whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments. Any conclusion must respect that common labels can conceal changed instruments or participation.[REF-10]

Comparative interpretation of comparability depends upon the conclusion should retain the material limitation and responsible body. For the learners concerned, the decisive consideration is whether for intervention and implementation, analysts should test this safeguard against the design and population evidence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes.

28

Uncertainty

For intervention and implementation, a public improvement claim depends on the policy, programme or institutional action alleged to have contributed to change. The principal risk is that announced design or expenditure can be treated as delivered exposure. Evidence producers should therefore verify implementation, reach, intensity and material variation.

The uncertainty test asks which sampling, measurement, model and missing-data limits apply. Authorities should publish intervals and qualitative constraints. The governing proposition is that precision of language should not exceed precision of evidence. Every claim should identify outcome, population, period, source and material limitation.

Review of uncertainty is credible only where it explains observed change, programme contribution and sole causal attribution are different propositions. The institutional consequence follows from whether public language should state which one the evidence addresses. The implication for claim strength and correction should be recorded. Applied to intervention and implementation, this requirement has a distinct attribution consequence.

29

Equity

Equity is material because the policy, programme or institutional action alleged to have contributed to change cannot be inferred from a favourable trend or completed activity. In this domain, announced design or expenditure can be treated as delivered exposure. The immediate safeguard is to verify implementation, reach, intensity and material variation.

A defensible account should establish which populations received the intervention and benefit and should analyse distribution and access responsibly, recognising that average improvement does not establish equitable improvement. Evidence inconsistent with the preferred explanation should remain visible.

The practical standard for equity concerns the implication for claim strength and correction should be recorded. The public account remains incomplete unless it explains how applied to intervention and implementation, this requirement has a distinct attribution consequence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

30

Implementation

Public responsibility for implementation begins with the public-interest meaning of intervention and implementation concerns the policy, programme or institutional action alleged to have contributed to change. For the learners concerned, the decisive consideration is whether if announced design or expenditure can be treated as delivered exposure, a credible result can be converted into an unsupported institutional claim. The required course is to verify implementation, reach, intensity and material variation. The implication for claim strength and correction should be recorded. Applied to intervention and implementation, this requirement has a distinct attribution consequence.[REF-12]

Under implementation, review concerns whether intended activity reached the population at sufficient quality and intensity. Public bodies should link finance, delivery, use and service condition, subject to the rule that an announced or funded programme is not evidence of exposure. Reasons for selecting the measure and comparison should be stated.

A defensible account of implementation distinguishes subgroup and local evidence should be capable of qualifying a headline average. A contrary reading would overlook that improvement should not be claimed universally where material populations deteriorated or remained unobserved. The implication for claim strength and correction should be recorded. Applied to intervention and implementation, this requirement has a distinct attribution consequence.

31

Alternative explanation

In assessing alternative explanation, authorities must determine the implication for claim strength and correction should be recorded. For the learners concerned, the decisive consideration is whether applied to intervention and implementation, this requirement has a distinct attribution consequence. Analysis of intervention and implementation should begin with the policy, programme or institutional action alleged to have contributed to change. A foreseeable failure arises where announced design or expenditure can be treated as delivered exposure. Authorities should verify implementation, reach, intensity and material variation.

The alternative explanation standard requires consideration of which other changes could plausibly produce the finding. Analysts should test or acknowledge concurrent factors, because chronology and association alone do not establish causation. Revisions and breaks should be adjacent to the affected comparison.

Review of alternative explanation is credible only where it explains the account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. The resulting interpretation should show why these conditions support different decisions and should not be collapsed into one judgement. The implication for claim strength and correction should be recorded. Applied to intervention and implementation, this requirement has a distinct attribution consequence.

32

Proportionality

Evidence concerning proportionality should establish intervention and implementation requires explicit metadata because it concerns the policy, programme or institutional action alleged to have contributed to change. A proportionate conclusion must also recognise that the risk is that announced design or expenditure can be treated as delivered exposure. A competent producer should verify implementation, reach, intensity and material variation. The implication for claim strength and correction should be recorded. Applied to intervention and implementation, this requirement has a distinct attribution consequence.

Review of proportionality should determine how strength of evidence relates to the consequence and certainty of the claim. It should use bounded language and distinguish contribution from attribution and acknowledge that higher-stakes claims require stronger support and clearer qualification. Personal information should remain protected while unequal effect remains visible.

proportionality requires a decision about the implication for claim strength and correction should be recorded. The evidence must therefore clarify how applied to intervention and implementation, this requirement has a distinct attribution consequence. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear.[REF-70]

33

Review and correction

Institutional action on review and correction should be tested against the public purpose of analysis on intervention and implementation is a responsible account of the policy, programme or institutional action alleged to have contributed to change. The evidence must therefore clarify how where announced design or expenditure can be treated as delivered exposure, public confidence and resource allocation can be distorted. Authorities should verify implementation, reach, intensity and material variation. The implication for claim strength and correction should be recorded. Applied to intervention and implementation, this requirement has a distinct attribution consequence.

For review and correction, bodies should identify how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction. Any conclusion must respect that quiet technical amendment is insufficient where the public relied on a stronger statement.

Institutional action on review and correction should be tested against claim strength should be revised as evidence changes. The evidence must therefore clarify how the implication for claim strength and correction should be recorded. Applied to intervention and implementation, this requirement has a distinct attribution consequence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion.

Part IV

Outcome and measurement

34

Claim proposition

The relevant claim concerns the participation, learning, equity or service condition said to have improved. The principal risk is that proxy, output or single indicator can be described as the complete educational outcome. Authorities should define construct, instrument, population and limitations.[REF-15]

35

Definition

For outcome and measurement, a public improvement claim depends on the participation, learning, equity or service condition said to have improved. The principal risk is that proxy, output or single indicator can be described as the complete educational outcome. Evidence producers should therefore define construct, instrument, population and limitations.

Review of definition is credible only where it explains the governing proposition is that improvement is not a self-defining measure. The public account remains incomplete unless it explains how every claim should identify outcome, population, period, source and material limitation. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. The definition test asks which intervention, outcome, population and period the statement concerns. Authorities should publish an exact claim and unit.

Public responsibility for definition begins with observed change, programme contribution and sole causal attribution are different propositions. A contrary reading would overlook that public language should state which one the evidence addresses. Any departure should be disclosed with its likely effect on inference. Within outcome and measurement, the principle bears on the claim described in this part.

36

Evidence source

Evidence source is material because the participation, learning, equity or service condition said to have improved cannot be inferred from a favourable trend or completed activity. In this domain, proxy, output or single indicator can be described as the complete educational outcome. The immediate safeguard is to define construct, instrument, population and limitations.

Review of evidence source is credible only where it explains the implication for claim strength and correction should be recorded. This matters because applied to outcome and measurement, this requirement has a distinct attribution consequence. A defensible account should establish which data and records support implementation and outcome and should identify provenance, coverage, date and quality, recognising that multiple reports from one source are not independent corroboration. Evidence inconsistent with the preferred explanation should remain visible.[REF-71]

A defensible account of evidence source distinguishes where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The institutional consequence follows from whether any departure should be disclosed with its likely effect on inference. Within outcome and measurement, the principle bears on the claim described in this part.

37

Design

The public-interest meaning of outcome and measurement concerns the participation, learning, equity or service condition said to have improved. If proxy, output or single indicator can be described as the complete educational outcome, a credible result can be converted into an unsupported institutional claim. The required course is to define construct, instrument, population and limitations.

The central question in design is public bodies should explain assumptions and threats, subject to the rule that no design eliminates the need for contextual judgement. For the learners concerned, the decisive consideration is whether reasons for selecting the measure and comparison should be stated. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. Under design, review concerns which comparison and identification strategy support attribution.

Institutional action on design should be tested against subgroup and local evidence should be capable of qualifying a headline average. A contrary reading would overlook that improvement should not be claimed universally where material populations deteriorated or remained unobserved. Any departure should be disclosed with its likely effect on inference. Within outcome and measurement, the principle bears on the claim described in this part.

38

Comparability

Analysis of outcome and measurement should begin with the participation, learning, equity or service condition said to have improved. A foreseeable failure arises where proxy, output or single indicator can be described as the complete educational outcome. Authorities should define construct, instrument, population and limitations.

Comparative interpretation of comparability depends upon revisions and breaks should be adjacent to the affected comparison. For the learners concerned, the decisive consideration is whether the implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. The comparability standard requires consideration of whether measures, populations and conditions are stable enough across observations. Analysts should document breaks and adjustments, because common labels can conceal changed instruments or participation.

comparability cannot be judged without identifying any departure should be disclosed with its likely effect on inference. The public account remains incomplete unless it explains how within outcome and measurement, the principle bears on the claim described in this part. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.

39

Uncertainty

Outcome and measurement requires explicit metadata because it concerns the participation, learning, equity or service condition said to have improved. The risk is that proxy, output or single indicator can be described as the complete educational outcome. A competent producer should define construct, instrument, population and limitations.[REF-72]

Comparative interpretation of uncertainty depends upon it should publish intervals and qualitative constraints and acknowledge that precision of language should not exceed precision of evidence. The evidence must therefore clarify how personal information should remain protected while unequal effect remains visible. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. Review of uncertainty should determine which sampling, measurement, model and missing-data limits apply.

uncertainty requires a decision about any departure should be disclosed with its likely effect on inference. A contrary reading would overlook that within outcome and measurement, the principle bears on the claim described in this part. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear.

40

Equity

The public purpose of analysis on outcome and measurement is a responsible account of the participation, learning, equity or service condition said to have improved. Where proxy, output or single indicator can be described as the complete educational outcome, public confidence and resource allocation can be distorted. Authorities should define construct, instrument, population and limitations.

The central question in equity is any conclusion must respect that average improvement does not establish equitable improvement. This matters because the implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. For equity, bodies should identify which populations received the intervention and benefit and should analyse distribution and access responsibly.

equity cannot be judged without identifying the final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. A contrary reading would overlook that claim strength should be revised as evidence changes. Any departure should be disclosed with its likely effect on inference. Within outcome and measurement, the principle bears on the claim described in this part.

41

Implementation

implementation requires a decision about for the learners concerned, the decisive consideration is whether evidence producers should therefore define construct, instrument, population and limitations. The resulting interpretation should show why the implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. For outcome and measurement, a public improvement claim depends on the participation, learning, equity or service condition said to have improved. Review of implementation is credible only where it explains the principal risk is that proxy, output or single indicator can be described as the complete educational outcome.

Evidence concerning implementation should establish for the learners concerned, the decisive consideration is whether authorities should link finance, delivery, use and service condition. This matters because the governing proposition is that an announced or funded programme is not evidence of exposure. Every claim should identify outcome, population, period, source and material limitation. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. Review of implementation is credible only where it explains the implementation test asks whether intended activity reached the population at sufficient quality and intensity.

implementation requires a decision about observed change, programme contribution and sole causal attribution are different propositions. The evidence must therefore clarify how public language should state which one the evidence addresses. The conclusion should retain the material limitation and responsible body. For outcome and measurement, analysts should test this safeguard against the design and population evidence.[REF-73]

42

Alternative explanation

Alternative explanation is material because the participation, learning, equity or service condition said to have improved cannot be inferred from a favourable trend or completed activity. In this domain, proxy, output or single indicator can be described as the complete educational outcome. The immediate safeguard is to define construct, instrument, population and limitations.

The central question in alternative explanation is a defensible account should establish which other changes could plausibly produce the finding and should test or acknowledge concurrent factors, recognising that chronology and association alone do not establish causation. The public account remains incomplete unless it explains how evidence inconsistent with the preferred explanation should remain visible. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence.

The practical standard for alternative explanation concerns the conclusion should retain the material limitation and responsible body. A contrary reading would overlook that for outcome and measurement, analysts should test this safeguard against the design and population evidence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

43

Proportionality

Public responsibility for proportionality begins with if proxy, output or single indicator can be described as the complete educational outcome, a credible result can be converted into an unsupported institutional claim. The resulting interpretation should show why the required course is to define construct, instrument, population and limitations. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. The public-interest meaning of outcome and measurement concerns the participation, learning, equity or service condition said to have improved.

proportionality cannot be judged without identifying public bodies should use bounded language and distinguish contribution from attribution, subject to the rule that higher-stakes claims require stronger support and clearer qualification. For the learners concerned, the decisive consideration is whether reasons for selecting the measure and comparison should be stated. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence. Under proportionality, review concerns how strength of evidence relates to the consequence and certainty of the claim.

For proportionality, the material distinction is between the conclusion should retain the material limitation and responsible body. A contrary reading would overlook that for outcome and measurement, analysts should test this safeguard against the design and population evidence. Subgroup and local evidence should be capable of qualifying a headline average. Improvement should not be claimed universally where material populations deteriorated or remained unobserved.

44

Review and correction

For review and correction, the material distinction is between analysis of outcome and measurement should begin with the participation, learning, equity or service condition said to have improved. For the learners concerned, the decisive consideration is whether a foreseeable failure arises where proxy, output or single indicator can be described as the complete educational outcome. Authorities should define construct, instrument, population and limitations. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence.

Public responsibility for review and correction begins with the review and correction standard requires consideration of how new evidence, error or changed method revises the claim. The public account remains incomplete unless it explains how analysts should retain versions, reasons and direct correction, because quiet technical amendment is insufficient where the public relied on a stronger statement. Revisions and breaks should be adjacent to the affected comparison. The implication for claim strength and correction should be recorded. Applied to outcome and measurement, this requirement has a distinct attribution consequence.[REF-77]

Evidence concerning review and correction should establish the conclusion should retain the material limitation and responsible body. A proportionate conclusion must also recognise that for outcome and measurement, analysts should test this safeguard against the design and population evidence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.

Part V

Comparison and counterfactual

45

Claim proposition

The relevant claim concerns the expected course of events in the absence of the intervention or under an alternative. The principal risk is that before-and-after change can be attributed without considering wider trends or selection. Authorities should use a defensible comparison or state the limits of contribution evidence.[REF-61]

46

Definition

Comparison and counterfactual requires explicit metadata because it concerns the expected course of events in the absence of the intervention or under an alternative. The risk is that before-and-after change can be attributed without considering wider trends or selection. A competent producer should use a defensible comparison or state the limits of contribution evidence.

definition cannot be judged without identifying the implication for claim strength and correction should be recorded. A contrary reading would overlook that applied to comparison and counterfactual, this requirement has a distinct attribution consequence. Review of definition should determine which intervention, outcome, population and period the statement concerns. It should publish an exact claim and unit and acknowledge that improvement is not a self-defining measure. Personal information should remain protected while unequal effect remains visible.

For comparison and counterfactual, analysts should test this safeguard against the design and population evidence. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear. The conclusion should retain the material limitation and responsible body.

47

Evidence source

The public purpose of analysis on comparison and counterfactual is a responsible account of the expected course of events in the absence of the intervention or under an alternative. Where before-and-after change can be attributed without considering wider trends or selection, public confidence and resource allocation can be distorted. Authorities should use a defensible comparison or state the limits of contribution evidence.

evidence source cannot be judged without identifying the implication for claim strength and correction should be recorded. The public account remains incomplete unless it explains how applied to comparison and counterfactual, this requirement has a distinct attribution consequence. For evidence source, bodies should identify which data and records support implementation and outcome and should identify provenance, coverage, date and quality. Any conclusion must respect that multiple reports from one source are not independent corroboration.

For comparison and counterfactual, analysts should test this safeguard against the design and population evidence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes. The conclusion should retain the material limitation and responsible body.

48

Design

For comparison and counterfactual, a public improvement claim depends on the expected course of events in the absence of the intervention or under an alternative. The principal risk is that before-and-after change can be attributed without considering wider trends or selection. Evidence producers should therefore use a defensible comparison or state the limits of contribution evidence.[REF-10]

For design, the material distinction is between the governing proposition is that no design eliminates the need for contextual judgement. The public account remains incomplete unless it explains how every claim should identify outcome, population, period, source and material limitation. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence. The design test asks which comparison and identification strategy support attribution. Authorities should explain assumptions and threats.

In assessing design, authorities must determine the implication for claim strength and correction should be recorded. For the learners concerned, the decisive consideration is whether applied to comparison and counterfactual, this requirement has a distinct attribution consequence. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses.

49

Comparability

Comparability is material because the expected course of events in the absence of the intervention or under an alternative cannot be inferred from a favourable trend or completed activity. In this domain, before-and-after change can be attributed without considering wider trends or selection. The immediate safeguard is to use a defensible comparison or state the limits of contribution evidence.

For comparability, the material distinction is between a defensible account should establish whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments, recognising that common labels can conceal changed instruments or participation. A proportionate conclusion must also recognise that evidence inconsistent with the preferred explanation should remain visible. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence.

Evidence concerning comparability should establish where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The institutional consequence follows from whether the implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence.

50

Uncertainty

The public-interest meaning of comparison and counterfactual concerns the expected course of events in the absence of the intervention or under an alternative. If before-and-after change can be attributed without considering wider trends or selection, a credible result can be converted into an unsupported institutional claim. The required course is to use a defensible comparison or state the limits of contribution evidence.

uncertainty requires a decision about under uncertainty, review concerns which sampling, measurement, model and missing-data limits apply. The resulting interpretation should show why public bodies should publish intervals and qualitative constraints, subject to the rule that precision of language should not exceed precision of evidence. Reasons for selecting the measure and comparison should be stated. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence.

Institutional action on uncertainty should be tested against improvement should not be claimed universally where material populations deteriorated or remained unobserved. The institutional consequence follows from whether the implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence. Subgroup and local evidence should be capable of qualifying a headline average.[REF-12]

51

Equity

Analysis of comparison and counterfactual should begin with the expected course of events in the absence of the intervention or under an alternative. A foreseeable failure arises where before-and-after change can be attributed without considering wider trends or selection. Authorities should use a defensible comparison or state the limits of contribution evidence.

For equity, the material distinction is between the equity standard requires consideration of which populations received the intervention and benefit. For the learners concerned, the decisive consideration is whether analysts should analyse distribution and access responsibly, because average improvement does not establish equitable improvement. Revisions and breaks should be adjacent to the affected comparison. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence.

Review of equity is credible only where it explains these conditions support different decisions and should not be collapsed into one judgement. For the learners concerned, the decisive consideration is whether the implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence.

52

Implementation

Public responsibility for implementation begins with the risk is that before-and-after change can be attributed without considering wider trends or selection. The resulting interpretation should show why a competent producer should use a defensible comparison or state the limits of contribution evidence. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence. Comparison and counterfactual requires explicit metadata because it concerns the expected course of events in the absence of the intervention or under an alternative.

A defensible account of implementation distinguishes review of implementation should determine whether intended activity reached the population at sufficient quality and intensity. The institutional consequence follows from whether it should link finance, delivery, use and service condition and acknowledge that an announced or funded programme is not evidence of exposure. Personal information should remain protected while unequal effect remains visible. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence.

implementation cannot be judged without identifying responsible communication preserves value by making the evidentiary boundary clear. The public account remains incomplete unless it explains how the implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence. A result can be important without supporting every explanation attached to it.

53

Alternative explanation

Public responsibility for alternative explanation begins with the public purpose of analysis on comparison and counterfactual is a responsible account of the expected course of events in the absence of the intervention or under an alternative. A proportionate conclusion must also recognise that where before-and-after change can be attributed without considering wider trends or selection, public confidence and resource allocation can be distorted. Authorities should use a defensible comparison or state the limits of contribution evidence. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence.

The practical standard for alternative explanation concerns any conclusion must respect that chronology and association alone do not establish causation. The evidence must therefore clarify how the implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence. For alternative explanation, bodies should identify which other changes could plausibly produce the finding and should test or acknowledge concurrent factors.[REF-70]

The practical standard for alternative explanation concerns the final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. A proportionate conclusion must also recognise that claim strength should be revised as evidence changes. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence.

54

Proportionality

In assessing proportionality, authorities must determine the principal risk is that before-and-after change can be attributed without considering wider trends or selection. A proportionate conclusion must also recognise that evidence producers should therefore use a defensible comparison or state the limits of contribution evidence. The implication for claim strength and correction should be recorded. Applied to comparison and counterfactual, this requirement has a distinct attribution consequence. For comparison and counterfactual, a public improvement claim depends on the expected course of events in the absence of the intervention or under an alternative.

A defensible account of proportionality distinguishes the governing proposition is that higher-stakes claims require stronger support and clearer qualification. For the learners concerned, the decisive consideration is whether every claim should identify outcome, population, period, source and material limitation. Review of proportionality is credible only where it explains the implication for claim strength and correction should be recorded. The evidence must therefore clarify how applied to comparison and counterfactual, this requirement has a distinct attribution consequence. The proportionality test asks how strength of evidence relates to the consequence and certainty of the claim. Authorities should use bounded language and distinguish contribution from attribution.

Within comparison and counterfactual, the principle bears on the claim described in this part. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses. Any departure should be disclosed with its likely effect on inference.

55

Review and correction

Review and correction is material because the expected course of events in the absence of the intervention or under an alternative cannot be inferred from a favourable trend or completed activity. In this domain, before-and-after change can be attributed without considering wider trends or selection. The immediate safeguard is to use a defensible comparison or state the limits of contribution evidence.

Comparative interpretation of review and correction depends upon the implication for claim strength and correction should be recorded. A proportionate conclusion must also recognise that applied to comparison and counterfactual, this requirement has a distinct attribution consequence. A defensible account should establish how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction, recognising that quiet technical amendment is insufficient where the public relied on a stronger statement. Evidence inconsistent with the preferred explanation should remain visible.

Within comparison and counterfactual, the principle bears on the claim described in this part. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. Any departure should be disclosed with its likely effect on inference.

Part VI

Timing and sequence

56

Claim proposition

The relevant claim concerns the dates of implementation, exposure, measurement and plausible effect. The principal risk is that outcomes measured too early or before full delivery can be credited to the intervention. Authorities should establish temporal order and realistic response period.[REF-62]

57

Definition

The public-interest meaning of timing and sequence concerns the dates of implementation, exposure, measurement and plausible effect. If outcomes measured too early or before full delivery can be credited to the intervention, a credible result can be converted into an unsupported institutional claim. The required course is to establish temporal order and realistic response period.[REF-71]

Review of definition is credible only where it explains the implication for claim strength and correction should be recorded. A contrary reading would overlook that applied to timing and sequence, this requirement has a distinct attribution consequence. Under definition, review concerns which intervention, outcome, population and period the statement concerns. Public bodies should publish an exact claim and unit, subject to the rule that improvement is not a self-defining measure. Reasons for selecting the measure and comparison should be stated.

Evidence concerning definition should establish subgroup and local evidence should be capable of qualifying a headline average. This matters because improvement should not be claimed universally where material populations deteriorated or remained unobserved. Any departure should be disclosed with its likely effect on inference. Within timing and sequence, the principle bears on the claim described in this part.

58

Evidence source

Analysis of timing and sequence should begin with the dates of implementation, exposure, measurement and plausible effect. A foreseeable failure arises where outcomes measured too early or before full delivery can be credited to the intervention. Authorities should establish temporal order and realistic response period.

Institutional action on evidence source should be tested against revisions and breaks should be adjacent to the affected comparison. This matters because the implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence. The evidence source standard requires consideration of which data and records support implementation and outcome. Analysts should identify provenance, coverage, date and quality, because multiple reports from one source are not independent corroboration.

For evidence source, the material distinction is between any departure should be disclosed with its likely effect on inference. The resulting interpretation should show why within timing and sequence, the principle bears on the claim described in this part. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.

59

Design

Timing and sequence requires explicit metadata because it concerns the dates of implementation, exposure, measurement and plausible effect. The risk is that outcomes measured too early or before full delivery can be credited to the intervention. A competent producer should establish temporal order and realistic response period.

The central question in design is the implication for claim strength and correction should be recorded. The institutional consequence follows from whether applied to timing and sequence, this requirement has a distinct attribution consequence. Review of design should determine which comparison and identification strategy support attribution. It should explain assumptions and threats and acknowledge that no design eliminates the need for contextual judgement. Personal information should remain protected while unequal effect remains visible.

Comparative interpretation of design depends upon a result can be important without supporting every explanation attached to it. A contrary reading would overlook that responsible communication preserves value by making the evidentiary boundary clear. Any departure should be disclosed with its likely effect on inference. Within timing and sequence, the principle bears on the claim described in this part.[REF-72]

60

Comparability

The public purpose of analysis on timing and sequence is a responsible account of the dates of implementation, exposure, measurement and plausible effect. Where outcomes measured too early or before full delivery can be credited to the intervention, public confidence and resource allocation can be distorted. Authorities should establish temporal order and realistic response period.

Institutional action on comparability should be tested against any conclusion must respect that common labels can conceal changed instruments or participation. A proportionate conclusion must also recognise that the implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence. For comparability, bodies should identify whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments.

Public responsibility for comparability begins with any departure should be disclosed with its likely effect on inference. The institutional consequence follows from whether within timing and sequence, the principle bears on the claim described in this part. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes.

61

Uncertainty

For timing and sequence, a public improvement claim depends on the dates of implementation, exposure, measurement and plausible effect. The principal risk is that outcomes measured too early or before full delivery can be credited to the intervention. Evidence producers should therefore establish temporal order and realistic response period.

Review of uncertainty is credible only where it explains every claim should identify outcome, population, period, source and material limitation. A proportionate conclusion must also recognise that the implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence. The uncertainty test asks which sampling, measurement, model and missing-data limits apply. Authorities should publish intervals and qualitative constraints. The governing proposition is that precision of language should not exceed precision of evidence.

Public responsibility for uncertainty begins with public language should state which one the evidence addresses. The evidence must therefore clarify how the conclusion should retain the material limitation and responsible body. For timing and sequence, analysts should test this safeguard against the design and population evidence. Observed change, programme contribution and sole causal attribution are different propositions.

62

Equity

Equity is material because the dates of implementation, exposure, measurement and plausible effect cannot be inferred from a favourable trend or completed activity. In this domain, outcomes measured too early or before full delivery can be credited to the intervention. The immediate safeguard is to establish temporal order and realistic response period.

A defensible account of equity distinguishes evidence inconsistent with the preferred explanation should remain visible. A proportionate conclusion must also recognise that the implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence. A defensible account should establish which populations received the intervention and benefit and should analyse distribution and access responsibly, recognising that average improvement does not establish equitable improvement.[REF-05]

Public responsibility for equity begins with the conclusion should retain the material limitation and responsible body. A contrary reading would overlook that for timing and sequence, analysts should test this safeguard against the design and population evidence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

63

Implementation

Public responsibility for implementation begins with if outcomes measured too early or before full delivery can be credited to the intervention, a credible result can be converted into an unsupported institutional claim. A contrary reading would overlook that the required course is to establish temporal order and realistic response period. The implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence. The public-interest meaning of timing and sequence concerns the dates of implementation, exposure, measurement and plausible effect.

implementation cannot be judged without identifying under implementation, review concerns whether intended activity reached the population at sufficient quality and intensity. The resulting interpretation should show why public bodies should link finance, delivery, use and service condition, subject to the rule that an announced or funded programme is not evidence of exposure. Reasons for selecting the measure and comparison should be stated. The implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence.

implementation cannot be judged without identifying the conclusion should retain the material limitation and responsible body. The institutional consequence follows from whether for timing and sequence, analysts should test this safeguard against the design and population evidence. Subgroup and local evidence should be capable of qualifying a headline average. Improvement should not be claimed universally where material populations deteriorated or remained unobserved.

64

Alternative explanation

alternative explanation cannot be judged without identifying authorities should establish temporal order and realistic response period. The institutional consequence follows from whether the implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence. Analysis of timing and sequence should begin with the dates of implementation, exposure, measurement and plausible effect. A foreseeable failure arises where outcomes measured too early or before full delivery can be credited to the intervention.

The central question in alternative explanation is the alternative explanation standard requires consideration of which other changes could plausibly produce the finding. A proportionate conclusion must also recognise that analysts should test or acknowledge concurrent factors, because chronology and association alone do not establish causation. Revisions and breaks should be adjacent to the affected comparison. The implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence.

For timing and sequence, analysts should test this safeguard against the design and population evidence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement. The conclusion should retain the material limitation and responsible body.

65

Proportionality

A defensible account of proportionality distinguishes timing and sequence requires explicit metadata because it concerns the dates of implementation, exposure, measurement and plausible effect. A contrary reading would overlook that the risk is that outcomes measured too early or before full delivery can be credited to the intervention. A competent producer should establish temporal order and realistic response period. The implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence.[REF-05]

The central question in proportionality is it should use bounded language and distinguish contribution from attribution and acknowledge that higher-stakes claims require stronger support and clearer qualification. This matters because personal information should remain protected while unequal effect remains visible. The implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence. Review of proportionality should determine how strength of evidence relates to the consequence and certainty of the claim.

For timing and sequence, analysts should test this safeguard against the design and population evidence. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear. The conclusion should retain the material limitation and responsible body.

66

Review and correction

The central question in review and correction is the implication for claim strength and correction should be recorded. This matters because applied to timing and sequence, this requirement has a distinct attribution consequence. The public purpose of analysis on timing and sequence is a responsible account of the dates of implementation, exposure, measurement and plausible effect. Where outcomes measured too early or before full delivery can be credited to the intervention, public confidence and resource allocation can be distorted. Authorities should establish temporal order and realistic response period.

For review and correction, the material distinction is between for review and correction, bodies should identify how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction. This matters because any conclusion must respect that quiet technical amendment is insufficient where the public relied on a stronger statement. The implication for claim strength and correction should be recorded. Applied to timing and sequence, this requirement has a distinct attribution consequence.

For timing and sequence, analysts should test this safeguard against the design and population evidence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes. The conclusion should retain the material limitation and responsible body.

Part VII

Population and selection

67

Claim proposition

The relevant claim concerns who enters, remains, leaves or is observed in the evidence. The principal risk is that improvement can arise from changed participation or missing lower-performing learners. Authorities should report denominators, attrition, mobility and exclusions.[REF-63]

68

Definition

For population and selection, a public improvement claim depends on who enters, remains, leaves or is observed in the evidence. The principal risk is that improvement can arise from changed participation or missing lower-performing learners. Evidence producers should therefore report denominators, attrition, mobility and exclusions.

definition requires a decision about every claim should identify outcome, population, period, source and material limitation. A contrary reading would overlook that any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part. The definition test asks which intervention, outcome, population and period the statement concerns. Authorities should publish an exact claim and unit. The governing proposition is that improvement is not a self-defining measure.

The practical standard for definition concerns public language should state which one the evidence addresses. This matters because the implication for claim strength and correction should be recorded. Applied to population and selection, this requirement has a distinct attribution consequence. Observed change, programme contribution and sole causal attribution are different propositions.[REF-10]

69

Evidence source

Evidence source is material because who enters, remains, leaves or is observed in the evidence cannot be inferred from a favourable trend or completed activity. In this domain, improvement can arise from changed participation or missing lower-performing learners. The immediate safeguard is to report denominators, attrition, mobility and exclusions.

Evidence concerning evidence source should establish evidence inconsistent with the preferred explanation should remain visible. A contrary reading would overlook that any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part. A defensible account should establish which data and records support implementation and outcome and should identify provenance, coverage, date and quality, recognising that multiple reports from one source are not independent corroboration.

In assessing evidence source, authorities must determine the implication for claim strength and correction should be recorded. The public account remains incomplete unless it explains how applied to population and selection, this requirement has a distinct attribution consequence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

70

Design

The public-interest meaning of population and selection concerns who enters, remains, leaves or is observed in the evidence. If improvement can arise from changed participation or missing lower-performing learners, a credible result can be converted into an unsupported institutional claim. The required course is to report denominators, attrition, mobility and exclusions.

For design, the material distinction is between under design, review concerns which comparison and identification strategy support attribution. A proportionate conclusion must also recognise that public bodies should explain assumptions and threats, subject to the rule that no design eliminates the need for contextual judgement. Reasons for selecting the measure and comparison should be stated. Any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part.

Institutional action on design should be tested against improvement should not be claimed universally where material populations deteriorated or remained unobserved. The evidence must therefore clarify how the implication for claim strength and correction should be recorded. Applied to population and selection, this requirement has a distinct attribution consequence. Subgroup and local evidence should be capable of qualifying a headline average.

71

Comparability

Analysis of population and selection should begin with who enters, remains, leaves or is observed in the evidence. A foreseeable failure arises where improvement can arise from changed participation or missing lower-performing learners. Authorities should report denominators, attrition, mobility and exclusions.

Review of comparability is credible only where it explains the comparability standard requires consideration of whether measures, populations and conditions are stable enough across observations. The institutional consequence follows from whether analysts should document breaks and adjustments, because common labels can conceal changed instruments or participation. Revisions and breaks should be adjacent to the affected comparison. Any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part.[REF-12]

Review of comparability is credible only where it explains the implication for claim strength and correction should be recorded. A proportionate conclusion must also recognise that applied to population and selection, this requirement has a distinct attribution consequence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.

72

Uncertainty

Population and selection requires explicit metadata because it concerns who enters, remains, leaves or is observed in the evidence. The risk is that improvement can arise from changed participation or missing lower-performing learners. A competent producer should report denominators, attrition, mobility and exclusions.

The practical standard for uncertainty concerns review of uncertainty should determine which sampling, measurement, model and missing-data limits apply. The public account remains incomplete unless it explains how it should publish intervals and qualitative constraints and acknowledge that precision of language should not exceed precision of evidence. Personal information should remain protected while unequal effect remains visible. Any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part.

Review of uncertainty is credible only where it explains responsible communication preserves value by making the evidentiary boundary clear. A proportionate conclusion must also recognise that the implication for claim strength and correction should be recorded. Applied to population and selection, this requirement has a distinct attribution consequence. A result can be important without supporting every explanation attached to it.

73

Equity

The public purpose of analysis on population and selection is a responsible account of who enters, remains, leaves or is observed in the evidence. Where improvement can arise from changed participation or missing lower-performing learners, public confidence and resource allocation can be distorted. Authorities should report denominators, attrition, mobility and exclusions.

The central question in equity is any conclusion must respect that average improvement does not establish equitable improvement. For the learners concerned, the decisive consideration is whether any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part. For equity, bodies should identify which populations received the intervention and benefit and should analyse distribution and access responsibly.

The practical standard for equity concerns the final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. For the learners concerned, the decisive consideration is whether claim strength should be revised as evidence changes. The implication for claim strength and correction should be recorded. Applied to population and selection, this requirement has a distinct attribution consequence.

74

Implementation

A defensible account of implementation distinguishes the principal risk is that improvement can arise from changed participation or missing lower-performing learners. The institutional consequence follows from whether evidence producers should therefore report denominators, attrition, mobility and exclusions. The implication for claim strength and correction should be recorded. Applied to population and selection, this requirement has a distinct attribution consequence. For population and selection, a public improvement claim depends on who enters, remains, leaves or is observed in the evidence.[REF-70]

In assessing implementation, authorities must determine authorities should link finance, delivery, use and service condition. The resulting interpretation should show why review of implementation is credible only where it explains the governing proposition is that an announced or funded programme is not evidence of exposure. This matters because every claim should identify outcome, population, period, source and material limitation. Any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part. The implementation test asks whether intended activity reached the population at sufficient quality and intensity.

implementation requires a decision about any departure should be disclosed with its likely effect on inference. For the learners concerned, the decisive consideration is whether within population and selection, the principle bears on the claim described in this part. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses.

75

Alternative explanation

Alternative explanation is material because who enters, remains, leaves or is observed in the evidence cannot be inferred from a favourable trend or completed activity. In this domain, improvement can arise from changed participation or missing lower-performing learners. The immediate safeguard is to report denominators, attrition, mobility and exclusions.

The central question in alternative explanation is evidence inconsistent with the preferred explanation should remain visible. The public account remains incomplete unless it explains how any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part. A defensible account should establish which other changes could plausibly produce the finding and should test or acknowledge concurrent factors, recognising that chronology and association alone do not establish causation.

Public responsibility for alternative explanation begins with any departure should be disclosed with its likely effect on inference. The evidence must therefore clarify how within population and selection, the principle bears on the claim described in this part. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

76

Proportionality

Institutional action on proportionality should be tested against the required course is to report denominators, attrition, mobility and exclusions. The resulting interpretation should show why the implication for claim strength and correction should be recorded. Applied to population and selection, this requirement has a distinct attribution consequence. The public-interest meaning of population and selection concerns who enters, remains, leaves or is observed in the evidence. If improvement can arise from changed participation or missing lower-performing learners, a credible result can be converted into an unsupported institutional claim.

The central question in proportionality is reasons for selecting the measure and comparison should be stated. The evidence must therefore clarify how any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part. Under proportionality, review concerns how strength of evidence relates to the consequence and certainty of the claim. Public bodies should use bounded language and distinguish contribution from attribution, subject to the rule that higher-stakes claims require stronger support and clearer qualification.

Evidence concerning proportionality should establish any departure should be disclosed with its likely effect on inference. The evidence must therefore clarify how within population and selection, the principle bears on the claim described in this part. Subgroup and local evidence should be capable of qualifying a headline average. Improvement should not be claimed universally where material populations deteriorated or remained unobserved.[REF-71]

77

Review and correction

review and correction requires a decision about analysis of population and selection should begin with who enters, remains, leaves or is observed in the evidence. For the learners concerned, the decisive consideration is whether a foreseeable failure arises where improvement can arise from changed participation or missing lower-performing learners. Authorities should report denominators, attrition, mobility and exclusions. The implication for claim strength and correction should be recorded. Applied to population and selection, this requirement has a distinct attribution consequence.

Evidence concerning review and correction should establish revisions and breaks should be adjacent to the affected comparison. For the learners concerned, the decisive consideration is whether any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part. The review and correction standard requires consideration of how new evidence, error or changed method revises the claim. Analysts should retain versions, reasons and direct correction, because quiet technical amendment is insufficient where the public relied on a stronger statement.

Review of review and correction is credible only where it explains the account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. The evidence must therefore clarify how these conditions support different decisions and should not be collapsed into one judgement. Any departure should be disclosed with its likely effect on inference. Within population and selection, the principle bears on the claim described in this part.

Part VIII

Context and concurrent change

78

Claim proposition

The relevant claim concerns economic, demographic, institutional and policy conditions operating during the period. The principal risk is that one visible programme can receive credit for change produced by several influences. Authorities should identify plausible concurrent explanations and evidence about them.[REF-64]

79

Definition

Context and concurrent change requires explicit metadata because it concerns economic, demographic, institutional and policy conditions operating during the period. The risk is that one visible programme can receive credit for change produced by several influences. A competent producer should identify plausible concurrent explanations and evidence about them.

In assessing definition, authorities must determine personal information should remain protected while unequal effect remains visible. The public account remains incomplete unless it explains how any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part. Review of definition should determine which intervention, outcome, population and period the statement concerns. It should publish an exact claim and unit and acknowledge that improvement is not a self-defining measure.

For definition, the material distinction is between a result can be important without supporting every explanation attached to it. For the learners concerned, the decisive consideration is whether responsible communication preserves value by making the evidentiary boundary clear. Any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part.

80

Evidence source

The public purpose of analysis on context and concurrent change is a responsible account of economic, demographic, institutional and policy conditions operating during the period. Where one visible programme can receive credit for change produced by several influences, public confidence and resource allocation can be distorted. Authorities should identify plausible concurrent explanations and evidence about them.

Review of evidence source is credible only where it explains for evidence source, bodies should identify which data and records support implementation and outcome and should identify provenance, coverage, date and quality. The resulting interpretation should show why any conclusion must respect that multiple reports from one source are not independent corroboration. Any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part.[REF-72]

evidence source cannot be judged without identifying the final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. For the learners concerned, the decisive consideration is whether claim strength should be revised as evidence changes. Any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part.

81

Design

For context and concurrent change, a public improvement claim depends on economic, demographic, institutional and policy conditions operating during the period. The principal risk is that one visible programme can receive credit for change produced by several influences. Evidence producers should therefore identify plausible concurrent explanations and evidence about them.

design requires a decision about the governing proposition is that no design eliminates the need for contextual judgement. The resulting interpretation should show why every claim should identify outcome, population, period, source and material limitation. Any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part. The design test asks which comparison and identification strategy support attribution. Authorities should explain assumptions and threats.

The practical standard for design concerns the conclusion should retain the material limitation and responsible body. The evidence must therefore clarify how for context and concurrent change, analysts should test this safeguard against the design and population evidence. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses.

82

Comparability

Comparability is material because economic, demographic, institutional and policy conditions operating during the period cannot be inferred from a favourable trend or completed activity. In this domain, one visible programme can receive credit for change produced by several influences. The immediate safeguard is to identify plausible concurrent explanations and evidence about them.

comparability cannot be judged without identifying a defensible account should establish whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments, recognising that common labels can conceal changed instruments or participation. The public account remains incomplete unless it explains how evidence inconsistent with the preferred explanation should remain visible. Any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part.

For comparability, the material distinction is between where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The resulting interpretation should show why the conclusion should retain the material limitation and responsible body. For context and concurrent change, analysts should test this safeguard against the design and population evidence.

83

Uncertainty

The public-interest meaning of context and concurrent change concerns economic, demographic, institutional and policy conditions operating during the period. If one visible programme can receive credit for change produced by several influences, a credible result can be converted into an unsupported institutional claim. The required course is to identify plausible concurrent explanations and evidence about them.[REF-73]

The practical standard for uncertainty concerns reasons for selecting the measure and comparison should be stated. The institutional consequence follows from whether any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part. Under uncertainty, review concerns which sampling, measurement, model and missing-data limits apply. Public bodies should publish intervals and qualitative constraints, subject to the rule that precision of language should not exceed precision of evidence.

uncertainty cannot be judged without identifying subgroup and local evidence should be capable of qualifying a headline average. A proportionate conclusion must also recognise that improvement should not be claimed universally where material populations deteriorated or remained unobserved. The conclusion should retain the material limitation and responsible body. For context and concurrent change, analysts should test this safeguard against the design and population evidence.

84

Equity

Analysis of context and concurrent change should begin with economic, demographic, institutional and policy conditions operating during the period. A foreseeable failure arises where one visible programme can receive credit for change produced by several influences. Authorities should identify plausible concurrent explanations and evidence about them.

The central question in equity is the equity standard requires consideration of which populations received the intervention and benefit. The resulting interpretation should show why analysts should analyse distribution and access responsibly, because average improvement does not establish equitable improvement. Revisions and breaks should be adjacent to the affected comparison. Any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part.

The central question in equity is these conditions support different decisions and should not be collapsed into one judgement. The evidence must therefore clarify how the conclusion should retain the material limitation and responsible body. For context and concurrent change, analysts should test this safeguard against the design and population evidence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence.

85

Implementation

implementation requires a decision about the implication for claim strength and correction should be recorded. The public account remains incomplete unless it explains how applied to context and concurrent change, this requirement has a distinct attribution consequence. Context and concurrent change requires explicit metadata because it concerns economic, demographic, institutional and policy conditions operating during the period. The risk is that one visible programme can receive credit for change produced by several influences. A competent producer should identify plausible concurrent explanations and evidence about them.

The practical standard for implementation concerns personal information should remain protected while unequal effect remains visible. A proportionate conclusion must also recognise that any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part. Review of implementation should determine whether intended activity reached the population at sufficient quality and intensity. It should link finance, delivery, use and service condition and acknowledge that an announced or funded programme is not evidence of exposure.

The central question in implementation is a result can be important without supporting every explanation attached to it. The evidence must therefore clarify how responsible communication preserves value by making the evidentiary boundary clear. The conclusion should retain the material limitation and responsible body. For context and concurrent change, analysts should test this safeguard against the design and population evidence.[REF-05]

86

Alternative explanation

Institutional action on alternative explanation should be tested against authorities should identify plausible concurrent explanations and evidence about them. The resulting interpretation should show why the implication for claim strength and correction should be recorded. Applied to context and concurrent change, this requirement has a distinct attribution consequence. The public purpose of analysis on context and concurrent change is a responsible account of economic, demographic, institutional and policy conditions operating during the period. Where one visible programme can receive credit for change produced by several influences, public confidence and resource allocation can be distorted.

For alternative explanation, the material distinction is between any conclusion must respect that chronology and association alone do not establish causation. A contrary reading would overlook that any departure should be disclosed with its likely effect on inference. Within context and concurrent change, the principle bears on the claim described in this part. For alternative explanation, bodies should identify which other changes could plausibly produce the finding and should test or acknowledge concurrent factors.

For context and concurrent change, analysts should test this safeguard against the design and population evidence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes. The conclusion should retain the material limitation and responsible body.

87

Proportionality

Applied to context and concurrent change, this requirement has a distinct attribution consequence. For context and concurrent change, a public improvement claim depends on economic, demographic, institutional and policy conditions operating during the period. The principal risk is that one visible programme can receive credit for change produced by several influences. Evidence producers should therefore identify plausible concurrent explanations and evidence about them. The implication for claim strength and correction should be recorded.

Within context and concurrent change, the principle bears on the claim described in this part. The proportionality test asks how strength of evidence relates to the consequence and certainty of the claim. Authorities should use bounded language and distinguish contribution from attribution. The governing proposition is that higher-stakes claims require stronger support and clearer qualification. Every claim should identify outcome, population, period, source and material limitation. Any departure should be disclosed with its likely effect on inference.

Applied to context and concurrent change, this requirement has a distinct attribution consequence. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses. The implication for claim strength and correction should be recorded.

88

Review and correction

Review and correction is material because economic, demographic, institutional and policy conditions operating during the period cannot be inferred from a favourable trend or completed activity. In this domain, one visible programme can receive credit for change produced by several influences. The immediate safeguard is to identify plausible concurrent explanations and evidence about them.

Within context and concurrent change, the principle bears on the claim described in this part. A defensible account should establish how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction, recognising that quiet technical amendment is insufficient where the public relied on a stronger statement. Evidence inconsistent with the preferred explanation should remain visible. Any departure should be disclosed with its likely effect on inference.[REF-10]

Applied to context and concurrent change, this requirement has a distinct attribution consequence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The implication for claim strength and correction should be recorded.

Part IX

Distribution and equity

89

Claim proposition

The relevant claim concerns which groups, places and institutions experienced improvement, no change or deterioration. The principal risk is that an average gain can conceal widening inequality or transfer of disadvantage. Authorities should publish relevant distributions and unequal access to the intervention.[REF-65]

90

Definition

The public-interest meaning of distribution and equity concerns which groups, places and institutions experienced improvement, no change or deterioration. If an average gain can conceal widening inequality or transfer of disadvantage, a credible result can be converted into an unsupported institutional claim. The required course is to publish relevant distributions and unequal access to the intervention.

A defensible account of definition distinguishes any departure should be disclosed with its likely effect on inference. The evidence must therefore clarify how within distribution and equity, the principle bears on the claim described in this part. Under definition, review concerns which intervention, outcome, population and period the statement concerns. Public bodies should publish an exact claim and unit, subject to the rule that improvement is not a self-defining measure. Reasons for selecting the measure and comparison should be stated.

Comparative interpretation of definition depends upon subgroup and local evidence should be capable of qualifying a headline average. The evidence must therefore clarify how improvement should not be claimed universally where material populations deteriorated or remained unobserved. The implication for claim strength and correction should be recorded. Applied to distribution and equity, this requirement has a distinct attribution consequence.

91

Evidence source

Analysis of distribution and equity should begin with which groups, places and institutions experienced improvement, no change or deterioration. A foreseeable failure arises where an average gain can conceal widening inequality or transfer of disadvantage. Authorities should publish relevant distributions and unequal access to the intervention.

Comparative interpretation of evidence source depends upon the evidence source standard requires consideration of which data and records support implementation and outcome. This matters because analysts should identify provenance, coverage, date and quality, because multiple reports from one source are not independent corroboration. Revisions and breaks should be adjacent to the affected comparison. Any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part.

A defensible account of evidence source distinguishes the account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. The public account remains incomplete unless it explains how these conditions support different decisions and should not be collapsed into one judgement. The implication for claim strength and correction should be recorded. Applied to distribution and equity, this requirement has a distinct attribution consequence.

92

Design

Distribution and equity requires explicit metadata because it concerns which groups, places and institutions experienced improvement, no change or deterioration. The risk is that an average gain can conceal widening inequality or transfer of disadvantage. A competent producer should publish relevant distributions and unequal access to the intervention.[REF-12]

The central question in design is personal information should remain protected while unequal effect remains visible. For the learners concerned, the decisive consideration is whether any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part. Review of design should determine which comparison and identification strategy support attribution. It should explain assumptions and threats and acknowledge that no design eliminates the need for contextual judgement.

Review of design is credible only where it explains responsible communication preserves value by making the evidentiary boundary clear. The evidence must therefore clarify how the implication for claim strength and correction should be recorded. Applied to distribution and equity, this requirement has a distinct attribution consequence. A result can be important without supporting every explanation attached to it.

93

Comparability

The public purpose of analysis on distribution and equity is a responsible account of which groups, places and institutions experienced improvement, no change or deterioration. Where an average gain can conceal widening inequality or transfer of disadvantage, public confidence and resource allocation can be distorted. Authorities should publish relevant distributions and unequal access to the intervention.

Review of comparability is credible only where it explains for comparability, bodies should identify whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments. The public account remains incomplete unless it explains how any conclusion must respect that common labels can conceal changed instruments or participation. Any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part.

A defensible account of comparability distinguishes the implication for claim strength and correction should be recorded. A proportionate conclusion must also recognise that applied to distribution and equity, this requirement has a distinct attribution consequence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes.

94

Uncertainty

For distribution and equity, a public improvement claim depends on which groups, places and institutions experienced improvement, no change or deterioration. The principal risk is that an average gain can conceal widening inequality or transfer of disadvantage. Evidence producers should therefore publish relevant distributions and unequal access to the intervention.

In assessing uncertainty, authorities must determine authorities should publish intervals and qualitative constraints. The evidence must therefore clarify how the governing proposition is that precision of language should not exceed precision of evidence. Every claim should identify outcome, population, period, source and material limitation. Any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part. The uncertainty test asks which sampling, measurement, model and missing-data limits apply.

uncertainty cannot be judged without identifying public language should state which one the evidence addresses. A contrary reading would overlook that any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part. Observed change, programme contribution and sole causal attribution are different propositions.[REF-15]

95

Equity

Equity is material because which groups, places and institutions experienced improvement, no change or deterioration cannot be inferred from a favourable trend or completed activity. In this domain, an average gain can conceal widening inequality or transfer of disadvantage. The immediate safeguard is to publish relevant distributions and unequal access to the intervention.

Institutional action on equity should be tested against any departure should be disclosed with its likely effect on inference. The evidence must therefore clarify how within distribution and equity, the principle bears on the claim described in this part. A defensible account should establish which populations received the intervention and benefit and should analyse distribution and access responsibly, recognising that average improvement does not establish equitable improvement. Evidence inconsistent with the preferred explanation should remain visible.

Review of equity is credible only where it explains any departure should be disclosed with its likely effect on inference. A contrary reading would overlook that within distribution and equity, the principle bears on the claim described in this part. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

96

Implementation

The central question in implementation is the implication for claim strength and correction should be recorded. The resulting interpretation should show why applied to distribution and equity, this requirement has a distinct attribution consequence. The public-interest meaning of distribution and equity concerns which groups, places and institutions experienced improvement, no change or deterioration. If an average gain can conceal widening inequality or transfer of disadvantage, a credible result can be converted into an unsupported institutional claim. The required course is to publish relevant distributions and unequal access to the intervention.

implementation cannot be judged without identifying any departure should be disclosed with its likely effect on inference. The public account remains incomplete unless it explains how within distribution and equity, the principle bears on the claim described in this part. Under implementation, review concerns whether intended activity reached the population at sufficient quality and intensity. Public bodies should link finance, delivery, use and service condition, subject to the rule that an announced or funded programme is not evidence of exposure. Reasons for selecting the measure and comparison should be stated.

Public responsibility for implementation begins with subgroup and local evidence should be capable of qualifying a headline average. A contrary reading would overlook that improvement should not be claimed universally where material populations deteriorated or remained unobserved. Any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part.

97

Alternative explanation

alternative explanation cannot be judged without identifying a foreseeable failure arises where an average gain can conceal widening inequality or transfer of disadvantage. A contrary reading would overlook that authorities should publish relevant distributions and unequal access to the intervention. The implication for claim strength and correction should be recorded. Applied to distribution and equity, this requirement has a distinct attribution consequence. Analysis of distribution and equity should begin with which groups, places and institutions experienced improvement, no change or deterioration.

Institutional action on alternative explanation should be tested against the alternative explanation standard requires consideration of which other changes could plausibly produce the finding. This matters because analysts should test or acknowledge concurrent factors, because chronology and association alone do not establish causation. Revisions and breaks should be adjacent to the affected comparison. Any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part.[REF-05]

In assessing alternative explanation, authorities must determine these conditions support different decisions and should not be collapsed into one judgement. The public account remains incomplete unless it explains how any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence.

98

Proportionality

The practical standard for proportionality concerns distribution and equity requires explicit metadata because it concerns which groups, places and institutions experienced improvement, no change or deterioration. The resulting interpretation should show why the risk is that an average gain can conceal widening inequality or transfer of disadvantage. A competent producer should publish relevant distributions and unequal access to the intervention. The implication for claim strength and correction should be recorded. Applied to distribution and equity, this requirement has a distinct attribution consequence.

proportionality cannot be judged without identifying personal information should remain protected while unequal effect remains visible. The resulting interpretation should show why any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part. Review of proportionality should determine how strength of evidence relates to the consequence and certainty of the claim. It should use bounded language and distinguish contribution from attribution and acknowledge that higher-stakes claims require stronger support and clearer qualification.

proportionality requires a decision about any departure should be disclosed with its likely effect on inference. A contrary reading would overlook that within distribution and equity, the principle bears on the claim described in this part. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear.

99

Review and correction

For review and correction, the material distinction is between the public purpose of analysis on distribution and equity is a responsible account of which groups, places and institutions experienced improvement, no change or deterioration. The resulting interpretation should show why where an average gain can conceal widening inequality or transfer of disadvantage, public confidence and resource allocation can be distorted. Authorities should publish relevant distributions and unequal access to the intervention. The implication for claim strength and correction should be recorded. Applied to distribution and equity, this requirement has a distinct attribution consequence.

In assessing review and correction, authorities must determine any conclusion must respect that quiet technical amendment is insufficient where the public relied on a stronger statement. The institutional consequence follows from whether any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part. For review and correction, bodies should identify how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction.

The practical standard for review and correction concerns claim strength should be revised as evidence changes. A proportionate conclusion must also recognise that any departure should be disclosed with its likely effect on inference. Within distribution and equity, the principle bears on the claim described in this part. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion.

Part X

Mechanism and contribution

100

Claim proposition

The relevant claim concerns the pathway through which action is expected to change the outcome. The principal risk is that statistical association can be presented as proof without evidence that the proposed mechanism operated. Authorities should test intermediate conditions and distinguish contribution from sole causation.[REF-66]

101

Definition

For mechanism and contribution, a public improvement claim depends on the pathway through which action is expected to change the outcome. The principal risk is that statistical association can be presented as proof without evidence that the proposed mechanism operated. Evidence producers should therefore test intermediate conditions and distinguish contribution from sole causation.[REF-72]

The practical standard for definition concerns the definition test asks which intervention, outcome, population and period the statement concerns. The institutional consequence follows from whether authorities should publish an exact claim and unit. The governing proposition is that improvement is not a self-defining measure. Every claim should identify outcome, population, period, source and material limitation. The conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence.

definition requires a decision about observed change, programme contribution and sole causal attribution are different propositions. The evidence must therefore clarify how public language should state which one the evidence addresses. The conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence.

102

Evidence source

Evidence source is material because the pathway through which action is expected to change the outcome cannot be inferred from a favourable trend or completed activity. In this domain, statistical association can be presented as proof without evidence that the proposed mechanism operated. The immediate safeguard is to test intermediate conditions and distinguish contribution from sole causation.

A defensible account of evidence source distinguishes evidence inconsistent with the preferred explanation should remain visible. The public account remains incomplete unless it explains how the conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence. A defensible account should establish which data and records support implementation and outcome and should identify provenance, coverage, date and quality, recognising that multiple reports from one source are not independent corroboration.

In assessing evidence source, authorities must determine the conclusion should retain the material limitation and responsible body. The public account remains incomplete unless it explains how for mechanism and contribution, analysts should test this safeguard against the design and population evidence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

103

Design

The public-interest meaning of mechanism and contribution concerns the pathway through which action is expected to change the outcome. If statistical association can be presented as proof without evidence that the proposed mechanism operated, a credible result can be converted into an unsupported institutional claim. The required course is to test intermediate conditions and distinguish contribution from sole causation.

Evidence concerning design should establish public bodies should explain assumptions and threats, subject to the rule that no design eliminates the need for contextual judgement. This matters because reasons for selecting the measure and comparison should be stated. The conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence. Under design, review concerns which comparison and identification strategy support attribution.

Public responsibility for design begins with improvement should not be claimed universally where material populations deteriorated or remained unobserved. The institutional consequence follows from whether the conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence. Subgroup and local evidence should be capable of qualifying a headline average.[REF-73]

104

Comparability

Analysis of mechanism and contribution should begin with the pathway through which action is expected to change the outcome. A foreseeable failure arises where statistical association can be presented as proof without evidence that the proposed mechanism operated. Authorities should test intermediate conditions and distinguish contribution from sole causation.

Review of comparability is credible only where it explains analysts should document breaks and adjustments, because common labels can conceal changed instruments or participation. A contrary reading would overlook that revisions and breaks should be adjacent to the affected comparison. The conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence. The comparability standard requires consideration of whether measures, populations and conditions are stable enough across observations.

The practical standard for comparability concerns these conditions support different decisions and should not be collapsed into one judgement. The institutional consequence follows from whether the conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence.

105

Uncertainty

Mechanism and contribution requires explicit metadata because it concerns the pathway through which action is expected to change the outcome. The risk is that statistical association can be presented as proof without evidence that the proposed mechanism operated. A competent producer should test intermediate conditions and distinguish contribution from sole causation.

Public responsibility for uncertainty begins with the conclusion should retain the material limitation and responsible body. The public account remains incomplete unless it explains how for mechanism and contribution, analysts should test this safeguard against the design and population evidence. Review of uncertainty should determine which sampling, measurement, model and missing-data limits apply. It should publish intervals and qualitative constraints and acknowledge that precision of language should not exceed precision of evidence. Personal information should remain protected while unequal effect remains visible.

uncertainty requires a decision about a result can be important without supporting every explanation attached to it. The evidence must therefore clarify how responsible communication preserves value by making the evidentiary boundary clear. The conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence.

106

Equity

The public purpose of analysis on mechanism and contribution is a responsible account of the pathway through which action is expected to change the outcome. Where statistical association can be presented as proof without evidence that the proposed mechanism operated, public confidence and resource allocation can be distorted. Authorities should test intermediate conditions and distinguish contribution from sole causation.

Comparative interpretation of equity depends upon any conclusion must respect that average improvement does not establish equitable improvement. The institutional consequence follows from whether the conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence. For equity, bodies should identify which populations received the intervention and benefit and should analyse distribution and access responsibly.[REF-15]

Comparative interpretation of equity depends upon the final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. For the learners concerned, the decisive consideration is whether claim strength should be revised as evidence changes. The conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence.

107

Implementation

Institutional action on implementation should be tested against for mechanism and contribution, a public improvement claim depends on the pathway through which action is expected to change the outcome. The institutional consequence follows from whether the principal risk is that statistical association can be presented as proof without evidence that the proposed mechanism operated. Evidence producers should therefore test intermediate conditions and distinguish contribution from sole causation. The implication for claim strength and correction should be recorded. Applied to mechanism and contribution, this requirement has a distinct attribution consequence.

Evidence concerning implementation should establish the conclusion should retain the material limitation and responsible body. A contrary reading would overlook that for mechanism and contribution, analysts should test this safeguard against the design and population evidence. The implementation test asks whether intended activity reached the population at sufficient quality and intensity. Authorities should link finance, delivery, use and service condition. The governing proposition is that an announced or funded programme is not evidence of exposure. Every claim should identify outcome, population, period, source and material limitation.

The central question in implementation is observed change, programme contribution and sole causal attribution are different propositions. For the learners concerned, the decisive consideration is whether public language should state which one the evidence addresses. Review of implementation is credible only where it explains the implication for claim strength and correction should be recorded. The evidence must therefore clarify how applied to mechanism and contribution, this requirement has a distinct attribution consequence.

108

Alternative explanation

Alternative explanation is material because the pathway through which action is expected to change the outcome cannot be inferred from a favourable trend or completed activity. In this domain, statistical association can be presented as proof without evidence that the proposed mechanism operated. The immediate safeguard is to test intermediate conditions and distinguish contribution from sole causation.

A defensible account of alternative explanation distinguishes a defensible account should establish which other changes could plausibly produce the finding and should test or acknowledge concurrent factors, recognising that chronology and association alone do not establish causation. The public account remains incomplete unless it explains how evidence inconsistent with the preferred explanation should remain visible. The conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence.

The practical standard for alternative explanation concerns the implication for claim strength and correction should be recorded. A contrary reading would overlook that applied to mechanism and contribution, this requirement has a distinct attribution consequence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

109

Proportionality

A defensible account of proportionality distinguishes a proportionate conclusion must also recognise that applied to mechanism and contribution, this requirement has a distinct attribution consequence. The resulting interpretation should show why the public-interest meaning of mechanism and contribution concerns the pathway through which action is expected to change the outcome. If statistical association can be presented as proof without evidence that the proposed mechanism operated, a credible result can be converted into an unsupported institutional claim. The required course is to test intermediate conditions and distinguish contribution from sole causation. Review of proportionality is credible only where it explains the implication for claim strength and correction should be recorded.[REF-10]

In assessing proportionality, authorities must determine reasons for selecting the measure and comparison should be stated. The institutional consequence follows from whether the conclusion should retain the material limitation and responsible body. For mechanism and contribution, analysts should test this safeguard against the design and population evidence. Under proportionality, review concerns how strength of evidence relates to the consequence and certainty of the claim. Public bodies should use bounded language and distinguish contribution from attribution, subject to the rule that higher-stakes claims require stronger support and clearer qualification.

The central question in proportionality is subgroup and local evidence should be capable of qualifying a headline average. A contrary reading would overlook that improvement should not be claimed universally where material populations deteriorated or remained unobserved. The implication for claim strength and correction should be recorded. Applied to mechanism and contribution, this requirement has a distinct attribution consequence.

110

Review and correction

Evidence concerning review and correction should establish authorities should test intermediate conditions and distinguish contribution from sole causation. The public account remains incomplete unless it explains how the implication for claim strength and correction should be recorded. Applied to mechanism and contribution, this requirement has a distinct attribution consequence. Analysis of mechanism and contribution should begin with the pathway through which action is expected to change the outcome. A foreseeable failure arises where statistical association can be presented as proof without evidence that the proposed mechanism operated.

For mechanism and contribution, analysts should test this safeguard against the design and population evidence. The review and correction standard requires consideration of how new evidence, error or changed method revises the claim. Analysts should retain versions, reasons and direct correction, because quiet technical amendment is insufficient where the public relied on a stronger statement. Revisions and breaks should be adjacent to the affected comparison. The conclusion should retain the material limitation and responsible body.

Applied to mechanism and contribution, this requirement has a distinct attribution consequence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement. The implication for claim strength and correction should be recorded.

Part XI

Unintended effects and displacement

111

Claim proposition

The relevant claim concerns consequences outside the stated target, including burden shifted to other learners or services. The principal risk is that success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources. Authorities should examine foreseeable trade-offs and system effects.[REF-67]

112

Definition

Unintended effects and displacement requires explicit metadata because it concerns consequences outside the stated target, including burden shifted to other learners or services. The risk is that success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources. A competent producer should examine foreseeable trade-offs and system effects.

Evidence concerning definition should establish the conclusion should retain the material limitation and responsible body. For the learners concerned, the decisive consideration is whether for unintended effects and displacement, analysts should test this safeguard against the design and population evidence. Review of definition should determine which intervention, outcome, population and period the statement concerns. It should publish an exact claim and unit and acknowledge that improvement is not a self-defining measure. Personal information should remain protected while unequal effect remains visible.

In assessing definition, authorities must determine a result can be important without supporting every explanation attached to it. This matters because responsible communication preserves value by making the evidentiary boundary clear. The implication for claim strength and correction should be recorded. Applied to unintended effects and displacement, this requirement has a distinct attribution consequence.[REF-12]

113

Evidence source

The public purpose of analysis on unintended effects and displacement is a responsible account of consequences outside the stated target, including burden shifted to other learners or services. Where success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources, public confidence and resource allocation can be distorted. Authorities should examine foreseeable trade-offs and system effects.

Evidence concerning evidence source should establish the conclusion should retain the material limitation and responsible body. A contrary reading would overlook that for unintended effects and displacement, analysts should test this safeguard against the design and population evidence. For evidence source, bodies should identify which data and records support implementation and outcome and should identify provenance, coverage, date and quality. Any conclusion must respect that multiple reports from one source are not independent corroboration.

evidence source requires a decision about claim strength should be revised as evidence changes. The public account remains incomplete unless it explains how the implication for claim strength and correction should be recorded. Applied to unintended effects and displacement, this requirement has a distinct attribution consequence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion.

114

Design

For unintended effects and displacement, a public improvement claim depends on consequences outside the stated target, including burden shifted to other learners or services. The principal risk is that success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources. Evidence producers should therefore examine foreseeable trade-offs and system effects.

Evidence concerning design should establish the conclusion should retain the material limitation and responsible body. A contrary reading would overlook that for unintended effects and displacement, analysts should test this safeguard against the design and population evidence. The design test asks which comparison and identification strategy support attribution. Authorities should explain assumptions and threats. The governing proposition is that no design eliminates the need for contextual judgement. Every claim should identify outcome, population, period, source and material limitation.

In assessing design, authorities must determine public language should state which one the evidence addresses. The evidence must therefore clarify how any departure should be disclosed with its likely effect on inference. Within unintended effects and displacement, the principle bears on the claim described in this part. Observed change, programme contribution and sole causal attribution are different propositions.

115

Comparability

Comparability is material because consequences outside the stated target, including burden shifted to other learners or services cannot be inferred from a favourable trend or completed activity. In this domain, success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources. The immediate safeguard is to examine foreseeable trade-offs and system effects.

comparability requires a decision about the conclusion should retain the material limitation and responsible body. The public account remains incomplete unless it explains how for unintended effects and displacement, analysts should test this safeguard against the design and population evidence. A defensible account should establish whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments, recognising that common labels can conceal changed instruments or participation. Evidence inconsistent with the preferred explanation should remain visible.[REF-70]

A defensible account of comparability distinguishes where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The evidence must therefore clarify how any departure should be disclosed with its likely effect on inference. Within unintended effects and displacement, the principle bears on the claim described in this part.

116

Uncertainty

The public-interest meaning of unintended effects and displacement concerns consequences outside the stated target, including burden shifted to other learners or services. If success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources, a credible result can be converted into an unsupported institutional claim. The required course is to examine foreseeable trade-offs and system effects.

In assessing uncertainty, authorities must determine the conclusion should retain the material limitation and responsible body. The institutional consequence follows from whether for unintended effects and displacement, analysts should test this safeguard against the design and population evidence. Under uncertainty, review concerns which sampling, measurement, model and missing-data limits apply. Public bodies should publish intervals and qualitative constraints, subject to the rule that precision of language should not exceed precision of evidence. Reasons for selecting the measure and comparison should be stated.

Evidence concerning uncertainty should establish any departure should be disclosed with its likely effect on inference. The resulting interpretation should show why within unintended effects and displacement, the principle bears on the claim described in this part. Subgroup and local evidence should be capable of qualifying a headline average. Improvement should not be claimed universally where material populations deteriorated or remained unobserved.

117

Equity

Analysis of unintended effects and displacement should begin with consequences outside the stated target, including burden shifted to other learners or services. A foreseeable failure arises where success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources. Authorities should examine foreseeable trade-offs and system effects.

Comparative interpretation of equity depends upon analysts should analyse distribution and access responsibly, because average improvement does not establish equitable improvement. The institutional consequence follows from whether revisions and breaks should be adjacent to the affected comparison. The conclusion should retain the material limitation and responsible body. For unintended effects and displacement, analysts should test this safeguard against the design and population evidence. The equity standard requires consideration of which populations received the intervention and benefit.

The central question in equity is the account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. The evidence must therefore clarify how these conditions support different decisions and should not be collapsed into one judgement. Any departure should be disclosed with its likely effect on inference. Within unintended effects and displacement, the principle bears on the claim described in this part.

118

Implementation

implementation requires a decision about unintended effects and displacement requires explicit metadata because it concerns consequences outside the stated target, including burden shifted to other learners or services. For the learners concerned, the decisive consideration is whether the risk is that success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources. A competent producer should examine foreseeable trade-offs and system effects. The implication for claim strength and correction should be recorded. Applied to unintended effects and displacement, this requirement has a distinct attribution consequence.[REF-71]

Evidence concerning implementation should establish review of implementation should determine whether intended activity reached the population at sufficient quality and intensity. The resulting interpretation should show why it should link finance, delivery, use and service condition and acknowledge that an announced or funded programme is not evidence of exposure. Personal information should remain protected while unequal effect remains visible. The conclusion should retain the material limitation and responsible body. For unintended effects and displacement, analysts should test this safeguard against the design and population evidence.

Institutional action on implementation should be tested against any departure should be disclosed with its likely effect on inference. The resulting interpretation should show why within unintended effects and displacement, the principle bears on the claim described in this part. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear.

119

Alternative explanation

Review of alternative explanation is credible only where it explains where success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources, public confidence and resource allocation can be distorted. The institutional consequence follows from whether authorities should examine foreseeable trade-offs and system effects. The implication for claim strength and correction should be recorded. Applied to unintended effects and displacement, this requirement has a distinct attribution consequence. The public purpose of analysis on unintended effects and displacement is a responsible account of consequences outside the stated target, including burden shifted to other learners or services.

alternative explanation requires a decision about the conclusion should retain the material limitation and responsible body. A contrary reading would overlook that for unintended effects and displacement, analysts should test this safeguard against the design and population evidence. For alternative explanation, bodies should identify which other changes could plausibly produce the finding and should test or acknowledge concurrent factors. Any conclusion must respect that chronology and association alone do not establish causation.

Institutional action on alternative explanation should be tested against any departure should be disclosed with its likely effect on inference. A proportionate conclusion must also recognise that within unintended effects and displacement, the principle bears on the claim described in this part. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes.

120

Proportionality

The central question in proportionality is evidence producers should therefore examine foreseeable trade-offs and system effects. The resulting interpretation should show why the implication for claim strength and correction should be recorded. Applied to unintended effects and displacement, this requirement has a distinct attribution consequence. For unintended effects and displacement, a public improvement claim depends on consequences outside the stated target, including burden shifted to other learners or services. The principal risk is that success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources.

Review of proportionality is credible only where it explains every claim should identify outcome, population, period, source and material limitation. The institutional consequence follows from whether the conclusion should retain the material limitation and responsible body. For unintended effects and displacement, analysts should test this safeguard against the design and population evidence. The proportionality test asks how strength of evidence relates to the consequence and certainty of the claim. Authorities should use bounded language and distinguish contribution from attribution. The governing proposition is that higher-stakes claims require stronger support and clearer qualification.

Evidence concerning proportionality should establish the conclusion should retain the material limitation and responsible body. This matters because for unintended effects and displacement, analysts should test this safeguard against the design and population evidence. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses.[REF-72]

121

Review and correction

Review and correction is material because consequences outside the stated target, including burden shifted to other learners or services cannot be inferred from a favourable trend or completed activity. In this domain, success on one measure can be achieved by narrowing curriculum, excluding participants or reallocating resources. The immediate safeguard is to examine foreseeable trade-offs and system effects.

Comparative interpretation of review and correction depends upon evidence inconsistent with the preferred explanation should remain visible. The resulting interpretation should show why the conclusion should retain the material limitation and responsible body. For unintended effects and displacement, analysts should test this safeguard against the design and population evidence. A defensible account should establish how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction, recognising that quiet technical amendment is insufficient where the public relied on a stronger statement.

review and correction requires a decision about where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The evidence must therefore clarify how the conclusion should retain the material limitation and responsible body. For unintended effects and displacement, analysts should test this safeguard against the design and population evidence.

Part XII

Scale, durability and transfer

122

Claim proposition

The relevant claim concerns whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. The principal risk is that pilot or short-term findings can be presented as system-wide and sustainable. Authorities should state coverage, duration, resources, context and limits on generalisation.[REF-69]

123

Definition

The public-interest meaning of scale, durability and transfer concerns whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. If pilot or short-term findings can be presented as system-wide and sustainable, a credible result can be converted into an unsupported institutional claim. The required course is to state coverage, duration, resources, context and limits on generalisation.

definition cannot be judged without identifying reasons for selecting the measure and comparison should be stated. The institutional consequence follows from whether the conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence. Under definition, review concerns which intervention, outcome, population and period the statement concerns. Public bodies should publish an exact claim and unit, subject to the rule that improvement is not a self-defining measure.

definition cannot be judged without identifying the conclusion should retain the material limitation and responsible body. A contrary reading would overlook that for scale, durability and transfer, analysts should test this safeguard against the design and population evidence. Subgroup and local evidence should be capable of qualifying a headline average. Improvement should not be claimed universally where material populations deteriorated or remained unobserved.

124

Evidence source

Analysis of scale, durability and transfer should begin with whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. A foreseeable failure arises where pilot or short-term findings can be presented as system-wide and sustainable. Authorities should state coverage, duration, resources, context and limits on generalisation.

Evidence concerning evidence source should establish revisions and breaks should be adjacent to the affected comparison. The institutional consequence follows from whether the conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence. The evidence source standard requires consideration of which data and records support implementation and outcome. Analysts should identify provenance, coverage, date and quality, because multiple reports from one source are not independent corroboration.[REF-73]

The central question in evidence source is these conditions support different decisions and should not be collapsed into one judgement. This matters because the conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence.

125

Design

Scale, durability and transfer requires explicit metadata because it concerns whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. The risk is that pilot or short-term findings can be presented as system-wide and sustainable. A competent producer should state coverage, duration, resources, context and limits on generalisation.

In assessing design, authorities must determine review of design should determine which comparison and identification strategy support attribution. For the learners concerned, the decisive consideration is whether it should explain assumptions and threats and acknowledge that no design eliminates the need for contextual judgement. Personal information should remain protected while unequal effect remains visible. The conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence.

Public responsibility for design begins with the conclusion should retain the material limitation and responsible body. The public account remains incomplete unless it explains how for scale, durability and transfer, analysts should test this safeguard against the design and population evidence. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear.

126

Comparability

The public purpose of analysis on scale, durability and transfer is a responsible account of whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. Where pilot or short-term findings can be presented as system-wide and sustainable, public confidence and resource allocation can be distorted. Authorities should state coverage, duration, resources, context and limits on generalisation.

For comparability, the material distinction is between for comparability, bodies should identify whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments. The resulting interpretation should show why any conclusion must respect that common labels can conceal changed instruments or participation. The conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence.

Comparative interpretation of comparability depends upon the final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. The resulting interpretation should show why claim strength should be revised as evidence changes. The conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence.

127

Uncertainty

For scale, durability and transfer, a public improvement claim depends on whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. The principal risk is that pilot or short-term findings can be presented as system-wide and sustainable. Evidence producers should therefore state coverage, duration, resources, context and limits on generalisation.[REF-05]

The practical standard for uncertainty concerns authorities should publish intervals and qualitative constraints. The evidence must therefore clarify how the governing proposition is that precision of language should not exceed precision of evidence. Every claim should identify outcome, population, period, source and material limitation. The conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence. The uncertainty test asks which sampling, measurement, model and missing-data limits apply.

A defensible account of uncertainty distinguishes public language should state which one the evidence addresses. For the learners concerned, the decisive consideration is whether the implication for claim strength and correction should be recorded. Applied to scale, durability and transfer, this requirement has a distinct attribution consequence. Observed change, programme contribution and sole causal attribution are different propositions.

128

Equity

Equity is material because whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions cannot be inferred from a favourable trend or completed activity. In this domain, pilot or short-term findings can be presented as system-wide and sustainable. The immediate safeguard is to state coverage, duration, resources, context and limits on generalisation.

Review of equity is credible only where it explains evidence inconsistent with the preferred explanation should remain visible. The evidence must therefore clarify how the conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence. A defensible account should establish which populations received the intervention and benefit and should analyse distribution and access responsibly, recognising that average improvement does not establish equitable improvement.

A defensible account of equity distinguishes the implication for claim strength and correction should be recorded. For the learners concerned, the decisive consideration is whether applied to scale, durability and transfer, this requirement has a distinct attribution consequence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

129

Implementation

In assessing implementation, authorities must determine the implication for claim strength and correction should be recorded. The resulting interpretation should show why applied to scale, durability and transfer, this requirement has a distinct attribution consequence. The public-interest meaning of scale, durability and transfer concerns whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. If pilot or short-term findings can be presented as system-wide and sustainable, a credible result can be converted into an unsupported institutional claim. The required course is to state coverage, duration, resources, context and limits on generalisation.

For implementation, the material distinction is between the conclusion should retain the material limitation and responsible body. The resulting interpretation should show why for scale, durability and transfer, analysts should test this safeguard against the design and population evidence. Under implementation, review concerns whether intended activity reached the population at sufficient quality and intensity. Public bodies should link finance, delivery, use and service condition, subject to the rule that an announced or funded programme is not evidence of exposure. Reasons for selecting the measure and comparison should be stated.

For implementation, the material distinction is between subgroup and local evidence should be capable of qualifying a headline average. This matters because improvement should not be claimed universally where material populations deteriorated or remained unobserved. The implication for claim strength and correction should be recorded. Applied to scale, durability and transfer, this requirement has a distinct attribution consequence.[REF-10]

130

Alternative explanation

Evidence concerning alternative explanation should establish authorities should state coverage, duration, resources, context and limits on generalisation. This matters because the implication for claim strength and correction should be recorded. Applied to scale, durability and transfer, this requirement has a distinct attribution consequence. Analysis of scale, durability and transfer should begin with whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. A foreseeable failure arises where pilot or short-term findings can be presented as system-wide and sustainable.

Public responsibility for alternative explanation begins with analysts should test or acknowledge concurrent factors, because chronology and association alone do not establish causation. The public account remains incomplete unless it explains how revisions and breaks should be adjacent to the affected comparison. The conclusion should retain the material limitation and responsible body. For scale, durability and transfer, analysts should test this safeguard against the design and population evidence. The alternative explanation standard requires consideration of which other changes could plausibly produce the finding.

alternative explanation cannot be judged without identifying the account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. For the learners concerned, the decisive consideration is whether these conditions support different decisions and should not be collapsed into one judgement. The implication for claim strength and correction should be recorded. Applied to scale, durability and transfer, this requirement has a distinct attribution consequence.

131

Proportionality

Review of proportionality is credible only where it explains the institutional consequence follows from whether applied to scale, durability and transfer, this requirement has a distinct attribution consequence. For the learners concerned, the decisive consideration is whether scale, durability and transfer requires explicit metadata because it concerns whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. The risk is that pilot or short-term findings can be presented as system-wide and sustainable. A competent producer should state coverage, duration, resources, context and limits on generalisation. Review of proportionality is credible only where it explains the implication for claim strength and correction should be recorded.

For scale, durability and transfer, analysts should test this safeguard against the design and population evidence. Review of proportionality should determine how strength of evidence relates to the consequence and certainty of the claim. It should use bounded language and distinguish contribution from attribution and acknowledge that higher-stakes claims require stronger support and clearer qualification. Personal information should remain protected while unequal effect remains visible. The conclusion should retain the material limitation and responsible body.

Evidence concerning proportionality should establish responsible communication preserves value by making the evidentiary boundary clear. The resulting interpretation should show why the implication for claim strength and correction should be recorded. Applied to scale, durability and transfer, this requirement has a distinct attribution consequence. A result can be important without supporting every explanation attached to it.

132

Review and correction

review and correction cannot be judged without identifying the implication for claim strength and correction should be recorded. The resulting interpretation should show why applied to scale, durability and transfer, this requirement has a distinct attribution consequence. The public purpose of analysis on scale, durability and transfer is a responsible account of whether an observed result persists, extends beyond selected settings and is feasible under ordinary conditions. Where pilot or short-term findings can be presented as system-wide and sustainable, public confidence and resource allocation can be distorted. Authorities should state coverage, duration, resources, context and limits on generalisation.

For scale, durability and transfer, analysts should test this safeguard against the design and population evidence. For review and correction, bodies should identify how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction. Any conclusion must respect that quiet technical amendment is insufficient where the public relied on a stronger statement. The conclusion should retain the material limitation and responsible body.[REF-12]

Comparative interpretation of review and correction depends upon claim strength should be revised as evidence changes. The institutional consequence follows from whether the implication for claim strength and correction should be recorded. Applied to scale, durability and transfer, this requirement has a distinct attribution consequence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion.

Part XIII

Uncertainty and sensitivity

133

Claim proposition

The relevant claim concerns the statistical, measurement and assumption limits affecting the estimate and attribution. The principal risk is that one specification or precise percentage can conceal a wide range of plausible results. Authorities should report uncertainty and examine reasonable alternative definitions or models.[REF-70]

134

Definition

For uncertainty and sensitivity, a public improvement claim depends on the statistical, measurement and assumption limits affecting the estimate and attribution. The principal risk is that one specification or precise percentage can conceal a wide range of plausible results. Evidence producers should therefore report uncertainty and examine reasonable alternative definitions or models.

definition requires a decision about authorities should publish an exact claim and unit. The public account remains incomplete unless it explains how the governing proposition is that improvement is not a self-defining measure. Every claim should identify outcome, population, period, source and material limitation. The implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. The definition test asks which intervention, outcome, population and period the statement concerns.

Comparative interpretation of definition depends upon observed change, programme contribution and sole causal attribution are different propositions. The evidence must therefore clarify how public language should state which one the evidence addresses. Any departure should be disclosed with its likely effect on inference. Within uncertainty and sensitivity, the principle bears on the claim described in this part.

135

Evidence source

Evidence source is material because the statistical, measurement and assumption limits affecting the estimate and attribution cannot be inferred from a favourable trend or completed activity. In this domain, one specification or precise percentage can conceal a wide range of plausible results. The immediate safeguard is to report uncertainty and examine reasonable alternative definitions or models.

In assessing evidence source, authorities must determine the implication for claim strength and correction should be recorded. The evidence must therefore clarify how applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. A defensible account should establish which data and records support implementation and outcome and should identify provenance, coverage, date and quality, recognising that multiple reports from one source are not independent corroboration. Evidence inconsistent with the preferred explanation should remain visible.

evidence source requires a decision about where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. The evidence must therefore clarify how any departure should be disclosed with its likely effect on inference. Within uncertainty and sensitivity, the principle bears on the claim described in this part.

136

Design

The public-interest meaning of uncertainty and sensitivity concerns the statistical, measurement and assumption limits affecting the estimate and attribution. If one specification or precise percentage can conceal a wide range of plausible results, a credible result can be converted into an unsupported institutional claim. The required course is to report uncertainty and examine reasonable alternative definitions or models.[REF-70]

The practical standard for design concerns under design, review concerns which comparison and identification strategy support attribution. This matters because public bodies should explain assumptions and threats, subject to the rule that no design eliminates the need for contextual judgement. Reasons for selecting the measure and comparison should be stated. The implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence.

A defensible account of design distinguishes improvement should not be claimed universally where material populations deteriorated or remained unobserved. The institutional consequence follows from whether any departure should be disclosed with its likely effect on inference. Within uncertainty and sensitivity, the principle bears on the claim described in this part. Subgroup and local evidence should be capable of qualifying a headline average.

137

Comparability

Analysis of uncertainty and sensitivity should begin with the statistical, measurement and assumption limits affecting the estimate and attribution. A foreseeable failure arises where one specification or precise percentage can conceal a wide range of plausible results. Authorities should report uncertainty and examine reasonable alternative definitions or models.

Evidence concerning comparability should establish revisions and breaks should be adjacent to the affected comparison. A proportionate conclusion must also recognise that the implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. The comparability standard requires consideration of whether measures, populations and conditions are stable enough across observations. Analysts should document breaks and adjustments, because common labels can conceal changed instruments or participation.

A defensible account of comparability distinguishes the account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. This matters because these conditions support different decisions and should not be collapsed into one judgement. Any departure should be disclosed with its likely effect on inference. Within uncertainty and sensitivity, the principle bears on the claim described in this part.

138

Uncertainty

Uncertainty and sensitivity requires explicit metadata because it concerns the statistical, measurement and assumption limits affecting the estimate and attribution. The risk is that one specification or precise percentage can conceal a wide range of plausible results. A competent producer should report uncertainty and examine reasonable alternative definitions or models.

uncertainty requires a decision about review of uncertainty should determine which sampling, measurement, model and missing-data limits apply. The institutional consequence follows from whether it should publish intervals and qualitative constraints and acknowledge that precision of language should not exceed precision of evidence. Personal information should remain protected while unequal effect remains visible. The implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence.

Public responsibility for uncertainty begins with any departure should be disclosed with its likely effect on inference. For the learners concerned, the decisive consideration is whether within uncertainty and sensitivity, the principle bears on the claim described in this part. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear.[REF-71]

139

Equity

The public purpose of analysis on uncertainty and sensitivity is a responsible account of the statistical, measurement and assumption limits affecting the estimate and attribution. Where one specification or precise percentage can conceal a wide range of plausible results, public confidence and resource allocation can be distorted. Authorities should report uncertainty and examine reasonable alternative definitions or models.

Institutional action on equity should be tested against any conclusion must respect that average improvement does not establish equitable improvement. The resulting interpretation should show why the implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. For equity, bodies should identify which populations received the intervention and benefit and should analyse distribution and access responsibly.

Public responsibility for equity begins with the final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. The public account remains incomplete unless it explains how claim strength should be revised as evidence changes. Any departure should be disclosed with its likely effect on inference. Within uncertainty and sensitivity, the principle bears on the claim described in this part.

140

Implementation

implementation requires a decision about for uncertainty and sensitivity, a public improvement claim depends on the statistical, measurement and assumption limits affecting the estimate and attribution. The institutional consequence follows from whether the principal risk is that one specification or precise percentage can conceal a wide range of plausible results. Evidence producers should therefore report uncertainty and examine reasonable alternative definitions or models. The implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence.

implementation requires a decision about the governing proposition is that an announced or funded programme is not evidence of exposure. A contrary reading would overlook that every claim should identify outcome, population, period, source and material limitation. The implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. The implementation test asks whether intended activity reached the population at sufficient quality and intensity. Authorities should link finance, delivery, use and service condition.

The central question in implementation is public language should state which one the evidence addresses. The evidence must therefore clarify how the conclusion should retain the material limitation and responsible body. For uncertainty and sensitivity, analysts should test this safeguard against the design and population evidence. Observed change, programme contribution and sole causal attribution are different propositions.

141

Alternative explanation

Alternative explanation is material because the statistical, measurement and assumption limits affecting the estimate and attribution cannot be inferred from a favourable trend or completed activity. In this domain, one specification or precise percentage can conceal a wide range of plausible results. The immediate safeguard is to report uncertainty and examine reasonable alternative definitions or models.

A defensible account of alternative explanation distinguishes evidence inconsistent with the preferred explanation should remain visible. For the learners concerned, the decisive consideration is whether the implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. A defensible account should establish which other changes could plausibly produce the finding and should test or acknowledge concurrent factors, recognising that chronology and association alone do not establish causation.[REF-72]

In assessing alternative explanation, authorities must determine the conclusion should retain the material limitation and responsible body. For the learners concerned, the decisive consideration is whether for uncertainty and sensitivity, analysts should test this safeguard against the design and population evidence. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly.

142

Proportionality

For proportionality, the material distinction is between the public-interest meaning of uncertainty and sensitivity concerns the statistical, measurement and assumption limits affecting the estimate and attribution. The institutional consequence follows from whether if one specification or precise percentage can conceal a wide range of plausible results, a credible result can be converted into an unsupported institutional claim. The required course is to report uncertainty and examine reasonable alternative definitions or models. The implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence.

Public responsibility for proportionality begins with public bodies should use bounded language and distinguish contribution from attribution, subject to the rule that higher-stakes claims require stronger support and clearer qualification. The institutional consequence follows from whether reasons for selecting the measure and comparison should be stated. The implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. Under proportionality, review concerns how strength of evidence relates to the consequence and certainty of the claim.

Evidence concerning proportionality should establish the conclusion should retain the material limitation and responsible body. This matters because for uncertainty and sensitivity, analysts should test this safeguard against the design and population evidence. Subgroup and local evidence should be capable of qualifying a headline average. Improvement should not be claimed universally where material populations deteriorated or remained unobserved.

143

Review and correction

In assessing review and correction, authorities must determine a foreseeable failure arises where one specification or precise percentage can conceal a wide range of plausible results. The evidence must therefore clarify how authorities should report uncertainty and examine reasonable alternative definitions or models. The implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. Analysis of uncertainty and sensitivity should begin with the statistical, measurement and assumption limits affecting the estimate and attribution.

Public responsibility for review and correction begins with revisions and breaks should be adjacent to the affected comparison. A proportionate conclusion must also recognise that the implication for claim strength and correction should be recorded. Applied to uncertainty and sensitivity, this requirement has a distinct attribution consequence. The review and correction standard requires consideration of how new evidence, error or changed method revises the claim. Analysts should retain versions, reasons and direct correction, because quiet technical amendment is insufficient where the public relied on a stronger statement.

review and correction cannot be judged without identifying the conclusion should retain the material limitation and responsible body. A proportionate conclusion must also recognise that for uncertainty and sensitivity, analysts should test this safeguard against the design and population evidence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.

Part XIV

Public communication and correction

144

Claim proposition

The relevant claim concerns the wording, evidence access, qualification and process for revising a public claim. The principal risk is that headline certainty can outlive later corrections or technical caveats. Authorities should place material limitations with the claim and correct affected audiences directly.[REF-71]

145

Definition

Public communication and correction requires explicit metadata because it concerns the wording, evidence access, qualification and process for revising a public claim. The risk is that headline certainty can outlive later corrections or technical caveats. A competent producer should place material limitations with the claim and correct affected audiences directly.[REF-73]

Evidence concerning definition should establish it should publish an exact claim and unit and acknowledge that improvement is not a self-defining measure. This matters because personal information should remain protected while unequal effect remains visible. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence. Review of definition should determine which intervention, outcome, population and period the statement concerns.

For public communication and correction, analysts should test this safeguard against the design and population evidence. A result can be important without supporting every explanation attached to it. Responsible communication preserves value by making the evidentiary boundary clear. The conclusion should retain the material limitation and responsible body.

146

Evidence source

The public purpose of analysis on public communication and correction is a responsible account of the wording, evidence access, qualification and process for revising a public claim. Where headline certainty can outlive later corrections or technical caveats, public confidence and resource allocation can be distorted. Authorities should place material limitations with the claim and correct affected audiences directly.

For evidence source, the material distinction is between any conclusion must respect that multiple reports from one source are not independent corroboration. For the learners concerned, the decisive consideration is whether the implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence. For evidence source, bodies should identify which data and records support implementation and outcome and should identify provenance, coverage, date and quality.

For public communication and correction, analysts should test this safeguard against the design and population evidence. The final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. Claim strength should be revised as evidence changes. The conclusion should retain the material limitation and responsible body.

147

Design

For public communication and correction, a public improvement claim depends on the wording, evidence access, qualification and process for revising a public claim. The principal risk is that headline certainty can outlive later corrections or technical caveats. Evidence producers should therefore place material limitations with the claim and correct affected audiences directly.

Public responsibility for design begins with authorities should explain assumptions and threats. This matters because the governing proposition is that no design eliminates the need for contextual judgement. Every claim should identify outcome, population, period, source and material limitation. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence. The design test asks which comparison and identification strategy support attribution.

The central question in design is public language should state which one the evidence addresses. This matters because the implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence. Observed change, programme contribution and sole causal attribution are different propositions.[REF-05]

148

Comparability

Comparability is material because the wording, evidence access, qualification and process for revising a public claim cannot be inferred from a favourable trend or completed activity. In this domain, headline certainty can outlive later corrections or technical caveats. The immediate safeguard is to place material limitations with the claim and correct affected audiences directly.

For comparability, the material distinction is between the implication for claim strength and correction should be recorded. A contrary reading would overlook that applied to public communication and correction, this requirement has a distinct attribution consequence. A defensible account should establish whether measures, populations and conditions are stable enough across observations and should document breaks and adjustments, recognising that common labels can conceal changed instruments or participation. Evidence inconsistent with the preferred explanation should remain visible.

Institutional action on comparability should be tested against where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. A contrary reading would overlook that the implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence.

149

Uncertainty

The public-interest meaning of public communication and correction concerns the wording, evidence access, qualification and process for revising a public claim. If headline certainty can outlive later corrections or technical caveats, a credible result can be converted into an unsupported institutional claim. The required course is to place material limitations with the claim and correct affected audiences directly.

uncertainty requires a decision about public bodies should publish intervals and qualitative constraints, subject to the rule that precision of language should not exceed precision of evidence. The institutional consequence follows from whether reasons for selecting the measure and comparison should be stated. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence. Under uncertainty, review concerns which sampling, measurement, model and missing-data limits apply.

For uncertainty, the material distinction is between subgroup and local evidence should be capable of qualifying a headline average. This matters because improvement should not be claimed universally where material populations deteriorated or remained unobserved. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence.

150

Equity

Analysis of public communication and correction should begin with the wording, evidence access, qualification and process for revising a public claim. A foreseeable failure arises where headline certainty can outlive later corrections or technical caveats. Authorities should place material limitations with the claim and correct affected audiences directly.

Review of equity is credible only where it explains analysts should analyse distribution and access responsibly, because average improvement does not establish equitable improvement. A proportionate conclusion must also recognise that revisions and breaks should be adjacent to the affected comparison. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence. The equity standard requires consideration of which populations received the intervention and benefit.

The practical standard for equity concerns the implication for claim strength and correction should be recorded. The public account remains incomplete unless it explains how applied to public communication and correction, this requirement has a distinct attribution consequence. The account should distinguish implementation failure, theory failure, measurement failure and insufficient evidence. These conditions support different decisions and should not be collapsed into one judgement.

151

Implementation

implementation requires a decision about public communication and correction requires explicit metadata because it concerns the wording, evidence access, qualification and process for revising a public claim. A proportionate conclusion must also recognise that the risk is that headline certainty can outlive later corrections or technical caveats. A competent producer should place material limitations with the claim and correct affected audiences directly. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence.

implementation requires a decision about personal information should remain protected while unequal effect remains visible. A contrary reading would overlook that the implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence. Review of implementation should determine whether intended activity reached the population at sufficient quality and intensity. It should link finance, delivery, use and service condition and acknowledge that an announced or funded programme is not evidence of exposure.

Institutional action on implementation should be tested against a result can be important without supporting every explanation attached to it. The resulting interpretation should show why responsible communication preserves value by making the evidentiary boundary clear. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence.

152

Alternative explanation

The practical standard for alternative explanation concerns where headline certainty can outlive later corrections or technical caveats, public confidence and resource allocation can be distorted. The evidence must therefore clarify how authorities should place material limitations with the claim and correct affected audiences directly. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence. The public purpose of analysis on public communication and correction is a responsible account of the wording, evidence access, qualification and process for revising a public claim.

The central question in alternative explanation is for alternative explanation, bodies should identify which other changes could plausibly produce the finding and should test or acknowledge concurrent factors. The evidence must therefore clarify how any conclusion must respect that chronology and association alone do not establish causation. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence.

Evidence concerning alternative explanation should establish the final account should explain which decision can reasonably follow, what remains uncertain and what evidence would change the conclusion. A proportionate conclusion must also recognise that claim strength should be revised as evidence changes. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence.

153

Proportionality

In assessing proportionality, authorities must determine for public communication and correction, a public improvement claim depends on the wording, evidence access, qualification and process for revising a public claim. The evidence must therefore clarify how the principal risk is that headline certainty can outlive later corrections or technical caveats. Evidence producers should therefore place material limitations with the claim and correct affected audiences directly. The implication for claim strength and correction should be recorded. Applied to public communication and correction, this requirement has a distinct attribution consequence.

The central question in proportionality is the implication for claim strength and correction should be recorded. The public account remains incomplete unless it explains how applied to public communication and correction, this requirement has a distinct attribution consequence. The proportionality test asks how strength of evidence relates to the consequence and certainty of the claim. Authorities should use bounded language and distinguish contribution from attribution. The governing proposition is that higher-stakes claims require stronger support and clearer qualification. Every claim should identify outcome, population, period, source and material limitation.

Within public communication and correction, the principle bears on the claim described in this part. Observed change, programme contribution and sole causal attribution are different propositions. Public language should state which one the evidence addresses. Any departure should be disclosed with its likely effect on inference.

154

Review and correction

Review and correction is material because the wording, evidence access, qualification and process for revising a public claim cannot be inferred from a favourable trend or completed activity. In this domain, headline certainty can outlive later corrections or technical caveats. The immediate safeguard is to place material limitations with the claim and correct affected audiences directly.

Applied to public communication and correction, this requirement has a distinct attribution consequence. A defensible account should establish how new evidence, error or changed method revises the claim and should retain versions, reasons and direct correction, recognising that quiet technical amendment is insufficient where the public relied on a stronger statement. Evidence inconsistent with the preferred explanation should remain visible. The implication for claim strength and correction should be recorded.

Within public communication and correction, the principle bears on the claim described in this part. Where causal identification is weak, a contribution account may still be useful if implementation, sequence, mechanism and alternative explanations are examined explicitly. Any departure should be disclosed with its likely effect on inference.

Part XV

Conclusions

155

A minimum public claim

Every claim should state what changed, for whom, when, against what baseline and on which evidence. It should distinguish descriptive result, contribution and attribution and place material limitations with the conclusion.

156

Evidence strength and public language

Language should be proportionate to design, implementation evidence, comparability and uncertainty. A bounded claim is not a weak claim; it is one whose public meaning can be tested and corrected.

157

Final conclusion

Educational improvement deserves accurate recognition. Public accountability is strengthened when institutions communicate important results without extending them beyond the population, outcome, period or causal inference supported by evidence.

References

  1. REF-05

    UNESCO and UNICEF. A Human Rights-Based Approach to Education for All. 2007.

    Rights-based duties concerning access, quality, participation and accountability.

    https://unesdoc.unesco.org/ark:/48223/pf0000154861
  2. REF-10

    UNESCO. Guidelines for Inclusion: Ensuring Access to Education for All. 2005.

    Inclusive system reform, barriers to participation and learner diversity.

    https://unesdoc.unesco.org/ark:/48223/pf0000140224
  3. REF-12

    United Nations General Assembly. Convention on the Rights of Persons with Disabilities. 2006.

    Inclusive education, accessibility and reasonable accommodation.

    https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-persons-disabilities
  4. REF-15

    United Nations Children’s Fund. Child Friendly Schools Manual. 2009.

    Learner-centred, inclusive, protective and community-linked school design.

    https://www.unicef.org/reports/child-friendly-schools-manual
  5. REF-61

    United Nations Educational, Scientific and Cultural Organization. Revised Recommendation concerning Technical and Vocational Education. 2001.

    International normative guidance on TVET aims, programmes, teachers, guidance and cooperation.

    https://unesdoc.unesco.org/ark:/48223/pf0000126050
  6. REF-62

    UNESCO and International Labour Organization. Technical and Vocational Education and Training for the Twenty-first Century: UNESCO and ILO Recommendations. 2002.

    Joined international guidance on TVET, work, lifelong learning and institutional responsibility.

    https://unesdoc.unesco.org/ark:/48223/pf0000126050
  7. REF-63

    International Labour Organization. Human Resources Development Recommendation, 2004 (No. 195). 2004.

    International labour standard on education, training, lifelong learning, competencies and recognition.

    https://normlex.ilo.org/dyn/nrmlx_en/f?p=NORMLEXPUB:12100:0::NO::P12100_ILO_CODE:R195
  8. REF-64

    European Parliament and Council of the European Union. Recommendation on the Establishment of the European Qualifications Framework for Lifelong Learning. 2008.

    Learning-outcomes-based reference levels expressed through knowledge, skills and competence.

    https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:32008H0506(01)
  9. REF-65

    European Parliament and Council of the European Union. Recommendation on the Establishment of a European Credit System for Vocational Education and Training. 2009.

    Units of learning outcomes, credit, assessment, validation, recognition and mobility.

    https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:32009H0708(02)
  10. REF-66

    European Parliament and Council of the European Union. Recommendation on the Establishment of a European Quality Assurance Reference Framework for Vocational Education and Training. 2009.

    European quality-assurance framework for TVET planning, implementation, evaluation and review.

    https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:32009H0708(01)
  11. REF-67

    Ministers Responsible for Higher Education in the Bologna Countries. The Framework for Qualifications of the European Higher Education Area. 2005.

    Qualifications framework using learning outcomes, cycles, workload and qualification descriptors.

    https://ehea.info/media.ehea.info/file/WG_Frameworks_qualification/85/2/Framework_qualificationsforEHEA-May2005_587852.pdf
  12. REF-69

    Ministers Responsible for Higher Education in the Bologna Countries. Budapest–Vienna Declaration on the European Higher Education Area. 2010.

    Contemporaneous statement on the European Higher Education Area and implementation of Bologna commitments.

    https://ehea.info/Upload/document/ministerial_declarations/Budapest_Vienna_Declaration_598640.pdf
  13. REF-70

    European Centre for the Development of Vocational Training. The Shift to Learning Outcomes: Policies and Practices in Europe. 2009.

    Comparative European analysis of learning-outcomes approaches in education and training.

    https://www.cedefop.europa.eu/files/3054_en.pdf
  14. REF-71

    European Centre for the Development of Vocational Training. Learning Outcomes Approaches in VET Curricula: A Comparative Analysis of Nine European Countries. 2010.

    Evidence on specification and use of learning outcomes in vocational curricula.

    https://www.cedefop.europa.eu/files/5506_en.pdf
  15. REF-72

    European Commission. New Skills for New Jobs: Anticipating and Matching Labour Market and Skills Needs. 2008.

    European policy on skills anticipation, matching, adaptability and labour-market change.

    https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52008DC0868
  16. REF-73

    European Ministers for Vocational Education and Training, European Social Partners and European Commission. The Bruges Communiqué on Enhanced European Cooperation in Vocational Education and Training for 2011–2020. 2010.

    European TVET cooperation priorities on quality, outcomes, mobility, inclusion and lifelong learning.

    https://www.cedefop.europa.eu/files/bruges_en.pdf
  17. REF-75

    United Nations General Assembly. Implementation of Agenda 21, the Programme for the Further Implementation of Agenda 21 and the Outcomes of the World Summit on Sustainable Development. 2011.

    Contemporaneous United Nations sustainable-development and conference-preparation context.

    https://undocs.org/A/RES/66/197
  18. REF-76

    United Nations Conference on Sustainable Development Preparatory Committee. The Future We Want: Zero Draft of the Outcome Document. 2012.

    Preparatory draft available by the cutoff, used only as evidence of current sustainable-development attention.

    https://sustainabledevelopment.un.org/content/documents/370The%20Future%20We%20Want%2010Jan%20clean.pdf
  19. REF-77

    European Commission. An Agenda for New Skills and Jobs: A European Contribution towards Full Employment. 2010.

    European policy concerning skills, qualifications, mobility and labour-market participation.

    https://eur-lex.europa.eu/legal-content/EN/TXT/?uri=CELEX:52010DC0682