Thematic Research Report

ICEQC-R-2005-05 — Strengthening Institutional Self-Review through Verifiable Improvement Evidence

A disciplined approach to institutional judgement, documented change and public accountability

Publication date
Research category
Quality Improvement Methods
Report archetype
Quality Improvement Study
Geographic scope
Global
Evidence cut-off date
Responsible body
ICEQC Research and Policy Directorate
International Council for Education Quality Certification

ICEQC-R-2005-05

Strengthening Institutional Self-Review through Verifiable Improvement Evidence

A disciplined approach to institutional judgement, documented change and public accountability

Publication date
Evidence cut-off date
Publication type
Thematic Research Report
Authoritative language
EN

Publication record

This is the controlled English edition. Evidence and institutional status are stated as at the evidence cut-off date.

Executive summary

Institutional self-review has value when it enables an education institution to make a reasoned judgement about its work, act on that judgement and demonstrate whether the intended condition changed. It loses value when it becomes an annual description, a collection of favourable activity counts or a preparatory exercise designed principally to satisfy an external visit. The distinction lies less in the volume of evidence than in the integrity of the chain connecting purpose, finding, decision, implementation and result.

This report sets out a method for self-review based on verifiable improvement evidence. It applies to schools, colleges and higher education institutions, subject to the different legal, academic and administrative arrangements governing each sector. It does not establish common institutional requirements across those sectors. It identifies the evidentiary disciplines that remain necessary whenever an institution examines its own provision: a defined question; an explicit criterion or intended condition; evidence proportionate to the claim; participation by those affected; a recorded judgement; a feasible improvement decision; and later verification of implementation and result.

The policy position available by 7 June 2005 supports a balanced conception. The 2001 Recommendation of the European Parliament and the Council on quality evaluation in school education encourages self-evaluation as a means of creating learning and improving schools, within a framework that also provides an external perspective and methodological support. It calls for the involvement of teachers, pupils, management, parents and experts, and warns against external evaluation restricted to administrative checks. The recommendation does not treat self-evaluation as institutional self-certification. It situates internal enquiry within public responsibility, participation and informed scrutiny.

The 2005 Bergen Communiqué extends the contemporaneous policy significance of internal institutional mechanisms in higher education. Ministers call for institutions to continue strengthening the quality of their activities through systematic internal arrangements directly related to external quality review, while also underlining institutional participation, student involvement, public responsibility and sustainable funding. This is a regional ministerial position for higher education, not a universal legal rule. Its relevance lies in the connection it makes between institutional responsibility and evidence that can withstand an outside view.

The 2005 Education for All Global Monitoring Report, published in 2004, provides the wider global context. It places the school at the centre of education improvement but observes that schools require capacity and continuing support. It distinguishes school improvement from a narrow response to externally imposed targets and emphasises teaching, learning, participation and institutional conditions. Self-review must accordingly examine the educational process and the distribution of experience, not merely compliance with administrative inputs.

The report distinguishes **institutional account**, **evaluative finding**, **improvement evidence** and **assurance evidence**. An institutional account describes provision and context. An evaluative finding compares evidence with a stated criterion and reaches a bounded judgement. Improvement evidence establishes whether an action was implemented and whether the selected condition changed. Assurance evidence permits another competent reader to understand the basis, limitations and consistency of the judgement. These functions can draw on the same records, but they are not interchangeable.

Self-review begins with a question that is capable of resolution. “How good is the institution?” is too broad to govern evidence. “Do first-year students in the two largest programmes receive the scheduled formative assessment and usable feedback early enough to revise subsequent work?” identifies population, provision, timing and educational purpose. It can be answered with curriculum documents, work samples, dates, student evidence and a defined comparison. The resulting judgement remains limited to the question; it does not become a general verdict on institutional quality.

Criteria come from legitimate sources: law, authorised curriculum, published institutional commitments, programme specifications, professional duties and properly adopted policy. They may also include an improvement objective set by the institution, provided that it does not displace a minimum public obligation. A self-review cannot declare success by lowering the criterion after evidence is known. Where the applicable requirement is unclear or several legitimate expectations conflict, the review records the ambiguity and seeks a competent interpretation.

Evidence is selected according to the proposition. Policy documents show intended arrangements; administrative records show transactions; surveys show reported experience; observation shows a condition at a time; assessment records show performance under specified tasks; interviews explain possible causes; financial records show authorised and actual use of resources. No source establishes more than it observes. Triangulation is valuable only where sources have sufficiently independent strengths and limitations. Three summaries copied from one underlying record remain one source.

Institutional evidence is particularly vulnerable to favourable selection. The institution knows its own work and can locate relevant records efficiently, but it also has interests in reputation, finance, enrolment and staff relations. A credible method addresses this condition openly. The review fixes the question and selection rule before examining outcomes, retains unfavourable and missing cases, separates evidence ownership from final judgement where feasible, and invites an informed challenge. These controls do not imply institutional bad faith. They protect the usefulness of institutional knowledge.

Participation is necessary for both validity and public interest. Learners can identify whether intended provision was accessible and usable; teachers and other staff can identify implementation constraints; leadership can explain authority and resource decisions; families or communities can provide evidence where relevant to access and experience. Participation does not mean that every view is treated as fact or that a majority determines a technical finding. The review records whose experience was sought, who was not reached, how disagreement was handled and which decisions remain with the competent authority.

Learning outcomes require restrained interpretation. A change in examination results can reflect intake, assessment difficulty, progression rules, missing candidates, teaching, attendance and wider conditions. Institutional self-review can use outcome evidence to identify a pattern and test a defined improvement proposition. It cannot infer institutional causation from a favourable trend alone. Comparative results require common definitions, stable populations and attention to uncertainty.

Improvement plans are not accepted as evidence of improvement. Each decision specifies the finding addressed, responsible role, authorised resources, actions, expected operating change, evidence, timetable, risk and review point. Milestones distinguish initiation, completion and operation. Training delivered, a policy approved or equipment purchased establishes activity; it does not establish changed teaching, learner access or institutional service.

Verification after action returns to the original criterion and population. It examines whether critical activities occurred, whether the intended group was reached, whether the condition changed, whether another group bore an unintended cost and whether the result is likely to continue. Where baseline and follow-up are not comparable, the institution reports current status and implementation without manufacturing a trend. Where the action was materially altered, the review assesses the implemented action rather than the original label.

External perspective has a methodological purpose. It can test whether criteria were applied consistently, whether contrary evidence was considered, whether the claimed result follows from the records and whether the institution has addressed a serious unresolved condition. It must not convert self-review into performance for an outsider. External challenge is strongest when it examines the reasoning and evidence chain, permits institutional response and leaves a clear record of disagreement.

Public reporting is proportionate to responsibility and risk. It states the questions reviewed, principal findings, material limitations, improvement decisions and later status. It protects personal information and does not publish small cells or case descriptions that permit identification. It distinguishes confirmed fact from institutional interpretation. A concise public account supported by an accessible method is more trustworthy than an extensive narrative in which weaknesses, missing evidence and unfinished action cannot be located.

The central conclusion is that self-review becomes credible through disciplined exposure to disconfirming evidence and subsequent verification. Institutional ownership is essential because improvement depends on those who understand and carry out the work. Ownership is not immunity from scrutiny. A strong institution can explain what it examined, why the evidence was adequate, what remained uncertain, what it changed and how another reader can see whether the educational condition improved.

Key findings

  • Institutional self-review is a reasoned process of enquiry and action. It is not a general institutional description, an annual compliance narrative or a claim that internal judgement is sufficient by itself.
  • A review question defines the population, period, provision and decision. Broad quality statements cannot determine what evidence is needed or what result would justify closure.
  • Criteria require a legitimate source and a stable meaning. An institution may set an improvement ambition but cannot use it to displace legal, curricular, rights-based or other applicable obligations.
  • Institutional account, evaluative finding, improvement evidence and assurance evidence perform different functions. Activity records do not establish evaluative conclusions, and favourable outcomes do not identify cause.
  • Evidence is proportionate to the claim. Documents, administrative records, observation, survey responses, interviews, work samples and outcomes each have distinct inferential limits.
  • Self-review must control favourable selection by fixing sampling and decision rules, retaining exceptions and missing cases, testing contrary explanations and permitting informed challenge.
  • Participation improves validity when it reaches the people affected and protects candour. It does not replace evidence standards or competent decision-making.
  • Learner outcomes require analysis of population, assessment, missingness, context and period. A change in results is not, by itself, evidence that an institutional action caused the change.
  • Improvement plans identify responsibility, resources, milestones, operating result, evidence and review date. A plan, meeting, purchase or training event is not a completed improvement.
  • Follow-up returns to the original finding and population. Changes to definitions, coverage, curriculum or assessment are disclosed and prevent unsupported trend claims.
  • External perspective tests reasoning, consistency and evidentiary sufficiency. It is not limited to administrative checking and does not remove institutional responsibility for improvement.
  • Public reporting separates fact, judgement, decision and unresolved uncertainty while protecting personal and sensitive information.
  • An institution demonstrates self-review capacity through repeated, verified improvement and correction of error, not through favourable language or the absence of recorded weakness.

Scope and method

This study concerns institutional self-review in school, vocational and higher education. It identifies common evidentiary disciplines while preserving sectoral differences in legal authority, curriculum, academic governance, learner age, programme structure and public accountability. The report does not assess named institutions, create sector requirements, authorise regulatory decisions or determine individual staff performance.

The evidence cut-off is 7 June 2005. The analysis uses international and regional policy instruments and official reports available by that date. The principal contemporary anchors are the 2001 European recommendation on school quality evaluation, the 2004 publication of the 2005 Education for All Global Monitoring Report and the Bergen ministerial communiqué of 20 May 2005. Rights instruments, international statistical principles and official comparative reports provide additional boundaries for participation, evidence integrity and outcome interpretation.

The method follows seven analytical operations: distinguish institutional description from evaluation; derive a criterion; formulate a bounded review question; construct a claim-to-evidence map; examine participation and bias; connect findings to controlled improvement action; and verify the result and its distribution. Worked examples are hypothetical. They demonstrate reasoning and calculation but do not describe an actual institution or jurisdiction.

No composite institutional-quality score is proposed. Education quality comprises conditions and outcomes that cannot be made interchangeable through arbitrary weights. Where aggregation is useful, the component values, definitions and exceptions remain visible. The method narrows conclusions when the evidence is incomplete rather than converting absence of evidence into a favourable result.

Part I

The purpose and standing of institutional self-review

1

Self-review as institutional enquiry

Self-review is a structured enquiry conducted by or under the authority of the institution whose work is examined. Its defining feature is not that every participant belongs to the institution. Its defining feature is that the institution accepts responsibility for the question, evidence, judgement and response.

The process is educational when it helps the institution understand how provision becomes learner experience and outcome. It is administrative when it records whether a procedure occurred. Both can be legitimate, but they answer different questions. A register showing that programme committees met can confirm governance activity; it cannot establish that a weakness in assessment was identified or corrected.

2

Four functions that require separation

An **institutional account** establishes context: programmes, learners, staff, resources, responsibilities and material change. An **evaluative finding** compares observed evidence with an explicit criterion. **Improvement evidence** follows a decision through implementation to an operating result. **Assurance evidence** permits scrutiny of the process and conclusion.

Confusion among the four creates characteristic errors. A detailed account is described as evaluation although it contains no judgement. A finding is treated as improvement although no action has operated. A favourable indicator is treated as assurance although its definition, coverage or selection cannot be examined. Separation allows the same document to serve several purposes without allowing one to stand in for another.

Table 1. Functions of institutional self-review evidence
FunctionPrimary questionTypical evidenceValid conclusionInvalid substitution
Institutional accountWhat provision, population and context existed?Registers, programme records, staffing, finance and contextual notesA dated description of the institutionDescription presented as quality judgement
Evaluative findingHow did a defined condition compare with a criterion?Selected records, observation, learner evidence and analysisA bounded judgement for the stated population and periodGeneral institutional rating from one measure
Improvement evidenceWas the response implemented and did the condition change?Action record, reach, operating verification and follow-upDecision to continue, adapt, expand or ceasePlan or expenditure presented as improvement
Assurance evidenceCan another informed party understand and test the basis?Method, sample, exceptions, sources, limitations and responseConfidence proportionate to transparency and verificationInstitutional assertion treated as independent confirmation

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC institutional self-review framework.

3

Improvement, accountability and learning

Improvement and accountability are not opposing purposes. An institution requires protected space to identify weakness and test change; learners and the public require truthful information about material conditions and unresolved risk. A system that punishes every recorded weakness encourages concealment. A system that keeps every weakness private removes accountability and may leave affected learners without remedy.

The balance depends on consequence. Early exploration of a teaching method can remain within professional enquiry. An unsafe facility, discriminatory barrier, unreliable award decision or prolonged absence of required teaching has a direct public-interest dimension and requires competent action. The review protocol identifies which matters may be addressed internally, which require notification and which require immediate protection.

4

The international policy setting in June 2005

The Dakar Framework connects quality and measurable learning outcomes with participation, national action, accountability and attention to learners in difficult circumstances. It does not prescribe one institutional self-review model. It establishes the public purpose against which an internal process is judged.[REF-01]

The European school recommendation gives explicit standing to self-evaluation as a means of school learning and improvement, balanced by an outside view, methodological support and participation. The instrument respects national responsibility for education-system organisation and acknowledges differences of context. It therefore supports principles rather than mechanical transfer of one national procedure.[REF-03]

The Bergen Communiqué addresses higher education within a defined regional reform process. Its call for systematic internal arrangements directly connected with external quality review, and for greater student involvement, reflects a contemporaneous expectation that institutional autonomy and public responsibility operate together. The commitments are not attributed beyond participating systems or converted into a global mandate.[REF-04]

5

The institution as a unit of analysis

An institution is rarely homogeneous. Programmes, sites, grades, languages, modes of study and learner groups can experience different conditions. An institution-wide average can therefore conceal the unit where improvement is required.

The unit of analysis follows the question. It may be a programme, course, class, service, cohort, award decision, facility or the institution as a whole. Findings are aggregated only when definitions and relevant conditions permit. A concern found in one programme does not support an adverse conclusion about all provision; a favourable institutional average does not resolve the programme concern.

6

Boundaries of institutional authority

Self-review identifies action within and beyond institutional control. Timetable organisation, local academic practice and access to existing resources may be institutionally controlled. Staffing establishment, public finance, national curriculum, major construction or statutory award rules may rest elsewhere.

An institution does not fail merely because it cannot unilaterally resolve an external constraint. It does, however, remain responsible for accurate identification, interim protection within its authority, documented referral and follow-up. The external authority's lack of response remains part of the improvement record.

Part II

From broad concern to a reviewable question

7

Why broad themes are insufficient

Themes such as teaching quality, student support, governance or resources are useful for organising a review programme but are not questions. They do not identify the population, expected condition, evidence or decision. They tend to produce broad descriptions and lists of activities.

A reviewable question contains an observable subject and a reason for enquiry. It can be answered within a defined period and leads to a decision. The institution may begin with a broad concern, but it narrows that concern before collecting evidence.

8

The question statement

The question statement records: the educational provision or service; intended population; period; criterion; observed concern; decision that the review will inform; and material exclusions. It avoids language that assumes the conclusion.

“Why are students dissatisfied with feedback?” assumes both dissatisfaction and a single causal explanation. A neutral question is: “What proportion of assessed work in the first semester received feedback within the published period, what did the feedback enable students to do, and which conditions explain material variation among programmes?” This formulation permits confirmation, qualification or rejection of the original concern.

9

Sources of criteria

Criteria can be binding, authorised or developmental. Binding criteria derive from applicable law, rights, safety duties or formally governing rules. Authorised criteria derive from approved curriculum, programme specifications, published student commitments and adopted institutional policy. Developmental criteria express an improvement objective beyond the current minimum.

The status is stated because consequences differ. Failure against a legal access duty is not managed in the same way as incomplete progress toward an aspirational objective. A developmental ambition cannot lower a binding obligation, and a local policy cannot be treated as law.

10

Interpreting qualitative criteria

Terms such as adequate, timely, effective, appropriate and sufficient require operational interpretation. The review identifies the object, population, period and educational consequence. “Timely feedback” may mean feedback before the next related task, within a published number of days or at another point justified by programme design.

Operational interpretation is not an opportunity to choose a convenient threshold after results are known. The institution records the source and rationale before analysis. Where professional judgement remains necessary, more than one informed reviewer examines boundary cases and records disagreement.

11

Multiple and conflicting expectations

An institution may face tension between curriculum breadth and available time, rapid return of assessment and depth of marking, access and finite specialist capacity, or confidentiality and participation. The self-review does not conceal the conflict in a single score.

The review identifies which expectations are mandatory, which harms are foreseeable, what options exist and who has authority to decide. It states the trade-off and any group bearing the cost. Where no lawful and educationally acceptable option is available within institutional resources, the evidence supports referral rather than a fictitious internal solution.

Table 2. Examples of reviewable institutional questions
Broad concernReviewable questionUnit and periodPrincipal evidenceDecision enabled
Learner accessCan learners with the identified mobility requirements reach every compulsory teaching space this term?Learner route and scheduled classTimetable, route inspection, learner evidence and accommodation recordImmediate access remedy and timetable revision
Assessment feedbackDid sampled first-semester work receive usable feedback before the next related task?Assessment item by programmeSubmission and return dates, marked work, task sequence and learner evidenceRevise timing, workload or feedback design
Curriculum coverageWere required practical components delivered to the graduating cohort?Required component and cohortProgramme specification, timetable, attendance and completion evidenceRestore omitted provision and examine award implications
Student supportDid referred learners obtain the defined support within the published response period?Referral episodeReferral, appointment, service and outcome recordAdjust capacity, triage or referral route
Teacher developmentWas the selected teaching practice used after supported development?Teacher practice in relevant classesDevelopment record, material conditions, observation and follow-upContinue, adapt or cease the support model
GovernanceWere material learner complaints considered by the authorised body and resolved with reasons?Complaint caseCase record, authority, elapsed time and decisionCorrect process and address recurring causes

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC question-design examples. Examples are illustrative and do not establish sector requirements.

12

Reviewability test

Before evidence collection, the institution asks whether the question has a legitimate criterion, identifiable population, feasible evidence, competent decision owner and realistic review period. If any element is absent, the question is revised or referred.

A question is not rejected merely because evidence is imperfect. The institution distinguishes what can be established now, what requires new evidence and what remains beyond reasonable inference. The reviewability test prevents extensive collection for a decision that no one is authorised or prepared to make.

Part III

Building the claim-to-evidence map

13

The proposition precedes the source

Evidence selection begins with the proposition to be tested. Starting from available records encourages the institution to report what is easy to count rather than what the question requires. A claim-to-evidence map lists each proposition, the most direct source, corroborating source, expected limitation and decision consequence.

For an access question, the proposition may be that every compulsory class was reachable by the learner group. Room inventories are insufficient. The map requires individual timetables, route and facility conditions, actual changes, and safe learner confirmation. Each source contributes a different part of the proposition.

14

Document evidence

Policies, programme specifications, course outlines, minutes and plans establish intended arrangements and authorised decisions. They have high relevance to what ought to occur, but limited relevance to what occurred.

Version and effective date matter. A policy approved after the reviewed event cannot be used as the criterion unless it restates an earlier applicable requirement. Undated documents are treated cautiously. Minutes record the matters and decisions reported to a meeting; silence in minutes does not prove that a problem did not exist.

15

Administrative records

Registers, timetables, staffing records, case files, financial transactions and assessment records provide systematic evidence where definitions and custody are stable. They can support full-population analysis and reduce recall error.

Administrative records are created for operational purposes. Their categories may not match the review question; missing entries may mean no event, no record or no requirement to record. The institution documents how duplicates, late entries, withdrawals, transfers and changes of status are treated.

16

Observation and work samples

Observation provides direct evidence of a condition or practice during the observed period. It is well suited to accessibility, facility function, teaching process and use of resources. It remains sensitive to timing, observer judgement and changed behaviour under observation.

Work samples can establish the task set, learner response, feedback and progression of work. They do not represent all learning unless selected through an appropriate rule. The sample includes ordinary and incomplete work, not only exemplary products. Names and sensitive information are protected.

17

Survey and interview evidence

Surveys estimate reported experience or opinion among respondents. Their validity depends on question wording, eligible population, response and mode. Satisfaction can identify an issue but is not a direct measure of curriculum, safety or learning.

Interviews permit explanation and exploration of cause. They are not converted into frequencies unless the selection and method support that use. Power relations affect candour: learners may not criticise teaching in the presence of teachers, and staff may not identify governance concerns to a direct manager. Safe arrangements and accurate reporting of non-participation are therefore part of the evidence design.

18

Assessment and outcome evidence

Assessment evidence requires a clear construct, common administration where comparison is intended, scoring quality and a defined population. Examination pass rates can change because entry rules, absences, assessment difficulty or marking change. Progression and completion are also affected by definitions and the treatment of transfers and repeaters.

International comparative evidence can provide system context but does not assign causal responsibility to one institution. The population sampled, age or grade basis, assessment domain and uncertainty remain essential to interpretation.[REF-09] [REF-10]

19

Financial and resource evidence

Budget allocation, commitment, expenditure and asset receipt are separate stages. Financial evidence can establish lawful and actual use of funds; it cannot alone establish that the educational service operated.

Self-review links expenditure to the specified action, physical or service receipt and intended population. Donated resources, staff time, maintenance and household costs are identified where material. A low recorded expenditure may reflect transferred burden rather than efficiency.

20

Evidence independence

Corroboration is strongest where sources have different error structures. A timetable and teacher interview may both derive from the same planned schedule; an observation adds evidence of delivery. A committee report summarising a student survey is not independent of the survey.

Institutional evidence does not become invalid because staff collected it. Independence is one quality among several. Directness, coverage, timing and transparency may make an internal record stronger than a brief outside observation. Consequence determines whether an additional independent check is necessary.

Table 3. Claim-to-evidence mapping
PropositionMost direct evidenceUseful corroborationPrincipal limitationPermitted conclusion
Required teaching occurredDated delivery record or observationTimetable, learner work and attendanceRecording quality and observation coverageDelivery for the observed or recorded population
Learners could access a serviceAccess episode and user evidenceFacility check, timetable and referral recordNon-response and disclosure constraintsAccess status for the defined users and period
Feedback was timelySubmission, return and next-task datesMarked work and learner reportLate entry or inconsistent date recordingTimeliness under the adopted definition
Governance considered a concernAgenda, paper, minute and decisionCase follow-up and participant confirmationMinutes may omit discussion detailConsideration and recorded decision, not effectiveness
A support action changed practiceComparable practice evidence before and afterImplementation, context and participant accountConcurrent change and observer effectChange following action; causation only if design supports it
Outcomes improvedComparable outcome measure and populationParticipation, assessment and context recordsIntake, missingness, test change and regressionDescriptive change with stated limitations

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC claim-to-evidence map.

21

The evidence sufficiency judgement

Evidence is sufficient when it is adequate for the specified claim and decision, not when every possible source has been collected. The judgement considers directness, coverage, consistency, timing, independence, missingness and consequence of error.

The institution records one of four positions: sufficient to decide; sufficient for a provisional decision with follow-up; insufficient and requiring specified evidence; or incapable of resolution through the proposed method. This prevents both unlimited collection and unsupported closure.

Part IV

Participation, candour and protection

22

Participation as evidence and governance

Participation serves two purposes. It provides evidence about educational experience that administrative records cannot show, and it gives affected people a legitimate place in institutional improvement. These purposes overlap but remain distinct. A participant can describe experience without holding decision authority; a governing body can hold authority without possessing direct experience.

The European school recommendation identifies teachers, pupils, management, parents and experts as relevant participants in external and self-evaluation. The right of children to express views in matters affecting them adds a rights basis, with weight appropriate to age and maturity.[REF-03] [REF-06]

23

Mapping affected groups

Participation begins with the population affected by the review question. Broad invitations commonly reach confident and available respondents while missing those facing language, disability, distance, cost, schedule or power barriers.

The institution maps groups, identifies the consequence of non-participation and chooses feasible routes. It reports invitations, participation and known exclusions. It does not claim representativeness merely because many responses were received.

24

Learner participation

Learners can provide direct evidence about access, teaching experience, feedback, workload, support and safety. Questions are limited to matters they can reasonably observe. Learners are not asked to determine professional competence, legal compliance or the cause of an outcome.

Participation arrangements protect against retaliation and unnecessary identification. Group discussions are not suitable for every concern. Where a disclosure indicates harm, the institution follows the applicable protection route rather than retaining it solely as research material.

25

Staff participation and professional conditions

Teachers and other staff hold essential evidence about curriculum, workload, resources, implementation and learner needs. Self-review can support professional learning when staff can identify problems without every statement becoming an individual performance allegation.

This protection is not immunity for misconduct or neglect. It is a separation of purposes. Evidence gathered for institutional improvement enters an individual process only under lawful authority, fair procedure and verified facts. The 1966 Recommendation concerning the Status of Teachers provides an important contemporaneous basis for consultation, professional responsibility and fair treatment.[REF-11]

26

Families, communities and external expertise

Families and communities can contribute evidence on access, cost, language, attendance and institutional communication. Their views are not presumed uniform. Participation arrangements recognise that some adults may not speak for all learners and that family preference cannot justify discrimination or an educationally inadequate condition.

External expertise is used where the institution lacks technical competence or where an outside challenge strengthens credibility. The expert's question, evidence access, method and interest are recorded. Reputation alone does not replace a transparent basis for judgement.

27

Disagreement

Disagreement can concern fact, criterion, interpretation or action. The review separates these forms. A disputed attendance count is resolved through source reconciliation; disagreement about whether the count is acceptable requires a criterion; disagreement about cause requires further analysis; disagreement about action requires a decision under the applicable authority.

Minority evidence is retained where it identifies a group-specific condition or plausible risk. Consensus language is not used to erase a material dissent. The final judgement states the principal alternative explanation and why it was accepted, rejected or left unresolved.

28

Confidentiality and disclosure

Personal information is collected only where necessary for the question and is accessed according to role. Review records separate identifiers from analytical data where feasible. Publication uses aggregation, suppression and careful narrative description.

Confidentiality has limits where immediate protection or law requires action. Participants receive an accurate explanation of those limits. An institution does not promise secrecy it cannot maintain, and it does not publish detail merely to demonstrate transparency.

Table 4. Participation design and safeguards
Participant groupDistinct evidence contributionCommon barrierSafeguardInterpretation limit
LearnersDirect experience of access, teaching and supportAuthority, age, fear and scheduleSafe route, clear purpose and protection responseExperience does not alone establish cause or compliance
TeachersCurriculum, practice, workload and implementationEmployment consequence and timePurpose separation, representative selection and protected discussionProfessional account requires corroboration for high-stakes claims
Other staffAdministration, facilities and service operationRole hierarchy and fragmented recordsRole-appropriate questions and confidential routePartial view of educational process
FamiliesAccess, cost, attendance and communicationLanguage, distance, literacy and representationMultiple modes and accessible informationRespondents may not represent all families or learners
LeadershipAuthority, resource and decision evidenceReputational interestDocumentary basis and contrary-evidence testExplanation is not independent verification
External expertTechnical judgement and methodological challengeLimited local context and possible interestDeclared role, evidence basis and institutional responseBrief review cannot represent all institutional activity

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC participation and protection framework.

29

Burden and proportionality

Participation has a cost in learner time, teacher time, administration and possible exposure. Repeated consultations on the same unresolved issue can reduce trust. The institution reuses valid evidence where definitions and consent permit, coordinates requests and explains what decision followed.

The depth of participation is proportionate to consequence and diversity. A minor timetable adjustment may require evidence from the affected class. A change to programme structure requires broader participation and an account of groups not reached. Proportionality reduces unnecessary burden without allowing leaders to consult only convenient voices.

Part V

Bias, missing evidence and the integrity of internal enquiry

30

The institutional evidence problem

An institution possesses detailed knowledge, continuing access and the ability to connect records across time. These are substantial advantages. The same institution also has interests in how findings affect reputation, finance, staff and learner confidence. A credible method retains the advantages while controlling the risk that evidence is selected or interpreted to protect the institution.

Integrity does not depend on a claim of neutrality. It depends on procedures that make the judgement open to correction: fixed questions, documented selection, preservation of exceptions, declared interests, separation of roles where feasible, and a recorded response to challenge.

31

Favourable case selection

Exemplary lessons, successful graduates and completed actions can illustrate practice but cannot estimate its prevalence. The review population includes ordinary, weak, missing and incomplete cases according to a rule established before results are known.

Purposive selection remains legitimate for some enquiries. A review of a rare access failure properly selects the affected cases. Its conclusion concerns those cases and the mechanism, not the institution-wide rate. The selection method and intended inference are stated together.

32

Convenient periods

Evidence from a single week can be affected by examinations, holidays, induction, staff absence or preparation for review. A period is selected because it represents the proposition or because its exceptional status is the subject of enquiry. It is not selected because records are most complete or results most favourable.

When the timing cannot be changed, the institution limits the conclusion. An observation during one intensive teaching week can establish practice during that week; it cannot establish routine use across the year.

33

Missingness

Missing evidence is classified as no event, no record, item not applicable, refusal, loss, inaccessible record or unknown. These states are not combined. A blank assessment-return date may mean that work was not returned, that the date was not recorded or that the learner did not submit. Each interpretation leads to a different finding.

The review reports the number eligible, the number observed and the reason for material loss. A favourable rate among complete records is not applied to missing cases without a stated and defensible assumption. Where the result is sensitive, plausible bounds are presented.

34

Confirmation bias

Review teams can favour evidence that supports an initial concern or a preferred action. The protocol requires at least one credible alternative explanation and identifies the evidence that would distinguish it. A falling completion rate, for example, may reflect weaker teaching, changing entry, altered award rules, more complete recording of withdrawals or an external shock.

The purpose is not to list every imaginable cause. It is to test explanations that are plausible, material and associated with different decisions. Where the evidence cannot distinguish them, the finding remains unresolved and the next enquiry is specified.

35

Deference and hierarchy

Leadership judgement has evidentiary value but can dominate an internal process. Junior staff, learners or distributed sites may withhold contrary information. Meeting agreement is therefore not treated as confirmation that no contrary evidence exists.

The institution uses routes appropriate to its context: independent facilitation, anonymous written evidence, separate discussions, rotation of review leadership or direct access to records. The safeguards are evaluated by whether material dissent became visible, not by whether participants reported satisfaction with the meeting.

36

Indicator displacement

An institution may improve the indicator without improving the educational condition. A response-time target can encourage premature closure; an attendance target can change recording; a pass-rate target can alter assessment or entry. The review identifies how the measure could be influenced and uses a balancing measure where risk is material.

For feedback timeliness, a balancing measure may examine usefulness and learner opportunity to apply feedback. For complaint closure, it may examine recurrence and unresolved remedy. For completion, it may examine progression rules, withdrawals and learning evidence.

37

Independence and role separation

The person responsible for an activity can supply essential evidence and explanation but need not make the sole final judgement. Where capacity permits, the review separates evidence custody, analysis, judgement and approval. Where it does not, another informed person examines selection and reasoning.

Role separation is proportional. A low-risk departmental enquiry may use peer challenge. A finding affecting learner rights, award integrity, safety or substantial public resources requires stronger independence and competent authority.

Table 5. Principal threats to internal review integrity
ThreatObservable signControlEvidence that control operatedResidual limitation
Favourable selectionOnly successful or complete cases appearFixed population and selection ruleCase list with inclusions, exclusions and missing casesFrame itself may be incomplete
Convenient periodObservation coincides with exceptional activityPeriod rationale and repeated evidence where neededDated schedule and comparison periodsSeasonal effects may remain
Confirmation biasNo alternative explanation consideredContrary-evidence questionRecorded alternatives and discriminating evidenceSome causes remain unobserved
Hierarchical influenceUniform views despite known role differencesProtected and separate participation routesParticipation and dissent recordFear may still reduce candour
Indicator displacementTarget improves while service concern persistsBalancing measure and source checkService and counter-indicator reported togetherNew behavioural responses can emerge
Reviewer interestReviewer owns the activity or preferred interventionDeclared interest and independent challengeChallenge and response recordSmall institutions may have limited separation

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC internal enquiry integrity register.

38

Correction of error

The institution provides a route for factual correction before a finding is finalised. Corrections identify the original record, new evidence, authorising person and effect on the judgement. An inconvenient correction is not rejected because the report is advanced.

After publication, material errors are corrected openly and promptly. The correction states what changed and whether decisions require reconsideration. The general principles of professional method, accountability for sources, correction and confidentiality in official statistics provide a relevant integrity model, while institutional self-review remains distinct from official statistics.[REF-05]

39

Record custody

Evidence records have an identified owner, date, version and retention period. Working copies are distinguished from authoritative records. Changes preserve a trace sufficient to explain the final figure or judgement.

Custody protects both integrity and confidentiality. Broad access is not synonymous with transparency. Reviewers receive the least personal detail needed, while the method and aggregate basis remain available for scrutiny.

Part VI

Reaching and recording an evaluative judgement

40

Finding before rating

The review first states what the evidence shows. It then compares that finding with the criterion. A rating, where used, follows these steps and does not replace them.

The finding identifies population, period, measure, variation, missingness and material exception. “Feedback is effective” is not a finding. “In the sample of 120 first-semester assignments, 83 were returned before the next related task; 19 were returned later, 8 had no recorded return date and 10 were not submitted” is a finding capable of examination.

41

The judgement statement

A judgement statement contains four elements: the criterion; the evidenced condition; the degree and distribution of difference; and the consequence for action. It separates verified fact from interpretation.

Where the criterion is met for most but not all learners, the judgement does not collapse the result into a majority pass. It identifies whether the exceptions are random, concentrated or connected with a protected or disadvantaged group. A small number can be material where the right or safety consequence is serious.

42

Degrees of evidentiary confidence

Confidence concerns the evidence supporting the judgement, not the favourability of the condition. A strong adverse finding can have high confidence; a favourable finding can have low confidence because coverage is poor.

The review uses plain positions: well supported; supported with material limitation; provisional; or unresolved. These positions are accompanied by the reason. A numerical confidence score is avoided because its precision and weighting would be difficult to justify across dissimilar questions.

43

Materiality

Materiality considers educational consequence, legal or protective significance, number and identity of learners, duration, recurrence and effect on awards or progression. It is not determined only by institutional size.

One incorrect award decision can be material to the learner and to award integrity. A minor delay affecting many records may be operationally significant without undermining educational validity. The judgement explains which dimension makes the finding material.

44

Distribution

Institution-wide averages are disaggregated where the question and lawful records permit. Relevant dimensions may include programme, level, site, mode, sex, disability, language, age or other national categories. The institution does not collect sensitive characteristics without purpose and safeguards.

Differences are interpreted with counts and denominators. A rate based on five cases is not compared rhetorically with a rate based on five hundred. Statistical uncertainty is considered where sampling is used; verified individual harm remains actionable even when a group rate is unstable.

45

Context without excuse

Context explains conditions that affect the interpretation or feasibility of action. It does not erase the criterion. Rapid enrolment growth, remote location or constrained finance may explain a staffing or facility gap; the gap and affected learners remain visible.

The judgement distinguishes institutional control, shared control and external control. It identifies action within authority and the evidence required for referral. This avoids both unfair attribution and institutional avoidance.

46

Cause and contribution

Self-review often needs a causal explanation to choose action. It distinguishes immediate mechanism from wider contributing conditions. A delay in feedback may arise immediately from the sequence and volume of assessment; staff capacity, course design and approval rules may contribute.

Cause is not assigned from sequence alone. The review tests whether the proposed mechanism was present, whether it preceded the condition, whether the pattern differs where the mechanism differs and whether plausible alternatives remain. The resulting wording reflects the strength of that test.

47

Rating systems

An institution may use categories to summarise findings, but the categories have definitions, evidence thresholds and exception rules. A colour or number cannot be the only recorded result. It is linked to the full judgement and action.

Aggregation across domains is approached cautiously. Safety, access, curriculum, teaching and outcomes are not fully substitutable. A high score in one cannot compensate for a serious failure in another. Where an overall category is required, non-compensable conditions and uncertainty are explicit.

Table 6. Judgement record
ElementRequired contentQuestion for challengeInadequate form
CriterionSource, status and operational meaningWas it applicable during the period?Unexplained use of “adequate” or “effective”
FindingPopulation, period, evidence, counts and variationAre missing and contrary cases included?General impression or selected example
ComparisonNature and extent of differenceWas the threshold fixed before results?Post-hoc standard or unsupported benchmark
DistributionGroups or units materially affectedDoes the average conceal an exception?Institution-wide rate only
ConfidenceSufficiency and principal limitationWhat evidence could change the judgement?Confidence implied by assertive wording
MaterialityConsequence, scale, duration and recurrenceWhy does the finding require this response?Materiality equated with large numbers only
ResponsibilityAuthority able to act or referIs responsibility assigned at the correct level?School or department blamed for system control

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC evaluative judgement record.

48

Review panel deliberation

Panel members examine the evidence independently before collective discussion where feasible. This reduces early conformity. The panel considers the proposed judgement, strongest contrary evidence, material missingness and effect of changing a key assumption.

The final record gives reasons. Where members cannot agree on a material question, the dissent and its evidence are preserved. A majority decision can authorise action, but it does not transform unresolved fact into consensus.

49

Institutional response

The unit whose work was reviewed receives the draft factual record and can provide correction, contextual evidence and an alternative interpretation. The response has a fixed period and cannot remove a supported adverse finding by negotiation.

The final report records material changes arising from response. It distinguishes correction of fact from acceptance of an explanation or proposed action. This makes the review fair without making the subject of review the sole judge of the evidence.

Part VII

Designing an improvement response

50

From finding to decision

An evaluative finding creates a decision requirement, not an automatic intervention. The institution first determines whether the condition requires immediate protection, correction, further enquiry, broader system action or no change. The decision is proportionate to materiality and confidence.

A weakly supported low-risk concern may justify a small test. A high-confidence safety or award-integrity finding requires immediate competent action. Uncertainty is not treated uniformly.

51

The improvement proposition

The institution states why the selected action is expected to change the finding. The proposition links action, mechanism, operating condition and intended result. Assumptions and external dependencies appear beside it.

For delayed feedback, reducing the number of simultaneous assessment deadlines may redistribute marking time and permit return before the next task. The proposition depends on actual workload, assessment sequence, staff capacity and no compensating increase elsewhere. A generic workshop does not address the mechanism unless evidence identifies knowledge or practice as the constraint.

52

Options and alternatives

The review considers more than one feasible response where the decision is material. Options are compared on expected educational effect, time, cost, reach, recurrent burden, institutional capability, risk and reversibility.

The lowest immediate cost is not always the most efficient response. A temporary arrangement can protect current learners but create recurring cost; a durable redesign can require more initial effort. The decision records why the selected option is preferable and why plausible alternatives were not selected.

53

Responsibility and authority

Every action has one accountable owner and named contributing roles. Accountability is assigned to the role with authority to make the required decision, not to the person most available to complete the form.

Where the institution depends on a ministry, governing body, employer, landlord or public service, the plan separates internal action, formal request and follow-up. It does not record the external decision as an institutional task already completed.

54

Resources and opportunity cost

The plan records finance, staff time, facilities, technical support, information and authority. It also identifies which existing activity will be reduced or delayed. An action is not costless because it uses current staff.

Teacher and learner time receive particular protection. Meetings, surveys and development activity can displace instruction or study. The plan schedules them so that the improvement process does not create a larger educational loss.

55

Implementation stages

The plan distinguishes authorised, initiated, completed and operating states. A revised policy is authorised when validly approved, initiated when relevant people and systems begin to use it, completed when specified transition tasks finish, and operating when the intended service functions.

Stage definitions prevent a meeting, purchase or publication from being reported as the final result. They also locate delay. An action may be approved on time but fail at staffing, communication, supply or actual use.

56

Reach and distribution

Reach identifies which intended units or people received the action. An institution-wide announcement is not evidence that learners requiring a service received it. Training offered is not training completed; training completed is not supported practice.

The plan specifies the intended population and monitors differential reach. A convenient pilot can exclude the programme where the original concern was greatest. If that is necessary for safety or feasibility, the limitation and later stage are explicit.

57

Adaptation

Implementation commonly requires adjustment. The record distinguishes minor operational adaptation from a change to the core mechanism or population. Minor adaptation can preserve the proposition; material change requires a revised proposition and interpretation.

The institution does not report fidelity to a plan that it did not implement. Nor does it treat every adjustment as failure. The question is whether the critical components operated and whether the changed action still addresses the finding.

58

Risk and safeguards

The plan considers safety, rights, privacy, academic or curricular integrity, staff conditions, financial control and burden transferred to learners or families. A favourable intended result cannot justify an unlawful or harmful means.

Risk controls are specific. “Monitor closely” is not a safeguard unless responsibility, evidence, timing and response threshold are defined. Temporary measures have expiry and review conditions so that they do not become an unexamined permanent standard.

Table 7. Improvement action charter
FieldRequired statementControl question
Finding addressedExact judgement and affected populationDoes the action respond to this finding?
Improvement propositionAction, mechanism, operating change and intended resultIs the causal link plausible and testable?
AuthorityAccountable owner and contributing rolesCan the owner authorise the required work?
ResourcesFinance, time, capability, facilities and external dependenciesWhat existing activity or resource is displaced?
StagesAuthorised, initiated, completed and operating definitionsCan administrative completion be distinguished from service?
ReachIntended units and evidence of receipt or participationCan exclusion or uneven implementation be detected?
SafeguardsMaterial risk, preventive control and responseDoes action protect rights, safety and integrity?
ReviewMeasure, baseline, follow-up period and decision ruleWhat evidence will support continuation, adaptation or cessation?

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC improvement action charter.

59

Decision rules

Before implementation, the institution states what result would justify continuation, adaptation, wider use or cessation. Rules include both service effect and safeguards. A result is not accepted if it is achieved through exclusion, weakened assessment or an unrecorded transfer of burden.

Decision rules can include a quantitative threshold where justified, but they also address distribution and critical exceptions. A target of 90 per cent timely feedback does not permit the remaining 10 per cent to be consistently concentrated in one programme or learner group.

60

The improvement register

The register links every material finding to an action, referral, accepted risk or reasoned decision that no action is required. It records status and elapsed time. Findings do not disappear when a report cycle ends.

Repeated actions for the same unresolved cause trigger escalation. The register distinguishes recurrence from a new case and permits leaders to see where local plans repeatedly fail because an institutional or system condition remains unchanged.

Part VIII

Verifying improvement

61

Return to the original finding

Follow-up begins with the original criterion, population and evidence. It does not replace them with the activity most readily reported. If the finding concerned learner access, follow-up measures access; expenditure and construction are supporting evidence.

Where the original evidence was weak, the institution improves the measurement but records the resulting comparability limit. A more complete follow-up can reveal previously unrecorded cases and should not be presented automatically as deterioration.

62

Implementation verification

Verification confirms whether critical action components occurred, for whom, when and under what adaptation. It uses the action charter rather than a general progress narrative.

Completion evidence is proportionate to consequence. A low-risk scheduling change may be confirmed through timetable and sampled delivery. Structural safety, assessment validity or substantial expenditure requires competent evidence and stronger separation of roles.

63

Operating result

An operating result is the first verified educational or institutional condition expected after implementation: a service accessible, a required component delivered, feedback returned in time, a case decided through the proper route or a teaching practice used under the specified conditions.

The result is defined closely enough to the action to support a management decision. It is distinct from longer outcomes. Restored access may later contribute to participation and learning; those effects require additional evidence and time.

64

Baseline comparability

Baseline and follow-up use the same definition, population and period, or the difference is adjusted and disclosed. Programme restructuring, changed assessment, new record systems, altered enrolment and staff turnover can create an apparent change.

Where a common subset is used, the review states that it represents continuing cases rather than the whole current institution. Where no valid comparison exists, the report presents current operating status and implementation, not a fabricated trend.

65

Outcome and contribution

The institution examines whether the operating change was followed by the intended educational outcome. A before-and-after change is descriptive. Stronger attribution requires evidence that the action preceded the change, the mechanism operated, competing explanations were examined and an appropriate comparison exists where feasible.

Urgent or universal service action may not permit an untreated comparison. This does not prevent action. It limits the causal claim. The institution can report that the condition improved following implementation and explain why contribution is plausible without claiming exclusive cause.

66

Unintended effects

Follow-up asks what else changed. Faster feedback may reduce depth; expanded access may increase staff workload beyond safe or sustainable levels; a new complaint route may expose confidential information; a revised timetable may disadvantage part-time learners.

Unintended effect evidence uses groups and measures relevant to the risk. Absence of complaints is not proof of absence where the action itself affects the ability to complain.

67

Sustainability

An operating result can depend on temporary finance, exceptional staff effort or a single individual. Sustainability review identifies recurrent cost, role ownership, maintenance, staff capability and the conditions needed after the initial period.

The review does not require certainty about indefinite continuation. It distinguishes stable operation, time-limited operation with an authorised continuation plan, and an at-risk result without secured means.

68

Decisions after verification

Continuation is justified where the action operated, the intended result is evident and risks remain acceptable. Adaptation is justified where the mechanism is plausible but reach, fidelity or context limited the result. Expansion requires evidence that capacity and relevant conditions exist in additional settings. Cessation is justified where the mechanism fails, harm is material or resources can be better directed.

An inconclusive result does not default to continuation. The institution records the value and burden of further evidence, the consequence of delay and any interim measure.

Table 8. Verification questions and decision consequences
Verification questionEvidenceFavourable conditionLimiting conditionDecision implication
Did critical action occur?Tasks, dates, responsible role and adaptationCritical components completeCore component absentDo not judge intended proposition; correct implementation
Was the intended population reached?Eligible and reached unitsCoverage sufficient and equitableMaterial exclusion or unknown reachAdapt route before wider use
Did the operating condition change?Comparable direct measureCriterion met or material gap reducedNo change or invalid comparisonRe-examine mechanism or measurement
Did longer outcome change?Suitable outcome evidence and contextChange consistent with mechanismTiming, assessment or population confoundedNarrow claim and continue observation if justified
Were there unintended effects?Balancing measures and affected-group evidenceNo material unresolved harmHarm, exclusion or transferred burdenCorrect, pause or cease affected component
Can the result continue?Recurrent resources, capability and maintenanceContinued means reasonably securedExceptional effort or unfunded liabilityTreat as temporary or at risk

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC improvement verification framework.

69

Case closure

A finding closes when the required operating condition is verified, the case is invalid or duplicated with reason, or a competent authority formally accepts a residual condition under a lawful and evidence-based decision. Referral, approval, expenditure or a stated intention does not close the finding.

Closure records the evidence date and any continuing monitoring. If the condition recurs, the institution reopens or links the case so that repeated temporary correction is not mistaken for durable improvement.

70

Institutional learning from multiple reviews

Individual reviews are periodically analysed for recurring causes, delayed decisions, failed actions and uneven service. This cross-case analysis can identify a governance, resource or information problem not visible in one review.

Aggregation does not erase open cases. A system lesson and an individual remedy proceed together. The institution records which changes have been made to review design itself, including removal of unused indicators, strengthened participation or revised decision authority.

Part IX

External perspective and institutional response

71

Purpose of an outside view

An outside view tests whether the institution's reasoning can be understood and whether its evidence supports the judgement. It can identify omitted populations, inconsistent criteria, favourable selection and conclusions that exceed the method. It can also confirm that an institution has reported a serious condition accurately and taken proportionate action.

The 2001 European recommendation describes external evaluation as a source of methodological support and outside perspective for continuous improvement, rather than a process confined to administrative checking. This distinction protects both institutional learning and public accountability.[REF-03]

72

Scope of external challenge

The outside reviewer examines the review question, criterion, population, source selection, missingness, analysis, judgement, improvement response and verification. The task is not necessarily to repeat the whole review. It identifies high-consequence or uncertain links and tests them directly.

The reviewer does not replace the institution's question with an unrelated preference after evidence collection. If a material omitted issue emerges, it is recorded as a separate finding requiring its own criterion and evidence.

73

Evidence access

Access is sufficient to test material claims while respecting confidentiality. The institution provides source registers, selection rules and de-identified records where possible. Personal files are not disclosed merely for convenience.

Where law or protection limits access, the reviewer records what could and could not be examined and how this affects confidence. A confidentiality restriction is not described as full verification.

74

Site visits and interviews

A visit can establish current conditions and allow direct discussion, but it remains a limited period. The institution does not stage exceptional activity as routine evidence. The reviewer samples against the published question and follows material exceptions.

Interviews test understanding, participation and implementation. Agreement among selected interviewees does not establish prevalence. The visit record distinguishes direct observation, participant report and reviewer interpretation.

75

Challenge and reply

External challenge is recorded as a question or proposition with an evidence basis. The institution has an opportunity to correct fact, provide further evidence and explain context. The reviewer then states whether the challenge is resolved, modifies the judgement or remains as a difference.

The process is not a negotiation over the favourability of language. Changes follow evidence or corrected interpretation. Material unresolved disagreement remains visible to the competent decision-maker.

76

Reliance on institutional evidence

External scrutiny can rely on institutional evidence where source custody, selection, definition and limitations are sound. Re-collecting every item can waste resources and weaken institutional responsibility. Reliance is greater where the institution has retained contrary cases, corrected errors and verified earlier actions.

Reliance is reduced where the review population cannot be reconstructed, material records are missing, sources derive from one interested account, or repeated internal findings closed without operating evidence. The response is targeted verification, not automatic rejection of all internal work.

77

Proportionality

The depth of external review follows consequence, novelty, evidentiary weakness and institutional history. A routine low-risk improvement can be tested through records and samples. Safety, learner rights, award integrity, substantial public expenditure or a contested outcome requires more direct and technically competent scrutiny.

Proportionality also limits burden. Institutions do not repeatedly supply the same verified evidence to several processes without a clear purpose. An external request identifies the decision it supports and the period for which existing evidence remains valid.

78

External judgement boundaries

An external reviewer reports within the evidence examined. A brief review of one programme cannot support an institution-wide conclusion without a valid sampling basis. Absence of a discovered problem does not prove that no problem exists.

The final record distinguishes confirmation of the self-review method, confirmation of a particular finding and direct external finding. These have different scopes and are not merged into a general endorsement.

Table 9. External perspective record
Review elementExternal testPossible resultConsequence
CriterionApplicable, legitimate and consistently interpretedConfirmed, clarified or disputedRevise judgement if criterion changes materially
Population and selectionComplete or appropriately sampledAdequate, limited or biasedNarrow conclusion or collect additional evidence
Source integrityTraceable, dated and suitable for claimReliable, qualified or unverifiedAdjust confidence and verification
Contrary evidenceMaterial exceptions and alternatives consideredAdequate or omittedReopen analysis where omission is material
Improvement responseAction addresses finding and authority is correctAligned, partial or unrelatedRevise action and responsibility
Operating resultDirect and comparable evidence supports closureVerified, provisional or absentClose, continue or reopen finding

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC external perspective record.

79

Use of the external report

The external report informs the competent institutional or public authority. It does not absolve that authority from making and recording the decision. Findings are prioritised by educational consequence and evidentiary confidence, not rhetorical strength.

The institution incorporates accepted corrections and actions into its improvement register. Rejected recommendations include reasons. Silence or general acknowledgement is not an institutional response.

Part X

Public reporting and accountability

80

The public purpose

Public reporting enables learners, families, staff, governing bodies and public authorities to understand what the institution examined, what it found and what remains to be done. Its purpose is not institutional promotion or publication of every working record.

The account is truthful, proportionate and connected to remedy. Material risks and unresolved failures are not concealed within an average or extensive descriptive text. Confidentiality is maintained without using privacy as a reason to omit aggregate conditions.

81

Fact, judgement and commitment

The public report distinguishes recorded fact, institutional calculation, evaluative judgement and future commitment. “Ninety of 110 sampled items met the published return period” is a recorded result. “Timeliness is inconsistent across programmes” is a judgement. “The assessment schedule will be revised by September” is a commitment.

These statements occupy separate fields or sentences. A commitment is not written in a form that implies completion, and an interpretation is not presented as directly observed fact.

82

Minimum public account

For each material review, the public account states the question, criterion, population and period; principal evidence; finding and distribution; limitation; decision and responsible authority; and current implementation or verification status.

Technical detail is proportionate. A short method note can support a simple full-population count. A sampled outcome comparison requires the sampling, response, uncertainty and assessment basis. Unsupported precision is not a sign of rigour.

83

Reporting unfavourable evidence

An institution that reports weakness credibly does not thereby establish general failure. It demonstrates that the condition was detected. The relevant accountability questions are whether the finding is accurate, learners are protected, responsibility is assigned and improvement is verified.

Conversely, the absence of adverse findings is not evidence of high quality when the review questions, selection or reporting rules make weakness unlikely to appear. The public report includes material missingness and open findings.

84

Comparative reporting

Institutions differ in programme, population, mission, resources and context. Public comparison requires common definitions and an appropriate basis. Raw outcomes can penalise institutions serving learners with greater prior disadvantage or encourage selection.

Comparison can still be useful when it identifies an enquiry rather than a verdict. The report presents counts, definitions, reference periods and uncertainty. It does not rank institutions on an internally constructed measure that was not designed for comparison.

85

Small numbers and confidentiality

Small groups may experience material exclusion while their public statistics are unstable or identifying. The authority acts on verified need but limits disclosure. Cells are combined or suppressed only where necessary, and totals are designed so that subtraction does not recover protected values.

Narrative examples are also reviewed for identification. Removing a name may be insufficient where programme, event and personal circumstance point to one individual.

86

Correction and version

The public account identifies the review period and date. Material correction states the original error, corrected result and effect on judgement or action. The earlier record is not silently replaced where it has informed a public decision.

Updates distinguish improved status from changed definition. An institution cannot show progress by removing difficult cases from the denominator or changing a threshold without restating the earlier result under the new basis where feasible.

87

Reporting incomplete improvement

Actions are labelled authorised, initiated, completed or operating. Delayed and at-risk actions remain visible with reason and next decision date. A referral to another authority is reported as referral, not remedy.

Where no valid action is available within current resources, the institution states the unresolved condition, interim protection and requested decision. Restraint is preferable to a nominal action that cannot change the service.

Table 10. Public self-review account
Public fieldContentRequired limitation
Review questionProvision, population, period and decisionExclusions stated
CriterionSource and operational meaningLegal or policy status distinguished
EvidenceSources, coverage and selectionMissingness and timing stated
FindingCounts, distribution and material exceptionNo generalisation beyond population
JudgementComparison with criterion and confidenceInterpretation separated from fact
ImprovementDecision, owner, resources and milestoneCommitment not presented as result
VerificationOperating evidence and review dateCausation limited to design
ProtectionDisclosure controls and protected action routeConfidentiality not used to hide aggregate failure

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC public self-review account.

88

Institutional annual synthesis

An annual synthesis groups reviews by educational function, reports open and closed material findings, identifies recurrence and states changes made to the self-review method. It does not reproduce every local record or convert diverse findings into one quality score.

The synthesis includes the areas not reviewed and the basis for selection. This prevents an annual report from implying comprehensive coverage where attention was limited to a small number of favourable topics.

Part XI

Application across education sectors

89

Common disciplines and sector differences

The common disciplines are question, criterion, evidence, participation, judgement, action and verification. Their application differs because authority, learner age, programme structure and educational purpose differ.

The report does not assume that a school self-review model can be transferred unchanged to a university, or that an academic programme review can govern a primary classroom. Sector application begins by identifying the competent body and the educational service.

90

School education

School review commonly examines curriculum delivery, instructional time, teaching conditions, learner attendance, inclusion, safety, support and learning evidence. Children require age-appropriate participation and protection. Families and public authorities may hold responsibilities not present in the same form in adult education.

The school is often dependent on district or national systems for staffing, curriculum, finance and facilities. Self-review therefore maps responsibility and keeps external constraints visible. The European recommendation's balanced treatment of self-evaluation, outside perspective and participation provides a relevant regional example, not a universal procedure.[REF-03]

91

Vocational education and training

Vocational provision adds workplace learning, occupational equipment, employer participation and transition to work or further learning. Review questions identify whether practical experience matches the approved programme and whether learners can complete it safely and equitably.

Employment outcomes can inform review but are affected by labour-market conditions, region, occupation and learner characteristics. An institution does not infer teaching quality from placement rates alone. Employer opinion likewise remains one source rather than the criterion for all educational purpose.

92

Higher education

Higher education self-review commonly operates through programme, department and institutional governance. Academic judgement, research, curriculum, assessment, student support and public responsibility interact. The institution must preserve the distinction between academic peer judgement and administrative evidence while requiring each to show its basis.

The Bergen Communiqué calls for systematic internal mechanisms connected to external review and emphasises institutional participation, student involvement and public responsibility. Application remains governed by national arrangements and institutional authority.[REF-04]

93

Multi-site and distributed provision

One institutional policy can operate differently across sites. Self-review samples or covers sites according to risk, size and difference, and does not infer consistent operation from the central office.

Remote sites may have weaker records and higher observation cost. This does not justify their systematic exclusion. The method uses proportionate local verification and reports access limitations.

94

Small institutions

Small institutions face difficulty separating reviewer and activity-owner roles, and public reporting can identify individuals through small numbers. Peer challenge, governing-body review or a competent external person can strengthen independence.

The method remains proportionate. A small institution need not reproduce the committee structure of a large one. It must still maintain traceable evidence, fair participation, reasoned judgement and verified action.

95

Large and complex institutions

Large institutions risk fragmentation and aggregation. Local units may use different definitions, while central reports obscure variation. A controlled core of definitions and review records is therefore necessary, with room for questions specific to programmes.

Central synthesis tests comparability before aggregation. It identifies units not reporting and unresolved cross-institutional causes. Local responsibility does not permit the governing institution to disregard repeated system patterns.

96

Institutions under rapid expansion

Expansion changes learner composition, staffing, facilities, programme load and administrative capacity. Historical rates may not describe the current population. Self-review uses cohort and capacity evidence and records structural breaks.

The Dakar and Millennium commitments give expansion a strong public purpose, but access without adequate educational conditions is insufficient. Review connects enrolment growth with instruction, support, completion and distribution.[REF-01] [REF-12]

97

Institutions in constrained or disrupted settings

Where records, staffing or safe access are limited, the institution prioritises essential service, protection and a small number of decision-relevant measures. It does not claim comprehensive review.

Evidence can be provisional while action remains necessary. The record states uncertainty, interim measure and follow-up. A reduced method is credible when its scope is explicit; it is not credible when missing evidence is converted into a favourable institutional account.

Table 11. Sector application boundaries
SettingCharacteristic review unitDistinct evidence needPrincipal safeguardInvalid transfer
SchoolClass, grade, service or schoolInstructional delivery, child participation and system supportProtection and responsibility mappingAdult participation model applied to children
Vocational provisionProgramme, workplace component and cohortPractical experience, equipment, safety and progressionSeparate education purpose from employer preferenceEmployment outcome treated as sole quality measure
Higher educationCourse, programme, department or institutionAcademic standards, assessment, student experience and governanceEvidence-based peer judgement and student participationAdministrative compliance treated as academic quality
Multi-site institutionSite and common programmeConsistency and local variationInclude remote and weak-record sitesCentral policy treated as proof of local operation
Small institutionWhole institution or functionTraceable evidence with limited role separationProportionate independent challengeLarge bureaucracy treated as evidence of rigour
Expanding institutionCohort, programme capacity and serviceStructural breaks, staffing and learner conditionsDo not trade minimum service for growthHistorical average applied to changed population

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC sector application framework.

98

Cross-sector learning

Institutions can learn from one another about evidence custody, participation, action tracking and verification. Transfer begins with the function of the method, not its label. A useful instrument is adapted to the receiving sector's authority, learners and educational purpose.

Claims that a practice is “good” require evidence of the condition it improved, the setting and resources. Dissemination includes limitations and failed applications, not only a favourable example.

Part XII

Governance, capability and conclusions

99

Governing responsibility

The governing body or equivalent authority approves the purpose, receives material findings, assures that competent decisions are made and monitors unresolved risk. It does not rewrite evidence to protect reputation or intervene in individual academic judgement without authority.

Governance receives a concise account of method, material findings, open actions, overdue cases and limitations. It asks whether the institution can detect and correct weakness, not whether every indicator is favourable.

100

Leadership responsibility

Leadership allocates time, access, analytical capability and protection for candour. It ensures that findings reach the role able to act and that cross-unit problems do not remain within isolated plans.

Leaders disclose their interest where they own the activity reviewed. Their response is evidence and decision input, not the sole conclusion. Retaliation against a good-faith participant undermines the reliability of future reviews and requires direct response.

101

Review leadership and competence

Review leaders require competence in question design, evidence selection, basic statistical interpretation, facilitation, confidentiality and improvement planning. Subject expertise is added according to the question.

Training is demonstrated through the quality of review work, not attendance alone. Institutions examine inter-reviewer agreement, correction patterns, missed risks and whether findings led to verifiable decisions.

102

The self-review schedule

The schedule is risk- and purpose-based. It includes recurring core matters, emerging concerns, follow-up and a balanced coverage of educational functions. It is not limited to areas likely to report success.

Frequency follows the rate at which evidence and risk can change. Safety or access issues require immediate review; curriculum and outcome questions may follow programme cycles; stable low-risk processes may be sampled less frequently.

103

Integration with ordinary institutional work

Self-review uses valid records already needed for teaching, support and governance. It adds evidence only where the decision requires it. This reduces parallel reporting and improves the incentive to maintain operational records.

Integration does not mean that administrative records determine every question. Learner experience, observation and work samples may remain necessary. The institution identifies where ordinary systems do not observe the educational condition.

104

Resources and independence

Review requires protected staff and learner time, access to records, analytical support and authority to report findings. An unfunded expectation can produce superficial reports and divert teaching time.

Independence is protected through terms of reference, access, recorded challenge and direct reporting of material concerns. It is not established merely by placing review in a separate office if leadership can suppress evidence without record.

105

Monitoring the self-review system

The institution examines the performance of self-review itself: questions completed; population coverage; corrections; material findings; time to decision; actions operating; recurrence; participant burden; and unresolved external dependencies.

A high number of reports is not necessarily favourable. A low number of weaknesses can reflect effective provision or weak detection. Interpretation combines quantity with evidence quality, scope and subsequent results.

106

Failure modes requiring intervention

Intervention is required where self-review repeatedly omits the same disadvantaged group, closes findings without operating evidence, changes criteria after results, suppresses material disagreement, or produces plans unrelated to findings.

Other warning signs include extensive description with no evaluative judgement, identical findings across dissimilar units, unexplained absence of missing data and a high rate of administrative completion with recurring service failure.

Table 12. Institutional capability for credible self-review
CapabilityEvidence of operationWarning signGoverning response
Question disciplineBounded questions linked to decisionsBroad themes and undirected data collectionRequire reviewability test
Evidence integrityTraceable sources, selection and missingnessExemplary evidence and unexplained omissionsReopen sample and analysis
ParticipationAffected groups reached through safe routesConvenient voices onlyRevise participation design and protection
JudgementCriteria, finding, distribution and confidenceRating without reasoningRequire full judgement record
ImprovementAction linked to mechanism and authorityGeneric plan or training responseReappraise options and ownership
VerificationOperating result and follow-up decisionClosure at approval or expenditureReopen finding until service evidence exists
Public accountabilityMaterial findings and limitations visiblePromotional narrative or silent correctionCorrect report and strengthen governance

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC institutional self-review capability framework.

107

Conditions for confidence

Confidence in institutional self-review accumulates when the institution identifies material weaknesses, preserves contrary evidence, corrects error, acts within authority, refers what it cannot control and verifies results. It is reduced by claims of comprehensiveness unsupported by coverage or by favourable conclusions unsupported by direct evidence.

No single review proves sustained capability. Consistency across questions and time matters. Evidence of learning includes changed instruments, improved definitions, shorter unresolved delays and abandonment of actions shown not to work.

108

Policy considerations

Public authorities can strengthen self-review by clarifying purpose, reducing duplicate reporting, protecting participation, providing methodological support and ensuring that institutions have a response route for conditions beyond their authority. They can also require public reporting proportionate to risk and safeguard personal information.

Authorities should avoid incentives that equate the absence of weakness with success. A mature system expects institutions to find and correct problems. External scrutiny is directed to the quality of that process and the material educational result.

109

Concluding judgement

Institutional self-review is credible when it produces a bounded judgement that another informed reader can trace, an improvement decision connected to the finding and later evidence that the intended condition changed. Its authority comes from method, candour and action, not from institutional assertion.

The institution remains the principal actor because educational improvement depends on people who organise and deliver the work. Public responsibility requires that ownership be accompanied by participation, outside challenge and visible treatment of material risk. The balance is not achieved through a uniform form. It is achieved through clear questions, legitimate criteria, verifiable evidence and decisions that remain open until the educational service operates.

110

Immediate institutional priorities

An institution establishing or revising self-review first identifies material open questions and applicable criteria. It creates a claim-to-evidence map, protects participation, records judgement and connects each material finding to an accountable decision.

It then tests a limited number of reviews through follow-up before expanding the schedule. The test of readiness is not whether templates are complete. It is whether the institution can detect an unfavourable condition, report it accurately, take proportionate action and show whether the condition improved.

The first cycle should also establish a dated baseline for unresolved findings, assign authority for decisions that cross departmental boundaries and identify which evidence can be published without exposing learners or staff. These arrangements make later comparison possible and prevent the review calendar from displacing urgent correction. Where evidence is incomplete, the institution records the limitation, the means of obtaining stronger evidence and the date on which the judgement will be reconsidered.

Part XIII

Evidence and implementation instrument A: Review question and criterion protocol

111

Initiating concern

The initiating concern is recorded in the words in which it first reached the institution, together with source and date. It may arise from routine evidence, learner or staff report, a governing question, external observation, changed provision or follow-up to an earlier finding. Recording the original concern permits later examination of whether the review merely confirmed its starting assumption.

The initiating record does not become a finding. Where immediate safety, rights or award integrity is at risk, protection and competent notification proceed before the full review. The need to act does not remove the need to establish facts; it changes their order.

112

Neutral reformulation

The review leader restates the concern as a neutral, answerable question. The question does not assume failure, cause or preferred remedy. It identifies the service, population, period and decision.

For example, “Poor teaching is causing students to leave the evening programme” contains two untested propositions. The neutral review asks: “How do continuation and withdrawal differ between evening and daytime entrants over the last three cohorts, what reasons are recorded or reported, and which institutionally controlled conditions are associated with the difference?” The first stage is descriptive; causal examination follows only where evidence permits.

113

Population and unit

The protocol defines who or what is eligible. Learners may be counted by person, enrolment, course registration or assessment attempt. Staff may be counted by person, post or teaching assignment. Facilities may be counted by room or scheduled use. The selected unit follows the proposition.

Entry, exit, transfer, deferral and repeat status are specified. A review of first-year continuation cannot silently exclude learners who withdrew before the census date if early withdrawal is part of the concern. Duplicate exposure is identified where one learner contributes several course records.

114

Period

The period covers the operation relevant to the question. It records calendar dates, academic term or cohort as applicable. Where a seasonal or cyclical effect is plausible, the design includes the relevant cycle or limits its conclusion.

The evidence cut-off for the review is stated. Events after that point may inform follow-up but do not alter the original finding without a clearly dated update.

115

Criterion register

Each criterion entry records the exact provision or intended condition, source, adoption or effective date, authority, population, operational interpretation and consequence of non-fulfilment. The entry distinguishes binding, authorised and developmental criteria.

Where a source uses qualitative language, the operational meaning is approved before outcome analysis. The institution retains the source text and identifies any local interpretation. It does not present the local interpretation as if it were the wording of the external instrument.

Table A.1. Criterion register
FieldEntry requirementReview control
Criterion identifierStable local referenceOne meaning throughout review and follow-up
Source and authorityLaw, curriculum, programme, published commitment or adopted objectiveStatus not overstated
Effective periodDate from which criterion appliedNo retrospective application without basis
PopulationLearners, provision, staff or service coveredExclusions explicit
Operational meaningObservable condition, unit and timeFixed before results are known
Exception ruleLawful or educationally justified variationException not used to conceal failure
ConsequenceAction required if condition differsProportionate to status and harm

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC criterion register.

116

Review exclusions

Exclusions identify matters deliberately outside scope and why. An assessment review may exclude the validity of the curriculum if that requires different expertise, but it cannot imply that curriculum validity was confirmed. Exclusions are examined for whether they remove the group or condition most likely to show weakness.

An exclusion caused by lack of access is reported as a limitation rather than a design choice. The review records who can authorise access and whether further enquiry is required.

117

Decision owner

The protocol names the role that will receive the judgement and the decisions within its authority. It also identifies required referral. A review without a decision owner can generate evidence without remedy.

Where the owner is responsible for the activity reviewed, an additional challenge or approval route is assigned according to consequence. The role arrangement is disclosed in the review record.

118

Evidence plan

For each proposition, the plan identifies direct source, supporting source, selection, collection period, responsible person, confidentiality requirement and expected limitation. It includes evidence capable of contradicting the initiating concern.

The plan is revised only for a recorded reason. If an unavailable source changes the inferential strength, the judgement is narrowed. The institution does not substitute a convenient survey for direct evidence and retain the original claim.

119

Analytical plan

The plan states the counts, rates, comparisons or qualitative analysis to be used. Denominators and missing categories are specified. Where judgement coding is required, the protocol defines categories and comparison of reviewers.

Exploratory analysis can identify an unanticipated pattern. It is labelled exploratory, and any new criterion or threshold is not represented as pre-specified. A material new question receives its own review record.

120

Approval and change control

The question protocol is approved by a role with appropriate authority before substantial evidence analysis. Changes record date, reason, approving role and effect on interpretation. Minor wording clarification is distinguished from a change to population, criterion or decision.

The final report includes the operative question and discloses material changes. This prevents a review from moving toward an easier question after encountering adverse evidence.

Part XIV

Evidence and implementation instrument B: Evidence register, selection and reconciliation

121

Evidence register purpose

The evidence register identifies every source used for a material finding. It enables another informed reviewer to locate the source, understand its purpose and determine which analytical result depends on it. It does not contain personal detail unnecessary for control.

Each source receives a stable identifier, title, owner, creation purpose, reference period, coverage, version, custody location, access restriction and linked proposition. Derived files identify the source records and transformation.

122

Source classes

Sources are classified as governing document, operational record, observation, participant evidence, work sample, assessment, financial record, contextual statistic or technical opinion. Classification assists interpretation but does not assign quality by itself.

An operational record can be comprehensive and weakly defined. A small observation can be direct and unrepresentative. The register separately assesses relevance, directness, coverage, timing and independence.

123

Frame construction

Where sampling or full-population analysis is intended, the institution constructs a list of eligible units. The frame date, source and known omissions are recorded. Duplicate programmes, inactive learners, closed sites and ambiguous status are resolved or retained as explicit categories.

Comparison between frame totals and authoritative institutional totals can reveal omissions. Agreement does not prove completeness if both derive from the same deficient source. Relevant local registers or service lists provide an additional check.

124

Selection methods

A census includes every eligible unit and is suitable where each case requires decision or the population is manageable. A probability sample supports estimation where every unit has a known selection chance. A purposive sample examines particular conditions or variation but does not estimate prevalence without additional basis.

Selection can combine methods. All high-risk cases may be included while routine cases are sampled. The resulting analysis reports the high-risk cases separately or applies appropriate design treatment. It does not present the combined unweighted sample as representative of the institution.

125

Sample size and consequence

Sample size depends on population, variation, precision, subgroup needs and clustering. The protocol avoids a universal minimum. A small qualitative enquiry can be sufficient to identify a mechanism; it cannot establish an institutional rate. A rare high-consequence failure may require complete case review even where prevalence is low.

The institution does not use a conventional number to create an appearance of statistical adequacy. It explains the precision or inferential limitation that follows from the actual design.

126

Non-response and unavailable records

Every selected unit has a disposition: observed, ineligible, duplicate, refused, unavailable, lost, not reached or pending. Item non-response is separate from whole-case non-response.

Follow-up effort is consistent enough to avoid concentrating completion among convenient respondents. Where replacement is permitted, the rule is fixed before collection and the original non-response remains visible. Replacing a remote or dissatisfied participant with an accessible one changes the evidence and cannot erase the first case.

127

Data reconciliation

Records are reconciled using identifiers, dates and event sequence. Totals are checked across sources that serve different purposes: enrolment and finance, timetable and attendance, assessment and award, referral and service.

Differences are not averaged away. The institution determines whether they reflect timing, definition, duplication, delayed entry or error. Unresolved differences remain in the limitation and can themselves identify a control weakness.

Table B.1. Evidence register fields
FieldPurposeExample of limitation recorded
Source identifier and titleTrace source through analysisLocal title changed during period
Owner and purposeExplain why record was createdRecord designed for payment, not service delivery
Reference periodAlign source with questionOne-day observation within annual claim
Population and coverageEstablish represented unitsPart-time learners absent from frame
Definition and unitPrevent incompatible comparisonTeacher counted by post rather than person
SelectionSupport inferencePurposive sample of disputed cases
MissingnessPreserve uncertaintyReturn date absent for 8 of 120 items
Version and correctionMaintain traceabilityLate corrections entered after first analysis
Access and protectionControl personal informationIdentifiable case restricted to protection role
Linked propositionPrevent unused collectionSource supports access, not learning outcome

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC evidence register.

128

Qualitative evidence coding

Qualitative material is coded against the review question, with allowance for unanticipated material themes. The code definition includes inclusion, exclusion and boundary examples. More than one reviewer examines a subset where judgement is material.

Frequency of a theme is reported only when selection and coding support a count. A vivid account is not described as common without evidence. A rare account can nevertheless require action where it identifies verified harm.

129

Work-sample control

The work-sample list records course or class, task, learner status, selection method, submission status and protection. Samples include missing and incomplete work according to the population rule. Exemplary display material is not substituted for the selected sample.

Reviewers assess only the proposition defined. A sample used to examine feedback timing does not automatically support a judgement about the academic standard of the task.

130

Observation control

Observation records date, duration, setting, observer role, activity, participants, applicable schedule and any departure from ordinary conditions. Criteria distinguish fact from professional judgement.

Where observation can affect behaviour, the review notes announcement and observer presence. Repetition across relevant periods is used where the condition varies. A missed visit is not coded as negative or favourable observation.

131

Survey control

The survey register records eligible population, invitation method, response, item wording, mode, period and exclusions. Reported percentages use the correct denominator and show counts. “No opinion” and skipped items are not combined with favourable responses.

Open invitations are described as responses received, not representative estimates. Where some groups have lower response, the interpretation considers whether their experience is likely to differ.

132

Derived measures

Every derived rate specifies numerator, denominator, exclusions and unit. Where a learner can appear several times, the report states whether the measure concerns learners or events. Rounding follows calculation and does not cause totals to be forced artificially to 100.

Changes in definition or source create a break. Results on either side may be reported but are not described as trend without a valid bridge.

133

Reconciliation sign-off

The analyst signs the evidence register as complete for the stated propositions. A separate reviewer confirms that each material table or judgement traces to registered evidence and that material exceptions have not been omitted.

Sign-off is not a declaration that the evidence is perfect. It confirms that known limitations and corrections are visible and that the evidence supports no broader claim than recorded.

Part XV

Evidence and implementation instrument C: Evaluative judgement casebook

134

Purpose and status

The cases show how evidence is converted into a bounded institutional judgement. They are hypothetical and do not establish required thresholds. Each case uses a different evidence problem so that the method is not reduced to one repeated decision form.

135

Case one: feedback timing and use

An institution has published that marked work will normally be returned within 15 working days. A review examines 160 assignments from four programmes during one semester. Twelve were not submitted and therefore did not enter the return-time denominator. Of 148 submitted items, 104 were returned within 15 working days, 28 later, and 16 had no reliable return date.

The observed timely rate among all submitted items is at least 70.3 per cent if every missing date was late and at most 81.1 per cent if every missing date was timely. Among records with known dates, the rate is 78.8 per cent. The report does not select the latter as the institutional rate without the missing-date qualification.

Programme results range from 58 to 94 per cent among known dates. Assessment sequencing shows that three programmes schedule major tasks in the same fortnight. Learner work indicates that feedback returned after the next related task could not inform that task. The judgement states that the published timing condition was not reliably demonstrated and that the educational consequence is concentrated in programmes with clustered deadlines. It does not conclude that staff commitment or feedback quality is generally poor.

136

Case two: accessible compulsory teaching

The institution lists 22 learners with recorded mobility requirements. Timetable and route inspection show that 20 can reach all compulsory classes. One learner has an authorised room change for two classes but the changed room remains inaccessible through the only open entrance. One learner's record states “arrangement agreed” without a timetable or route.

The institution-wide access rate of 20 out of 22 does not satisfy the criterion that compulsory teaching be accessible to each affected learner. The two cases require immediate verification and remedy. Small numbers limit public detail but not institutional action.

The finding also identifies a control weakness: accommodation is recorded as agreed before its operation is checked. The improvement response therefore addresses both the two individual routes and the verification point in the accommodation process.

137

Case three: withdrawal from an evening programme

Administrative data show an 18 per cent first-year withdrawal rate in the evening programme and 9 per cent in the daytime programme. Entry qualifications, age, employment and missing exit reasons differ between groups. A survey reaches 42 per cent of evening withdrawers, who frequently report timetable and transport difficulties.

The evidence supports a material difference in recorded withdrawal and identifies plausible access constraints among respondents. It does not establish that timetable and transport caused the whole rate difference, because respondent selection and learner characteristics may contribute.

The decision is a targeted timetable and access enquiry, with a small reversible schedule adjustment in one intake and direct follow-up. The institution does not delay all action until perfect causal evidence is available, nor present the pilot as a proven retention intervention.

138

Case four: practical curriculum component

A programme specification requires three supervised practical experiences before award. Records for 96 final-year learners show 91 with three verified experiences, 3 with two and 2 with ambiguous records. Award decisions have already been entered for all 96.

The review does not report 94.8 per cent completion as generally satisfactory. The criterion applies to each award. The five cases require immediate record and provision review under the competent academic authority. Ambiguous evidence is not coded as completion.

The system finding concerns reconciliation between practical completion and award approval. Action includes correction or provision for affected learners and a control that prevents final approval before the required evidence is verified.

139

Case five: student-support response

A support service reports that 92 per cent of referrals were “closed” within ten days. Examination of 120 cases shows that closure includes appointment offered, learner contacted, referral transferred and service completed. Only 54 cases record that the defined support took place; 19 are transfers without confirmation; 31 are offers with no recorded attendance; and 16 lack a clear outcome.

The published indicator measures administrative handling, not service received. The finding is not that only 54 learners benefited, because non-attendance and missing records require further classification. It is that the institution cannot establish reach from the closure measure.

Improvement first revises case states and confirms responsibility after transfer. A later review can assess response time and reach separately.

140

Case six: apparent improvement in completion

Programme completion rises from 72 to 81 per cent between two cohorts. During the same period, entry selection changes, the definition of completion excludes deferred learners, and a low-completion pathway closes. Recalculation under the earlier definition and common programme scope produces 75 per cent for the later cohort.

The evidence supports improvement of three percentage points on the more comparable basis, subject to cohort difference. It does not support the published nine-point improvement as a like-for-like change. The institution corrects the public comparison and marks the definition break.

141

Case seven: teaching observation after development

Twenty teachers receive support focused on structured learner questioning. Observation is completed for 15 before and after the support; three have only follow-up evidence and two are unavailable. Eleven of the comparable 15 show the defined practice at follow-up, compared with four at baseline.

The result supports increased observed use among the 15 teachers with paired observations. It does not establish an institution-wide prevalence or a learning effect. Observation timing, selection and possible observer response remain limitations.

The institution continues targeted support, observes the missing and follow-up-only cases and examines learner participation before considering broader expansion.

Table C.1. Case judgements and prohibited overstatement
CaseSupported judgementFurther actionProhibited overstatement
Feedback timingPublished return period not reliably demonstrated; variation concentratedCorrect records and workload sequenceFeedback is generally ineffective
Compulsory accessTwo learners lack verified full accessImmediate remedy and process verificationAccess is acceptable for 91 per cent
Evening withdrawalRecorded difference and plausible respondent-reported barriersTargeted enquiry and reversible testTimetable caused the full withdrawal gap
Practical componentFive award records lack verified fulfilmentCompetent case review and award control95 per cent compliance is sufficient
Support responseClosure measure cannot establish service reachRedefine states and confirm transfer outcomeOnly completed-state cases received benefit
Completion trendComparable improvement smaller than published differenceCorrect definition and continue cohort analysisNine-point like-for-like improvement
Teaching practicePaired sample shows increased observed useComplete coverage and examine participationSupport caused institution-wide learning gain

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC hypothetical judgement casebook.

142

Lessons from the cases

The cases show why denominators, criteria and event states matter. A favourable majority cannot compensate for an individual minimum obligation; an administrative status may not correspond to service; and a trend can be created by definition change.

They also show that limited evidence can still support proportionate action. The response is to align the claim and decision with the evidence, not to withhold every decision until a comprehensive evaluation is possible.

Part XVI

Evidence and implementation instrument D: Improvement and verification instruments

143

Finding-to-action statement

The instrument begins with the final judgement, confidence and affected population. It then records the condition to change and the immediate mechanism. It does not restate the broad review theme.

The action owner confirms authority. Where action crosses units, each contribution and the final accountable role are identified. An unresolved authority question is referred before deadlines are assigned.

144

Option appraisal

At least two feasible options are considered for material actions unless immediate protection leaves only one safe course. The appraisal records expected operating result, time, full resource demand, reach, recurrent implication, risk and reversibility.

The option not selected remains in the record with reason. This allows later review to determine whether the decision was reasonable on the evidence then available, even if the result is unfavourable.

145

Milestone definition

Milestones are observable states rather than verbs such as “progress” or “develop.” They state the item, population, date and evidence. “Assessment schedule approved for all first-year programmes and reflected in the operative timetables” is verifiable. “Improve assessment coordination” is not.

Operating status is always included. Completion of preparation is insufficient where learners have not received the intended service.

146

Implementation log

The log records planned date, actual date, responsible role, evidence, reach, adaptation, unresolved issue and next decision. Entries are made close to the event and can be corrected with a trace.

The log is not a narrative of effort. It permits analysis of where the action stopped and whether the implemented form retained the intended mechanism.

Table D.1. Implementation log
StagePlanned conditionActual evidenceReachAdaptation or varianceDecision
AuthorisedValid decision and resourcesApproval reference and available meansUnits covered by authorityLimits or conditions of approvalProceed, revise or refer
InitiatedCritical work has begunDated task evidenceIntended units entering actionTiming or route changeContinue or correct
CompletedSpecified tasks finishedDeliverable or verified transactionUnits receiving completed actionMissing componentVerify operation or reopen work
OperatingIntended service functionsDirect service evidenceEligible population reachedUneven access or instabilityContinue, adapt, expand or cease

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC implementation log.

147

Baseline sheet

The baseline sheet records measure, definition, population, source, period, value, missingness and limitation. It also records whether evidence was reconstructed after the action began. A reconstructed baseline remains usable for some decisions but does not carry the same confidence as a pre-action measure.

Where the action begins urgently before a full baseline, the institution records the safest available pre-action evidence and establishes current status as soon as possible. It does not delay protection solely to improve evaluation design.

148

Reach account

Reach uses the eligible population as denominator and distinguishes offered, received, completed and using states where relevant. The appropriate state depends on the proposition.

Reasons for non-reach are classified. The institution examines whether access, schedule, language, disability, cost or selection created uneven participation. It does not attribute every non-participation to learner choice without evidence.

149

Operating verification

The verifier examines the service under ordinary conditions and within a period suited to the mechanism. A new timetable is checked during operation; a revised assessment sequence is checked across the relevant task cycle; a support process is checked from referral to service.

The source and verifier are proportionate to consequence. The person who completed a low-risk action may record operation, subject to sample review. A high-consequence result requires competent and sufficiently independent confirmation.

150

Outcome sheet

The outcome sheet preserves the original measure and adds any balancing measures. It shows baseline and follow-up counts, denominators, missingness, definition change and relevant concurrent events.

Interpretation uses the weakest necessary causal wording. “Improved after implementation” is not replaced by “was caused by” unless the design warrants that statement.

151

Sustainability review

The sustainability record covers recurrent finance, workload, technical competence, maintenance, governance ownership and dependence on a particular person. It identifies the period for which continuation is reasonably secured.

An action supported by temporary funds is not necessarily rejected. It is classified accurately, with a decision date before resources end and protection for affected learners.

152

Closure certificate

Closure identifies the original finding, action, operating evidence, distribution, unresolved risk, responsible approval and date. It records whether longer outcome monitoring continues. “No further action” requires a reason.

Closure can be reversed if evidence proves incorrect or the condition recurs. The register links recurrence to the earlier case, preventing repeated short-term correction from appearing as separate success.

Part XVII

Evidence and implementation instrument E: Quantitative interpretation for institutional review

153

Counts before percentages

Institutional reports show counts with percentages unless disclosure risk prevents it. Counts reveal the scale and help readers identify results based on small populations. The eligible denominator is defined before calculation.

If 36 of 40 respondents report access to a service, the respondent rate is 90 per cent. If 60 learners were eligible, the evidence establishes confirmed access for 60 per cent of the eligible population and a 90 per cent favourable response among respondents. Neither figure alone describes the unknown 20.

154

Person, event and registration rates

One person can generate several events or registrations. A feedback review may report assignments returned, learners receiving at least one timely return, or courses meeting the return condition. These measures answer different questions.

The unit appears in the title, numerator and denominator. Event rates are not described as the proportion of learners affected. Where clustering by learner or course matters, the analysis acknowledges that events are not independent observations.

155

Percentage points and relative change

A rate rising from 50 to 60 per cent increases by 10 percentage points and by 20 per cent relative to the original rate. The report names the calculation. It does not use “10 per cent improvement” where the intended meaning is 10 percentage points.

Relative change can appear large when the baseline is small. Both baseline and follow-up values remain visible.

156

Cohort measures

Cohort entry is fixed according to the review question. Continuation may mean re-enrolment in the next period; completion may require fulfilment within a stated duration. Transfers, deferrals, deaths and unknown destinations are shown separately where material.

Apparent completion can rise when slow completers are removed or the observation window changes. Every cohort table states entry basis, outcome window and treatment of status changes.

157

Cross-sectional and longitudinal evidence

Cross-sectional evidence compares different populations at one or more points. Longitudinal evidence follows the same people or units. A repeated cross-section can describe institutional change but may reflect changing composition. A panel can describe change among continuing cases but may lose those most affected by failure.

The institution selects the design suited to the decision and, where useful, reports both. It does not describe a panel result as if it covered all current learners.

158

Missing-data bounds

Simple bounds show whether missingness could alter a conclusion. If 70 of 90 observed cases meet a condition and 10 are unknown, the all-case rate lies between 70 and 80 per cent. A threshold of 75 per cent cannot be declared met or unmet without an assumption or further evidence.

Bounds can be wide and uninformative; that is itself evidence that record quality limits the decision. A modelled adjustment is used only where its assumptions are stated and defensible.

159

Average, median and distribution

An average response time can be affected by a few very long cases. The median describes the middle case but can conceal the same tail. Institutional review therefore reports an appropriate distribution: median, range or intervals, and cases beyond a service threshold.

Where the obligation concerns every urgent case, a favourable average cannot compensate for one dangerous delay. Distribution is interpreted against the criterion.

160

Small groups

Small-group rates can change sharply with one case. The report uses counts, avoids rhetorical comparison and considers multi-period aggregation only where the service and population remain comparable. Aggregation must not delay remedy for verified individuals.

Public disclosure is limited where a small cell could identify a person. Internal decision-makers receive the evidence needed under appropriate access.

161

Sampling uncertainty

A sample estimate differs from the full population because of selection. Where a probability sample is used, the report provides an appropriate measure of sampling uncertainty and reflects stratification or clustering. The calculation does not address non-response, measurement error or bias in the frame.

Where a non-probability sample is used, a conventional margin of error is not attached. The report describes selection and limits the inference.

162

Practical and statistical importance

A statistically detectable difference may be too small to alter an educational decision, while a large educational difference in a small population may not meet a conventional significance rule. The institution considers effect, consequence, confidence and population together.

Formal tests are used when their assumptions and decision value are clear. They do not replace an account of counts, definitions and distribution.

163

Benchmarking

A benchmark is used only where the definition, population and period are comparable and where the benchmark has a legitimate relation to the review criterion. An external average is not automatically a minimum standard or a target.

Internal comparison over time can be more useful but remains sensitive to changed programmes, assessment or intake. The report identifies every material break.

164

Composite measures

Combining indicators requires explicit weights and rules for missingness and compensation. Because access, safety, academic integrity and outcomes are not interchangeable, this report favours component findings.

If an institution uses a composite for screening, it retains all components and defines conditions that cannot be compensated. The composite does not become the sole basis for closure or public ranking.

Table E.1. Quantitative interpretation controls
IssueRequired disclosureFrequent errorCorrective treatment
DenominatorEligible, observed and excluded countsRespondents treated as all eligibleShow response and all-case status
UnitPerson, event, course or caseEvent rate described as learner rateName and preserve unit
ChangeBaseline, follow-up and calculationPercentage points called per centState both forms where useful
CohortEntry rule, window and status treatmentDifferent exposure periods comparedUse common definition or mark break
MissingnessNumber, reason and sensitivityBlank coded favourablyKeep unknown or show bounds
DistributionRange, intervals and critical tailAverage used aloneReport criterion-relevant exceptions
SamplingDesign and uncertaintyMargin attached to convenience sampleLimit inference to observed sample
BenchmarkDefinition, source and statusAverage treated as universal standardUse as context or justified criterion only

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC quantitative interpretation controls.

165

Worked calculation: response-time distribution

A service handles 200 valid cases. It resolves 120 within 5 days, 50 in 6–10 days, 20 in 11–20 days and 10 after 20 days. The proportion within 10 days is 170/200, or 85 per cent. The proportion beyond 20 days is 5 per cent.

If all 10 cases beyond 20 days are low-risk, the tail may indicate capacity or process inefficiency. If two are urgent protection cases, the same aggregate has a different consequence. Severity and timing must therefore be analysed together.

166

Worked calculation: changed population

A programme reports 80 completions among 100 entrants in the later cohort and 63 among 90 in the earlier cohort. The completion rates are 80 and 70 per cent. Further review shows that 10 entrants in the later cohort transferred from a closed programme after completing half the curriculum; 9 completed.

The published all-entrant rate remains a valid description if the entry rule includes transfers, but it is not directly comparable with the earlier cohort without examining entry stage. Excluding the transfers solely because they improve the result would also be invalid. The report presents the structural change and, where useful, a common-entry subgroup.

Part XVIII

Evidence and implementation instrument F: Challenge, dissent and correction protocol

167

Challenge questions

Every draft finding is tested through five questions: What is the strongest contrary evidence? Which eligible cases are absent? What other explanation would lead to a different action? Which assumption most affects the result? What harm could follow if the judgement is wrong?

Answers are recorded beside the judgement. “None known” is acceptable only after the relevant evidence and participants have been considered.

168

Factual correction

The subject unit receives the source extract and proposed fact. It can identify an incorrect record, date, population or calculation. Correction requires evidence and preserves the original analytical trace.

Disagreement with the implication is not classified as factual correction. It enters interpretation or action response.

169

Alternative interpretation

An alternative interpretation states the same evidence and proposes a different meaning. The review panel evaluates which interpretation better accounts for the pattern and limitations. It may retain both where evidence cannot resolve the difference.

The final record avoids artificial certainty. An unresolved alternative can change confidence, action or the need for further evidence.

170

Dissent

Material dissent records the proposition, evidence, author or represented role, and consequence for judgement. Personal information is protected where disclosure creates risk. Anonymous assertion without evidence can still identify an enquiry but is not treated as established fact.

Dissent is not measured only by number of supporters. A single technically supported objection can invalidate a calculation or reveal a serious case.

171

Conflict of interest

Reviewers disclose responsibility for the activity, financial or professional interest, close reporting relationship and prior advocacy of an option. The review authority decides whether declaration, limited participation or replacement is necessary.

An interest does not automatically disqualify a knowledgeable participant. It changes the role and need for challenge.

172

Correction after decision

When new evidence materially changes a finding, the institution reassesses every dependent action and public statement. It records whether an action remains justified on other grounds. Correction is not delayed to protect consistency with an earlier report.

Table F.1. Challenge disposition
Challenge typeEvidence requiredDecisionRecord
Factual errorCorrect source or calculationCorrect, reject or seek verificationBefore and after value with reason
Scope objectionPopulation or exclusion evidenceExpand, narrow or retain scopeEffect on inference
Alternative explanationPlausible mechanism and relevant evidenceAccept, reject or leave unresolvedConfidence and action consequence
Method objectionDemonstrated selection, measurement or analytical weaknessRevise method or qualifyReanalysis and limitation
Action objectionAuthority, feasibility, cost or harm evidenceChange or retain decisionOption appraisal and reasons

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC challenge disposition protocol.

173

Closure of challenge

A challenge closes with a reasoned disposition and notification to the relevant participant where appropriate. Repeated challenges on the same unresolved weakness are linked and escalated.

The institution monitors whether people can use the correction route without adverse treatment. A formal route with no credible use is not assumed effective.

Part XIX

Evidence and implementation instrument G: Public reporting model

174

Summary finding

The summary states the question, population and principal finding in direct language. It avoids adjectives of institutional excellence and does not begin with activity undertaken.

Where the finding is provisional, that status appears in the first account rather than a distant note.

175

Evidence note

The evidence note gives sources, period, coverage, selection and principal limitation. It is concise but sufficient to prevent an internal sample or respondent result from appearing comprehensive.

Material definition changes appear beside the affected result.

176

Improvement status

The public status uses the controlled states: finding confirmed; decision authorised; action initiated; specified tasks completed; operating result verified; or closed with reason. Dates and responsible authority accompany material open actions.

An institution may group minor actions, but it does not group a high-consequence overdue case into a favourable aggregate.

177

Outcome statement

The outcome statement identifies baseline, follow-up and population. It uses causal wording only where design supports it. Where no comparable baseline exists, current status is reported.

The statement includes distribution where a group remained underserved.

178

Limitation and unresolved matter

The report states the limitation most likely to change interpretation. It also identifies unresolved conditions outside institutional authority and the action requested from the competent body.

The phrase “work is ongoing” is insufficient without the service affected, responsible role and next decision date.

Table G.1. Illustrative public record
Public elementIllustrative entry
QuestionWhether all compulsory first-year practical components were delivered to the 2004/05 cohort
CriterionApproved programme specification in force for the cohort
EvidenceTimetables, attendance and verified completion records for all 186 enrolled learners
Finding174 learners have complete verified provision; 8 missed one component with a scheduled replacement; 4 records remain unresolved
JudgementThe requirement is not yet demonstrated for 12 learners; four cases require immediate record and provision review
DecisionReplacement provision authorised for eight; competent academic review opened for four
StatusReplacement initiated; operating completion not yet verified
LimitationCompletion record and attendance record disagree in four cases

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC hypothetical public reporting model.

179

Annual synthesis fields

The annual synthesis reports material questions reviewed, coverage by educational function, open findings at start, new findings, operating closures, other reasoned closures, overdue actions, recurrence and material limitations. It states areas not reviewed.

These fields are interpreted together. More findings can reflect better detection; fewer closures can reflect stronger operating evidence. The synthesis therefore avoids an unsupported performance score.

Part XX

Evidence and implementation instrument H: Controlled terminology

180

Institutional self-review

A structured enquiry owned by the institution into a defined aspect of its educational work, producing evidence, judgement and a response. The term does not imply that institutional evidence is sufficient for every external or public decision.

181

Review question

A bounded question identifying provision, population, period and decision. A theme or domain is not a review question.

182

Criterion

The applicable or adopted condition against which evidence is judged. Its source, authority and operational meaning are stated.

183

Institutional account

A dated description of provision, population, resources, responsibilities and context. It is not an evaluative judgement by itself.

184

Evaluative finding

An evidence-based statement comparing an observed condition with a criterion for a defined population and period.

185

Improvement evidence

Evidence that a response was implemented, reached its intended population and changed the defined operating condition or outcome.

186

Assurance evidence

Evidence that permits another informed person to understand and test the basis, method, limitations and consistency of a judgement.

187

Direct evidence

Evidence observing the condition stated in the proposition. Directness depends on the claim: an invoice is direct evidence of a transaction and indirect evidence of learner access.

188

Corroboration

Support from an additional source or method with sufficiently different errors. Repeated summary of one source is not corroboration.

189

Population

All persons, events, units or records to which the question and intended conclusion apply.

190

Frame

The operational list from which cases are selected or coverage is assessed. A frame can omit eligible units or include ineligible ones.

191

Missingness

The absence of a case, response or item from the evidence. Its reason and possible relation to the finding are examined.

192

Materiality

The significance of a finding considering educational consequence, rights or safety, number and identity of people affected, duration, recurrence and integrity of provision or awards.

193

Confidence

The degree to which the available evidence supports the stated judgement. It is separate from whether the judgement is favourable or adverse.

194

Participation

The involvement of affected and responsible persons in supplying evidence, interpretation or decision, with roles and safeguards made clear.

195

Contrary evidence

Evidence inconsistent with the proposed finding or preferred explanation. It is examined and retained where material.

196

Improvement proposition

The stated relationship between an action, its mechanism, the operating condition expected to change and the intended result.

197

Reach

The eligible people or units that actually received the action or service. Offer, receipt, completion and use are distinguished where material.

198

Implementation fidelity

The extent to which critical components of the specified action operated. Fidelity does not require every incidental detail to remain unchanged.

199

Operating result

The first verified educational or institutional service condition expected from the action. It is distinct from administrative completion and from a longer outcome.

200

External perspective

An informed view from outside the activity or institution that tests criteria, evidence, reasoning or result. It does not imply comprehensive independent re-performance.

201

Provisional judgement

A judgement sufficient for a limited and reviewable decision but subject to a stated uncertainty and further evidence.

202

Structural break

A discontinuity in definition, population, source, assessment or provision that limits comparison over time.

203

Balancing measure

A measure selected to detect a material unintended effect or indicator displacement associated with an improvement action.

204

Case closure

A reasoned decision that the operating result is verified, the case is invalid or duplicated, or residual status has been lawfully accepted by competent authority. Referral or planned action is not closure.

205

Public account

A proportionate report of question, evidence, finding, judgement, action, verification and limitation. It protects individuals while preserving material institutional accountability.

Part XXI

Evidence and implementation instrument I: Selecting and governing the institutional review programme

206

From individual concern to review programme

An institution cannot examine every educational activity at the same depth each year. It therefore maintains a review programme that combines recurring essential questions, follow-up to open findings, emerging evidence and periodic examination of less visible functions. Selection is recorded so that unreviewed areas do not disappear from governance attention.

The programme is not a list of reports due. It is an account of which educational risks and improvement questions require institutional judgement, when evidence will be sufficiently mature and who can act on the result.

207

Sources for selection

Selection draws on learner access and outcomes, complaints and appeals, staff evidence, material change, new or expanding programmes, previous review limitations, unresolved actions, resource shifts and external policy or legal change. It also includes a rotation across curriculum, teaching, assessment, support, governance and educational conditions.

An indicator can initiate a question but does not predetermine the finding. A falling rate, rising complaint count or unusual variation is screened for definition, coverage and consequence before a full review is commissioned.

208

Essential recurring questions

Some questions recur because the consequence of failure is high or the condition changes frequently. They can concern safe and non-discriminatory access, delivery of required curriculum, integrity of assessment and awards, protection, and operation of essential learner services.

Recurrence does not require identical annual collection. A stable control can be tested through a risk-based sample, while known weak or changed areas receive complete or more frequent review. The institution explains why the current coverage is sufficient.

209

Material change

A new programme, site, timetable, assessment regime, information system, partnership or rapid increase in enrolment can change assumptions on which earlier evidence depended. The programme schedule therefore includes review before implementation where risk can be prevented and after operation where actual effect must be observed.

The review period is long enough for the relevant mechanism to operate but not so long that a serious condition persists without detection. Interim evidence and safeguards are used where mature outcome evidence will take time.

210

Follow-up priority

Open material findings take priority over repeated diagnosis of the same condition. The programme identifies the date at which operating evidence is expected and the authority responsible for delay.

A delayed action can require a revised decision rather than another progress report. Where the original criterion remains unmet, the institution considers interim protection, additional authority or cessation of an ineffective course.

211

Visibility and selection bias

Units with better records, active leadership or easy access can receive more review attention. The programme tests for under-coverage of remote sites, small programmes, part-time or evening provision, minority languages and groups less able to raise concerns.

Absence of complaints is not treated as evidence of low risk where the route is inaccessible or confidence is weak. Selection includes direct enquiry into conditions that ordinary reporting may not reveal.

212

Review burden

The programme estimates time required from learners, teachers, administrators, analysts and decision-makers. Reviews sharing a population or source are coordinated, subject to confidentiality and purpose. Unused evidence requirements are removed.

Burden is assessed against decision value. A repeated survey that produces no changed decision is revised or ceased. A high-consequence review is not omitted solely because it is difficult, but its method is made as direct and proportionate as possible.

213

Analytical capability

The schedule matches questions to available expertise. Statistical comparison, assessment validity, facility safety, accessibility and confidential protection evidence may require different competence. The institution obtains support or narrows the question where competence is unavailable.

Capability planning includes time for evidence preparation, challenge and follow-up, not only collection. Commissioning many reviews without capacity to decide or verify creates an appearance of activity and an expanding unresolved stock.

Table I.1. Review programme selection record
Selection considerationEvidenceScheduling consequenceSafeguard against misuse
Immediate safety or rights riskVerified case or credible protective concernImmediate competent action and reviewProtection precedes comprehensive enquiry
Open material findingImprovement register and due dateFollow-up before new review of same issueNo closure through repeated planning
Material institutional changeApproved change and affected populationBaseline before change and early operating reviewDo not infer result from implementation plan
Unusual outcome patternDefined indicator, counts and comparisonScreening followed by bounded questionIndicator does not assign cause
Under-reviewed populationCoverage map and participation evidencePlanned direct or sampled enquiryWeak reporting is not low need
Recurring essential controlPrior results and change riskPeriodic complete or sampled testFrequency tied to consequence and stability
Developmental enquiryInstitutional objective and feasible mechanismScheduled after essential obligationsAspiration does not displace minimum service

Source and methodological notes are stated immediately below the table in the authoritative Markdown text.

Source: ICEQC institutional review programme selection record.

214

Sequencing

The programme sequences diagnosis, action and verification. It avoids commissioning a broad new review while the evidence required to determine the previous action is due. Where related questions can be resolved together, their criteria and populations remain distinguishable.

Sequencing also protects the academic year. Collection and observation occur when the service operates normally, unless disruption itself is the subject. Decisions are timed early enough to affect the next relevant cycle.

215

Governing approval

The competent governing authority approves the programme, material changes and deferred high-consequence questions. Approval includes a statement that resources and decision routes are sufficient for the planned work.

Leadership can add an urgent review, but it records what other work is displaced. Removal of a review after adverse preliminary evidence requires explicit reason and governing visibility.

216

Programme monitoring

Monitoring shows questions scheduled, initiated, decided and verified; populations covered; open material findings; overdue decisions; participant burden; and correction. It distinguishes review completion from improvement verification.

An annual count of completed reviews is interpreted with scope and consequence. One rigorous review that restores a material service can have greater public value than many descriptive reports.

217

Periodic programme evaluation

The institution periodically asks whether the review programme detects material conditions in time, represents its diverse provision, supports decisions and reduces recurrence. It examines missed cases, corrections after external challenge, unused findings and actions that never reached operation.

The result can change selection, authority, evidence definitions and capability. Programme evaluation is not presented as proof that all educational provision is sound. It concerns the institution's capacity to enquire and respond.

218

Minimum annual statement

The minimum annual statement identifies the educational functions and populations reviewed, the basis for selection, material areas not reviewed, open findings carried forward, operating results verified and significant limitations in evidence or capacity.

This statement allows the public and governing authority to understand coverage without publication of sensitive case detail. It also prevents a favourable annual narrative from implying that silence equals confirmed quality.

219

Programme conclusion

A credible programme gives priority to consequence, unfinished remedy and under-observed populations. It reserves sufficient capacity for verification. It does not use fixed rotation to postpone an urgent condition or repeated data collection to avoid a difficult decision.

The proper institutional measure is whether review attention reaches the educational conditions most in need of judgement and whether those judgements produce verified change. Scheduling discipline is therefore part of self-review integrity, not a separate administrative exercise.

220

Selection under uncertainty

Evidence available at programme-planning stage will often be incomplete. The institution records whether the uncertainty concerns the existence, scale, cause or consequence of a condition. It then selects the least burdensome enquiry capable of resolving the uncertainty that matters to the decision. A brief screening may establish whether a full review is warranted; it is not published as the final evaluation.

Uncertainty is treated differently where delay can cause harm. A credible indication of unsafe access, discriminatory exclusion or invalid award practice leads to precaution and competent examination even when prevalence is unknown. By contrast, an unverified difference in a low-consequence indicator may first require confirmation of definition and source. This distinction prevents the review programme from giving equal urgency to every unexplained variation while also preventing incomplete records from postponing protective action.

The selection record states what evidence would cause the institution to increase, reduce or end the enquiry. A review that continues despite repeated evidence that its question cannot inform a decision is closed or redesigned. A review that uncovers a wider material condition is expanded through an approved change, with the original and enlarged populations kept distinct.

References

  1. REF-01

    World Education Forum. The Dakar Framework for Action: Education for All — Meeting Our Collective Commitments. 2000. ED-2000/WS/27.

    The public commitment to quality, measurable learning outcomes, participation, accountability and national action for Education for All.

    https://unesdoc.unesco.org/ark:/48223/pf0000121147
  2. REF-02

    Education for All Global Monitoring Report Team. Education for All: The Quality Imperative. 2004. EFA Global Monitoring Report 2005.

    The contemporaneous account of education quality, school improvement, enabling inputs, teaching and learning, outcomes, capacity and distribution.

    https://unesdoc.unesco.org/ark:/48223/pf0000137333
  3. REF-03

    European Parliament and Council of the European Union. Recommendation of 12 February 2001 on European Cooperation in Quality Evaluation in School Education. 2001. 2001/166/EC.

    The policy basis for transparent evaluation, school self-evaluation, balanced external perspective, stakeholder participation, methodological support and improvement use.

    https://eur-lex.europa.eu/legal-content/EN/TXT/PDF/?uri=CELEX:32001H0166
  4. REF-04

    Conference of European Ministers Responsible for Higher Education. The European Higher Education Area — Achieving the Goals: Bergen Communiqué. 2005.

    The contemporaneous ministerial position on institutional responsibility, systematic internal mechanisms, external perspective, student involvement, public responsibility and stocktaking.

    https://ehea.info/media.ehea.info/file/2005_Bergen/52/0/2005_Bergen_Communique_english_580520.pdf
  5. REF-05

    United Nations Economic and Social Council. Fundamental Principles of Official Statistics. 1994. E/RES/1994/29.

    Principles of relevance, impartiality, professional method, source accountability, correction and confidentiality applicable to institutional evidence.

    https://unstats.un.org/unsd/dnss/gp/fundprinciples.aspx
  6. REF-06

    United Nations. Convention on the Rights of the Child. 1989. A/RES/44/25. Articles 2, 12, 28 and 29.

    The rights basis for non-discrimination, participation, access and the substantive aims of education.

    https://www.ohchr.org/en/instruments-mechanisms/instruments/convention-rights-child
  7. REF-07

    United Nations Committee on Economic, Social and Cultural Rights. General Comment No. 13: The Right to Education. 1999. E/C.12/1999/10.

    The availability, accessibility, acceptability and adaptability framework used to test whether self-review remains directed to the right to education.

    https://docstore.ohchr.org/SelfServices/FilesHandler.ashx?enc=4slQ6QSmlBEDzFEovLCuW1AVC1NkPsgUedPlF1vfPMJb2C7KRvOaewo5P54LEjsHEpeN01Dr2U7Zw%2BK5%2F3WZKUclog1%2BBe3TC8O6zK4NNSgWPJ0yZhtq61OlL
  8. REF-08

    World Bank. World Development Report 2004: Making Services Work for Poor People. 2003.

    The distinction between financed inputs, provider action and services received, and the role of information and accountability in public service delivery.

    https://documents1.worldbank.org/curated/en/832891468338681960/pdf/268950WDR00PUB0ces0work0poor0people.pdf
  9. REF-09

    UNESCO Institute for Statistics. Global Education Digest 2004: Comparing Education Statistics Across the World. 2004.

    International definitions, coverage principles and comparability limitations relevant to institutional use of system and outcome indicators.

    https://uis.unesco.org/sites/default/files/documents/global-education-digest-2004-comparing-education-statistics-across-the-world-en_0.pdf
  10. REF-10

    Organisation for Economic Co-operation and Development. Learning for Tomorrow's World: First Results from PISA 2003. 2004.

    Comparative learning evidence and the limitations of attributing institutional performance from aggregate outcomes without context and design.

    https://www.oecd.org/education/school/programmeforinternationalstudentassessmentpisa/34002216.pdf
  11. REF-11

    International Labour Organization and United Nations Educational, Scientific and Cultural Organization. Recommendation concerning the Status of Teachers. 1966.

    Principles concerning professional responsibility, consultation, teacher participation, working conditions and fair evaluation.

    https://www.ilo.org/ilo-unesco-recommendation-concerning-status-teachers-1966
  12. REF-12

    United Nations General Assembly. United Nations Millennium Declaration. 2000. A/RES/55/2.

    The global public-interest setting for equitable development, accountability and universal primary education.

    https://undocs.org/A/RES/55/2