QualityHero platform logo
Back to blogQuality Assurance

Running an FE Quality Calibration Event

Ensure your internal quality judgements are consistent and accurate. This guide offers a practical model for running a calibration event to strengthen your SAR.

20 August 2026

Running an FE Quality Calibration Event

Reliable self-assessment depends on consistent and accurate quality judgements. If one manager's 'strong standard' is another's 'expected standard', your Self-Assessment Report (SAR) becomes unreliable, and your Quality Improvement Plan (QIP) may target the wrong areas. Running a structured calibration event is a powerful way to ensure everyone is applying your quality framework and the inspection toolkit consistently.

This process builds a shared understanding of what quality looks like across different provision types and evidence sources - from observations of teaching and training to work scrutiny and learner voice. It moves your team from isolated judgements to a shared, validated view of provider performance.

Preparing for Your Calibration Event

Thorough preparation is key to a successful session. The goal is to create a focused, evidence-based discussion, not a debate based on opinion. Before the event, you should:

  • Select your evidence packs: Curate anonymised evidence for discussion. A pack might include a set of teaching observation notes, examples of learner work, a summary of learner voice feedback, or data on achievement for a specific group.
  • Choose a range of quality: Ensure your packs represent different levels of performance. Include evidence you believe might demonstrate 'needs attention', 'expected standard', and 'strong standard' to facilitate a rich discussion.
  • Define a clear focus: Don't try to calibrate everything at once. You might focus one session on judgements for 'curriculum, teaching and training' in apprenticeships, and another on 'participation and development' for adult learning.
  • Brief your participants: Send the evidence packs to participants in advance. Clearly communicate the purpose of the event: to calibrate judgements and improve consistency, not to re-judge a colleague's original assessment. This fosters a safe, professional environment.

Structuring the Calibration Session

A structured approach ensures fairness and efficiency. A proven model involves moving from individual reflection to a whole-group consensus, facilitated by a quality lead.

  • Step 1: Individual Review: Allow 15-20 minutes for participants to independently review the first evidence pack. They should make a private judgement against the relevant criteria from the Further Education and Skills Inspection Toolkit and note their key reasons.
  • Step 2: Small Group Discussion: In small groups of three or four, participants share their initial judgements and, more importantly, their reasoning. The discussion should focus on connecting specific pieces of evidence to the toolkit's criteria.
  • Step 3: Whole Group Moderation: A facilitator asks each small group to share its consensus judgement and rationale. The facilitator guides a whole-group discussion, highlighting areas of agreement and exploring differences. The aim is to arrive at a 'best-fit' judgement for the evidence presented.
  • Step 4: Document the Rationale: This is the most critical step. Document precisely why the final moderated judgement was reached. For example: "We judged this as 'strong standard' because the evidence showed that almost all learners made rapid progress from their starting points and the curriculum was expertly sequenced." This documented rationale becomes a powerful reference point.

Key Questions to Guide Discussion

To keep the conversation evaluative and focused on the toolkit's requirements, facilitators and participants should use probing questions:

  • What is the direct impact on the learners or apprentices?
  • How typical might this evidence be? What other evidence would we need to confirm our judgement?
  • Which evaluation area does this evidence relate to most strongly?
  • What specific evidence stops this from being 'strong standard'? Or what makes it more than 'expected standard'?
  • Where in the toolkit criteria is this strength or weakness described?
  • Are we making assumptions, or is the judgement based purely on the evidence provided?

Using the Outcomes to Drive Improvement

A calibration event is only worthwhile if it leads to action. The insights gained should be used to strengthen your quality assurance processes and inform improvement strategies.

  • Update internal guidance: Use the documented rationales as exemplars in your internal quality handbooks or staff training materials. This helps socialise the shared understanding of quality standards.
  • Target professional development: If the session reveals common misunderstandings about an aspect of the toolkit or a specific teaching strategy, this provides a clear mandate for targeted CPD.
  • Refine your Self-Assessment Report: The calibrated judgements and detailed rationales should be used to update your live SAR, increasing the accuracy of your self-grading.
  • Schedule future sessions: Make calibration a regular part of your quality cycle. Covering different provision types and evaluation areas throughout the year maintains consistency and continuously develops staff expertise.

Where this fits in QualityHero

Creating a robust, evidence-based self-assessment is the core purpose of quality calibration. QualityHero's SAR module enables you to record these calibrated judgements and their detailed rationales directly against each provision type. Evidence from the event can be linked to your judgements, and any required actions can be logged and tracked in the QIP module, ensuring that insights from your professional dialogue lead to tangible improvements.

#Quality Assurance#Self-Assessment#Professional Development#Leadership

Want this in your workspace?

QualityHero turns insights like this into actions, evidence and governance-ready reports.