New

Now in Claude, ChatGPT, Cursor & more with our MCP server

Back to blog
Programme evaluation9 min read

Evaluating Pilot Training Programmes: Course Evaluation in the Age of EASA Competency-Based Training

Flight training moved from counting hours to certifying competencies a decade ago. If your evaluation instrument still asks cadets to rate teaching on a five-point scale, it is measuring a paradigm the training has already left behind.

Koji Education Team

Product ·

Flight training has spent the last decade doing something most of higher education has only talked about: moving from counting hours to certifying competencies. If your course-evaluation instrument still asks pilot cadets to rate "the quality of teaching" on a five-point scale, it is measuring a paradigm the training itself has already left behind.

The short answer: Under the European Union Aviation Safety Agency (EASA) framework, modern pilot training is increasingly competency-based, and its Approved Training Organisations (ATOs) are assessed on whether cadets demonstrate defined competencies and observable behaviours — not on satisfaction. Course evaluation for aviation programmes should therefore mirror the training model: it should ask whether instruction, feedback, debrief, and assessment actually built the competencies, and it should treat cadet feedback as one strand of a safety-critical evidence system. A single Likert average is not just weak; it is misaligned with how the sector defines quality.

The paradigm shift the evaluation has to catch up with

Competency-Based Training and Assessment (CBTA) is the International Civil Aviation Organization's preferred method for designing pilot training, precisely because it links training to demonstrable performance and supports reliable mutual recognition of licences (IATA, CBTA). EASA has codified competency-based routes into its Part-FCL licensing framework — the Multi-Pilot Licence (MPL), the Airline Pilot Standard Multi-Crew Cooperation course (APS-MCC), and the Competency-Based Instrument Rating (CB-IR) are structured around competencies rather than fixed flight hours (EASA, "Training for Success – Leading the way with CBTA").

As one industry summary puts it, for a training school adopting CBTA "is not a documentation change — it is a methodology change": lesson plans, instructor calibration, debrief structure, grading rubrics, and quality-assurance loops all have to be rebuilt around competencies and observable behaviours. If the training has been rebuilt around competencies but the evaluation of the training still asks generic satisfaction questions, the feedback loop is broken at exactly the point where it should be tightest.

Why generic course evaluation fails an ATO

Aviation training has features that make an averaged satisfaction score close to useless as quality evidence.

The outcome is safety, and the standard is behavioural. ICAO's competency framework describes pilot competencies with associated observable behaviours; an ATO's own quality system is built to evidence that cadets demonstrate them. "How satisfied were you with the course" simply does not map onto that vocabulary. Evaluation should be asking, in the trainee's own words, where instruction and debrief did or did not build a specific competency — situational awareness, workload management, communication, decision-making.

Instructor calibration is a live variable. CBTA depends on instructors and examiners applying grading rubrics consistently. Cadets are often the first to notice when one instructor debriefs against the competencies and another gives an unstructured "that was fine." That is high-value quality intelligence — and it lives in narrative feedback, not in a mean.

Simulator, aircraft, and ground phases are different learning environments. A programme average blends full-flight-simulator sessions, real aircraft sorties, and classroom theory into one number, hiding phase-specific problems. The signal a head of training needs is phase-anchored.

The cohort is small and high-stakes. ATO cohorts are small; a "mean of the satisfaction scores" over a handful of cadets is statistically fragile and easy to over-read. Qualitative depth per cadet is far more defensible than a precise-looking average of twelve responses.

Attrition and confidence are part of the outcome. Ab-initio pilot training is expensive and demanding, and a cadet's growing (or collapsing) confidence in their own decision-making is a signal worth capturing long before a check-ride. Whether a cadet feels the programme is deliberately building their resilience and airmanship — not just their stick-and-rudder skills — is exactly the kind of formative, forward-looking evidence a static end-of-course form is structurally incapable of surfacing in time to act on.

Where Koji fits

Koji for Education aligns course evaluation with the competency paradigm rather than fighting it. Its AI-moderated conversational interviews can be framed around the competencies the programme certifies: when a cadet raises a weakness in, say, threat-and-error management, the interviewer probes for the specific sortie, the debrief they received, and what would have helped — producing evidence a training review can act on, rather than a rating it can only file. The six structured question types (open-ended, scale, single-choice, multiple-choice, ranking, yes/no) let a programme combine a light quantitative backbone with the qualitative depth CBTA quality assurance actually needs.

Because the interviewer is a standardised AI moderator, it does not vary the way human focus-group facilitators do — a useful property in a sector obsessed with calibration and consistency. Automatic thematic analysis lets a head of training see, across a cohort, that debrief quality dips in one simulator phase or with one instructor group — the phase- and instructor-anchored variance a mean conceals. Formative, mid-course collection catches problems while the cohort can still benefit. And EU-appropriate, GDPR-compliant data handling suits organisations operating across European jurisdictions. Koji surfaces and structures the cadet's perspective; it does not certify competence — that remains the examiner's job, and Koji's evidence is designed to feed the ATO's quality loop, not to substitute for it.

But doesn't safety-critical training already have rigorous evaluation?

The strongest objection is that aviation is the last place that needs help evaluating training: ATOs already run compliance-monitoring and safety-management systems, examiners already assess competence, and the regulatory oversight is intense. Isn't cadet course evaluation a soft add-on to a hard system?

Two things. First, formal competence assessment answers "did the cadet meet the standard"; it does not answer "was the training that got them there well designed and consistently delivered." Those are different questions, and the second is exactly where the learner's perspective is a leading indicator — cadets experience instructor inconsistency, thin briefings, or a rushed phase long before it shows up in a check-ride failure. Second, the regulatory system's rigour is a reason to make the feedback rigorous too: soft satisfaction surveys stapled onto a hard safety culture are a mismatch that trained professionals rightly distrust. The honest limitations: cadet feedback is perception, not performance data, and must be triangulated with assessment outcomes, instructor records, and safety data; and small cohorts mean qualitative synthesis, not spurious statistical precision, is the right analytical stance. Used that way — as structured, competency-aligned qualitative evidence feeding an existing quality system — course evaluation strengthens the loop rather than diluting it.

Programmes whose parent organisations also run wider stakeholder or customer research can apply the same conversational interview engine through the main koji.so platform, keeping the method consistent across training and commercial feedback.

Frequently asked questions

What is CBTA and why does it matter for course evaluation?

Competency-Based Training and Assessment is ICAO's preferred approach to pilot training, structuring learning and assessment around demonstrable competencies and observable behaviours rather than flight hours. It matters for evaluation because feedback should be framed around whether the training built those competencies, not around generic satisfaction.

Does EASA require student feedback from Approved Training Organisations?

EASA's framework requires ATOs to operate a compliance-monitoring and quality system. While it does not prescribe a specific student-feedback tool, systematic trainee feedback is a common and expected input to that quality loop, and it must be handled consistently with the organisation's management system.

Why is an averaged satisfaction score a poor fit for pilot training?

Because the sector defines quality behaviourally and by phase, and cohorts are small. A single mean blends simulator, aircraft, and ground phases and different instructors into one fragile number, hiding the phase- and instructor-level variance a head of training must act on.

How does Koji align evaluation with competency-based training?

Koji frames its adaptive interviews around the competencies a programme certifies, probing for the specific sortie or debrief behind a cadet's comment, and uses automatic thematic analysis to localise where instruction or debrief is weak — feeding an ATO's existing quality system with structured qualitative evidence.

Can cadet feedback be trusted given small cohort sizes?

Small cohorts make averaged scores statistically fragile, which is a reason to prefer qualitative depth per cadet over precise-looking means, and to triangulate feedback with assessment outcomes and instructor records rather than treating any single source as definitive.