Is Your Course Actually Aligned? Using Biggs's Constructive Alignment as an Evaluation Lens
Biggs's constructive alignment says learning outcomes, teaching activities, and assessment must point the same way. Here is how to turn that theory into a course-evaluation instrument that diagnoses misalignment students can feel but rarely name.
Koji Education Team
Product
The short answer
A satisfaction score tells you whether students enjoyed a course; it does not tell you whether the course was coherent. Constructive alignment — John Biggs's proposition that intended learning outcomes, teaching and learning activities, and assessment tasks must all target the same verbs — gives evaluation a structural question that a Likert average cannot: did the parts of this course actually point in the same direction? Used as an evaluation lens, constructive alignment reframes "students were unhappy with assessment" into a diagnosable claim: the assessment measured something the teaching never developed. That is a finding a programme team can act on.
What the research says
Biggs introduced constructive alignment in a 1996 paper in Higher Education ("Enhancing teaching through constructive alignment"). The idea marries two traditions: constructivism (students construct meaning through what they do, not through what is transmitted to them) and instructional design (alignment of objectives, methods, and assessment). The core move is deceptively simple. You specify intended learning outcomes (ILOs) as verbs — analyse, design, justify, apply — and then you engineer teaching activities that require students to practise exactly those verbs, and assessment tasks that require students to demonstrate them. When the verbs match across all three, the system is aligned; the assessment "traps" the intended learning rather than rewarding a proxy for it.
Biggs's contention, developed further in Biggs and Tang's Teaching for Quality Learning at University, is that alignment does much of the motivational work we usually attribute to the teacher's charisma. In a misaligned course, the rational student studies the assessment, not the outcomes — because the assessment is what is rewarded. If the ILOs promise "critically evaluate competing models" but the exam rewards recall, students correctly infer that recall is the real game. Biggs frames this as students being "entrapped" by the assessment either into surface learning (memorising) or deep learning (engaging with meaning), depending on how the task is built. Alignment is the mechanism that makes deep engagement the path of least resistance.
The framework has become one of the most cited concepts in higher-education pedagogy and underpins outcomes-based quality assurance across the European Higher Education Area — the Bologna "learning outcomes" architecture is, in effect, constructive alignment operationalised at system scale. But it is not uncontested. Loughlin, Lygo-Baker and Lindberg-Sand's 2021 paper in the European Journal of Higher Education ("Reclaiming constructive alignment") argues the concept has been flattened in practice into a bureaucratic template — a box-ticking exercise where outcomes are written to match pre-existing assessments rather than to drive design. They insist constructive alignment "is neither the panacea, nor the unalloyed evil" of the literature, but a heuristic that only works when it genuinely informs design decisions rather than retrofitting them. That critique matters for evaluation: it means you cannot verify alignment from the paperwork alone. You have to ask the people who experienced the course whether it felt aligned.
Why it matters for course evaluation in practice
Most course-evaluation instruments ask students to rate teaching quality, clarity, workload, and overall satisfaction. None of these directly surfaces misalignment — yet misalignment is one of the most common and most fixable sources of student dissatisfaction. When students write "the exam had nothing to do with the lectures" or "we were never taught how to do the assignment," they are reporting a broken alignment link, not a bad teacher. A well-designed evaluation converts that diffuse frustration into a specific structural diagnosis.
Concretely, an alignment-aware evaluation probes three links separately:
- Outcome-to-teaching: did the teaching and learning activities give you genuine practice in the things the course said you would learn to do?
- Teaching-to-assessment: did the assessment ask you to demonstrate the things you actually practised?
- Outcome-to-assessment: did the assessment reward the stated outcomes — or did it reward something else (speed, memory, guessing the marker)?
A course can score well on general teaching quality and still fail one of these links badly. Surfacing which link failed is far more useful to a programme director than an aggregate 3.8/5, because each broken link implies a different remedy: rewrite the activities, redesign the assessment, or rewrite the outcomes so they stop over-promising.
Limitations and honest caveats
A PhD reader will raise several objections, and they are right to.
Students are not the authority on alignment. Students experience the consequences of misalignment but may misattribute them. A student who found an assessment hard because it demanded genuine higher-order thinking (a desirable difficulty) may report it as "unfair" or "unaligned" when the course was, in fact, working as designed. Constructive alignment predicts this: aligned deep-learning tasks are often less comfortable than surface ones. So alignment evidence from students must be triangulated with an expert review of the ILO-assessment mapping, never taken at face value.
Alignment is necessary, not sufficient. A perfectly aligned course built around trivial outcomes is still a weak course. Alignment tells you the machine is internally consistent; it says nothing about whether the outcomes are worth having. Goal-free and connoisseurship approaches to evaluation exist precisely to catch what alignment misses.
The "reclaiming" critique applies to evaluation too. If you reduce alignment to a checklist ("Does each ILO map to an assessment? Y/N"), you get the same hollow compliance Loughlin and colleagues warn about. The evaluative value is in the qualitative gap between the documented map and the lived experience — which means you need rich open-text data, not just a mapping matrix.
Generalisability. Constructive alignment was developed largely in on-campus, outcomes-based Western higher education. Its salience varies across disciplines (it fits competency-based professional programmes better than open-ended humanities seminars) and across cultures where the outcomes-assessment relationship is understood differently. Treat it as a lens that fits many courses well, not a universal law.
How Koji incorporates this
Koji is built to collect exactly the kind of structured-plus-narrative evidence that alignment evaluation needs, and to keep the three alignment links distinct rather than collapsing them into one satisfaction number.
- Link-specific structured questions. Instead of a single "assessment quality" item, a Koji evaluation can pose three separate
scalequestions — one per alignment link — so misalignment localises to the outcome-to-teaching, teaching-to-assessment, or outcome-to-assessment join. The reporting keeps them disaggregated. - AI-moderated conversational probing. When a student rates the teaching-to-assessment link low, Koji's AI moderator follows up in the moment — "Can you give an example of something the assessment asked for that the teaching didn't prepare you for?" — turning a bare rating into an actionable, example-grounded account of the specific gap. This is the diagnostic depth a fixed questionnaire cannot reach, and it is designed to distinguish genuine misalignment from a desirable difficulty that merely felt hard.
- Automatic thematic analysis clusters open-text responses so that recurring alignment complaints ("the group project wasn't marked on anything we did in seminars") surface as a named theme with prevalence, rather than being lost in a comment dump.
- Triangulation across cohorts and instruments. Because alignment evidence from students must be checked against an expert ILO-assessment map, Koji supports pairing the student-facing evaluation with a staff or peer instrument, so the programme team can compare the documented alignment with the experienced alignment — the exact gap Loughlin and colleagues say matters most.
- Closing-the-loop action tracking records which alignment link a team decided to repair and lets the next cycle test whether the fix worked.
Koji is designed to mitigate the flattening critique, not to certify alignment: it produces the rich qualitative signal that tells you whether a documented map is real, without ever claiming a good score proves the course is well-designed.
The same AI-moderated interview engine powers Koji's core research platform at koji.so for product and customer research — the education product simply points those conversational probes at teaching, learning, and assessment.
Related Resources
- Is Your Course Pushing Students Toward Deep or Surface Learning? The R-SPQ-2F as an Evaluation Lens
- Course Evaluations Measure Reaction, Not Learning: What Kirkpatrick's Four Levels Reveal
- Validity Is About the Use, Not the Instrument: Applying Kane's Argument-Based Framework to Course Evaluation
- Evaluating Active Learning: The ICAP Framework as a Course-Evaluation Lens
- Teacher Clarity Predicts Learning Better Than Charisma: What Course Evaluations Should Measure
- Desirable Difficulties: Why the Teaching That Improves Learning Often Lowers Satisfaction
References
- Biggs, J. (1996). Enhancing teaching through constructive alignment. Higher Education, 32(3), 347–364. https://doi.org/10.1007/BF00138871
- Biggs, J., & Tang, C. (2011). Teaching for Quality Learning at University (4th ed.). Open University Press / SRHE.
- Loughlin, C., Lygo-Baker, S., & Lindberg-Sand, Å. (2021). Reclaiming constructive alignment. European Journal of Higher Education, 11(2), 119–136. https://doi.org/10.1080/21568235.2020.1816197
Related articles
Validity Is About the Use, Not the Instrument: Applying Kane's Argument-Based Framework to Course Evaluation
Asking whether course evaluations are valid is the wrong question. Kane's argument-based framework asks whether a specific interpretation and use of the scores is justified. We rebuild the SET debate as an interpretation-use argument, expose where each inference breaks, and show how Koji strengthens the weak links.
Teacher Clarity Predicts Learning Better Than Charisma: What Course Evaluations Should Measure
A meta-analysis of 144 effects and 73,000+ students shows teacher clarity explains roughly 13% of the variance in student learning. Here is what that means for the items you put on a course evaluation.
Course Evaluations Measure Reaction, Not Learning: What Kirkpatrick's Four Levels Reveal
Kirkpatrick's four-level model explains why an end-of-term course evaluation is a Level-1 'reaction' measure — and why decades of meta-analytic evidence show reaction correlates almost nothing with actual learning.
Is Your Course Pushing Students Toward Deep or Surface Learning? The R-SPQ-2F as an Evaluation Lens
Most course evaluations ask whether students liked the teaching. The deep/surface approaches tradition asks a more consequential question: did the course lead students to engage meaningfully or just memorise to pass? The R-SPQ-2F instrument makes that measurable.