
Introduction
When final examinations were cancelled for International Baccalaureate students in the United Arab Emirates and parts of the Gulf, schools faced a test that extended far beyond student knowledge.
They faced a test of their assessment systems.
Under the Non Exam Contingency Measure, commonly known as NECM, final results depended heavily on externally assessed coursework, internal assessment evidence, teacher predicted grades, and established quality assurance processes.
The situation created understandable anxiety. Students had prepared for examinations. Families were concerned about university admissions. Teachers questioned how grades would be calculated without the final papers that normally carry significant weight.
Yet the final outcomes revealed an important lesson.
The UAE recorded an average Diploma Programme score of approximately 34.5, significantly above the global average. This did not happen simply because examinations were removed. It reflected the ability of many schools to produce credible academic evidence before the examination period began.
Schools with strong internal assessment practices, accurate predicted grades, consistent marking standards, and well documented student progress were better prepared for disruption.
The NECM experience therefore offers a valuable assessment design lesson for every school, even those that may never face an examination cancellation.
A resilient assessment system should be capable of demonstrating what students know without depending on one final event.
NECM Was More Than an Emergency Grading Process
It is tempting to view NECM only as a temporary response to exceptional circumstances.
That interpretation misses its wider significance.
NECM exposed the strengths and weaknesses of existing school assessment systems. It showed whether teachers had collected reliable evidence throughout the course, whether predicted grades were based on academic judgement, and whether internal assessments genuinely reflected student ability.
In a traditional examination model, weaknesses in school assessment can remain hidden. A final examination may compensate for inconsistent classroom marking, poor tracking, or weak moderation.
Once the examination is removed, the quality of the school’s assessment culture becomes visible.
Schools that maintained strong coursework standards had a reliable evidence base. Schools that treated internal assessments as administrative requirements had far less confidence in their data.
The central lesson is simple.
Assessment resilience is built before a disruption occurs.
It cannot be created during the final weeks of a programme.
Lesson One: Internal Assessments Must Be Authentic
Internal assessments are sometimes treated as projects that students need to complete to satisfy programme requirements.
A strong school treats them differently.
An internal assessment should be a valid demonstration of subject knowledge, analytical thinking, research ability, communication, and independent application.
When the task is authentic, the final submission tells the school something meaningful about the student’s level of achievement.
When the task is over supported, heavily edited, or completed through repeated teacher intervention, the final grade becomes less reliable.
This is especially important when coursework may contribute significantly to a student’s final result.
Schools should ask several questions.
Did the student produce the central ideas independently?
Can the student explain the reasoning behind the work?
Does the quality of the final submission match the student’s performance in class?
Were teacher comments focused on guidance rather than correction?
Was academic integrity checked throughout the process?
A polished submission is not automatically strong evidence. Authenticity matters more than presentation.
The purpose of an internal assessment is not to create perfect work. It is to capture the student’s actual level of understanding.
Lesson Two: Predicted Grades Must Be Evidence Based
Predicted grades are often described as professional judgements. That description is accurate, but professional judgement should not mean personal intuition.
A reliable predicted grade must be supported by multiple forms of evidence.
These may include timed assessments, internal assessment performance, class tests, essays, oral responses, practical work, mock examinations, and patterns of improvement.
No single piece of evidence should determine the final prediction.
A teacher who relies mainly on recent performance may overlook long term patterns. A teacher who focuses only on coursework may overestimate students who perform well with extended preparation but struggle under timed conditions.
The strongest predictions are based on triangulation.
This means comparing different types of evidence and identifying the grade that most consistently represents the student’s current performance.
Schools should also separate aspiration from prediction.
A target grade describes what a student may achieve with further progress. A predicted grade describes the most likely outcome based on current evidence.
Confusing the two can produce inflated predictions, undermine trust, and create problems for students when final results are issued.
Lesson Three: Moderation Cannot Be Optional
Even experienced teachers interpret criteria differently.
One teacher may reward conceptual depth. Another may focus more heavily on structure, accuracy, or technical vocabulary. These differences are natural, but they must be managed.
Moderation creates consistency.
It allows teachers to compare samples, discuss criterion interpretations, challenge assumptions, and agree on common standards.
Schools with a strong moderation culture do not wait until predicted grades are due. They embed moderation throughout the academic year.
This may include reviewing anonymised student work, conducting cross class comparisons, comparing current samples with previous externally moderated work, and discussing cases near grade boundaries.
Effective moderation is not designed to force every teacher to assign identical marks immediately.
Its purpose is to reduce unexplained variation and strengthen the reasoning behind each judgement.
The process should be documented.
A school should be able to explain which work was reviewed, which teachers participated, what differences were identified, and how final decisions were reached.
When disruption occurs, this documentation becomes evidence that the school’s grades are credible.
Lesson Four: Assessment Evidence Must Be Collected Over Time
A resilient assessment system does not depend on one mock examination conducted near the end of the course.
It builds a portfolio of evidence across several months.
This protects students from the consequences of one poor day. It also prevents one unusually strong performance from creating an inaccurate picture.
Evidence collected over time reveals patterns.
It shows whether students can retain knowledge, transfer concepts, respond to feedback, manage different task formats, and improve their performance independently.
Schools should review the balance of their assessment calendars.
Too many tasks in the final term create pressure for students and produce rushed evidence. Too few formal tasks during the earlier stages make accurate prediction difficult.
A well designed calendar distributes meaningful assessment across the programme.
Each task should have a clear purpose.
Some tasks should diagnose gaps. Some should provide practice. Others should contribute to formal grade evidence.
Not every classroom activity needs to produce a mark. However, every reported grade should come from an assessment that measures something important.
Lesson Five: Coursework Standards Must Be Consistent
Strong results under NECM did not prove that examinations are unnecessary.
They showed that coursework can carry greater responsibility when it is designed and evaluated carefully.
For coursework to become reliable evidence, schools need consistent expectations across subjects, classes, and year groups.
Students should know what independent work looks like. Teachers should understand how much support is appropriate. Departments should follow common rules for drafts, feedback, deadlines, extensions, and academic integrity.
Inconsistency creates unfairness.
One student may receive extensive feedback across several drafts, while another receives only general guidance. One department may strictly enforce deadlines, while another allows unlimited extensions. One teacher may correct language and structure, while another avoids direct intervention.
These differences affect outcomes.
A school wide coursework policy should define the boundaries clearly.
It should explain what teachers may comment on, how many feedback opportunities are permitted, how missed deadlines are managed, and how authenticity is verified.
Consistency does not mean every subject must use the same process. Subject differences matter.
However, the principles of fairness, independence, transparency, and academic integrity should remain common.
Lesson Six: Assessment Data Should Support Action
Many schools collect large amounts of assessment data without using it effectively.
Grades are entered into systems, reports are generated, and averages are discussed. Yet little changes in the classroom.
A strong assessment system connects data with intervention.
When a student’s performance declines, teachers should know why. The cause may be conceptual misunderstanding, weak examination technique, inconsistent attendance, poor time management, language difficulty, or excessive dependence on teacher support.
Different problems require different responses.
A general instruction to work harder is rarely useful.
Schools should create regular opportunities for teachers to examine student evidence and decide what action follows.
These discussions should focus on questions such as:
-
Which students are performing below their established pattern?
-
Which assessment criteria are producing the weakest results?
-
Are predicted grades supported by recent evidence?
-
Which students show a major gap between coursework and timed performance?
-
What intervention will be provided before the next assessment?
Assessment data becomes valuable only when it changes teaching, feedback, support, or student behaviour.
How to Audit Your School’s Assessment Design
The NECM experience provides a practical framework for evaluating assessment resilience.
A school does not need to wait for an external disruption to conduct this review.
The following audit can be completed by programme leaders, coordinators, department heads, and senior leadership teams.
Audit Area One: Quality of Evidence
Begin by identifying the evidence used to determine student achievement.
For each subject, list the assessments that contribute to predicted grades and internal reporting.
Then evaluate whether the evidence is sufficiently varied.
A strong evidence base should include a combination of timed work, extended coursework, application tasks, subject specific assessments, and internally moderated samples.
If a predicted grade depends mainly on one mock examination or one major project, the system is vulnerable.
The school should also examine whether each task reflects the demands of the final qualification.
An assessment may be engaging and educational without producing valid grade evidence. Validity depends on whether the task measures the skills and knowledge represented by the official criteria.
Audit Area Two: Prediction Accuracy
Compare predicted grades with final grades from previous cohorts.
The purpose is not to demand perfect accuracy. Student performance can change, and external assessment always introduces uncertainty.
The aim is to identify patterns.
Does one subject consistently predict too generously?
Does another department repeatedly underestimate students?
Are certain teachers more accurate than others?
Do predicted grades become inflated during university application periods?
Schools should investigate patterns rather than blame individuals.
Persistent over prediction may indicate weak standardisation, pressure from families, confusion between target and predicted grades, or an overreliance on supported coursework.
Persistent under prediction may reflect excessive caution or poor recognition of student progress.
Prediction accuracy should become a professional learning tool.
Audit Area Three: Moderation Quality
Review how moderation operates within each department.
Ask whether teachers moderate only final internal assessments or whether they also review regular classroom assessments.
Check whether departments use common task specific expectations and whether teachers discuss grade boundary cases.
Moderation should include more than agreement.
A group of teachers can agree on an inaccurate standard. Departments should compare their work with externally marked samples, official subject reports, criterion guidance, and previous moderation feedback.
The school should also confirm that moderation occurs before grades are reported, not after decisions have already been communicated.
Audit Area Four: Coursework Integrity
Review the full student journey from task introduction to final submission.
Identify how teachers record feedback, supervise research, verify authenticity, and monitor drafting.
The school should look for signs of excessive support.
These may include final submissions that differ sharply from classroom performance, sophisticated work that students cannot explain orally, or similar structures appearing across several submissions.
Artificial intelligence has made this area even more important.
Schools need clear expectations for acceptable assistance, source acknowledgement, idea generation, language correction, and independent authorship.
The objective is not simply to catch misconduct.
It is to design processes that make authentic student thinking visible.
Audit Area Five: Assessment Calendar
Map every major assessment across the academic year.
Look for periods of overload, long gaps without meaningful evidence, and excessive concentration near reporting deadlines.
A balanced calendar should protect both assessment quality and student wellbeing.
When several subjects schedule major tasks at the same time, students may submit work that reflects exhaustion rather than ability.
A coordinated calendar also gives teachers enough time to provide useful feedback, conduct moderation, and plan intervention.
Assessment scheduling is therefore an academic quality issue, not merely an administrative task.
Audit Area Six: Student Assessment Literacy
Students perform more reliably when they understand how assessment works.
They should know what the criteria mean, how different tasks contribute to judgement, what evidence supports a predicted grade, and how feedback should be used.
Students should also be able to evaluate their own work.
A school with strong assessment literacy does not keep standards hidden until marking is complete. Teachers use annotated samples, criterion discussions, reflection activities, and guided self assessment to make quality visible.
This does not mean teaching students to imitate model answers.
It means helping them recognise the features of strong reasoning, communication, analysis, and subject understanding.
Building an Assessment System That Survives Disruption
The most important lesson from NECM is not that coursework should replace examinations.
It is that no school should rely entirely on one assessment event to understand student achievement.
A resilient system combines several elements.
It gathers evidence over time. It protects authenticity. It trains teachers to apply standards consistently. It reviews prediction accuracy. It uses moderation as professional learning. It turns assessment data into action.
Such a system benefits schools even when examinations proceed normally.
Students receive earlier support. Teachers make more confident decisions. Families receive more accurate reports. Leaders gain a clearer picture of teaching quality. Predicted grades become more trustworthy.
Most importantly, assessment becomes part of learning rather than a judgement added at the end.
Conclusion
The NECM experience revealed a truth that applies far beyond the UAE and Gulf region.
Assessment resilience is not created by an emergency policy. It is created through the everyday practices of teachers, departments, coordinators, and school leaders.
Schools that produced strong, credible outcomes without final examinations were not simply fortunate. They had systems capable of showing what students knew through internal assessments, coursework, predicted grades, and accumulated academic evidence.
The lesson for other schools is not to prepare for cancelled examinations.
It is to build an assessment culture that remains credible under any condition.
Every school should be able to answer three questions with confidence.
What evidence supports each student’s grade?
How do we know that the evidence is authentic and consistently judged?
Would our assessment decisions remain defensible if the final examination disappeared?
If the answer to any of these questions is uncertain, the assessment system requires attention.
The strongest schools do not wait for disruption to expose weaknesses.
They audit, moderate, document, and improve before the system is tested.
