Defence beats detection
Classrooms that added a short live defence saw the steepest fall in integrity flags. No AI-detection software was used anywhere in the study, and none was needed.
2026 Edition · Published 5 May 2026
What happened to integrity flags, teacher marking time and student reasoning scores when 240 partner classrooms moved from written submissions to defence-based assessment.
240 classrooms · 3 countries · 2 academic terms · 2025 to 2026
Key finding
Across 240 NASCA partner classrooms that replaced take-home written submissions with defence-based assessment, integrity flags fell from 23 percent of submissions to 4 percent, while measured reasoning scores rose 17 percent over two terms. Marking time per class fell by roughly a fifth once the rubric was in steady use.
The study covers two academic terms across primary, middle and senior classrooms in India, the UAE and the United States. The rubric used is published openly at /assessment/rubric.
| Stage | Classrooms | Flags before (%) | Flags after (%) | Reasoning before | Reasoning after |
|---|---|---|---|---|---|
| Primary (Grades 1-5) | 74 | 11 | 3 | 1.9 | 2.3 |
| Middle (Grades 6-8) | 88 | 24 | 4 | 2.1 | 2.5 |
| Senior (Grades 9-12) | 78 | 33 | 6 | 2.4 | 2.8 |
| All stages | 240 | 23 | 4 | 2.1 | 2.5 |
| Redesign move | Share of classrooms ranking it first (%) |
|---|---|
| Live defence of the submitted work | 34 |
| Process evidence graded alongside the artefact | 26 |
| Task anchored to local, un-Googleable context | 18 |
| In-class checkpoint before submission | 14 |
| Peer critique round | 8 |
Every table on this page is available as a single CSV file: download the dataset.
Classrooms that added a short live defence saw the steepest fall in integrity flags. No AI-detection software was used anywhere in the study, and none was needed.
Term one marking time rose slightly as teachers learned the rubric. From term two, marking time per class was about a fifth lower than the written-submission baseline.
Senior classrooms started with the highest flag rate and posted the largest reasoning gain, suggesting the redesign converts avoidance behaviour into visible thinking.
In the NASCA 240-classroom study, teacher-raised integrity flags fell from 23 percent of submissions to 4 percent after moving to defence-based assessment, without using any AI-detection software.
Only at first. Marking time rose slightly in term one while teachers learned the rubric, then settled about 21 percent below the written-submission baseline from term two onward.
A short live defence of the submitted work. 34 percent of classrooms ranked it the single biggest contributor, ahead of grading process evidence at 26 percent.
NASCA Research Desk with the World STEM Federation (2026). AI-Resilient Assessment: Evidence from 240 Classrooms. NASCA. https://www.nasca.edu.in/research/reports/ai-resilient-assessment-2026
Licensed CC BY 4.0. Quote the figures freely, with attribution and a link back to this page. Media and researchers may request the raw tabulation and the survey instrument.
We share the instrument, the cleaned dataset and an interview with the research desk on request.