Detection-led responses face well-documented limits: validity and fairness failures (bias against non-native writers), notable error rates, erosion of trust, and distraction from assessment design. Detection should be a limited, situational tool β not a strategy of first resort. The constructive question is not "how do we prevent students from using AI?" but "how do we enable them to use it th
Beyond Detection: authentic assessment in an AI-mediated world
Kickbusch, Ashford-Rowe, Kemp, Boreland & Huijser (2025) argue the dominant institutional response to generative AI in assessment β surveillance and AI detection β misdiagnoses the problem: in an AI-mediated world, authenticity cannot be policed into existence; it must be redesigned. They reconceptualise authenticity as constructed where AI is expected, declared, and scrutinised, and offer discipline-agnostic "design for learning" patterns that position AI as a collaborator rather than a cheating application.
The case against detection
Detection-led responses face well-documented limits: validity and fairness failures (bias against non-native writers), notable error rates, erosion of trust, and distraction from assessment design. Detection should be a limited, situational tool β not a strategy of first resort. The constructive question is not "how do we prevent students from using AI?" but "how do we enable them to use it thoughtfully, responsibly, and effectively in contexts that mirror their future work?" Excluding AI from assessment creates an inauthentic scenario: the authentic professional justifies when, how, and why they use tools, and critically evaluates their outputs.
Authenticity as a four-dimensional continuum
Not a binary but a continuum across four intersecting dimensions:
1. Taskβcontext alignment with contemporary professional practice β judgement, decision-making, and problem-solving under uncertainty, not superficial workplace replication
2. Foregrounding professional judgement and ethics β sustainable assessment (Boud & Soler 2016), collaboration (Boud & Bearman 2024), UNESCO 2023 capability framing
3. Visibility of process β iteration, critique, rationale; polished outputs can mask superficial understanding, so assessment must reveal the "messiness" of authentic professional work
4. Appropriate use of tools (including AI) within human decision-making β tools as enablers of higher-order capability, not substitutes for it
Stage-appropriate authenticity: early units get constrained, well-scaffolded tasks; later units open complexity, uncertainty, and stakeholder engagement.
Design patterns ("design for learning moves")
Critique, adapt, verify AI outputs: business students interrogate a chatbot-generated market analysis; pre-service teachers evaluate AI-produced lesson plans for inclusivity and pedagogical soundness; journalism students edit an AI news brief to identify bias; health students appraise AI diagnostic recommendationsProcess transparency artefacts: process logs, AI prompt records, drafts showing iterations β "behind the scenes" evidence submitted alongside the final outputReflective commentaries: explain key decisions, justify tool use, account for changes, with explicit criteria for depth, criticality, and ethical awarenessOral defences / annotated portfolios / recorded walkthroughs: probe reasoning in real time, mirroring professional practices like pitching and peer reviewSelf-critique and peer feedback for feedback literacy (Boud & Molloy 2012)Progressive release across a programme: transparency artefacts + short defences β collaboration and negotiated briefs β capstones with external stakeholders and negotiated criteriaChallenges and institutional responsibilities
Equity: unequal access to tools deepens divides; institutional provision (fenced AI deployments) reduces back-channel inequality; authentic formats can create new barriers (workload, carer/employment constraints) β mitigate with workload modelling, staged scaffolding, modality choiceEthics and bias: tools reproduce cultural stereotypes and can be fluent yet unfaithful (Bender et al. 2021); institutions should run privacy/data-protection impact assessments (PIA/DPIA) for assessment AI, vet tools against privacy/bias/accessibility criteria, and standardise prompt-log conventions that evidence process without exposing personal dataLoad and feasibility: process artefacts and defences raise workload; needs modelling and scaffoldsStaff development: design-led collaboration rather than superficial tool trainingConnected Concepts
AI LiteracyAssessment ValidityMetacognitionSelf Regulated LearningGenerative AIHigher EdConnected Articles
Authentic Assessment β Authentic AssessmentAuthentic Products Authenticated Processes 2026 β From authentic products to authenticated processes: authentic assessment in AI-rich higher educationCare Full Feedback GenAI β The care-full craft of feedback in an age of generative AITool Invariant Framework Agentic AI β A Tool-Invariant Framework for Teaching and Assessing Computational Methods in the Age of Agentic AIA4l Analytics Pipeline β Generalizing a Highly Configurable Analytics Pipeline to Replicate and Support Educational Research Across Multiple D...Aaai2026 Prompting Literacy K12 β Learning to Use AI for Learning: Teaching Responsible Use of AI Chatbot to K-12 Students Through an AI Literacy ModuleAcademiclaw Student Agent Benchmark β AcademiClaw: When Students Set Challenges for AI AgentsAccess Not Enough AI Tutoring 2026 β Access is Not Enough: Human Support Improves Engagement with AI TutoringAdapt Adaptive Lesson Plan Transformer β AdaPT: Adaptive Lesson Plan Transformer for Cross-Regional and Differentiated InstructionAdaptive Pretesting Retention β Do Gains from Generative AI-Enabled Adaptive Pretesting Persist? Evidence from a Retention StudyAffective Text Wearable Student Health β A Formative Study of Brief Affective Text as a Complement to Wearable Sensing for Longitudinal Student Health MonitoringAgency Gap AI Writing β The agency gap in AI-supported writing: how reactive and proactive agent designs shape multimodal reasoningAgent Voice Accents K12 Group Learning β Exploring How Agent Voice Accents Shape Human-AI Collaboration in K-12 Group LearningAgentic AI Education Scoping Review β Agentic AI in Education: A Scoping Review of Research Landscape, Capabilities, and the Frontier Agent ParadigmAgentic AI Pedagogical Best Practice 2026 β Agentic AI and Pedagogical Best Practice: The Tension Between Automation and LearningAgentic Education Coding β Agentic Education with AI Coding AssistantsAgentic Literacy Debt β Agentic Literacy Debt: A Structural Problem the AI Literacy Field Has Not Yet NamedAgents That Teach Incidental Learning β Agents That Teach: Designing Incidental Learning Back into AI-Assisted Software DevelopmentAI Adoption Training Public Sector β The Main Barrier to AI Adoption in the Public Sector is Lack of TrainingAI Adult Learning Guidelines Dis2026 β Guidelines for Designing AI Technologies to Support Adult LearningAI Agents Constructive Conflict Design Education 2026 β Enacting Constructive Conflicts with AI Agents to Enhance Reconsideration among Novice Interaction DesignersAI Assessment Scale Reform β A bit of chaos and madness": The AI Assessment Scale and the work of assessment reformAI Assistance Discretionary Feedback β AI Assistance for Discretionary Work: Increasing Feedback Provision in Higher EducationAI Assisted Learning Modes Eeg β An exploratory behavioral and electroencephalographic study of artificial intelligence-assisted learning modes in hig...AI Assisted Se Curriculum Syllabus Analysis 2026 β Mapping the Emerging Curriculum for AI-Assisted Software Engineering via Syllabus AnalysisCitation
Kickbusch, S., Ashford-Rowe, K., Kemp, A., Boreland, J., & Huijser, H. (2025). Beyond Detection: Redesigning Authentic Assessment in an AI-Mediated World. Education Sciences, 15(11), 1537. DOI