🧠 AI Ed Wiki

Assessment designed to inform ongoing instruction and learning, as opposed to summative evaluation. AI systems can generate, validate, and adapt formative assessment items at scale, though quality varies dramatically across assessment types.

AI-Generated Formative Items

Multiple-Choice Questions (CODE-GEN)

Duan et al. (2026) demonstrate that agentic AI can reliably generate MCQs for coding comprehension when validated across seven pedagogical dimensions. Success rates reach 98.6% for concept alignment and 79.9% for feedback quality—suggesting that AI is strongest on verifiable dimensions and weakest on instructional-judgment dimensions.

Automated Essay Scoring (MASS)

Kamalov et al. (2026) implement a multi-agent framework (MASS) for essay scoring. Preliminary results show improved consistency over stand-alone LLMs, though interpretability of multi-agent scoring decisions remains an open challenge.

Curriculum-Grounded Feedback (LearnLens)

Zhao et al. (2025) present LearnLens, a modular LLM system for science education feedback that addresses three persistent problems in AI formative assessment:

1. Error-aware assessment — captures nuanced reasoning errors rather than surface mistakes

2. Topic-linked memory chains — replaces noisy similarity-based RAG with structured curriculum-grounded retrieval

3. Educator-in-the-loop — teacher customisation and oversight, not full automation

Key differentiator: LearnLens uses a structured, topic-linked memory chain rather than traditional RAG similarity search, improving relevance and reducing noise. This connects to the broader tension in Human In The Loop AI: scalable automation with expert validation.

Design Trade-offs

DimensionAI SuitabilityHuman Requirement
Factual correctnessHighLow
Concept alignmentHighMedium
Distractor qualityLowHigh
Feedback depthLowHigh
Rubric consistencyMediumMedium

Risk: Assessment as Surveillance

Formative assessment systems can shift from learning-support tools to behavior-monitoring infrastructure. The same data streams that enable adaptive tutoring can enable punitive tracking if governance is weak.

Connected Concepts

  • AI Literacy
  • Higher Ed
  • Automated Grading
  • Student Experience
  • Scaffolding
  • STEM Education
  • Personalized Learning
  • Teacher Role
  • Adaptive Learning
  • Feedback Loop
  • Generative AI
  • Intelligent Tutoring
  • Connected Articles

  • AI Changing Teaching Workflows
  • AI Coaching RL Skill Development
  • AI Generated Feedback Higher Ed
  • AI Learning Tools Engineering Education Needs
  • Assessment Team Problem Solving Computing Education
  • Authentic Assessment
  • Automated Formative Assessments A Level Sciences
  • Automated Grading Linux Bash Examinations Large Language Models
  • Becerra Aicofe Feedback 2026
  • Buggy GenAI Code Student Responses
  • Code Review GenAI Cs1
  • Cognitive Offloading LLM Synthesis Writing
  • Correct Answer Trap AI Tutor
  • Correct Answer Trap Misconceptions
  • Critical Engagement Code Completion