📄 Research Article
Evaluating Interactivity: Toward Automated Assessment of AI-Generated Explorable Explanations
While LLMs now enable rapid generation of learning materials like Generative AI, evaluating the pedagogical quality of these materials remains an open challenge. This paper proposes an automated assessment framework for evaluating interactivity in AI-generated explorable explanations — dynamic, learner-driven content that students can manipulate to discover concepts. The framework addresses the gap between content generation speed and quality assurance, providing metrics for Formative Assessment of learning designs. This connects to Learning Analytics approaches for understanding how students engage with AI-produced educational content in Higher Ed settings.
Key Findings
Method in Brief
Existing benchmarks give limited insight into dynamic interaction behaviors such as learner-controlled state transitions and context-sensitive system responses — the factors that critically shape learners' conceptual understanding. EE-Eval addresses this by framing interactivity as testable behavioral models rather than an emergent byproduct of LLM generation. The resulting FSM comparison supports pedagogically grounded, actionable Human AI Collaboration in creating interactive educational content.
Implications for AI in Education
For educators and tool builders, EE-Eval offers a diagnostic lens: instead of asking only whether generated content runs correctly, one can ask whether the interaction logic a Generative AI system produced actually serves the intended learning goals. By externalizing interaction logic into an inspectable graph, the framework transforms evaluation into a reflective diagnostic tool for the increasingly common practice of generating Active Learning materials with LLMs, supporting quality assurance at scale.
Connected Concepts
Connected Articles
Citation
Xiaozao Wang, Zhewei Wang, Hongyi Wen (2026). Evaluating Interactivity: Toward Automated Assessment of AI-Generated Explorable Explanations. arXiv:2606.31012.