Efficacy studies โ controlled evaluations of AI education tools โ are the wiki's evidence backbone: genai-policy-prompting-rct, access-not-enough-ai-tutoring-2026, genai-can-harm-teaching-rct-2026, and lets-chat-chatbot-outreach-2026 are pre-registered RCTs spanning K-12 and higher education (RCT, learning-gains).
Related Pages
๐ 46 other pages tagged efficacy-study
- A meta-analysis of the effect of generative AI on productivity and learning in programming
- A review of intervention designs of LLM Integration in Undergraduate Computer Science Education
- AI Assistance for Discretionary Work: Increasing Feedback Provision in Higher Education
- AI Assistance Reduces Persistence and Hurts Independent Performance
- AI in K-12 Evidence Base
- AI Learning Transfer
- AI Tutor Effectiveness Review
- AI-Driven Assessment of Human Tutors: Linking Training Performance to Real-Life Practice
- An Interpretable Closed-Loop Intelligent Tutoring System for Multimodal Affective Feedback in Asynchronous Presentation Training
- Artificial intelligence in vocational education and training: A systematic review of educational purposes, theoretical conceptualizations, and empirical effectiveness
- Assessing the Impact and Underlying Pathways of Sequenced AI Feedback on Student Learning
- Automated Grading of Handwritten Mathematics Using Vision-Capable LLMs
- Benchmark
- Cognitive offloading and the speedup illusion in human-AI interaction
- Combating Harms of Generative AI in CS1 with Code Review Interviews and a Flipped Classroom
- Confidence-Aware Automated Assessment of Student-Drawn Scientific Models
- Cross-Subject Predictive Validity for Learning Outcomes of Delayed Start Behavior
- Designing a mobile chatbot-based learning journaling system for intrinsic motivation and engagement
- Educational LLM Alignment
- Effects of an AI-supported inquiry model on AI literacy and authentic performance: A quasi-experimental study with preservice teachers
- Evidence of a Cognitive Shift in AI Education: How Students Are Rethinking Human Intelligence?
- Experiential Versus Instructional Approaches for Eliciting Metacognitive Awareness in AI-Assisted Learning
- Explaining Too Much? Understanding How Large Language Model Reasoning Traces Influence Performance and Metacognition
- From Heuristics to Analytics: Forecasting Effort and Progress in Online Learning
- Generative AI Availability, Grades, and Student Satisfaction at a Large University
- How AI Is Changing Teaching Workflows
- Improving Hybrid Human-AI Tutoring by Differentiating Human Tutor Roles Based on Student Needs
- Is Solving Better Than Evaluating GenAI Solutions?
- LaTA: A Drop-in, FERPA-Compliant Local-LLM Autograder for Upper-Division STEM Coursework
- Learning Analytics
- Little Impact of ChatGPT Availability on High School Student Test Score Performance
- Modernizing Ground Truth: Four Shifts Toward Improving Reliability and Validity in AI in Education
- NSMQ Riddles: A Benchmark of Scientific and Mathematical Riddles for Quizzing Large Language Models
- Position: Adopting AI in Practice Does Not Guarantee the Productivity Boost
- Quality-Conditioned Agreement in Automated Short Answer Scoring: Mid-Range Degradation and the Impact of Task-Specific Adaptation
- REC-CBM: Rubric-Aware Error-Correction Concept Bottleneck Models for Trustworthy Open-Ended Grading
- Rethinking Scaffolding in LLM Tutors: The Interactional Mismatch Between Benchmarks and Real-World Deployments
- Self-Efficacy and Favorability Shape Learning from Tutoring Systems and Paper Practice
- Simulating Students' Java Programming Errors with Large Language Models
- Structured AI Demonstrations and Student LLM Use in Engineering Mechanics: Study Design and Preliminary Results
- The Effects of Structured LLM-Generated Feedback on Programming Assignment Performance
- The Environmental Cost of LLMs in AIED: Reporting and Practices
- The Illusion of Competence: Self-Perceived Digital Literacy and AI Readiness Among European Secondary Students
- The Main Barrier to AI Adoption in the Public Sector is Lack of Training
- The Missing Evaluation Axis: What 10,000 Student Submissions Reveal About AI Tutor Effectiveness
- The Tutoring Effectiveness Index: Predicting LLM Math Tutor Quality from Four Conversation Signals