🧠 AI Ed Wiki

🏷️ reinforcement-learning

9 pages tagged with reinforcement-learning(8 articles, 1 concepts)

📄 ResidencyRL: Reinforcement Learning in Simulated Clinical Environments
> **Synthesis:** Liévin et al. (2026) present **ResidencyRL**, a reinforcement learning method for training clinical AI agents through simulated multi-turn clinical encounters (up to 60 dialogue turns…
2026-08-13 · simulation, health-education, llm, professional-training, dialogue-tutoring
🏷️ Simulation
> **Simulation** — the use of modeled environments, agents, or scenarios to support learning through practice and feedback in contexts that are safe, repeatable, and often otherwise inaccessible. Simu…
2026-08-12 · active-learning, adaptive-learning, pedagogical-agent, skill-development, experiential-learning
📄 EduQwen: Pedagogical RL
> **EduQwen: Pedagogical RL** — A multi-stage optimization strategy combining reinforcement learning (DAPO) and supervised fine-tuning (SFT) to enhance the pedagogical knowledge of open-source LLMs, p…
2026-07-29 · llm, pedagogical-safety, pedagogical-llm-training, open-source, rag
📄 Representation Robustness under Executable Reasoning Constraints in Large Language Models for Mathematical Problem Solving
This study probes how sensitive [[llm]] mathematical problem solving is to the surface representation of an item — a question with direct bearing on [[assessment-validity]] when LLMs are used for scor…
2026-07-24 · llm, stem-education, benchmark, assessment-validity, rag
📄 Q-Learning Lab: Teaching Reinforcement Learning Through Learner-Generated Trace Analysis
> Presents Q-Learning Lab, a single-file tool that makes the Bellman update concrete by letting undergraduates inspect how each value is computed and why actions are chosen, through learner-generated …
2026-07-14 · active-learning, higher-ed, stem-education, self-regulated-learning, scaffolding
📄 Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues
A key part of developing large language model (LLM)-powered, automated tutoring tools is student simulation, i.e., using LLMs to role-play as students, which can facilitate tutor model evaluation and …
2026-05-29 · intelligent-tutoring, llm, student-experience, learning-analytics, personalized-learning
📄 Pedagogical Safety in Educational Reinforcement Learning
> As reinforcement learning personalizes instruction in intelligent tutoring systems, there is no formal framework for pedagogical safety — a critical gap. > First formal framework for defining and de…
2026-05-08 · intelligent-tutoring, pedagogical-safety, adaptive-learning, adaptive-learning-systems, metacognition

Related Tags

llm (8)rag (4)intelligent-tutoring (4)active-learning (2)adaptive-learning (2)pedagogical-safety (2)stem-education (2)benchmark (2)scaffolding (2)personalized-learning (2)simulation (1)health-education (1)professional-training (1)dialogue-tutoring (1)trust-calibration (1)