🏷️ reinforcement-learning
9 pages tagged with reinforcement-learning(8 articles, 1 concepts)
📄 ResidencyRL: Reinforcement Learning in Simulated Clinical Environments
> **Synthesis:** Liévin et al. (2026) present **ResidencyRL**, a reinforcement learning method for training clinical AI agents through simulated multi-turn clinical encounters (up to 60 dialogue turns…
🏷️ Simulation
> **Simulation** — the use of modeled environments, agents, or scenarios to support learning through practice and feedback in contexts that are safe, repeatable, and often otherwise inaccessible. Simu…
📄 EduQwen: Pedagogical RL
> **EduQwen: Pedagogical RL** — A multi-stage optimization strategy combining reinforcement learning (DAPO) and supervised fine-tuning (SFT) to enhance the pedagogical knowledge of open-source LLMs, p…
📄 Representation Robustness under Executable Reasoning Constraints in Large Language Models for Mathematical Problem Solving
This study probes how sensitive [[llm]] mathematical problem solving is to the surface representation of an item — a question with direct bearing on [[assessment-validity]] when LLMs are used for scor…
📄 Q-Learning Lab: Teaching Reinforcement Learning Through Learner-Generated Trace Analysis
> Presents Q-Learning Lab, a single-file tool that makes the Bellman update concrete by letting undergraduates inspect how each value is computed and why actions are chosen, through learner-generated …
📄 Special-R1: Reinforcement Learning for Special Education — Aligning LLM Tutors to Diverse Learners through Disability-Adaptive Training
> **Authors:** Unggi Lee, Jihoi Na, Yeil Jeong, Haeun Park, Yeonju Jang (2026)…
📄 The Tutoring Effectiveness Index: Predicting LLM Math Tutor Quality from Four Conversation Signals
> **Authors:** Shim Jaechang, Unggi Lee (2026) — CIKM 2026…
📄 Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues
A key part of developing large language model (LLM)-powered, automated tutoring tools is student simulation, i.e., using LLMs to role-play as students, which can facilitate tutor model evaluation and …
2026-05-29 · intelligent-tutoring, llm, student-experience, learning-analytics, personalized-learning
📄 Pedagogical Safety in Educational Reinforcement Learning
> As reinforcement learning personalizes instruction in intelligent tutoring systems, there is no formal framework for pedagogical safety — a critical gap. > First formal framework for defining and de…