🏷️ educational-measurement
10 pages tagged with educational-measurement(6 articles, 4 concepts)
🏷️ Research Methods in AIED
> **Research methods in AIED** — the set of empirical designs, data-collection strategies, and analytic techniques researchers use to study AI in education: whether and how AI tools support (or harm) …
📄 Multimodal Item Parameter Estimation using Simulated Response Probabilities
> **Synthesis:** This paper fine-tunes a multimodal large language model (Qwen3.5-based) to reconstruct multiple-choice model (MCM) and three-parameter logistic (3PL) item characteristic curves. By le…
🏷️ Cognitive Diagnosis
> **Cognitive diagnosis** — the inference of a learner's latent knowledge state — the specific concepts, skills, and misconceptions they have or lack — from their responses or behavior. It is the asse…
2026-08-12 · student-modeling, knowledge-tracing, assessment, intelligent-tutoring, learning-analytics
🏷️ Dual-Process Theory
> **Dual-process theory** — the account of cognition as operating through two interacting systems: a fast, automatic, intuitive System 1 and a slower, effortful, analytical System 2. In education, dua…
📄 From Evaluated Models to Evaluation Aids: A Multi-Evidence Study of LLM-Based Difficulty Calibration for Programming Examinations
> **Synthesis:** Yan, Xiong, Li & Chen (2026) reposition LLMs from benchmark targets to auxiliary evidence sources for interpreting programming-exam difficulty, showing that AI difficulty estimates co…
2026-08-11 · computing-education, programming-education, assessment, automated-assessment, llm-evaluation
📄 AI-based scoring systematically underestimates conceptual understanding of linguistically weak students' explanations in physics
> **Authors:** Markus S. Feser, Paul L. Tschisgale (Leibniz Institute for Science and Mathematics Education, Kiel, Germany)…
📄 ICLE++: Modeling Fine-Grained Traits for Holistic Essay Scoring
Introduces ICLE++, a new annotated corpus of persuasive student essays that addresses critical limitations of the dominant ASAP benchmark in [[automated-essay-scoring]] research. Unlike ASAP — used by…
📄 Analyzing Undergraduate Problem-Solving in Physics Through Interaction With an AI Chatbot
> **Synthesis:** A custom Socratic AI chatbot deployed in a large-enrollment introductory mechanics course with 150 first-year STEM majors, demonstrating that AI-driven Socratic dialogue can foster ex…
🏷️ AI Ed Evaluation
> **AI-ed evaluation** — the body of methods, benchmarks, and criteria used to assess whether AI education tools (LLM-based tutors, automated graders, feedback systems, agents) actually work — not jus…
📄 Gen-AI-tecture: using generative AI to support architectural students in design tasks
Kapsalis (2026) presents one of the first empirical studies of generative AI integration in architectural design education, using a locally executed, discipline-specific tool within a mixed-methods fo…