🧠 AI Ed Wiki

🏷️ educational-measurement

10 pages tagged with educational-measurement(6 articles, 4 concepts)

🏷️ Research Methods in AIED
> **Research methods in AIED** — the set of empirical designs, data-collection strategies, and analytic techniques researchers use to study AI in education: whether and how AI tools support (or harm) …
2026-08-13 · ai-education, efficacy-study, rct, benchmark, methodology
📄 Multimodal Item Parameter Estimation using Simulated Response Probabilities
> **Synthesis:** This paper fine-tunes a multimodal large language model (Qwen3.5-based) to reconstruct multiple-choice model (MCM) and three-parameter logistic (3PL) item characteristic curves. By le…
2026-08-12 · item-response-theory, student-modeling, llm, multimodal, automated-assessment
🏷️ Cognitive Diagnosis
> **Cognitive diagnosis** — the inference of a learner's latent knowledge state — the specific concepts, skills, and misconceptions they have or lack — from their responses or behavior. It is the asse…
2026-08-12 · student-modeling, knowledge-tracing, assessment, intelligent-tutoring, learning-analytics
🏷️ Dual-Process Theory
> **Dual-process theory** — the account of cognition as operating through two interacting systems: a fast, automatic, intuitive System 1 and a slower, effortful, analytical System 2. In education, dua…
2026-08-12 · cognitive-load-theory, metacognition, cognitive-offloading, decision-making
📄 From Evaluated Models to Evaluation Aids: A Multi-Evidence Study of LLM-Based Difficulty Calibration for Programming Examinations
> **Synthesis:** Yan, Xiong, Li & Chen (2026) reposition LLMs from benchmark targets to auxiliary evidence sources for interpreting programming-exam difficulty, showing that AI difficulty estimates co…
2026-08-11 · computing-education, programming-education, assessment, automated-assessment, llm-evaluation
📄 AI-based scoring systematically underestimates conceptual understanding of linguistically weak students' explanations in physics
> **Authors:** Markus S. Feser, Paul L. Tschisgale (Leibniz Institute for Science and Mathematics Education, Kiel, Germany)…
2026-07-31 · assessment-validity, automated-grading, bias-mitigation, equity, multilingual-learning
📄 ICLE++: Modeling Fine-Grained Traits for Holistic Essay Scoring
Introduces ICLE++, a new annotated corpus of persuasive student essays that addresses critical limitations of the dominant ASAP benchmark in [[automated-essay-scoring]] research. Unlike ASAP — used by…
2026-07-31 · automated-essay-scoring, automated-grading, benchmark, formative-assessment, higher-ed
📄 Analyzing Undergraduate Problem-Solving in Physics Through Interaction With an AI Chatbot
> **Synthesis:** A custom Socratic AI chatbot deployed in a large-enrollment introductory mechanics course with 150 first-year STEM majors, demonstrating that AI-driven Socratic dialogue can foster ex…
2026-07-29 · socratic-method, physics-education, generative-ai, intelligent-tutoring, socratic-questioning
🏷️ AI Ed Evaluation
> **AI-ed evaluation** — the body of methods, benchmarks, and criteria used to assess whether AI education tools (LLM-based tutors, automated graders, feedback systems, agents) actually work — not jus…
2026-05-29 · llm, assessment, benchmark, formative-assessment, teacher-role
📄 Gen-AI-tecture: using generative AI to support architectural students in design tasks
Kapsalis (2026) presents one of the first empirical studies of generative AI integration in architectural design education, using a locally executed, discipline-specific tool within a mixed-methods fo…
2026-05-21 · generative-ai, higher-ed, student-experience, creative-thinking, ai-literacy

Related Tags

higher-ed (4)benchmark (3)llm (3)automated-assessment (3)assessment (3)learning-analytics (3)generative-ai (3)evaluation (2)item-response-theory (2)student-modeling (2)psychometrically-aware-ai (2)intelligent-tutoring (2)assessment-validity (2)automated-grading (2)equity (2)