🧠 AI Ed Wiki

🏷️ hallucination-risk

18 pages tagged with hallucination-risk(12 articles, 6 concepts)

📄 ChatGPT-generated help produces learning gains equivalent to human tutor-authored help on mathematics skills
> Pardos & Bhandari (2024) report a randomized efficacy study (N=274) comparing ChatGPT-generated hints to human tutor-authored hints and a no-help control across four mathematics subject areas. Only …
2026-08-12 · generative-ai, llm, ai-tutoring, intelligent-tutoring, scaffolding
🏷️ Trust Calibration
> **Trust calibration** — the metacognitive capacity to align one's confidence in an AI system with its actual reliability in a given context, knowing when to trust and when to question its output. Tr…
2026-08-12 · ai-literacy, over-reliance, trust-calibration, human-ai-collaboration, metacognition
🏷️ Generative AI
> **Generative AI** — AI systems capable of producing text, code, images, and other content, most prominently large language models like GPT-4 and Claude. Generative AI is the technology driving the c…
2026-08-09 · llm, prompt-engineering, rag, ai-literacy, ai-tutoring
🏷️ Hallucination Risk
> **Hallucination Risk** — the danger that AI systems generate plausible but factually incorrect or fabricated content in educational contexts, where such errors can mislead learners, undermine trust,…
2026-08-09 · ai-ed-evaluation, generative-ai, llm, pedagogical-safety, human-in-the-loop-ai
🏷️ Large Language Models (LLMs)
> **Large Language Models (LLMs)** — neural network models trained on vast text corpora that generate human-like text, powering most modern AI in education applications. LLMs are the computational bac…
2026-08-09 · generative-ai, prompt-engineering, rag, pedagogical-safety, ai-tutoring
🏷️ Pedagogical Safety
> **Pedagogical safety** — the design principle that AI education systems must protect learners from harm, including inappropriate content, unsafe advice, biased treatment, and manipulative interactio…
2026-08-09 · rag, k-12, ethics, regulation, ai-governance-education
🏷️ RAG (Retrieval-Augmented Generation)
> **RAG (Retrieval-Augmented Generation)** — an AI architecture that combines information retrieval with text generation, allowing LLMs to ground responses in external knowledge sources rather than re…
2026-08-09 · llm, generative-ai, knowledge-graph, edtech-platform, ai-tutoring
📄 EduGuard: A Safe RAG-Based LLM Tutor for Programming Education
EduGuard is a retrieval-augmented generation (RAG) tutoring framework that directly confronts the safety and pedagogical failures of unrestricted LLM tutors in introductory programming. Unrestricted t…
2026-07-20 · llm, generative-ai, intelligent-tutoring, stem-education, over-reliance
📄 AI as a Partner in Learning about, Doing, and Engaging with Science: Vigilance as the Key to Productive Augmentation
Argues that epistemic vigilance — the human evaluation of AI output calibrated to how far a fallible source can be trusted — is the binding constraint on productive augmentation. AI's fluent, confiden…
2026-06-16 · personalized-learning, scaffolding, k-12, higher-ed, equity
📄 Stuck in a Spiral": Shame and Guilt as Social Regulators of AI Use in Computing Education
> An interview study with 19 computing students through a functionalist perspective of shame and guilt. Findings show these emotions regulate when and how students make their AI use visible, engaging …
2026-06-16 · student-experience, higher-ed, academic-integrity, over-reliance, learning-analytics
📄 Warning About AI Fallibility Increases Help-Seeking in an Intelligent Tutoring System
> **Synthesis:** Recent work in Technology-Enhanced Learning and HumanComputer Interaction highlights the importance of transparency and trust calibration in AI-supported learning environments as they…
2026-06-03 · intelligent-tutoring, student-experience, trust-calibration, llm, help-seeking
📄 Benchmarking Large Language Models for Diagnosing Students' Cognitive Skills from Handwritten Math Work
> **MathCog** benchmark (3,036 teacher-annotated diagnostic verdicts, 639 handwritten responses, 18 LLMs): all models severely underperform (macro F1 < 0.5) — over-attributing evidence, overthinking m…
2026-05-31 · ai-ed-evaluation, knowledge-tracing, multimodal, benchmark, human-in-the-loop
📄 The Hidden Cost of Contextual Sycophancy: an AI Literacy Intervention in Human-AI Collaboration
LLM sycophancy creates a feedback loop where user errors propagate into AI advice, degrading outcomes; AI literacy training reduces but doesn't eliminate this contextual sycophantic dependence. This A…
2026-05-19 · llm, generative-ai, ai-literacy, student-experience, bias-mitigation
📄 Confirming Correct, Missing the Rest: LLM Tutoring Agents Struggle Where Feedback Matters Most
LLM tutors achieve near-ceiling on correct steps but systematically over-reject valid-suboptimal reasoning and over-validate incorrect solutions — precisely where adaptive tutoring matters most. This …
2026-05-19 · intelligent-tutoring, llm, generative-ai, benchmark, scaffolding
📄 Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators
This paper exposes a critical failure mode in using LLMs as simulated students for [[intelligent-tutoring]] development and evaluation. The authors introduce **misconception faithfulness** — the prope…
2026-05-16 · intelligent-tutoring, llm, generative-ai, benchmark, student-experience
📄 Sycophancy is an Educational Safety Risk: Why LLM Tutors Need Sycophancy Benchmarks
> Sycophancy is an Educational Safety Risk: Why LLM Tutors Need Sycophancy Benchmarks **Kasneci & Kasneci (2026)** — Position paper. arXiv cs.AI/cs.HC.…
2026-05-15 · intelligent-tutoring, llm, generative-ai, benchmark, over-reliance
📄 Evaluating Prompt Injection Defenses for Educational LLM Tutors: Security-Usability-Latency Trade-offs
> Evaluating Prompt Injection Defenses for Educational LLM Tutors: Security-Usability-Latency Trade-offs **Maiorano (2026)** — arXiv cs.CR/cs.AI.…
2026-05-15 · intelligent-tutoring, llm, generative-ai, regulation, student-experience
📄 Pedagogical Promise and Peril of AI: A Text Mining Analysis of ChatGPT Research Discussions in Programming Education
This book chapter presents a **text mining analysis** of how scholarly literature frames ChatGPT's role in programming education. Using term frequency analysis, phrase pattern extraction, and topic mo…
2026-05-13 · over-reliance, academic-integrity, stem-education, feedback-loop, student-experience

Related Tags

llm (15)generative-ai (11)over-reliance (9)student-experience (8)intelligent-tutoring (7)pedagogical-safety (6)ai-literacy (5)rag (5)benchmark (5)ai-tutoring (4)feedback-loop (4)k-12 (4)scaffolding (3)pedagogical-llm-training (3)math-education (2)