📄 Research Article
Explaining Too Much? Understanding How Large Language Model Reasoning Traces Influence Performance and Metacognition
LLM Reasoning Traces & Metacognition
This preregistered between-subjects study (N=559) provides the first rigorous evidence that LLM reasoning traces — increasingly common in AI interfaces — do not improve performance and can actively impair it. More critically, they create a dangerous metacognitive blind spot: participants substantially overestimate their performance regardless of trace format.
Key Findings
Connection to AIED
These findings have profound implications for Intelligent Tutoring and AI feedback systems. If students feel more confident after seeing AI reasoning but don't actually learn better, then simply exposing AI reasoning in educational interfaces may create an Over Reliance trap. The paper's recommendation — that calibration should be scaffolded by interactions that elicit users' own reasoning first — directly aligns with Self Regulated Learning principles and cognitive offloading research showing that AI use can reduce active engagement.
Contrast with Assessment Governance
While GenAI assessment governance focuses on when to allow AI in evaluation, this paper addresses how AI explanations affect learning — suggesting that even well-designed AI transparency features can backfire without metacognitive scaffolding.
Connected Concepts
Connected Articles
Citation
Fernandes, D., Buschek, D., Tankelevitch, L., Kosch, T., & Welsch, R. (2026). Explaining too much? Understanding how large language model reasoning traces influence performance and metacognition. arXiv:2605.25856. cs.HC.