Are LLM-based Chatbots Good Enough to Support Computer Science Students in Multiple-Choice Exercises?

Created: 2026-06-16 | Tags: higher-edllmautomated-gradingstudent-experiencestem-education

Markos Stamatakis, Omkar Gavali, Joshua Berger, Christian Wartena, Anett Hoppe, Ralph Ewerth (2026) โ€” arXiv preprint ๐Ÿ“„ Full text (arXiv)

Investigates LLM chatbots' performance on 70 MCQs for a university CS lecture on interactive visual data analysis, comparing with student performance. GPT-4o and GPT-5 significantly outperformed smaller models. A user study in two courses showed that presenting ChatGPT answers with explanations did NOT generally improve student performance.

Key Contributions

Related Pages

Citation

APA: Markos Stamatakis, Omkar Gavali, Joshua Berger, Christian Wartena, Anett Hoppe, Ralph Ewerth (2026). Are LLM-based Chatbots Good Enough to Support Computer Science Students in Multiple-Choice Exercises?. arXiv:2606.15919. arXiv preprint.