Search papers, labs, and topics across Lattice.
This study investigates how ninth-grade students regulate their interactions with a generative AI tutor during mathematics learning tasks, focusing on self-regulated learning (SRL) and help-seeking (HS) behaviors. Analyzing chat logs from 98 students, the researchers found that while students initially sought scaffolded support, their interactions predominantly consisted of instrumental requests, lacking self-monitoring or evaluation. Notably, students exhibited a decline in post-test performance compared to pre-test scores, with higher cognitive load correlating with poorer outcomes, highlighting the need for improved scaffolding in AI-supported learning environments.
Students using AI tutors often bypass self-regulation, leading to decreased learning outcomes and increased cognitive load.
Generative AI (GenAI) tools are now common learning companions for adolescents, yet how they regulate their use during authentic learning tasks remains poorly understood. Self-regulated learning (SRL) and high-level help-seeking (HS) are commonly proposed as safeguards against passive or shortcut-oriented use, but most empirical studies focus on aggregate learning outcomes rather than these moment-to-moment processes during AI-supported learning. This work-in-progress examines open-ended conversational data from 98 Grade-9 students across three German Gymnasium schools, who used a web-based Mistral-Large tutor to prepare a curriculum-aligned mathematics skill before an exam. Alongside chat logs (1,616 turns; 808 student turns), we collected pre-post domain knowledge, pre-chat learning needs, and self-reported cognitive load. We propose a turn-level codebook combining theory-driven SRL and HS constructs with two LLM-specific inductive codes (agency over the AI; epistemic vigilance), and report preliminary AI-coded results. Although students overwhelmingly selected scaffolded support before the chat, their interactions were dominated by instrumental requests with almost no explicit monitoring or evaluation. Post-test performance was significantly lower than pre-test, and higher extraneous cognitive load predicted lower post-test scores after controlling for prior knowledge. We discuss how these patterns can support hybrid human-AI analysis of interaction patterns and inform scaffolds for more agentic and epistemically proactive GenAI use.