Search papers, labs, and topics across Lattice.
This paper introduces the CoRe-3 competency model, which delineates three essential skills鈥擣raming, Judging, and Steering鈥攏ecessary for effectively utilizing generative AI in educational contexts. By assessing these skills separately, the model reveals how students can better engage with AI-generated outputs, addressing the limitations of traditional performance metrics that fail to capture the nuances of AI interaction. The findings demonstrate that each skill can be independently measured and that their interrelations can provide insights into students' competencies in AI-assisted reasoning tasks.
Students' ability to effectively engage with generative AI hinges on mastering distinct skills, not just prompting.
Generative AI makes answers easy and understanding hard, and uncritical use invites cognitive offloading. Schools still measure unaided performance, yet the real task is to produce good work with AI: framing an ill-defined task, judging the output, and steering the model toward a better result. This ability is rarely assessed in its own right; where measured, it collapses into one "prompting" score that cannot diagnose why AI use succeeds or fails. We propose CoRe-3 (Co-Reasoning), a competency model factoring productive AI use into three assessable skills we abbreviate FJS: Framing (specifying an ill-defined task before invoking AI), Judging (evaluating output for errors and unstated assumptions), and Steering (iteratively redirecting the model). Its distinguishing claim is the separation of pre-generation Framing from post-generation Steering, with Judging as the gate between. We ground the skills in theory, state five testable propositions, and instantiate them in CoReasoningLab, an open platform that presents flawed AI output and scores them independently. Over simulated learners (generated and graded by different models), the skills dissociate: each tracks its own manipulated competence while staying flat in the others, and grades become correlated when one competence is shared across all three (convergent and discriminant validity), across grader backends from two providers. Human-rater agreement and outcomes are next; we release the instrument, data, and protocol.