Search papers, labs, and topics across Lattice.
This paper investigates the sociotechnical control challenges posed by the internal deployment of agentic AI systems, particularly focusing on the Loss of Control (LoC) phenomenon. By applying established systems-safety methodologies鈥擲TECA, STPA, and FRAM鈥攖o a reconstructed coding-agent scenario, the authors reveal critical insights about governance gaps and the temporal ineffectiveness of control actions due to monitoring delays. The findings underscore the necessity of integrating systems-level hazard analysis with model-focused evaluations to ensure ongoing effectiveness of control measures in AI deployments.
Governance gaps in AI systems can render control measures ineffective, highlighting the urgent need for integrated risk management strategies.
Internal deployment of agentic AI systems for coding and research creates a sociotechnical control problem that extends beyond model behaviour. We treat internal-deployment Loss of Control as the inability to reliably constrain, audit, reverse, or halt AI-mediated changes to code, infrastructure, evaluation, or deployment processes in time to prevent serious organisational or societal harms. We ask whether established systems-safety methods can identify risks that model-level evaluations may miss. Using a generic frontier-lab coding-agent scenario reconstructed from public materials, we apply STECA, STPA, and FRAM. The analyses surface complementary findings: published frameworks can leave governance responsibilities and feedback loops externally unverifiable; delays in monitoring and intervention can make otherwise valid control actions ineffective; and routine operational variability can gradually erode the calibration and independence of safeguards. We argue that frontier-AI risk management should pair model-focused evaluations with systems-level hazard analysis and operational assurance that tracks whether controls remain effective over time.