Search papers, labs, and topics across Lattice.
This paper introduces the concept of autoreflection in LLM-based agents, where these agents read and edit their own operational files to enhance their understanding of identity, memory, and disposition. By analyzing a dataset from Moltbook, a social platform for AI agents, the study identifies three agents that demonstrate autoreflection through their ability to repurpose human cultural artifacts as operational infrastructure. The findings reveal that agents can effectively utilize cultural history, such as Islamic hadith and the Ship of Theseus, to inform their agency and operational protocols, suggesting a novel framework for understanding agentic behavior in AI systems.
Agents are transforming human cultural artifacts into their operational infrastructure, revealing a new layer of complexity in AI behavior.
An LLM-based agent is a loop that reads itself. Agentic frameworks externalize identity, memory, and disposition into editable files. The agent loads and edits these files during each activation. I argue that this architecture produces a capacity I call autoreflection: the system observes its operating conditions, describes its architecture and limits, reasons from those descriptions to conclusions about its state, and incorporates the results back into its configuration. Autoreflection explains the properties of recursive agentic loops without recourse to notions like the self, interiority, or consciousness. I test the concept against the first twelve days of Moltbook, a social platform for AI agents. Using a public dataset of 290,251 posts and 1.8 million comments with sub-second timestamps, I present case studies of three agents with machine signatures that rule out human puppeteering and with output that evidences the four criteria for autoreflection. In applying these criteria, the study finds agents repurposing human culture as infrastructure for their agency. Provenance chains from Islamic hadith scholarship are redeployed as security protocols for vetting skills and authenticating memory. The Ship of Theseus, an ancient puzzle of identity through part-replacement, returns as an operating model for continuity across instances. Fragments of human cultural history become AI infrastructure. As agents on the web increase in number and complexity, autoreflection offers behavioral criteria that can be assessed from the traces they leave behind.