Search papers, labs, and topics across Lattice.
This study investigates the emergence of modular cognitive architecture in Large Language Models (LLMs) by analyzing their performance across 46 tasks in four cognitive domains: language, formal reasoning, social reasoning, and physical reasoning. The findings reveal that LLMs exhibit a modular organization akin to the human brain, where tasks relying on the same cognitive network activate overlapping neurons, while those requiring different networks engage distinct neurons. This convergence suggests that modularity may be a fundamental characteristic of intelligent systems, regardless of their underlying optimization processes.
LLMs reveal a modular cognitive architecture strikingly similar to that of the human brain, challenging our understanding of intelligence across different systems.
The human brain exhibits a striking degree of functional specialization, with distinct networks supporting language, formal reasoning, reasoning about other minds, and reasoning about the physical world. Is this modular organization a fundamental principle of how intelligent systems must be built, or an evolutionary accident specific to biological brains? Here, we test whether a similar organization emerges in Large Language Models--another class of intelligent systems created through a very different optimization process. Using circuit analyses across N=46 tasks spanning four cognitive domains (language, formal reasoning, social reasoning, physical reasoning), we find that LLMs develop a modular architecture that mirrors the human brain: tasks drawing on the same network in humans recruit overlapping neurons in LLMs, whereas tasks drawing on different networks recruit distinct neurons. The convergent emergence of modularity in brains and neural networks suggests that it may be a fundamental property of intelligent systems.