Skip to content
Conference Open access

Emergence of Social Reasoning through Adaptive Evolution of LLMs

Aug 2026 · IEEE Symposium on Artificial Life · 0 citations

TL;DR

A constructive, artificial-life approach that treats LLMs as model organisms investigating the emergent mechanisms of cognitive functions through biological adaptive evolution rather than static analysis is proposed, as a step toward understanding the human-AI societies now taking form.

Abstract

Large language models (LLMs) exhibit advanced social reasoning capabilities like Theory of Mind (ToM), yet the dynamic acquisition process remains underexplored. We propose a constructive, artificial-life approach that treats LLMs as “model organisms,” investigating the emergent mechanisms of cognitive functions through biological adaptive evolution rather than static analysis, as a step toward understanding the human-AI societies now taking form. By applying a genetic algorithm to evolve LoRA adapters from a behaviorally degraded state, in which task performance is reduced to near-random levels while latent knowledge remains in the frozen weights, we analyzed this evolutionary process at both behavioral and mechanistic levels. At the behavioral level, comparing adaptations in knowledge-intensive (MMLU) and social reasoning (ToMBench) environments revealed that task characteristics dictated fitness landscape ruggedness. An asymmetric generalization was observed: while moderate adaptation to broad knowledge partially bolstered heuristic social reasoning, excessive specialization created an evolutionary trade-off constraining deep inferential capabilities. At the mechanistic level, a Sparse Autoencoder (SAE) revealed the dynamic refinement of reasoning mechanisms during ToM evolution. The evolved individual’s strategy underwent a stepwise transition from superficial linguistic cues to mental state concepts, ultimately specializing in ToM-related conceptual representations. This stepwise acquisition trajectory, alongside the compensatory reasoning observed in the knowledge-intensive environment, suggests a structural generality in the adaptive acquisition of higher-order cognitive capabilities. Data/Code available at: https://doi.org/10.5281/zenodo.20790937

Read PDF

Similar papers

Open access Sep 2026

Large Language Models Predict Human Social Behavior via Interpretable Mechanisms

MindEvolve is introduced, an autonomous workflow designed to predict behavior in social interactions by generating interpretable symbolic models of cognition, and provides a roadmap for advancing LLM-based cognitive modeling toward human-expert-level theory construction.

Ying-Ying Ye, You-Le Fang, Xiao-Xue Gao et al. · 0 citations
Open access Aug 2026

Counter-Inferential Behavior in Natural and Artificial Intelligent Agents

This study explores the emergence of counter-inferential behavior in natural and artificial cognitive systems, that is, patterns in which agents misattrib-ute empirical success or suppress adaptation, leading to epistemic rigidity or mal-adaptive stability. We analyze archetypal scenarios in which such behavior arises:...

Serge Dolgikh · 0 citations
Preprint Aug 2026

Assessing mentalization in humans and large language models

Different capacities for mentalization across LLMs are demonstrated, and cognitive computational modeling is highlighted as a formal method for assessing comparative intelligence across humans and machines.

Aamir Sohail, Xintong Zhong, Arkady Konovalov et al. · 0 citations
#artificial intelligence Preprint Oct 2026

Benchmarking Psychological Dynamics in Generative Agents

Large language models (LLMs) are increasingly deployed to simulate human behavior, acting as computational replicas of human subjects. Yet the lived psychological experience of humans is difficult to benchmark, particularly as it unfolds over time. We introduce a psychometric benchmark for computational replicas: perso...

S. Vaid, Ashley Whillans · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.