A century of psychology has found that the trait words people use to describe one another vary, but the relational structure among those traits, which ones go together and which oppose, is strikingly consistent across raters and cultures. We test whether the LLM (Qwen 2.5-7B-Instruct) reproduces this structure in its internal trait representations. From millions of crowd-sourced personality ratings of fictional characters, we build a human implicit-personality matrix over hundreds of traits; from contrastive model activations, we build a matching matrix over the same traits. The two relational structures align strongly (Mantel r = 0.77), and the agreement holds trait by trait as well as in aggregate. Two dominant axes of the model's trait representations recover the social and intellectual dimensions long known to organize human personality impressions, social warmth and intellectual competence. On held-out dialogue, projecting model activations onto these directions yields personality profiles that agree with human ratings. This work establishes a framework that enables comprehensive, human-grounded comparison between internal model trait geometry and the shared structure of human personality impressions.
This work proposes A-B-D to infer traits bottom-up from behavioral data (B-data), namely how agents act on their environment and communicate with users, as recorded in existing trajectories, and offers a new lens for understanding AI personality.
Hao-Kai Zhao, Jie Gao, Yunze Xiao et al.· 0 citations
Do large language models (LLMs) exhibit coherent personality profiles, and are these profiles culturally neutral? This study used two validated personality inventories, the Big Five Inventory–2 and the Dark Triad Dirty Dozen, to assess five state-of-the-art LLMs (ChatGPT-4o, Gemini 2.0 Flash, DeepSeek-V3, Ernie 3.5 a...
Yu-Zhan Hang, C. Soto, B. Lignier et al.· Royal Society Open Science· 0 citations
Large language models (LLMs) are used to simulate human participants in psychological research. We asked whether LLMs that reproduce human evaluations of a whistleblower's moral character also reproduce the motive attributions that accompany them. Five LLMs and two human samples (N = 125 and N = 742) evaluated a physic...
This work proposes PRISM (Persona Reasoning with Inverse SFL-based Modeling), a psycholinguistically grounded framework that reformulates persona fidelity evaluation as a structured inverse inference task, providing a more reliable framework for persona fidelity evaluation.
Meng-Fan Li, Ze-Sheng Wei, Xuan-Hua Shi et al.· 1 citation
It is found that models align more towards deep acting, and that Conscientiousness and Emotional Stability consistently predict this preference, and entropy analysis confirms that persona reliably influences the output and varies across models and emotions.
With $2.1 million funding from Google.org, the open-source Public Transit Intelligence Hub will unify public transit monitoring, operations, and passenger communication.
Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.