Skip to content

Implicit Personality Representations in Humans and LLMs

Sep 2026 · 0 citations · 41 references
Computer Science

Abstract

A century of psychology has found that the trait words people use to describe one another vary, but the relational structure among those traits, which ones go together and which oppose, is strikingly consistent across raters and cultures. We test whether the LLM (Qwen 2.5-7B-Instruct) reproduces this structure in its internal trait representations. From millions of crowd-sourced personality ratings of fictional characters, we build a human implicit-personality matrix over hundreds of traits; from contrastive model activations, we build a matching matrix over the same traits. The two relational structures align strongly (Mantel r = 0.77), and the agreement holds trait by trait as well as in aggregate. Two dominant axes of the model's trait representations recover the social and intellectual dimensions long known to organize human personality impressions, social warmth and intellectual competence. On held-out dialogue, projecting model activations onto these directions yields personality profiles that agree with human ratings. This work establishes a framework that enables comprehensive, human-grounded comparison between internal model trait geometry and the shared structure of human personality impressions.

View source

Similar papers

#natural language process... Preprint Sep 2026

On the Behavioral Traits of LLM Agents

This work proposes A-B-D to infer traits bottom-up from behavioral data (B-data), namely how agents act on their environment and communicate with users, as recorded in existing trajectories, and offers a new lens for understanding AI personality.

Hao-Kai Zhao, Jie Gao, Yunze Xiao et al. · 0 citations
#large language models Open access Sep 2026

Cultural bias or universal traits? Exploring personality-like profiles of large language models

Do large language models (LLMs) exhibit coherent personality profiles, and are these profiles culturally neutral? This study used two validated personality inventories, the Big Five Inventory–2 and the Dark Triad Dirty Dozen, to assess five state-of-the-art LLMs (ChatGPT-4o, Gemini 2.0 Flash, DeepSeek-V3, Ernie 3.5 a...

Yu-Zhan Hang, C. Soto, B. Lignier et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Human-like moral judgments conceal divergent motive attributions in large language models

Large language models (LLMs) are used to simulate human participants in psychological research. We asked whether LLMs that reproduce human evaluations of a whistleblower's moral character also reproduce the motive attributions that accompany them. Five LLMs and two human samples (N = 125 and N = 742) evaluated a physic...

Xiao-Yan Wu, J. Dreher · 0 citations
Preprint Aug 2026

Do LLMs Understand Personality? Rethinking Persona Fidelity Evaluation through Structured Behavioral Inference

This work proposes PRISM (Persona Reasoning with Inverse SFL-based Modeling), a psycholinguistically grounded framework that reformulates persona fidelity evaluation as a structured inverse inference task, providing a more reliable framework for persona fidelity evaluation.

Meng-Fan Li, Ze-Sheng Wei, Xuan-Hua Shi et al. · 1 citation
#natural language process... Preprint Aug 2026

Emotional Labor Strategy Preferences in LLM Personas

It is found that models align more towards deep acting, and that Conscientiousness and Emotional Stability consistently predict this preference, and entropy analysis confirms that persona reliably influences the output and varies across models and emotions.

Mohammad Saim, Tianyu Jiang · 0 citations

Related blog posts

MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.