A deep learning approach to personality assessment
Generalizing across items and expanding the reach of survey-based research
Bibliographic Data
| ID | 6203594 |
|---|---|
| Authors | Suhaib Abdurahman (0000-0001-5615-0129, University of Southern California), Huy Vu (Stony Brook University), Wanling Zou (University of Pennsylvania), Lyle Ungar (0000-0003-2047-1443, University of Pennsylvania), Sudeep Bhatia (0000-0001-6068-684X, University of Pennsylvania) |
| Year | 2024 |
| Volume | 126 |
| Issue | 2 |
| Pages | 312-331 |
| Publication date | 2024-02-01 |
| Peer Reviewed | Yes |
| Open Access | Yes |
| Type | ARTICLE |
| Venue | Journal of Personality and Social Psychology (JOURNAL) |
| Journal identifiers | ISSN: 0022-3514 • E-ISSN: 1939-1315 |
| Publisher | American Psychological Association (APA) (PUBLISHER) |
| DOI | 10.1037/pspp0000480 |
| PMID | 37676124 |
| OpenAlex | W4386497319 |
| Language | EN |
| Citations received | 8 |
Traditional methods of personality assessment, and survey-based research in general, cannot make inferences about new items that have not been surveyed previously. This limits the amount of information that can be obtained from a given survey. In this article, we tackle this problem by leveraging recent advances in statistical natural language processing. Specifically, we extract "embedding" representations of questionnaire items from deep neural networks, trained on large-scale English language data. These embeddings allow us to construct a high-dimensional space of items, in which linguistically similar items are located near each other. We combine item embeddings with machine learning algorithms to extrapolate participant ratings of personality items to completely new items that have not been rated by any participants. The accuracy of our approach is on par with incentivized human judges given an identical task, indicating that it predicts ratings of new personality items as accurately as people do. Our approach is also capable of identifying psychological constructs associated with questionnaire items and can accurately cluster items into their constructs based only on their language content. Overall, our results show how representations of linguistic personality descriptors obtained from deep language models can be used to model and predict a large variety of traits, scales, and constructs. In doing so, they showcase a new scalable and cost-effective method for psychological measurement. (PsycInfo Database Record (c) 2024 APA, all rights reserved)
Big Five personality traits · Cognitive psychology · Construct (python library) · Machine learning · Natural language processing · Personality · PsycINFO · Task (project management) · Variety (cybernetics) · Artificial Intelligence · Computer Science · Personality Traits and Psychology · Psychology · Social Psychology
Comparison of Psychological Data between Populations, A Measurement Perspective
Ensuring Transparency and Trust in Supervised-Machine-Learning Studies
Neural language models as content analysis tools in psychology
Explainable artificial intelligence in cognitive learning psychology
Human Expertise and Large Language Model Embeddings in the Content Validity Assessment of Personality Tests
Neural Network Analysis of Psychological Data
The general factor of personality (GFP) in natural language
Semantic embeddings reveal and address taxonomic incommensurability in psychological measurement
| Unique citing works | 8 |
|---|---|
| Citations per year | 8 |
| Citation span | 2025 - 2026 (2) |
| Citation velocity | current |
| Highly cited | No |
| Citation types | Neutral: 7 |