Assessing the Quality of Personality Measures Generated by Large Language Models
ready for submission · PNAS
Can an LLM provide a low-cost, scalable measure of Big Five traits from student letters,
and do these measures meaningfully predict future outcomes? We treat the LLM as an
additional informant in a multi-informant framework and show that AI ratings retain
distinct predictive power alongside teacher, student, and guardian reports.