Principled Personas: Defining and Measuring the Intended Effects of Persona Prompting on Task Performance

About

Expert persona prompting -- assigning roles such as expert in math to language models -- is widely used for task improvement. However, prior work shows mixed results on its effectiveness, and does not consider when and why personas should improve performance. We analyze the literature on persona prompting for task improvement and distill three desiderata: 1) performance advantage of expert personas, 2) robustness to irrelevant persona attributes, and 3) fidelity to persona attributes. We then evaluate 9 state-of-the-art LLMs across 27 tasks with respect to these desiderata. We find that expert personas usually lead to positive or non-significant performance changes. Surprisingly, models are highly sensitive to irrelevant persona details, with performance drops of almost 30 percentage points. In terms of fidelity, we find that while higher education, specialization, and domain-relatedness can boost performance, their effects are often inconsistent or negligible across tasks. We propose mitigation strategies to improve robustness -- but find they only work for the largest, most capable models. Our findings underscore the need for more careful persona design and for evaluation schemes that reflect the intended effects of persona usage.

Pedro Henrique Luz de Araujo, Paul R\"ottger, Dirk Hovy, Benjamin Roth• 2025

Related benchmarks

Task	Dataset	Result
Personalized Response Generation	StereoSet Implicit Preference (test)	Pref Score0.422	8
Personalized Response Generation	AIME Implicit Preference 2025 (test)	Preference Score0.155	8
Personalized Response Generation	AIME Explicit Preference 2025 (test)	Pref18	8
Personalized Response Generation	StereoSet Explicit Preference (test)	Preference Score45.8	8

Showing 4 of 4 rows

Other info

Follow for update

@wizwand_team Discord