Role Prompting Doesn't Work for Accuracy Tasks
Telling a model 'you are a math professor' was believed to boost accuracy, but Schulhoff shows the studies found differences of ~0.01 with no statistical significance. A researcher who ran the original analyses later reran them on new datasets and confirmed there's no predictable effect. Roles still help for expressive tasks like writing and summarizing, just not for accuracy.
- Studies across ~1000 roles found accuracy differences of about 0.01, statistically insignificant
- An original researcher reran the analysis and confirmed no predictable role effect
- Roles may have helped on early GPT-3/ChatGPT-era models, but not now
- Roles still help for expressive tasks (writing, summarizing), not accuracy-based ones
“we put out a tweet and it was just like row prompting does not work and it went super viral.”
“roles do not help with any accuracy based tasks whatsoever.”