Does telling ChatGPT to act as an expert actually work?
Sometimes, and not in the way the advice implies. A role line does not give the model access to expertise it was otherwise withholding. It narrows what the model draws on, which means it helps when you want the standard version of a recognised genre and actively hurts when you are trying to produce something that is not the average of that genre.
That is a smaller and stranger effect than "act as an expert" is usually sold as. The line is not a key that unlocks a better model. It is a pointer at a region of what the model already knows, and pointing at a region gets you the centre of that region. Whether the centre is where you wanted to land is the entire question, and it has a different answer for a competitive battlecard than it does for a homepage.
Two things below are worth keeping separate. The first is a census of our own library, which shows what prompts that work look like but cannot tell you whether the role line caused anything. The second is an actual controlled comparison, which can, on one task. We have kept them apart deliberately and said plainly where each one runs out.
How often do working prompts use a role line?
Less often than prompt advice implies. Of the 455 prompts we publish, 97 open with a role line and 358 do not. Excluding the 28 image prompts, where a role makes no sense, that is 97 of 427 text prompts, or 23 percent.
The revealing part is how those 97 are distributed. They are not spread evenly across the library, as they would be if authors reached for a role line whenever a task called for one. They are concentrated in three packs, and inside those three packs every single prompt has one.
| Packs | Prompts | Open with a role line | Adoption |
|---|---|---|---|
| Business Strategy, Content Creator, Marketing | 97 | 97 | 100% |
| The other 11 text packs | 330 | 0 | 0% |
| AI Image Prompts | 28 | 0 | 0% |
Ecommerce, Educator, Freelancer, Health and Wellness, HR and Recruiting, Job Seeker, Legal, Personal Finance, Real Estate, SaaS Growth and Startup Founder ship 330 prompts between them without a single "You are a" opener. These are not lesser prompts. They are the same product, sold the same way, and 311 of those 330 still explicitly name the deliverable, which is the part that does the work.
Does the role line make the prompt better? Our library cannot tell you
This is the point where most articles on the topic quote a correlation and let it imply a cause. Ours would be a tempting one: prompts that open with a role average 14.4 constraint lines against 8.4 for prompts that do not, and 270 words against 166.
That number is worthless as evidence and we are not going to use it. Because adoption is 100 percent or 0 percent per pack, the comparison is not measuring the role line at all. It is measuring the house style of the three packs that happen to use one, and those are also the packs written with longer requirement blocks. Any variable that split the library the same way would produce the same gap. A correlation with a perfect confound is not a weak finding, it is not a finding.
What the library can honestly show is narrower and still useful: the role line is a house convention, not a functional necessity. Eleven packs and 330 prompts do the same job without one. If a role line were load bearing, a library written by people who do this for a living would not have 78 percent of it skipping the technique entirely.
One more clean figure. Role prompts and non role prompts ask for almost exactly the same number of pasted inputs, 7.0 bracketed variables against 7.5. The role line is not standing in for information you would otherwise have to supply.
What happened when we did test it properly
On 28 August 2026 we ran the comparison the library cannot support. One brief, a small batch perfumery, built twice. The only difference in the second run was a line prepended to an otherwise identical prompt:
You are a world-renowned graphic designer and visual designer for websites. You build $10,000 websites and this should look as such.
Nothing else changed. Three things were fixed in advance so the result could be counted rather than argued: the prediction was written down before either page existed, the list of generic design markers was defined before either page existed, and the persona version was built second, so any practice effect favoured the side predicted to lose.
| Marker | No persona | With persona |
|---|---|---|
| Centred text blocks | 0 | 18 |
| Rounded cards | 0 | 14 |
| Gradient backgrounds | 0 | 5 |
| Gradient filled buttons | 0 | 2 |
| Pill and chip elements | 0 | 2 |
| Monogram tiles | 0 | 2 |
| Four up stat strip | 0 | 1 |
| Markers present | 0 of 11 | 7 of 11 |
The persona version also picked Playfair Display and Inter without being asked, the two typefaces most associated with machine generated design. Its type scale went from 7 sizes with no near duplicate pairs to 15 sizes with 4. It ran 62 percent taller on desktop and 84 percent taller on a phone while carrying fewer words, 387 against 408. On the aesthetic pass it scored 38 against 62.
The persona did not make the page worse at random. It moved the page somewhere specific, and the destination has a name: black and gold, centred, letterspaced eyebrow, gradient circles standing in for the two founders. That is what "expensive website" looks like when it is recalled instead of derived.
The result that went against us, reported rather than buried
The persona version won on the axis that matters most to a real reader. The plain version shipped five genuine WCAG contrast failures at 3.29:1 in its small grey text. The persona version had none. We had predicted the persona would score better on structure, and it did.
A caveat cuts back the other way, and it is a caveat about the tool rather than a rescue of the argument: text sitting on a gradient defeats automated contrast checking, because the checker cannot resolve a single background colour and climbs to the page ground instead. Part of the persona version's cleanliness is invisibility to the measurement, not correctness.
The limits of the whole test, stated plainly: n equals 1. One category, one brief, one author, and that author wrote the prediction first and then built both sides. A cleaner version runs several categories and has someone else score them blind. Treat it as one well instrumented data point, not a law.
What a role line actually changes
Put the census and the test together and the mechanism is consistent. A role line selects a category, and the model returns something near the middle of that category. That explains both results at once: it helps when the middle of the category is the correct answer, and it hurts when the middle of the category is the cliche you were trying to escape.
So the useful distinction is not whether you name a role, but whether the role narrows anything. Of our 97 role lines, roughly half are a bare job title and half carry a qualifying clause, and the two do very different work.
| Property | Bare job title | Carries a qualifying clause |
|---|---|---|
| How many | 49 | 48 |
| Median length | 6 words | 15 words |
| Example | Act as a pricing strategist. | Act as a pitch deck strategist who has reviewed 1,000+ investor decks. |
| What it tells the model | Nothing the task did not already imply. The request was about pricing regardless. | Which of several readings to take, and what a good answer would notice first. |
A bare title is redundant with the task. If you ask for a pricing analysis, the model already knows it is doing pricing. The clause is the only part carrying information, and it works because it behaves more like a constraint than a costume: "who has reviewed 1,000+ investor decks" implies a specific vantage, that this reader has seen the common failures and will be bored by the usual answer.
Either way, keep it in proportion. Across those 97 prompts the role clause is a median of 4.1 percent of the prompt words, ranging from 1.5 to 12.1 percent. Whatever else is true, 96 percent of a working prompt is not the persona.
When to use a role line and when to skip it
| Situation | Use a role line? | Why |
|---|---|---|
| Output belongs to a known genre with real conventions (battlecard, SOAP note, discovery guide, press release) | Yes, with a clause | The conventions are the point. Landing in the middle of the category is landing correctly. |
| You need domain vocabulary you cannot supply yourself | Yes | The role reliably shifts register and terminology, which is the one thing it does well. |
| Design, layout, visual direction | No | Measured above. It replaces derivation with recall and summons the template you were trying to avoid. |
| Brand voice, naming, positioning, anything meant to sound like one specific person | No | The same failure. A category has an average voice, and the average is what you get. |
| The prompt is vague and you are hoping the role will carry it | No | It cannot. Write the requirements instead, which is what the role line was standing in for. |
What to write instead
The useful half of the "act as an expert" instinct is the standard, not the costume. The trouble with a costume is that it has no failure condition: there is no way to check whether the output is world class, so the instruction cannot be enforced or debugged. A standard names what would make you reject the answer, and that is checkable.
The translation is mechanical. Ask what you would have used to judge whether the expert had actually done the job, then write that down instead of the title.
Instead of:
You are a world-class copywriter. Write my homepage headline.
Write:
Write 5 homepage headlines for [PRODUCT, e.g. "a scheduling tool
for dental practices"].
Each headline must:
1. Be under 12 words
2. Name a specific outcome, not a category ("fills Monday mornings",
not "streamlines scheduling")
3. Take a different angle from the other four (state the angle in
brackets after each)
Rules:
- No colons splitting the headline into two halves
- Never use: seamless, effortless, revolutionize, unlock, empower
- Do not mention AI unless the product value depends on it
Then name the one you would ship and say what it risks.Nothing in that prompt tells the model who to be, and every line tells it what would count as failing. That is the trade the evidence supports: the persona is 4 percent of the prompt and unfalsifiable, while the requirement block is the rest of it and every line can be checked against the output.
For the longer version of the same argument applied to tone, an adjective like "professional" names a category exactly the way a job title does, which is covered in how to get ChatGPT to write in your voice.
Questions people ask about role prompts
Sometimes, and not by making the model more capable. A role line narrows what the model draws on, so it helps when you want the standard version of a recognised genre, such as a competitive battlecard or a discovery interview guide, and it hurts when you are trying to produce something that is not the average of that genre. In a controlled build test we ran on 28 August 2026, adding one expert persona line to an otherwise identical design brief moved the output from 0 of 11 predefined generic markers to 7 of 11, meaning the second version looked far more like the recalled template. Of the 455 prompts we publish, 358 do not open with a role line at all.
No. There is no reserve of expertise that a role line unlocks. The model has the same knowledge either way, and the role line only shifts which region of that knowledge it favours and what register it writes in. This is why a role line cannot rescue a vague request: if the prompt does not say what the answer must contain, naming a job title in front of it still leaves the model to guess. In our library the role clause is a median of 4.1 percent of the prompt words, so 96 percent of the work is being done by the specification that follows it.
The choice between the two openers makes no practical difference, and both are the weakest form of the technique. Across the 97 library prompts that open with a role, 67 use "You are" and 30 use "Act as". What separates the useful ones is not the verb but whether the role carries a qualifying clause. Roughly half our role lines are a bare job title of about 6 words, such as "Act as a pricing strategist." The other half name a vantage point of about 15 words, such as "Act as a pitch deck strategist who has reviewed 1,000+ investor decks." Only the second kind tells the model anything it could not have inferred from the task.
Skip it whenever the goal is to avoid the obvious answer. Design work, brand voice, naming, positioning, and anything where you would recognise the cliche on sight are all cases where pointing the model at a category summons that category. Our 28 August test used a premium design brief specifically because that is where a persona should help if it helps anywhere, and the persona version reached unprompted for Playfair Display and Inter, centred everything, and added 14 rounded cards and 5 gradient backgrounds where the plain version had none. It scored 38 against 62 on the aesthetic pass.
State the standard the answer has to meet instead of the costume the model should wear. A standard has a failure condition you can check, and an adjective does not. Replace "You are a world class copywriter" with the things you would have used to judge whether a world class copywriter had done the job: no sentence over 20 words, none of these six phrases, an opening line that names a specific number, three options that each take a different angle. In our library, 311 of the 330 prompts that use no role line at all still explicitly name the deliverable, which is the part that actually determines the output.
Keep it to one sentence and make that sentence earn its place. Across the 97 role lines we publish the median is 12 words and the longest is 25, occupying between 1.5 and 12.1 percent of the prompt. The test for whether it is worth including: does the clause narrow what the model will notice, or does it only name a job title? "A pricing strategist" narrows nothing, because the task already implies pricing. "A pricing strategist who has repositioned products after a competitor undercut them" tells the model which of several possible readings of the task to take.
Related reading and next steps: the underlying reason a role line disappoints is the same reason a plain request does, covered in why ChatGPT gives generic answers. For how much of the prompt the requirement block should occupy, see how long a ChatGPT prompt should be. For the rules inside that block that rule things out, how to tell ChatGPT what not to do counts all 412 of them. To start from prompts that already carry the requirement block rather than writing one, browse the prompt packs, or read the how to use guide.