09.02Segmentation and Audience UnderstandingAvailable

Persona Development

Build personas where every attribute traces to evidence or is labelled as interpretation, and thin evidence produces a shorter persona, not an invented one.

Free. Works on Claude, ChatGPT, Gemini or any assistant that accepts a skill file.

What this skill does

The method, encoded.

A persona is a compression device. It exists so a product manager choosing between two features can hold an audience in mind without re-reading a 90-page report. When it works it is one of the most useful things research produces. It fails in a specific way: the format invites colour. A persona with a name, a face, a morning commute and a favourite coffee feels more real than one without, so the colour gets added, and almost none of it was measured. What ends up on the wall is evidence and invention with no visible seam, and nobody downstream can tell which is which.

This skill supplies the discipline that keeps a persona usable and honest. Every attribute is either traceable to a named source with a base, or visibly marked as the researcher's interpretation with a confidence level. Invented biography, fabricated verbatim, quotes reattributed from a participant to a persona name, and stock photographs presented as participants are all prohibited outright. Elements earn their place by changing a decision, which excludes most of what fills a standard template.

You get personas with sourced elements, real quotes with participant identifiers, sizes, the variation inside each group, the counter-evidence, an explicit list of what was not established, and a fieldwork date so the set cannot decay silently.

Best used for

  • Turning validated segments into artefacts teams use day to day
  • Building personas from depth interviews where no segmentation exists
  • Auditing inherited personas by trying to source every statement
  • Reducing a persona set to the number that are genuinely distinguishable
  • Keeping evidence annotation on the artefact when it travels without its report
  • Refusing to invent biography, quotes or photographs to fill a template
  • Refreshing personas whose evidence base has decayed

Typical inputs

What you give it.

A validated segmentation with sizes, stability results and typing tool accuracy, Qualitative corpus with participant identifiers, Source references and bases for every intended attribute, The decisions the personas will inform, and who makes them, Linked behavioural or transactional data (optional), Verified verbatim with participant identifiers (optional), Journey, service or usage-context evidence (optional), Population prevalence data for sizing (optional), An existing persona set for audit or continuity (optional)

Typical outputs

What you get back.

Persona faces with provenance, size, evidence date and re-validation trigger, Source and base attached to every element on the face of the artefact, Interpretation visibly separated from observation, with confidence per claim, Evidence sheet per persona with full references matching the face, Distinguishability record showing substitution and decision tests, and merges, Variation-within and counter-evidence lines on every persona, A "not established" block naming gaps and decisions the persona should not inform, Persona set summary table with sizes and defining differences, Statement of populations, occasions and needs the set does not cover

Method coverage

What the skill works through.

  1. What a persona is for, and what it is not for
  2. The strict evidence rule: traceable or labelled
  3. The prohibitions: invented biography, fabricated quotes, reattributed verbatim, stock participants
  4. Personas from segmentation versus personas from qualitative work, and the confidence each carries
  5. The seven elements that earn their place
  6. The decoration test: would a different value change any decision?
  7. Attaching source and base to the face of the artefact
  8. How many personas: the substitution test and the decision test
  9. Sizing personas so the vivid minority does not win
  10. Variation within the persona, and the compression damage it prevents
  11. Writing what the persona does not tell you
  12. Persona decay and the re-validation trigger
  13. Auditing an inherited persona set by sourcing every statement
  14. When the honest output is a shorter persona
  15. Where a researcher's judgement is required

Download

Free skill. One file.

Enter your email once. Every skill you download after that takes a single click.

How to install

Add the skill file and the five kernel protocols to a Claude Project, a ChatGPT Project, a Gemini Gem, or paste them at the top of any assistant conversation. Then give it your real research material, not a description of it.

Download skill

Questions

Common questions.

What should a persona actually contain?

Seven elements earn their place in almost every case, because each one changes a decision: the goal the person is pursuing, the job to be done, the context and constraints they operate under, their current behaviour, the criteria they use when choosing, the barrier that stops or slows them, and the unmet need. Anything beyond that must justify itself against a named decision. The test is specific: would a different value on this attribute change any decision the persona is used for? Favourite coffee, car, family composition and weekend hobbies almost never pass it, unless the category makes them load-bearing.

Is it acceptable to invent details to make a persona feel realistic?

No. This is the central prohibition. Invented age, occupation, household, daily routine and possessions are indistinguishable from measured ones once they are on the page, so downstream users design for a fiction while believing they are following research. If an element has no evidence, write "not established" and leave it. A persona with four evidenced elements and three visible gaps is honest and useful; the same persona with the gaps filled in plausibly is a fabrication that will outlive the study.

Can I write a quote for a persona?

Only if a real participant said it, and it should carry that participant's identifier. A composite sentence in quotation marks is a fabricated quote no matter how faithfully it summarises the theme, and a real quote reassigned from participant P07 to a persona called "Sarah" is misattribution. Where no suitable real quote exists for a point, the persona carries no quote there. In practice, real quotes with identifiers give a persona far more credibility and more emotional weight than invented ones, which tend to read as written by a marketer.

How many personas should we have?

As many as are genuinely distinguishable, which is usually three to five and often fewer than requested. Run two tests. The substitution test: take a claim from persona A and put it in persona B; if it still reads as true, they are not distinguishable on that element, and personas surviving substitution on fewer than two or three elements should be merged. The decision test: would the team make a different choice for A than for B? Beyond four or five, teams stop distinguishing and default to the one they remember, so the extra personas cost attention and change nothing.

Should personas be built from a segmentation or from qualitative research?

Either, and they carry different confidence. A persona built on a validated segmentation can carry population sizes, tested differences and declarative statements. A persona built on depth interviews carries a described pattern with participant counts, uses hedged language, and does not carry percentages, because qualitative evidence does not support population estimates. A set built from both should say, element by element, which stream each claim came from, or the mixed artefact silently promotes the weaker evidence to the confidence of the stronger.

How do I check whether existing personas are evidence-based?

Take each statement and try to source it. Produce a marked-up version showing sourced claims, unsourced claims and claims contradicted by the data. In most inherited sets, somewhere between a third and two-thirds of the statements have no traceable source. The marked-up version is more persuasive than a rebuild, because it names the specific claims currently circulating as fact, and it usually makes the case for re-evidencing on its own.

Why do personas stop being used?

Three common reasons. They contain nothing decision-relevant, so the team has no occasion to consult them. They are not distinguishable, so the team defaults to one. Or they read as fiction to people who know the customers, which happens when they are full of invented character. A fourth reason is silent decay: a persona looks identical on the wall in year four, and nothing on the artefact says when the evidence was collected. Put the fieldwork date and a re-validation trigger on the face of every persona.

What is compression damage?

A persona is an average with a face, and the people it hides are the ones the design will fail. A team designs for the persona and misses the third of the group who behave differently. The fix is two lines on every persona: what varies within this group on the elements that matter, and what evidence cut against this characterisation. It is a small addition that prevents the largest downstream error, and it is never present in a fabricated persona, which makes it a useful audit signal.

Can AI generate personas?

It generates them extremely fluently, and that is the problem. A persona template has slots, and a language model fills them plausibly whether or not evidence exists, producing a uniformly complete set with no gaps anywhere. Uniform completeness is a warning sign rather than a mark of quality: real evidence bases are uneven. Use AI for the mechanical work of extracting elements from a corpus and attaching source references, and check the output by trying to source every statement in it.

What if the client asks for names and photographs?

Offer the strongest honest alternative rather than refusing flat. Label personas by their defining difference rather than a first name, so nothing invites the reader to invent the rest. Use no photograph that could be read as a picture of a participant; an abstract illustration or none at all is safer, and any image used is credited as an illustration. Carry the human weight with real, identified quotes instead. Teams generally find the real quotes more compelling than the name and face they asked for.

Research where people already are.
Analyse it where you already work.

Yazi helps researchers conduct surveys, AI interviews and longitudinal research directly through WhatsApp.

New Report on SA Gambling Impact
Check It Out