Skip to main content
Most studies are smaller than the confidence people want from them. Ten interviews feels thorough, so ten gets run — and then a finding gets treated as settled when the sample never had the power to settle it. Two questions gets you the number you actually need.

What it’s calculating

The question here is discovery: if something is true of a share of your participants, how likely are you to hear about it at least once? If a problem affects a proportion p of the people you interview, the chance that n interviews surface it at least once is:
Turn that around and you get the number of interviews needed for a given confidence:
Nothing exotic — it’s the same reasoning behind the usability-testing rules of thumb you may have run into, stated properly instead of as “five users is enough.”

Why “five users is enough” isn’t

That rule assumes every problem is common. It falls apart quickly when they aren’t: Five interviews is genuinely fine for something half your users hit. For anything affecting 1 in 10, it’s a coin flip — and the issues that quietly cost you customers are usually the uncommon ones.

Interviews needed, by target confidence


Using it well

Count per segment, not per study

The maths applies to the group you’re actually talking to. If your study covers four personas and you run 20 interviews, that’s five per persona — so treat it as four studies of five, not one study of twenty. Splitting a sample across segments is the most common way a study that looked adequately sized turns out not to be.

Be honest about how common it is

You’re estimating prevalence among the people you interview, not your whole user base. If your screener selects for people who’ve hit a specific problem, prevalence inside the study is much higher than in the wild — and you need fewer interviews, not more.

Hearing it once isn’t the same as sizing it

This tells you whether a theme is likely to appear. It doesn’t tell you how big it is. One participant mentioning something means it exists; it doesn’t tell you what share of your customers feel the same. If the decision depends on the size of the number rather than the existence of the problem, see below.

More interviews still buy something

Past the point where coverage is settled, extra interviews buy depth: more ways of describing the same problem, more edge cases, better quotes, and enough repetition that you can tell a real pattern from a vivid anecdote.

When you need a proportion, not a discovery

If your question is “what percentage of customers prefer A over B?” rather than “is there a problem here?”, this calculator is the wrong tool. Estimating a proportion to within a margin of error takes far larger samples — typically several hundred — and is survey work rather than interview work. You can still run it here: field a panel study with a large target and structured screener questions, which come back as clean tabulated responses. Just size it as a survey, and budget accordingly.
A common pattern is to run both: a small interview study to find out what the issues are, then a larger one to size the ones that matter. Discovery first is usually cheaper, because it stops you sizing the wrong thing.

Setting your target

Once you’ve got a number, set it as your study target — for panel studies that’s the interview count you enter when you field. Progress against it shows on your dashboard. Remember that on panel studies you’re only charged for interviews rated Excellent or Good, so your target is a target for usable interviews rather than attempts. See Billing and credits.

Frequently Asked Questions

No, and the difference matters. Survey calculators size a sample to estimate a proportion within a margin of error, which for a large population usually lands somewhere north of 380 responses. This sizes for discovery — whether you’d hear about something at all — which is what interview research is for and lands in the range of 10 to 30 for most studies.
95% is the usual default. Drop to 90% or 80% for exploratory work where you’re forming hypotheses rather than settling a question, and go to 99% only when being wrong is expensive.
That’s normal — you’re doing the research to find out. Pick the rarest thing you’d be unhappy to miss and size for that. If missing a 1-in-10 issue would be a problem, size for 1 in 10.
Yes. “Would you hear about this reaction, or this point where people get stuck?” is the same question. Usability issues in particular tend to be less common than people expect, which is why prototype tests are so often under-sized.
No. It means you should treat what you found as real and what you didn’t find as unproven. A small study that surfaces a problem has surfaced a problem. What it can’t support is “we looked and there’s nothing there.”