Ask a hiring team whether their interviews are structured and most will say yes. They have a question list. Everyone uses it. That counts, doesn't it?
It counts for something. It is also one component out of fifteen, and it is not the one that does the most work.
When researchers say structured interviews are the strongest predictor of job performance, they mean interviews that were structured in specific, measurable ways. The 2022 meta-analysis by Sackett, Zhang, Berry, and Lievens put structured interviews at an estimated validity of 0.42 and unstructured interviews at 0.19. That gap is the whole reason to care about structure. But the studies behind the 0.42 did not just hand interviewers a question list. They changed how questions were built, how answers were scored, and how the scores were combined.
The useful question is which parts of structure a team is actually doing.
What does a structured interview consist of?
A structured interview is a bundle of design choices, not a single technique. The most-cited breakdown comes from Campion, Palmer, and Campion, whose 1997 review in Personnel Psychology identified 15 components. They fall into two groups: choices about the questions, and choices about the evaluation.
Content: how the questions are built
- Base the questions on a job analysis. Know what the job requires before you write anything.
- Ask every candidate the same questions.
- Limit unplanned prompting and follow-up. Follow-ups are fine when they are planned; free improvisation is where consistency leaks.
- Use better question types. Behavioral ("tell me about a time") and situational ("what would you do if") questions beat opinion questions ("how do you handle stress").
- Use more questions, or a longer interview. More evidence, more reliability.
- Control what else the interviewer knows. Resumes, test scores, and referrals shape impressions before the first question.
- Hold candidate questions until after the scored portion.
Evaluation: how the answers are judged
- Rate each answer, not the whole interview. One global impression is where halo effects live.
- Use detailed anchored rating scales. Write down what a 2 sounds like and what a 4 sounds like.
- Take detailed notes.
- Use more than one interviewer.
- Use the same interviewers across candidates.
- Do not discuss candidates between interviews.
- Train the interviewers.
- Combine scores by a fixed rule, not by gut feel.
The authors' conclusion was blunt for an academic paper: there is no good rationale for a fully unstructured interview. Every component on the list can be added on its own, and each one buys some reliability, validity, or defensibility.
Which components do interviewers actually use?
This is where the "yes, we're structured" answer falls apart.
Roulin, Bourdage, and Wingate (2019) surveyed 131 professional interviewers about seven groups of structure components. On a 5-point scale, interviewers reported high use of the parts that feel like good conversation: follow-up questioning (4.03), question consistency (3.99), rapport building (3.89), and note taking (3.85).
Standardized evaluation, meaning set criteria, numeric scales, and anchored scoring, came in at 2.66.
That is a pattern, not just a number. Interviewers keep the parts of structure that look like interviewing and skip the parts that look like grading. They ask the same questions, then go back to arguing about impressions in the debrief. Chapman and Zweig (2005) found the same shape a decade earlier, and found something else worth knowing: fewer than 34% of their interviewers had any formal interview training.
Lievens and De Paepe (2004) put a finer point on it. In their sample, only 3% of interviewers reported the highest level of structure: all main and follow-up questions set in advance, with each answer rated against example answers. Almost everyone was somewhere in the middle, using enough structure to feel organized and not enough to change the decision.
Standard questions with unstandardized scoring gets you most of the work and a fraction of the benefit.
Which components matter most?
Not all fifteen carry equal weight, and you don't need all fifteen to see a difference.
Campion and colleagues singled out the content components that do the most work as job analysis, same-question discipline, and better question types. On the evaluation side: rating each answer, anchored scales, and interviewer training.
There is a simple logic to why the evaluation components matter so much. An interview rating is a judgment about what the candidate managed to say in one social setting, filtered through an interviewer's memory, mood, and first impression, rather than a direct reading of the candidate's ability. Huffcutt, Van Iddekinge, and Roth (2011) call that middle layer interviewee performance. Structure on the question side helps the candidate bring the right evidence forward. Structure on the evaluation side stops the interviewer from scoring charm, polish, or the fact that the candidate reminded them of themselves.
Highhouse and Brooks (2023) make the same point from the noise angle. Two interviewers see the same answer and score it differently. The same interviewer scores differently at 9 a.m. and 4 p.m. Breaking one vague judgment about whether the team likes a candidate into several narrow ones, about whether they named specifics, explained their own approach, and said what made it hard, shrinks that noise. That only works if each narrow judgment is scored on its own, against a written standard.
What does the training component actually change?
Training gets listed and then ignored, so it is worth being specific about what it does.
Roulin and colleagues found that formal interview training correlated with question consistency (r = .33), note taking (r = .28), and standardized evaluation (r = .39). The strongest link was to the component interviewers use least. Training is how the scoring habit gets built.
Experience doesn't do the same job. In the same study, years of interviewing experience were negatively related to question sophistication (r = -.30). Doing interviews for a long time does not make an interviewer more structured. It can harden the habits they started with.
That matches what most hiring leaders have seen. The veteran who just knows is usually the one whose scores drift the most, because nobody has ever asked them to write down what a 3 means.
A structured interview checklist you can run this week
Take your last hiring loop and go down the list. Be honest about the difference between having something and doing it every time.
| Component | We do this every time | Sometimes | No |
|---|---|---|---|
| Questions come from an analysis of the job | |||
| Every candidate gets the same questions | |||
| Follow-ups are planned, not improvised | |||
| Questions are behavioral or situational, not opinion | |||
| Enough questions to cover the key skills | |||
| Interviewers do not see test scores or referral notes first | |||
| Candidate questions wait until after scoring | |||
| Each answer gets its own score | |||
| Scores use written anchors | |||
| Interviewers take detailed notes | |||
| More than one interviewer | |||
| Same interviewers for every candidate | |||
| No hallway debriefs between interviews | |||
| Interviewers have been trained on the scoring | |||
| Scores combine by a fixed rule |
Most teams check six or seven boxes in the first column, and they cluster in the top half of the table. If your "no" column is mostly in the bottom half, you are in the same place as the 131 interviewers in the 2019 study: organized on the way in, improvised on the way out.
Where to start if you are missing the scoring half
Two changes cover most of the gap.
Score each answer against written anchors before you discuss anything. The anchors don't have to be elaborate. For each question, write what an empty answer sounds like and what a strong one sounds like. Have every interviewer score alone, in writing, before the debrief. Highhouse and Brooks cite assessment-center research where consensus discussion added no validity beyond the underlying measures, while a simple average of independent scores did. The discussion is still useful. It just has to come after the scores exist.
Train on the anchors, not on the theory. Lievens and De Paepe recommend vivid cases over validity statistics. The most effective training session is the one where five interviewers score the same three answers, see that they disagree by a full point, and argue it out. That is the moment the scoring habit forms.
One caveat the research is honest about: structure has variability. Sackett's group reports an 80% credibility interval for structured interviews of 0.18 to 0.66. Implementation is the difference. A team that adopts the question half and skips the evaluation half should expect to land at the low end, and should not be surprised when the debrief still comes down to who liked whom.
The interviewer keeps the conversation
None of this turns an interview into an interrogation. Rapport building is one of the seven components in the 2019 study, and interviewers were right to keep it. The warm open, the follow-up that chases something interesting, the read of the person all stay.
What changes is small and specific. Every candidate for the role answers the same job-based questions, and every answer gets scored against the same written cues, before anyone compares notes. That is the difference between an interview that feels structured and one that predicts.
How Ratio fits
Most of the fifteen components are hard to do by hand for every role: the job analysis, the question writing, the anchors, and the discipline of scoring each answer alone. Ratio builds that structure from whatever you have on the job, whether that is a job description, intake notes, or a hiring-manager conversation. The skills that matter for the role, one question per skill, a planned follow-up for when an answer runs thin, and a written scoring guide for every answer. Interviewers still run the conversation and score the candidate. The parts teams usually skip are the parts they no longer have to build. Book a working session and we will build the interview plan for one of your open roles in front of you.
References
- Campion, M. A., Palmer, D. K., & Campion, J. E. (1997). A review of structure in the selection interview. Personnel Psychology, 50(3), 655-702. https://doi.org/10.1111/j.1744-6570.1997.tb00709.x
- Chapman, D. S., & Zweig, D. I. (2005). Developing a nomological network for interview structure: Antecedents and consequences of the structured selection interview. Personnel Psychology, 58(3), 673-702. https://doi.org/10.1111/j.1744-6570.2005.00516.x
- Highhouse, S., & Brooks, M. E. (2023). Improving workplace judgments by reducing noise: Lessons learned from a century of selection research. Annual Review of Organizational Psychology and Organizational Behavior, 10, 519-533. https://doi.org/10.1146/annurev-orgpsych-120920-050708
- Huffcutt, A. I., Van Iddekinge, C. H., & Roth, P. L. (2011). Understanding applicant behavior in employment interviews: A theoretical model of interviewee performance. Human Resource Management Review, 21(4), 353-367.
- Lievens, F., & De Paepe, A. (2004). An empirical investigation of interviewer-related factors that discourage the use of high structure interviews. Journal of Organizational Behavior, 25(1), 29-46. https://doi.org/10.1002/job.246
- Roulin, N., Bourdage, J. S., & Wingate, T. G. (2019). Who is conducting "better" employment interviews? Antecedents of structured interview components use. Personnel Assessment and Decisions, 5(1), 37-48. https://doi.org/10.25035/pad.2019.01.002
- Sackett, P. R., Zhang, C., Berry, C. M., & Lievens, F. (2022). Revisiting meta-analytic estimates of validity in personnel selection: Addressing systematic overcorrection for restriction of range. Journal of Applied Psychology, 107(11), 2040-2068. https://doi.org/10.1037/apl0000994
Frequently asked questions
What makes an interview structured?
An interview is structured to the degree it uses the components research identifies, not because it has a question list. Campion, Palmer, and Campion (1997) name 15, split between how questions are built (job analysis, the same questions for every candidate, planned rather than improvised follow-ups, behavioral or situational question types) and how answers are judged (rating each answer, anchored rating scales, detailed notes, trained interviewers, combining scores by a fixed rule). Structure is a matter of degree, and each component can be added on its own.
What are the 15 components of a structured interview?
Seven cover content: base questions on a job analysis, ask every candidate the same questions, limit unplanned follow-up, use behavioral or situational question types, use enough questions, control what else the interviewer knows, and hold candidate questions until after the scored portion. Eight cover evaluation: rate each answer rather than the whole interview, use detailed anchored rating scales, take detailed notes, use more than one interviewer, use the same interviewers across candidates, avoid discussing candidates between interviews, train the interviewers, and combine scores by a fixed rule (Campion, Palmer, & Campion, 1997).
Which structured interview components matter most?
On the content side: job analysis, asking every candidate the same questions, and using behavioral or situational question types. On the evaluation side: rating each answer on its own, anchored rating scales, and interviewer training. The evaluation components matter disproportionately because an interview rating is a judgment about what a candidate managed to say in one social setting, filtered through the interviewer's memory, mood, and first impression.
Do most companies actually run structured interviews?
Most run the visible half. In Roulin, Bourdage, and Wingate's 2019 survey of 131 professional interviewers, follow-up questioning scored 4.03 out of 5, question consistency 3.99, rapport building 3.89, and note taking 3.85, while standardized evaluation came in at 2.66. Lievens and De Paepe (2004) found only 3% of interviewers reported the highest level of structure. Standard questions with unstandardized scoring is the most common shape in practice.
Does interviewer experience make someone better at structured interviewing?
No. Roulin, Bourdage, and Wingate (2019) found years of interviewing experience negatively related to question sophistication (r = -.30), while formal training correlated with standardized evaluation at r = .39. Fewer than 34% of interviewers in Chapman and Zweig's (2005) sample had any formal interview training. Experience hardens whatever habits an interviewer started with.
How many of the 15 components do we need?
There is no threshold in the research, and the components are additive rather than all-or-nothing. The practical guidance is to know which ones you skipped, because the evaluation components are both the ones teams skip most and the ones that do the most to reduce noise. Adding written anchors and independent scoring before the debrief covers the largest gap for most teams.
Was this useful?