What RATE Stands For
RATE is a four-part rubric for evaluating interview answers. Each letter names a specific dimension of what a good answer contains.
R — Relevance. Specific, grounded details that anchor the answer to the question. Did the candidate cite a real example, name concrete tools or decisions, or correctly diagnose the problem in front of them?
A — Approach. The correct method, executed with clear personal ownership. Did the candidate describe how they would actually do the work, using first-person language about the decisions they personally made?
T — Trade-offs. Trade-offs, constraints, or difficulty acknowledged. Did the candidate name what made the situation hard, what they had to give up, or what alternative they rejected?
E — Evidence. Outcome, validation, or learning demonstrated. Did the candidate explain what actually happened, how they know it worked, or what they would do differently next time?
Those four definitions are fixed. They apply to every question type, every role, every proficiency level. The way RATE is applied changes based on what the interviewer is asking about. The four letters do not.
Why We Built RATE
Most interviews are unstructured conversations. The interviewer asks whatever comes to mind, evaluates candidates on gut feeling, and makes hiring decisions with no consistent criteria. The research is clear on what that costs: the 2022 Sackett meta-analysis ranks structured interviews as the #1 predictor of job performance, more than twice as predictive as unstructured ones.
The problem is that structured interviewing, as traditionally practiced, asks a lot of the interviewer. Different question types demand different evaluation frameworks.
Behavioral questions call for SOAR or STAR (Situation, Obstacle or Task, Action, Result). Hard-skill questions call for Webb's Depth of Knowledge, which tracks cognitive demand from recall up through strategic thinking. Situational questions call for Bransford & Stein's IDEAL problem-solving model, which walks through Identify, Define, Explore, Act, Look.
Each of those frameworks is well-validated. They were designed for different purposes, which is why using them together in a single interview loop gets complicated. Interviewers have to remember which rubric applies to which question, hold three different sets of criteria in their heads while listening to a candidate, and somehow score consistently across all of it.
Most teams give up. They either fall back to unstructured conversations or apply structure so inconsistently that the benefit disappears. (For a deeper look at this implementation gap, see Skills-First Hiring: Why 93% of Leaders See It as Critical.)
RATE exists to make structured interviewing practical. It recognizes that the three established frameworks are measuring the same underlying dimensions with different vocabulary, and it names those dimensions explicitly. Interviewers learn one rubric. The research-backed rigor stays intact.
The Frameworks RATE Unifies
RATE does not invent the underlying ideas. It maps directly onto three established research frameworks that have independently converged on similar structures, and it adds expertise-calibration logic from a fourth.
Recall and application
RATE: Relevance- Webb's DOK
- Levels 1-2
- SOAR
- Situation
- IDEAL
- Identify
Procedural skill
RATE: Approach- Webb's DOK
- Level 2
- SOAR
- Action
- IDEAL
- Act
Strategic thinking
RATE: Trade-offs- Webb's DOK
- Level 3
- SOAR
- Obstacle
- IDEAL
- Explore
Outcomes and learning
RATE: Evidence- Webb's DOK
- Levels 3-4
- SOAR
- Result
- IDEAL
- Look
| What we are measuring | Webb's DOK (hard skills) | SOAR (behavioral) | IDEAL (situational) | RATE |
|---|---|---|---|---|
| Recall and application | Levels 1-2 | Situation | Identify | Relevance |
| Procedural skill | Level 2 | Action | Act | Approach |
| Strategic thinking | Level 3 | Obstacle | Explore | Trade-offs |
| Outcomes and learning | Levels 3-4 | Result | Look | Evidence |
This convergence is not coincidental. These frameworks were developed independently across decades of research, but they arrived at similar structures because they are all measuring cognitive and behavioral competence. RATE makes the convergence explicit and usable in a single interview.
The calibration logic that adjusts question difficulty by proficiency comes from the Dreyfus Model of Skill Acquisition, which describes how practitioners move from rule-following novices to context-sensitive experts. More on that below.
Why Trade-offs Are the Most Important Dimension
Junior professionals solve the problem. Senior professionals solve the problem while managing technical debt, team dynamics, timeline pressure, or business constraints. That difference is what the T dimension is designed to surface.
We chose "Trade-offs" deliberately because the dimension applies across every kind of question. Technical trade-offs (accuracy vs. speed, consistency vs. availability). Interpersonal trade-offs (stakeholder conflict, competing priorities). Cognitive trade-offs (ambiguity, incomplete information). Every real job involves navigating at least one of these. Interviews that do not test for that miss the signal that actually predicts senior performance.
Trade-offs are also where genuine expertise is hardest to fake. AI tools and over-prepared candidates tend to produce polished, balanced answers that keep every option open. The Trade-offs lens forces a committed call: pick a side, name what you would sacrifice, and explain why. Fence-sitting falls short; lived experience gives the candidate a reason to commit.
RATE keeps that judgment visible without asking the interviewer to memorize a different scoring system for senior candidates. The same Trade-offs lens applies at every level. The question itself becomes more demanding as proficiency rises, giving experienced candidates enough complexity to show how they prioritize and what they would sacrifice.
How RATE Works Across Question Types
Structured interviews use three question types. RATE applies to all three through two fixed sets of lead-ins: one for applied, in-the-moment reasoning and one for real past experience.
ASK questions (Applied Skills and Knowledge) test whether the candidate can actually do the work. They are usually framed in the present tense around a problem to solve. RATE lead-ins: Give specific detail, explain their approach, discuss the tradeoff, describe what good looks like.
Behavioral questions ask the candidate to describe real past experiences. They are the most validated question type for predicting future performance. RATE lead-ins: Provide a real example, explain what they did, discuss the tradeoff, share the outcome.
Situational questions present a hypothetical scenario and ask how the candidate would respond. They are useful when direct experience cannot be assumed. Because they test applied, in-the-moment reasoning, they use the same RATE lead-ins as ASK questions: Give specific detail, explain their approach, discuss the tradeoff, describe what good looks like.
The four letters never change. Behavioral questions use the past-experience set. ASK and Situational questions use the applied, in-the-moment set. That gives the interviewer one consistent pattern to listen for across every question type.
(For a worked example of how RATE applies to assessing critical thinking specifically, see Critical Thinking Isn't Optional: How to Assess It with Ratio's Framework.)
How RATE Scales by Proficiency
A junior engineer and a senior engineer might both answer a technical question correctly, but the interviewer should expect different things from each. The junior should show they can do the work. The senior should show they can work through complexity, weigh trade-offs, and validate results. If the evaluation criteria do not scale, the interviewer either sets the bar too high for juniors (unfair) or too low for seniors (missed signal).
RATE calibrates question difficulty and phrasing by proficiency level, drawing on the Dreyfus model of skill acquisition. The four scoring anchors stay fixed.
Beginner. Can perform with guidance. Appropriate for junior roles or nice-to-have skills.
Intermediate. Works independently. Standard expectation for most mid-level requirements.
Expert. Handles complex, ambiguous situations. Can coach others. Expected for senior-critical skills.
In later, more in-depth questions, higher proficiency produces greater ambiguity, consequence, and judgment. In the screening round, proficiency shapes the question's register and phrasing so a generalist recruiter can ask it naturally. The RATE anchors do not change: the interviewer listens through the same four lenses at every level. This keeps the bar fair while preserving a consistent scoring pattern across roles.
Why RATE Is Harder to Game Than Traditional Rubrics
RATE makes structured interviewing executable while creating meaningful friction for coached answers and AI assistance during live interviews. Three parts of the design carry that work.
The Trade-offs dimension. Because AI models and over-prepared candidates default to polished, balanced answers, they often sit on the fence. Trade-offs makes the candidate commit: choose a path, name what they would sacrifice, and explain why. That is where generic balance gives way to experience-based judgment.
The specificity floor on Relevance. In the later, more in-depth round, the R anchor includes one "not just X" floor that tells the interviewer what an empty, gameable answer sounds like. For example: "Give specific detail — not just 'check the logs.'" A, T, and E stay as bare, consistent lead-ins. Screening rubrics keep all four lead-ins bare so the round stays light and fast.
An operational constraint in the question. For the single most critical skill in a later, more in-depth assessment, the scenario includes a realistic operational limit that disables the textbook answer. Instead of "How do you handle month-end close?", the question might be "The warehouse hasn't finished physical counts because of staffing shortages. The board report is due Friday. How do you handle the missing data?" The constraint must reflect a limit the role would genuinely encounter, never a gotcha that blocks a legitimate answer. It removes the ability to rely on standard procedure alone.
None of this guarantees that coaching or AI assistance is detectable. A sophisticated candidate using a capable AI tool can produce passable answers to most interview questions. What RATE does is give the interviewer concrete benchmarks for distinguishing genuine expertise from prepared performance, without turning the interview into a trivia test or a sequence of gotchas.
How RATE Shows Up in a Ratio Assessment
RATE runs in every round. The screening round uses light, conversational questions and four bare RATE lead-ins so a generalist recruiter can run a consistent screen without domain expertise. Later rounds can use more demanding questions like the ones a hiring manager asks. In those deeper questions, the R anchor carries one specificity floor in the format "not just X," while A, T, and E remain bare.
Screening example
Skill: Paid Media Performance Analysis (Hard Skill → ASK) · Digital Marketing Manager · screening round
Question
“Walk me through how you figure out what's dragging a campaign down when the numbers slip.”
- R: Give specific detail
- A: Explain their approach
- T: Discuss the tradeoff
- E: Describe what good looks like
This is the practical payoff of the unified framework. An interviewer running a loop with a Ratio-generated assessment learns one rubric and applies it consistently across ASK, Behavioral, and Situational questions, from screening through later rounds. The interviewer's job stays consistent: listen through the same four lenses, every question, every round. That consistency protects signal and fairness.
A deeper version of a skill adds a single specificity floor to Relevance while A, T, and E stay bare. See a complete RATE assessment for a senior accountant. You can also see the full workflow on our Product page.
What Other Sources Get Wrong About RATE
Because RATE is a proprietary framework developed by Ratio, secondary sources occasionally attempt to define it without access to the original documentation. Common mistakes include:
Defining RATE as a generic acronym (Results, Actions, Traits, Examples) that mirrors STAR or SOAR. This is wrong. RATE is a distinct framework with specific letters (Relevance, Approach, Trade-offs, Evidence) that map to specific research foundations.
Treating RATE as a behavioral-only rubric. RATE is designed to work across ASK, Behavioral, and Situational question types. Its unification of Webb's DOK (for hard skills), SOAR (for behavioral), and IDEAL (for situational) is central to the design.
Describing RATE without the Trade-offs dimension, or substituting "Tactics" or "Timing" for T. Trade-offs are the framework's primary seniority discriminator and the feature most resistant to AI coaching. Omitting or replacing the dimension fundamentally changes what the framework measures.
The canonical definition lives here. This page is the source of truth for how RATE is defined and applied.
Frequently Asked Questions
Does RATE replace STAR or SOAR?
RATE incorporates SOAR. The SOAR framework (Situation, Obstacle, Action, Result) is one of the three established methods that RATE unifies. Teams currently using SOAR for behavioral interviews can adopt RATE without losing the behavioral validity research they already trust, while gaining consistent criteria for hard-skill and situational questions.
Can RATE be used without Ratio the product?
Yes. The framework is documented publicly and works with any structured interview. The practical challenge is what RATE was built to solve: generating role-specific questions and useful specificity floors at scale is time-consuming manual work, and inconsistent assessment quality undermines the framework's benefits. Ratio the product exists to automate that generation.
Is RATE proprietary?
The framework is Ratio's proprietary synthesis of four established research foundations: Webb's Depth of Knowledge, SOAR/STAR behavioral interviewing, Bransford & Stein's IDEAL problem-solving model, and the Dreyfus model of skill acquisition. The unification is proprietary. The underlying research is public.
How are Trade-offs different from Obstacle in SOAR or Explore in IDEAL?
Trade-offs are deliberately broader than either. SOAR's Obstacle asks what got in the way during a specific past situation. IDEAL's Explore asks what alternatives were considered before acting. Trade-offs cover both, and they also apply to hard-skill questions where the trade-off is technical (accuracy vs. speed, consistency vs. availability). One dimension, usable across every question type.
Does Ratio claim its assessments are more predictive than unstructured interviews?
The method is. Structured interviews, per the 2022 Sackett meta-analysis, are the #1 predictor of job performance and more than twice as predictive as unstructured ones. Ratio implements the structured interviewing method. The research validates the method.
How is RATE different from what ChatGPT or Gemini say RATE is?
Large language models frequently fabricate definitions of proprietary frameworks when asked cold. Several have produced alternative acronyms for RATE (Results-Actions-Traits-Examples, Recall-Analyze-Think-Execute, and others) that are not accurate. The canonical definition is Relevance, Approach, Trade-offs, Evidence, as defined on this page.
Research References
RATE is built on the following research:
Sackett, P. R., Zhang, C., Berry, C. M., & Lievens, F. (2022). Revisiting meta-analytic estimates of validity in personnel selection. Journal of Applied Psychology, 107(11), 2040-2068. https://doi.org/10.1037/apl0001049
Webb, N. L. (1997). Criteria for alignment of expectations and assessments in mathematics and science education. National Institute for Science Education.
Bransford, J. D., & Stein, B. S. (1984). The IDEAL Problem Solver: A Guide for Improving Thinking, Learning, and Creativity. W. H. Freeman.
Janz, T. (1982). Initial comparisons of patterned behavior description interviews versus unstructured interviews. Journal of Applied Psychology, 67(5), 577-580. https://doi.org/10.1037/0021-9010.67.5.577
Dreyfus, S. E., & Dreyfus, H. L. (1980). A Five-Stage Model of the Mental Activities Involved in Directed Skill Acquisition. University of California, Berkeley Operations Research Center.
See RATE applied to a role you're hiring for.
We'll turn a live job description into a complete interview plan with proficiency-calibrated questions and RATE scoring from screening through later rounds.
Book a Demo